Senior Data Platform Engineer
Gemini•3h ago
United StatesRemoteFull-timeSenior Level5+ yrs exp
Visa-friendly
Top focus
Platform EngineerSenior Data EngineerData EngineerVp DataData Warehouse Engineer
- About the Company
- Gemini is a global crypto and Web3 platform founded by Cameron and Tyler Winklevoss in 2014, offering a wide range of simple, reliable
- secure crypto products and services to individuals and institutions in over 70 countries. Our mission is to unlock the next era of financial, creative
- personal freedom by providing trusted access to the decentralized future. We envision a world where crypto reshapes the global financial system, internet
- money to create greater choice, independence
- opportunity for all — bridging traditional finance with the emerging cryptoeconomy in a way that is more open, fair
- secure. As a publicly traded company, Gemini is poised to accelerate this vision with greater scale, reach
- The Department: Platform
- Our Platform organization’s purpose is to enable Gemini to scale effectively and empower our engineering teams to focus on building innovative financial products and experiences for individuals around the world. Platform focuses around building a scalable and secure foundations platform, enabling Engineering to deploy, validate
- operate their services in production, improve resiliency of the service and increase organizational efficiency by reducing operational toil and increase system efficiency through architectural evolution.
- The Platform team engages directly with our other engineering teams to onboard them onto our platform systems, reviewing and recommending design and architectural decisions
- guiding our engineering teams on how to implement the tooling provided by the larger Platform organization required to ensure systems can scale and react to changing conditions, with continuous improvement loops.
- The Role: Senior Data Platform Engineer
- As a Senior Data Platform Engineer, you'll be an SRE for our data infrastructure: responsible for the reliability, automation
- operability of the database and datastore fleet, not just tuning any single engine. You'll work closely with data engineering and product engineering teams to build self-service, automated foundations that let teams provision, scale
- operate their own datastores safely.
- Your role centers on deep relational database expertise (PostgreSQL, Aurora) - internals, replication, failover, backup/recovery
- query performance - paired with the automation and tooling to run that expertise at scale rather than by hand. You'll extend that same automation-first discipline to the other datastores in our fleet - NoSQL/document, columnar, key-value
- streaming systems - so that scaling, failover
- provisioning are repeatable and largely self-service. You'll drive improvements to observability, incident response
- uptime posture through proactive ownership of the platform
- you'll bring SRE practices (SLOs, error budgets, toil reduction, blameless postmortems) to how we run data infrastructure.
- This role is ideal for someone who thinks in systems and automation first, thrives on cross-team collaboration
- wants to reduce operational toil across a diverse set of database technologies in a fast-paced, cloud-native environment
Responsibilities
- Automation and Reliability Engineering: Build Infrastructure as Code (IaC), CLI tools
- CI/CD-driven automation that make database provisioning, scaling, failover
- deployment self-service, consistent
- repeatable across environments - this is the core of the role, not a supporting activity.
- Database Scaling and Optimization: Serve as the team's depth on relational database systems (e.g., Amazon Aurora, PostgreSQL) - replication topologies, failover, backup/recovery
- query/engine-level performance - ensuring high performance and availability under growing workloads.
- Fleet-Wide Infrastructure Design: Extend that operational rigor to the rest of the datastore fleet - document, key-value
- columnar systems - applying the right paradigm to the right workload and building common tooling and guardrails across all of them.
- High Availability, Observability
- SRE Practice: Define and track SLOs/error budgets, build proactive monitoring and alerting, implement high-availability architectures
- participate in the on-call rotation to troubleshoot and resolve production issues quickly.
- Pipeline Integration: Collaborate with data and product engineering teams to integrate with upstream and downstream pipelines - both real-time and batch - via message queues (e.g., Kafka), ETL workflows
- processing frameworks.
- Performance Tuning and Troubleshooting: Identify and resolve performance bottlenecks at both the query and infrastructure levels across engines. Establish alerting, observability
- incident response procedures that reduce MTTR and maintain service health.
- Toil Reduction and Operational Excellence: Continuously identify and automate away repetitive operational work
- contribute to shared documentation, incident retrospectives, and platform playbooks to improve team effectiveness and reliability of operations
Qualifications
- 5 years of experience in the field.
- Deep, specialist-level experience managing and scaling relational databases - cloud-native systems like PostgreSQL, Amazon Aurora
- similar - including replication, failover, backup/recovery
- query/engine performance tuning.
- Demonstrated SRE mindset: experience building automation, self-service tooling, and guardrails that eliminate manual, repetitive database operations rather than performing them by hand.
- Hands-on experience with at least one non-relational paradigm in production (e.g., NoSQL, columnar, document, key-value), and working knowledge of when to apply each.
- Familiarity with cloud-based data platforms and services such as AWS RDS, Redshift, EMR, Google BigQuery, or Databricks.
- Experience in an infrastructure as code environment (Terraform), developing automated solutions to solve support and operational issues.
- Proficiency writing scripts, CLIs, or services that increase developer productivity and reduce operational toil, in languages like Python, Go, etc.
- Understanding of CI/CD, observability tooling, SLOs/error budgets, and incident response in production environments.
- Experience integrating with data pipelines and real-time messaging systems like Kafka or Kinesis.
- Comfortable participating in on-call rotations and owning uptime and recovery responsibilities across multiple database technologies.
- Strong communication and collaboration skills; able to work effectively across infrastructure, data, and product teams.
- It Pays to Work Here
- The compensation & benefits package for this role includes:
- Competitive starting pay
- A discretionary annual bonus
- Long-term incentive in the form of a new hire equity grant
- Comprehensive health plans
- 401K with company matching
- Paid Parental Leave
- Flexible time off
- Salary Range : The base salary range for this role is between $126,000 - $180,000 in the State of New York. This range is not inclusive of our discretionary bonus or equity package. When determining a candidate’s compensation, we consider a number of factors including skillset, experience, job scope, and current market data.
- In the United States, we offer a hybrid work approach at our hub offices, balancing the benefits of in-person collaboration with the flexibility of remote work. Expectations may vary by location and role, so candidates are encouraged to connect with their recruiter to learn more about the specific policy for the role. Employees who do not live near one of our hubs are part of our remote workforce. All employees, however, are required to onboard in-person at one of our office locations.
- At Gemini, we strive to build diverse teams that reflect the people we want to empower through our products
- we are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity
- Veteran status. Equal Opportunity is the Law
- Gemini is proud to be an equal opportunity workplace. If you have a specific need that requires accommodation, please let a member of the People Team know.
- #LI-AA1
Required skills
PythonGoReactPostgreSQLBigQueryDatabricksKafkaAWSTerraformCI/CD