Locations: Spain, Bulgaria, or Portugal (relocation assistance may be provided).
Employment: Full-Time.
About the Company/Product:
A global B2B product company developing a powerful Integration Platform as a Service (iPaaS) that uses AI and machine learning to enable organizations to seamlessly connect data sources, cloud applications, and enterprise systems through low-code/no-code automation.
The platform is trusted by more than 400,000 customers worldwide, including leading companies such as Visa, Goldman Sachs, Cisco, Amazon, HubSpot, and L’Oréal.
The engineering culture emphasizes technical excellence, strong ownership, and close collaboration across globally distributed, remote-first teams.
What You’ll Do:
We are hiring a Senior DBA/SRE to own the health, performance, and day-to-day operations of the PostgreSQL fleet behind the platform. The fleet is ~180 Aurora PostgreSQL clusters across production data centers worldwide. Some databases hold several terabytes, most clusters already run PostgreSQL 17, and the fleet is set to grow to a multiple of its current size. Your responsibilities may include:
Keeping the Aurora PostgreSQL fleet healthy across all data centers: building indexes on large and partitioned tables without downtime, analyzing plans, and moving data.
Handling Aurora configuration (parameters, replicas, failover, capacity) and cost work: rightsizing, merging low-activity databases, and evaluating serverless.
Taking ownership of the PgBouncer layer, which currently needs an owner and a roadmap.
Reviewing schemas and data models with product teams, and advising on partitioning, retention, and shared vs. dedicated clusters, backing recommendations with data.
Working on the performance of the main application and the search queries behind AI features, including research toward a 10,000 jobs/second target.
Packaging the team's existing zero-downtime upgrade and switchover procedures into versioned, released automation, along with provisioning of databases and users, grants, and schema sync across data centers (Terraform, Ansible, GitOps).
Deploying tooling for monitoring and managing the PostgreSQL fleet in Kubernetes, and developing monitoring and alerting (exporters, VictoriaMetrics).
Handling database incidents end-to-end, including root-cause analysis and follow-up fixes, and taking part in resilience projects: a cross-region DR pilot, application reconnect after failover, and CDC pipelines that survive upgrades and failovers.
Contributing to time-boxed evaluations of new database technologies that end with a decision on what the company adopts.
Participating in the on-call rotation, with off-rotation time going to uninterrupted project work.
Core Tech Stack: Aurora PostgreSQL, PgBouncer, AWS, Kubernetes, Terraform, Ansible, Linux, VictoriaMetrics, Vault.
What You Have:
5+ years of experience operating PostgreSQL in production (DBA, DBRE, or SRE with a database focus) on databases under real load.
Deep experience with database schemas: design, debugging, and troubleshooting performance issues in production.
SQL proficiency, query plans, indexing strategy, and performance work on multi-TB tables.
Expertise in the ecosystem around PostgreSQL, not only the database itself: streaming and logical replication, HA and failover setups (Patroni or equivalent), connection poolers (PgBouncer), major-version upgrades and migrations with mini
Анонимная аналитика
Мы используем анонимную аналитику, чтобы улучшать поиск вакансий и работу сайта.