Senior PostgreSQL DBA/SRE (iPaaS, remote, FTE)
Агентство / HR ресурс NEWHR ( new.hr )
Опыт работы более 5 лет
Locations: Spain, Bulgaria, or Portugal (relocation assistance may be provided).
Employment: Full-Time.
About the Company/Product:
- A global B2B product company developing a powerful Integration Platform as a Service (iPaaS) that uses AI and machine learning to enable organizations to seamlessly connect data sources, cloud applications, and enterprise systems through low-code/no-code automation.
- The platform is trusted by more than 400,000 customers worldwide, including leading companies such as Visa, Goldman Sachs, Cisco, Amazon, HubSpot, and L’Oréal.
- The engineering culture emphasizes technical excellence, strong ownership, and close collaboration across globally distributed, remote-first teams.
What You’ll Do:
We are hiring a Senior DBA/SRE to own the health, performance, and day-to-day operations of the PostgreSQL fleet behind the platform. The fleet is ~180 Aurora PostgreSQL clusters across production data centers worldwide. Some databases hold several terabytes, most clusters already run PostgreSQL 17, and the fleet is set to grow to a multiple of its current size. Your responsibilities may include:
- Keeping the Aurora PostgreSQL fleet healthy across all data centers: building indexes on large and partitioned tables without downtime, analyzing plans, and moving data.
- Handling Aurora configuration (parameters, replicas, failover, capacity) and cost work: rightsizing, merging low-activity databases, and evaluating serverless.
- Taking ownership of the PgBouncer layer, which currently needs an owner and a roadmap.
- Reviewing schemas and data models with product teams, and advising on partitioning, retention, and shared vs. dedicated clusters, backing recommendations with data.
- Working on the performance of the main application and the search queries behind AI features, including research toward a 10,000 jobs/second target.
- Packaging the team's existing zero-downtime upgrade and switchover procedures into versioned, released automation, along with provisioning of databases and users, grants, and schema sync across data centers (Terraform, Ansible, GitOps).
- Deploying tooling for monitoring and managing the PostgreSQL fleet in Kubernetes, and developing monitoring and alerting (exporters, VictoriaMetrics).
- Handling database incidents end-to-end, including root-cause analysis and follow-up fixes, and taking part in resilience projects: a cross-region DR pilot, application reconnect after failover, and CDC pipelines that survive upgrades and failovers.
- Contributing to time-boxed evaluations of new database technologies that end with a decision on what the company adopts.
- Participating in the on-call rotation, with off-rotation time going to uninterrupted project work.
Core Tech Stack: Aurora PostgreSQL, PgBouncer, AWS, Kubernetes, Terraform, Ansible, Linux, VictoriaMetrics, Vault.
What You Have:
- 5+ years of experience operating PostgreSQL in production (DBA, DBRE, or SRE with a database focus) on databases under real load.
- Deep experience with database schemas: design, debugging, and troubleshooting performance issues in production.
- SQL proficiency, query plans, indexing strategy, and performance work on multi-TB tables.
- Expertise in the ecosystem around PostgreSQL, not only the database itself: streaming and logical replication, HA and failover setups (Patroni or equivalent), connection poolers (PgBouncer), major-version upgrades and migrations with minimal downtime.
- Linux fundamentals: processes and memory, OOM behavior, networking basics.
- Scripting skills (Bash, Python, or Go) and an automation-first attitude: you prefer to script a repeated manual task when possible.
- Readiness to work actively with Kubernetes: deploying tooling for monitoring and managing the PostgreSQL fleet.
- Strong RCA and incident management skills.
- Good level of English — at least Upper-intermediate (B2).
Nice-to-Haves:
- Experience with Infrastructure as Code (Terraform) and cloud infrastructure (AWS, ideally hands-on Aurora or RDS).
- Experience with different storage systems: ClickHouse (strong plus), OpenSearch, Elasticsearch, Redis, Valkey, MemoryDB, Neo4j.
- Exposure to Kubernetes basics, GitOps, ArgoCD, or Vault.
- Contributions to open-source database or infrastructure projects.
- Openness to using AI assistants in daily work.
What We Offer:
- Real scale: every database behind a global automation platform, processing billions of events daily.
- Influence on the stack: your analysis in technology reviews shapes which databases the company adopts.
- Top-tier engineering team.
- Competitive base salary + stock options.
- Fully remote work with flexible hours from Spain, Bulgaria, or Portugal.
- Possible assistance with relocation to Spain or Bulgaria (visa support & relocation package).
- Full-time employment with benefits.
- Health and life insurance.
- Flexible vacation policy and paid sick leave.
- Company-provided equipment.
- Company-sponsored team-bonding activities.
- Other competitive benefits depending on location (e.g., wellness programs, education budget, referral bonuses, etc.).
