Trumid 2026 Best Places to Work
Trumid Logo

Trumid

Senior Database Reliability Engineer (DBRE)

Posted 17 Minutes Ago
Be an Early Applicant
Easy Apply
Remote or Hybrid
Hiring Remotely in USA
225K-265K Annually
Senior level
Easy Apply
Remote or Hybrid
Hiring Remotely in USA
225K-265K Annually
Senior level
Own the durability, recoverability, performance, and security of a production PostgreSQL/RDS fleet supporting a live trading platform. Lead replication, failover, backup and restore, disaster-recovery drills, data lifecycle management, access control, encryption, and database observability. Investigate engine-level performance issues including WAL contention, replica lag, bloat, and locking. Build infrastructure and AI-assisted operational tooling while documenting runbooks and reliability decisions.
The summary above was generated by AI

About us.

Trumid is a dynamic fintech revolutionizing the landscape of fixed income trading. With intelligent, easy-to-use, electronic solutions, we are rapidly growing and seeking exceptional talent to help redefine the boundaries of technology and finance.

Founded in 2014 by a team of fixed income market experts, Trumid has quickly become one of the top three corporate bond e-trading platforms in the U.S. Today, over 1,300 traders from an extensive and expanding client network of 890+ buy-and sell-side institutions transact on Trumid monthly.

With a rich history of innovation and a unique ability to innovate at scale, we collaborate closely with our clients, iterating quickly toward optimal solutions. With market share and client engagement at all-time highs and our pace of product development faster than ever, this is an exciting and transformative time at Trumid.

Our business model thrives on participation, and so does our company culture. We rely on every team member’s contribution to help us accomplish our goals. To succeed at Trumid, you must be curious, passionate about your craft, ambitious, collaborative, and driven. Learn more at www.trumid.com

The opportunity.

This role owns data resilience and continuity, and the scope test is simple: if losing it loses data, or makes data unavailable, it’s yours. The job exists so that data-layer failure modes — storage contention, replica lag, region loss — are found and retired in drills, not discovered in production. Our reliability doctrine is to assume failure and concentrate statefulness into a small core of systems proven against specific failure modes. Postgres is the heart of that core: everything around it gets to be disruptible because the data layer is not.

You’d join the databases side of our SRE team, working alongside deep Postgres expertise. Some of what you’d walk into:

  • A production RDS fleet backing a live trading venue, with performance work that goes deep: we’ve characterized WAL-write contention under concurrent commits down to the fsync level, and are weighing group-commit tuning, dedicated log volumes, and storage-class changes against actual measurements.
  • Disaster recovery as an engineering discipline: automated cross-region failover with promotion measured in minutes, and restore paths (snapshot, point-in-time, logical) validated by timed, documented drills on a fixed cadence.
  • Database observability and AI-assisted tooling: engine performance telemetry exported into Prometheus and Grafana, modular database health-check skills, and an automated reviewer for schema-migration PRs.

What you’ll do?

  • Own the durability, recoverability, and performance of the Postgres/RDS fleet across every environment: replication, failover, backup and restore, and storage behavior under load.
  • Make recovery provable: restore-tested coverage of every production database, timed failover drills, and measured RPO/RTO per tier — evidence, not assertion.
  • Own the data lifecycle end to end: retention and cleanup policies that preserve recoverability, access control at the data layer, encryption posture, and knowing where sensitive data lives.
  • Hunt performance pathologies at the engine level: lock contention, WAL throughput, replica lag, bloat, index hygiene, write amplification.
  • Build database observability with the team, and extend our AI-assisted operations tooling (health-check skills, migration PR review).

About you.

  • Five or more years running production PostgreSQL at meaningful scale — ideally on RDS or Aurora — with depth in the internals: replication, WAL mechanics, MVCC and vacuum behavior, query planning and performance.
  • You’ve owned the full lifecycle of data somewhere, not just the query path: retention, backup and restore, access control, and security posture.
  • Disaster-recovery experience you can talk through concretely — failovers you designed, drills you ran, and what they changed.
  • Infrastructure fluency: infrastructure-as-code (Terraform or similar), scripting (Python, bash, SQL), Linux, and cloud storage/IOPS characteristics.
  • An SRE sensibility: SLOs, blameless postmortems, and a preference for rehearsed over improvised.
  • Clear writing — you leave runbooks and decision records behind you.

Nice to have.

  • Warehouse and pipeline experience (BigQuery, AlloyDB, Kafka-based pipelines) — over time we intend to extend the same reliability guarantees beyond Postgres to every store that holds business data.
  • Experience in regulated or fintech environments; exposure to data classification and compliance review.
  • Kubernetes; Prometheus/Grafana/ELK-style observability stacks.
  • Interest in AI-assisted operations tooling

Employee benefits.

  • Highly competitive compensation
  • Fully paid medical, dental and vision coverage
  • Team-oriented and collaborative company culture
  • Lively and dynamic office space with fully stocked kitchen

In compliance with New York City Pay Transparency Law, the base salary range for this role in New York City is between $225,000 – $265,000. This range does not include discretionary bonus or other forms of compensation or benefits offered in connection with this job. Several factors are considered when determining a candidate’s compensation.

Trumid is an equal opportunity employer.

Please note: All communication from Trumid’s recruiting team comes from @trumid.com email addresses. We conduct remote interviews via Zoom only. We will never ask you to purchase equipment, download software (other than Zoom), or share sensitive personal information during the hiring process.

Similar Jobs at Trumid

17 Minutes Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
225K-250K Annually
Senior level
225K-250K Annually
Senior level
Fintech • Information Technology • Software • Financial Services
Own end-to-end network reliability across data centers, AWS, and GCP, including routing, switching, DNS, load balancing, service mesh, VPNs, circuits, and cross-connects. Automate network configuration with Ansible, Terraform, GitOps, CI validation, and IPAM. Engineer failover, observability, security hardening, client connectivity, and Kubernetes ingress while participating in an interrupt-contained on-call rotation.
Top Skills: AnsibleAWSBgpCisco Ios-XeDnsElkEnvoyGCPGitopsGrafanaHaproxyInfrahubIpamKubernetesNautobotNetboxNginxPalo AltoPanoramaPrometheusTerraformTlsVpn
8 Days Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
150K-175K Annually
Mid level
150K-175K Annually
Mid level
Fintech • Information Technology • Software • Financial Services
Own technology compliance and security operations across AI governance, SOC 2, risk assessments, incident response, business continuity, vulnerability management, access reviews, and penetration testing. Manage client security questionnaires, vendor assessments, third-party risk, and audit coordination while monitoring security tools and coordinating remediation with engineering. Advise clients, auditors, regulators, legal teams, and compliance stakeholders on technical security matters.
Top Skills: Ai GatewayAWSCrowdstrike Falcon CompleteEu Ai ActGCPIso 27001Llm Monitoring PlatformsMfaNist Ai RmfObservability ToolsSoc 2SsoVanta
28 Days Ago
Easy Apply
Remote or Hybrid
USA
Easy Apply
175K-250K Annually
Senior level
175K-250K Annually
Senior level
Fintech • Information Technology • Software • Financial Services
Design, build, test, and maintain low-latency, high-throughput trading services on an Aeron-based platform. Expand trading protocol features, migrate workflows to the new stack, optimize performance and scalability, support production reliability, and collaborate with cross-functional teams while using AI-assisted development tools.
Top Skills: AeronCi/CdDockerFixGrpcJvmKafkaKubernetesPekkoScalaTcp/IpZio

What you need to know about the Los Angeles Tech Scene

Los Angeles is a global leader in entertainment, so it’s no surprise that many of the biggest players in streaming, digital media and game development call the city home. But the city boasts plenty of non-entertainment innovation as well, with tech companies spanning verticals like AI, fintech, e-commerce and biotech. With major universities like Caltech, UCLA, USC and the nearby UC Irvine, the city has a steady supply of top-flight tech and engineering talent — not counting the graduates flocking to Los Angeles from across the world to enjoy its beaches, culture and year-round temperate climate.

Key Facts About Los Angeles Tech

  • Number of Tech Workers: 375,800; 5.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Snap, Netflix, SpaceX, Disney, Google
  • Key Industries: Artificial intelligence, adtech, media, software, game development
  • Funding Landscape: $11.6 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Strong Ventures, Fifth Wall, Upfront Ventures, Mucker Capital, Kittyhawk Ventures
  • Research Centers and Universities: California Institute of Technology, UCLA, University of Southern California, UC Irvine, Pepperdine, California Institute for Immunology and Immunotherapy, Center for Quantum Science and Engineering

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account