Databento
Databento
Databento oferă infrastructură pentru date de piață prin API-uri dedicate datelor financiare în timp real și istorice. Platforma sa furnizează informații normalizate despre contracte futures, opțiuni, acțiuni și alte clase de active, inclusiv fluxuri live, seturi de date istorice, detalii despre acțiuni corporative și redarea completă a registrelor de ordine ale burselor. Datele provin din centre de colocare pentru a susține accesul cu latență redusă, iar instrumentele și integrările pentru dezvoltatori, compatibile cu Python și C++, ajută echipele să creeze și să opereze aplicații pentru date de piață. Databento deservește peste 3.000 de companii și startupuri, oferind opțiuni flexibile de acces și tarifare.

Site Reliability Engineer — Remote, United States

Maintain reliability, performance, and observability for Databento’s financial market-data platform. Improve deployments, incident response, and backend infrastructure across API and platform services.

Descriere

  • Own uptime, SLAs, and SLOs for API and platform services
  • Establish reliability and operational standards for developers
  • Build and maintain logging, metrics, and tracing observability
  • Design and operate highly available deployments and containerized environments
  • Profile and optimize Python applications for throughput, latency, and cost
  • Diagnose production issues at the operating-system level with strace, perf, eBPF, ss, and gdb
  • Strengthen deployment processes and CI/CD workflows
  • Join the on-call rotation, lead incident response, and facilitate post-incident reviews
  • Determine necessary fixes and carry projects from initial concept through completion
  • Contribute to petabyte-scale data processing, customer management and billing, and query systems that power APIs

Cerințe

  • Mid-level or senior individual contributor
  • Professional experience in SRE, DevOps, or backend engineering, ideally within a trading firm, technology company, or high-growth startup
  • Practical experience with observability tools for logs, metrics, and traces, including Prometheus, OpenTelemetry, VictoriaMetrics, Jaeger, Logstash, Loki, or Vector
  • Experience with containerization and highly available deployment using technologies such as Docker, Podman, Docker Compose, Docker Swarm, Kubernetes, or k3s
  • Advanced Python skills, including application development and performance tuning
  • Proficiency with Linux debugging and profiling tools such as strace, perf, eBPF, ss, and gdb
  • Demonstrated measurable results in a recent position
  • Experience applying alerting and incident-response practices is advantageous
  • Familiarity with configuration management or infrastructure-as-code tools such as Ansible or Terraform is beneficial
  • Experience with HTTP benchmarking, load testing, and capacity planning is a bonus
  • Database schema design and query optimization experience is desirable
  • Strong communication skills and a dependable work ethic in a remote environment
  • Interest in financial data or algorithmic trading

Beneficii

  • Equal employment opportunity and protection against discrimination
  • Workplace accommodations available upon request
  • Option to opt out of AI-powered Talent Matching

Locuri de muncă similare

Terumo Medical Corporation

Region Manager, Terumo Interventional Systems Sales

Terumo Medical Corporation

Lead medical device sales across a North Central New Jersey region, managing field teams and hospital relationships. Drive regional revenue, sales performance, and compliant promotion of Terumo Interventional Systems products.

Deschide