InterviewHack.ai
Start free
Jobs / Replit

Engineering Manager, Site Reliability Engineering

Replit·Foster City, CARemotelead$20,833–$27,083 USD/mo

In short

  • →Liderar equipo SRE en una plataforma de desarrollo con IA a escala global.
  • →Enfocado en observabilidad, pruebas de carga, gestión de incidentes y mejora continua del rendimiento.
  • →Papel técnico y de liderazgo: se espera participación directa en análisis de fallos y diseño de sistemas.

Fluent in English

Apply on company site ↗Share on WhatsApp

In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.

🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →

What they ask for

  • ✓Liderazgo demostrado en equipos de ingeniería.
  • ✓Experiencia en sistemas distribuidos y plataformas de confiabilidad.
  • ✓Capacidad para diagnosticar problemas de rendimiento y fiabilidad con datos reales.
  • ✓Habilidades para construir y escalar equipos técnicos con propiedad sostenible.
  • ✓Experiencia en herramientas de observabilidad, pruebas de carga y gestión de incidentes.
  • ✓Capacidad para trabajar con equipos distribuidos y fomentar el desarrollo de líderes técnicos.

Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.

KubernetesobservabilitymetricslogstracesalertingSLOsload testingperformance engineeringAI coding tools

Who should you write to at Replit?

Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.

Compensation: $250K – $325K • Offers Equity. Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. ABOUT THE ROLE Replit enables people to build software with AI. The systems underneath that experience must support safe production changes, measurable reliability, and predictable performance as usage grows. This Engineering Manager will lead SRE across observability, incident management, load testing, performance engineering, cloud cost and capacity, and rollout infrastructure. You'll lead and grow an existing team that builds and operates production platforms and works hands-on across application and infrastructure boundaries. This is a software-building leadership role, not simply an incident-management function. You'll help teams ship safely, understand production behavior, and remove performance bottlenecks through concrete engineering improvements. You should be comfortable going deep on a rollout failure or performance investigation while developing technical leaders and sustainable ownership across a distributed team. WHAT YOU'LL DO - Observability. Build and operate metrics, logs, traces, and alerting capabilities. Help teams establish meaningful SLOs and use production telemetry to diagnose problems and verify improvements. - Incident Management. Own incident tooling and practices, coordinate cross-team response, and turn incident reviews into engineering improvements that reduce recovery time and repeat failures. - Load Testing. Build and maintain load/failure testing capabilities. Validate critical paths under expected demand, quantify headroom, and test recovery and production readiness with service owners. - Performance Engineering. Lead deep engagements with internal teams on SLOs and end-to-end performance. Use profiling, telemetry, and load tests to identify bottlenecks and deliver improvements with service owners—not just recommendations. - Stay technically engaged. Review designs and production changes, debug difficult failure modes, and use AI coding tools—including Replit—to prototype and automate. Apply rigorous review and verification to AI-generated changes. - Build and grow a high-ownership engineering team. Coach engineers, develop technical leaders, manage performance, and hire against agreed needs. Make distributed collaboration, mentoring, and backup coverage deliberate rather than relying on a few permanent escalation points. - Measure outcomes and close the loop. Track rollout safety, recovery time, repeat incidents, critical-path latency/throughput, test coverage, and improvements arising from cost/capacity analysis. Agree success measures and continuing ownership with partner teams. WHAT YOU'LL BRING - Demonstrated engineering management. You have led and developed engineers, made prioritization and performance decisions, hired thoughtfully, and delivered through a team—not only acted as its strongest individual contributor. - Software-oriented production systems depth. You have built and operated distributed systems or reliability platforms and can reason across deployment behavior, Kubernetes, telemetry, service dependencies, and recovery mechanisms. - Safe-change and performance judgment. You have led consequential migrations or incidents and used measurement to diagnose reliability or performance problems. You can distinguish symptoms from causes and validate fixes under realistic conditions. - Platform-product and cross-team judgment. You can build capabilities other teams adopt, lead hands-on engagements without absorbing every service's operations, and make clear tradeoffs among reliability, performance, engineering effort, and cost. NICE TO HAVE - Experience with GitOps or progressive-delivery platforms such as Harness, ArgoCD, or Kargo. - Experience w

More jobs like this

Remote DevOps Engineer jobsRemote Engineering Manager jobs

Looking for something similar?

Leave your email and we'll alert you when matching jobs appear.

Don't apply unprepared

We research who's interviewing you, tailor your CV and rehearse you live — first one free.

InterviewHack.ai

Prepare for the exact interview: who's interviewing you, a tailored CV, and a real coach.

Product

JobsCompanies hiringFree ATS checkerInterview-English checkSalary checkLATAM salary reportFree coursesBlogTailored CVSpoken practiceIt's free

Remote jobs

ReactPythonFull-StackLATAMArgentinaMexicoSee all →

Prepare

Spoken practiceFrontendBackendAI EngineerBy companySell with your CV

Company

For employersAboutContactPrivacyTerms

© 2026 InterviewHack.ai · Your CV is yours. Never used to train anything. · A product of IA-PTY

Similar open roles

Staff Software Engineer, Agentic Ads

Replit · Foster City, CA

→

Engineering Manager, Cloud Infrastructure

Replit · Foster City, CA

→

Support Engineer I (NYC, Weekend Shift)

Replit · NYC (SoHo)

→

Support Engineer I (FC, Weekend Shift)

Replit · Foster City, CA

→