InterviewHack.ai
Empezar gratis
Vacantes / GT

Senior Site Reliability Engineer (SRE) | Feeld

GT · BrazilRemotosenior

En corto

  • ▸Ingeniero de Confiabilidad de Sitios (SRE) senior encargado de la observabilidad y resiliencia de viajes de usuario clave.
  • ▸Trabaja dentro de un equipo autónomo distribuido, respondiendo a incidentes críticos y mejorando sistemas de monitoreo y alertas.
  • ▸Destaca el enfoque en la combinación de experiencia de ingeniería backend con SRE, operando de forma autónoma en un entorno global.

No se requiere inglés

Postularme en la empresa ↗Compartir por WhatsApp

En ~1 minuto te damos: quién te entrevista, las preguntas probables con respuestas desde tu CV, y tu CV adaptado a esta vacante. Gratis, sin tarjeta.

¿Qué piden?

  • ✓Experiencia sólida como ingeniero backend/software con nivel senior en Node.js y TypeScript.
  • ✓Experiencia práctica en roles SRE o de ingeniería de producción.
  • ✓Experiencia directa con AWS, Cloudflare y CloudWatch.
  • ✓Habilidad para investigar y responder a incidentes críticos de producción de forma autónoma.
  • ✓Conocimiento sólido en monitoreo, registro, trazado y alertas.
  • ✓Capacidad para definir y mantener SLIs/SLOs para servicios esenciales.

¿No cumplís todo? Es lo normal — tu dossier gratis te dice qué gaps tenés y cómo cubrirlos en la entrevista.

Node.jsTypeScriptAWSCloudflareCloudWatchSentryReact NativeMetricsLoggingTracing

¿A quién escribirle en GT?

Tu dossier gratis identifica a las personas que te entrevistarían — con su background, qué valoran y cómo escribirles para destacar antes de aplicar.

GT was founded in 2019 by a former Apple, Nest, and Google executive. GT’s mission is to connect the world’s best talent with product careers offered by high-growth companies in the UK, USA, Canada, Germany, and the Netherlands. On behalf of Feeld , GT is looking for a Senior Site Reliability Engineer (SRE) to join a fast-growing consumer mobile product in the online dating space. About the Client Founded in 2014 as a dating app, Feeld gathered millions of users in one place to create a safer and more inclusive space online for everyone open to experiencing people and relationships in a new way. Their mission is to elevate the human experience of sexuality and relationships and create a world where everyone is more intimately connected to each other and themselves. About the Project You’ll join a consumer mobile product with an engineering and product organization of around 50 people distributed across Europe and the US. The team works in small, autonomous product squads, each responsible for a specific area of the product and critical user journeys. Because the team operates across multiple regions without a full follow-the-sun model, strong observability, monitoring and reliable incident response are essential. From a technical perspective, the team is focused on building reliable, observable systems that allow engineers to identify issues early, understand their impact and respond quickly when incidents occur. js, TypeScript, AWS, Cloudflare, CloudWatch, Sentry. React Native is used on the mobile side. Team: Cross-functional product squads of approximately 6–8 people, distributed across Europe, the US and LATAM. js and TypeScript . The ideal profile is someone who started in backend/software engineering and has moved into SRE or reliability-focused work, combining a strong understanding of application code with hands-on experience in observability, monitoring and production incident management. You will be embedded within a product squad and take ownership of the reliability and observability of critical user journeys. An important part of the role is being able to interpret production signals, identify when something is going wrong and begin mitigating incidents independently while bringing in the wider engineering team when needed. Responsibilities: Own observability for critical product and user journeys within your squad. Define, build and maintain meaningful metrics, dashboards and alerts. Define and maintain SLIs/SLOs for key services and product-level metrics. Improve monitoring, logging, tracing and alerting across the squad’s systems. Act as the first responder for critical P0/P1 production incidents, including out-of-hours incidents. Investigate production signals, identify potential root causes and begin mitigating issues independently. Coordinate with other engineers when broader support or escalation is required. Participate in incident triage, mitigation and postmortems. Identify recurring reliability issues and drive improvements to infrastructure, tooling and incident-response processes. Work closely with backend and product engineers in a distributed, autonomous squad. js and TypeScript . Hands-on experience working in an SRE, Production Engineering or similar reliability-focused role . Strong production experience with AWS . Experience with Cloudflare and CloudWatch. Experience with monitoring and observability across metrics, logging, tracing and alerting . Practical experience responding to production incidents, including triage, mitigation and postmortems . Ability to interpret monitoring signals and independently investigate and begin resolving production issues.

Más empleos como este

Empleos remotos de DevOps EngineerEmpleos remotos para LATAMEmpleos remotos en Brazil

No apliques sin prepararte

Investigamos quién te entrevista, adaptamos tu CV y te ensayamos en vivo — gratis la primera.

InterviewHack.ai

Preparate para la entrevista exacta: quién te entrevista, tu CV a medida y coach real.

Producto

VacantesRevisar CV (ATS) gratis¿Cómo suena tu inglés?¿Te pagan bien?Reporte de sueldos LATAMCursos gratisBlogCV a medidaCoach realPrecios

Empleos remotos

ReactPythonFull-StackLATAMMéxicoVer todas →

Preparate

Practicá con coachFrontendBackendAI EngineerPor empresaVendete con tu CV

Empresa

Buscás talentoAcerca deContactoPrivacidadTérminos

© 2026 InterviewHack.ai · Tu CV es tuyo. Nunca se usa para entrenar nada. · Un producto de IA-PTY

Vacantes similares activas

Informatica Admin

NTT DATA · LATAM

→