InterviewHack.ai
Start free
Jobs / Yuno

Site Reliability Engineer

Yuno·EuropeRemotesenior

In short

  • →Ingeniero de confiabilidad de plataforma para infraestructura de agentes de IA a escala en AWS.
  • →Diseñas arquitectura, mensajería asíncrona, observabilidad y estrategia de fiabilidad para sistemas críticos.
  • →El rol es líder técnico en fiabilidad, con foco en IA, autosuficiencia y resiliencia proactiva.

Experiencia con inglés en entornos técnicos, dominio de documentación y colaboración en in

Apply on company site ↗Share on WhatsApp
✓ Free to start✓ Runs in your browser✓ First dossier, no card✓ Ready in ~1 minute

In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Your first dossier is free.

The questions they'll ask you

1. ¿Cómo diseñarías una capa de mensajería asíncrona confiable para un sistema de agentes de IA en producción?

2. ¿Qué métricas usarías para definir SLOs en un sistema que procesa pagos globales a alta escala?

3. Describe un caso de chaos engineering que hayas implementado y cómo mejoró la resiliencia del sistema.

🔒 +7 more questions

No card. Upload your resume and the full dossier is ready in ~1 minute.

🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →

💵 USD · Remote · No visa

Not finding what you want? Try Micro1

Micro1 places engineers directly at US companies paying in USD. One vetting, multiple offers — no cold applying.

Get matched by Micro1 →
📬Jobs picked for YOUR resume, every morning on WhatsApp. Free: text “vacantes” and the bot sends your daily matches. Subscribe →

What they ask for

  • ✓7+ años de experiencia en ingeniería de confiabilidad o sistemas distribuidos.
  • ✓Experiencia comprobada en arquitecturas basadas en eventos y mensajería asíncrona (Kafka, SQS, etc).
  • ✓Dominio de IaC (Terraform, AWS CDK) y gestión de infraestructura en AWS.
  • ✓Conocimiento profundo de SLO, error budgets y prácticas de postmortem sin culpa.
  • ✓Experiencia en diseño y evolución de plataformas escalables con alto volumen de transacciones.
  • ✓Capacidad para liderar decisiones técnicas y mentorizar a equipos de ingeniería.

Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.

AWSKafkaTerraformAWS CDKPrometheusGrafanaOpenTelemetryPostgreSQLDockerGit

Who should you write to at Yuno?

Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.

Remote · Full Time · Individual Contributor · +7 Years of Experience Site Reliability Engineer Who We Are Yuno is the AI-native operating system of global commerce, powering the financial infrastructure of enterprise merchants, banks, and wallets. Through a single API, Yuno connects them to pay-ins, payouts, fraud prevention, KYC/KYB, and stablecoins globally, so they can operate everywhere. Agnostic by design and connected to 1,000+ payment methods and 460+ integrations in 190+ countries, Yuno optimizes acceptance rates, reduces costs, and strengthens security through specialized AI agents that learn from every transaction. Global brands including McDonald's, NetEase Games, GoFundMe, and Rappi run their payments on Yuno. About The Role Yuno is looking for a Staff Site Reliability Engineer to set the technical direction for reliability across our infrastructure — starting with the platform that provisions, deploys, and manages AI agents at scale on AWS, the system powering payments across 190+ countries. The platform is in production and growing, and we need the most senior reliability voice in the room to evolve the architecture and make sure it stays reliable, observable, and ready to scale. This is not a "maintain what exists" role, and it's not a single-system role. You'll own the reliability strategy — driving architectural decisions, designing event-driven communication, defining how we measure and defend reliability, and setting the standards other engineering teams build on. How AI Shows Up in This Role The platform you own is Yuno's AI agent infrastructure — provisioning and deploying AI agents at scale, plus the agents that route payments and prevent fraud. Keeping the AI-native layer reliable is the core of the role AI is our default execution layer: you're encouraged to use AI-assisted tooling across automation, runbooks, incident analysis, and root-cause investigations, and to help define how the wider engineering org adopts it. We care how you use it, not whether you do Your Contribution Will Be Reliability strategy and standards — define the SLO culture, error-budget policy, and incident practices that scale across engineering teams, turning reliability from firefighting into a measurable, org-wide discipline Platform architecture and evolution — drive architectural decisions as the platform matures; the deciding voice on choosing technologies, designing systems, and when to evolve the infrastructure Messaging and event-driven architecture — design and own the messaging layer for inter-service communication, replacing synchronous patterns with durable, reliable async messaging Infrastructure and deployment — own the cloud infrastructure, automate provisioning with IaC, and ensure the platform scales reliably as transaction volume grows Observability — build the monitoring, tracing, and alerting that keeps the platform healthy; when something breaks at 3am, your dashboards and alerts should explain why before anyone has to dig Incident leadership and mentorship — act as the senior escalation point for the hardest production problems, run blameless postmortems and root-cause analyses that turn into permanent fixes, and raise the reliability bar by mentoring senior and mid-level engineers Chaos engineering mindset — continuous fault injection and resilience experiments that surface weaknesses before they turn into incidents, plus identifying and proposing resilience patterns to prevent those failures from reaching production. What Success Looks Like Within your first 6–12 months, you've set the reliability strategy for the platform, driven at least one major architectural evolution (event-driven messaging, streaming reliability, or observability), and engineering teams have adopted the SLO and error-budget framework you defined. You're the person Yuno trusts with the hardest reliability calls. Skills You Need Minimum Qualifications Event-driven architecture and messaging systems — you've designed and owned systems

More jobs like this

Remote DevOps Engineer jobs

Looking for something similar?

Leave your email and we'll alert you when matching jobs appear.

Don't apply unprepared

We research who's interviewing you, tailor your CV and rehearse you live — first one free.

InterviewHack.ai

Prepare for the exact interview: who's interviewing you, a tailored CV, and a real coach.

Product

JobsCompanies hiringAll free toolsResume verdict (Jev)Free cover letterInterview questions by roleTechnical assessment simulator"Tell me about yourself" answerFree ATS checkerInterview-English checkSalary checkSalary negotiation scriptFree STAR answerLinkedIn headline + AboutLATAM salary reportFree coursesBlogTailored CVSpoken practicePricingAffiliates — 30%

Remote jobs

ReactPythonFull-StackLATAMArgentinaMexicoSee all →

Prepare

Spoken practiceFrontendBackendAI EngineerBy companySell with your CV

Company

For employersAboutContactPrivacyTerms

© 2026 InterviewHack.ai · Your CV is yours. Never used to train anything. · A product of IA-PTY

Similar open roles

Staff Engineer - Money, Risk & Payment Ancillaries

Yuno · Europe

→

Staff Engineer - Client Experience

Yuno · Europe

→

Staff AI Engineer, Payments Intelligence

Yuno · Europe

→

Staff Engineer - Trust & Vault

Yuno · Europe

→