InterviewHack.ai
Start free
Jobs / Onapsis

Site Reliability Engineer II

Onapsis·Dallas, Texas, United Statesmid

In short

  • →Ingeniero de Confiabilidad de Sitios (SRE) en un equipo global que mantiene plataformas de seguridad críticas en AWS.
  • →Días típicos: resolver incidentes en producción, escribir código para automatizar operaciones y mejorar monitoreo con observabilidad.
  • →Destacado: se valora el uso de asistentes de IA para programar, depurar y prototipar rápidamente.

Experience using AI coding assistants for daily tasks, debugging, and rapid prototyping

Apply on company site ↗Share on WhatsApp
✓ Free to start✓ Runs in your browser✓ First dossier, no card✓ Ready in ~1 minute

In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Your first dossier is free.

The questions they'll ask you

1. ¿Cómo usarías Python y Bash para automatizar la detección de fallos en un servicio de RabbitMQ en producción?

2. ¿Cómo definirías un SLO para un servicio de monitoreo de seguridad con un error budget de 10% mensual?

3. ¿Qué métricas y traces incluirías en un sistema de observabilidad para una aplicación de seguridad en EKS con alta latencia?

🔒 +7 more questions

No card. Upload your resume and the full dossier is ready in ~1 minute.

🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →

💵 USD · Remote · No visa

Not finding what you want? Try Micro1

Micro1 places engineers directly at US companies paying in USD. One vetting, multiple offers — no cold applying.

Get matched by Micro1 →
📬Jobs picked for YOUR resume, every morning on WhatsApp. Free: text “vacantes” and the bot sends your daily matches. Subscribe →

What they ask for

  • ✓2-4 años en roles de SRE, DevOps o Ingeniería en la nube con sistemas distribuidos
  • ✓Conocimiento básico de Java y Python
  • ✓Conocimiento intermedio en prácticas de observabilidad (APM, error budgets, monitoreo, logging, tracing)
  • ✓Entendimiento de SLI, SLO, SLA
  • ✓Conocimiento básico de Terraform e IaC
  • ✓Experiencia con Linux (Debian/Ubuntu/OpenSuse), RabbitMQ, PostgreSQL y AWS (EC2, EKS)

Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.

AWSEC2EKSTerraformKubernetesBashPythonJavaPostgreSQLRabbitMQ

Who should you write to at Onapsis?

Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.

About the job The world’s most critical--and at-risk--business applications have been neglected for far too long. Onapsis eliminates this blind spot by providing cybersecurity solutions dedicated to business-critical applications. Onapsis helps nearly 30% of the Forbes Global 100 understand the threats and risks across their SAP and Oracle landscapes, whether running on-premises, in the cloud, or in a hybrid environment. We are looking for a Site Reliability Engineer II to join our global engineering team. In this role, you will apply software engineering principles to operations, ensuring our cloud platform and distributed security products remain highly available, scalable, and resilient. You will be directly accountable for monitoring, inspecting, troubleshooting, and resolving service and product issues while continuously working with engineering partners to improve telemetry and related operation automations. Rather than performing manual operational maintenance, you will write code, build automation, and design observability frameworks that eliminate toil and prevent system failures. You will work side-by-side with product development teams to embed reliability into the software lifecycle from day one. What you will be doing, your legacy: As a member of our global SRE team, you will take shared full-stack ownership of Onapsis production environments, balancing active operational response with modern software engineering principles. You will develop a deep, end-to-end understanding of our system architecture, technical dependencies, and service behaviors to maximize the performance, scalability, and resilience of our enterprise security platforms. Your time will be split between managing live production environments and driving engineering initiatives that ensure long-term system stability. When troubleshooting live incidents, you will diagnose complex issues across distributed cloud services and stateful infrastructure. Between operational cycles, you will shift into a software engineering mindset—designing, developing, and maintaining custom automation tooling to eliminate repetitive toil, optimize monitoring telemetry, and increase operational efficiency. The ideal Site Reliability Engineer II brings a strong foundational toolkit spanning the software development lifecycle (SDLC), Linux systems administration, core networking protocols, and cloud computing (AWS preferred). You excel at viewing operational friction through a software lens, leveraging coding and automation to transform system vulnerabilities and outages into permanently solved engineering problems. Requirements: • Bachelor's or Master’s degree in Computer Science or related fields or equivalent experience. • 2-4 years experience in SRE, DevOps, or Cloud Engineering role supporting production distributed systems in cloud environments • Beginner knowledge of programming skills such as Java and Python • Intermediate Knowledge of the SRE Observability practices such as Application Performance Monitoring, Error Budget definition and tracking, and ensuring monitoring, logging and tracing are connected. • Understanding of SLI, SLO, SLA methodologies • Beginner knowledge of Infrastructure as Code and Configuration Management using Terraform • Beginner knowledge of Containers and related orchestration platforms such as Kubernetes • Intermediate knowledge of orchestrating technical operations through scripting and automation using Bash or Python. • Experience using AI coding assistants for daily tasks, debugging, and rapid prototyping • Experience managing Linux OS (Debian/Ubuntu/OpenSuse), Queue servers (RabbitMQ), Database instances (PostgreSQL) and AWS cloud computing infrastructure such as EC2 instances or EKS • Intermediate kn

More jobs like this

Remote DevOps Engineer jobs

Looking for something similar?

Leave your email and we'll alert you when matching jobs appear.

Don't apply unprepared

We research who's interviewing you, tailor your CV and rehearse you live — first one free.

InterviewHack.ai

Prepare for the exact interview: who's interviewing you, a tailored CV, and a real coach.

Product

JobsCompanies hiringAll free toolsResume verdict (Jev)Free cover letterInterview questions by roleTechnical assessment simulator"Tell me about yourself" answerFree ATS checkerInterview-English checkSalary checkSalary negotiation scriptFree STAR answerLinkedIn headline + AboutLATAM salary reportFree coursesBlogTailored CVSpoken practicePricingAffiliates — 30%

Remote jobs

ReactPythonFull-StackLATAMArgentinaMexicoSee all →

Prepare

Spoken practiceFrontendBackendAI EngineerBy companySell with your CV

Company

For employersAboutContactPrivacyTerms

© 2026 InterviewHack.ai · Your CV is yours. Never used to train anything. · A product of IA-PTY

Similar open roles

Junior Technical Support Specialist Level 1 - Night Shift

Onapsis · Bucharest, Bucharest, Romania

→

Software Engineer II - Python

Onapsis · Bucharest, Bucharest, Romania

→

Software Developer Engineer in Test III (SDET) - Python & JavaScript

Onapsis · Bucharest, Romania

→

Marketing Associate (Contractor)

Onapsis · United Kingdom

→