InterviewHack.ai
Empezar gratis
Vacantes / Fact Finder

Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

Fact Finder · Berlinsenior

En corto

  • ▸Ingeniero SRE senior que construye y gestiona una nube privada con Kubernetes y Harvester en infraestructura propia.
  • ▸Trabaja en equipo con desarrollo para automatizar, observar y hacer sostenible el sistema; impacto directo en 2.000 tiendas online.
  • ▸No es SRE en la nube gestionada: se construye una plataforma on-prem desde cero con capacidades híbridas.

Fluent English required

Postularme en la empresa ↗Compartir por WhatsApp

En ~1 minuto te damos: quién te entrevista, las preguntas probables con respuestas desde tu CV, y tu CV adaptado a esta vacante. Gratis, sin tarjeta.

🎧¿Llegás a la entrevista? Llevá el copiloto. Nuestra extensión escucha la entrevista en vivo y te muestra anclas de 3-4 palabras desde tu CV y tu preparación — mirás, conectás, hablás. Gratis. Ver la extensión →

¿Qué piden?

  • ✓Experiencia probada con Kubernetes en producción, instalado y mantenido en servidores propios (kubeadm, RKE2, k3s)
  • ✓Práctica real de SRE: SLOs, error budgets, gestión de incidentes y on-call
  • ✓Experiencia con GitOps (Argo CD o Flux) para automatización de infraestructura
  • ✓Habilidades sólidas en observabilidad: métricas, logs, traces y alertas confiables
  • ✓Instinto fuerte para la automatización y eliminación de toil
  • ✓Enfoque colaborativo y servicial hacia equipos de desarrollo

¿No cumplís todo? Es lo normal — tu dossier gratis te dice qué gaps tenés y cómo cubrirlos en la entrevista.

KubernetesHarvesterKubeVirtArgo CDFluxPrometheusGrafanaLonghornCephvSphere/ESXi

¿A quién escribirle en Fact Finder?

Tu dossier gratis identifica a las personas que te entrevistarían — con su background, qué valoran y cómo escribirles para destacar antes de aplicar.

Introduction At a glance Location & work model : Berlin, hybrid Tech stack: Kubernetes on our own servers, Harvester ( KubeVirt ), Argo CD/Flux, Prometheus/Grafana, Longhorn/Ceph Team: A growing SRE team – you report to our CTPO for now and to the Team Lead SRE we're hiring next; two system administrators in Pforzheim run the physical hardware Process: Intro call · take-home task (~2h) · 90-min tech interview with our developers · leadership conversation · meet the team Languages: Fluent English required; German is a plus, not a must Why this role is special Most SRE jobs today mean clicking around a managed cloud console. This one doesn't. We run our own hardware in Frankfurt and are building a modern private cloud platform on Kubernetes and Harvester – on-prem by default, with elastic burst into the public cloud and the option to go cloud-only later. You won't inherit a finished SRE practice: you'll help define it, side by side with our Berlin development teams – and you won't do it alone, a Team Lead SRE hire is coming next. SRE here is an enabling discipline: you build what our developers need to ship reliably, while two system administrators in Pforzheim run the physical hardware. And the impact is direct – our product discovery technology powers more than 2,000 European online shops (Intersport, SPAR, Douglas and more), handling billions of shopper queries a year. When product discovery is slow or down, our customers lose revenue in real time. , or automating away a piece of toil. By day 90 you've shipped visible improvements and know where you want to take the platform next. Your mission Define and own SLOs, SLIs and error budgets; drive data-informed reliability decisions Lead incident response end-to-end: fast detection, clear communication, blameless postmortems – and reduce whole classes of incidents structurally, not case by case Eliminate toil through automation and GitOps; evolve our observability (metrics, logs, traces, alerting, runbooks) across two different stacks Help build our custom Kubernetes operator (CRDs) that makes stateful search clusters declarative, self-healing and safely upgradable – and roll out the auto-scaling (HPA/VPA, KEDA, cluster auto scaler) today's architecture makes hard Plan capacity, performance and cost across on-premises and cloud – including the large-catalogue and peak-season loads our merchants care about – and use AI tools wherever they measurably speed up diagnosis and operations g.

Más empleos como este

Empleos remotos de DevOps Engineer

No apliques sin prepararte

Investigamos quién te entrevista, adaptamos tu CV y te ensayamos en vivo — gratis la primera.

InterviewHack.ai

Preparate para la entrevista exacta: quién te entrevista, tu CV a medida y coach real.

Producto

VacantesRevisar CV (ATS) gratis¿Cómo suena tu inglés?¿Te pagan bien?Reporte de sueldos LATAMCursos gratisBlogCV a medidaPráctica habladaEs gratis

Empleos remotos

ReactPythonFull-StackLATAMArgentinaMéxicoVer todas →

Preparate

Práctica habladaFrontendBackendAI EngineerPor empresaVendete con tu CV

Empresa

Buscás talentoAcerca deContactoPrivacidadTérminos

© 2026 InterviewHack.ai · Tu CV es tuyo. Nunca se usa para entrenar nada. · Un producto de IA-PTY

Vacantes similares activas

(Senior) Customer Success Manager (m/w/d) – B2B SaaS / eCommerce – Berlin, München, Pforzheim (Hybrid)

Fact Finder · Berlin

→

Senior Site Reliability Engineer (all genders)

Fact Finder · Berlin

→

Team Lead - Site Reliability Engineering (all genders)

Fact Finder · Berlin

→

Interim project - Senior DevOps Consultant (f/m/d)

Climate · Berlin HQ

→