InterviewHack.ai
Start free
Jobs / Cloudflare

Systems Engineering, Metrics and Alerting

Cloudflare·Hybridmid

In short

  • →Diseñar y operar la plataforma de observabilidad de Cloudflare, centrada en métricas y alertas.
  • →Trabajar en sistemas distribuidos a gran escala que manejan tráfico global en tiempo real.
  • →Destaca la oportunidad de influir directamente en la fiabilidad y eficiencia del Internet global.

Fluent English required for collaboration across global engineering teams.

Apply on company site ↗Share on WhatsApp

In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.

🎧Land the interview? Bring the copilot. Our free extension listens to the live interview and flashes 3-4-word anchors from your resume and prep — glance, connect, talk. Get the extension →

What they ask for

  • ✓Experiencia en diseño y operación de sistemas distribuidos.
  • ✓Conocimiento profundo de métricas, alertas y observabilidad.
  • ✓Experiencia con lenguajes de programación como Go, Python o Rust.
  • ✓Capacidad para resolver problemas de escalabilidad en pipelines críticos.
  • ✓Habilidades de colaboración y mentoría en entornos de ingeniería.
  • ✓Inclinación hacia la curiosidad técnica y la innovación constante.

Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.

GoPythonRustPrometheusOpenTelemetryGrafanaDatadogKafkaSnowflakeKubernetes

Who should you write to at Cloudflare?

Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations: • London • Lisbon About the Department Production Engineering is responsible for the world’s most reliable, observable, performant, and safe network ecosystem. Our customers rely on our products and systems to safely modify, troubleshoot, and release products without external impact. Our external customers rely on us to provide seamless and predictable incident, traffic, policy management, resulting in the fastest and safest network services in the world. We are accountable for the overall performance of internal and external facing services, guiding our product teams to optimal configurations and maximum efficiency. From the moment that a packet enters the Cloudflare ecosystem, we know exactly what its expected purpose and behaviour is and we are capable of determining and exposing anomalous behaviour. The Cloudflare network makes it possible to solve challenges at massive scale and efficiency which would be impossible for almost any other organization. About the Team This role is for the internal Observability Team, responsible for the observability platform and stack to make our engineering teams productive. This includes (but is not limited to) areas like metrics, alerting, error tracking, logging, tracing, and more. In this role, you can expect to: • Design, deliver, and operate software and a platform that progresses Cloudflare's Observability competency • Solve scaling bottlenecks in critical services in our Metrics & Alerting pipeline • Work on highly distributed and scalable systems • Participate in the constant cycle of knowledge sharing and mentoring <li style="font-size: 12p

Looking for something similar?

Leave your email and we'll alert you when matching jobs appear.

Don't apply unprepared

We research who's interviewing you, tailor your CV and rehearse you live — first one free.

InterviewHack.ai

Prepare for the exact interview: who's interviewing you, a tailored CV, and a real coach.

Product

JobsFree ATS checkerInterview-English checkSalary checkLATAM salary reportFree coursesBlogTailored CVSpoken practiceIt's free

Remote jobs

ReactPythonFull-StackLATAMArgentinaMexicoSee all →

Prepare

Spoken practiceFrontendBackendAI EngineerBy companySell with your CV

Company

For employersAboutContactPrivacyTerms

© 2026 InterviewHack.ai · Your CV is yours. Never used to train anything. · A product of IA-PTY

Similar open roles

Senior Named Account Executive, Telcos & SPs

Cloudflare · Hybrid

→

Systems Engineer, MCP Portals

Cloudflare · Hybrid

→

Senior Supply Chain Program Manager

Cloudflare · Hybrid

→

Software Engineer

Cloudflare · Hybrid

→