InterviewHack.ai
Start free
Jobs / Lyft

Software Engineer, Observability

Lyft · Toronto, Canadamid

In short

  • ▸Ingeniero de infraestructura enfocado en observabilidad con foco en métricas, logging y monitoreo a gran escala.
  • ▸Día a día: desarrollas herramientas, mantienes sistemas de monitoreo y colaboras con equipos para mejorar la confiabilidad y rendimiento.
  • ▸Destacado: trabajas con tecnologías de código abierto como Envoy Proxy y contribuyes a proyectos open source del ecosistema.

Bilingual (English required for collaboration with global teams)

Apply on company site ↗Share on WhatsApp

In ~1 minute you get: who interviews you, the likely questions answered from your CV, and your CV tailored to this job. Free, no card.

What they ask for

  • ✓3+ años en desarrollo de software, automatización y sistemas.
  • ✓Grado en Ciencias de la Computación o experiencia equivalente.
  • ✓Experiencia en Go o Python para código en producción.
  • ✓Operación en AWS con servicios gestionados.
  • ✓Construcción y mantenimiento de infraestructura de observabilidad.
  • ✓Familiaridad con Kubernetes y entornos multi-cluster en producción.

Don't tick every box? That's normal — your free dossier shows your gaps and how to cover them in the interview.

GoPythonAWSPrometheusGrafanaLokiKubernetesEnvoy ProxyOpen-source tracingAlerting frameworks

Who should you write to at Lyft?

Your free dossier identifies the people who'd interview you — their background, what they value, and how to reach out so you stand out before applying.

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Our Infrastructure team is passionate about building software to solve problems at massive scale. We do this often, and when we believe our solution is worth sharing with the community, such as Envoy Proxy , we open source our ideas for the benefit of others. As an Observability team member, you are responsible for the operation and maintenance of our logging and metrics infrastructure. You ensure all teams at Lyft are aware of the operational health of their products by monitoring system availability and take a holistic view of our platform performance. You build software and platforms to automate infrastructure platform operations and management. By measuring and monitoring our operations you find opportunities to improve our systems in order to push our platform forward. You provide our partners with the support they need to help them build robust large scale distributed systems. We count on the reliability of our infrastructure to empower Lyft teams to provide our customers rich experiences that are highly available with rock solid performance to ensure our transportation platform continues to connect people and places. As we grow our team, we are seeking experienced Infrastructure Engineer to ensure that as our Infrastructure continues to scale, our platform continues to provide an essential and dependable service that transports millions of people every day. Specifically we are searching for someone who brings fresh perspectives, enjoys collaborating with cross-functional teams in order to continually improve our products and services for our customers. Responsibilities: • Maintain, improve, and develop tooling and systems that enhance the reliability, scalability, and efficiency of our platform. • Assist engineering teams in defining service-level objectives (SLOs) and provide the necessary tooling to monitor and balance feature development speed and reliability. • Maintain and analyze metrics from operating systems, control planes, and applications to assist in fault detection and performance enhancement. • Collaborate with cross-functional engineering teams to enhance Lyft's observability and meet developers' needs, ensuring alignment with design and production readiness reviews, platform management, and capacity planning. • Keep and maintain our documentation at a world-class level by documenting infrastructure operations processes and insights. • Identifying repeatable actions, and automating repetitive tasks. • Participate in our team's on-call rotations, respond to incidents, and support other teams to mitigate customer-impacting events. Experience: • 3+ years of experience working on teams responsible for software development, automation, and systems engineering. • Bachelor's Degree or equivalent experience in Computer Science or a relevant discipline. • Proficiency in creating production-ready code in one or more high-level languages, such as Go or Python. • Experience operating infrastructure in public cloud environments, such as AWS, including familiarity with Managed Services. • Experience in building and maintaining observability infrastructure to support robust monitoring and analysis. • Familiarity with Kubernetes and managing multi-cluster environments in production settings. • Proven track record with modern Observability stack, including proficiency in Prometheus, Grafana, Loki, and other open-source tracing and alerting frameworks. Benefits: • Extended health and dental coverage options, along with life insurance and disability benefits • M

Don't apply unprepared

We research who's interviewing you, tailor your CV and rehearse you live — first one free.

InterviewHack.ai

Prepare for the exact interview: who's interviewing you, a tailored CV, and a real coach.

Product

JobsFree ATS checkerSalary checkLATAM salary reportFree coursesBlogTailored CVReal coachPricing

Remote jobs

ReactPythonFull-StackLATAMMexicoSee all →

Prepare

Practice with a coachFrontendBackendAI EngineerBy companySell with your CV

Company

For employersAboutContactPrivacyTerms

© 2026 InterviewHack.ai · Your CV is yours. Never used to train anything.

Similar open roles

Counsel, Product & Commercial (Rideshare)

Lyft · San Francisco, CA

→

Software Engineer, Fulfillment Core Services

Lyft · Seattle, WA

→

Business Operations Specialist

Lyft · Nashville, TN

→

Fleet Operations Associate (Day Shift)

Lyft · Nashville, TN

→