PB✓
PBridge

Full-time jobsCanada

Software Engineer, Observability

lyft · Toronto, Canada · Full-time

About this role

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Our Infrastructure team is passionate about building software to solve problems at massive scale. We do this often, and when we believe our solution is worth sharing with the community, such as Envoy Proxy , we open source our ideas for the benefit of others.

As an Observability team member, you are responsible for the operation and maintenance of our logging and metrics infrastructure. You ensure all teams at Lyft are aware of the operational health of their products by monitoring system availability and take a holistic view of our platform performance. You build software and platforms to automate infrastructure platform operations and management. By measuring and monitoring our operations you find opportunities to improve our systems in order to push our platform forward. You provide our partners with the support they need to help them build robust large scale distributed systems.

We count on the reliability of our infrastructure to empower Lyft teams to provide our customers rich experiences that are highly available with rock solid performance to ensure our transportation platform continues to connect people and places. As we grow our team, we are seeking experienced Infrastructure Engineer to ensure that as our Infrastructure continues to scale, our platform continues to provide an essential and dependable service that transports millions of people every day. Specifically we are searching for someone who brings fresh perspectives, enjoys collaborating with cross-functional teams in order to continually improve our products and services for our customers.

Responsibilities:

• Maintain, improve, and develop tooling and systems that enhance the reliability, scalability, and efficiency of our platform.

• Assist engineering teams in defining service-level objectives (SLOs) and provide the necessary tooling to monitor and balance feature development speed and reliability.

• Maintain and analyze metrics from operating systems, control planes, and applications to assist in fault detection and performance enhancement.

• Collaborate with cross-functional engineering teams to enhance Lyft's observability and meet developers' needs, ensuring alignment with design and production readiness reviews, platform management, and capacity planning.

• Keep and maintain our documentation at a world-class level by documenting infrastructure operations processes and insights.

• Identifying repeatable actions, and automating repetitive tasks.

• Participate in our team's on-call rotations, respond to incidents, and support other teams to mitigate customer-impacting events.

Experience:

• 3+ years of experience working on teams responsible for software development, automation, and systems engineering.

• Bachelor's Degree or equivalent experience in Computer Science or a relevant discipline.

Tired of applying one by one?

Our Career Success Team finds roles in Canada that fit you, tailors your CV to each, and submits the applications — tracked end to end. You just show up to interviews.

We apply, you interview →