Senior Software Engineer - Reliability, Infrastructure, and Tooling
This is the employer's own posting, not a copy on a job board.
What we know
Is it still open?
Confirmed still open
Last checked 9h ago — checked against the employer's own applicant tracking system, which is the company answering directly.
We re-read the employer's own applicant tracking system and the posting was still there. That is the company answering directly.
How old is it?
Posted 180d ago
The date the source published, not the day we noticed it (2026-03-18). Last seen at its source 3h ago.
Is it remote?
Marked remote on the employer's board
Their board carries a remote setting on this posting — a field they filled in, not wording we read. The location field names somewhere specific, which is usually where the team or the entity sits.
Who may apply?
North America, EMEA
The description states no restriction of its own. This is the source's own tag.
Pay not stated
Similar roles pay $161.8k–208.7k/yr
Middle 50% of 785 listings that do state pay — Engineering · Senior · North America · USD/year. This employer has published no salary; this is what comparable listings we hold disclose, never converted between currencies or periods. How this is calculated.
Skills named in the ad
Recognised terms only, from a fixed vocabulary — this is what CV matching compares against.
Carried by 1 source
-
ashby employer's own board first seen 39d ago · last seen 3h ago
The listing
About LiveKit
LiveKit is building the infrastructure layer for the agentic era of computing. Our platform gives developers everything they need to build, test, deploy, scale, and observe AI agents in production. Founded in 2021, LiveKit powers voice and agentic AI applications for OpenAI, Salesforce, Spotify, Meta, and tens of thousands of other developers, collectively facilitating billions of calls each year.
About This Role
We’re hiring Senior Software Engineers to join our team focused on infrastructure development and reliability engineering. This is not an “ops” team by any means. We partner with product dev teams to co-design and develop along with them to ensure that LiveKit systems are reliable, maintainable, and secure. We work on internal tooling to provide a smooth experience for our product dev teams to own and run their workloads on top of our infrastructure and to meet our strict reliability requirements.
Like all teams, we own the ops and maintenance for the systems we work on, but where possible we automate away what we can. Our team also facilitates a healthy oncall rotation and incident management practices, but the rotation is shared with product dev team members to ensure the important production perspective that oncall provides isn’t isolated to just our team.
We support the full range of LiveKit products which provide a fascinating landscape of problems to be solved because they are much more demanding than a simple web app. It includes real time media workloads, hosting of customer agent code in secure sandboxes, and advanced networking requirements, all of which keep us on our toes.
You'll Thrive Here If You:
You are tenacious with investigating tricky system level issues.
You like the art of observability including quantifying reliability and visualizing it efficiently.
You understand the delicate balance between moving fast now and moving fast later.
You are able to communicate effectively with partner teams and tactfully handle sometimes contentious topics.
You get satisfaction from clean, DRY, error-resistant configuration even when the underlying systems are complex and diverse.
You look at the world in terms of signals and control systems.
What You'll Do
Ramp on LiveKit's global architecture — CockroachDB, NATS, Nebula, Kubernetes — and map where reliability debt lives
Ship product SRE work directly in the product codebase: load balancing, load shedding, instrumentation, scalability, efficiency
Build and extend common tooling so product teams can self-service reliability without Infra as a bottleneck
Participate in the on-call rotation and help resolve recurring reliability patterns
Bring informed systems opinions that improve how the team makes architectural decisions
Who You Are
Experience building non-trivial applications (high concurrency, complex control loops, etc).
Strong experience with Kubernetes (or equivalent, Borg, etc).
Experience with Linux internals and networking.
Experience making use of observability tools to debug tricky problems.
Experience running large scale globally distributed systems and working with the complex configuration management problems and technical debt that come along with it.
Experience with complex production incident handling.
Experience running open source tooling (e.g. Kafka, Clickhouse, etc).
Nice to Have
Data engineering and analytics.
Global layer 3 networking.
Experience working with systems that handle long lived load like media.
Google SRE or equivalent high-scale background.
Dealing with compliance frameworks (e.g. PCI).
Our Commitment to You
An opportunity to build something truly impactful to the world
Contribute to open source alongside world-class engineers
Competitive salary and equity package
Health, dental, and vision benefits
Flexible vacation policy
LiveKit is an equal opportunity employer and does not discriminate on the basis of any characteristic protected by applicable law. If you require a reasonable accommodation during the application or interview process, please contact recruiting@livekit.io.