Software Engineer
Senior Site Reliability Engineer
Role
We are hiring a dedicated SRE to take real ownership of operational excellence across Cloud Infrastructure at Synthesia. You will own domains end to end, build automation and tooling, and raise the bar on how reliably we run our systems. This is a role that combines operational and engineering work to grow with the team, not simply a ticket-queue job.
What you’ll own
- Incident management and operational excellence — on-call quality, response, post-mortems, and driving down incident count, time-to-detect, and time-to-resolve.
- Automation & reliability engineering — automate low-frequency, high-consequence operations (e.g., certificate renewals), and prioritize automation based on risk and blast radius.
- Platform domain — long-term ownership of a domain such as Temporal, observability, or Kubernetes operations, partnering with engineers building in it.
- Vendor & third-party management — own key external relationships and integrations (e.g., LLM API providers, third-party services) with structure, automation, and resilience.
- FinOps — own cloud and platform cost visibility and efficiency, mapping usage to billing.
What success looks like (first 12 months)
- Critical operational knowledge documented and shared — no single point of failure for vendor, cost, or incident response.
- Measurable reliability gains: fewer SEV1–SEV3 incidents per quarter, faster customer-impact resolution, and more incidents caught by monitoring before customers notice.
- High-risk manual processes automated and self-documenting.
What we’re looking for
- Strong production operations experience on AWS and Kubernetes; comfortable with MongoDB and scripting/automation in Python.
- Operations-and-reliability mindset — pride in systems that run quietly; engineer the problem away rather than absorb it manually.
- Sound judgment on incidents and risk; calm and clear under pressure.
- Influences through relationships and evidence, not escalation; comfortable owning a domain and partnering across teams.
- Bonus: vendor/cost management exposure, Temporal, observability tooling.
How we think about this role
We don’t letterbox engineers. You’ll have a clear primary mission (operational excellence) but real domain ownership and the mandate to build — not a fixed lane. The role may evolve as the team grows.
Apply Now
Job Details
- Location
- Remote (US East Coast preferred, for timezone coverage)
- Experience Level
- Senior
- Salary Range
- Not specified
Company Overview
S
Synthesia
Synthesia is the #1 AI Video Platform for Business. Our AI Video Generator enables anyone to create professional videos without mics, cameras, actors, or stu...
View all jobs at Synthesia