Open roleExternal
Site Reliability Engineer
- london, england, United Kingdom
- Job type not listed
- £90 - £150 Per Hour
Job Description
Senior Site Reliability Engineer (SRE)
We’re partnered with one of the UK’s fastest-growing AI companies, building cloud-native video intelligence products deployed across a rapidly growing fleet of connected devices.
They’re looking for a Senior SRE to own reliability across both their AWS platform and edge infrastructure — combining software engineering, automation and production infrastructure at serious scale.
What You’ll Be Doing
- Build software and tooling to improve the reliability and scalability of production systems
- Own and evolve large-scale AWS & Kubernetes environments
- Build observability, monitoring and telemetry across cloud and edge infrastructure
- Improve deployment, CI/CD and incident response processes
- Automate provisioning and lifecycle management across connected devices
- Identify reliability bottlenecks and engineer them out of the platform
- Improve developer experience through internal tooling and automation
What They’re Looking For
- Strong SRE or software engineering background
- Production coding experience with Python or Go
- Strong Infrastructure-as-Code and automation experience
- Solid understanding of observability, incident management and production reliability
- Edge / IoT experience is a bonus, not essential
- SRE role where you'll actually write software, not just manage infrastructure
- Reliability challenges spanning cloud + thousands of physical devices
- Real-world AI systems with meaningful scale and availability requirements
- Small engineering team with significant ownership and architectural influence
- Strong commercial traction, significant funding and meaningful equity
If you’re an SRE who enjoys building systems rather than babysitting them, this is worth a conversation.
#J-18808-Ljbffr

