Senior Infrastructure Engineer (AI-First)
London, United KingdomPosted Jun 3, 2026
Senior Infrastructure Engineer (AI-First) LocationLondonEmployment TypeFull timeLocation TypeHybridDepartmentEngineeringAbout MUBIMUBI is a global streaming service, production company and film distributor dedicated to elevating great cinema. To make this possible, we create, curate, acquire and champion visionary films, bringing them to audiences all over the world. We have a team of brilliant, dedicated and passionate people to help bring our mission to life. From London to New York, Istanbul to Paris, and Berlin to Mexico - we work together to realize MUBI’s vision.That’s where you come in! Join our global team and help us make great cinema accessible to everyone, everywhere.About the RoleThis isn't a traditional DevOps role, and it's not an ML infrastructure role. You're an infrastructure engineer who builds and ships software — internal tools, services, and agentic automation that solve real infrastructure problems. Think automated incident response, intelligent observability, MCP servers that connect infrastructure systems, and support ticket triaging. You apply agentic patterns to the messy operational world, not train models or manage GPUs.We're looking for someone senior — you architect systems, mentor other engineers, and lead by example. You've built real applications and tools in the infra space before, not just configured existing ones. You don't hesitate to poke around unfamiliar systems, take full ownership of problems, and drive them to completion.Our infrastructure team operates on SRE principles: we build reliable, automated systems and eliminate toil. We also build and maintain the platforms that let the rest of engineering deploy, scale, and monitor their services. You'll own problems end-to-end, from bare-metal CDN servers to Kubernetes clusters to the AI-powered workflows that tie everything together. You use AI tools in your own workflow too — not as a novelty, but as a core part of how you build.Small team, high autonomy, direct impact.Hybrid in London (2-3 days/week). Remote considered for strong candidates outside London.Where you'll have impact:Build & Ship — Agentic InfrastructureDesign and build agentic workflows for infrastructure operations — incident response, observability, support ticket triagingBuild and maintain MCP servers and AI-powered automation that connects infrastructure systemsAssess and own the security surface of agentic tools and workflows you build — understand what you're exposing and how to lock it downBuild internal tools and services that make the engineering team fasterPrototype rapidly using AI-assisted development workflowsCollaborate with product engineers — you're a peer, not a support functionWrite production-quality code, not just glue scriptsPlatform & InfrastructureDesign, run, and evolve our EKS clusters — the primary platform for all MUBI servicesBuild and improve CI/CD pipelines (Jenkins, ArgoCD, Helm)Operate and extend our in-house bare-metal CDN spanning multiple global locationsManage cloud infrastructure on AWS using Terraform and ChefBuild and maintain Kubernetes operators and controllers to automate platform workflowsDesign and operate data pipelines and event-driven architectures using KafkaMaintain and improve observability (ELK, Prometheus, Grafana, Datadog)Define and uphold SLO commitments — own availability and reliability targetsAutomate everything worth automatingTech stackCloud: AWS, Kubernetes (currently EKS)IaC: Terraform, ChefCI/CD: Jenkins, ArgoCD, Helm, KustomizeObservability: ELK Stack, Prometheus with Thanos, Grafana, DatadogStreaming/Messaging: KafkaServers: Nginx (with Lua), bare-metal CDNDatabases: MariaDB, PostgreSQLAI/Agents: Claude, MCP, agentic workflow toolingLanguages: Language-agnostic — strong fundamentals matter more than specific languagesWhat you'll bring5+ years in infrastructure, platform engineering, or DevOpsInterest or experience in building agentic workflows, AI-powered automation, or working with LLM APIsStrong software...