Skip to main content
Search
Find Jobs For
Location
Search
Close
Search
Sign up for job alerts
Go to Saved Jobs
Staff Site Reliability Engineer
Staff Site Reliability Engineer
Prague, Hlavní město Praha
Saved Role
Save role
Apply now
Function Tech
Team Engineering
Role Type Permanent
Work Location Hybrid
Date Posted 08/20/2025
Explore this location
Take a look at the map to discover what’s nearby.
Veeam is the Data and AI Trust Company, specializing in helping organizations ensure their data and AI are fully understood, secured, and resilient to enable the acceleration of safe AI at scale. As the market leader in both data resilience and data security posture management, Veeam is built for the convergence of identity, data, security, and AI risk. Headquartered in Seattle with offices in more than 30 countries, Veeam protects over 550,000 customers worldwide, who trust Veeam to keep their businesses running. Join us as we go fearlessly forward together, growing, learning, and making a real impact for some of the world’s biggest brands.
About the Role
Veeam is launching a global Site Reliability Engineering (SRE) function to support the rollout and operation of our new SaaS offering: the Veeam Data Cloud. As a Staff Site Reliability Engineer, you will serve as a hands-on technical leader within the SRE team, guiding senior engineers, influencing product development teams, and ensuring the systems we operate are built to be reliable, scalable, and observable from the ground up.
You will drive strategic initiatives, mentor others in the practice of SRE, and help define architectural best practices across our platform. This role is pivotal in aligning teams, enforcing high standards, and scaling SRE principles globally within Veeam.
What You’ll Do
Act as a technical authority in your area, mentoring senior engineers and guiding design choices that improve service reliability and resilience
Lead the definition and enforcement of SLIs, SLOs, and error budgets; drive adherence across engineering teams
Ensure metrics, logs, and traces provide deep, actionable insights across systems
Lead complex incident responses, postmortems, and systemic reliability improvements
Promote and enforce a blameless culture of learning and continuous improvement
Lead initiatives in infrastructure as code, deployment automation, and resilience testing
Influence the development and adoption of chaos engineering practices and release validation frameworks
Collaborate with cross-functional teams - including engineering, product, platform, and security - to align strategies, establish reliability standards, and ensure resilient, production-ready architecture from the outset
Provide architectural guidance and advocate for engineering rigor and consistency
Represent the SRE team in technical leadership forums and product planning discussions
What You’ll Bring
8+ years of experience in a Software Engineering or SRE role, including technical leadership
Demonstrated experience mentoring and guiding senior engineers
Deep expertise in building distributed systems on public cloud (Azure preferred)
Strong skills in programming (e.g., JS, Go, Typescript, Java, or C#)
Hands-on experience with observability tooling (e.g., Prometheus, Grafana, OpenTelemetry)
Mastery of infrastructure automation tools (Terraform, Pulumi) and container orchestration (Kubernetes)
Ability to communicate clearly across geographies and disciplines
Bonus Skills
Familiarity with global compliance standards (ISO, SOC 2, GDPR, FedRAMP, CMMC)
What You’ll Get
25 vacation days, 4 sick days, 21 paid medical leave days, plus 4 extra global VeeaMe Days for self-care and 24 paid volunteer hours annually through Veeam Cares
Premium private medical insurance for employees and...
Want jobs like this matched to you?
Swoopd scores fresh postings against your résumé so you only see the matches that matter.