Site Reliability Engineer

Office, IndiaFull-timePosted Jul 22, 2026

Our Mission

At Palo Alto Networks®, we’re united by a shared mission—to protect our digital way of life. We thrive at the intersection of innovation and impact, solving real-world problems with cutting-edge technology and bold thinking. Here, everyone has a voice, and every idea counts. If you’re ready to do the most meaningful work of your career alongside people who are just as passionate as you are, you’re in the right place.

Who We Are

In order to be the cybersecurity partner of choice, we must trailblaze the path and shape the future of our industry. This is something our employees work at each day and is defined by our values: Disruption, Collaboration, Execution, Integrity, and Inclusion. We weave AI into the fabric of everything we do and use it to augment the impact every individual can have. If you are passionate about solving real-world problems and ideating beside the best and the brightest, we invite you to join us!

We believe collaboration thrives in person. That’s why most of our teams work from the office full time, with flexibility when it’s needed. This model supports real-time problem-solving, stronger relationships, and the kind of precision that drives great outcomes.

Job Summary

The Team

Engineering - Our engineering team is at the core of our products and connected directly to the mission of preventing cyberattacks. We are constantly innovating — challenging the way we, and the industry, think about cybersecurity. Our engineers don’t shy away from building products to solve problems no one has pursued before. We define the industry instead of waiting for directions. We need individuals who feel comfortable in ambiguity, excited by the prospect of a challenge, and empowered by the unknown risks facing our everyday lives that are only enabled by a secure digital environment.

Job Summary

We are seeking a highly skilled Site Reliability Engineer (SRE) to join our team. As an SRE, you will play a pivotal role in ensuring the reliability, scalability, and performance of our cloud-based infrastructure. You will be responsible for driving and improving the Incident Management processes, with a focus on triaging and ensuring the reliability of CyberArk’s SaaS services and underlying AWS infrastructure. You will collaborate closely with development, operations, and other teams to implement and maintain efficient and resilient systems.

Key Responsibilities

  • Drive incident response processes and troubleshoot complex issues, ensuring timely resolution of outages.
  • Establish monitoring, logging, and alerting best practices using tools like Datadog, Site24x7 etc.
  • Build essential tooling to improve reliability of systems and automated remediation of issues.
  • Be a part of the on-call rotation 365x24x7.
  • Create and maintain documentation for infrastructure, processes, and incident management protocols (SOPs).
  • Utilize Infrastructure as Code (IaC) tools such as Terraform and Ansible to automate provisioning, configuration, and deployment.
  • Continuously optimize system performance, identify bottlenecks, and implement strategies to improve scalability and efficiency.
  • Identify and implement strategies to reduce cloud costs while maintaining performance and reliability.
  • Adhere to security best practices and implement measures to protect infrastructure and data.
  • Implement AI based automations and productivity improvement solutions and share best practices.
  • Work effectively with cross-functional teams to understand business requirements and provide technical guidance.

Qualifications

Required Qualifications

  • 2-3 years of experience as a Site Reliability Engineer/Cloud Engineer.
  • Strong proficiency in AWS cloud services (e.g., EC2, S3, VPC, RDS, EKS, ECS, CloudFormation).
  • Strong scripting skills (Python, PowerShell, CDK, Shell scripting).
  • Understanding of infrastructure as code tools (Terraform, Ansible) and AWX Tower for Ansible automation.
  • Knowledge of containerization (Docker) and orchestration platforms (Kubernetes).
  • Expertise in CI/CD pipelines and automation tools (Jenkins, GitHub).
  • Exposure to monitoring and alerting tools (CloudWatch, Datadog, ELK, Grafana, Site24x7).
  • Experience documenting Standard Operating Procedures (SOP) and Root Cause Analyses (RCAs).
  • Strong communication skills and ability to work in shifts (24x7).
  • Familiarity with AI assisted software development and productivity improvements.

Preferred Qualifications

  • AWS Certification.
  • Security Certification.

Our Commitment

We’re trailblazers that dream big, take risks, and challenge cybersecurity’s status quo. It’s simple: we can’t accomplish our mission without diverse teams innovating, together.

We are committed to providing reasonable accommodations for all qualified individuals with a disability. If you require assistance or accommodation due to a disability or special need, please contact us at  accommodations@paloaltonetworks.com.

Palo Alto Networks is an equal opportunity employer. We celebrate diversity in our workplace, and all qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or other legally protected characteristics.

All your information will be kept confidential according to EEO guidelines.

Is role eligible for Immigration Sponsorship? No. Please note that we will not sponsor applicants for work visas for this position.

Want jobs like this matched to you?

Swoopd scores fresh postings against your résumé so you only see the matches that matter.

Get started free