Senior Site Reliability Engineer

OPOWER
Reston, VA$81.1k–$187kPosted Jul 22, 2026
Skip to main content. sitemap Profile Sign Out View More Jobs Senior Site Reliability Engineer Reston, VA, United StatesAustin, TX, United States Job Identification 340647 Job Category Product and Research Posting Date 07/21/2026, 02:38 PM Role Individual Contributor Job Type Regular Employee Does this position require a security clearance? Yes Years 3 to 5+ years Additional Info Visa / work permit sponsorship is not available for this position Applicants are required to read, write, and speak the following languages English Job Description Capacity Ingestion and Management: -       Participates and listens in on discussions for the design and architecture of infrastructure and/or service according to terms for reliability and functionality. -       Assists team members responding to infrastructure demands and capacity increases to support current and future workloads. -       Supports collaborations with the software development team to contribute to the development of reliable and scalable infrastructures based on detailed deployment requirements. -       Participates in identifying opportunities for prototyping and provides support for prototyping initiatives (e.g., testing new applications or infrastructures, assisting in onboarding). Incident and Service Lifecycle Management: -       Assists in data collection, triage, and redirection to maintain and optimize operations and infrastructure reliability. -       Monitors services and maintains up-to-date knowledge of their performance. -       Supports incident response and/or maintenance tasks (e.g., software installs, version upgrades, and security updates, backup and recovery) under supervision. -       Assists in providing health and performance reporting and takes appropriate actions based on trends in data. -       May perform provisioning according to established procedures to support infrastructure, applications, and services. -       May perform decommissioning (e.g., shutting down servers, removing data from databases) according to established procedures to remove objects that are no longer needed. Automation: -       Assists in identifying opportunities for automation and assessing potential benefits. -       Supports the development of automation or scripts to provide solutions, gather metrics, monitor, analyze, mitigate, or remediate issues/defects within infrastructures. -       Follows detailed instructions to conduct testing to ensure automation performs tasks correctly and produces expected results, with supervision. Technical Communication and Guidance: -       Communicates basic information about the scale, capacity, security, and performance attributes of services and technology within immediate team. -       Assists in identifying and communicating basic infrastructure, feature, and tool changes within immediate team. Troubleshooting and Resolution: -       Provides operational support for technology, escalating routine, low-impact incidents and other issues arising within Oracle services. -       Participates in on-call shifts to address issues. -       Assists with resolving technical issues, performing investigations, and debugging products in order to reach SLOs (service level objectives), with supervision. -       Follows detailed instructions to document incidents and perform root cause analyses according to standard reporting methods. -       Participates in post-mortem procedures to prevent incident reoccurrence. Innovation and Improvement: -       Assists in experimenting with new tools and technologies to improve...

Want jobs like this matched to you?

Swoopd scores fresh postings against your résumé so you only see the matches that matter.

Get started free