Skip to main content.
sitemap
Profile
Sign Out
View More Jobs
Senior Site Reliability Engineer
Reston, VA, United StatesAustin, TX, United States
Job Identification
340647
Job Category
Product and Research
Posting Date
07/21/2026, 02:38 PM
Role
Individual Contributor
Job Type
Regular Employee
Does this position require a security clearance?
Yes
Years
3 to 5+ years
Additional Info
Visa / work permit sponsorship is not available for this position
Applicants are required to read, write, and speak the following languages
English
Job Description
Capacity Ingestion and Management:
- Participates and listens in on discussions for the design and architecture of infrastructure and/or service according to terms for reliability and functionality.
- Assists team members responding to infrastructure demands and capacity increases to support current and future workloads.
- Supports collaborations with the software development team to contribute to the development of reliable and scalable infrastructures based on detailed deployment requirements.
- Participates in identifying opportunities for prototyping and provides support for prototyping initiatives (e.g., testing new applications or infrastructures, assisting in onboarding).
Incident and Service Lifecycle Management:
- Assists in data collection, triage, and redirection to maintain and optimize operations and infrastructure reliability.
- Monitors services and maintains up-to-date knowledge of their performance.
- Supports incident response and/or maintenance tasks (e.g., software installs, version upgrades, and security updates, backup and recovery) under supervision.
- Assists in providing health and performance reporting and takes appropriate actions based on trends in data.
- May perform provisioning according to established procedures to support infrastructure, applications, and services.
- May perform decommissioning (e.g., shutting down servers, removing data from databases) according to established procedures to remove objects that are no longer needed.
Automation:
- Assists in identifying opportunities for automation and assessing potential benefits.
- Supports the development of automation or scripts to provide solutions, gather metrics, monitor, analyze, mitigate, or remediate issues/defects within infrastructures.
- Follows detailed instructions to conduct testing to ensure automation performs tasks correctly and produces expected results, with supervision.
Technical Communication and Guidance:
- Communicates basic information about the scale, capacity, security, and performance attributes of services and technology within immediate team.
- Assists in identifying and communicating basic infrastructure, feature, and tool changes within immediate team.
Troubleshooting and Resolution:
- Provides operational support for technology, escalating routine, low-impact incidents and other issues arising within Oracle services.
- Participates in on-call shifts to address issues.
- Assists with resolving technical issues, performing investigations, and debugging products in order to reach SLOs (service level objectives), with supervision.
- Follows detailed instructions to document incidents and perform root cause analyses according to standard reporting methods.
- Participates in post-mortem procedures to prevent incident reoccurrence.
Innovation and Improvement:
- Assists in experimenting with new tools and technologies to improve...
Want jobs like this matched to you?
Swoopd scores fresh postings against your résumé so you only see the matches that matter.