Senior Core Infrastructure Engineer
The Oracle Cloud Infrastructure (OCI) team can provide you the opportunity to build and operate a suite of massive scale, integrated cloud services in a broadly distributed, multi-tenant cloud environment. OCI is committed to providing the best in cloud products that meet the needs of our customers who are tackling some of the world’s biggest challenges.
We offer unique opportunities for smart, hands-on engineers with the expertise and passion to solve difficult problems in distributed highly available services and virtualized infrastructure. At every level, our engineers have a significant technical and business impact designing and building innovative new systems to power our customer’s business critical applications. Oracle's Infrastructure Cloud Object Storage team is Hiring Software Engineers, level commensurate with demonstrated achievements in the past and experience.
The object store team is responsible for a performant, scalable, highly available, and durable object store built from the ground up. We believe this team and mission sits squarely at the center of Oracle's future and is an integral part of Oracle's public cloud efforts.
We have engineering work at every layer of the stack, from REST APIs to distributed systems to file systems, and we are looking for teammates who are interested in being part of a team that innovates from top to bottom. We believe that the only way to succeed is to own every part of the problem, and so we are creating a team that controls its own destiny. You will own development of new components and features, from initial concepts through design, implementation, test, and operation.
Key Responsibilities
Design, develop, and optimize Background Services (garbage collection, Object Lifecycle Management, Object Inventory, Scanner, Publisher, Snapshot Processing, etc.) to meet large-scale performance and scalability requirements.
Drive performance engineering initiatives, including profiling, benchmarking, capacity planning, and performance optimization of distributed background services.
Design and implement scalable processing pipelines capable of handling large-scale metadata, object lifecycle operations, and background workflows under metadata services umbrella.
Develop and enhance end to end tests for Background Services to improve end-to-end validation and production readiness, including scenarios that cannot be validated through traditional tests.
Build and execute performance, stress, endurance, and fault-injection tests to validate system reliability, scalability, and correctness.
Improve operational readiness by developing dashboards, telemetry, alerts, automation, runbooks, and operational tooling for production support.
Investigate and resolve complex production and performance issues, perform root cause analysis, and implement long-term fixes.
Collaborate closely with Object Storage, Dataplane and infrastructure teams to optimize end-to-end system performance and ensure seamless integration across dependent services.
Participate in operational support rotations and contribute to continuous improvements in system reliability, availability, automation, and operational excellence.
Core Responsibilities
Planning & Execution
Independently plan and execute assigned projects while meeting program milestones and delivery timelines.
Balance feature development, performance optimization, operational qualification, and production readiness activities based on project priorities.
Work closely with cross-functional teams, including Dataplane, Oak, Infrastructure, and SRE teams, to deliver scalable and reliable metadata service solutions.
Partner with stakeholders to define performance targets, operational readiness criteria, and feature delivery plans.
Analyze complex distributed systems, identify performance bottlenecks and reliability issues, and implement scalable solutions.
Drive root cause analysis for production and test issues and contribute long-term engineering improvements.
Stay current with distributed systems, cloud infrastructure, and large-scale storage technologies, applying industry best practices to metadata services.
Share technical knowledge, mentor team members, and contribute to engineering best practices.
Continuously improve engineering processes, automation, testing frameworks, observability, and operational workflows to enhance the efficiency and reliability of Background Services.
Career Level - IC3