Principal Core Infrastructure Engineer
Oracle · Nashville, TN, United States
About The Role
Oracle Cloud Infrastructure builds and operates large-scale cloud services in a distributed, multi-tenant environment. The SPLAT team owns critical platform services that provide secure API routing, service registration, traffic management, private connectivity, authentication, and operational controls for OCI services.
SPLAT sits in the request path for more than 300 OCI control planes and processes hundreds of billions of API requests each month. Our engineering challenges span high-throughput Java services, distributed systems, networking, security, observability, capacity management, deployment automation, and production reliability.
We are looking for a hands-on technical leader who can own substantial systems and cross-service initiatives. You will lead architecture and delivery, write and review production code, resolve complex operational problems, influence partner teams, and develop other engineers. The role requires strong distributed-systems judgment and the ability to make progress when requirements, ownership, or failure domains are not initially clear.
Architecture and Delivery
- Lead significant systems and initiatives from problem definition and design through implementation, rollout, adoption, and production validation.
- Translate scalability, security, reliability, and business requirements into clear technical designs and execution plans.
- Make sound tradeoffs involving availability, consistency, latency, throughput, durability, cost, and operational complexity.
- Design for partial failures, retries, duplicate requests, mixed-version deployments, dependency degradation, and regional disruption.
- Write and review secure, maintainable, well-tested Java code.
- Define service contracts, compatibility requirements, migration plans, validation strategies, and rollback criteria.
Scalability and Operational Excellence
- Establish capacity models, performance objectives, scaling strategies, and load-testing plans for high-throughput services.
- Design effective throttling, load shedding, backpressure, caching, concurrency, and failure-recovery mechanisms.
- Define useful service indicators, objectives, metrics, alarms, dashboards, runbooks, and deployment safeguards.
- Lead complex incident investigations and convert recurring failures or manual procedures into automation and preventive engineering improvements.
- Serve as a technical escalation point for problems that cross application, infrastructure, network, or organizational boundaries.
Technical Leadership
- Provide architectural direction in one or more critical areas such as routing, authentication, private connectivity, runtime performance, observability, or deployment infrastructure.
- Decompose broad initiatives so multiple engineers can own meaningful work while maintaining architectural consistency.
- Mentor engineers through design, code review, delivery, and incident response.
- Raise engineering quality through reusable systems, tools, standards, and operational practices.
- Contribute to hiring and help identify architectural investments, platform gaps, and reliability risks for the team roadmap.
Cross-Team Execution
- Align SPLAT and partner teams on technical decisions, responsibilities, dependencies, and rollout plans.
- Communicate complex designs, tradeoffs, risks, and progress clearly to engineers and leaders.
- Make progress under ambiguity by separating facts, assumptions, reversible decisions, and external dependencies.
- Adjust direction when production evidence or new technical information invalidates earlier assumptions.
- Use modern development and AI-assisted tools responsibly to improve engineering quality and productivity.
Minimum Qualifications
- Bachelor’s degree in Computer Science, Computer Engineering, or a related field, or equivalent practical experience.
- 8+ years of experience designing, building, and operating production backend or platform services.
- Strong development experience in Java or another modern object-oriented language.
- Strong understanding of distributed systems, concurrency, fault tolerance, and production operations.
- Experience leading substantial technical initiatives across multiple engineers or teams.
- Demonstrated ability to diagnose complex production issues and improve service reliability.
- Strong written and verbal communication skills.
Preferred Qualifications
- Experience with high-throughput, low-latency services, HTTP, networking, proxies, or API gateways.
- Experience with authentication, authorization, TLS, certificates, private connectivity, or multi-tenant security.
- Experience with throttling, load shedding, caching, and capacity planning.
- Experience with cloud infrastructure, Kubernetes, infrastructure as code, CI/CD, and deployment automation.
- Experience improving a broader engineering organization through mentoring, shared tooling, or technical standards.
Skills
- Software Design and Development
- Backend Programming Languages
- Distributed Systems
- System Design
Disclaimer
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $114,600 to $234,600 per annum. May be eligible for bonus, equity, and compensation deferral.
Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.
Oracle US offers a comprehensive benefits package which includes the following
- Medical, dental, and vision insurance, including expert medical opinion
- Short term disability and long term disability
- Life insurance and AD&D
- Supplemental life insurance (Employee/Spouse/Child)
- Health care and dependent care Flexible Spending Accounts
- Pre-tax commuter and parking benefits
- 401(k) Savings and Investment Plan with company match
- Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
- 11 paid holidays
- Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
- Paid parental leave
- Adoption assistance
- Employee Stock Purchase Plan
- Financial planning and group legal
- Voluntary benefits including auto, homeowner and pet insurance
The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted. Career Level - IC4
Similar roles you might like
See all →This is an external listing. JobSpring does not represent or verify the employer. Report this listing
