Principal Core Infrastructure Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+12 more
Job description
Architecture and Delivery
- Lead significant systems and initiatives from problem definition and design through implementation, rollout, adoption, and production validation.
- Translate scalability, security, reliability, and business requirements into clear technical designs and execution plans.
- Make sound tradeoffs involving availability, consistency, latency, throughput, durability, cost, and operational complexity.
- Design for partial failures, retries, duplicate requests, mixed-version deployments, dependency degradation, and regional disruption.
- Write and review secure, maintainable, well-tested Java code.
- Define service contracts, compatibility requirements, migration plans, validation strategies, and rollback criteria.
Scalability and Operational Excellence
- Establish capacity models, performance objectives, scaling strategies, and load-testing plans for high-throughput services.
- Design effective throttling, load shedding, backpressure, caching, concurrency, and failure-recovery mechanisms.
- Define useful service indicators, objectives, metrics, alarms, dashboards, runbooks, and deployment safeguards.
- Lead complex incident investigations and convert recurring failures or manual procedures into automation and preventive engineering improvements.
- Serve as a technical escalation point for problems that cross application, infrastructure, network, or organizational boundaries.
Technical Leadership
- Provide architectural direction in one or more critical areas such as routing, authentication, private connectivity, runtime performance, observability, or deployment infrastructure.
- Decompose broad initiatives so multiple engineers can own meaningful work while maintaining architectural consistency.
- Mentor engineers through design, code review, delivery, and incident response.
- Raise engineering quality through reusable systems, tools, standards, and operational practices.
- Contribute to hiring and help identify architectural investments, platform gaps, and reliability risks for the team roadmap.
Cross-Team Execution
- Align SPLAT and partner teams on technical decisions, responsibilities, dependencies, and rollout plans.
- Communicate complex designs, tradeoffs, risks, and progress clearly to engineers and leaders.
- Make progress under ambiguity by separating facts, assumptions, reversible decisions, and external dependencies.
- Adjust direction when production evidence or new technical information invalidates earlier assumptions.
- Use modern development and AI-assisted tools responsibly to improve engineering quality and productivity.
Requirements
- Bachelor’s degree in Computer Science, Computer Engineering, or a related field, or equivalent practical experience.
- 8+ years of experience designing, building, and operating production backend or platform services.
- Strong development experience in Java or another modern object-oriented language.
- Strong understanding of distributed systems, concurrency, fault tolerance, and production operations.
- Experience leading substantial technical initiatives across multiple engineers or teams.
- Demonstrated ability to diagnose complex production issues and improve service reliability.
- Strong written and verbal communication skills., * Experience with high-throughput, low-latency services, HTTP, networking, proxies, or API gateways.
- Experience with authentication, authorization, TLS, certificates, private connectivity, or multi-tenant security.
- Experience with throttling, load shedding, caching, and capacity planning.
- Experience with cloud infrastructure, Kubernetes, infrastructure as code, CI/CD, and deployment automation.
- Experience improving a broader engineering organization through mentoring, shared tooling, or technical standards.
Skills
- Software Design and Development
- Backend Programming Languages
- Distributed Systems
- System Design
Benefits & conditions
US: Hiring Range in USD from: $114,600 to $234,600 per annum. May be eligible for bonus, equity, and compensation deferral.
Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle’s differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.
Oracle US offers a comprehensive benefits package which includes the following
- Medical, dental, and vision insurance, including expert medical opinion
- Short term disability and long term disability
- Life insurance and AD&D
- Supplemental life insurance (Employee/Spouse/Child)
- Health care and dependent care Flexible Spending Accounts
- Pre-tax commuter and parking benefits
- 401(k) Savings and Investment Plan with company match
- Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
- 11 paid holidays
- Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
- Paid parental leave
- Adoption assistance
- Employee Stock Purchase Plan
- Financial planning and group legal
- Voluntary benefits including auto, homeowner and pet insurance
The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - IC4
About the company
Oracle Cloud Infrastructure builds and operates large-scale cloud services in a distributed, multi-tenant environment. The SPLAT team owns critical platform services that provide secure API routing, service registration, traffic management, private connectivity, authentication, and operational controls for OCI services.
SPLAT sits in the request path for more than 300 OCI control planes and processes hundreds of billions of API requests each month. Our engineering challenges span high-throughput Java services, distributed systems, networking, security, observability, capacity management, deployment automation, and production reliability., Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.
True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
93 Java Interview Questions You Should Prepare For
20 Essential Tools For Backend Development
The Best X (Twitter) Accounts for Developers
7 Important Tips That Every Software Developer Should Know