> Markdown version of [/jobs/ext/2299075-principal-core-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/2299075-principal-core-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Principal Core Infrastructure Engineer - **Company:** Oracle - **Location:** Nashville, TN, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, User Authentication, Cloud Computing, Code Review, Computer Engineering, Continuous Integration, Distributed Systems, Fault Tolerance, Load Testing, Routing, Object-Oriented Software Development, Oracle (Applications), Cloud Services, Runbook, Software Engineering, Strategies of Testing, Concurrency, Caching, Backend, Kubernetes, Infrastructure Automation Frameworks, Information Technology, Low Latency, Deployment Automation, Api Gateway - **Published:** August 29, 2026 - **Apply:** https://eeho.fa.us2.oraclecloud.com/hcmUI/CandidateExperience/en/sites/CX_1/requisitions/preview/342842 ## About the Role * Bachelor's degree in Computer Science, Computer Engineering, or a related field, or equivalent practical experience. * 8+ years of experience designing, building, and operating production backend or platform services. * Strong development experience in Java or another modern object-oriented language. * Strong understanding of distributed systems, concurrency, fault tolerance, and production operations. * Experience leading substantial technical initiatives across multiple engineers or teams. * Demonstrated ability to diagnose complex production issues and improve service reliability. * Strong written and verbal communication skills., * Experience with high-throughput, low-latency services, HTTP, networking, proxies, or API gateways. * Experience with authentication, authorization, TLS, certificates, private connectivity, or multi-tenant security. * Experience with throttling, load shedding, caching, and capacity planning. * Experience with cloud infrastructure, Kubernetes, infrastructure as code, CI/CD, and deployment automation. * Experience improving a broader engineering organization through mentoring, shared tooling, or technical standards., * Software Design and Development * Backend Programming Languages * Distributed Systems * System Design Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives. ## Description * Lead significant systems and initiatives from problem definition and design through implementation, rollout, adoption, and production validation. * Translate scalability, security, reliability, and business requirements into clear technical designs and execution plans. * Make sound tradeoffs involving availability, consistency, latency, throughput, durability, cost, and operational complexity. * Design for partial failures, retries, duplicate requests, mixed-version deployments, dependency degradation, and regional disruption. * Write and review secure, maintainable, well-tested Java code. * Define service contracts, compatibility requirements, migration plans, validation strategies, and rollback criteria. Scalability and Operational Excellence * Establish capacity models, performance objectives, scaling strategies, and load-testing plans for high-throughput services. * Design effective throttling, load shedding, backpressure, caching, concurrency, and failure-recovery mechanisms. * Define useful service indicators, objectives, metrics, alarms, dashboards, runbooks, and deployment safeguards. * Lead complex incident investigations and convert recurring failures or manual procedures into automation and preventive engineering improvements. * Serve as a technical escalation point for problems that cross application, infrastructure, network, or organizational boundaries. Technical Leadership * Provide architectural direction in one or more critical areas such as routing, authentication, private connectivity, runtime performance, observability, or deployment infrastructure. * Decompose broad initiatives so multiple engineers can own meaningful work while maintaining architectural consistency. * Mentor engineers through design, code review, delivery, and incident response. * Raise engineering quality through reusable systems, tools, standards, and operational practices. * Contribute to hiring and help identify architectural investments, platform gaps, and reliability risks for the team roadmap. Cross-Team Execution * Align SPLAT and partner teams on technical decisions, responsibilities, dependencies, and rollout plans. * Communicate complex designs, tradeoffs, risks, and progress clearly to engineers and leaders. * Make progress under ambiguity by separating facts, assumptions, reversible decisions, and external dependencies. * Adjust direction when production evidence or new technical information invalidates earlier assumptions. * Use modern development and AI-assisted tools responsibly to improve engineering quality and productivity. ## Related Videos - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [HTTP headers that make your website go faster](https://www.wearedevelopers.com/videos/1676-http-headers-that-make-your-website-go-faster) - [Technical Documentation - How Can I Write Them Better and Why Should I Care?](https://www.wearedevelopers.com/videos/681-technical-documentation-how-can-i-write-them-better-and-why-should-i-care) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Bridging AI and Nomad: a Go-based MCP Server for Cluster Control](https://www.wearedevelopers.com/videos/2063-bridging-ai-and-nomad-a-go-based-mcp-server-for-cluster-control) - [Event based cache invalidation in GraphQL](https://www.wearedevelopers.com/videos/433-event-based-cache-invalidation-in-graphql) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [93 Java Interview Questions You Should Prepare For](https://www.wearedevelopers.com/magazine/14-93-java-interview-questions-you-should-prepare-for) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)