Lead Software Engineer- Inference platform engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+22 more
Job description
Be an integral part of an agile team that’s constantly pushing the envelope to enhance, build, and deliver top-notch technology products., As a Senior Lead Software Engineer at JPMorgan Chase within the Corporate Sector, Infrastructure Platforms team, you are an integral part of an agile team that works to enhance, build, and deliver trusted market-leading technology products in a secure, stable, and scalable way. Drive significant business impact through your capabilities and contributions, and apply deep technical expertise and problem-solving methodologies to tackle a diverse array of challenges that span multiple technologies and applications., * Provide technical guidance aligned to business goals, collaborating with engineers, contractors, and vendors.
- Develop secure, high-quality production code; review, debug, and improve others’ code.
- Influence product design, application functionality, and technical operations through informed decisions.
- Champion firmwide SDLC frameworks, tools, and best practices.
- Promote a culture of diversity, Opportunity, inclusion, and respect.
- Architect and deploy secure, scalable cloud platforms optimized for AI/ML workloads.
- Partner with AI teams to translate compute needs into infrastructure requirements.
- Monitor, manage, and optimize cloud resources for performance and cost efficiency.
- Build CI/CD, automation, and infrastructure-as-code to streamline ML deployment and operations.
- Drives adoption and governance of approved AI-assisted engineering practices across teams to improve code quality, delivery speed, and operational outcomes (e.g., AI-assisted code review/refactoring, test acceleration, release readiness, incident/root-cause analysis), while establishing measurable validation standards (secure coding, peer review, automated testing) and promoting reuse of proven patterns and automation within the SDLC/TLM toolchain.
- Applies knowledge of tools within the Software Development Life Cycle toolchain, including approved AI-assisted development and automation capabilities, to improve the value realized by automation at scale.
Requirements
- Formal training or certification on software engineering concepts and 5+ years applied experience
- Hands-on experience in system design, application development, testing, and operational stability.
- Strong experience with Kubernetes and containerization (Docker), including cluster operations and production troubleshooting.
- Proficiency in at least one programming language, such as Python, Go, Java, or C#.
- Ability to independently tackle design and functionality problems with minimal oversight.
- Strong knowledge of cloud computing delivery models (IaaS, PaaS, SaaS) and deployment models (Public, Private, Hybrid Cloud).
- Foundational understanding of machine learning concepts, including transformer architecture, ML training, and inference.
- Experience with Infrastructure as Code.
- Deep understanding of cloud component architecture: Microservices, IaaS, Storage, Security, and routing/switching technologies.
- Demonstrated experience leading effective use of enterprise-authorized AI-assisted software development tools within the work environment (e.g., for coding, code review, test acceleration, troubleshooting) with the ability to set team expectations for validating AI outputs for correctness, performance, and security
- Strong understanding of responsible AI use in engineering workflows, including data sensitivity considerations, secure handling of inputs/outputs, and adherence to resiliency and security expectations; experience coaching senior engineers/leads on compliant usage patterns and controls.
Preferred qualifications, capabilities, and skills
- Foundational understanding of NVIDIA GPU infrastructure software (e.g., DCGM, BCM, Dynamo Inference).
- Proficiency with observability tools like Prometheus and Grafana.
- Experience in ML Ops and related tooling, including MLflow.
- Experience building and scaling a high-performance LLM inference platform - leveraging vLLM and GPU/CPU serving optimization - to deliver low-latency, high-throughput, production-grade model serving.
Benefits & conditions
We offer a competitive total rewards package including base salary determined based on the role, experience, skill set and location. Those in eligible roles may receive commission-based pay and/or discretionary incentive compensation, paid in the form of cash and/or forfeitable equity, awarded in recognition of individual achievements and contributions. We also offer a range of benefits and programs to meet employee needs, based on eligibility. These benefits include comprehensive health care coverage, on-site health and wellness centers, a retirement savings plan, backup childcare, tuition reimbursement, mental health support, financial coaching and more. Additional details about total compensation and benefits will be provided during the hiring process.
About the company
hackajob is collaborating with J.P. Morgan to connect them with exceptional professionals for this role., JPMorganChase, one of the oldest financial institutions, offers innovative financial solutions to millions of consumers, small businesses and many of the world’s most prominent corporate, institutional and government clients under the J.P. Morgan and Chase brands. Our history spans over 200 years and today we are a leader in investment banking, consumer and small business banking, commercial banking, financial transaction processing and asset management., Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we’re setting our businesses, clients, customers and employees up for success.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud
How to Become an AI Engineer
Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence
Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production