LLM Application Engineer

OpenKyber LLC
United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Continuous Integration Graph Database Python (Programming Language) Software Deployment Software Engineering Systems Integration Google Cloud
+5 more
Large Language Models Multi-Agent Systems Software Application Programming Machine Learning Operations Restful APIs

Job description

We are seeking a highly skilled and experienced Senior GenAI Engineer who will be instrumental in designing, developing, and deploying modern generative AI solutions, leveraging expertise in foundation models, agentic AI systems, and full-stack application development. You will work across the entire AI lifecycle, from research and prototyping to production deployment and monitoring, contributing to high-impact transformative projects.

What You’ll Get to Do:

  • Build applications leveraging foundation models and advanced GenAI techniques (Agentic RAG, KG-RAG)
  • Develop autonomous agents using modern agentic frameworks
  • Implement MCP integrations and Python-based full-stack solutions
  • Create scalable REST APIs and deploy AI workloads on major clouds
  • Apply LLMOps/MLOps best practices for CI/CD and monitoring
  • Work across teams, mentor others, and drive innovation in embeddings, knowledge graphs, and ontologies

Requirements

  • Degree in CS, AI/ML, or related field
  • Overall 10+ Years experience and 5+ years of AI/ML-focused software engineering
  • Production experience with LLMs and agentic systems
  • Strong MCP experience (MCP experience means building or integrating AI systems that use the Model Context Protocol to connect LLMs with tools, APIs, and data sources) and expert-level Python
  • Experience with embeddings at scale, knowledge graphs, ontology extraction, and advanced RAG
  • Full-stack skills (Python back end + modern front end)
  • Cloud deployment experience (AWS, Azure, or Google Cloud Platform)
  • Experience with LLMOps tools and strong problem-solving/communication skills
  • Willingness to work in the Dallas office 5 days/week.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:30 min

Building components of a real-world LLM lifecycle

Maxim Salnikov Maxim Salnikov · LIVE

2:59 min

The danger of adopting default REST APIs

Stefan Priebsch · WWC 2021

6:10 min

Unlocking free learning credits via Google Cloud Innovators

Asrar Asrar · WWC 2024

4:20 min

Structuring connected data using graph database architectural fundamentals

Martin O'hanlon · LIVE

3:25 min

Understanding foundational concepts of LLMs, agents, and MCPs

Perf + AI

3:34 min

Approaching language models as scalable synchronous rest APIs

Patrick Koss Patrick Koss · WWC 2025

Videos

See all

Related articles

See all