AI OS Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+8 more
Job description
We are seeking an experienced AI OS Engineer for a high-impact, 6-month contract initiative. In this role, you will lead the architecture and integration of our next-generation AI Operating System (AI OS)-a core orchestration framework designed to seamlessly manage autonomous agents, multi-LLM routing, context memory systems, tool execution, and local-to-cloud compute pipelines., * Design, build, and deploy agentic workflows, dynamic task schedulers, and execution runtime environments powering internal AI applications.
- Implement robust retrieval systems, long-term state persistence, vector databases (e.g., pgvector, Qdrant, Pinecone), and hybrid-search mechanisms to optimize agent context windows.
- Architect multi-model routing layers (e.g., Anthropic, OpenAI, open-source foundation models) for cost-efficiency, fallback management, and low-latency inference.
- Develop secure sandbox environments for tool execution, code generation, API calls, and agent safety protocols.
- Build evaluation harnesses to track model drift, execution accuracy, hallucination rates, and latency bottlenecks.
- Containerize and deploy AI OS infrastructure on cloud environments (AWS / GCP / Azure) using CI/CD pipelines., * Finalize system architecture, set up local/cloud runtime execution environments, and deploy the core orchestration layer.
- Integrate multi-agent tool execution, long-term memory state persistence, and guardrail protocols.
- Conduct system-wide evaluation harness benchmarking, latency/cost optimization, and handoff documentation for internal engineering teams.
Requirements
- 5+ years of production software engineering experience, with 2+ years focused on building agentic frameworks, multi-agent orchestrations, or LLM infrastructure.
- Advanced proficiency in Python, TypeScript/Node.js, and modern async execution models.
- Hands-on expertise with agent architectures and orchestration frameworks (e.g., LangGraph, AutoGen, CrewAI, LlamaIndex, or custom in-house runtimes).
- Proven track record working with vector databases, embedding systems, and hybrid RAG implementations.
- Direct experience with Docker, Kubernetes, vLLM / Triton inference engines, and cloud platforms (AWS Sagemaker, GCP Vertex AI, or Azure ML).
- Mastery of RESTful/gRPC APIs, message queues (Kafka, RabbitMQ, Redis), and microservice architectures., * Experience with local LLM serving, quantization methods (AWQ, GGUF), and self-hosted foundation models (Llama, Mistral).
- Deep understanding of sandboxed execution environments (e.g., WebAssembly, Docker-in-Docker, E2B) for safe AI agent tool execution.
- Prior contract experience operating in fast-paced, 6-month delivery cycles with clear milestone check-ins.
Benefits & conditions
Medical, dental, and vision coverage 401(k) retirement plan Voluntary benefits and insurance options Referral bonus opportunities Pre-tax commuter benefits
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
What is Agentic Programming and Why Should Developers Care?
From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path
MLOps And AI Driven Development
Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?