AI Engineer

TWENTY TECHNOLOGIES INC.
United States
1 day ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Automated Storage and Retrieval Systems Microsoft Azure Cloud Computing Database Theory Software Debugging Graph Database Python (Programming Language) Language Modeling Tensorflow Software Deployment
+15 more
Software Engineering Systems Integration Management of Software Versions Google Cloud Pytorch Large Language Models Prompt Engineering Deep Learning Kubernetes Information Technology Low Latency ONNX (Open Neural Network Exchange) Format Machine Learning Operations TensorRT Docker

Job description

You’ll build and ship language-model-powered systems that strengthen Twenty’s mission-critical cyber capabilities for U.S. national security. You’ll own the end-to-end workflow-from curating specialized datasets and post-training models to deploying reliable inference and retrieval systems in production. You’ll partner closely with product and engineering to translate real operational needs into high-performing AI features, operating across cloud and on-premises environments where speed, correctness, and security matter., * Create, clean, and maintain high-quality training and evaluation datasets for specialized AI use cases.

  • Fine-tune language models (small specialized through medium foundation models) for mission needs.
  • Implement post-training and alignment approaches to improve task performance and reliability.
  • Build retrieval-augmented generation (RAG) systems that ground model outputs in external knowledge.
  • Develop and optimize model serving infrastructure for production deployments.
  • Design evaluation frameworks and test harnesses to measure quality, latency, and regressions.
  • Integrate AI capabilities into applications and workflows using modern orchestration frameworks.
  • Collaborate with cross-functional partners to identify high-leverage use cases and deliver solutions.
  • Produce clear technical documentation for models, datasets, and operational processes.

Requirements

  • You’re motivated by real-world outcomes and want your work to directly impact national security missions.
  • You care about rigor: clean data, measurable evaluation, and repeatable experiments beat demos.
  • You balance research curiosity with product instincts-you ship, observe, iterate, and harden.
  • You’re comfortable working across cloud and on-premises constraints and adapting to the environment.
  • You communicate clearly with engineers and non-ML partners, and you write documentation people use.
  • You think in systems: models, retrieval, infrastructure, and feedback loops all have to work together.
  • You thrive in fast-moving teams with high standards, direct feedback, and high ownership., * You have 4+ years of professional software development experience building and supporting ML/AI-enabled applications.
  • You have strong Python skills and deep learning experience with PyTorch, TensorFlow, or JAX.
  • You have hands-on experience with LLM post-training methods (e.g., continued pre-training, SFT, RLHF, DPO, PPO, GRPO).
  • You have experience curating, cleaning, and preprocessing datasets for training and evaluation.
  • You have working knowledge of relational, graph, and vector database concepts.
  • You have experience designing or using evaluation metrics and testing procedures for LLMs and agents.
  • You have experience integrating LLM/agent systems using frameworks like Pydantic-AI, LangChain/LangGraph, or CrewAI.
  • You have a Bachelor’s degree in Computer Science, Software Engineering, or a related field (or equivalent practical experience).

Nice To Have

  • You have deployed models to production and supported them through real-world usage and incidents.
  • You have experience with distributed training systems and performance debugging at scale.
  • You have implemented quantization or other optimization techniques to improve inference efficiency.
  • You have strong prompt engineering and model alignment instincts for reliability and control.
  • You have experience building MLOps/LLMOps/AgentOps practices (versioning, rollout, monitoring).

Tech Environment (You Might Work With)

  • Deep learning stacks: PyTorch, TensorFlow, JAX
  • LLMOps and serving: vLLM, TensorRT, ONNX
  • Retrieval and storage: pgvector, ChromaDB, Pinecone, Milvus, Weaviate; relational/graph databases
  • Orchestration: Pydantic-AI, LangChain/LangGraph, CrewAI
  • Infra: Docker, Kubernetes; cloud platforms (AWS, GCP, Azure)
  • Experiment and artifact tracking: dataset/prompt/model versioning

Security / Work Environment

Must be eligible to obtain and maintain a U.S. Government security clearance.

Benefits & conditions

What’s on the table:

  • Health. Medical, dental, and vision plan options. Life / AD&D, disability coverage options.
  • Family. Paid parental leave for eligible full-time employees. 12 weeks for birthing parents, 4 for non-birthing parents, 6 weeks for adoptive, foster, or intended parents through surrogacy.
  • Vacation. Paid holidays and flexible PTO. Take what you need.
  • Retirement. 401(k) with pre-tax and Roth options. HSA/FSA options, dependent care FSA.

About the company

America is under sustained cyber attack. Our adversaries infiltrate our networks, steal our IP, and degrade the digital infrastructure that modern life runs on. They’ve learned-correctly-that those attacks rarely produce consequences.

Twenty was founded to change that, by making our adversaries think twice before they attack us. Our vision is American and allied primacy in cyberspace-a future where they cannot contest us, deterrence is assured, and the free world remains secure.

Founded in 2024, Twenty Technologies (www.twenty.io) industrializes offensive cyber operations for the U.S. and its allies. Headquartered in Arlington, Virginia, Twenty has raised $168M from Khosla Ventures, Accel, Caffeinated Capital, Friends & Family Capital, Point72 Ventures, General Catalyst, and In-Q-Tel.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · World Congress 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl · World Congress 2024

Videos

See all

Related articles

See all