Senior AI/LLM Engineer (Python) - Remote

KAKE,INC
United States
18 days ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Part-time (≤ 32 hours)
Experience level
Expert
Working hours
Regular working hours
Languages
English
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Data Infrastructure Data Stores Distributed Systems Python (Programming Language) PostgreSQL Machine Learning Open Source Technology Redis Tensorflow Software Deployment
+13 more
Software Engineering Data Streaming Enterprise Software Applications Pytorch Large Language Models Prompt Engineering Backend Fastapi Containerization Scikit Learn Integration Tests Apache Kafka Docker

Job description

Senior AI/LLM Engineers with strong Python experience, skilled in designing, building, and productionizing LLM-powered applications and AI systems at scale. It’s a great fit for people who enjoy working at the intersection of software engineering and applied AI, and who take ownership from prototyping through production deployment.

What you’ll build and own

  • Design, build, and deploy LLM-powered features and applications using Python.
  • Develop and maintain backend services and APIs (e.g., FastAPI) that expose AI/LLM capabilities to other systems.
  • Build and optimize RAG pipelines, including embeddings, vector search, and retrieval strategies.
  • Design, test, and iterate on prompts, agents, and orchestration flows using frameworks such as LangChain, LlamaIndex, or similar.
  • Integrate with LLM providers and APIs (e.g., OpenAI, Anthropic, open-source models) and manage tradeoffs around cost, latency, and quality.
  • Evaluate model outputs systematically, building tooling and metrics to test accuracy, safety, and regression across iterations.
  • Work with containerized environments and data infrastructure (e.g., PostgreSQL, Redis, vector databases) to support reliable AI systems in production.
  • Collaborate with cross-functional stakeholders to translate ambiguous product needs into technically sound AI solutions.

Requirements

  • Strong proficiency in Python and experience with FastAPI or similar backend frameworks.
  • Experience working with LLM APIs (e.g., OpenAI, Anthropic, or similar) and frameworks such as LangChain or LlamaIndex.
  • Experience with RAG architectures, embeddings, and vector databases (e.g., Pinecone, Weaviate, pgvector, or similar).
  • Hands-on experience with Docker and containerized development environments.
  • Experience working with PostgreSQL, Redis, or similar data stores.
  • Strong experience writing functional and integration tests, including evaluation frameworks for AI/LLM output quality.
  • Excellent written and verbal communication skills in English.
  • Ability to work independently in a remote, fast-paced environment.

Nice-to-Have

  • Experience fine-tuning or evaluating open-source LLMs.
  • Familiarity with prompt engineering best practices and agentic workflows.
  • Experience with distributed systems, streaming (e.g., Kafka), or large-scale applications.
  • Background in machine learning fundamentals (e.g., scikit-learn, PyTorch, or TensorFlow).
  • Comfortable working flexible hours to overlap with distributed teams across different time zones.

Benefits & conditions

Additional

  • US Timezone Overlap: 5h-6h daily PST

Why Join Kake?

The icing on the Kake:

  • Competitive Pay in USD: Work globally, get paid globally.
  • Fully Remote: Simply put, we trust you.
  • Better Me Fund: We invest in your personal growth and passions.
  • Compassion is Badass: Join a community that invests in social good.

About the company

Kake is a remote-first company and a people-first global community of senior engineers. Kake engineers are behind some of the world’s most innovative products (Brands you’ve heard!), from Fortune 500 to fast-growing companies. We believe it’s not where your table is, but what you bring to the table that matters. Our community spans 45,000+ engineers across 55+ countries; join a culture where great people stay, grow, and thrive (and love eating kake!).

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

2:37 min

Optimizing technical profiles for AI sourcing and recruitment

Mina Golesorkhi Mina Golesorkhi · World Congress 2026 Europe

3:42 min

Comparing in-memory and Redis storage for cache scalability

Simone Sanfratello · World Congress 2022

Videos

See all

Related articles

See all