AI / ML Engineer

Insight Global
Boston, MA, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Software Debugging Monitoring of Systems Python (Programming Language) Machine Learning Enterprise Messaging Systems Natural Language Processing Search Technologies Flask (Web Framework) Large Language Models Multi-Agent Systems
+12 more
Generative AI Backend Fastapi Event Driven Architecture Containerization Kubernetes Data Analytics Restful APIs Artificial Intelligence Markup Language (AIML) GPT Docker Microservices

Job description

The AI ML Engineer will join an existing development team to enhance and expand a complex, dynamic application. The role requires strong communication skills and technical expertise across the stack. You will collaborate with global teams spanning multiple time zones and actively contribute to ongoing feature development. This role will be sitting out of India and require to work from 12:00-8:00pm IST.

Requirements

A highly skilled Senior Python engineer and LLM engineer

  • Executing both planning and hands-on technical work independently.

  • Collaborating effectively with Product Owners and other stakeholders to solve complex problems.

  • Working cross-functionally to contribute to impactful solutions across teams

  • Continuously developing your technical expertise and staying up-to-date with new technologies.

  • Being passionate, intellectually curious, and driven to expand your skills and knowledge.

  • Using a data-driven approach to solve technical challenges and make informed decisions.

  • Applying your systems-level thinking, integrating both data science and engineering principles.

  • Taking full ownership of the features and projects you work on, delivering high-quality solutions on your own., Strong experience with Python, particularly in building REST APIs using frameworks like FastAPI or Flask.

  • Expertise in microservices architecture and deployment in containerized environments (e.g., Docker, Kubernetes).

  • Strong knowledge in AI, machine learning, and natural language processing

  • Strong experience working with key LLM models APIs (e.g. OpenAI, Anthropic) and LLM Frameworks (e.g. LangChain, LlamaIndex)

  • Experience with MCP, Model Context Protocol.

  • Understanding of multi-agent systems and their applications in complex problem-solving scenarios.

  • Experience with RAG concepts and fundamentals (vectorDBs, semantic search, etc.)

  • Expertise in implementing RAG systems that combine knowledge bases with generative AI models.

  • Experience with prompt writing for various use cases

  • Experience with generative solutions released to prod, at scale, beyond POCs

  • Proficiency with server-side events, event-driven architectures, and messaging systems.

  • Strong problem-solving skills and experience debugging and optimizing backend systems.

  • Solid understanding of security best practices for backend systems, including authentication and data protection. - Experience with LLM guardrails

  • Experience with LLM monitoring and observability

  • Experience developing AI/ML technologies within large and business critical applications

Benefits & conditions

$16.00 to $17.00 per year annual salary

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:30 min

Building components of a real-world LLM lifecycle

Maxim Salnikov Maxim Salnikov · LIVE

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · WWC 2024

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:33 min

Connecting frontends via a FastAPI proxy backend layer

Saoussen Chaabnia Saoussen Chaabnia · Europe 2026 Virtual

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · WWC 2025

Videos

See all

Related articles

See all