Artificial Intelligence Senior Associate

Fasttek Global
United States
1 day ago
Apply on www.fasttek.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
1 year minimum
Working hours
Regular working hours

Tech stack

HTML Artificial Intelligence Amazon Web Services Business Analytics Applications Data Analysis Microsoft Azure BigQuery Spreadsheets Cloud Computing Cloud Database Concurrent Computing Continuous Integration
+31 more
Data Integration Database Queries Distributed Systems Github Python (Programming Language) Machine Learning Natural Language Processing Role-Based Access Control Salesforce.Com Service-Oriented Architecture Software Engineering SQL Databases SQL Server Agent Management of Software Versions Reinforcement Learning Google Cloud Flask (Web Framework) Large Language Models Multi-Agent Systems Deep Learning Generative AI Build Server Backend Fastapi Containerization Kubernetes Information Technology Operational Systems Api Design Software Version Control Docker

Job description

  • Understand business requirements and develop AI algorithms, models and programs to solve complex problems, generate recommendations, extract patterns, make predictions, interpret sensor data (images, sound), orchestrate automation and enable self-service capabilities
  • Perform large-scale experimentation and develop data driven applications that translate data into actionable intelligence
  • Drive innovative applications of Artificial Intelligence tools and techniques such as deep learning, generative AI, natural language processing, image processing, cognitive automation, intelligent process automation, reinforcement learning, virtual assistants and specialized programming
  • Research and optimize AI technologies to enhance efficiency and accuracy of data analysis and create more efficient automation, * Architect and deploy the production multi-agent orchestration layer (interpreter/orchestrator, NL-to-SQL agent, visualization agent, RCA/RAG agent, report composition agent, notification agent), using modern agent frameworks with state management and checkpointing rather than ad-hoc loops.
  • Design and productionize RAG pipelines (chunking, embeddings, hybrid retrieval, reranking) grounded in approved schemas, engineering documentation, and historical issue records.
  • Own BigQuery integration and enforce safe, least-privilege, validated execution of LLM-generated SQL. Build CI/CD, containerization, and infrastructure-as-code for deploying agent services on GCP (Cloud Run/GKE, Vertex AI).
  • Implement evaluation pipelines and observability/tracing for every agent (golden datasets, LLM-as-judge scoring, regression alerts) so quality is measurable, not assumed.
  • Implement guardrails, prompt-injection defenses, and human-in-the-loop approval checkpoints to ensure correctness and safety before any output triggers downstream action. Design cost/latency optimization strategies, including tiered model routing (cheap filter models vs. high-capability deep-dive models) and caching.
  • Integrate validated outputs with operational systems (Salesforce ticketing, driver/site-manager notifications) and report export pipelines (PDF/HTML/spreadsheet).
  • Collaborate with data scientists to productionize prototypes (anomaly detection, diagnostic agents) into scalable, monitored services.
  • Establish versioning, testing, and safe rollout practices (canary/shadow deployments) for evolving agent logic.

Requirements

  • Google Cloud Platform, * Bachelor’s or Master’s degree in Computer Science, Software Engineering, or related field (or equivalent practical experience).
  • 3+ years building production software systems, including 1-2+ years on ML/AI or LLM-based applications. Proven experience designing and deploying multi-agent or multi-service architectures in production - not just notebooks or demos.
  • As one 2026 hiring analysis puts it, the job is closer to distributed systems engineering with a probabilistic component than it is to ML research or prompt tweaking .
  • Strong Python proficiency, including async/concurrent programming, and experience with backend frameworks (FastAPI, Flask).
  • Hands-on experience with agent orchestration frameworks - LangGraph, CrewAI, LlamaIndex, or equivalent - for building stateful, multi-step, tool-using agent workflows.
  • Practical experience building RAG pipelines: vector databases (pgvector, Pinecone, Weaviate, or Qdrant), embeddings, chunking strategies, and retrieval evaluation. Cloud deployment experience, ideally Google Cloud Platform (BigQuery, Cloud Run/GKE, Vertex AI, Pub/Sub) or equivalent AWS/Azure services.
  • Strong SQL skills and experience with cloud data warehouses. Containerization and CI/CD experience (Docker, Kubernetes, GitHub Actions/Cloud Build).
  • Experience building evaluation and observability pipelines for LLM/agent systems - offline eval sets, LLM-as-judge scoring, and tracing tools (LangSmith, Langfuse, OpenTelemetry, or equivalent) to track task success, latency, and cost.
  • Understanding of LLM safety practices: guardrails, output validation, prompt-injection defense, and safe execution of AI-generated code/SQL (sandboxing, least privilege). Solid software engineering fundamentals: API design, testing, version control, security best practices.

Experience Preferred:

  • Experience with cost optimization and model routing - designing tiered pipelines that route between low-cost and high-capability models based on task complexity, and modeling per-conversation or per-task cost at scale.
  • Experience deploying agentic systems with human-in-the-loop or multi-checkpoint validation workflows for high-reliability/high-stakes use cases.
  • Experience with automotive, EV charging, IoT, or connected-vehicle telemetry data.
  • Familiarity with Model Context Protocol (MCP) or similar standards for tool/data integration across agents.
  • Prior experience in a startup or 0-to-1 product environment, comfortable with ambiguity and fast-evolving requirements.

Education Required:

  • Bachelor’s Degree

Education Preferred:

  • Master’s Degree

Benefits & conditions

At FastTek Global, Our Purpose is Our People and Our Planet. We come to work each day and are reminded we are helping people find their success stories. Also, Doing the right thing is our mantra. We act responsibly, give back to the communities we serve and have a little fun along the way. We have been doing this with pride, dedication and plain, old-fashioned hard work for 24 years! FastTek Global is financially strong, privately held company that is 100% consultant and client focused. We’ve differentiated ourselves by being fast, flexible, creative and honest. Throw out everything you’ve heard, seen, or felt about every other IT Consulting company. We do unique things and we do them for Fortune 10, Fortune 500, and technology start-up companies. Our benefits are second to none and thanks to our flexible benefit options you can choose the benefits you need or want, options include:

  • Medical and Dental (FastTek pays majority of the medical program)
  • Vision
  • Personal Time Off (PTO) Program
  • Long Term Disability (100% paid)
  • Life Insurance (100% paid)
  • 401(k) with immediate vesting and 3% (of salary) dollar-for-dollar match

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.fasttek.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

2:21 min

Projecting external HTML content using default and named slots

Rowdy Rabouw Rowdy Rabouw · World Congress 2022

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

3:51 min

Shifting skill requirements for entry-level software engineers

Taroon Mandhana Taroon Mandhana +1 · World Congress 2026 Europe

6:12 min

Streaming HTML content natively using declarative processing instructions

Chris Heilmann +2 · LIVE

Videos

See all

Related articles

See all