Ai Engineer

Green Key Resources
New York, NY, United States
3 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Software as a Service Cloud Computing Continuous Integration Python (Programming Language) Azure Machine Learning Software Construction Cloud Platform System Large Language Models Technical Debt Information Technology Low Latency
+2 more
Performance Monitor Machine Learning Operations

Job description

Experteer Overview In this role you will design and deploy advanced AI systems to tackle complex business problems within a hybrid work setup. You will work with cross-functional teams to translate AI opportunities into scalable solutions, while mentoring engineers and guiding architectural decisions. You’ll build and maintain ML pipelines, evaluate performance, and optimize cost and latency for large-scale AI inference. This position focuses on LLMs, document intelligence, and cloud-based platforms to drive innovation in SaaS and FinTech contexts. Compensation / Benefits * Architect and implement AI systems focused on workflow automation and document intelligence * Lead technical design and decision-making processes aligned with business goals * Collaborate with stakeholders to ensure AI projects deliver measurable outcomes * Develop and maintain ML pipelines for reliable deployment * Establish evaluation metrics and frameworks to monitor performance * Enhance team processes and infrastructure to reduce technical debt * Ensure adherence to software engineering best practices (CI/CD, testing, documentation) * Optimize cost and latency for large-scale AI inference using advanced methodologies Tasks * Bachelor of Science degree in Computer Science, Engineering, or a related field * Minimum of 6 years of experience in developing and deploying production-grade AI/ML systems * Proficiency in Python and cloud-native development patterns for AI/ML workloads * Expertise in LLM systems, document intelligence, and scalable ML platform design * Strong knowledge of MLOps practices (training, deployment, monitoring) * Solid understanding of statistics, experimentation, and data quality principles * Proven ability to mentor engineers and lead technical design decisions * Experience optimizing cost and latency for large-scale AI inference systems Key requirements *

Requirements

large-scale to reduce technical debt * Ensure adherence to software engineering best practices (CI/CD, testing, documentation) * Optimize cost and latency for large-scale AI inference using advanced methodologies Tasks * Bachelor of Science degree in Computer Science, Engineering, or a related field * Minimum of 6 years of experience in developing and deploying production-grade AI/ML systems * Proficiency in Python and cloud-native development patterns for AI/ML workloads * Expertise in LLM systems, document intelligence, and scalable ML platform design * Strong knowledge of MLOps practices (training, deployment, monitoring) * Solid understanding of statistics, experimentation, and data quality principles * Proven ability to mentor engineers and lead technical design decisions * Experience optimizing cost and latency for large-scale AI inference systems Key requirements *

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on us.experteer.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

5:48 min

Balancing delivery latency with stream reliability and scale

Phil Cluff · LIVE

3:06 min

The lack of responsibility over technical debt

Stefan Priebsch · WWC 2021

2:27 min

Introduction to WebAssembly in a cloud computing context

Edo Edo · WWC 2024

4:35 min

Learning resources and community engagement for AI engineers

Alfonso Graziano Alfonso Graziano · Coffee With Developers

3:37 min

Accessing API documentation and testing remote driving latency

Alexandru Ciinaru Alexandru Ciinaru +3 · WWC 2025

Videos

See all

Related articles

See all