Sr. LLM / Python Test Engineer

BCI-IT, Inc.
Weehawken, NJ, United States
2 days ago
Apply on www.disabledperson.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Agile Methodology Artificial Intelligence Amazon Web Services Software Applications Automation of Tests Microsoft Azure Software Quality Continuous Integration Database Testing DevOps Monitoring of Systems
+30 more
Apache JMeter Python (Programming Language) Load Testing Microsoft UI Automation NoSQL Object-Oriented Software Development Systems Development Life Cycle Standard Sql Software Safety Selenium SQL Databases Test Data Datadog Cloud Platform System Large Language Models Prompt Engineering Model Validation Generative AI Backend Pandas Pytest Playwright Virtual Agents Cloudwatch Restful APIs Splunk New Relic (SaaS) Api Management SDET Microservices

Job description

BCI is seeking a self-driven Test Engineer with experience testing AI-powered applications, Generative AI solutions, and Agentic AI systems. This role is responsible for designing and executing quality engineering activities, developing automation solutions, validating APIs and integrations, supporting CI/CD and release readiness, and promoting a quality-first engineering mindset across traditional and AI-driven platforms. Role is hybrid. Must already be local to the NY/NJ area. No relocation. Must be a US Citizen, GC, Perm Resident, EAD - OPT/STEM. No H1 or sponsorship. No C2C or subcontractors., * Design and execute functional, integration, system, regression, and automation testing activities.

  • Apply Shift-Left testing practices by participating in requirement reviews and defining test scenarios early in the SDLC.
  • Build, maintain, and enhance UI automation frameworks using Python, PyTest, and Playwright.
  • Perform manual and automated API testing using industry-standard tools and frameworks.
  • Collaborate with developers, product owners, business analysts, and DevOps teams to ensure product quality.
  • Work with SQL and NoSQL databases for backend validation and data testing.
  • Support CI/CD pipelines and ensure automated tests are integrated into deployment processes.
  • Create test plans, test cases, test data strategies, and quality metrics dashboards.
  • Manage defect lifecycle activities and provide clear reporting of quality risks.
  • Develop AI test datasets, benchmarks, evaluation frameworks, and automated validation approaches.
  • Experience with performance and load testing using JMeter or similar tools.
  • Knowledge of AI observability, model monitoring, and evaluation platforms.
  • Experience with monitoring tools such as New Relic, Datadog, Splunk, or CloudWatch.

Requirements

  • 8+ years of software quality engineering, test automation, or SDET experience in Agile development environments.

  • Strong proficiency in Python, Pytest, Pandas, and object-oriented programming principles.

  • 6+ years of experience with UI automation frameworks such as Playwright or Selenium.

  • 6+ years of experience testing RESTful APIs and microservices.
  • Hands-on experience with CI/CD pipelines and DevOps practices.
  • Experience working with SQL and NoSQL databases.
  • Experience with AWS or other cloud-based platforms.
  • Understanding prompt engineering, prompt validation, hallucination detection, bias testing, guardrail validation, and AI output evaluation.
  • Knowledge of LLMs, Retrieval-Augmented Generation (RAG), AI agents, MCP, and AI-assisted workflows.
  • Ability to develop QA strategies, test frameworks, and automation solutions for AI-driven applications.
  • Analytical, problem-solving mindset with continuous learning and quality-first engineering habits., * Experience testing Generative AI and Agentic AI applications.
  • Ability to validate AI outputs for accuracy, reliability, and business relevance.
  • Knowledge of AI agent workflows, prompt testing, hallucination detection, and model evaluation.
  • Understanding AI safety, guardrails, security, compliance, and responsible AI practices.
  • Experience with LangChain, LangGraph, MCP, OpenAI, Anthropic, Azure OpenAI, Kiro or Amazon Bedrock.
  • Familiarity with AI governance, risk management, and AI quality metrics.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.disabledperson.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:06 min

Elevating the QA engineering role for complex challenges

Ondřej Gróf Ondřej Gróf · World Congress 2026 Europe

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes · LIVE

Videos

See all

Related articles

See all