Embedded AI Engineer

Bright Vision Technologies
Gilbert, AZ, United States
about 1 month ago

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$176,800.0 - $187,200.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Systems Engineering C++ (Programming Language) Computer Engineering Data Centers Embedded Software Firmware Python (Programming Language) Machine Learning Tensorflow Runbook Software Engineering
+7 more
Management of Software Versions Graphics Processing Unit (GPU) Large Language Models Information Technology ONNX (Open Neural Network Exchange) Format Free and Open-Source Software Machine Learning Operations

Job description

This role is part of Bright Vision Technologies’ in-house Statement of Work (SOW) engagement. The client, end customer, and employer for this position is Bright Vision Technologies - there is no third-party client, vendor, or implementation partner involved. We do not engage in C2C, 1099, or third-party arrangements for this role BUT STRICTLY NO C2C/1099/3RD PARTY COMPANIES. ALL OUR ROLES ARE W2 AND NO 3RD PARTY BROKERING PLEASE. Candidates must be willing to work directly as a full-time W2 employee of Bright Vision Technologies and contribute to our in-house SOW deliverables. No new H1B sponsorship is available for this role. However, candidates who are currently on a valid H1B visa and require a transfer are welcome to apply. We will support H1B transfers for qualified candidates. For every role, a technical coding assessment is mandatory. Please apply only if you are confident in your technical abilities and hands-on experience., We are looking for an Embedded AI Engineer to design, optimize, and deploy machine learning models that run efficiently on resource-constrained edge devices, including mobile platforms, embedded systems, and specialized accelerators. The role requires deep expertise in model compression, quantization, and hardware-aware optimization, along with strong systems engineering skills to ship reliable AI capabilities outside the data center. The ideal candidate has shipped edge AI in production environments where compute, memory, energy, and connectivity constraints fundamentally shape the engineering trade-offs. Key Responsibilities

  • Design and implement edge AI solutions optimized for diverse hardware including mobile SoCs, NPUs, and embedded accelerators.
  • Apply quantization, pruning, distillation, and architectural optimization to fit models within edge constraints.
  • Tune model performance for latency, energy efficiency, and memory footprint on target hardware.
  • Build cross-platform inference runtimes leveraging frameworks such as TensorFlow Lite, ONNX Runtime, and Core ML.
  • Optimize models for specific accelerator backends including DSPs, NPUs, and mobile GPUs.
  • Implement on-device model update, versioning, and rollback workflows that allow safe staged rollouts to large device populations and rapid recovery if a model release behaves unexpectedly in the field.
  • Design hybrid edge-cloud architectures that gracefully degrade based on connectivity and device capability.
  • Build telemetry pipelines that respect privacy while enabling continuous improvement.
  • Collaborate with hardware, firmware, and product teams to align AI capabilities with device constraints.
  • Implement secure execution paths, model protection, and integrity verification on edge devices.
  • Develop benchmarking suites that characterize accuracy, latency, and energy trade-offs across devices.
  • Drive responsible AI considerations including on-device privacy and bias evaluation.
  • Maintain comprehensive, current technical documentation - including architecture diagrams, design decisions, configuration references, runbooks, and operational procedures - so that the system remains supportable, auditable, and easy to onboard new engineers onto over time.
  • Stay current with edge AI hardware and software developments, regularly review release notes and community discussions, and translate noteworthy advances into concrete recommendations and adoption proposals for the team., Job Title: Embedded Software Engineer Location: Chandler AZ, 85286 (100% Onsite) Duration: 12 Months Salary Range: $85.00 - $90.00/Hour on W2 (Without Benefits). Applicants mu…
  • 8 days ago

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field.
  • Six or more years of experience in ML engineering, with significant work on edge or mobile AI.
  • Strong proficiency in Python and C++.
  • Hands-on experience with model compression, quantization, and pruning techniques.
  • Experience with at least one major edge inference framework.
  • Solid understanding of mobile and embedded hardware architectures.
  • Experience deploying ML models to production on mobile or embedded platforms.
  • Strong performance engineering and profiling skills.
  • Familiarity with on-device privacy and security considerations.
  • Strong communication and cross-functional collaboration skills., * Experience with custom NPU or DSP toolchains.
  • Familiarity with federated learning or on-device personalization.
  • Exposure to safety-critical or industrial edge deployments.
  • Open-source contributions to edge AI frameworks.
  • Experience optimizing LLMs for on-device inference.

About the company

Bright Vision Technologies is a forward-thinking software development company dedicated to building innovative solutions that help businesses automate and optimize their operations. We leverage cutting-edge technologies to create scalable, secure, and user-friendly applications., Bright Vision Technologies is a forward-thinking software development company dedicated to building innovative solutions that help businesses automate and optimize their operations…

  • 19 hours ago

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:39 min

Fundamentals of tensors and the TensorFlow library

Håkan Silfvernagel · LIVE

2:19 min

Orchestrating over-the-air firmware updates for vehicle modules

Denis Grahovac · WWC 2021

2:50 min

Introduction and the value of runbooks

Hila Fish · WWC 2023

1:56 min

Optimizing AI processing capabilities for edge microcontrollers

Stephan Gillich Stephan Gillich +3 · WWC 2024

3:55 min

Evaluating central server APIs against edge deployment models

Hauke Brammer · WWC 2021

2:20 min

Utilizing custom firmware for variable torque manipulation

Daniel Meilak Daniel Meilak +1 · WWC Europe 2026

Videos

See all

Related articles

See all