Computer Vision Engineer

Panoptyc, Inc.
United States
18 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Compensation
$184,000.0
Working hours
Regular working hours

Tech stack

Clean Code Principles Artificial Intelligence Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Computer Vision Image Analysis Data Files Machine Learning Language Modeling Object Detection Open Source Technology
+17 more
Sensor Fusion Software Deployment Software Engineering Jupyter Notebook Pytorch Deep Learning Kubernetes ONNX (Open Neural Network Exchange) Format AWS Fargate Free and Open-Source Software Machine Learning Operations TensorRT Software Version Control Data Pipelines Mixed Reality Docker Data Generation

Job description

Be an Early Applicant Remote Hiring Remotely in USA Senior level Remote Hiring Remotely in USA Senior level Design, train, and deploy production computer vision and vision-language models for retail product recognition. Optimize models for edge devices, build data pipelines and annotation workflows, fine-tune open-source VLMs, develop VLA pipelines, mentor engineers, and drive CV infrastructure and MLOps best practices. The summary above was generated by AI Computer Vision Engineer

Panoptyc is seeking an exceptional Senior Computer Vision Engineer to architect and train cutting-edge models for retail object recognition and drive our edge deployment strategy., You’ll be joining our awesome team of hardware, full-stack and CV engineers developing our next generation computer vision capabilities, building and optimizing models that power real-world retail applications. This role demands someone who can move seamlessly from training custom YOLO architectures to deploying optimized models on edge devices - and from fine-tuning open-source VLMs to building VLA pipelines that reason about and act on what they see. What You’ll Do

  • Model Development: Design, train, and iterate on custom object detection models specifically tuned for retail environments, inventory tracking, and product recognition
  • VLM & VLA Integration: Fine-tune and deploy open-source vision-language models (LLaVA, Qwen-VL, InternVL, PaliGemma, etc.) for product understanding, zero-shot classification, and scene reasoning; build vision-language-action pipelines that translate visual understanding into downstream decisions
  • Edge Optimization: Take state-of-the-art models and make them blazingly fast for edge deployment through quantization, pruning, and architectural optimization
  • Dataset Engineering: Build robust data pipelines and annotation workflows to continuously improve model performance on diverse retail scenarios
  • Research & Innovation: Stay ahead of the curve on CV and VLM research, prototype new architectures, and determine what’s actually production-ready versus academic noise
  • Technical Leadership: Mentor engineers, establish best practices for model development, and drive technical decisions around our CV infrastructure, Computer Vision * Machine Learning * Software Design, train, and deploy custom object-detection models for retail; fine-tune and integrate vision-language models; optimize models for edge devices; build dataset and annotation pipelines; prototype research ideas; and provide technical leadership and mentoring for CV infrastructure and production ML systems. Top Skills: BedrockDockerEc2EcsFargateInternvlKubernetesLitgptLlama.CppLlavaMlflowNvidia JetsonOnnxOnnx RuntimePaligemmaPyTorchQwen-VlS3SagemakerSglangTensorrtTransformersUnslothVllmWeights & BiasesYoloYolo-E NVIDIA

Senior Deep Learning and Computer Vision Engineer - Autonomous Vehicles

3 Days Ago In-Office or Remote 184K-357K Annually Senior level 184K-357K Annually Senior level Artificial Intelligence * Computer Vision * Hardware * Robotics * Metaverse Design, implement, and productionize state-of-the-art deep learning and computer vision models for autonomous vehicles. Define and collect training datasets, build training pipelines and real-time inference runtimes, collaborate with researchers to turn experiments into robust, deployable systems, and optimize architectures for multi-sensor fusion and efficient deployment. Top Skills: C++Computer VisionDeep LearningLidarNvidia GpusPythonPyTorchSelf-Supervised LearningTensorFlowTensorrtUnsupervised Learning OpenSpace

Sr. Computer Vision Engineer

10 Days Ago Remote United States Senior level Senior level Artificial Intelligence * Computer Vision * PropTech Lead design and implementation of computer vision systems for mapping, localization, image analysis, and 3D reconstruction on construction site data. Develop and optimize SLAM/SfM/VIO, deep learning perception models, and multi-modal approaches. Prototype, test, oversee datasets and collaborate with platform engineers to productionize scalable, robust solutions. Top Skills: LlmsOpenglPhotogrammetryPyqtPyTorchSfmSlamTransformersVioVllms

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Requirements

  • 4+ years of hands-on computer vision engineering, with a proven track record of shipping models to production
  • Deep expertise with YOLO and YOLO-E architectures - you’ve trained them, tuned them, and know their quirks intimately
  • Hands-on experience with open-source VLMs (LLaVA, Qwen-VL, InternVL, PaliGemma, or similar) - fine-tuning, evaluation, and production deployment
  • Familiarity with VLA frameworks and applying vision-language-action models to real-world perception and decision tasks
  • Edge deployment mastery - experience with TensorRT, ONNX Runtime, or similar frameworks for optimizing models for constrained devices, including quantized VLMs
  • Strong software engineering fundamentals - clean code, version control, CI/CD for ML, and the ability to build maintainable systems
  • Production ML experience - you understand the difference between a Jupyter notebook and a production-grade ML system

Preferred Qualifications

  • Experience developing solutions deployed to the NVIDIA Jetson family of products
  • Experience with retail, inventory management, or similar product-focused CV applications
  • Background with PyTorch and modern training frameworks (Transformers, LitGPT, Unsloth, etc.)
  • Experience running VLM inference efficiently (vLLM, llama.cpp, SGLang, or similar)
  • Familiarity with synthetic data generation and data augmentation techniques
  • Knowledge of model versioning and experiment tracking (MLflow, Weights & Biases, etc.)
  • Publications or open-source contributions in computer vision or multimodal AI
  • Experience with AWS: EC2, ECS, Fargate, S3, Bedrock, SageMaker, etc.

Technical Stack

While we value expertise over specific tools, you’ll likely work with: PyTorch, YOLO variants, open-source VLMs, TensorRT, ONNX, vLLM, Docker, Kubernetes, and various MLOps tooling.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.ashbyhq.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Open-source community and machine learning frameworks

Gian Marco Iodice Gian Marco Iodice · WWC 2025

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 · WWC 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl · WWC 2024

4:41 min

Replacing PyTorch with ONNX runtime for AWS Lambda deployments

Marek Suppa · LIVE

Videos

See all

Related articles

See all