> Markdown version of [/jobs/ext/1992728-senior-decision-intelligence-engineer-nba](https://www.wearedevelopers.com/jobs/ext/1992728-senior-decision-intelligence-engineer-nba). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Decision Intelligence Engineer (NBA) - **Company:** Humana Inc. - **Location:** Richmond, VA, United States (Remote available) - **Experience:** Expert - **Salary:** $106,900.0 - $147,000.0 - **Contract:** Permanent contract - **Skills:** Cable Modem, Computer Programming, Data Integration, Discrete Event Simulation, Human-Computer Interaction, Internet Hosting Service, Integer Programming, Internet Services, Linear Programming, Recommender Systems, Tensorflow, Server Administration, Software Engineering, Reinforcement Learning, Pytorch, Multi-Agent Systems, Backend, Vue.js, Data Lakes, Pyspark, Machine Learning Operations, Front End Software Development, Markov, Dynamic Programming, Data Pipelines, Workday, Databricks - **Published:** August 8, 2026 - **Apply:** https://www.juju.com/job/00000000gm6y2u ## About the Role + 5+ years (post undergraduate level) of software engineering or quantitative research experience building and operating large-scale production systems, with emphasis on data-intensive platforms, recommendation systems, optimization engines, or simulation frameworks serving millions of users. + 2+ years (post graduate level) of software engineering or quantitative research experience building and operating large-scale production systems, with emphasis on data-intensive platforms, recommendation systems, optimization engines, or simulation frameworks serving millions of users. + 2+ years of hands-on experience implementing reinforcement learning, operations research methods, or simulation-driven decision systems in production. Relevant backgrounds include policy gradient and value-based RL (PPO, A3C, DQN, CQL), stochastic dynamic programming, discrete-event simulation, or large-scale combinatorial or constrained optimization. + Deep familiarity with Markov Decision Processes, Bellman-equation-based value estimation, reward or objective shaping, exploration-exploitation tradeoffs, and constraint formulation in real-world decision systems. + Demonstrated ability to diagnose failure modes in learned or optimized policies: instability, poor credit assignment across long horizons, and distributional shift across large populations. + Proficiency in Python 3.x; experience with PyTorch or TensorFlow for policy network or learned model implementation. + Experience with Ray RLlib or equivalent distributed computation frameworks for large-scale training or optimization. + Experience with Databricks, PySpark, and Delta Lake for large-scale ML or data pipelines processing tens of millions of records. + Experience with MLflow for experiment tracking, model registry, and artifact management. + Experience with shipping systems that operate reliably under production load, not just research or prototype work. Preferred Qualifications + Experience with multi-agent RL frameworks (PettingZoo or equivalent) or multi-agent simulation and coordination methods. + Familiarity with operations research methods applicable to constrained sequential decisioning: linear programming, mixed-integer programming, Lagrangian relaxation, or constraint programming. + Experience operating decision or optimization systems in regulated domains (healthcare, finance, or insurance) where member safety, auditability, and explainability are requirements. + Experience building simulation environments using Gymnasium, SimPy, AnyLogic, or equivalent frameworks for policy evaluation and backtesting. + Familiarity with event-driven feedback loops and how disposition signals feed retraining or re-optimization pipelines. + OpenTelemetry instrumentation experience for ML or optimization pipeline observability. **Work Style:** Remote/Hybrid - Preferably Boston, MA. Occasional travel to Humana's Tech Hubs for training or meetings may be required. **Work Hours** : Typical business hours are Monday-Friday, 8 hours/day, 5 days/week-- some flexibility might be possible, depending on business needs. Very minimal travel might be required for training, meetings, and/or conferences Work at Home Requirements WAH requirements: Must have the ability to provide a high-speed DSL or cable modem for a home office. Associates or contractors who live and work from home in the state of California will be provided with payment for their internet expense. A minimum standard speed for optimal performance of 25x10 (25mpbs download x 10mpbs upload) is required. ## Description Become a part of our caring community and help us put health first. We are looking for a skilled Decision Intelligence Engineer to design, train, and improve the reinforcement learning policy at the heart of Humana's Next Best Action platform. This role is hands-on and research-oriented. You will design and evaluate decision-making algorithms, and instrument training pipelines. Additionally, you will collaborate with data and platform engineers. Furthermore, you will ensure the system operates correctly within the constraints of clinical eligibility rules and program-specific objectives. The Senior Decision Intelligence Engineer Is involved in all stages of software development, including front-end development, back-end development, database integrations, network and hosting management, user interface, user experience, and back-end server management. Begins to influence department's strategy. Makes decisions on moderately complex to complex issues regarding technical approach for project components, andwork is performed without direction. Exercises considerable latitude in determining objectives and approaches to assignments., As part of our hiring process, we will be using on-demand technology provided by Hire Vue, a third-party vendor. This technology provides our team of recruiters and hiring managers with an enhanced method for decision-making through on-demand candidate assessments. If you are selected to move forward from your application prescreen, you will receive correspondence inviting you to participate in an on-demand assessment with pre-determined questions. You should anticipate the assessment to take approximately 10-15 minutes. Your on-demand assessment will be reviewed, and you will subsequently be informed if you will be moving forward to next round of interviews. SSN Task via Workday Should you be extended a formal employment offer you will receive a request to enter your SSN into our Workday system to scan for duplicate profiles. Work at Home Requirements: To ensure Home or Hybrid Home/Office employees' ability to work effectively, the self-provided internet service of Home or Hybrid Home/Office employees must meet the following criteria: At minimum, a download speed of 25 Mbps and an upload speed of 10 Mbps is required; wireless, wired cable or DSL connection is suggested. In certain roles, the minimum recommended internet speed required by Humana may not be sufficient for business needs. Humana reserves the right to require associates to upgrade their internet service if necessary. Work from a dedicated space lacking ongoing interruptions to protect member PHI / HIPAA information. Travel: While this is a remote position, occasional travel to Humana's offices for training or meetings may be required. Scheduled Weekly Hours 40 ## Related Videos - [Beyond the 9–5: Designing Work Around Humans](https://www.wearedevelopers.com/videos/1321-beyond-the-9-5-designing-work-around-humans) - [Destigmatizing the Workplace: Building Real Inclusion](https://www.wearedevelopers.com/videos/1492-destigmatizing-the-workplace-building-real-inclusion) - [Lessons learned from building a thriving Vue.js SaaS application](https://www.wearedevelopers.com/videos/1666-lessons-learned-from-building-a-thriving-vue-js-saas-application) - [Rules, Heuristics, or LLMs? Lessons from Solving the Same Problem Twice](https://www.wearedevelopers.com/videos/100112-rules-heuristics-or-llms-lessons-from-solving-the-same-problem-twice) - [What if your HR software adapted to you, not the other way around?](https://www.wearedevelopers.com/videos/100259-what-if-your-hr-software-adapted-to-you-not-the-other-way-around) - [The Future of Employee Wellbeing: Benefits, Trust & Performance](https://www.wearedevelopers.com/videos/1808-the-future-of-employee-wellbeing-benefits-trust-performance) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Best Paying Remote Jobs](https://www.wearedevelopers.com/magazine/255-best-paying-remote-jobs) - [Remote Work: Best Practices for Developers](https://www.wearedevelopers.com/magazine/315-remote-work-best-practices-for-developers) - [Remote, Hybrid, or In-Office: What’s Really Best for Developers?](https://www.wearedevelopers.com/magazine/638-remote-hybrid-or-in-office-what-s-really-best-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)