×
Register Here to Apply for Jobs or Post Jobs. X

Senior Decision Intelligence Engineer

Job in Annapolis, Anne Arundel County, Maryland, 21401, USA
Listing for: Humana
Full Time position
Listed on 2026-09-17
Job specializations:
  • Software Development
    Software Engineer, Machine Learning/ ML Engineer, DevOps, AI Engineer (Applied/Software)
Job Description & How to Apply Below
** Become a part of our caring community*
* The Senior Machine Learning Engineer, Decision Intelligence is a hands-on individual contributor responsible for building, deploying, and operating ML and decisioning pipelines for the NBA Decision Intelligence Platform.

This role focuses on production pipeline development, MLOps, feature engineering, scoring workflows, monitoring, and optimization-aware decisioning. You will help ensure the platform selects the right action for the right member while respecting clinical eligibility, suppression rules, channel constraints, program goals, and operational capacity.

You will work closely with ML engineers, data engineers, platform engineers, product owners, and decision engine teams to deliver reliable, scalable, and auditable production systems.

Additional

Job Description

** Required Qualifications*
* + Bachelors in computer science or relevant field

+ 5+ years (post undergraduate level) of software engineering or quantitative research experience building and operating large-scale production systems, with emphasis on data-intensive platforms, recommendation systems, optimization engines, or simulation frameworks serving millions of users.

+ 2+ years (post graduate level) of software engineering or quantitative research experience building and operating large-scale production systems, with emphasis on data-intensive platforms, recommendation systems, optimization engines, or simulation frameworks serving millions of users.

+ 2+ years of hands-on experience implementing reinforcement learning, operations research methods, or simulation-driven decision systems in production. Relevant backgrounds include policy gradient and value-based RL (PPO, A3C, DQN, CQL), stochastic dynamic programming, discrete-event simulation, or large-scale combinatorial or constrained optimization.

+ Deep familiarity with Markov Decision Processes, Bellman-equation-based value estimation, reward or objective shaping, exploration-exploitation tradeoffs, and constraint formulation in real-world decision systems.

+ Demonstrated ability to diagnose failure modes in learned or optimized policies: instability, poor credit assignment across long horizons, and distributional shift across large populations.

+ Proficiency in Python 3.x; experience with PyTorch or Tensor Flow for policy network or learned model implementation.

+

Experience with Ray RLlib or equivalent distributed computation frameworks for large-scale training or optimization.

+

Experience with Databricks, PySpark, and Delta Lake for large-scale ML or data pipelines processing tens of millions of records.

+

Experience with MLflow for experiment tracking, model registry, and artifact management.

+

Experience with shipping systems that operate reliably under production load, not just research or prototype work.

** Preferred Qualifications*
* +

Experience with multi-agent RL frameworks (Petting Zoo or equivalent) or multi-agent simulation and coordination methods.

+ Familiarity with operations research methods applicable to constrained sequential decisioning: linear programming, mixed-integer programming, Lagrangian relaxation, or constraint programming.

+ Experience operating decision or optimization systems in regulated domains (healthcare, finance, or insurance) where member safety, auditability, and explainability are requirements.

+ Experience building simulation environments using Gymnasium, Sim Py, Any Logic, or equivalent frameworks for policy evaluation and backtesting.

+ Familiarity with event-driven feedback loops and how disposition signals feed retraining or re-optimization pipelines.

+ Open Telemetry instrumentation experience for ML or optimization pipeline observability.

** Use your skills to make an impact*
* ** Additional Information*
* ** Work Style:
** Remote/Hybrid - Preferably Boston, MA.

Occasional travel to Humana's offices for training or meetings may be required.

** Work Hours** :
Typical business hours are Monday-Friday, 8 hours/day, 5 days/week-- some flexibility might be possible, depending on business needs.

Very minimal travel might be required for training, meetings, and/or conferences

** Interview Format*
* As part of our hiring process, we will be using on-demand technology provided by Hire Vue, a third-party vendor. This technology provides our team of recruiters and hiring managers with an enhanced method for decision-making through on-demand candidate assessments.

If you are selected to move forward from your application prescreen, you will…
Position Requirements
10+ Years work experience
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary