Develops and productionizes advanced AI and machine learning solutions for clinical NLP applications. Responsibilities include fine-tuning large language models, building GraphRAG and semantic retrieval systems, optimizing embeddings, developing information extraction and summarization methods, and scaling distributed multi-GPU training. The role applies reinforcement learning alignment techniques, translates research into production systems, collaborates cross-functionally, and requires strong expertise in NLP, deep learning, transformers, cloud ML platforms, and healthcare data.
Requisition Number: 2381960
Optum is a global organization that delivers care, aided by technology to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care, pharmacy benefits, data and resources they need to feel their best. Here, you will find a culture guided by inclusion, talented peers, comprehensive benefits and career development opportunities. Come make an impact on the communities we serve as you help us advance health optimization on a global scale. Join us to start Caring. Connecting. Growing together.
The AI/ML Engineer is a primary driver in the design and development of state-of-the-art Artificial Intelligence solutions for medical applications. The Principal Data Scientist ML/DL works closely with senior data scientists, machine learning engineers, software engineers and subject matter experts on current company technologies and forward-looking projects. They are the drivers of new research and solution implementation in the creation of novel artificial intelligence approaches.
Primary responsibilities include the enhancement of existing company NLP technologies and extension of those systems in new cloud-based applications. Emphasis is on development of novel machine/deep learning techniques for information extraction and synthesis. They translate research code into clinical NLP solutions deployed at scale in production environments including statistical methods, deep learning, and large language model technologies. Work will involve all aspects of methods development from initial PoC implementation to performance characterization and production launch of new methods.
The successful candidate will have a strong history of publication in Machine/Deep Learning with an emphasis on Natural Language Processing, Information Retrieval and/or Information Extraction. Exposure to recent research literature and the ability to effectively implement new technologies is key. The successful candidate will have proven success in taking machine/deep learning solutions to production environments. Strong technical skills are required.
Primary Responsibilities:
Required Qualifications:
Preferred Qualifications:
Bonus Skills
At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.
Optum is a global organization that delivers care, aided by technology to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care, pharmacy benefits, data and resources they need to feel their best. Here, you will find a culture guided by inclusion, talented peers, comprehensive benefits and career development opportunities. Come make an impact on the communities we serve as you help us advance health optimization on a global scale. Join us to start Caring. Connecting. Growing together.
The AI/ML Engineer is a primary driver in the design and development of state-of-the-art Artificial Intelligence solutions for medical applications. The Principal Data Scientist ML/DL works closely with senior data scientists, machine learning engineers, software engineers and subject matter experts on current company technologies and forward-looking projects. They are the drivers of new research and solution implementation in the creation of novel artificial intelligence approaches.
Primary responsibilities include the enhancement of existing company NLP technologies and extension of those systems in new cloud-based applications. Emphasis is on development of novel machine/deep learning techniques for information extraction and synthesis. They translate research code into clinical NLP solutions deployed at scale in production environments including statistical methods, deep learning, and large language model technologies. Work will involve all aspects of methods development from initial PoC implementation to performance characterization and production launch of new methods.
The successful candidate will have a strong history of publication in Machine/Deep Learning with an emphasis on Natural Language Processing, Information Retrieval and/or Information Extraction. Exposure to recent research literature and the ability to effectively implement new technologies is key. The successful candidate will have proven success in taking machine/deep learning solutions to production environments. Strong technical skills are required.
Primary Responsibilities:
- Develop end-to-end training and fine-tuning of Large Language Models (LLMs), including both open-source (e.g., Qwen, LLaMA, Mistral) and closed-source (e.g., OpenAI, Gemini, Anthropic) ecosystems
- Demonstrated publication record in AI domain especially relating to text extraction and summarization
- Architect and implement GraphRAG pipelines, including knowledge graph representation and retrieval for enhanced contextual grounding
- Design, train, and optimize semantic and dense vector embeddings for document understanding, search, and retrieval
- Develop semantic retrieval systems with advanced document segmentation and indexing strategies
- Build and scale distributed training environments using NCCL and InfiniBand for multi-GPU and multi-node training
- Apply reinforcement learning techniques (e.g., RLHF, RLAIF) to align model behavior with human preferences and domain-specific goals
- Collaborate with cross-functional teams to translate business needs into AI-driven solutions and deploy them in production environments
- Comply with the terms and conditions of the employment contract, company policies and procedures, and any and all directives (such as, but not limited to, transfer and/or re-assignment to different work locations, change in teams and/or work shifts, policies in regards to flexibility of work benefits and/or work environment, alternative work arrangements, and other decisions that may arise due to the changing business environment). The Company may adopt, vary or rescind these policies and directives in its absolute discretion and without any limitation (implied or otherwise) on its ability to do so
Required Qualifications:
- Bachelor's degree in computer science, machine Learning, or related field or equivalent certification
- 5+ years of experience in applied AI/ML with statistics, with a strong track record of delivering production-grade models
- Experience with PyTorch
- Experience with Hybrid NLP solutions that combine symbolic and machine learning approaches
- Deep expertise in: NLP, Fundamental machine learning, deep learning, transformer, state space-based architecture Azure ML and/or AWS
- Deep knowledge and extensive experience with Machine/Deep Learning frameworks including transformer architectures, state space models, large language models, and agentic approaches
- Knowledge of algorithms and techniques within a computational domain with emphasis on text processing
- Solid in Python coding, SQL and database queries, data preparation, and analysis
- Exploratory Data Analysis (EDA)
- Embedding models (e.g., BGE, E5, SimCSE)
- Semantic search and vector databases (e.g., FAISS, Weaviate, Milvus)
- Model fusion and ensemble techniques (stacking, boosting, gating)
Preferred Qualifications:
- Document segmentation and preprocessing (OCR, layout parsing)
- Optimization algorithms (Bayesian, Particle Swarm, Genetic Algorithms)
- Reinforcement learning (e.g., RLHF, PPO, DPO, GRPO), Supervised Fine Tuning (SFT), LoRA, QLoRA, axolotl
- Prompt optimization framework (AutoPrompt, GreaterPrompt, DSPy), GEPA
Bonus Skills
- Experience with healthcare data and medical coding systems (e.g., CPT, CM, PCS)
- Familiarity with regulatory and compliance frameworks in AI deployment
- Contributions to open-source AI projects or published research. And/Or ability to take research papers to poc - production
At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.
Similar Jobs at Optum
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Designs and operates scalable Databricks, Airflow, Snowflake, and Azure data platforms while developing AI/ML, Generative AI, RAG, API, and MLOps solutions. Responsibilities include building pipelines, vector databases, intelligent workflows, model deployment and monitoring, governance controls, and cloud-native AI services. The role collaborates with data, analytics, product, and platform teams, contributes to architecture and technical decisions, resolves production issues, and evaluates emerging AI technologies.
Top Skills:
Apache AirflowSparkAutogenAzure Ai ServicesAzure Api ManagementAzure App ServiceAzure FunctionsAzure Machine LearningAzure OpenaiCi/CdCrewaiDatabricksEmbeddingsGenerative AiGitInfrastructure As CodeLangchainLarge Language ModelsMlopsPrompt EngineeringPysparkPythonRagRest ApisSemantic KernelSnowflakeVector Databases
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Supports the design, development, testing, deployment, and monitoring of production AI/ML solutions. Responsibilities include data preparation, exploratory analysis, feature engineering, model training, validation, drift analysis, experiment tracking, documentation, Responsible AI, and collaboration with engineering, product, data science, and UX teams. The role also contributes to scalable cloud ML infrastructure, production pipelines, model observability, architecture governance, and mentoring other ML professionals.
Top Skills:
AWSAzureBigQueryDatabricksDataflowEvent HubsFeature StoresGoogle Cloud PlatformKafkaMlflowNumpyOnnxPandasPub/SubPythonPyTorchScikit-LearnSQLSynapseTensorFlowVector Databases
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Design and develop autonomous agentic AI and multi-agent systems for MarTech use cases. Build LLM-powered applications, orchestration workflows, reusable tools, enterprise integrations, and human-in-the-loop governance. Implement enterprise-scale RAG pipelines involving document ingestion, embeddings, retrieval, ranking, and vector databases. Evaluate and optimize models for quality, latency, cost, and performance while developing monitoring and evaluation frameworks. Full-stack development includes Java, React, Next.js, Python, REST APIs, SQL, and Azure cloud services.
Top Skills:
Agentic AiAi Orchestration FrameworksAzure Blob StorageAzure Cloud PlatformAzure Cosmos DbAzure Kubernetes Service (Aks)JavaLarge Language Models (Llms)Next.JsPrompt EngineeringPythonReactRest ApisRetrieval-Augmented Generation (Rag)SQLVector Databases
What you need to know about the Kolkata Tech Scene
When considering the industries shaping India's tech scene, gaming might not immediately come to mind. However, in the last decade, increased internet usage and greater access to mobile devices have catapulted the industry to new heights, with Kolkata-based companies like Virtualinfocom, Red Apple Technologies and Digitoonz, at the forefront, driving the design and animation of new gaming titles for players.

