Senior AI/ML Engineer
Hyderabad, Telangana
Check the employer’s page before spending time on an application. This is a saved copy of the posting.
- Pay
- Salary not listed in the saved posting
- Work setup
- Unconfirmed
- Employment
- Unconfirmed
What you’ll work on
Full postingLead AI-Driven DevOps Delivery- Drive end-to-end DevOps execution for AI-powered platforms with a focus on automation, scalability, and reliability
Support Core AI Ops & Operational Excellence- Own and enhance operational support for critical AI/ML systems, ensuring high availability, performance, and governed execution
From the employer’s posting
Primary Responsibilities: Lead AI-Driven DevOps Delivery- Drive end-to-end DevOps execution for AI-powered platforms with a focus on automation, scalability, and reliability Support Core AI Ops & Operational Excellence- Own and enhance operational support for critical AI/ML systems, ensuring high availability, performance, and governed execution
Lead AI-Driven DevOps Delivery- Drive end-to-end DevOps execution for AI-powered platforms with a focus on automation, scalability, and reliability Support Core AI Ops & Operational Excellence- Own and enhance operational support for critical AI/ML systems, ensuring high availability, performance, and governed execution Advanced Observability & Monitoring- Design, build, and optimize dashboards using Grafana, Dynatrace, and Splunk to provide real-time insights into system health, model performance, and platform stability
What you’ll bring
All qualificationsCore experience
- 8+ years of experience in DevOps / AI Ops / platform engineering
- Experience with CI/CD pipeline development and automation
- Proficiency in REST APIs, JSON, and YAML for integrations
- 6+ years of hands-on experience with Splunk, Grafana, and Dynatrace dashboards
- Solid experience supporting production deployments and enterprise applications
- 6+ years of experience configuring and managing alerts in monitoring tools
Qualification wording
8+ years of experience in DevOps / AI Ops / platform engineering
Experience with CI/CD pipeline development and automation
Proficiency in REST APIs, JSON, and YAML for integrations
6+ years of hands-on experience with Splunk, Grafana, and Dynatrace dashboards
Solid experience supporting production deployments and enterprise applications
6+ years of experience configuring and managing alerts in monitoring tools
Tools in this posting
- SQL
- Azure
- Grafana
- Kubernetes
Source — Tool mentions in context
- Proficiency in REST APIs, JSON, and YAML for integrations - Working knowledge of SQL for data analysis and troubleshooting At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.
- CI/CD Automation for AI/ML Pipelines- Continuously improve and maintain CI/CD pipelines, enabling automated builds, testing, validation, and secure deployment of AI models and services - Cloud Infrastructure & Platform Engineering (Azure)- Manage and optimize Azure-based infrastructure, including compute, storage, networking, and Kubernetes environments to support scalable AI workloads - API Engineering & Integration Enablement- Develop and maintain Postman collections and API frameworks for seamless integration, validation, and testing of AI services
- Solid experience supporting production deployments and enterprise applications - Solid understanding of cloud platforms (Azure preferred) and distributed systems - Solid foundation in object-oriented programming and code-level debugging
- Support Core AI Ops & Operational Excellence- Own and enhance operational support for critical AI/ML systems, ensuring high availability, performance, and governed execution - Advanced Observability & Monitoring- Design, build, and optimize dashboards using Grafana, Dynatrace, and Splunk to provide real-time insights into system health, model performance, and platform stability - Intelligent Alerting & Incident Prevention- Implement and fine-tune proactive alerting strategies leveraging AIOps principles to detect anomalies, reduce noise, and enable predictive issue resolution
- 8+ years of experience in DevOps / AI Ops / platform engineering - 6+ years of hands-on experience with Splunk, Grafana, and Dynatrace dashboards - 6+ years of experience configuring and managing alerts in monitoring tools
Job description
Optum is a global organization that delivers care, aided by technology to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care, pharmacy benefits, data and resources they need to feel their best. Here, you will find a culture guided by inclusion, talented peers, comprehensive benefits and career development opportunities. Come make an impact on the communities we serve as you help us advance health optimization on a global scale. Join us to start Caring. Connecting. Growing together.
Primary Responsibilities:
- Lead AI-Driven DevOps Delivery
- Drive end-to-end DevOps execution for AI-powered platforms with a focus on automation, scalability, and reliability
- Support Core AI Ops & Operational Excellence
- Own and enhance operational support for critical AI/ML systems, ensuring high availability, performance, and governed execution
- Advanced Observability & Monitoring
- Design, build, and optimize dashboards using Grafana, Dynatrace, and Splunk to provide real-time insights into system health, model performance, and platform stability
- Intelligent Alerting & Incident Prevention
- Implement and fine-tune proactive alerting strategies leveraging AIOps principles to detect anomalies, reduce noise, and enable predictive issue resolution
- CI/CD Automation for AI/ML Pipelines
- Continuously improve and maintain CI/CD pipelines, enabling automated builds, testing, validation, and secure deployment of AI models and services
- Cloud Infrastructure & Platform Engineering (Azure)
- Manage and optimize Azure-based infrastructure, including compute, storage, networking, and Kubernetes environments to support scalable AI workloads
- API Engineering & Integration Enablement
- Develop and maintain Postman collections and API frameworks for seamless integration, validation, and testing of AI services
- AI-Driven Incident Management
- Lead daily incident operations using automation and analytics to accelerate triage, resolution, and root cause identification
- War Room Leadership for Critical Incidents
- Orchestrate high-priority incident war rooms, ensuring cross-team coordination, rapid decision-making, and minimal business impact
- Deep Technical Troubleshooting & Code-Level Debugging
- Diagnose and resolve P1/P2 incidents by analyzing application code, infrastructure, and data pipelines, ensuring end-to-end issue resolution
- Human-in-the-Loop Governance
Ensure all AI-driven actions are governed with human oversight, maintaining control, compliance, and auditability across operations
- Comply with the terms and conditions of the employment contract, company policies and procedures, and any and all directives (such as, but not limited to, transfer and/or re-assignment to different work locations, change in teams and/or work shifts, policies in regards to flexibility of work benefits and/or work environment, alternative work arrangements, and other decisions that may arise due to the changing business environment). The Company may adopt, vary or rescind these policies and directives in its absolute discretion and without any limitation (implied or otherwise) on its ability to do so
Required Qualifications:
- Graduate degree or equivalent experience
- 8+ years of experience in DevOps / AI Ops / platform engineering
- 6+ years of hands-on experience with Splunk, Grafana, and Dynatrace dashboards
- 6+ years of experience configuring and managing alerts in monitoring tools
- Hands-on AI experience
- Experience with CI/CD pipeline development and automation
- Solid experience supporting production deployments and enterprise applications
- Solid understanding of cloud platforms (Azure preferred) and distributed systems
- Solid foundation in object-oriented programming and code-level debugging
- Exposure to AI Ops concepts such as predictive monitoring and automated remediation
- Proficiency in REST APIs, JSON, and YAML for integrations
- Working knowledge of SQL for data analysis and troubleshooting
At UnitedHealth Group, our mission is to help people live healthier lives and make the health system work better for everyone. We believe everyone-of every race, gender, sexuality, age, location and income-deserves the opportunity to live their healthiest life. Today, however, there are still far too many barriers to good health which are disproportionately experienced by people of color, historically marginalized groups and those with lower incomes. We are committed to mitigating our impact on the environment and enabling and delivering equitable care that addresses health disparities and improves health outcomes - an enterprise priority reflected in our mission.
Your next step
Check the employer’s posting for the current role and application details.
Already applied? Track this application
Source & posting history
Source notes
Source excerptsSelected passages from the saved posting. Check the full description for conditions and exceptions.
- Pay
No pay amount identified in the saved description.
- Location & working pattern
Hyderabad, Telangana
Working pattern and location restrictions need checking in the full posting.
- Work authorization
No clear work-authorization passage found. Eligibility is unconfirmed.
- Status in our records
- Unknown — awaiting fresh evidence
- First seen by us
- Aug 31, 2026
- Recorded sightings
- 14
- Last seen by us
- Sep 2, 2026
These dates show when we found the listing. Check the employer’s website to confirm it is still accepting applications.
Report an errorSee how this role fits your experience
Add your resume to compare the role’s scope, tools and requirements with your experience.