Sr Specialist - AI Operations
About this opportunity
Working with Us
Challenging. Meaningful. Life-changing. Those aren’t words that are usually associated with a job. But working at Bristol Myers Squibb is anything but usual. Here, uniquely interesting work happens every day, in every department. From optimizing a production line to the latest breakthroughs in cell therapy, this is work that transforms the lives of patients, and the careers of those who do it. You’ll get the chance to grow and thrive through opportunities uncommon in scale and scope, alongside high-achieving teams. Take your career farther than you thought possible.
Bristol Myers Squibb recognizes the importance of balance and flexibility in our work environment. We offer a wide variety of competitive benefits, services and programs that provide our employees with the resources to pursue their goals, both at work and in their personal lives. Read more: careers.bms.com/working-with-us .
Position Overview
The Senior Specialist of AI Operations for AI Applications plays a critical role in supporting, maintaining, and optimizing AI applications for BMS’s PDS. This role combines technical expertise in IT operations, cloud engineering, automation, and AIOps to ensure AI applications run reliably, securely, and at scale. The ideal candidate is a hands-on engineer with strong problem-solving skills and experience supporting production machine learning or data-intensive systems.
Based on your function, department or individual position, you will have the opportunity to discuss with your Manager the option to work remotely up to 50% of the time, over a two-week period, with the flexibility to choose the days that align with your collaboration needs.
Key Responsibilities
AI Application Support & Production Operations
Provide operational support for production AI applications, ensuring availability, performance, and stability.
Troubleshoot and resolve issues across data pipelines, model deployments, microservices, and integration layers.
Implement monitoring, logging, and alerting solutions tailored to AI workload behavior.
Maintain runbooks, operational documentation, and incident response procedures.
Automation & AIOps Engineering
Deploy and manage containerized workloads using Kubernetes or similar orchestration frameworks.
Support and enhance CI/CD and AI orchestration used for model training, testing, deployment, and monitoring.
Automate operational tasks such as environment provisioning, configuration, scaling, and failover processes.
Partner with AI engineers to ensure smooth handoff from development to production operations.
System Monitoring, Reliability & Performance
Develop and maintain reliability engineering practices for AI systems, including SLO/SLI definitions and capacity planning.
Conduct root cause analysis (RCA) for incidents and drive continuous improvement to prevent recurrence.
Optimize system performance, cost efficiency, and resource utilization.
Develop AI Native observability dashboards
Cross-Functional Collaboration
Work closely with AI/ML teams, software engineering, data engineering, product management, and IT security.
Qualifications
Required
Bachelor’s degree in Computer Science, Information Technology, Engineering, or related discipline (or equivalent experience).
3+ years of experience in IT operations, site reliability engineering (SRE), cloud engineering, or DevOps.
Hands-on experience with cloud environments (AWS), including compute, networking, IAM, and storage.
Strong proficiency with Kubernetes, Docker, and containerized application architectures.
Experience with automation tools and IaC frameworks (Terraform, CloudFormation, Ansible, etc.).
Familiarity with LLM/ML tools and workflows (e.g., MLflow, Kubeflow, SageMaker, Vertex AI, Databricks).
Proficient in scripting languages such as Python, Bash, or PowerShell.
Strong troubleshooting skills across distributed systems, databases, APIs, and pipelines.
Preferred
Experience supporting AI applications or generative AI workloads (LLMs, vector databases, embedding pipelines).
Knowledge of observability platforms (Langsmith, Prometheus, Grafana, ELK/EFK, Datadog, New Relic, Splunk).
Familiarity with ITIL practices, change management, and incident management frameworks.
Background in highly regulated environments (healthcare, finance, government, etc.).
Certifications in cloud (AWS/GCP/Azure), DevOps, or MLOps.
Key Competencies
Technical Depth: Strong understanding of distributed systems, cloud architecture, and AI/ML operational workflows.
Problem Solving: Ability to diagnose complex issues and design durable solutions.
Automation Mindset: Focused on eliminating manual toil and improving operational efficiency.
Collaboration: Comfortable working across data, engineering, security, and product teams.
Reliability Focus: Dedicated to building robust, resilient, and well-monitored systems.
About the Role’s Impact
This role is central to ensuring that AI applications operate smoothly in production, enabling the organization to confidently scale its AI initiatives. The Senior Engineer will help build and refine the technological foundation that supports AI innovation and operational excellence.
Why you should apply
You will help patients in their fight against serious diseases
You will be part of a company that encourages excellence and innovation, respects diversity, develops leaders and values its employees.
You’ll get a competitive salary and a great benefits package including, but not only, an annual bonus, pension contribution, family health insurance, 27 days of annual leave , access to BMS Cruiserath on-site gym and life assurance
#LI-Hybrid
If you come across a role that intrigues you but doesn’t perfectly line up with your resume, we encourage you to apply anyway. You could be one step away from work that will transform your life and career.
Uniquely Interesting Work, Life-changing Careers
With a single vision as inspiring as “Transforming patients’ lives through science™ ”, every BMS employee plays an integral role in work that goes far beyond ordinary. Each of us is empowered to apply our individual talents and unique perspectives in a supportive culture, promoting global participation in clinical trials, while our shared values of passion, innovation, urgency, accountability, inclusion and integrity bring out the highest potential of each of our colleagues.
On-site Protocol
BMS has an occupancy structure that determines where an employee is required to conduct their work. This structure includes site-essential, site-by-design, field-based and remote-by-design jobs. The occupancy type that you are assigned is determined by the nature and responsibilities of your role:
Site-essential roles require 100% of shifts onsite at your assigned facility. Site-by-design roles may be eligible for a hybrid work model with at least 50% onsite at your assigned facility. For these roles, onsite presence is considered an essential job function and is critical to collaboration, innovation, productivity, and a positive Company culture. For field-based and remote-by-design roles the ability to physically travel to visit customers, patients or business partners and to attend meetings on behalf of BMS as directed is an essential job function.
Supporting People with Disabilities
BMS is dedicated to ensuring that people with disabilities can excel through a transparent recruitment process, reasonable workplace accommodations/adjustments and ongoing support in their roles. Applicants can request a reasonable workplace accommodation/adjustment prior to accepting a job offer. If you require reasonable accommodations/adjustments in completing this application, or in any part of the recruitment process, direct your inquiries to adastaffingsupport@bms.com . Visit careers.bms.com/ eeo -accessibility to access our complete Equal Employment Opportunity statement.
Candidate Rights
BMS will consider for employment qualified applicants with arrest and conviction records, pursuant to applicable laws in your area.
If you live in or expect to work from Los Angeles County if hired for this position, please visit this page for important additional information: https://careers.bms.com/california-residents/
Data Protection
We will never request payments, financial information, or social security numbers during our application or recruitment process. Learn more about protecting yourself at https://careers.bms.com/fraud-protection .
Any data processed in connection with role applications will be treated in accordance with applicable data privacy policies and regulations.
If you believe that the job posting is missing information required by local law or incorrect in any way, please contact BMS at TAEnablement@bms.com . Please provide the Job Title and Requisition number so we can review. Communications related to your application should not be sent to this email and you will not receive a response. Inquiries related to the status of your application should be directed to Chat with Ripley.
R1605220 : Sr Specialist - AI Operations
Job details
How this role compares
Computed from every other active Information Technology role in our database, not just this employer's listings.
We currently track 378 comparable Senior Information Technology roles across 33 biopharma companies.
Salary context
45 of 378 peers report a salary range (USD, annualized)
Peers share this role's job function and a matching or adjacent seniority level -- not necessarily the same therapeutic area or country.
Where these roles are based
Top locations among the 378 comparable roles
+ 15 more countries
Seniority mix
378 of 378 peers have a known seniority level
Therapeutic area mix
0 of 378 peers have a known therapeutic area; the rest are genuinely unlabeled, not hidden
No peers with a known therapeutic area yet.
Similar opportunities
The closest matches from our peer group, ranked by how similar they are, not how well you'd qualify for them -- treat this as market context, not a guaranteed shortlist; a weak match is labeled as one below.
How we calculate "similar"
No black box, no LLM guesswork: a deterministic score built from four normalized attributes. Here's this role's own peer group at different match levels, so you can see the mechanism, not just the result.
Every comparison starts from the same 100-point budget: 25 for working in the same function, 40 for the same therapeutic area, 20 for the same or adjacent seniority, 15 for the same country. A dimension we can't confirm on both sides contributes nothing, never a guess, never a free pass.
0 points, never a partial guess. A role we know almost nothing about beyond its function bottoms out at 25%; it never inflates to 100% just because there's little to compare against. Seniority uses a defined ladder (Associate → Manager → Associate Director → Senior → Principal → Director → Senior Director → Executive/VP) so "Director" and "Senior Director" count as adjacent, but "Director" and "Executive/VP" do not.