Machine Learning Engineer, AI Studio
About this opportunity
Career Category
Information Systems
Job Description
Machine Learning Engineer, AI Studio
CAREER LEVEL: GCF 4 – Senior Associate
CAREER TRACK: Individual Contributor
PRIMARY SCOPE: Independent ownership of defined production ML/AI components
ORGANIZATION: Applied AI | AI Studio
ABOUT AMGEN
Amgen harnesses the best of biology and technology to fight the world’s toughest diseases and make people’s lives easier, fuller and longer. We discover, develop, manufacture and deliver innovative medicines to help millions of patients. Amgen helped establish the biotechnology industry more than 40 years ago and remains at the cutting edge of innovation, using technology and human genetic data to push beyond what is known today.
ABOUT THE ROLE
Role Description:
The Machine Learning Engineer offers a unique opportunity to join a fun, innovative engineering team within the AI & Data Science (AI&D) - organization. We are the Applied AI team (AI Studio). AI Studio is Amgen’s enterprise engine for turning high-value business challenges into scalable AI products. We partner with key business partners across the company to identify the right opportunities, shape them into actionable use cases, and design, build, and launch AI products responsibly. Our work spans the full lifecycle from early discovery and rapid prototyping to production deployment, reuse across the enterprise, and measurable business impact. Y ou will independently own defined production components within enterprise AI products and automation solutions.
Your remit may include a model or inference service, data or knowledge pipeline, retrieval component , agent tool, evaluation module, API, workflow or monitoring capability. You will design, release, diagnose and support the component while connecting technical measures to user and workflow outcomes. Within Applied AI, AI Studio turns prioritized business demand into governed, reusable AI assets with accountable ownership and measurable value across software, data, automation, machine learning, Generative AI, RAG, bounded agents, evaluation, observability and lifecycle operations.
Roles & Responsibilities:
Define component boundaries, intended use, acceptance criteria, non-functional requirements, decision consequences, support expectations and technical estimates with product and architecture partners.
Design and implement maintainable Python, SQL, API, data, model, retrieval, agent-tool and workflow components with clear contracts, configuration, testing, error handling and documentation.
Apply EDA, feature engineering, supervised or unsupervised methods, baselines, cross-validation, leakage prevention, calibration, subgroup, threshold, explainability and error analysis where relevant.
Build GenAI, NLP, RAG and bounded agent components using structured output, embeddings, hybrid search, reranking, provenance, citations, permissions, approvals, retries and recoverable failure behavior .
Engineer batch or event-driven data, document, feature, embedding, label and evaluation pipelines with schema validation, lineage, provenance, access control and consistency checks.
Define representative evaluation for model quality, uncertainty, retrieval, grounding, citations, task success, tool correctness, safety, latency, cost and user impact.
Release and support components using cloud services, containers, CI/CD, versioning, monitoring, rollback, incident response and runbooks; lead diagnosis of moderately complex failures.
Apply security, privacy, Responsible AI, validation, auditability, human oversight and applicable GxP controls; contribute reusable assets and guide Associate engineers on familiar work.
Basic Qualifications and Experience:
Bachelor’s/ Master’s degree and 5 to 9 years of Computer Science, IT or related field experience .
Functional Skills:
Production software and AI/ML system design: Python and SQL modules, APIs, background jobs, event flows, testing, performance, observability, source control and maintainable failure semantics.
Statistics, modeling and experimentation: EDA, feature engineering, classification, regression, clustering, ensembles, cross-validation, leakage prevention, calibration, uncertainty and decision-aware error analysis.
GenAI, RAG, knowledge and agents: Prompt and context management, structured output, chunking, embeddings, hybrid retrieval, reranking, citations, access-aware retrieval, tool schemas and human approval.
Data, knowledge and cloud-scale systems: Batch and event-driven pipelines, contracts, lineage, provenance, relational/document/graph/vector stores, APIs, containers, Spark or Databricks and cloud-native services.
AI evaluation and MLOps / LLMOps : Representative measures, versioning, CI/CD, release gates, quality and drift monitoring, incidents, rollback, runbooks, component support and lifecycle traceability.
Responsible and regulated delivery: Least privilege, privacy, prompt-injection safeguards, bias and robustness checks, intended-use documentation, human oversight, auditability and GxP evidence.
Must-Have Skills:
Demonstrated ownership of at least one production software, data, ML, GenAI or automation component .
Strong hands-on proficiency in Python and SQL, with sound software-engineering and testing practices.
Strong capability in at least one of classical ML, GenAI/RAG/ agents or MLOps /platform engineering, with working knowledge of adjacent areas.
Good-to-Have Skills:
Advanced ML and deep learning: Experience with PyTorch , TensorFlow, Hugging Face, scikit-learn, XGBoost , PyMC , computer vision, NLP, GNNs, causal inference or uncertainty estimation.
Advanced GenAI and knowledge systems: Experience with LangChain , LangGraph , LlamaIndex , Semantic Kernel, AutoGen , CrewAI , hybrid retrieval, knowledge graphs, graph RAG or evidence verification.
Cloud, data and MLOps : Experience with AWS, Bedrock or SageMaker, Databricks, Spark, Kubernetes, infrastructure as code, MLflow , Airflow, Kubeflow or GitHub Actions.
Agents and human-AI workflows: Familiarity with MCP-style integration, agent tracing, adversarial testing, durable workflows, permissions, human review, BI or process automation.
Regulated enterprise delivery: Experience in healthcare, life sciences, GxP , validated systems or another regulated or high-impact environment.
Soft Skills:
Independent problem solving and sound component-level technical judgment.
Clear communication of assumptions, evidence, trade-offs, risks and support implications.
Strong collaboration with business SMEs, product, architecture, software, data, platform, evaluation and control partners.
Ownership, reliability and disciplined follow-through from design through production support.
Ability to guide junior engineers and learn new tools through evidence-based experimentation.
EQUAL OPPORTUNITY STATEMENT
Amgen is an Equal Opportunity employer and will consider you without regard to your race, color , religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability status, or any other basis protected by applicable law.
.
Job details
How this role compares
Computed from every other active Information Technology role in our database, not just this employer's listings.
We currently track 1097 comparable Information Technology roles across 55 biopharma companies.
Salary context
143 of 1097 peers report a salary range (USD, annualized)
Peers share this role's job function. This posting doesn't list a seniority level, so peers aren't narrowed by seniority either -- the range below may span more levels than usual.
Where these roles are based
Top locations among the 1097 comparable roles
+ 23 more countries
Seniority mix
612 of 1097 peers have a known seniority level
Therapeutic area mix
1 of 1097 peers have a known therapeutic area; the rest are genuinely unlabeled, not hidden
Similar opportunities
The closest matches from our peer group, ranked by how similar they are, not how well you'd qualify for them -- treat this as market context, not a guaranteed shortlist; a weak match is labeled as one below.
How we calculate "similar"
No black box, no LLM guesswork: a deterministic score built from four normalized attributes. Here's this role's own peer group at different match levels, so you can see the mechanism, not just the result.
Every comparison starts from the same 100-point budget: 25 for working in the same function, 40 for the same therapeutic area, 20 for the same or adjacent seniority, 15 for the same country. A dimension we can't confirm on both sides contributes nothing, never a guess, never a free pass.
0 points, never a partial guess. A role we know almost nothing about beyond its function bottoms out at 25%; it never inflates to 100% just because there's little to compare against. Seniority uses a defined ladder (Associate → Manager → Associate Director → Senior → Principal → Director → Senior Director → Executive/VP) so "Director" and "Senior Director" count as adjacent, but "Director" and "Executive/VP" do not.