AbbVie Posted September 15, 2026

Data Engineer, AI Enablement

North Chicago, IL Full-time
Notify me about similar jobs

AbbVie is the source of truth for this posting and owns the application process. We surface normalized context and market comparison you won't find on the original listing.

About this opportunity

Company Description

About AbbVie

AbbVie's mission is to discover and deliver innovative medicines and solutions that solve serious health issues today and address the medical challenges of tomorrow. We strive to have a remarkable impact on people's lives across several key therapeutic areas including immunology, oncology and neuroscience - and products and services in our Allergan Aesthetics portfolio. For more information about AbbVie, please visit us at  www.abbvie.com . Follow @abbvie on  LinkedIn,   Facebook ,  Instagram ,  X  and  YouTube.

Job Description

AbbVie’s Business Technology Solutions (BTS) Information Research (IR) organization is seeking a Data Engineer, AI Enablement to help deliver trusted, well-structured, AI-ready data products within ARCH, AbbVie’s R&D Convergence Hub. As part of the DELOS team, Data Exploration and Linked Outcome Solutions, this role helps build the reliable data foundations needed to advance analytics, reporting, knowledge graph capabilities, machine learning, and AI-enabled use cases across R&D.

In this role, you will independently design, develop, and operate scalable data pipelines and curated data products that make high-value research data easier to find, connect, understand, and use. The work spans data curation, normalization, modeling, metadata, lineage, quality controls, governance, documentation, and publication to the ARCH knowledge graph. Rather than developing AI models directly, you will ensure that data science, AI engineering, and research partners have the reliable, accessible, and appropriately governed data they need to deliver trusted outcomes.

Working closely with R&D stakeholders, data scientists, machine learning engineers, platform teams, architects, and data owners, you will help translate scientific and business needs into dependable data solutions. You will also help scale delivery by providing technical guidance to contracted engineers supporting the same data products, translating requirements into clear work, reviewing outputs, helping remove barriers, and ensuring results meet agreed quality, documentation, and acceptance standards.

Under the direction of the Associate Director – Data Strategy, AI & Knowledge Enablement, this role is an opportunity to contribute at the center of AbbVie’s R&D data transformation. The data foundations you build will help determine which analytics, knowledge graph, and AI use cases are possible across research, and how confidently the organization can use their output to support scientific decision-making.

Responsibilities 

AI-Ready Data Product Engineering: Design, build, and operate curated, reusable data products that make high-value R&D data easier to find, connect, understand, and use. Collect, integrate, normalize, model, and transform data from databases, applications, APIs, licensed external sources, and other systems into ARCH and related data environments.

Trusted Data Foundation Enablement: Establish reliable, scalable data foundations that support analytics, reporting, knowledge graph capabilities, machine learning, and AI-enabled use cases. Ensure data assets are structured, documented, accessible, governed, traceable, and fit for downstream consumption.

AI, RAG & Knowledge Graph Readiness: Prepare data and documents for AI and knowledge discovery use cases by cleaning, standardizing, enriching, labeling, organizing metadata, supporting chunking, and embedding workflows, and producing vector database-ready assets. Enable publication of curated data to the ARCH knowledge graph.

Data Quality, Governance & Documentation: Apply data quality and governance practices, including accuracy and completeness checks, metadata, lineage, access controls, privacy, license terms, assumptions, quality rules, and appropriate-use guidance so data consumers can understand and trust the assets they use.

Technical Coordination & Delivery Support: Collaborate with data scientists, machine learning engineers, software engineers, platform teams, architects, data owners, and R&D stakeholders to translate scientific and business requirements into usable AI-ready data products. Provide technical guidance to contracted engineers, clarify work, review outputs, help remove barriers, and support delivery against agreed quality and acceptance standards.

Operational Reliability & Continuous Improvement: Monitor pipeline performance, data freshness, cost, failures, and delivery issues; troubleshoot and resolve problems before they impact data consumers. Contribute to reusable engineering patterns, automation, process improvements, and consistent ways of working across data product workflows.

Compliance & Standards: Follow applicable Corporate and Divisional policies, including GxP compliance, data security, software development lifecycle practices, data governance standards, and relevant regulatory or contractual requirements.

Qualifications

Required:

Bachelor’s Degree with 5 years of experience; OR Master’s Degree with 4 years of experience in information technology, data engineering, data management, analytics, life sciences, or a related field.

Hands-on experience designing, developing, and operating production data pipelines and curated data products using SQL, Python, ETL/ELT patterns, and workflow orchestration tools such as Airflow.

Working knowledge of modern data platforms, data integration, data warehousing or lakehouse patterns, distributed SQL or big data environments, cloud infrastructure, and analytics enablement.

Experience preparing data for downstream analytics, machine learning, knowledge graph, or retrieval use cases, including cleaning, standardization, enrichment, structuring, metadata organization, and support for embedding or vector-search workflows.

Experience applying data quality, metadata management, governance, lineage, documentation, and data modeling practices to support trusted, reusable data products.

Experience collaborating with cross-functional business, scientific, technical, platform, vendor, contractor, or managed-services teams to translate requirements and deliver fit-for-purpose data assets.

Ability to operate with a high degree of autonomy, manage priorities across concurrent workstreams, modify approach when needed, escalate open issues, and keep stakeholders informed through clear written and verbal communication.

Demonstrated ability to learn, understand, and apply new data engineering, platform, and AI-enablement technologies, and to serve as a technical resource for others.

Experience providing technical input, clarifying requirements, and reviewing outputs from contracted, vendor, or managed-services engineers without direct reporting authority.

Strong communication, planning, and organizational skills, with the ability to explain technical concepts and keep stakeholders informed.

Data product engineering mindset, with the ability to shape reusable, well-structured data assets that are practical, scalable, and fit for analytics and AI-enabled use.

Data curation and stewardship mindset, with attention to quality, metadata, lineage, governance, standards, documentation, and appropriate use.

Technical fluency across data platforms, pipelines, integration patterns, orchestration, cloud environments, and data delivery practices sufficient to work effectively with engineering and platform teams.

Operational discipline across monitoring, troubleshooting, prioritization, issue resolution, automation, reusable patterns, and continuous improvement.

Technical coordination and influence, with the ability to clarify priorities, guide work, review outputs, resolve ambiguity, and coordinate across internal and external contributors.

Stakeholder communication, with the ability to frame tradeoffs, risks, dependencies, and progress in a clear and practical way for technical, scientific, and business audiences.

Preferred:

Pharmaceutical or healthcare industry experience preferred.

Experience supporting research, discovery, translational, clinical, scientific, or other life sciences data environments.

Familiarity with graph databases, knowledge graphs, ontology-based data structures, semantic data, metadata-driven data products, or linked-data concepts.

Experience working with AWS-based, cloud-based, lakehouse, or modern data platform technologies such as Databricks, Spark, Snowflake, Neo4j, or similar tools.

Experience working with regulated data environments, including data governance, documentation, security, privacy, license terms, or compliance expectations.

Exposure to analytics, machine learning, retrieval-augmented generation (RAG), embeddings, vector databases, AI-search patterns, or AI-ready data product delivery.

Familiarity with Agile practices or planning tools such as Jira, including backlog refinement, sprint planning, prioritization, acceptance criteria, and delivery tracking.

Additional Information

​Applicable only to applicants applying to a position in any location with pay disclosure requirements under state or local law: ​

The compensation range described below is the range of possible base pay compensation that the Company believes in good faith it will pay for this role at the time of this posting based on the job grade for this position. Individual compensation paid within this range will depend on many factors including geographic location, and we may ultimately pay more or less than the posted range. This range may be modified in the future. ​

We offer a comprehensive package of benefits including paid time off (vacation, holidays, sick), medical/dental/vision insurance and 401(k) to eligible employees.​

This job is eligible to participate in our short-term incentive programs. ​

Note: No amount of pay is considered to be wages or compensation until such amount is earned, vested, and determinable. The amount and availability of  any bonus, commission, incentive, benefits, or any other form of compensation and benefits that are allocable to a particular employee remains in the Company's sole and absolute discretion unless and until paid and may be modified at the Company’s sole and absolute discretion, consistent with applicable law. ​

AbbVie is an equal opportunity employer and is committed to operating with integrity, driving innovation, transforming lives and serving our community.  Equal Opportunity Employer/Veterans/Disabled.

US & Puerto Rico only - to learn more, visit  https://www.abbvie.com/join-us/equal-employment-opportunity-employer.html

US & Puerto Rico applicants seeking a reasonable accommodation, click here to learn more:

https://www.abbvie.com/join-us/reasonable-accommodations.html

Job details

Seniority
Not listed
Function
Data & Digital
Therapeutic area
Not listed
Location
North Chicago, IL
Employment type
Full-time

How this role compares

Computed from every other active Data & Digital role in our database, not just this employer's listings.

We currently track 244 comparable Data & Digital roles across 43 biopharma companies.

244Comparable roles tracked
230Currently active
43Companies hiring similar roles
18Countries represented

Salary context

74 of 244 peers report a salary range (USD, annualized)

Peers share this role's job function. This posting doesn't list a seniority level, so peers aren't narrowed by seniority either -- the range below may span more levels than usual.

This roleSubject $84,500/yr – $162,000/yr
Lowest disclosed · Agentic AI Data Engineer - CMC Data Integration · Lilly $65,250/yr – $169,400/yr
Peer group range $117,325 – $339,950 (median $199,100)

Where these roles are based

Top locations among the 244 comparable roles

United States92
India80
France16
Spain15
United Kingdom10
Germany6

+ 12 more countries

Seniority mix

143 of 244 peers have a known seniority level

Senior40
Manager27
Associate Director24
Principal16
Director15
Executive/VP10
Senior Director5
Associate4
Intern/Fellow/Postdoc2

Therapeutic area mix

4 of 244 peers have a known therapeutic area; the rest are genuinely unlabeled, not hidden

Immunology3
Ophthalmology1

Similar opportunities

The closest matches from our peer group, ranked by how similar they are, not how well you'd qualify for them -- treat this as market context, not a guaranteed shortlist; a weak match is labeled as one below.

40%similar
Lilly Indianapolis, Indiana, United States of America
Same function Same country
40%similar
Novartis Cambridge (USA), United States
Same function Same country
40%similar
Regeneron Pharmaceuticals, Inc (USA) Tarrytown, United States Executive/VP
Same function Same country
40%similar
Regeneron Pharmaceuticals, Inc (USA) Cambridge, United States Manager
Same function Same country
40%similar
Regeneron Pharmaceuticals, Inc (USA) Warren, United States Principal
Same function Same country
40%similar
Regeneron Pharmaceuticals, Inc (USA) Cambridge, United States Manager
Same function Same country

Notify me about similar jobs

Get an email when we spot other openings like this one – same job function, comparable seniority, roles you'd actually want to see.

How we calculate "similar"

No black box, no LLM guesswork: a deterministic score built from four normalized attributes. Here's this role's own peer group at different match levels, so you can see the mechanism, not just the result.

Every comparison starts from the same 100-point budget: 25 for working in the same function, 40 for the same therapeutic area, 20 for the same or adjacent seniority, 15 for the same country. A dimension we can't confirm on both sides contributes nothing, never a guess, never a free pass.

40%
Agentic AI Data Engineer - CMC Data Integration
Lilly · Indianapolis, Indiana, United States of America · Seniority not listed
Function Therapeutic area Seniority Country
40%
Manager Statistical Programming
Regeneron Pharmaceuticals, Inc (USA) · Cambridge, United States · Manager
Function Therapeutic area Seniority Country
40%
Senior Manager, Statistical Programming
Regeneron Pharmaceuticals, Inc (USA) · Cambridge, United States · Manager
Function Therapeutic area Seniority Country
40%
Manager, Applied AI, Advanced Informatics
Regeneron Pharmaceuticals, Inc (USA) · Tarrytown, United States · Manager
Function Therapeutic area Seniority Country
Unmatched or unknown dimensions score exactly the same: 0 points, never a partial guess. A role we know almost nothing about beyond its function bottoms out at 25%; it never inflates to 100% just because there's little to compare against. Seniority uses a defined ladder (Associate → Manager → Associate Director → Senior → Principal → Director → Senior Director → Executive/VP) so "Director" and "Senior Director" count as adjacent, but "Director" and "Executive/VP" do not.