Lilly Posted August 10, 2026

Consultant/Sr. Consultant, Consumer Analytics Data Engineering

Bangalore, Karnātaka, India FULL_TIME
Data & Digital Senior

Lilly is the source of truth for this posting and owns the application process. We surface normalized context and market comparison you won't find on the original listing.

About this opportunity

At Lilly, the work is demanding because patients are waiting. We unite caring with discovery to help make life better for people around the world, knowing that every decision, every detail, and every day matters. Headquartered in Indianapolis, Indiana, our over 50,000 employees around the globe take on complex challenges to discover and deliver life-changing medicines, strengthen how health is understood and managed, and support the communities we serve. This is hard, urgent, selfless work, but it’s work worth doing. If you’re driven by purpose and ready to bring your best to work that truly matters for patients, we invite you to join us.

As Eli Lilly strives to achieve its purpose of making life better for patients, we have been building up our in-house ‘Consumer Experience’ function, which designs and executes the next-generation marketing campaigns aimed at informing and educating consumers (or patients) directly.

We are seeking a highly skilled Consultant/Sr. Consultant to work with consumer analytics data engineering initiatives at Lilly, Bengaluru. This role is primarily focused on the Databricks platform and AWS ecosystem, designing and maintaining robust data pipelines, lakehouse architectures, and semantic layers that power advanced analytical solutions and AI/agentic workflows for consumer insights. The ideal candidate brings deep hands-on Databricks expertise, good AWS data engineering skills, and a working knowledge of semantic layer design and AI agent development. The role will be a blend of technical expertise along with business/domain integration.

Job Responsibilities:

Design, develop, and implement scalable ETL/ELT pipelines for extracting, transforming, and loading consumer data from various sources (e.g., CRM, marketing platforms, DCM, GA4, digital channels), with Databricks/AWS as the primary execution platform.

Develop and manage end-to-end solutions on Databricks including Unity Catalog, Delta Live Tables, Databricks Workflows, Databricks SQL, etc.; own platform governance covering schemas, permissions, and data lineage.

Design multi-hop lakehouse architectures (Bronze / Silver / Gold) using Delta Lake; optimize Spark compute, cluster configurations, and Auto Loader for performance and cost efficiency.

Leverage AWS data services, S3, Glue, Lambda and Redshift, in conjunction with Databricks to build reliable, end-to-end consumer data flows.

Architect and optimize data models and schemas to support complex analytical queries and reporting requirements related to consumer behaviour, preferences, and engagement.

Publish semantic layers (metrics definitions, certified datasets, business logic) consumed by downstream BI tools and AI agents; build and deploy agentic workflows using Databricks AI Functions or similar frameworks.

Ensure data quality, integrity, and governance across all consumer data assets by implementing validation rules, schema evolution controls, and monitoring processes through Unity Catalog.

Collaborate with data scientists, business analysts, and marketing teams to understand data needs and translate them into technical data engineering solutions; partner to productionize ML models and feature stores on Databricks.

Implement automation for data ingestion, processing, and delivery with a focus on efficiency, reliability, and SLA adherence.

Troubleshoot and resolve data-related issues, performing root cause analysis and implementing corrective actions.

Stay current with Databricks platform updates, AWS data services, and emerging best practices in data engineering and AI-driven analytics.

Job Qualifications:

Bachelor's or Master's degree in Computer Science, Engineering, Information Technology, or a related quantitative field.

3-8 years of data engineering experience with hands-on production experience on Databricks.

Deep knowledge of Databricks platform architecture, Unity Catalog, Delta Lake, Databricks Workflows, Databricks SQL, and cluster/compute management. (Preferred)

Experience designing lakehouse architectures (medallion/multi-hop patterns) at scale and with ETL/ELT orchestration tools. (Preferred)

Experience with AWS data stack: S3, Glue, Redshift, Lambda, and IAM.

Strong proficiency in Python and PySpark; advanced SQL for complex analytical workloads.

Familiarity with Software Development Life Cycle (SDLC) practices, including version control (Git), CI/CD pipelines, code reviews, and agile development methodologies. (Preferred)

Strong understanding of data warehousing concepts, dimensional modeling, and data governance principles.

Hands-on experience building semantic layers (e.g., dbt metrics, Databricks AI/BI semantic layer) and creating AI agents or agentic pipelines (Databricks AI Functions etc.) .

Excellent problem-solving, communication, and stakeholder collaboration skills.

Lilly is dedicated to helping individuals with disabilities to actively engage in the workforce, ensuring equal opportunities when vying for positions. If you require accommodation to submit a resume for a position at Lilly, please complete the accommodation request form ( https://careers.lilly.com/us/en/workplace-accommodation ) for further assistance. Please note this is for individuals to request an accommodation as part of the application process and any other correspondence will not receive a response.

Lilly does not discriminate on the basis of age, race, color, religion, gender, sexual orientation, gender identity, gender expression, national origin, protected veteran status, disability or any other legally protected status.

#WeAreLilly

Job details

Seniority
Senior
Function
Data & Digital
Therapeutic area
Not listed
Location
Bangalore, Karnātaka, India
Employment type
FULL_TIME

How this role compares

Computed from every other active Data & Digital role in our database, not just this employer's listings.

We currently track 92 comparable Senior Data & Digital roles across 23 biopharma companies.

92Comparable roles tracked
87Currently active
23Companies hiring similar roles
13Countries represented

Salary context

16 of 92 peers report a salary range (USD, annualized)

Peers share this role's job function and a matching or adjacent seniority level -- not necessarily the same therapeutic area or country.

This roleSubject Not listed on this posting
Lowest disclosed · Senior Associate, Metadata Engineering · Pfizer $79,400/yr – $132,400/yr
Highest disclosed · Associate Director, AI Engineer - Remote · Novartis $176,400/yr – $327,600/yr
Peer group range $105,900 – $252,000 (median $184,800)

Where these roles are based

Top locations among the 92 comparable roles

India47
United States23
United Kingdom5
Japan3
Ireland3
Spain2

+ 7 more countries

Seniority mix

92 of 92 peers have a known seniority level

Senior56
Associate Director20
Principal16

Therapeutic area mix

2 of 92 peers have a known therapeutic area; the rest are genuinely unlabeled, not hidden

Immunology1
Oncology1

Similar opportunities

The closest matches from our peer group, ranked by how similar they are, not how well you'd qualify for them -- treat this as market context, not a guaranteed shortlist; a weak match is labeled as one below.

60%similar
Novartis Hyderabad (Office), India Senior
Same function Same seniority Same country
60%similar
Novartis Hyderabad (Office), India Senior
Same function Same seniority Same country
60%similar
Regeneron India Private Limited Hyderabad, India Senior
Same function Same seniority Same country
60%similar
Regeneron India Private Limited Hyderabad, India Senior
Same function Same seniority Same country
60%similar
Regeneron India Private Limited Hyderabad, India Senior
Same function Same seniority Same country
60%similar
Regeneron India Private Limited Hyderabad, India Senior
Same function Same seniority Same country

How we calculate "similar"

No black box, no LLM guesswork: a deterministic score built from four normalized attributes. Here's this role's own peer group at different match levels, so you can see the mechanism, not just the result.

Every comparison starts from the same 100-point budget: 25 for working in the same function, 40 for the same therapeutic area, 20 for the same or adjacent seniority, 15 for the same country. A dimension we can't confirm on both sides contributes nothing, never a guess, never a free pass.

60%
Senior Expert (Rapid Prototyping)
Novartis · Hyderabad (Office), India · Senior
Function Therapeutic area Seniority Country
60%
Senior Specialist–Full Stack Developer- Clinical Data Operation
Regeneron India Private Limited · Hyderabad, India · Senior
Function Therapeutic area Seniority Country
60%
Senior Data Engineer - Data & ML Engineering
Bristol-Myers Squibb Business Services India Private Limited · Hyderabad, India · Senior
Function Therapeutic area Seniority Country
60%
Senior Stats Programmer (VAX)
Sanofi Healthcare India Private Limited · Hyderabad, India · Senior
Function Therapeutic area Seniority Country
Unmatched or unknown dimensions score exactly the same: 0 points, never a partial guess. A role we know almost nothing about beyond its function bottoms out at 25%; it never inflates to 100% just because there's little to compare against. Seniority uses a defined ladder (Associate → Manager → Associate Director → Senior → Principal → Director → Senior Director → Executive/VP) so "Director" and "Senior Director" count as adjacent, but "Director" and "Executive/VP" do not.