Roche Posted September 29, 2026

Kubernetes Reliability Engineer

Mississauga, Ontario, Canada FULL_TIME
Notify me about similar jobs

Roche is the source of truth for this posting and owns the application process. We surface normalized context and market comparison you won't find on the original listing.

About this opportunity

At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections,  where you are valued, accepted and respected for who you are, allowing you to thrive both personally and professionally. This is how we aim to prevent, stop and cure diseases and ensure everyone has access to healthcare today and for generations to come. Join Roche, where every voice matters.

The Position

Kubernetes Reliability Engineer 

A healthier future. It’s what drives us to innovate. To continuously advance science and ensure everyone has access to the healthcare they need today and for generations to come. Creating a world where we all have more time with the people we love. That’s what makes us Roche.

The CaaS IT Infrastructure Engineer is a highly skilled expert responsible for solving complex business problems using advanced cloud native technologies. The engineer will build and maintain a Kubernetes-based infrastructure, enabling the modernization of business applications and processes.

This role combines software and systems engineering to optimize systems, increase efficiency, and eliminate operational work through automation.

The Opportunity:

You will be part of the global CaaS infrastructure team at a leading healthcare company, working with members across different regions. The team's mandate is to deliver, maintain, and continuously improve a highly available Kubernetes platform across hybrid cloud deployments, including on-premise data centers and public clouds like AWS. In this role, you will apply software engineering principles to operations to build and run massively distributed, fault-tolerant systems, focusing heavily on automation, security, and observability.

Scope: Engages in and improves autonomously the whole lifecycle of platforms and services, from inception and design through deployment, operation, and retirement. Applies software engineering principles to build and manage large-scale IT infrastructure products, abstracting away complexity by providing self-service tools and APIs for developers. Designs, implements, and maintains CI/CD pipelines and develops self-healing features.

Service Reliability and Optimization: Focus on capacity planning and launch reviews for services before they go live. Perform blameless postmortems and proactive identification of potential outages to foster iterative improvements

Accountability/Problem Solving: Resolves complex problems in a global Kubernetes-based infrastructure through in-depth evaluation of variable factors, including inter-organizational impact, balanced with effective consultative engagement of key stakeholders. Leads end-to-end design of infrastructure solutions and maintains component standards. Evaluates promising solutions via Proof of Concept (PoCs) and feasibility studies across multiple areas, and serves as an internal escalation point for major incidents

Stakeholder Management: Acts as a bridge between engineering and operations. Communicates and presents complex information and potential solutions to cross-functional teams and the business in non-technical terms. Represents the organization as a prime contact on initiatives and interacts with senior internal and external personnel. Uses deep knowledge to influence IT infrastructure vendor product evaluations and collaborates with multiple IT partners (e.g. Enterprise Architects, Solution Owners) to integrate feedback. Mentors and shares DevOps culture, guiding developers on how to create and deploy cloud-native applications

Impact/Strategy: Provides technical leadership and direction for small-to-medium sized initiatives (projects, lifecycle work, PoCs). Ensures solutions comply with Quality/Regulatory standards and that designs adhere to the organization’s Technical Architecture Framework (TAF) policies and directions. Assists in planning technology projects, estimating engineering resources, dependencies, risks and timelines for successful delivery

Complexity: Demonstrates the ability to lead geo-distributed initiatives across different locations and cultural backgrounds through influence and mentorship, providing specialized guidance to drive success and cohesion.

Business/Technical ability: Applies extensive cloud native technical expertise, acting as a recognized expert in Kubernetes and maintaining in-depth knowledge across related cloud native technologies (containers, AWS, etc.). Demonstrates a detailed understanding of how IT infrastructure impacts respective Roche business processes and outcomes.

Who you are : 

Education / Experience

Bachelor’s degree in Computer Science, Mathematics, Physics or related field, and 2-5 years of relevant experience.

Without degree: 4-7 years of relevant experience.

Master’s degree: 1-3 years of relevant experience.

Technical Skills

Kubernetes & Containers: Strong hands-on experience navigating, managing, and hardening Kubernetes clusters and containers, including knowledge of distributed storage. A Certified Kubernetes Administrator (CKA) certification is a strong plus. Knowledge of tools like Rancher or Portworx is beneficial

Infrastructure as Code (IaC): Hands-on experience delivering and managing infrastructure automation using tools like Ansible and Terraform

Scripting & Software Engineering: Proficiency in scripting and programming languages, primarily Python, Bash, or Go, including experience with test automation (e.g., pytest) and APIs deployment and management

CI/CD Tools: Expert knowledge of implementing software delivery pipelines using tools (e.g., Jenkins, Rundeck, or GitLab)

Systems & Networking: Strong understanding of Linux operating systems and core networking principles, including DNS, load balancing, firewalls, routing, and service meshes.

Observability: Experience configuring logging, metrics, and monitoring tools, specifically focusing on setting up alerts based on symptoms rather than waiting for system outages

Cloud Infrastructure: Experience with public cloud platforms, with a strong preference for AWS, specifically involving managed services for compute, networking, security, and identity (e.g., EKS, VPC, IAM).

General and Operational Knowledge

You have a proven experience applying best practices in an always-up, always-available service environment utilizing Scrum and Agile methodologies

You demonstrate a deep understanding of Technical Architecture Frameworks (TAF) and Quality/Regulatory compliance standards

Additional Qualifications

You have excellent problem-solving skills, decision-making ability, and sound judgment

You demonstrated a strong team-oriented mindset with the ability to function independently with low supervision and navigate ambiguity.

You are highly fluent in oral and written English communication skills are required.

You have the ability to work across multiple time zones.

You demonstrate strong customer & delivery focus with the ability to act as an analyst, seamlessly transforming complex stakeholder needs into actionable technical requirements.

You possess strong practice of sustainable incident response, including managing ITSM processes and leading audit evidence collection.

Relocation benefits are not available for this position.

The expected salary range for this position based on the primary location of Mississauga is 105,560.00 and 138,547.50 of hiring range. Actual pay will be determined based on experience, qualifications, and other job-related factors as determined by the company.

We use artificial intelligence to screen, assess or select applicants for this role.

This posting is for an existing vacancy at Hoffmann-La Roche Ltd.

Who we are

A healthier future drives us to innovate. Together, more than 100’000 employees across the globe are dedicated to advance science, ensuring everyone has access to healthcare today and for generations to come. Our efforts result in more than 26 million people treated with our medicines and over 30 billion tests conducted using our Diagnostics products. We empower each other to explore new possibilities, foster creativity, and keep our ambitions high, so we can deliver life-changing healthcare solutions that make a global impact.

Let’s build a healthier future, together.

Roche is an Equal Opportunity Employer.

Job details

Seniority
Not listed
Function
Information Technology
Therapeutic area
Not listed
Location
Mississauga, Ontario, Canada
Employment type
FULL_TIME

How this role compares

Computed from every other active Information Technology role in our database, not just this employer's listings.

We currently track 1167 comparable Information Technology roles across 74 biopharma companies.

1167Comparable roles tracked
1083Currently active
74Companies hiring similar roles
35Countries represented

Salary context

162 of 1167 peers report a salary range (USD, annualized)

Peers share this role's job function. This posting doesn't list a seniority level, so peers aren't narrowed by seniority either -- the range below may span more levels than usual.

This roleSubject Not listed on this posting
Lowest disclosed · 2027 Business Technology Solutions Intern - Cybersecurity (Undergraduate) · AbbVie $21/hr – $37/hr (≈ $43,680–$76,960/yr)
Peer group range $60,320 – $400,000 (median $185,000)

Where these roles are based

Top locations among the 1167 comparable roles

India488
United States310
Spain90
Poland71
Portugal22
China17

+ 29 more countries

Seniority mix

667 of 1167 peers have a known seniority level

Senior272
Manager143
Principal67
Associate Director55
Associate49
Director41
Intern/Fellow/Postdoc21
Senior Director13
Executive/VP6

Therapeutic area mix

2 of 1167 peers have a known therapeutic area; the rest are genuinely unlabeled, not hidden

Oncology1
Vaccines & Infectious Disease1

Similar opportunities

The closest matches from our peer group, ranked by how similar they are, not how well you'd qualify for them -- treat this as market context, not a guaranteed shortlist; a weak match is labeled as one below.

40%similar
ID Biomedical Corporation of Quebec Ste Foy, Canada
Same function Same country
40%similar
Amgen Canada Inc. Remote, Canada
Same function Same country
40%similar
Amgen Canada Inc. Burnaby, Canada Senior
Same function Same country
40%similar
Roche Laval, Quebec, Canada
Same function Same country
40%similar
Roche Ontario, Ontario, Canada Manager
Same function Same country
40%similar
Roche Mississauga, Ontario, Canada Principal
Same function Same country

Notify me about similar jobs

Get an email when we spot other openings like this one – same job function, comparable seniority, roles you'd actually want to see.

How we calculate "similar"

No black box, no LLM guesswork: a deterministic score built from four normalized attributes. Here's this role's own peer group at different match levels, so you can see the mechanism, not just the result.

Every comparison starts from the same 100-point budget: 25 for working in the same function, 40 for the same therapeutic area, 20 for the same or adjacent seniority, 15 for the same country. A dimension we can't confirm on both sides contributes nothing, never a guess, never a free pass.

40%
Directeur(rice), Sciences et technologies de la fabrication (MSAT)
ID Biomedical Corporation of Quebec · Ste Foy, Canada · Seniority not listed
Function Therapeutic area Seniority Country
40%
RDT Business Partner
Roche · Laval, Quebec, Canada · Seniority not listed
Function Therapeutic area Seniority Country
40%
PT Learning Technology & Adoption Lead
Roche · Mississauga, Ontario, Canada · Seniority not listed
Function Therapeutic area Seniority Country
40%
Chef Technologie et Procédé
ID Biomedical Corporation of Quebec · Quebec, Canada · Seniority not listed
Function Therapeutic area Seniority Country
Unmatched or unknown dimensions score exactly the same: 0 points, never a partial guess. A role we know almost nothing about beyond its function bottoms out at 25%; it never inflates to 100% just because there's little to compare against. Seniority uses a defined ladder (Associate → Manager → Associate Director → Senior → Principal → Director → Senior Director → Executive/VP) so "Director" and "Senior Director" count as adjacent, but "Director" and "Executive/VP" do not.