Description
At Roche, you can be authentically yourself and will be valued for your unique qualities. Our culture encourages personal expression, open dialogue, and genuine connections. Here, you are appreciated, accepted, and respected for who you are—creating an environment where you can grow both personally and professionally. Together, we aim to prevent, stop, and cure diseases and ensure that everyone has access to healthcare—today and in the future. Join Roche, where every voice matters.
The Position
The DevOps Infrastructure Engineer takes ownership and drives strategic technical initiatives, focusing on building, optimizing, and maintaining end\-to\-end lifecycle processes and developed orchestrations. In this role, you will challenge the status quo, foster collaboration, and mentor junior team members while focusing on making end\-to\-end processes more robust, resilient, and effortless for internal customers to consume. You will lead the development of custom operational tooling and workflow orchestrations (using engines like Temporal.io), manage infrastructure as code (IaC) for large workloads, and ensure robust security, monitoring, and logging practices are in place to maintain optimal system health and performance.
**Job Responsibilities:**
* Scope / (Content Leadership): Takes ownership of ambiguous topics ("grey zones") and successfully drives small to medium initiatives (Project/Squad/Product). Builds, hardens, and maintains end\-to\-end lifecycle processes and developed orchestrations. Develops custom internal tooling and self\-service abstractions to ease service consumption for customers, manages infrastructure using IaC principles, and implements robust security and monitoring practices.
* Accountability/Problem Solving: Performs complex troubleshooting across distributed environments, coordinates across multiple teams, and takes ownership of decisions that involve risk. Focuses on making end\-to\-end processes more robust, identifying process gaps at the company level, and bringing in outside best practices to eliminate operational friction and failure points.
* Stakeholder Management: Prioritizes and communicates effectively with stakeholders. Maintains a "bigger picture" view to thoroughly understand how various technical components fit into the overall product strategy, developer experience, and organizational goals.
* Impact/Strategy: Acts as a technical mentor to junior engineers and actively shares knowledge in internal and external events (e.g., DevOps Days, Tech Talks). Aligns personal and team development plans with the global strategy and participates in initiatives reaching outside the regular project scope.
* Complexity / (Product Size): Operates at the Product or Site level, managing complex infrastructures, workflow orchestrations, CI/CD pipelines, and scripts for large workloads and products in Production environments.
* Business / Technical ability: Possesses a deep understanding of the broader technical landscape beyond specific components. Formulates and implements technical strategies aligned with organizational objectives, investing time in initiatives that boost creativity, ease of service consumption, and innovation while maintaining the highest engineering quality.
**Qualifications:**
**Education / Experience:**
* Demonstrated experience in Production environments managing large workloads, complex workflow orchestrations, and end\-to\-end platform lifecycles.
* Proven track record of mentoring junior engineers, fostering good collaboration, and skillfully navigating group dynamics.
* Experience identifying gaps in Agile and DevOps principles at the company level and successfully implementing external best practices to harden operational workflows.
**Technical Skills:**
* Knowledge of stateful workflow orchestration engines, with direct experience or strong familiarity with Temporal.io (or similar systems like Cadence).
* Advanced knowledge of at least one programming language (Python, Go, Ruby, Rust, JavaScript) used for custom tooling and workflow SDKs, alongside advanced terminal skills (Bash/PowerShell scripting, networking tools, process monitoring).
* High proficiency in Infrastructure as Code (Terraform, Ansible, CDK), advanced Docker containerization, and Version Control (Git) with the ability to troubleshoot complex issues and build reliable automation pipelines.
* Expertise in CI/CD platforms (Gitlab CI, Jenkins, or Github Actions) and advanced monitoring, logging, metrics, and tracing tools (Datadog, Elastic Stack, Grafana, Jaeger).
* Knowledge of cloud platforms (AWS, Azure, GCP) and Linux operating systems (Ubuntu/Debian, SUSE, RHEL) is considered a strong plus (nice to have).
**Additional Qualifications:**
* Strong communication and stakeholder management skills, with a developer\-first mindset focused on simplifying how internal users consume platform services.
* Proactive mindset dedicated to continuous improvement, identifying technical opportunities to make systems more robust, and independently leading complex troubleshooting efforts across multiple teams.
Who We Are
A healthier future drives us to innovate. More than 100,000 employees worldwide work together to advance scientific progress and ensure that everyone has access to healthcare—today and for future generations. Through our commitment, over 26 million people are treated with our medicines and more than 30 billion tests are performed annually using our diagnostics products. We encourage each other to explore new possibilities, foster creativity, and set ambitious goals to deliver life-changing healthcare solutions.
Together, we can shape a healthier future.
Roche is an equal opportunity employer.