At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted and respected for who …
Oliver Wyman – Platform / Infrastructure Engineer (m/f/d) – Quotient AI Specialist – Madrid
Madrid, Community of Madrid, Spain · Hybrid
Mid level$15B raised
Company:Oliver Wyman Description: Who we are Oliver Wyman is a global management consulting firm. We work with clients through their most defining moments, combining deep industry knowledge, specialized expertise, and th…
Skills: Python, Machine Learning, LLM, AI Agent Design, Software Engineering
At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted and respected for who …
Skills: Platform Engineering, Enterprise Printing Solutions, Microsoft 365, Azure, Power Platform
About Cigna Healthcare Cigna Healthcare is a global health service company dedicated to transforming healthcare. With roots in the U.S. and operations in over 30 countries, we serve more than 180 million customers and pa…
Oliver Wyman – Platform / Infrastructure Engineer (m/f/d) – Quotient AI Specialist – Madrid
Madrid, Community of Madrid, Spain · Hybrid
Senior$15B raised
Company:Oliver Wyman Description: Who we are Oliver Wyman is a global management consulting firm. We work with clients through their most defining moments, combining deep industry knowledge, specialized expertise, and th…
¡Comienza a desarrollar tu carrera profesional en Mapfre! ¿Estás list@ para hackear tu futuro (de forma legal, claro)? Dentro de nuestro Plan de Talento Joven #CreceConNosotros2026, te abrimos la puerta a varias becas en…
Skills: Cybersecurity, Python, Bash, Data Analysis, Ethical Hacking
3153 - Ingeniero/a Junior Desarrollador de Sistemas Autónomos para Monodon
Madrid, Community of Madrid, Spain · On-site
$30k–$35k/yr
Entry level
La empresa Navantia S.A., S.M.E. realiza la convocatoria indicada para su centro de Madrid. El plazo de inscripción de candidaturas finalizará el día 15/09/2026 a las 12:00 horas. En el siguiente enlace: Bases personal f…
3600 - Ingeniero/a Junior en Automatización Inteligente
Madrid, Community of Madrid, Spain · On-site
$30k–$35k/yr
Entry level
La empresa Navantia S.A., S.M.E. realiza la convocatoria indicada para su centro de Madrid. El plazo de inscripción de candidaturas finalizará el día 15/09/2026 a las 12:00 horas. En el siguiente enlace: Bases personal f…
Skills: RPA, UiPath, Power Platform, Python, Power Automate
La empresa Navantia S.A., S.M.E. realiza la convocatoria indicada para su centro de Madrid. El plazo de inscripción de candidaturas finalizará el día 15/09/2026 a las 12:00 horas. En el siguiente enlace: Bases personal f…
Skills: SAP HCM, SuccessFactors, Technical Support, Functional Support, System Maintenance
3382 - Ingeniero/a Especialista en Gestión de Vulnerabilidades
Madrid, Community of Madrid, Spain · On-site
Senior
La empresa Navantia S.A., S.M.E. realiza la convocatoria indicada para su centro de Madrid. El plazo de inscripción de candidaturas finalizará el día 15/09/2026 a las 12:00 horas. En el siguiente enlace: Bases personal f…
3732 - Ingeniero/a Junior IT en Herramientas de Ingeniería
Madrid, Community of Madrid, Spain · On-site
$30k–$35k/yr
Entry level
La empresa Navantia S.A., S.M.E. realiza la convocatoria indicada para su centro de Madrid. El plazo de inscripción de candidaturas finalizará el día 15/09/2026 a las 12:00 horas. En el siguiente enlace: Bases personal f…
Skills: CAD, PLM, NX, Teamcenter, Active Workspace
Grupo de Seguros El Corte Inglés
Ingeniero/a de Inteligencia Artificial Generativa
Madrid, Community of Madrid, Spain · Hybrid
Mid level
¿Buscas desarrollar tu talento y un lugar dónde tu esfuerzo sea reconocido? Te ofrecemos un entorno dinámico, cercano y de futuro. Aquí, cada día es una nueva oportunidad para aprender, crecer y brillar. Imagina formar p…
Skills: Generative AI, Microsoft Copilot, OpenAI, Claude, Process automation
Accenture Song es la unidad de Accenture que combina creatividad, datos, tecnología y diseño para crear experiencias de cliente relevantes y escalables. Dentro de este contexto, nos gustaría crecer en el equipo con un/a …
Más sobre esta oportunidad ¡Dale un cambio a tu vida y sigue desarrollando tu carrera profesional en Mapfre en nuestras Áreas Corporativas! Formarás parte de un entorno internacional y que se encuentra muy cerca de la es…
Skills: IT Auditing, Cybersecurity, Data Analysis, Artificial Intelligence, Cloud Computing
3791 - 2 vacantes de Ingeniero/a Especialista en operaciones de ciberdefensa/SOC
Madrid, Community of Madrid, Spain · On-site
Mid level
La empresa Navantia S.A., S.M.E. realiza la convocatoria indicada para su centro de Madrid. El plazo de inscripción de candidaturas finalizará el día 15/09/2026 a las 12:00 horas. En el siguiente enlace: Bases personal f…
Oliver Wyman – Front-end / AI Content Creator (m/f/d) – Quotient AI Specialist – Madrid
Madrid, Community of Madrid, Spain · Hybrid
Mid level$15B raised
Company:Oliver Wyman Description: Who we are Oliver Wyman is a global management consulting firm. We work with clients through their most defining moments, combining deep industry knowledge, specialized expertise, and th…
Skills: AI content creation, Front-end development, Video production, Audio production, Visual storytelling
Capgemini Engineering, líder mundial en servicios de ingeniería, reúne a equipos de ingenieros/as, científicos/as y arquitectos/as para ayudar a las empresas más innovadoras del mundo a liberar su potencial y contribuir …
Skills: Generative AI, Data Science, Machine Learning, LLMs, RAG architectures
Oliver Wyman – (Senior) Technical Business Partner (m/f/d) – Quotient AI Specialist – Madrid
Madrid, Community of Madrid, Spain · Hybrid
Senior+$15B raised
Company:Oliver Wyman Description: Who we are Oliver Wyman is a global management consulting firm. We work with clients through their most defining moments, combining deep industry knowledge, specialized expertise, and th…
Skills: Artificial Intelligence, Technical Architecture, Business Development, Consulting, Solution Design
At Accenture, we believe in technology as the engine for the total reinvention of the company. We work with the leading platforms and partners in the market to drive our clients’ businesses through digitalization, AI, an…
Skills: IBM System I, AS/400, COBOL 400, System administration, Technical support
En VISEO buscamos ampliar nuestro Equipo de profesionales Microsoft. Si te gustan los nuevos retos, y quieres progresar en tu carrera profesional con un Equipo de expertos con tecnologías a la última, únete al equipo VIS…
Skills: React, Node.js, TypeScript, Full stack development, Unit testing
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
Full-time
bachelor degree, postgraduate degree
Posted 26d ago
~40 hrs/week
Responsibilities
You will oversee and advance the workload orchestration tech stack for HPC and AI platforms, ensuring efficient scheduling and resource utilization. This involves deploying and maintaining orchestration tools like SLURM and Run:ai while bridging the gap between compute capacity and scientific execution.
Requirements
Candidates must have at least 5 years of systems engineering experience with a focus on workload scheduling and cluster optimization. A degree in a technical discipline and deep familiarity with Linux, distributed systems, and containerized GPU scheduling are required.
Full job description
At Roche you can show up as yourself, embraced for the unique qualities you bring. Our culture encourages personal expression, open dialogue, and genuine connections, where you are valued, accepted and respected for who you are, allowing you to thrive both personally and professionally. This is how we aim to prevent, stop and cure diseases and ensure everyone has access to healthcare today and for generations to come. Join Roche, where every voice matters.
The Position
Job description
As a Workload Orchestration Engineer within the Accelerated Compute Engineering (ACE) team, you will be responsible for overseeing and advancing our workload orchestration tech stack across both our High-Performance Computing (HPC) and industry-leading AI Factory platforms. With the rapid expansion of our compute infrastructure, efficiently scheduling, managing, and maximizing the utilization of our CPU and GPU environments is paramount.
You will own the deployment, configuration, and fine-tuning of orchestration platforms that schedule massive, parallel computational workloads. By implementing robust scheduling policies for traditional scientific workflows and modern containerized AI workloads, you will bridge the gap between heavy compute capacity and efficient execution. Your work will directly ensure that Roche’s researchers, data scientists, and engineers can seamlessly run large-scale AI model training and computational science simulations at scale.
Description of the area
Hosting and Infrastructure (HI) provides mission-critical on-premise infrastructure, cloud hosting, connectivity, and technology products that enable all functions at every Roche site to develop, innovate, connect, and deliver compliant digital products across the Roche Enterprise.
The Value Streams - Accelerated Compute Engineering (ACE) Team is focused on driving both customer success and platform success by acting as a center of excellence and delivery for the High Performance Compute and AI Infrastructure supporting AI and HPC use cases across Roche. This team facilitates seamless onboarding and adoption for business vertical customers needing accelerated compute—helping those infrastructure consumers with needs optimized for high availability, seamless data transfer, flexibility, speed, and the rapidly changing needs of AI—helping achieve rapid time-to-value.
Job Responsibilities
Orchestration Stack Deployment & Governance
Design, implement, and maintain the SLURM Workload Manager ecosystem across our HPC cluster architectures, ensuring high availability and optimal resource distribution.
Deploy and manage Run:ai as the core orchestration and virtualization layer for the AI Factory, enabling fractional GPU allocation and dynamic resource allocation.
Evaluate, architect, and implement SLURM Slinky integrations where required to seamlessly bridge Kubernetes-based AI orchestration with traditional HPC cluster resources.
Containerization & Workload Optimization
Define best practices and frameworks for containerized scientific execution, utilizing Singularity/Apptainer and/or Enroot to provide secure, reproducible performance environments for HPC.
Translate user and workload requirements into optimized scheduling parameters (e.g., topology-aware scheduling, multi-node scaling).
Actively profile and tune scheduling queues, quality-of-service (QoS) parameters, and fair-share policies to maximize multi-tenant efficiency.
Platform Reliability & Telemetry
Partner with Observability Engineers to implement continuous monitoring, telemetry, and reporting dashboards to track scheduler efficiency, queue wait times, and hardware utilization rates.
Troubleshoot complex workload failures, including distributed training synchronization issues, MPI communication bottlenecks, and driver incompatibilities.
Maintain configuration-as-code models for the scheduling tier, leveraging automation to deploy cluster policies uniformly.
Qualifications
Education / Experience
Bachelor’s or an advanced degree in Computer Science, Applied Mathematics, Computational Engineering, or a similar technical discipline.
5+ years of systems engineering experience, with a heavy emphasis on workload scheduling, resource management, and cluster optimization for multi-tenant environments.
Deep technical familiarity with Enterprise Linux operating systems and distributed systems architecture.
HPC Scheduling & Tooling: Expert-level proficiency in administering SLURM, including complex partition designs, accounting, and plug-in management. Highly proficient with Singularity for container runtime execution.
AI Orchestration: Hands-on experience or deep architectural understanding of Run:ai, Kubernetes, and containerized GPU scheduling paradigms.
Infrastructure Literacy: Solid understanding of high-speed interconnects (InfiniBand, RoCE) and multi-node communication architectures (MPI, NCCL) as they relate to job placement.
Automation: Proficiency in automating scheduler configurations and telemetry gathering, or infrastructure automation tooling.
Leadership & Mindset:
Lean & Agile Mindset: Highly focused on driving efficiency, reducing idle compute time, and creating frictionless pathways for user workload submissions.
Collaboration & Advocacy: Outstanding capability to translate scientific and AI model workflow challenges into scalable scheduler configurations.
Intellectual Curiosity: A strong passion for remaining ahead of industry trends regarding GPU slicing, fractionalization, and the convergence of AI workloads with traditional HPC schedulers.
Who we are
A healthier future drives us to innovate. Together, more than 100’000 employees across the globe are dedicated to advance science, ensuring everyone has access to healthcare today and for generations to come. Our efforts result in more than 26 million people treated with our medicines and over 30 billion tests conducted using our Diagnostics products. We empower each other to explore new possibilities, foster creativity, and keep our ambitions high, so we can deliver life-changing healthcare solutions that make a global impact.
Roche is a global pioneer in pharmaceuticals and diagnostics focused on advancing science to improve people’s lives. The combined strengths of pharmaceuticals and diagnostics under one roof have made Roche the leader in personalised healthcare – a strategy that aims to fit the right treatment to each patient in the best way possible.
Roche is the world’s largest biotech company, with truly differentiated medicines in oncology, immunology, infectious diseases, ophthalmology and diseases of the central nervous system. Roche is also the world leader in in vitro diagnostics and tissue-based cancer diagnostics, and a frontrunner in diabetes management.
Founded in 1896, Roche continues to search for better ways to prevent, diagnose and treat diseases and make a sustainable contribution to society. The company also aims to improve patient access to medical innovations by working with all relevant stakeholders. Thirty medicines developed by Roche are included in the World Health Organization Model Lists of Essential Medicines, among them life-saving antibiotics, antimalarials and cancer medicines. Roche has been recognised as the Group Leader in sustainability within the Pharmaceuticals, Biotechnology & Life Sciences Industry ten years in a row by the Dow Jones Sustainability Indices (DJSI).
For more information, please visit https://careers.roche.com
Read our community guidelines here:
https://www.roche.com/some-guidelines.htm
#Roche #Biotechnology #Pharmaceuticals #Diagnostics #Healthcare #PersonalisedHealthcare #GreatPlaceToWork #Innovation
Offices: Grenzacherstrasse, Switzerland 🇨🇭 , 4070, CH · 4414 Lake Boone Trail, Raleigh, NC 27607, US · 7070 Mississauga Rd, Mississauga, ON L5N 5M8, CA · Calle Dionisio Derteano 144, San Isidro, Lima 15047, PE · 333 Buckhead Ave NE, Atlanta, GA 30305, US
biotechnologyinnovationpersonalized healthcaregreat place to workdiagnosticspharmaceuticalsresearchhealthcarepersonalised healthcarediabetes care
How many Software jobs are open in Madrid, Spain right now?
There are currently 2,597 open software positions in Madrid, Spain listed on Clera. New openings are added daily as companies post roles.
Which companies are hiring for Software roles in Madrid, Spain?
Companies currently hiring include Grupo TECDATA Engineering, knowmad mood, Capgemini, Indra Group USA, Telefónica Tech, among others. Browse the listings above to see every active employer.
Are there remote or hybrid Software jobs in Madrid, Spain?
Yes — 1964 of the 2597 open software positions offer remote or hybrid work (537 remote, 1427 hybrid).
How do I apply for Software jobs in Madrid, Spain?
Each listing links directly to the employer's application page. Apply early — fresh listings get the most recruiter attention in the first two weeks.