Volver a los resultados · Barcelona

Oferta verificada hace 17 horas

Senior AI Operations Engineer

AstraZeneca·Barcelona
Salario no indicado
Resumen Vox
  • Rol principal: Lider técnico en operaciones de IA, responsable de diseñar, construir y mantener plataformas y herramientas en Azure y AWS, promoviendo la automatización y la mejora continua.
  • Requisitos clave: Experiencia en ingeniería de fiabilidad del sitio, diseño de sistemas de observabilidad, automatización, respuesta a incidentes y liderazgo técnico en entornos de nube.
  • Condiciones y beneficios: Posición senior con enfoque en ingeniería, mentoría a equipos, colaboración con diferentes funciones, y participación en la transformación digital de la empresa.
  • Habilidades técnicas: Diseñar y mantener paneles, alertas, telemetría, y sistemas de análisis impulsados por IA, contribuyendo a la infraestructura y productos de ingeniería.
Candidatarse en el origenSales de VoxJobs hacia tecnoempleo.com — la candidatura se hace directamente en la empresa. tecnoempleo

Descripción de la oferta

Do you have expertise in and passion for AI-powered operational excellence Would you like to apply your expertise to impact a company that follows the science and turns ideas into life changing medicines If so AstraZeneca might be the one for you ABOUT ASTRAZENECA AstraZeneca is a global innovation-driven BioPharmaceutical business that focuses on the discovery development and commercialisation of prescription medicines for some of the worlds most serious diseases. But were more than one of the worlds leading pharmaceutical companies. At AstraZeneca were dedicated to being a Great Place to Work. Where you are empowered to push the boundaries of science and unleash your entrepreneurial spirit. Theres no better place to make a difference to medicine patients and society. An inclusive culture that champions diversity and collaboration. Always committed to lifelong learning growth and development. ABOUT OUR ENTERPRISE AI PLATFORMS SERVICES TEAM A place to do important work. We connect across the whole business to power each function to better influence patient outcomes and improve their lives. Impactful and valuable this is where you come to raise your profile and do good for others. Play an increasingly crucial role in driving disruptive transformation on our journey to becoming a digital and data-led enterprise. Unleash the power of our latest innovations in data machine learning and technology to turn complex information into life-changing and practical insights. Work in synergy with leading experts in our specialist communities. Here we bring the brightest minds to bear with access to cutting-edge techniques and the opportunity to be part of novel solutions. Its up to us to drive the outcomes forward inventing and building expanding our knowledge to identify the next opportunity. An inclusive team we bring together diverse areas - different functions as well as external partners. Pooling from an unrivalled source of knowledge we share learn and challenge. It powers us to decode business needs and apply our technical know-how to add greater value. Rise to the challenge of shaping the future of an evolving business in the technology space. ABOUT THE ROLE The Enterprise AI Platforms Services Team are responsible for building and running the platforms tooling and infrastructure that powers AstraZenecas ambition to use AI in every step of the value chain from discovering new compounds to patient safety systems. We are looking for a Senior AI Operations Engineer to be a hands-on technical leader within our AI/ML platform operations function. Reporting to the AI for Operations Lead you will be the primary executor and technical driver across our growing platform estate built on Azure and AWS. The ideal candidate will have deep current experience in site reliability engineering and will combine strong individual contribution with mentoring and technical leadership of L1 and L2 engineers. You will operate within a three-tier operations structure (L1 Runbook Operators L2 Site Reliability Engineers L3 Product Engineering interface) and be the senior hands-on practitioner who builds the automation designs the observability and drives the continuous improvement flywheel day-to-day. You will be a key contributor to AI-augmented operations - building and refining AI-powered tools for runbook querying incident pattern analysis and automated diagnosis. This is not a traditional support engineer role. This is a senior technical position for someone who sees operations as an engineering discipline and who can design build and instrument systems to the highest standard while coaching others to do the same. KEY ACCOUNTABILITIES Technical Execution and Design - Design build and maintain the centralised observability layer - instrumentation dashboards alerting rules and telemetry pipelines using platforms such as Datadog New Relic Grafana or Splunk - Architect and implement AI-augmented operations tooling conversational runbook interfaces AI-driven incident analysis and pattern recognition systems - Lead complex incident response perform root-cause analysis and produce actionable post-mortems - Contribute patches and instrumentation to product engineering codebases ensuring they are architecturally sound and address root causes - Contribute to vendor evaluation to assess and recommend appropriate technologies and automation tooling. Automation and Continuous Improvement - Build and maintain the automation that powers the continuous improvement cycle every incident results in a runbook an automation or a patch - Own the technical execution of automation investments prioritised by frequency resolution time and blast radius - Proactively identify and eliminate toil maintaining the standard of no more than 50 of SRE time on reactive work - Develop and maintain runbooks ensuring they are accurate current and progressively automated Operational Readiness and Platform Onboarding - Implement the Operational Readiness Gate - assess platforms against defined criteria before they transition from product engineering to operations ownership - Shape platform operability during development - contributing instrumentation health checks and observability hooks before platforms enter BAU - Provide technical input to platform architecture reviews from an operational perspective Mentoring and Team Contribution - Mentor and coach L1 operators and junior L2 SREs building their technical capability and engineering mindset - Identify operators with engineering aptitude and support their development pathway from L1 to L2 - Contribute to the operations review by preparing flywheel metrics L1 resolution rate repeat incident rate automation coverage and toil budget compliance - Foster a culture where SREs are engineers whose product is operational excellence Stakeholder Collaboration - Partner with product engineering teams on post-mortems handover assessments and operational obligation delivery - Represent operational requirements in technical design discussions - Communicate incident findings operational risks and improvement proposals clearly to the Operations Lead and wider team CANDIDATE KNOWLEDGE SKILLS AND EXPERIENCE Essential - BSc/MSc degree in Computer Science or related quantitative or analytical field - Significant hands-on experience as a Site Reliability Engineer or platform operations engineer at scale - you build and run systems not just direct others - Strong expertise in observability platforms (Datadog New Relic Grafana Splunk or equivalent) including dashboard design alerting strategies and telemetry pipeline implementation - Strong working knowledge of OpenTelemetry distributed tracing and structured logging standards - Proven track record of designing and implementing automation that materially reduces operational toil - strong scripting skills in Python and Bash with the ability to build robust tooling beyond one-off scripts - Experience assessing platform readiness and contributing to operational handover processes - Strong hands-on skills with cloud infrastructure (Azure and/or AWS) including container orchestration serverless architectures and managed services - Experience running or contributing significantly to post-mortem processes and translating findings into preventive engineering work - Ability to implement precise technical solutions - you can instrument a system configure meaningful alerts and build automation that intervenes at the right point - Experience mentoring junior engineers and contributing to team development Desirable - Experience applying AI/ML to operational challenges - intelligent alerting automated diagnosis predictive incident detection or conversational operations interfaces - Familiarity with the AstraZeneca technology estate or regulated pharmaceutical environments - Experience operating platforms that serve AI/ML workloads (LLM inference model serving data pipelines) - ITIL SRE or operational excellence certifications or equivalent practical frameworks - Experience working within multi-tier support structures with clear escalation paths - Infrastructure-as-code expertise (Terraform CloudFormation) - Awareness to GxP and audit trail awareness. Personal Qualities - An engineering mindset applied to operations - you see every repeated manual task as an automation opportunity - Comfort operating at the boundary between deep technical work and collaborative team contribution - Creative collaborative and resilient - Strong communicator who can explain technical complexity clearly to both engineers and leadership - A bias toward systems thinking - you address root causes not symptoms - Self-directed with the ability to manage competing priorities across multiple platforms When we put unexpected teams in the same room we unleash bold thinking with the power to inspire life-changing medicines. In-person working gives us the platform we need to connect work at pace and challenge perceptions. Thats why we work on average a minimum of three days per week from the office. But that doesnt mean were not flexible. We balance the expectation of being in the office while respecting individual flexibility. Join us in our unique and ambitious world.

Panel de transparencia

Fuente original
tecnoempleo
Publicada
20 jul 2026 · fecha real
Última verificación
hace 17 horas
Puntuación de calidad
35/100
Salario indicado0
Empresa identificada0
applyUrl0
postedAt15
Descripción completa20

Semejantes

Ofertas cercanas a esta.

¿Algo mal con este anuncio? Denunciar oferta fraudulenta o desactualizada