About Us We're living through a Cambrian explosion of intelligence: new models and new chips, each specialised for different tasks, are arriving all at once. The result is a new era for AI, one of radical heterogeneity. the company is the Intelligent Systems Company. We believe the next generation of AI won't be defined by any single model or chip, but by intelligent systems in which hardware and intelligence co-evolve. We are building the infrastructure that unifies heterogeneous compute across the full stack. This opens a new axis of scaling intelligence: a dynamic system that tailors itself to what each workload actually needs, whether that's speed, cost, precision, or whatever unit comes next. The last era scaled on a different bet: one bigger model, more of the same chip, more data. That bet is running into structural limits. Frontier models offer extraordinary capability at unsustainable cost, one that today's monolithic infrastructure was never designed to serve. Our founding principle is that intelligence comes from many specialised systems working together, not from any single component. We build the software orchestration layer that co-evolves models, workflows and silicon into one system, delivering inference tailored to every workload, and demonstrating orders-of-magnitude leaps in capability and cost. Because our software spans the full stack, our engineering team works directly with heterogeneous accelerators and frontier silicon, including Cerebras, d-Matrix, Intel, NVIDIA, AMD, Normal Computing, Tenstorrent, GreatSky, and Mixx. We are not stopping at today's chips: each new generation of silicon unlocks algorithms that couldn't run before, and we intend to be first to them, every time. If we get it right, it will belong to everyone building on it - not to any single vendor. In our latest funding round, we raised $100M, led by Atomico with participation from Plural, DCVC and the UK Sovereign AI Fund's first investment. With this, we are building the infrastructure for the next era of intelligence. We are engineers and scientists based in London, working across the full depth of the stack. We are curious, intellectually honest, and building what doesn't exist yet. If you thrive on uncharted territory and are energised by the scale of the challenge, we'd love to hear from you. About the Role the company believes that the next generation of AI infrastructure will be built from a much more diverse set of hardware than the clusters of today. Scaling that infrastructure creates a new physical systems problem: different accelerators bring different power densities, cooling requirements, rack constraints, network topologies, vendor dependencies, and deployment models. We want to find ways to turn those increasingly complex inputs into physical infrastructure that can be deployed repeatedly and scaled. This role owns the company's physical compute infrastructure, from the first assessment of a potential deployment through to the reliable operation of the resulting cluster. You will scope sites, coordinate the design cluster infrastructure and deployment across vendors and facilities, and lead commissioning of new hardware. As our footprint grows, you will build and lead the team responsible for keeping that infrastructure healthy, observable, and available, establishing the operational systems that turn individual deployments into a reliable fleet. What You'll Do Scope new cluster deployments across site selection, power, cooling, rack layout, networking, storage, fibre connectivity, and capacity requirements Coordinate delivery and commissioning across facilities, utilities, OEMs, networking and storage vendors, ensuring the physical and technical dependencies come together Build the operational infrastructure for the fleet, including hardware telemetry, health monitoring, maintenance, incident response, capacity planning, and hardware lifecycle management Build and lead the physical infrastructure team responsible for deploying new clusters and maintaining the reliability and availability of the resources we operate What Sets You Apart Experience designing, deploying, or operating high-density compute infrastructure, HPC clusters, AI infrastructure, or similarly complex physical systems Strong understanding of how power, cooling, rack design, networking, storage, and compute interact to determine the capabilities and constraints of a cluster Demonstrated ability to drive complex infrastructure deployments across technical teams, facilities, hardware vendors, network providers, and other external stakeholders Experience building and leading infrastructure teams, with strong operational instincts around reliability, observability, capacity, maintenance, and failure management What We Offer Competitive Salary, determined by skills and experience Equity & Ownership Private healthcare We offer Visa sponsorship and relocation benefits to hire the best in the world We work in person at our London office. You'll have the tools, space and setup to do your best work, and if you have specific needs, just tell us We're committed to building an inclusive workplace where everyone feels welcome, and believe in equal opportunities for all.
Aug 26, 2026
Full time
About Us We're living through a Cambrian explosion of intelligence: new models and new chips, each specialised for different tasks, are arriving all at once. The result is a new era for AI, one of radical heterogeneity. the company is the Intelligent Systems Company. We believe the next generation of AI won't be defined by any single model or chip, but by intelligent systems in which hardware and intelligence co-evolve. We are building the infrastructure that unifies heterogeneous compute across the full stack. This opens a new axis of scaling intelligence: a dynamic system that tailors itself to what each workload actually needs, whether that's speed, cost, precision, or whatever unit comes next. The last era scaled on a different bet: one bigger model, more of the same chip, more data. That bet is running into structural limits. Frontier models offer extraordinary capability at unsustainable cost, one that today's monolithic infrastructure was never designed to serve. Our founding principle is that intelligence comes from many specialised systems working together, not from any single component. We build the software orchestration layer that co-evolves models, workflows and silicon into one system, delivering inference tailored to every workload, and demonstrating orders-of-magnitude leaps in capability and cost. Because our software spans the full stack, our engineering team works directly with heterogeneous accelerators and frontier silicon, including Cerebras, d-Matrix, Intel, NVIDIA, AMD, Normal Computing, Tenstorrent, GreatSky, and Mixx. We are not stopping at today's chips: each new generation of silicon unlocks algorithms that couldn't run before, and we intend to be first to them, every time. If we get it right, it will belong to everyone building on it - not to any single vendor. In our latest funding round, we raised $100M, led by Atomico with participation from Plural, DCVC and the UK Sovereign AI Fund's first investment. With this, we are building the infrastructure for the next era of intelligence. We are engineers and scientists based in London, working across the full depth of the stack. We are curious, intellectually honest, and building what doesn't exist yet. If you thrive on uncharted territory and are energised by the scale of the challenge, we'd love to hear from you. About the Role the company believes that the next generation of AI infrastructure will be built from a much more diverse set of hardware than the clusters of today. Scaling that infrastructure creates a new physical systems problem: different accelerators bring different power densities, cooling requirements, rack constraints, network topologies, vendor dependencies, and deployment models. We want to find ways to turn those increasingly complex inputs into physical infrastructure that can be deployed repeatedly and scaled. This role owns the company's physical compute infrastructure, from the first assessment of a potential deployment through to the reliable operation of the resulting cluster. You will scope sites, coordinate the design cluster infrastructure and deployment across vendors and facilities, and lead commissioning of new hardware. As our footprint grows, you will build and lead the team responsible for keeping that infrastructure healthy, observable, and available, establishing the operational systems that turn individual deployments into a reliable fleet. What You'll Do Scope new cluster deployments across site selection, power, cooling, rack layout, networking, storage, fibre connectivity, and capacity requirements Coordinate delivery and commissioning across facilities, utilities, OEMs, networking and storage vendors, ensuring the physical and technical dependencies come together Build the operational infrastructure for the fleet, including hardware telemetry, health monitoring, maintenance, incident response, capacity planning, and hardware lifecycle management Build and lead the physical infrastructure team responsible for deploying new clusters and maintaining the reliability and availability of the resources we operate What Sets You Apart Experience designing, deploying, or operating high-density compute infrastructure, HPC clusters, AI infrastructure, or similarly complex physical systems Strong understanding of how power, cooling, rack design, networking, storage, and compute interact to determine the capabilities and constraints of a cluster Demonstrated ability to drive complex infrastructure deployments across technical teams, facilities, hardware vendors, network providers, and other external stakeholders Experience building and leading infrastructure teams, with strong operational instincts around reliability, observability, capacity, maintenance, and failure management What We Offer Competitive Salary, determined by skills and experience Equity & Ownership Private healthcare We offer Visa sponsorship and relocation benefits to hire the best in the world We work in person at our London office. You'll have the tools, space and setup to do your best work, and if you have specific needs, just tell us We're committed to building an inclusive workplace where everyone feels welcome, and believe in equal opportunities for all.
Job Overview: We are building a central AI control plane to support the safe, scalable, and efficient use of AI across Arm's engineering teams. This hands-on, technical leadership, platform engineering role will help direct and own delivery of the runtime platforms that make AI services reliable, secure, observable and supportable at Arm scale. You will work across Kubernetes, cloud, identity, secrets, networking, telemetry, incident management and automation to provide the production foundation for Arm's AI platform. Participate in production support and our paid on-call rota for high impact incident response. Production AI runtime platforms: Build/deploy and operate the infrastructure for centrally hosted AI platform services, including MCP server infrastructure, model gateway services and supporting control-plane components. Design runtime patterns for isolation, scalability, secure execution, capacity management and cost-aware operation. Automate provisioning, configuration, upgrades and lifecycle management using infrastructure-as-code and GitOps patterns. Reliability, observability and support: Define and implement service-level indicators, service-level objectives, alerting, dashboards, runbooks and support workflows. Incident response, post-incident review and vendor outage handling for AI services embedded in engineering workflows. Build telemetry that helps Arm understand AI platform health, usage, performance, cost and operational risk. Secure operations by default: Ensure platform components meet production readiness, security and compliance expectations. Help make secure AI usage the default by providing reliable paved paths rather than manual or fragmented infrastructure. Responsibilities: Technical leadership - this role will become a recognised expert and lead engineer on our AI platform's technology stack. Build, operate and continuously improve the production infrastructure for Arm's AI platform services, using both custom and third party solutions. Own reliability, scalability, monitoring, alerting, incident response, runbooks and operational readiness for AI platform components. Develop automation for provisioning, deployment, configuration, backup, recovery, patching, upgrades and lifecycle management. Implement secure runtime patterns, including workload isolation, secrets management, identity integration, network controls and auditability. Contribute towards platform roadmap planning and prioritisation. Required Skills and Experience: Expert level experience deploying and operating production infrastructure or platform services in a Linux-based environment using Kubernetes, containers, cloud or private-cloud platforms, infrastructure-as-code and CI/CD/GitOps tooling. Very strong understanding of security fundamentals for production platforms, including identity, secrets, access control, network segmentation, vulnerability management and audit logging. Familiarity with AI platform concepts such as model routing, MCP servers, agentic workflows, RAG systems or LLM observability. Very strong automation and scripting skills, for example Terraform, Go, Python or similar. Experience with incident management, problem management, demand forecasting, and production readiness practices. "Nice To Have" Skills and Experience: Experience operating internal developer platforms, AI platforms, model gateways, MCP infrastructure or other shared engineering platforms. Experience with service mesh, policy-as-code, workload identity, sandboxing, secure runtime environments or multi-tenant platform designs. Experience with regulated or security-sensitive engineering environments. Working knowledge of Open AI tools, products and APIs! In Return: Joining Arm means being part of an elite team that is high-reaching and determined to achieve outstanding results. You will collaborate with industry leaders and innovators, contributing to flawless service builds that improve our global operations. We provide an encouraging and inclusive environment that encourages continuous learning and professional growth. Your success is our success, and together we can build services that compete on the world stage! Accommodations at Arm At Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email . To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process. Hybrid Working at Arm Arm's approach to hybrid working is designed to create a working environment that supports both high performance and personal wellbeing. We believe in bringing people together face to face to enable us to work at pace, whilst recognizing the value of flexibility. Within that framework, we empower groups/teams to determine their own hybrid working patterns, depending on the work and the team's needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you. Equal Opportunities at Arm Arm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don't discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
Aug 25, 2026
Full time
Job Overview: We are building a central AI control plane to support the safe, scalable, and efficient use of AI across Arm's engineering teams. This hands-on, technical leadership, platform engineering role will help direct and own delivery of the runtime platforms that make AI services reliable, secure, observable and supportable at Arm scale. You will work across Kubernetes, cloud, identity, secrets, networking, telemetry, incident management and automation to provide the production foundation for Arm's AI platform. Participate in production support and our paid on-call rota for high impact incident response. Production AI runtime platforms: Build/deploy and operate the infrastructure for centrally hosted AI platform services, including MCP server infrastructure, model gateway services and supporting control-plane components. Design runtime patterns for isolation, scalability, secure execution, capacity management and cost-aware operation. Automate provisioning, configuration, upgrades and lifecycle management using infrastructure-as-code and GitOps patterns. Reliability, observability and support: Define and implement service-level indicators, service-level objectives, alerting, dashboards, runbooks and support workflows. Incident response, post-incident review and vendor outage handling for AI services embedded in engineering workflows. Build telemetry that helps Arm understand AI platform health, usage, performance, cost and operational risk. Secure operations by default: Ensure platform components meet production readiness, security and compliance expectations. Help make secure AI usage the default by providing reliable paved paths rather than manual or fragmented infrastructure. Responsibilities: Technical leadership - this role will become a recognised expert and lead engineer on our AI platform's technology stack. Build, operate and continuously improve the production infrastructure for Arm's AI platform services, using both custom and third party solutions. Own reliability, scalability, monitoring, alerting, incident response, runbooks and operational readiness for AI platform components. Develop automation for provisioning, deployment, configuration, backup, recovery, patching, upgrades and lifecycle management. Implement secure runtime patterns, including workload isolation, secrets management, identity integration, network controls and auditability. Contribute towards platform roadmap planning and prioritisation. Required Skills and Experience: Expert level experience deploying and operating production infrastructure or platform services in a Linux-based environment using Kubernetes, containers, cloud or private-cloud platforms, infrastructure-as-code and CI/CD/GitOps tooling. Very strong understanding of security fundamentals for production platforms, including identity, secrets, access control, network segmentation, vulnerability management and audit logging. Familiarity with AI platform concepts such as model routing, MCP servers, agentic workflows, RAG systems or LLM observability. Very strong automation and scripting skills, for example Terraform, Go, Python or similar. Experience with incident management, problem management, demand forecasting, and production readiness practices. "Nice To Have" Skills and Experience: Experience operating internal developer platforms, AI platforms, model gateways, MCP infrastructure or other shared engineering platforms. Experience with service mesh, policy-as-code, workload identity, sandboxing, secure runtime environments or multi-tenant platform designs. Experience with regulated or security-sensitive engineering environments. Working knowledge of Open AI tools, products and APIs! In Return: Joining Arm means being part of an elite team that is high-reaching and determined to achieve outstanding results. You will collaborate with industry leaders and innovators, contributing to flawless service builds that improve our global operations. We provide an encouraging and inclusive environment that encourages continuous learning and professional growth. Your success is our success, and together we can build services that compete on the world stage! Accommodations at Arm At Arm, we want to build extraordinary teams. If you need an adjustment or an accommodation during the recruitment process, please email . To note, by sending us the requested information, you consent to its use by Arm to arrange for appropriate accommodations. All accommodation or adjustment requests will be treated with confidentiality, and information concerning these requests will only be disclosed as necessary to provide the accommodation. Although this is not an exhaustive list, examples of support include breaks between interviews, having documents read aloud, or office accessibility. Please email us about anything we can do to accommodate you during the recruitment process. Hybrid Working at Arm Arm's approach to hybrid working is designed to create a working environment that supports both high performance and personal wellbeing. We believe in bringing people together face to face to enable us to work at pace, whilst recognizing the value of flexibility. Within that framework, we empower groups/teams to determine their own hybrid working patterns, depending on the work and the team's needs. Details of what this means for each role will be shared upon application. In some cases, the flexibility we can offer is limited by local legal, regulatory, tax, or other considerations, and where this is the case, we will collaborate with you to find the best solution. Please talk to us to find out more about what this could look like for you. Equal Opportunities at Arm Arm is an equal opportunity employer, committed to providing an environment of mutual respect where equal opportunities are available to all applicants and colleagues. We are a diverse organization of dedicated and innovative individuals, and don't discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
Rivan Industries () is a synthetic fuel company designed to decarbonise heavy industry. We aim to make synthetic fuel cheaper than fossil fuels and sustain life on earth by keeping the CO 2 locked underground. We design and manufacture modular synthetic fuel plants consisting of a DAC system, an Alkaline Electrolyser, and a Sabatier Reactor, whilst vertically integrating with off-grid DC solar and the European gas-grid. Here's our entire business plan. We've recently deployed the UK's largest synthetic fuel plant and are planning to 1000x this scale in the next 2 years. We need robust, scalable control systems that can be deployed across a growing fleet of synthetic fuel plants. We are looking for someone who can manage our existing pilot deployments and operations, and grow our controls systems to deploy 1000s of synthetic fuel plants ASAP. The role Own functional safety for our systems - ensuring compliance with existing risk controls (HAZOP/HAZID, Cause & Effect, Fault Tree) Manage the performance and reliability of deployed control systems Design control architectures that can scale from pilot to thousands of deployments Standardise PLC, SCADA, and telemetry patterns across sites Develop remote diagnostics and fleet-wide monitoring capability Core skills Lead the design, commissioning, and lifecycle ownership of distributed control systems in hazardous or high-energy process environments Experience implementing safety functions under IEC 61508 (or related standards), including SRS development, SIL verification, and change management Delivered systems compliant with LVD, Machinery Directive, and CE/UKCA IEC 61131-3 (ST/LD/FBD, Codesys), HMI design and development Integration and troubleshooting of industrial communications (Modbus TCP/RTU, CANbus, OPC-UA) in multi-device, multi-site systems Hands on commissioning and fault finding experience on complex industrial systems Factory automation experience Python 3 Solidworks Electrical / ECAD tools Specifics £70-100k Salary depending on experience Significant share options as part of the early team In-person work at our HQ in Bermondsey, South London Extensive relocation support, including £6000 per year extra to live close to our HQ Private health insurance Unlimited time off We encourage exceptional applicants from all backgrounds to apply for this role, even if they do not meet all the requirements listed.
Aug 25, 2026
Full time
Rivan Industries () is a synthetic fuel company designed to decarbonise heavy industry. We aim to make synthetic fuel cheaper than fossil fuels and sustain life on earth by keeping the CO 2 locked underground. We design and manufacture modular synthetic fuel plants consisting of a DAC system, an Alkaline Electrolyser, and a Sabatier Reactor, whilst vertically integrating with off-grid DC solar and the European gas-grid. Here's our entire business plan. We've recently deployed the UK's largest synthetic fuel plant and are planning to 1000x this scale in the next 2 years. We need robust, scalable control systems that can be deployed across a growing fleet of synthetic fuel plants. We are looking for someone who can manage our existing pilot deployments and operations, and grow our controls systems to deploy 1000s of synthetic fuel plants ASAP. The role Own functional safety for our systems - ensuring compliance with existing risk controls (HAZOP/HAZID, Cause & Effect, Fault Tree) Manage the performance and reliability of deployed control systems Design control architectures that can scale from pilot to thousands of deployments Standardise PLC, SCADA, and telemetry patterns across sites Develop remote diagnostics and fleet-wide monitoring capability Core skills Lead the design, commissioning, and lifecycle ownership of distributed control systems in hazardous or high-energy process environments Experience implementing safety functions under IEC 61508 (or related standards), including SRS development, SIL verification, and change management Delivered systems compliant with LVD, Machinery Directive, and CE/UKCA IEC 61131-3 (ST/LD/FBD, Codesys), HMI design and development Integration and troubleshooting of industrial communications (Modbus TCP/RTU, CANbus, OPC-UA) in multi-device, multi-site systems Hands on commissioning and fault finding experience on complex industrial systems Factory automation experience Python 3 Solidworks Electrical / ECAD tools Specifics £70-100k Salary depending on experience Significant share options as part of the early team In-person work at our HQ in Bermondsey, South London Extensive relocation support, including £6000 per year extra to live close to our HQ Private health insurance Unlimited time off We encourage exceptional applicants from all backgrounds to apply for this role, even if they do not meet all the requirements listed.
# SRE Managing Consultant - Cloud Operating ModelManchester, LondonApply for this job Permanent Experienced Professionals Strategy & Transformation ID 428177-en\_GB Capgemini Invent At Capgemini Invent, we believe difference drives change. As inventive transformation consultants, we blend our strategic, creative and scientific capabilities, collaborating closely with clients to deliver cutting-edge solutions. Join us to drive transformation tailored to our client's challenges of today and tomorrow. Informed and validated by science and data. Superpowered by creativity and design. All underpinned by technology created with purpose. Your Role As an SRE Consultant (Manager) at Capgemini Invent you will be part of our Cloud Advisory capability within the wider Business Technology capability unit. Our cloud advisory capability aims to inspire, lead and support organisations on their journey of adopting cloud for creating business and revenue models, generating growth, ensuring regulatory compliance and reducing their carbon footprint. Specifically In your role you will help clients build and embed reliability as an engineering discipline, shifting from ticket-led operations to measurable, product-aligned service performance. You will assess and shape the operating model, ways of working, and governance required to run resilient cloud and hybrid platforms at scale, partnering with engineering, operations, security and product leaders to improve outcomes across availability, reliability, scalability and operational performance.This will include: SRE Operating Model & Ways of Working : Define and implement SRE ways of working and engagement patterns, aligning reliability practices with existing ITSM/ITIL processes (e.g., incident, problem, release and change) and modern engineering delivery. Reliability Measures (SLIs/SLOs) & Error Budgets : Establish service measures and targets (SLIs/SLOs) and introduce Error Budgets to enable data-driven trade-offs between reliability and delivery velocity. Observability & Operational Insight: Shape observability approaches (metrics/logs/traces) and operational monitoring models that make reliability risks visible and actionable, improving operational decision-making. Incident Excellence & Continuous Learning: Design incident analysis and improvement loops, including practical approaches that strengthen incident response and drive learning through post-incident improvement actions. Toil Reduction Through Engineering & Automation: Identify high-friction operational work and prioritise engineering-led automation to reduce manual effort, improve repeatability, and increase operational scalability. SRE Capability Assessment & Roadmaps : Assess SRE maturity/capabilities (e.g., availability, reliability, scalability, complexity and operational performance) and create a phased roadmap from assessment through recommendations and implementation. Cross-discipline Enablement (DevSecOps / Platform / SRE): Improve collaboration across engineering disciplines by standardising processes and enabling platform and delivery capabilities that embed resiliency into application and infrastructure layers. Client Advisory Leadership: Lead advisory engagements, guide senior stakeholders through decisions on reliability investment, and coach teams to adopt new practices and measures sustainably (including training/enablement where needed).# part of your role you will also have the opportunity to contribute to the business and your own personal growth, through activities that form part of the following categories: Business Development - Leading/contributing to proposals, RFPs, bids, proposition development, client pitch contribution, client hosting at events. Internal Contribution - Campaign development, internal think-tanks, whitepapers, practice development (operations, recruitment, team events & activities), offering development. Learning & Development - Training to support your career development and the skills Your Profile Extensive experience in client facing consulting and advisory roles, operating credibly with senior stakeholders and shaping complex transformation engagements. Proven ability to lead and own advisory engagements end to end, building trusted senior client relationships and actively contributing to shaping, selling, and expanding consulting work. Currently working in a major Consulting firm, and/or in industry but having a Consulting background Proven ability to be successful in a matrixed organisation, and to enlist support and commitment from peers in selling and delivering consulting solutions. Experience of proposition building and delivery. Demonstrated business development capability, leveraging personal networks and firm relationships to originate, shape, and grow SRE, cloud, and operational resilience consulting opportunities. Experience working with at least one major cloud service provider (AWS, Microsoft Azure, or Google Cloud Platform), applying SRE and operating model principles in advisory, transformation, or large scale delivery contexts; associate level certifications are desirable but not mandatory. Design, establish, and evolve SRE led centres of excellence (e.g. Reliability, Observability, or Operational Excellence), setting enterprise level standards for SLIs/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms. Exposure to modern observability tooling and ecosystems (e.g. Datadog, Dynatrace, Prometheus, OpenTelemetry, Loki), with a strong understanding of how metrics, logs, and traces are applied to inform reliability strategy, incident management, and operational decision making. Security Check (SC) Clearance To be successfully appointed to this role, it is a requirement to obtain Security Check (SC) clearance. ()To obtain SC clearance, the successful applicant must have resided continuously within the United Kingdom for the last 5 years, along with other criteria and requirements.Throughout the recruitment process, you will be asked questions about your security clearance eligibility such as, but not limited to, country of residence and nationality.Some posts are restricted to sole UK Nationals for security reasons; therefore you may be asked about your citizenship in the application process. What You'll Love About Working Here Join the close-knit, rapidly growing Cloud Transformation Tribe at Capgemini Invent, where you'll play a key role in helping top organisations unlock the full potential of their cloud and infrastructure investments. As part of our team, you'll work on impactful projects that drive innovation and efficiency, collaborating closely with experts in a supportive, agile environment that values growth, learning, and teamwork. If you're excited to be part of a dynamic group making real transformations in cloud technology, Capgemini Invent is the place to grow.We provide a host of opportunities for learning and certification through internal and partner led programmes and hold monthly showcases of our digital transformation initiatives, sharing knowledge and showing off how the power of technology is impacting our clients Need To Know At Capgemini we don't just believe in Diversity & Inclusion, we actively go out to making it a working reality. Driven by our core values and Active Inclusion Campaign, we build environments where you can bring you whole self to work.We aim to build an environment where employees can enjoy a positive work-life balance. We embed hybrid working in all that we do and make flexible working arrangements the day-to-day reality for our people. All UK employees are eligible to request flexible working arrangements.Employee wellbeing is vitally important to us as an organisation. We see a healthy and happy workforce a critical component for us to achieve our organisational ambitions.To help support wellbeing we have trained 'Mental Health Champions' across each of our business areas. We have also invested in wellbeing apps such as Thrive and Peppy. CSR We're also focused on using tech to have a positive social impact. So, we're working to reduce our own carbon footprint and improve everyone's access to a digital world. It's something we're really serious about. In fact, we were even named as one of the world's most ethical companies by the Ethisphere Institute for the 10th year. When you join Capgemini, you'll join a team that does the right thing. We are a Disability Confident Employer Capgemini is proud to be a Disability Confident Employer (Level 2) under the UK Government's Disability Confident scheme. As part of our commitment to inclusive recruitment, we will offer an interview to all candidates who:Declare they have a disability, and Meet the minimum essential criteria for the role.Please opt in during the application process.# you will have London, Manchester or Glasgow as an office base location, you must be fully flexible in terms of assignment location, as these roles may involve periods of time away from home at short notice.We offer a remuneration package which includes flexible benefits options for you to choose to suit your own personal circumstances and a variable element dependent grade and on company and personal performance.
Aug 24, 2026
Full time
# SRE Managing Consultant - Cloud Operating ModelManchester, LondonApply for this job Permanent Experienced Professionals Strategy & Transformation ID 428177-en\_GB Capgemini Invent At Capgemini Invent, we believe difference drives change. As inventive transformation consultants, we blend our strategic, creative and scientific capabilities, collaborating closely with clients to deliver cutting-edge solutions. Join us to drive transformation tailored to our client's challenges of today and tomorrow. Informed and validated by science and data. Superpowered by creativity and design. All underpinned by technology created with purpose. Your Role As an SRE Consultant (Manager) at Capgemini Invent you will be part of our Cloud Advisory capability within the wider Business Technology capability unit. Our cloud advisory capability aims to inspire, lead and support organisations on their journey of adopting cloud for creating business and revenue models, generating growth, ensuring regulatory compliance and reducing their carbon footprint. Specifically In your role you will help clients build and embed reliability as an engineering discipline, shifting from ticket-led operations to measurable, product-aligned service performance. You will assess and shape the operating model, ways of working, and governance required to run resilient cloud and hybrid platforms at scale, partnering with engineering, operations, security and product leaders to improve outcomes across availability, reliability, scalability and operational performance.This will include: SRE Operating Model & Ways of Working : Define and implement SRE ways of working and engagement patterns, aligning reliability practices with existing ITSM/ITIL processes (e.g., incident, problem, release and change) and modern engineering delivery. Reliability Measures (SLIs/SLOs) & Error Budgets : Establish service measures and targets (SLIs/SLOs) and introduce Error Budgets to enable data-driven trade-offs between reliability and delivery velocity. Observability & Operational Insight: Shape observability approaches (metrics/logs/traces) and operational monitoring models that make reliability risks visible and actionable, improving operational decision-making. Incident Excellence & Continuous Learning: Design incident analysis and improvement loops, including practical approaches that strengthen incident response and drive learning through post-incident improvement actions. Toil Reduction Through Engineering & Automation: Identify high-friction operational work and prioritise engineering-led automation to reduce manual effort, improve repeatability, and increase operational scalability. SRE Capability Assessment & Roadmaps : Assess SRE maturity/capabilities (e.g., availability, reliability, scalability, complexity and operational performance) and create a phased roadmap from assessment through recommendations and implementation. Cross-discipline Enablement (DevSecOps / Platform / SRE): Improve collaboration across engineering disciplines by standardising processes and enabling platform and delivery capabilities that embed resiliency into application and infrastructure layers. Client Advisory Leadership: Lead advisory engagements, guide senior stakeholders through decisions on reliability investment, and coach teams to adopt new practices and measures sustainably (including training/enablement where needed).# part of your role you will also have the opportunity to contribute to the business and your own personal growth, through activities that form part of the following categories: Business Development - Leading/contributing to proposals, RFPs, bids, proposition development, client pitch contribution, client hosting at events. Internal Contribution - Campaign development, internal think-tanks, whitepapers, practice development (operations, recruitment, team events & activities), offering development. Learning & Development - Training to support your career development and the skills Your Profile Extensive experience in client facing consulting and advisory roles, operating credibly with senior stakeholders and shaping complex transformation engagements. Proven ability to lead and own advisory engagements end to end, building trusted senior client relationships and actively contributing to shaping, selling, and expanding consulting work. Currently working in a major Consulting firm, and/or in industry but having a Consulting background Proven ability to be successful in a matrixed organisation, and to enlist support and commitment from peers in selling and delivering consulting solutions. Experience of proposition building and delivery. Demonstrated business development capability, leveraging personal networks and firm relationships to originate, shape, and grow SRE, cloud, and operational resilience consulting opportunities. Experience working with at least one major cloud service provider (AWS, Microsoft Azure, or Google Cloud Platform), applying SRE and operating model principles in advisory, transformation, or large scale delivery contexts; associate level certifications are desirable but not mandatory. Design, establish, and evolve SRE led centres of excellence (e.g. Reliability, Observability, or Operational Excellence), setting enterprise level standards for SLIs/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms. Exposure to modern observability tooling and ecosystems (e.g. Datadog, Dynatrace, Prometheus, OpenTelemetry, Loki), with a strong understanding of how metrics, logs, and traces are applied to inform reliability strategy, incident management, and operational decision making. Security Check (SC) Clearance To be successfully appointed to this role, it is a requirement to obtain Security Check (SC) clearance. ()To obtain SC clearance, the successful applicant must have resided continuously within the United Kingdom for the last 5 years, along with other criteria and requirements.Throughout the recruitment process, you will be asked questions about your security clearance eligibility such as, but not limited to, country of residence and nationality.Some posts are restricted to sole UK Nationals for security reasons; therefore you may be asked about your citizenship in the application process. What You'll Love About Working Here Join the close-knit, rapidly growing Cloud Transformation Tribe at Capgemini Invent, where you'll play a key role in helping top organisations unlock the full potential of their cloud and infrastructure investments. As part of our team, you'll work on impactful projects that drive innovation and efficiency, collaborating closely with experts in a supportive, agile environment that values growth, learning, and teamwork. If you're excited to be part of a dynamic group making real transformations in cloud technology, Capgemini Invent is the place to grow.We provide a host of opportunities for learning and certification through internal and partner led programmes and hold monthly showcases of our digital transformation initiatives, sharing knowledge and showing off how the power of technology is impacting our clients Need To Know At Capgemini we don't just believe in Diversity & Inclusion, we actively go out to making it a working reality. Driven by our core values and Active Inclusion Campaign, we build environments where you can bring you whole self to work.We aim to build an environment where employees can enjoy a positive work-life balance. We embed hybrid working in all that we do and make flexible working arrangements the day-to-day reality for our people. All UK employees are eligible to request flexible working arrangements.Employee wellbeing is vitally important to us as an organisation. We see a healthy and happy workforce a critical component for us to achieve our organisational ambitions.To help support wellbeing we have trained 'Mental Health Champions' across each of our business areas. We have also invested in wellbeing apps such as Thrive and Peppy. CSR We're also focused on using tech to have a positive social impact. So, we're working to reduce our own carbon footprint and improve everyone's access to a digital world. It's something we're really serious about. In fact, we were even named as one of the world's most ethical companies by the Ethisphere Institute for the 10th year. When you join Capgemini, you'll join a team that does the right thing. We are a Disability Confident Employer Capgemini is proud to be a Disability Confident Employer (Level 2) under the UK Government's Disability Confident scheme. As part of our commitment to inclusive recruitment, we will offer an interview to all candidates who:Declare they have a disability, and Meet the minimum essential criteria for the role.Please opt in during the application process.# you will have London, Manchester or Glasgow as an office base location, you must be fully flexible in terms of assignment location, as these roles may involve periods of time away from home at short notice.We offer a remuneration package which includes flexible benefits options for you to choose to suit your own personal circumstances and a variable element dependent grade and on company and personal performance.
# SRE Managing Consultant - Cloud Operating ModelManchester, LondonApply for this job Permanent Experienced Professionals Strategy & Transformation ID 428177-en\_GB Capgemini Invent At Capgemini Invent, we believe difference drives change. As inventive transformation consultants, we blend our strategic, creative and scientific capabilities, collaborating closely with clients to deliver cutting-edge solutions. Join us to drive transformation tailored to our client's challenges of today and tomorrow. Informed and validated by science and data. Superpowered by creativity and design. All underpinned by technology created with purpose. Your Role As an SRE Consultant (Manager) at Capgemini Invent you will be part of our Cloud Advisory capability within the wider Business Technology capability unit. Our cloud advisory capability aims to inspire, lead and support organisations on their journey of adopting cloud for creating business and revenue models, generating growth, ensuring regulatory compliance and reducing their carbon footprint. Specifically In your role you will help clients build and embed reliability as an engineering discipline, shifting from ticket-led operations to measurable, product-aligned service performance. You will assess and shape the operating model, ways of working, and governance required to run resilient cloud and hybrid platforms at scale, partnering with engineering, operations, security and product leaders to improve outcomes across availability, reliability, scalability and operational performance.This will include: SRE Operating Model & Ways of Working : Define and implement SRE ways of working and engagement patterns, aligning reliability practices with existing ITSM/ITIL processes (e.g., incident, problem, release and change) and modern engineering delivery. Reliability Measures (SLIs/SLOs) & Error Budgets : Establish service measures and targets (SLIs/SLOs) and introduce Error Budgets to enable data-driven trade-offs between reliability and delivery velocity. Observability & Operational Insight: Shape observability approaches (metrics/logs/traces) and operational monitoring models that make reliability risks visible and actionable, improving operational decision-making. Incident Excellence & Continuous Learning: Design incident analysis and improvement loops, including practical approaches that strengthen incident response and drive learning through post-incident improvement actions. Toil Reduction Through Engineering & Automation: Identify high-friction operational work and prioritise engineering-led automation to reduce manual effort, improve repeatability, and increase operational scalability. SRE Capability Assessment & Roadmaps : Assess SRE maturity/capabilities (e.g., availability, reliability, scalability, complexity and operational performance) and create a phased roadmap from assessment through recommendations and implementation. Cross-discipline Enablement (DevSecOps / Platform / SRE): Improve collaboration across engineering disciplines by standardising processes and enabling platform and delivery capabilities that embed resiliency into application and infrastructure layers. Client Advisory Leadership: Lead advisory engagements, guide senior stakeholders through decisions on reliability investment, and coach teams to adopt new practices and measures sustainably (including training/enablement where needed).# part of your role you will also have the opportunity to contribute to the business and your own personal growth, through activities that form part of the following categories: Business Development - Leading/contributing to proposals, RFPs, bids, proposition development, client pitch contribution, client hosting at events. Internal Contribution - Campaign development, internal think-tanks, whitepapers, practice development (operations, recruitment, team events & activities), offering development. Learning & Development - Training to support your career development and the skills Your Profile Extensive experience in client facing consulting and advisory roles, operating credibly with senior stakeholders and shaping complex transformation engagements. Proven ability to lead and own advisory engagements end to end, building trusted senior client relationships and actively contributing to shaping, selling, and expanding consulting work. Currently working in a major Consulting firm, and/or in industry but having a Consulting background Proven ability to be successful in a matrixed organisation, and to enlist support and commitment from peers in selling and delivering consulting solutions. Experience of proposition building and delivery. Demonstrated business development capability, leveraging personal networks and firm relationships to originate, shape, and grow SRE, cloud, and operational resilience consulting opportunities. Experience working with at least one major cloud service provider (AWS, Microsoft Azure, or Google Cloud Platform), applying SRE and operating model principles in advisory, transformation, or large scale delivery contexts; associate level certifications are desirable but not mandatory. Design, establish, and evolve SRE led centres of excellence (e.g. Reliability, Observability, or Operational Excellence), setting enterprise level standards for SLIs/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms. Exposure to modern observability tooling and ecosystems (e.g. Datadog, Dynatrace, Prometheus, OpenTelemetry, Loki), with a strong understanding of how metrics, logs, and traces are applied to inform reliability strategy, incident management, and operational decision making. Security Check (SC) Clearance To be successfully appointed to this role, it is a requirement to obtain Security Check (SC) clearance. ()To obtain SC clearance, the successful applicant must have resided continuously within the United Kingdom for the last 5 years, along with other criteria and requirements.Throughout the recruitment process, you will be asked questions about your security clearance eligibility such as, but not limited to, country of residence and nationality.Some posts are restricted to sole UK Nationals for security reasons; therefore you may be asked about your citizenship in the application process. What You'll Love About Working Here Join the close-knit, rapidly growing Cloud Transformation Tribe at Capgemini Invent, where you'll play a key role in helping top organisations unlock the full potential of their cloud and infrastructure investments. As part of our team, you'll work on impactful projects that drive innovation and efficiency, collaborating closely with experts in a supportive, agile environment that values growth, learning, and teamwork. If you're excited to be part of a dynamic group making real transformations in cloud technology, Capgemini Invent is the place to grow.We provide a host of opportunities for learning and certification through internal and partner led programmes and hold monthly showcases of our digital transformation initiatives, sharing knowledge and showing off how the power of technology is impacting our clients Need To Know At Capgemini we don't just believe in Diversity & Inclusion, we actively go out to making it a working reality. Driven by our core values and Active Inclusion Campaign, we build environments where you can bring you whole self to work.We aim to build an environment where employees can enjoy a positive work-life balance. We embed hybrid working in all that we do and make flexible working arrangements the day-to-day reality for our people. All UK employees are eligible to request flexible working arrangements.Employee wellbeing is vitally important to us as an organisation. We see a healthy and happy workforce a critical component for us to achieve our organisational ambitions.To help support wellbeing we have trained 'Mental Health Champions' across each of our business areas. We have also invested in wellbeing apps such as Thrive and Peppy. CSR We're also focused on using tech to have a positive social impact. So, we're working to reduce our own carbon footprint and improve everyone's access to a digital world. It's something we're really serious about. In fact, we were even named as one of the world's most ethical companies by the Ethisphere Institute for the 10th year. When you join Capgemini, you'll join a team that does the right thing. We are a Disability Confident Employer Capgemini is proud to be a Disability Confident Employer (Level 2) under the UK Government's Disability Confident scheme. As part of our commitment to inclusive recruitment, we will offer an interview to all candidates who:Declare they have a disability, and Meet the minimum essential criteria for the role.Please opt in during the application process.# you will have London, Manchester or Glasgow as an office base location, you must be fully flexible in terms of assignment location, as these roles may involve periods of time away from home at short notice.We offer a remuneration package which includes flexible benefits options for you to choose to suit your own personal circumstances and a variable element dependent grade and on company and personal performance.
Aug 24, 2026
Full time
# SRE Managing Consultant - Cloud Operating ModelManchester, LondonApply for this job Permanent Experienced Professionals Strategy & Transformation ID 428177-en\_GB Capgemini Invent At Capgemini Invent, we believe difference drives change. As inventive transformation consultants, we blend our strategic, creative and scientific capabilities, collaborating closely with clients to deliver cutting-edge solutions. Join us to drive transformation tailored to our client's challenges of today and tomorrow. Informed and validated by science and data. Superpowered by creativity and design. All underpinned by technology created with purpose. Your Role As an SRE Consultant (Manager) at Capgemini Invent you will be part of our Cloud Advisory capability within the wider Business Technology capability unit. Our cloud advisory capability aims to inspire, lead and support organisations on their journey of adopting cloud for creating business and revenue models, generating growth, ensuring regulatory compliance and reducing their carbon footprint. Specifically In your role you will help clients build and embed reliability as an engineering discipline, shifting from ticket-led operations to measurable, product-aligned service performance. You will assess and shape the operating model, ways of working, and governance required to run resilient cloud and hybrid platforms at scale, partnering with engineering, operations, security and product leaders to improve outcomes across availability, reliability, scalability and operational performance.This will include: SRE Operating Model & Ways of Working : Define and implement SRE ways of working and engagement patterns, aligning reliability practices with existing ITSM/ITIL processes (e.g., incident, problem, release and change) and modern engineering delivery. Reliability Measures (SLIs/SLOs) & Error Budgets : Establish service measures and targets (SLIs/SLOs) and introduce Error Budgets to enable data-driven trade-offs between reliability and delivery velocity. Observability & Operational Insight: Shape observability approaches (metrics/logs/traces) and operational monitoring models that make reliability risks visible and actionable, improving operational decision-making. Incident Excellence & Continuous Learning: Design incident analysis and improvement loops, including practical approaches that strengthen incident response and drive learning through post-incident improvement actions. Toil Reduction Through Engineering & Automation: Identify high-friction operational work and prioritise engineering-led automation to reduce manual effort, improve repeatability, and increase operational scalability. SRE Capability Assessment & Roadmaps : Assess SRE maturity/capabilities (e.g., availability, reliability, scalability, complexity and operational performance) and create a phased roadmap from assessment through recommendations and implementation. Cross-discipline Enablement (DevSecOps / Platform / SRE): Improve collaboration across engineering disciplines by standardising processes and enabling platform and delivery capabilities that embed resiliency into application and infrastructure layers. Client Advisory Leadership: Lead advisory engagements, guide senior stakeholders through decisions on reliability investment, and coach teams to adopt new practices and measures sustainably (including training/enablement where needed).# part of your role you will also have the opportunity to contribute to the business and your own personal growth, through activities that form part of the following categories: Business Development - Leading/contributing to proposals, RFPs, bids, proposition development, client pitch contribution, client hosting at events. Internal Contribution - Campaign development, internal think-tanks, whitepapers, practice development (operations, recruitment, team events & activities), offering development. Learning & Development - Training to support your career development and the skills Your Profile Extensive experience in client facing consulting and advisory roles, operating credibly with senior stakeholders and shaping complex transformation engagements. Proven ability to lead and own advisory engagements end to end, building trusted senior client relationships and actively contributing to shaping, selling, and expanding consulting work. Currently working in a major Consulting firm, and/or in industry but having a Consulting background Proven ability to be successful in a matrixed organisation, and to enlist support and commitment from peers in selling and delivering consulting solutions. Experience of proposition building and delivery. Demonstrated business development capability, leveraging personal networks and firm relationships to originate, shape, and grow SRE, cloud, and operational resilience consulting opportunities. Experience working with at least one major cloud service provider (AWS, Microsoft Azure, or Google Cloud Platform), applying SRE and operating model principles in advisory, transformation, or large scale delivery contexts; associate level certifications are desirable but not mandatory. Design, establish, and evolve SRE led centres of excellence (e.g. Reliability, Observability, or Operational Excellence), setting enterprise level standards for SLIs/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms. Exposure to modern observability tooling and ecosystems (e.g. Datadog, Dynatrace, Prometheus, OpenTelemetry, Loki), with a strong understanding of how metrics, logs, and traces are applied to inform reliability strategy, incident management, and operational decision making. Security Check (SC) Clearance To be successfully appointed to this role, it is a requirement to obtain Security Check (SC) clearance. ()To obtain SC clearance, the successful applicant must have resided continuously within the United Kingdom for the last 5 years, along with other criteria and requirements.Throughout the recruitment process, you will be asked questions about your security clearance eligibility such as, but not limited to, country of residence and nationality.Some posts are restricted to sole UK Nationals for security reasons; therefore you may be asked about your citizenship in the application process. What You'll Love About Working Here Join the close-knit, rapidly growing Cloud Transformation Tribe at Capgemini Invent, where you'll play a key role in helping top organisations unlock the full potential of their cloud and infrastructure investments. As part of our team, you'll work on impactful projects that drive innovation and efficiency, collaborating closely with experts in a supportive, agile environment that values growth, learning, and teamwork. If you're excited to be part of a dynamic group making real transformations in cloud technology, Capgemini Invent is the place to grow.We provide a host of opportunities for learning and certification through internal and partner led programmes and hold monthly showcases of our digital transformation initiatives, sharing knowledge and showing off how the power of technology is impacting our clients Need To Know At Capgemini we don't just believe in Diversity & Inclusion, we actively go out to making it a working reality. Driven by our core values and Active Inclusion Campaign, we build environments where you can bring you whole self to work.We aim to build an environment where employees can enjoy a positive work-life balance. We embed hybrid working in all that we do and make flexible working arrangements the day-to-day reality for our people. All UK employees are eligible to request flexible working arrangements.Employee wellbeing is vitally important to us as an organisation. We see a healthy and happy workforce a critical component for us to achieve our organisational ambitions.To help support wellbeing we have trained 'Mental Health Champions' across each of our business areas. We have also invested in wellbeing apps such as Thrive and Peppy. CSR We're also focused on using tech to have a positive social impact. So, we're working to reduce our own carbon footprint and improve everyone's access to a digital world. It's something we're really serious about. In fact, we were even named as one of the world's most ethical companies by the Ethisphere Institute for the 10th year. When you join Capgemini, you'll join a team that does the right thing. We are a Disability Confident Employer Capgemini is proud to be a Disability Confident Employer (Level 2) under the UK Government's Disability Confident scheme. As part of our commitment to inclusive recruitment, we will offer an interview to all candidates who:Declare they have a disability, and Meet the minimum essential criteria for the role.Please opt in during the application process.# you will have London, Manchester or Glasgow as an office base location, you must be fully flexible in terms of assignment location, as these roles may involve periods of time away from home at short notice.We offer a remuneration package which includes flexible benefits options for you to choose to suit your own personal circumstances and a variable element dependent grade and on company and personal performance.
Role: Product & Account Lead Contract: Contract 6-12 months, Hybrid (ASAP start) Salary: Competitive Location: Greater Oxford Area A note from the Founders Oxford Dynamics is at an inflection point. We have spent five years building AI and robotic systems for some of the most demanding environments in the world. We are now taking that technology into the private sector, and life sciences is where we are starting. The decisions we make now will define not just how fast we grow, but who we become. You will work closely with all the team. You will be trusted with judgment calls. You will influence the business. And you will see the impact of your work every day in the work we do. If you are excited by ownership, pace and purpose - and by building something that genuinely matters - we would love to hear from you. Who We Are Founded in 2020, Oxford Dynamics (OD) is a fast-growing UK deep-tech company developing AI and robotic systems designed to operate in mission-critical environments. Our flagship AVIS (A Very Intelligent System) AI framework fuses multi-modal data - text, imagery, telemetry and sensor feeds - enabling operators to interrogate complex information at speed and make better decisions under pressure. Our STRIDER robotic platform performs autonomous tasks in hazardous environments, protecting people while extending operational reach. Our ambition is simple but demanding: to converge AI and robotics so machines can sense, understand and act in complex, real-world environments. Alongside our defence and security work, we are building a private sector business applying the same technology to scientific and industrial problems. Our biotech partnership is the flagship of that effort - using AVIS to help researchers navigate vast, fragmented scientific data and get to answers faster. What you will be doing here / why this role matters Oxford Dynamics is a small team who rely on a collaborative and positive approach, and so the right attitude for this role is equally as important as experience. We are at an important stage and time in our growth, and as our Product & Account Lead you will be an essential part of our success. This is a role with a name on it. You will own our biotech account outright - the client relationship, the product, the team that builds it, and how the account performs commercially. You will be the person the client calls and the person our engineers look to for direction. There is no established playbook here; you will be writing it. If you are excited by owning a client relationship, shaping a product, and growing an account from first deployment to long-term partnership, this role is built for you! Role Summary Oxford Dynamics seeks a Product & Account Lead to own our biotech account end to end. You will hold the client relationship, set product priorities from client and market needs, direct the internal delivery team, and carry responsibility for the commercial performance and growth of the account. The role bridges cutting-edge AI and practical deployment, making sure what we build lands with real users, gets adopted, and turns into a scalable long-term partnership. This is a private sector role. No security clearance is required. Key Responsibilities Own the Biotech Account Act as the primary point of contact for the client, building trust at every level from day-to-day users through to senior stakeholders. Set and manage priorities with the client, keeping expectations, scope and commitments clear and realistic on both sides. Run communication end to end: regular updates, structured reviews, honest escalation when something slips, and no surprises. Own delivery into the account, from what is promised through to what actually ships and gets used. Take responsibility for overall account performance - health, satisfaction, renewal risk and commercial outcomes - and report on it to leadership. Lead the Product and Team Manage the internal team delivering the account, setting direction, sequencing work and keeping momentum going. Translate client feedback and market needs into a clear product roadmap, deciding what gets built next and why. Coordinate resources across engineering, AI, design and commercial functions, flagging constraints early and resolving competing demands. Ensure everyone is aligned on responsibilities and outcomes, so each person knows what they own and what good looks like. Partner with Tech Leads on design and architecture trade-offs, aligning what is technically ambitious with what the client actually needs. Drive Delivery and Scale Own project plans, milestones and deadlines, managing scope, risk and change control to deliver on time and within cost. Identify growth opportunities within the account and pursue them, expanding usage into new teams, workflows and use cases. Increase adoption and revenue, tracking the metrics that show the product is genuinely embedded in the client's work. Solve problems independently, making decisions with incomplete information and escalating with options rather than open questions. Build a scalable long-term account, replacing one-off effort with repeatable process so the relationship grows without breaking. Reporting and Ways of Working Run lightweight Agile ceremonies appropriate to the team, with clear acceptance criteria and measurable outcomes for each increment. Produce crisp client and leadership reporting, acting as the single point of contact to unblock issues across internal teams, suppliers and stakeholders. Oversee vendor and subcontractor deliverables, ensuring integrations meet contractual, performance and quality requirements. Uphold data handling, confidentiality and quality standards appropriate to working with sensitive commercial and scientific data. Skills and Experience 5+ years managing AI projects with multiple stakeholders, delivering production releases in fast-moving environments. Proven ownership of a client relationship end to end, including commercial responsibility for how that account performs. Background in biotech, pharma or life sciences - understanding the requirements of our customers and speaking their language. Track record translating customer and market needs into product priorities that engineering teams can execute. Experience leading or coordinating technical teams, with the credibility to discuss trade-offs with engineers. Comfortable around modern AI and software stacks - LLMs, RAG, data pipelines, Python and React or equivalent - without needing to write the code yourself. Strong scope, requirements and budget control, with practical command of milestones, risk and change control. Genuinely self-directed: comfortable operating without a playbook and building the process as you go. Preferred / Bonus Experience growing an account through expansion, upsell or renewal, with measurable adoption and revenue outcomes. Portfolio of shipped AI or software products with measurable outcomes and references you can point to. Familiarity with scientific or research data, or with regulated environments. Comfort presenting to senior stakeholders and contributing to business development activities and proposals. Soft Skills High-agency ownership from first conversation to long-term partnership, with a bias to clarity, iteration and decision-making under uncertainty. Excellent communication across technical and non-technical audiences - scientists, engineers and executives alike - enabling fast, safe progress across disciplines. Commercial instinct: able to spot where value sits for the client and turn that into growth for the account. Calm, structured problem-solving during incidents and changing constraints, with clear follow-through on actions and learning. What We Offer Join the most exciting growth area in the UK: AI and Robotics! Every member of the Oxford Dynamics team has a major impact on the products and services we provide. Regardless of job title, you'll get to make a real difference and learn from colleagues about all areas of our business. Benefits include: Competitive Day Rate - outside IR35 Flexible working hours Hybrid working model Generous 29 days holiday in addition to public holidays (Full Time Equivalent) Oxford Dynamics is committed to creating an inclusive team experience for all. Regardless of race, gender, religion, sexual orientation, age, disability, or parental status, we believe our work is at its best when everyone feels free to be their authentic self.
Aug 24, 2026
Full time
Role: Product & Account Lead Contract: Contract 6-12 months, Hybrid (ASAP start) Salary: Competitive Location: Greater Oxford Area A note from the Founders Oxford Dynamics is at an inflection point. We have spent five years building AI and robotic systems for some of the most demanding environments in the world. We are now taking that technology into the private sector, and life sciences is where we are starting. The decisions we make now will define not just how fast we grow, but who we become. You will work closely with all the team. You will be trusted with judgment calls. You will influence the business. And you will see the impact of your work every day in the work we do. If you are excited by ownership, pace and purpose - and by building something that genuinely matters - we would love to hear from you. Who We Are Founded in 2020, Oxford Dynamics (OD) is a fast-growing UK deep-tech company developing AI and robotic systems designed to operate in mission-critical environments. Our flagship AVIS (A Very Intelligent System) AI framework fuses multi-modal data - text, imagery, telemetry and sensor feeds - enabling operators to interrogate complex information at speed and make better decisions under pressure. Our STRIDER robotic platform performs autonomous tasks in hazardous environments, protecting people while extending operational reach. Our ambition is simple but demanding: to converge AI and robotics so machines can sense, understand and act in complex, real-world environments. Alongside our defence and security work, we are building a private sector business applying the same technology to scientific and industrial problems. Our biotech partnership is the flagship of that effort - using AVIS to help researchers navigate vast, fragmented scientific data and get to answers faster. What you will be doing here / why this role matters Oxford Dynamics is a small team who rely on a collaborative and positive approach, and so the right attitude for this role is equally as important as experience. We are at an important stage and time in our growth, and as our Product & Account Lead you will be an essential part of our success. This is a role with a name on it. You will own our biotech account outright - the client relationship, the product, the team that builds it, and how the account performs commercially. You will be the person the client calls and the person our engineers look to for direction. There is no established playbook here; you will be writing it. If you are excited by owning a client relationship, shaping a product, and growing an account from first deployment to long-term partnership, this role is built for you! Role Summary Oxford Dynamics seeks a Product & Account Lead to own our biotech account end to end. You will hold the client relationship, set product priorities from client and market needs, direct the internal delivery team, and carry responsibility for the commercial performance and growth of the account. The role bridges cutting-edge AI and practical deployment, making sure what we build lands with real users, gets adopted, and turns into a scalable long-term partnership. This is a private sector role. No security clearance is required. Key Responsibilities Own the Biotech Account Act as the primary point of contact for the client, building trust at every level from day-to-day users through to senior stakeholders. Set and manage priorities with the client, keeping expectations, scope and commitments clear and realistic on both sides. Run communication end to end: regular updates, structured reviews, honest escalation when something slips, and no surprises. Own delivery into the account, from what is promised through to what actually ships and gets used. Take responsibility for overall account performance - health, satisfaction, renewal risk and commercial outcomes - and report on it to leadership. Lead the Product and Team Manage the internal team delivering the account, setting direction, sequencing work and keeping momentum going. Translate client feedback and market needs into a clear product roadmap, deciding what gets built next and why. Coordinate resources across engineering, AI, design and commercial functions, flagging constraints early and resolving competing demands. Ensure everyone is aligned on responsibilities and outcomes, so each person knows what they own and what good looks like. Partner with Tech Leads on design and architecture trade-offs, aligning what is technically ambitious with what the client actually needs. Drive Delivery and Scale Own project plans, milestones and deadlines, managing scope, risk and change control to deliver on time and within cost. Identify growth opportunities within the account and pursue them, expanding usage into new teams, workflows and use cases. Increase adoption and revenue, tracking the metrics that show the product is genuinely embedded in the client's work. Solve problems independently, making decisions with incomplete information and escalating with options rather than open questions. Build a scalable long-term account, replacing one-off effort with repeatable process so the relationship grows without breaking. Reporting and Ways of Working Run lightweight Agile ceremonies appropriate to the team, with clear acceptance criteria and measurable outcomes for each increment. Produce crisp client and leadership reporting, acting as the single point of contact to unblock issues across internal teams, suppliers and stakeholders. Oversee vendor and subcontractor deliverables, ensuring integrations meet contractual, performance and quality requirements. Uphold data handling, confidentiality and quality standards appropriate to working with sensitive commercial and scientific data. Skills and Experience 5+ years managing AI projects with multiple stakeholders, delivering production releases in fast-moving environments. Proven ownership of a client relationship end to end, including commercial responsibility for how that account performs. Background in biotech, pharma or life sciences - understanding the requirements of our customers and speaking their language. Track record translating customer and market needs into product priorities that engineering teams can execute. Experience leading or coordinating technical teams, with the credibility to discuss trade-offs with engineers. Comfortable around modern AI and software stacks - LLMs, RAG, data pipelines, Python and React or equivalent - without needing to write the code yourself. Strong scope, requirements and budget control, with practical command of milestones, risk and change control. Genuinely self-directed: comfortable operating without a playbook and building the process as you go. Preferred / Bonus Experience growing an account through expansion, upsell or renewal, with measurable adoption and revenue outcomes. Portfolio of shipped AI or software products with measurable outcomes and references you can point to. Familiarity with scientific or research data, or with regulated environments. Comfort presenting to senior stakeholders and contributing to business development activities and proposals. Soft Skills High-agency ownership from first conversation to long-term partnership, with a bias to clarity, iteration and decision-making under uncertainty. Excellent communication across technical and non-technical audiences - scientists, engineers and executives alike - enabling fast, safe progress across disciplines. Commercial instinct: able to spot where value sits for the client and turn that into growth for the account. Calm, structured problem-solving during incidents and changing constraints, with clear follow-through on actions and learning. What We Offer Join the most exciting growth area in the UK: AI and Robotics! Every member of the Oxford Dynamics team has a major impact on the products and services we provide. Regardless of job title, you'll get to make a real difference and learn from colleagues about all areas of our business. Benefits include: Competitive Day Rate - outside IR35 Flexible working hours Hybrid working model Generous 29 days holiday in addition to public holidays (Full Time Equivalent) Oxford Dynamics is committed to creating an inclusive team experience for all. Regardless of race, gender, religion, sexual orientation, age, disability, or parental status, we believe our work is at its best when everyone feels free to be their authentic self.
About the company - the company's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role the company's Safeguards organization builds the systems that keep Claude safe to use at scale. The Account Compromise team owns one specific slice of that work: protecting the people and organizations who use Claude from losing control of their accounts, and preventing compromised accounts and credentials from being used to abuse our platform. Account takeover, credential stuffing, phishing-driven session theft, leaked API keys, and resold access all cause real harm - to the customers whose accounts are hijacked, and to the wider ecosystem when stolen access is used to route abusive traffic through legitimate accounts. We are looking for a staff+ engineer to set the technical direction for this work. You will own the architecture of how we detect, contain, and remediate account compromise across Claude and the Claude Developer Platform, making the design decisions that other engineers and teams build on top of. You will move between deep technical work and adversarial problem-solving: threat modelling how attackers will adapt, scoping multi-month projects from ambiguous starting points, and leading complex live investigations when they escalate. This is a hard problem space. Attackers iterate quickly, the signals separating a compromised account from an unusual but legitimate one are subtle, and every defence carries a cost to real users if it fires incorrectly. You will be building the playbook, and the judgement calls about where to draw those lines will largely be yours to make and defend. You will operate with high autonomy - owning detection coverage and incident leadership, driving alignment with Security, Product, and Policy teams, and shaping how the company approaches this problem globally rather than executing against someone else's roadmap. Key responsibilities Set the technical direction and own the architecture for account compromise detection, response, and remediation across Claude and the Claude Developer Platform Independently scope and lead complex, multi-month engineering projects from an ambiguous starting point through to production systems that operate reliably under adversarial pressure Build and evolve detection systems that identify account takeover, credential abuse, and compromised API keys in near real time Design automated response flows that cut off attacker access while minimising disruption to legitimate users Lead investigations into significant compromise incidents end to end, then convert what you learn into durable, automated defences Threat model how attackers are likely to adapt, and prioritise the team's work against that view rather than only against incidents already observed Drive cross-organisational alignment on account security direction with Security, Product, Support, Policy and other partners. Define how the team measures success and hold the work to those measures Set technical standards for the domain and raise the bar for other engineers through code review, design review, and mentorship Surface patterns from compromise cases to research and product teams so that protections improve upstream Minimum qualifications 10 + years experience designing, building, and operating detection, anti-fraud, anti-abuse, or security systems in production A track record of independently scoping and delivering complex, ambiguous, multi-month technical projects Experience making architectural decisions in an adversarial domain that other engineers and teams then build on Proficiency in Python and SQL, with strong software engineering fundamentals and hands-on coding ability Experience leading investigations into account-based abuse or security incidents, and translating findings into automated detection Ability to reason rigorously about large behavioural or telemetry datasets, and to distinguish attacker behaviour from unusual but legitimate use Strong written communication and a track record of driving alignment across multiple teams and stakeholders Sound judgement about the tradeoff between stopping bad actors and disrupting legitimate users, and the ability to explain and defend where you have drawn that line Preferred qualifications Significant engineering experience in trust and safety, platform integrity, fraud, or detection and response, including time as a technical lead or mentor Deep familiarity with account attack techniques Experience with authentication and identity systems, including OAuth, single sign-on, multi-factor authentication, device binding, and risk-based authentication Experience applying machine learning to fraud or abuse detection, alongside a clear sense of when simpler rules-based approaches are the better choice Experience with cloud data tooling such as BigQuery, Spark, dbt, Airflow or similar Experience building tooling for operational or investigative teams, and partnering closely with the people who use it Interest in AI safety, and in the specific ways account compromise intersects with model misuse Representative projects - Design the architecture for real-time login risk scoring, and get agreement across Security, Product, and Safeguards on where step-up authentication should and should not fire Design detection for compromised API keys, together with security controls to limit exposure Rebuild product features to regain access quickly when customers get compromised Write the threat model that sets the team's roadmap for the next year, and bring the rest of the organisation along with it Instrument a false positive review loop that measures how often the team's defences affect legitimate users, and drive that number down The annual compensation range for this role is listed below. For sales roles, the range provided is the role's On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: £325,000 - £390,000 GBP Logistics Minimum education: Bachelor's degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that the company recruiters only contact you email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of the company. Be cautious of emails from other domains. Legitimate the company recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links-visit the directly for confirmed position openings. How we're different - We're different We believe that the highest-impact AI research will be big science. At the company we work as a single cohesive team on just a few large-scale research efforts. And we value impact - advancing our long-term goals of steerable, trustworthy AI - rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to the company, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute . click apply for full job details
Aug 24, 2026
Full time
About the company - the company's mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the role the company's Safeguards organization builds the systems that keep Claude safe to use at scale. The Account Compromise team owns one specific slice of that work: protecting the people and organizations who use Claude from losing control of their accounts, and preventing compromised accounts and credentials from being used to abuse our platform. Account takeover, credential stuffing, phishing-driven session theft, leaked API keys, and resold access all cause real harm - to the customers whose accounts are hijacked, and to the wider ecosystem when stolen access is used to route abusive traffic through legitimate accounts. We are looking for a staff+ engineer to set the technical direction for this work. You will own the architecture of how we detect, contain, and remediate account compromise across Claude and the Claude Developer Platform, making the design decisions that other engineers and teams build on top of. You will move between deep technical work and adversarial problem-solving: threat modelling how attackers will adapt, scoping multi-month projects from ambiguous starting points, and leading complex live investigations when they escalate. This is a hard problem space. Attackers iterate quickly, the signals separating a compromised account from an unusual but legitimate one are subtle, and every defence carries a cost to real users if it fires incorrectly. You will be building the playbook, and the judgement calls about where to draw those lines will largely be yours to make and defend. You will operate with high autonomy - owning detection coverage and incident leadership, driving alignment with Security, Product, and Policy teams, and shaping how the company approaches this problem globally rather than executing against someone else's roadmap. Key responsibilities Set the technical direction and own the architecture for account compromise detection, response, and remediation across Claude and the Claude Developer Platform Independently scope and lead complex, multi-month engineering projects from an ambiguous starting point through to production systems that operate reliably under adversarial pressure Build and evolve detection systems that identify account takeover, credential abuse, and compromised API keys in near real time Design automated response flows that cut off attacker access while minimising disruption to legitimate users Lead investigations into significant compromise incidents end to end, then convert what you learn into durable, automated defences Threat model how attackers are likely to adapt, and prioritise the team's work against that view rather than only against incidents already observed Drive cross-organisational alignment on account security direction with Security, Product, Support, Policy and other partners. Define how the team measures success and hold the work to those measures Set technical standards for the domain and raise the bar for other engineers through code review, design review, and mentorship Surface patterns from compromise cases to research and product teams so that protections improve upstream Minimum qualifications 10 + years experience designing, building, and operating detection, anti-fraud, anti-abuse, or security systems in production A track record of independently scoping and delivering complex, ambiguous, multi-month technical projects Experience making architectural decisions in an adversarial domain that other engineers and teams then build on Proficiency in Python and SQL, with strong software engineering fundamentals and hands-on coding ability Experience leading investigations into account-based abuse or security incidents, and translating findings into automated detection Ability to reason rigorously about large behavioural or telemetry datasets, and to distinguish attacker behaviour from unusual but legitimate use Strong written communication and a track record of driving alignment across multiple teams and stakeholders Sound judgement about the tradeoff between stopping bad actors and disrupting legitimate users, and the ability to explain and defend where you have drawn that line Preferred qualifications Significant engineering experience in trust and safety, platform integrity, fraud, or detection and response, including time as a technical lead or mentor Deep familiarity with account attack techniques Experience with authentication and identity systems, including OAuth, single sign-on, multi-factor authentication, device binding, and risk-based authentication Experience applying machine learning to fraud or abuse detection, alongside a clear sense of when simpler rules-based approaches are the better choice Experience with cloud data tooling such as BigQuery, Spark, dbt, Airflow or similar Experience building tooling for operational or investigative teams, and partnering closely with the people who use it Interest in AI safety, and in the specific ways account compromise intersects with model misuse Representative projects - Design the architecture for real-time login risk scoring, and get agreement across Security, Product, and Safeguards on where step-up authentication should and should not fire Design detection for compromised API keys, together with security controls to limit exposure Rebuild product features to regain access quickly when customers get compromised Write the threat model that sets the team's roadmap for the next year, and bring the rest of the organisation along with it Instrument a false positive review loop that measures how often the team's defences affect legitimate users, and drive that number down The annual compensation range for this role is listed below. For sales roles, the range provided is the role's On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: £325,000 - £390,000 GBP Logistics Minimum education: Bachelor's degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that the company recruiters only contact you email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of the company. Be cautious of emails from other domains. Legitimate the company recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links-visit the directly for confirmed position openings. How we're different - We're different We believe that the highest-impact AI research will be big science. At the company we work as a single cohesive team on just a few large-scale research efforts. And we value impact - advancing our long-term goals of steerable, trustworthy AI - rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to the company, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute . click apply for full job details
We're building something better This role is in NatWest Boxed At NatWest Boxed, we're a high performing team that's always on the lookout for exceptional people. We encourage an inclusive and diverse working environment, where teamwork, input and collaboration is valued no matter your level. Job description This role is based in the United Kingdom and as such all normal working days must be carried out in the United Kingdom. Join us as a Platform Engineering Manager Identifying and solve engineering-wide challenges as they appear We'll look to you to deliver exceptionally high-quality work and empower engineers to make a big impact Investing in NatWest Boxed's overall platform architecture Ensuring the security of our systems and working practice What you'll do As a Platform Engineering Manager, you will take ownership of our data ecosystem, guiding a high-performing engineering squad to architect and maintain a robust platform that empowers every business unit to effectively ingest, manage, and leverage data for varied strategic requirements. Day-to-day, you'll be: Leading, mentoring and growing a high-performing team of platform engineers, shaping our hiring strategy and fostering a culture of innovation and ownership Defining the strategic vision and technical roadmap for our data platform, focusing on self-service, reliability, scalability, security, and user experience Critically evaluating architectural proposals and technical designs from the team, ensuring alignment with long-term business goals, platform vision, architecture and security Championing engineering excellence by embedding best practices for self-service platform engineering, CI/CD, and SRE, while owning the operational health of our platform through SLOs Collaborating with Product, Architecture, Analytics and Security teams to deliver a secure, robust platform that enables all groups to easily manage data within defined standards Driving projects from inception to delivery, taking into consideration the needs of your customers whilst fitting into the wider Platform strategy Engaging with clients and vendors on integration patterns and implementation The skills you'll need We're looking for someone with p roven experience leading, managing, and scaling high-performing data mesh architectures as well as the ability to connect technical architecture decisions directly to business outcomes; you understand the cost, risk, and revenue impact. It would be great if you had experience with FinTech / Banking and hands-on experience with programming languages such as Python or Go. You'll also have: A strong technical background in cloud-native technologies (e.g., AWS, GCP) and event streaming (Kafka). Deep understanding of modern platform engineering principles, including infrastructure-as-code (e.g., Terraform, Crossplane), data streaming platforms, data transformation, observability (e.g., Datadog, OpenTelemetry) and self-service platform design A strategic mindset with experience designing, building, and operating scalable, distributed systems. Excellent communication and stakeholder management skills, with the ability to articulate complex technical concepts to diverse audiences internal and external
Aug 21, 2026
Full time
We're building something better This role is in NatWest Boxed At NatWest Boxed, we're a high performing team that's always on the lookout for exceptional people. We encourage an inclusive and diverse working environment, where teamwork, input and collaboration is valued no matter your level. Job description This role is based in the United Kingdom and as such all normal working days must be carried out in the United Kingdom. Join us as a Platform Engineering Manager Identifying and solve engineering-wide challenges as they appear We'll look to you to deliver exceptionally high-quality work and empower engineers to make a big impact Investing in NatWest Boxed's overall platform architecture Ensuring the security of our systems and working practice What you'll do As a Platform Engineering Manager, you will take ownership of our data ecosystem, guiding a high-performing engineering squad to architect and maintain a robust platform that empowers every business unit to effectively ingest, manage, and leverage data for varied strategic requirements. Day-to-day, you'll be: Leading, mentoring and growing a high-performing team of platform engineers, shaping our hiring strategy and fostering a culture of innovation and ownership Defining the strategic vision and technical roadmap for our data platform, focusing on self-service, reliability, scalability, security, and user experience Critically evaluating architectural proposals and technical designs from the team, ensuring alignment with long-term business goals, platform vision, architecture and security Championing engineering excellence by embedding best practices for self-service platform engineering, CI/CD, and SRE, while owning the operational health of our platform through SLOs Collaborating with Product, Architecture, Analytics and Security teams to deliver a secure, robust platform that enables all groups to easily manage data within defined standards Driving projects from inception to delivery, taking into consideration the needs of your customers whilst fitting into the wider Platform strategy Engaging with clients and vendors on integration patterns and implementation The skills you'll need We're looking for someone with p roven experience leading, managing, and scaling high-performing data mesh architectures as well as the ability to connect technical architecture decisions directly to business outcomes; you understand the cost, risk, and revenue impact. It would be great if you had experience with FinTech / Banking and hands-on experience with programming languages such as Python or Go. You'll also have: A strong technical background in cloud-native technologies (e.g., AWS, GCP) and event streaming (Kafka). Deep understanding of modern platform engineering principles, including infrastructure-as-code (e.g., Terraform, Crossplane), data streaming platforms, data transformation, observability (e.g., Datadog, OpenTelemetry) and self-service platform design A strategic mindset with experience designing, building, and operating scalable, distributed systems. Excellent communication and stakeholder management skills, with the ability to articulate complex technical concepts to diverse audiences internal and external
We don't just believe in better. We make it happen. Better content. Better products. And better careers. Working in Tech, Product or Data at Sky is about building the next and the new. From broadband to broadcast, streaming to mobile, Sky Stream to Sky Glass, we never stand still. We optimise and innovate. We turn big ideas into the products, content and services millions of people love. And we do it all right here at Sky. Sky is looking for a talented and dedicated Senior Software Engineer in RDK Triage Engineering team to support the release cycles associated with deployed software on CPE devices. These devices use Reference Design Kit (RDK), which is deployed on over 100+ million devices spanning video, broadband and home security. The Release Triage Software Engineer function is critical to the successful development and deployment of new features and fixes as the function ensures that these changes get deployed to the field. Sky's development environment is advanced and highly integrated. It uses industry-standard tools that are combined effectively to support a fast-moving, agile development cycle. The combination of these tools running on cloud infrastructure, coupled with effective use of Open-Source code, allows Sky to deliver features and products against aggressive timelines. What You'll Do: As a key engineer of the team, you will be responsible for leading on-time, high-quality releases across a large number of devices. You will also be responsible for the complete release quality and triage lifecycle, which includes deployment, triage, mitigation, and tool development for software release operations. The position will require daily collaboration with the Development, Release, and QA teams. You will assess and ensure the release quality of the RDK software with Key performance metrics as well as incidents from the field. Also, identify new tools, processes, etc., necessary to improve the software release triage engineering process. You will manage risks and resolve issues that affect release scope, schedule and quality. Your daily tasks: Investigate customer-impacting field issues, identify root cause, and guide development teams toward the right fix Use strong problem-solving and debugging skills to isolate faults and resolve issues quickly Monitor business intelligence and telemetry dashboards to catch emerging customer issues early and act on them Collaborate closely with internal engineering teams and external partners, communicating findings clearly Contribute to and follow Agile development practices Stay current with new tools, technologies, and best practices across development, data analysis, cloud platforms, and monitoring Help scope requirements and contribute ideas to improve triage tooling and workflows Support and mentor teammates, sharing knowledge and best practices as you grow What You'll Bring: Experience in software engineering, systems debugging, or technical support, with a strong problem-solving mindset Experience with Linux and at least one programming language (C, C++, Python, or similar) Comfort working with networking & video/digital TV concepts (WiFi, broadband, or video technologies a plus, but not required) Exposure to data analysis or monitoring tools (e.g., Elastic, Splunk, Datadog, or similar) is a plus Understanding of software development practices, version control, and Agile methodologies Strong communication skills and the ability to work well with cross-functional and global teams A curious, adaptable mindset - comfortable learning new tools and technologies on the job Team overview: The Global RDK Field Triage Team within the RDK development team plays a critical role in ensuring the quality and reliability of these products. The team supports software deployment on CPE (Customer Premises Equipment) devices, collaborates with the release and engineering teams during integration, and leads issue investigation and live data monitoring. Our mission is to ensure optimal product performance and customer satisfaction. This role involves close collaboration with colleagues across the UK, India, and the US, offering a dynamic and truly international working environment. Global Product We're the Global Product. We're the team behind your favourite Sky products, and the platforms that power them. We make every moment magical, everywhere. Our team is made up of self-motivated, big thinkers who have a knack for solving problems and find new ways to captivate millions of customers by putting them at the heart of everything we do. From Sky Glass, Sky Q, Peacock and NOW to news and sports apps, we make entertainment even better, and we can't wait to get started on what's next. Benefits and perks There's one thing people can't stop talking about when it comes to life at Sky: the perks . Here's a taster: Free Sky TV or NOW package, including Sky Sports and Sky Cinema Pension package with up to 9% employer contribution Private healthcare with mental health support Aviva Digital GP and dental insurance Discounts on Sky products, including Sky Mobile, Sky Broadband, Sky Glass and Sky Protect Sharesave and Tech schemes A range of Sky VIP rewards and experiences How you'll work We've adopted a hybrid working approach to give more flexibility on where and how we work. The hybrid working expectations for this role are 2 days in the office per week. Your office base Our Sky Group HQ. Equipped with state-of-the-art technology and workspaces, there's plenty of space to see your big ideas come to life. Here you'll find 13 subsidised restaurants and cafes. You can re-energise at our gym, catch the latest films at our cinema, get your car washed and even get pampered at our beauty salon . Our Osterley Campus is just a 10-minute walk from Syon Lane train station, or you can get one of our free shuttle buses from Osterley, Gunnersbury and Ealing Broadway stations. Plus, there's free onsite parking available for cars, motorbikes and bicycles. Who we are We're Sky, a leading media and entertainment company who connect millions with entertainment, sports, news and arts through innovative products and services. Working with us means you'll be bringing the joy of a better experience to more people, every day. All so we can do better and deliver better for our customers, colleagues and society. We're an equal opportunity employer and value diversity at our company. We're a Disability Confident Accredited Employer, and welcome and encourage applications from all candidates. We will look to ensure a fair and consistent experience for all and will make reasonable adjustments to support you where appropriate . Please flag any adjustments you need as early as you can. Just so you know: if your application is successful, we'll ask you to complete a criminal record check. And depending on the role you have applied for and the nature of any convictions you may have, we might have to withdraw the offer. To be eligible for this role you are required to have the appropriate right to work in the UK. Please be aware Sky does not offer sponsorship for this position. To find out more about working with us, search on social media.
Aug 20, 2026
Full time
We don't just believe in better. We make it happen. Better content. Better products. And better careers. Working in Tech, Product or Data at Sky is about building the next and the new. From broadband to broadcast, streaming to mobile, Sky Stream to Sky Glass, we never stand still. We optimise and innovate. We turn big ideas into the products, content and services millions of people love. And we do it all right here at Sky. Sky is looking for a talented and dedicated Senior Software Engineer in RDK Triage Engineering team to support the release cycles associated with deployed software on CPE devices. These devices use Reference Design Kit (RDK), which is deployed on over 100+ million devices spanning video, broadband and home security. The Release Triage Software Engineer function is critical to the successful development and deployment of new features and fixes as the function ensures that these changes get deployed to the field. Sky's development environment is advanced and highly integrated. It uses industry-standard tools that are combined effectively to support a fast-moving, agile development cycle. The combination of these tools running on cloud infrastructure, coupled with effective use of Open-Source code, allows Sky to deliver features and products against aggressive timelines. What You'll Do: As a key engineer of the team, you will be responsible for leading on-time, high-quality releases across a large number of devices. You will also be responsible for the complete release quality and triage lifecycle, which includes deployment, triage, mitigation, and tool development for software release operations. The position will require daily collaboration with the Development, Release, and QA teams. You will assess and ensure the release quality of the RDK software with Key performance metrics as well as incidents from the field. Also, identify new tools, processes, etc., necessary to improve the software release triage engineering process. You will manage risks and resolve issues that affect release scope, schedule and quality. Your daily tasks: Investigate customer-impacting field issues, identify root cause, and guide development teams toward the right fix Use strong problem-solving and debugging skills to isolate faults and resolve issues quickly Monitor business intelligence and telemetry dashboards to catch emerging customer issues early and act on them Collaborate closely with internal engineering teams and external partners, communicating findings clearly Contribute to and follow Agile development practices Stay current with new tools, technologies, and best practices across development, data analysis, cloud platforms, and monitoring Help scope requirements and contribute ideas to improve triage tooling and workflows Support and mentor teammates, sharing knowledge and best practices as you grow What You'll Bring: Experience in software engineering, systems debugging, or technical support, with a strong problem-solving mindset Experience with Linux and at least one programming language (C, C++, Python, or similar) Comfort working with networking & video/digital TV concepts (WiFi, broadband, or video technologies a plus, but not required) Exposure to data analysis or monitoring tools (e.g., Elastic, Splunk, Datadog, or similar) is a plus Understanding of software development practices, version control, and Agile methodologies Strong communication skills and the ability to work well with cross-functional and global teams A curious, adaptable mindset - comfortable learning new tools and technologies on the job Team overview: The Global RDK Field Triage Team within the RDK development team plays a critical role in ensuring the quality and reliability of these products. The team supports software deployment on CPE (Customer Premises Equipment) devices, collaborates with the release and engineering teams during integration, and leads issue investigation and live data monitoring. Our mission is to ensure optimal product performance and customer satisfaction. This role involves close collaboration with colleagues across the UK, India, and the US, offering a dynamic and truly international working environment. Global Product We're the Global Product. We're the team behind your favourite Sky products, and the platforms that power them. We make every moment magical, everywhere. Our team is made up of self-motivated, big thinkers who have a knack for solving problems and find new ways to captivate millions of customers by putting them at the heart of everything we do. From Sky Glass, Sky Q, Peacock and NOW to news and sports apps, we make entertainment even better, and we can't wait to get started on what's next. Benefits and perks There's one thing people can't stop talking about when it comes to life at Sky: the perks . Here's a taster: Free Sky TV or NOW package, including Sky Sports and Sky Cinema Pension package with up to 9% employer contribution Private healthcare with mental health support Aviva Digital GP and dental insurance Discounts on Sky products, including Sky Mobile, Sky Broadband, Sky Glass and Sky Protect Sharesave and Tech schemes A range of Sky VIP rewards and experiences How you'll work We've adopted a hybrid working approach to give more flexibility on where and how we work. The hybrid working expectations for this role are 2 days in the office per week. Your office base Our Sky Group HQ. Equipped with state-of-the-art technology and workspaces, there's plenty of space to see your big ideas come to life. Here you'll find 13 subsidised restaurants and cafes. You can re-energise at our gym, catch the latest films at our cinema, get your car washed and even get pampered at our beauty salon . Our Osterley Campus is just a 10-minute walk from Syon Lane train station, or you can get one of our free shuttle buses from Osterley, Gunnersbury and Ealing Broadway stations. Plus, there's free onsite parking available for cars, motorbikes and bicycles. Who we are We're Sky, a leading media and entertainment company who connect millions with entertainment, sports, news and arts through innovative products and services. Working with us means you'll be bringing the joy of a better experience to more people, every day. All so we can do better and deliver better for our customers, colleagues and society. We're an equal opportunity employer and value diversity at our company. We're a Disability Confident Accredited Employer, and welcome and encourage applications from all candidates. We will look to ensure a fair and consistent experience for all and will make reasonable adjustments to support you where appropriate . Please flag any adjustments you need as early as you can. Just so you know: if your application is successful, we'll ask you to complete a criminal record check. And depending on the role you have applied for and the nature of any convictions you may have, we might have to withdraw the offer. To be eligible for this role you are required to have the appropriate right to work in the UK. Please be aware Sky does not offer sponsorship for this position. To find out more about working with us, search on social media.
Lead Design Engineer - UX/UI and AI Role StarCompliance is seeking a Lead Design Engineer - AI Product Experience who will shape StarCompliance is seeking a Lead Design Engineer - AI Product Experience who will shape how user experiences are designed, prototyped, and delivered across our engineering organisation. This role is responsible for bringing together user experience design, frontend engineering, and AI-assisted development to create scalable product experience standards that can be consistently applied by both engineers and AI agents. This is not a traditional UX design or UI engineering role. The successful candidate willoperateas a highly technical, hands on design engineering leader focused on evolving our design system for agentic consumption, creating working prototypes in place of static design handovers, and ensuring that accessibility, usability, and consistency are embedded directly into our frontend foundations. The role will work closely with Software Engineering, Product, Architecture, Platform Engineering, and AI Enablement teams toestablishreusable components, interaction patterns, agent instructions, and governance that support faster, more consistent, and increasingly autonomous product delivery. How We Think About AI AtStarCompliance, AI is not a side experiment or an isolatedspecialistcapability. We treat it as a foundational part of modern software engineering and SaaS platform development. Actively use AI-assisted engineering tools as part of their daily workflows. Apply AI to improve design quality, development velocity, automation depth, operational insight, and overall engineering effectiveness. Ensure AI-generated outputs arevalidated, auditable, and compliant with security, privacy, and regulatory standards. Key Responsibilities Own the strategy, architecture, and ongoing evolution ofStarCompliance'senterprise design system, ensuring it supports consistent delivery across products and teams. Develop the design system into an AI-consumable platform that can be reliably used by engineers, coding assistants, and autonomous development agents. Define reusable components, design tokens, interaction patterns, application layouts, and accessibility standards. Create clear documentation, examples, constraints, and validation mechanisms that enable reliable agent-generated user interfaces. Maintain alignment between the design system, shared frontend foundations, andStarCompliance'swider technical architecture. Translate product opportunities, user needs, and business requirements into working prototypes rather than relying on static design handovers. Use code and AI-assisted development tools to create realistic product experiences that can be tested and refined with users. Work with Product, customers, subject matter experts, and engineers to explore workflows,validatesolutions, and improve concepts before production delivery. Define interaction models for complex compliance workflows, data intensive applications, and AI enabled product capabilities. Agentic Product Development Define how AI agents should interpret product intent and generate user interfaces using approved components and experience standards. Create andmaintainagent instructions, reference implementations, rules, and machine readable context for frontend development. Partner with AI Enablement and Engineering teams to embed design system knowledge into agentic development workflows. Establish governance for agent-generated interfaces, including accessibility validation, visual regression testing, andcomponentconformance. Review agent outputs and improve the components, guidance, and contextrequiredto produce reliable results. Product Experience & Accessibility Ensure user centric design principlesremainembedded within an increasingly automated software development process. Champion accessible and inclusive design, ensuring accessibility requirements are built into components, patterns, and automated validation. Use customer feedback, product analytics, usability evaluation, and telemetry to improve productexperiences. Ensure delivery speed does not come at the expense of usability, consistency, maintainability, or user trust. Technical Leadership & Governance Lead the Design System Forum, bringing together Product, Engineering, and Architecture to govern shared standards andprioritisethe continued evolution of the design system. Lead and develop a blended team of UI Engineers across permanent and consultancy partner resources, with responsibility for performance, workload allocation, delivery, and capability development. Act asStarCompliance'stechnical authority for design engineering and frontend experience,providinghands on guidance andrepresentingthe discipline within relevant governance forums. Coach Product Managers and engineers in prototype led discovery, AI assisted development, and the effective use of shared frontend foundations. Skills and Experience Strong background in design engineering, frontend engineering, or user experience design. Hands on experience building modern React and TypeScript applications. Experience creating or governing enterprise design systems, includingcomponentarchitecture, design tokens, and interaction patterns. Practical experience using AI assisted development tools and supporting agentic software delivery workflows. Strong understanding of accessibility, inclusive design, and working prototype development. Strong stakeholder communication, technical leadership,and peoplemanagement capability. Experience within SaaS, enterprise, or regulated software environments. Knowledge of Storybook, visual regression testing, automated accessibility testing, andfrontendquality tooling. Experience managing permanent employees and consultancy partner resources. Familiarity with product analytics, usability testing, andbehaviouraltelemetry. StarCompliance Background Checks All positions require pre employment screening due to employees potentially having access to highly sensitive and confidential information involving finance and compliance; candidates must be trustworthy and have a heightened sensitivity to protecting confidential financial, professional information. To be eligible for employment with StarCompliance, candidates must undergo a rigorous background investigation with checks including, but not limited to, criminal record history, consumer credit, employment history, qualifications, and education checks. Equal Opportunity Employer Statement We prohibit discrimination and harassment of any kind based on race, sex, religion, sexual orientation, national origin, disability, genetic information, pregnancy, gender identity or expression, marital/civil union/domestic partnership status, veteran status or any other protected characteristic as outlined by country, state, or local laws. This policy applies to all employment practices within our organisation, including hiring, recruiting, promotion, termination, layoff, recall, leave of absence, compensation, benefits, training, and apprenticeship. StarCompliance makes hiring decisions based solely on qualifications, merit, and business needs at the time. For more information, please request a copy of our Equal Opportunities Policy.
Jul 31, 2026
Full time
Lead Design Engineer - UX/UI and AI Role StarCompliance is seeking a Lead Design Engineer - AI Product Experience who will shape StarCompliance is seeking a Lead Design Engineer - AI Product Experience who will shape how user experiences are designed, prototyped, and delivered across our engineering organisation. This role is responsible for bringing together user experience design, frontend engineering, and AI-assisted development to create scalable product experience standards that can be consistently applied by both engineers and AI agents. This is not a traditional UX design or UI engineering role. The successful candidate willoperateas a highly technical, hands on design engineering leader focused on evolving our design system for agentic consumption, creating working prototypes in place of static design handovers, and ensuring that accessibility, usability, and consistency are embedded directly into our frontend foundations. The role will work closely with Software Engineering, Product, Architecture, Platform Engineering, and AI Enablement teams toestablishreusable components, interaction patterns, agent instructions, and governance that support faster, more consistent, and increasingly autonomous product delivery. How We Think About AI AtStarCompliance, AI is not a side experiment or an isolatedspecialistcapability. We treat it as a foundational part of modern software engineering and SaaS platform development. Actively use AI-assisted engineering tools as part of their daily workflows. Apply AI to improve design quality, development velocity, automation depth, operational insight, and overall engineering effectiveness. Ensure AI-generated outputs arevalidated, auditable, and compliant with security, privacy, and regulatory standards. Key Responsibilities Own the strategy, architecture, and ongoing evolution ofStarCompliance'senterprise design system, ensuring it supports consistent delivery across products and teams. Develop the design system into an AI-consumable platform that can be reliably used by engineers, coding assistants, and autonomous development agents. Define reusable components, design tokens, interaction patterns, application layouts, and accessibility standards. Create clear documentation, examples, constraints, and validation mechanisms that enable reliable agent-generated user interfaces. Maintain alignment between the design system, shared frontend foundations, andStarCompliance'swider technical architecture. Translate product opportunities, user needs, and business requirements into working prototypes rather than relying on static design handovers. Use code and AI-assisted development tools to create realistic product experiences that can be tested and refined with users. Work with Product, customers, subject matter experts, and engineers to explore workflows,validatesolutions, and improve concepts before production delivery. Define interaction models for complex compliance workflows, data intensive applications, and AI enabled product capabilities. Agentic Product Development Define how AI agents should interpret product intent and generate user interfaces using approved components and experience standards. Create andmaintainagent instructions, reference implementations, rules, and machine readable context for frontend development. Partner with AI Enablement and Engineering teams to embed design system knowledge into agentic development workflows. Establish governance for agent-generated interfaces, including accessibility validation, visual regression testing, andcomponentconformance. Review agent outputs and improve the components, guidance, and contextrequiredto produce reliable results. Product Experience & Accessibility Ensure user centric design principlesremainembedded within an increasingly automated software development process. Champion accessible and inclusive design, ensuring accessibility requirements are built into components, patterns, and automated validation. Use customer feedback, product analytics, usability evaluation, and telemetry to improve productexperiences. Ensure delivery speed does not come at the expense of usability, consistency, maintainability, or user trust. Technical Leadership & Governance Lead the Design System Forum, bringing together Product, Engineering, and Architecture to govern shared standards andprioritisethe continued evolution of the design system. Lead and develop a blended team of UI Engineers across permanent and consultancy partner resources, with responsibility for performance, workload allocation, delivery, and capability development. Act asStarCompliance'stechnical authority for design engineering and frontend experience,providinghands on guidance andrepresentingthe discipline within relevant governance forums. Coach Product Managers and engineers in prototype led discovery, AI assisted development, and the effective use of shared frontend foundations. Skills and Experience Strong background in design engineering, frontend engineering, or user experience design. Hands on experience building modern React and TypeScript applications. Experience creating or governing enterprise design systems, includingcomponentarchitecture, design tokens, and interaction patterns. Practical experience using AI assisted development tools and supporting agentic software delivery workflows. Strong understanding of accessibility, inclusive design, and working prototype development. Strong stakeholder communication, technical leadership,and peoplemanagement capability. Experience within SaaS, enterprise, or regulated software environments. Knowledge of Storybook, visual regression testing, automated accessibility testing, andfrontendquality tooling. Experience managing permanent employees and consultancy partner resources. Familiarity with product analytics, usability testing, andbehaviouraltelemetry. StarCompliance Background Checks All positions require pre employment screening due to employees potentially having access to highly sensitive and confidential information involving finance and compliance; candidates must be trustworthy and have a heightened sensitivity to protecting confidential financial, professional information. To be eligible for employment with StarCompliance, candidates must undergo a rigorous background investigation with checks including, but not limited to, criminal record history, consumer credit, employment history, qualifications, and education checks. Equal Opportunity Employer Statement We prohibit discrimination and harassment of any kind based on race, sex, religion, sexual orientation, national origin, disability, genetic information, pregnancy, gender identity or expression, marital/civil union/domestic partnership status, veteran status or any other protected characteristic as outlined by country, state, or local laws. This policy applies to all employment practices within our organisation, including hiring, recruiting, promotion, termination, layoff, recall, leave of absence, compensation, benefits, training, and apprenticeship. StarCompliance makes hiring decisions based solely on qualifications, merit, and business needs at the time. For more information, please request a copy of our Equal Opportunities Policy.
Director Data & Analytics API Experience - Data ServicesSkip to main contentBy clicking "Accept All Cookies," you agree to the storing of cookies on your device to enhance site navigation, analyze site usage, and assist in our marketing efforts.# CareersDirector Data & Analytics API Experience - Data Services page is loaded Director Data & Analytics API Experience - Data ServicesApplylocations: Belfast - Millennium Housetime type: Full timeposted on: Posted Todayjob requisition id: 34591 Director Data & Analytics API Experience - Data Services Background: CME Group runs some of the world's largest and most important trading venues, including derivatives exchanges, FX/FI platforms. The trading participants include leading investment banks, hedge funds, proprietary trading firms, investment managers, and retail traders, as well as additional stakeholders such as independent software vendors (ISVs) and data redistributors. To support our unique data and analytics products, we are now building a unified API-first distribution framework, simplifying consumption and value extraction. The Role: We are seeking a key stakeholder for the CMEG External Data-Analytics (DA) APIs and distribution services. This is a high-impact, individual contributor (IC) role for an individual who can navigate complex enterprise systems and deliver a modern, developer-first experience. This customer-facing role serves as the technical link between our APIs and the customer, acting as a high-level practitioner who crafts SDKs and delivery mechanisms to make complex market data easy to consume. You will translate direct customer feedback into refined requirements for the product team while solving integration hurdles firsthand. By bridging these gaps, you will ensure our technical ecosystem is both highly functional and seamless for clients to adopt. Our API First Vision Our external API is the singular point where our strategy, technology, and customer experience converge to create value. Our corporate DA strategy focuses on "activation," which involves externalizing proven internal IP and exposing it through a simple, consumable API, as well as adding net-new IP jointly developed with our customers. The API is our "single, unified distribution channel". It serves as the "handover point" where internal engineering ends and the customer or partner ecosystem begins. The API provides a "single 'front door'" for customers. By using a "unified API", barriers to consumption are removed, integration time is reduced, development costs lowered, and a consistent experience provided across all data latencies (Real-time, Recent History, and Deep History), data types (time series, metadata, unstructured) and services. Internal users are recognized as ultra-valuable sources of feedback. By mandating both internal and external DA workflows through the same SDK, we create a shared experience for both sets of users. In summary, our API is the "final product" and the "primary tool" moving CMEG up the value curve from a raw data producer to an invaluable source of insights for our customers. What you Will be Doing: The focus of the work is on activation and will be organized across four strategic pillars designed to continue to build the foundation, enhance the customer experience, and align all our public offerings. Developer Experience + "North Star" Activation: Build a high-performance Python SDK that mimics the industry's best "one-line-to-insight" user experiences. Your goal is to move the user from installation to action in seconds. + Lighthouse Deliverable: This SDK will serve as the technical and UX benchmark for how all future CMEG IP is externalized and consumed. It will be the first deliverable and form your MVP. + Internal User Parity: The SDK must provide a 'single front door' for internal users, allowing them to 'pip install' and access BigQuery datasets through the same idiomatic Python patterns designed for customers. + Release Open-Source SDKs: Develop and maintain a suite of SDKs in popular languages like Python and C++ to make integrating with our APIs as simple as possible. + Public Schema Registry: Implement a publicly accessible registry (e.g., Protobuf, JSON or Avro) that allow internal teams and partners to generate their own client software. Developer Collateral + Simplify Public Messaging: Partner with internal stakeholders to clearly communicate CME Group's D&A value proposition. + Technical Documentation: Evaluate, implement and maintain an API documentation library and migrate existing fragmented solutions to it. + Publish Reference Docs: Create first-class, detailed technical reference documentation that serves as the definitive contract for all API behaviour and usage. + Accelerate Time-to-Value: Develop practical tutorials, code samples, and dashboard templates to empower customers to self-serve and validate products rapidly. Data Plane + Expand Cloud Distribution: Work with our customers and product teams to extend our existing data offerings to meet customer cloud demand. Abstracting cloud specific technologies away and provisioning a common messaging layer. + Abstraction Layers: Work with our customers to better understand if/how we can improve data delivery for deep history and data sharing services so that accessing large datasets feels identical to a standard, programmatic API call + Enforce Schema Consistency: Use a centralized schema registry to ensure identical data models and field names across different services and latency tiers. Control Plane + Normalize Authentication: Work with our central technology team to ensure a unified CMEG credential system that retires legacy approaches and hides cloud-specific complexity from the user. + Downstream Distribution Technologies: Support external facing distribution technologies to provision products, including Bobsled, SFTP, Curl, email, etc. + Implement Group Wide Telemetry Tracking: Work with business intelligence teams to build the capabilities to track user activity With the goal to generate insights into user activity and product value. + Full-Stack Execution: Be a key stakeholder in the "plumbing" of the API lifecycle, including GitHub governance, and CI/CD pipelines for PyPi publishing. Unified API Strategy + Technical audit: Work with stakeholders across the organization to conduct a survey of existing APIs (RPC over HTTP, JSON, XML Query, REST, WebSocket, S3, Pub-Sub, etc) across the business, to identify fragmentation, authentication gaps, and areas for consolidation. Survey peers and partners to identify best-practice in API design and user experience. + Establish Versioning Models: Work with stakeholders to define and document the hybrid versioning model to manage API updates while ensuring stability for partners and customers. Time Line Major milestones for the work are split into three, Short (1yr): Publish the MVP of the SDK on github and PyPI, including support for Looker dashboards and other upstream uses. Medium (2-3yr). Design and publish a normalized multi cloud distribution framework. Long (3-5yrs): Work as a stakeholder across the organization to streamline the process of normalizing APIs across CME Group Data Services and the wider organization. Who are we looking for? Exceptional Achievers & Problem Solvers: We are looking for high-capacity individuals who are driven to be elite at what they do. You must possess a strong technical foundation and a genuine interest in the mechanics of capital markets-understanding how data is used through futures, options, and cash markets. API-First Mindset: You are obsessed with developer experience (DX). You believe the API is the product and should be as easy to use as a "pip install" command. API-First Mindset: You are obsessed with developer experience (DX). You believe the API is the product and should be as easy to use as a "pip install" command. Cloud-Native Builders: Experience using, designing, and building SDKs and APIs in the cloud Polymathic Ability: You are comfortable bridging the gap between business strategy, engineering, and technical writing. Python Expertise: Highly proficient in Python, with the ability to write clean, idiomatic, and robust code for public consumption. Technical Stack Environment: Google Cloud Platform (GCP), Apigee. Protocols: WebSockets, REST, GCS. Tooling: Python (Expert), OAuth 2.0, GitHub, Protobuf/Avro, SQL. Company Benefits: Bonus Programme Equity Programme Employee Stock Purchase Plan (ESPP) Private Medical and Dental coverage Mental Health Benefit Programme Group Pension Plan Income Protection Life Assurance Cycle To Work EV Car Benefit Scheme Gym Membership Family Leave Education Assistance - MBA/Advanced Degree/Bachelor Degree Ongoing Employee Development Training/Certification Hybrid CME Group: Where Futures are Made CME Group is the world's leading derivatives marketplace. But who we are goes deeper than that. Here, you can impact markets worldwide. Transform industries. And build a career by shaping tomorrow. We invest in your success and you own it - all while working alongside a team of leading experts who inspire you in ways big and small. Problem solvers, difference makers, trailblazers. Those are our people. And we're looking for more.At CME Group, we embrace our employees' unique experiences and skills to ensure that everyone's perspectives are acknowledged and valued. As an equal-opportunity employer, we consider all potential employees without regard to any protected characteristic. Important Notice: Recruitment fraud is on the rise, with scammers using misleading promises of job offers and interviews to solicit money and personal information from job seekers. CME Group adheres to established procedures designed to maintain trust, confidence and security throughout our recruitment process. Learn more here.
Jul 31, 2026
Full time
Director Data & Analytics API Experience - Data ServicesSkip to main contentBy clicking "Accept All Cookies," you agree to the storing of cookies on your device to enhance site navigation, analyze site usage, and assist in our marketing efforts.# CareersDirector Data & Analytics API Experience - Data Services page is loaded Director Data & Analytics API Experience - Data ServicesApplylocations: Belfast - Millennium Housetime type: Full timeposted on: Posted Todayjob requisition id: 34591 Director Data & Analytics API Experience - Data Services Background: CME Group runs some of the world's largest and most important trading venues, including derivatives exchanges, FX/FI platforms. The trading participants include leading investment banks, hedge funds, proprietary trading firms, investment managers, and retail traders, as well as additional stakeholders such as independent software vendors (ISVs) and data redistributors. To support our unique data and analytics products, we are now building a unified API-first distribution framework, simplifying consumption and value extraction. The Role: We are seeking a key stakeholder for the CMEG External Data-Analytics (DA) APIs and distribution services. This is a high-impact, individual contributor (IC) role for an individual who can navigate complex enterprise systems and deliver a modern, developer-first experience. This customer-facing role serves as the technical link between our APIs and the customer, acting as a high-level practitioner who crafts SDKs and delivery mechanisms to make complex market data easy to consume. You will translate direct customer feedback into refined requirements for the product team while solving integration hurdles firsthand. By bridging these gaps, you will ensure our technical ecosystem is both highly functional and seamless for clients to adopt. Our API First Vision Our external API is the singular point where our strategy, technology, and customer experience converge to create value. Our corporate DA strategy focuses on "activation," which involves externalizing proven internal IP and exposing it through a simple, consumable API, as well as adding net-new IP jointly developed with our customers. The API is our "single, unified distribution channel". It serves as the "handover point" where internal engineering ends and the customer or partner ecosystem begins. The API provides a "single 'front door'" for customers. By using a "unified API", barriers to consumption are removed, integration time is reduced, development costs lowered, and a consistent experience provided across all data latencies (Real-time, Recent History, and Deep History), data types (time series, metadata, unstructured) and services. Internal users are recognized as ultra-valuable sources of feedback. By mandating both internal and external DA workflows through the same SDK, we create a shared experience for both sets of users. In summary, our API is the "final product" and the "primary tool" moving CMEG up the value curve from a raw data producer to an invaluable source of insights for our customers. What you Will be Doing: The focus of the work is on activation and will be organized across four strategic pillars designed to continue to build the foundation, enhance the customer experience, and align all our public offerings. Developer Experience + "North Star" Activation: Build a high-performance Python SDK that mimics the industry's best "one-line-to-insight" user experiences. Your goal is to move the user from installation to action in seconds. + Lighthouse Deliverable: This SDK will serve as the technical and UX benchmark for how all future CMEG IP is externalized and consumed. It will be the first deliverable and form your MVP. + Internal User Parity: The SDK must provide a 'single front door' for internal users, allowing them to 'pip install' and access BigQuery datasets through the same idiomatic Python patterns designed for customers. + Release Open-Source SDKs: Develop and maintain a suite of SDKs in popular languages like Python and C++ to make integrating with our APIs as simple as possible. + Public Schema Registry: Implement a publicly accessible registry (e.g., Protobuf, JSON or Avro) that allow internal teams and partners to generate their own client software. Developer Collateral + Simplify Public Messaging: Partner with internal stakeholders to clearly communicate CME Group's D&A value proposition. + Technical Documentation: Evaluate, implement and maintain an API documentation library and migrate existing fragmented solutions to it. + Publish Reference Docs: Create first-class, detailed technical reference documentation that serves as the definitive contract for all API behaviour and usage. + Accelerate Time-to-Value: Develop practical tutorials, code samples, and dashboard templates to empower customers to self-serve and validate products rapidly. Data Plane + Expand Cloud Distribution: Work with our customers and product teams to extend our existing data offerings to meet customer cloud demand. Abstracting cloud specific technologies away and provisioning a common messaging layer. + Abstraction Layers: Work with our customers to better understand if/how we can improve data delivery for deep history and data sharing services so that accessing large datasets feels identical to a standard, programmatic API call + Enforce Schema Consistency: Use a centralized schema registry to ensure identical data models and field names across different services and latency tiers. Control Plane + Normalize Authentication: Work with our central technology team to ensure a unified CMEG credential system that retires legacy approaches and hides cloud-specific complexity from the user. + Downstream Distribution Technologies: Support external facing distribution technologies to provision products, including Bobsled, SFTP, Curl, email, etc. + Implement Group Wide Telemetry Tracking: Work with business intelligence teams to build the capabilities to track user activity With the goal to generate insights into user activity and product value. + Full-Stack Execution: Be a key stakeholder in the "plumbing" of the API lifecycle, including GitHub governance, and CI/CD pipelines for PyPi publishing. Unified API Strategy + Technical audit: Work with stakeholders across the organization to conduct a survey of existing APIs (RPC over HTTP, JSON, XML Query, REST, WebSocket, S3, Pub-Sub, etc) across the business, to identify fragmentation, authentication gaps, and areas for consolidation. Survey peers and partners to identify best-practice in API design and user experience. + Establish Versioning Models: Work with stakeholders to define and document the hybrid versioning model to manage API updates while ensuring stability for partners and customers. Time Line Major milestones for the work are split into three, Short (1yr): Publish the MVP of the SDK on github and PyPI, including support for Looker dashboards and other upstream uses. Medium (2-3yr). Design and publish a normalized multi cloud distribution framework. Long (3-5yrs): Work as a stakeholder across the organization to streamline the process of normalizing APIs across CME Group Data Services and the wider organization. Who are we looking for? Exceptional Achievers & Problem Solvers: We are looking for high-capacity individuals who are driven to be elite at what they do. You must possess a strong technical foundation and a genuine interest in the mechanics of capital markets-understanding how data is used through futures, options, and cash markets. API-First Mindset: You are obsessed with developer experience (DX). You believe the API is the product and should be as easy to use as a "pip install" command. API-First Mindset: You are obsessed with developer experience (DX). You believe the API is the product and should be as easy to use as a "pip install" command. Cloud-Native Builders: Experience using, designing, and building SDKs and APIs in the cloud Polymathic Ability: You are comfortable bridging the gap between business strategy, engineering, and technical writing. Python Expertise: Highly proficient in Python, with the ability to write clean, idiomatic, and robust code for public consumption. Technical Stack Environment: Google Cloud Platform (GCP), Apigee. Protocols: WebSockets, REST, GCS. Tooling: Python (Expert), OAuth 2.0, GitHub, Protobuf/Avro, SQL. Company Benefits: Bonus Programme Equity Programme Employee Stock Purchase Plan (ESPP) Private Medical and Dental coverage Mental Health Benefit Programme Group Pension Plan Income Protection Life Assurance Cycle To Work EV Car Benefit Scheme Gym Membership Family Leave Education Assistance - MBA/Advanced Degree/Bachelor Degree Ongoing Employee Development Training/Certification Hybrid CME Group: Where Futures are Made CME Group is the world's leading derivatives marketplace. But who we are goes deeper than that. Here, you can impact markets worldwide. Transform industries. And build a career by shaping tomorrow. We invest in your success and you own it - all while working alongside a team of leading experts who inspire you in ways big and small. Problem solvers, difference makers, trailblazers. Those are our people. And we're looking for more.At CME Group, we embrace our employees' unique experiences and skills to ensure that everyone's perspectives are acknowledged and valued. As an equal-opportunity employer, we consider all potential employees without regard to any protected characteristic. Important Notice: Recruitment fraud is on the rise, with scammers using misleading promises of job offers and interviews to solicit money and personal information from job seekers. CME Group adheres to established procedures designed to maintain trust, confidence and security throughout our recruitment process. Learn more here.
The Role We're looking for an experienced Irrigation Manager to lead the planning, operation and continuous improvement of irrigation across our large-scale vegetable farming operation. Reporting to the relevant Farm Director, you'll play a key role in ensuring our crops receive the right water and nutrient applications to maximise yield, quality and sustainability. You'll oversee our irrigation infrastructure, lead a dedicated team, and work closely with Agronomy, Production and Engineering teams to optimise water use, improve system performance and support the delivery of high-quality crops. This is an excellent opportunity for someone with strong technical knowledge, leadership skills and a passion for modern agriculture to make a real impact within one of the UK's leading fresh produce businesses. What You'll Be Doing Develop and manage irrigation programmes based on crop requirements, soil moisture, weather conditions and agronomic recommendations. Monitor and optimise water usage to maximise crop performance, quality and resource efficiency. Oversee the operation, maintenance and continuous improvement of irrigation infrastructure, including pumps, reservoirs, pipelines, filtration, automated systems and fertigation equipment. Manage water abstraction, storage and distribution while ensuring compliance with environmental regulations and water licences. Lead, coach and develop the irrigation team, planning workloads and supporting continuous improvement. Analyse irrigation data, soil moisture monitoring and telemetry to improve system performance and efficiency. Coordinate preventative maintenance programmes, contractor activities and accurate operational reporting. Ensure all irrigation activities comply with health, safety, environmental and food safety standards. What We're Looking For Degree, diploma or equivalent qualification in Agriculture, Agricultural Engineering, Irrigation Technology, Agronomy or a related discipline. At least 5 years' experience managing irrigation systems within commercial vegetable, horticultural or large-scale farming operations. Experience managing irrigation systems within commercial vegetable, horticultural or large-scale farming operations. Strong knowledge of irrigation scheduling, crop water requirements, fertigation and water management. Experience with irrigation infrastructure, including pumps, filtration systems and automated controls. Proven leadership skills with experience managing teams and contractors. Excellent planning, organisational and problem-solving skills. Good IT skills, including Microsoft Office and irrigation management software. A proactive, continuous improvement mindset with a passion for sustainable farming. Technical Experience Irrigation scheduling, water management and fertigation systems. Drip irrigation, pump stations, reservoirs, filtration and distribution systems. Automated irrigation controls, telemetry, soil moisture sensors and weather monitoring technology. Water abstraction, environmental compliance and resource management. Preventative maintenance, asset management and performance reporting. Why work for Barfoots? Investors In People Silver Award status 24/7 On-line GP Company pension scheme Life Assurance Employee Assistance Program Benefits Platform Development opportunities Discounted leisure membership Discounted vegetable box scheme Cycle to work scheme Free onsite parking Approved training centre for Highfield qualifications Rapidly growing company Committed to Sustainability - Using our waste to power our site!
Jul 30, 2026
Full time
The Role We're looking for an experienced Irrigation Manager to lead the planning, operation and continuous improvement of irrigation across our large-scale vegetable farming operation. Reporting to the relevant Farm Director, you'll play a key role in ensuring our crops receive the right water and nutrient applications to maximise yield, quality and sustainability. You'll oversee our irrigation infrastructure, lead a dedicated team, and work closely with Agronomy, Production and Engineering teams to optimise water use, improve system performance and support the delivery of high-quality crops. This is an excellent opportunity for someone with strong technical knowledge, leadership skills and a passion for modern agriculture to make a real impact within one of the UK's leading fresh produce businesses. What You'll Be Doing Develop and manage irrigation programmes based on crop requirements, soil moisture, weather conditions and agronomic recommendations. Monitor and optimise water usage to maximise crop performance, quality and resource efficiency. Oversee the operation, maintenance and continuous improvement of irrigation infrastructure, including pumps, reservoirs, pipelines, filtration, automated systems and fertigation equipment. Manage water abstraction, storage and distribution while ensuring compliance with environmental regulations and water licences. Lead, coach and develop the irrigation team, planning workloads and supporting continuous improvement. Analyse irrigation data, soil moisture monitoring and telemetry to improve system performance and efficiency. Coordinate preventative maintenance programmes, contractor activities and accurate operational reporting. Ensure all irrigation activities comply with health, safety, environmental and food safety standards. What We're Looking For Degree, diploma or equivalent qualification in Agriculture, Agricultural Engineering, Irrigation Technology, Agronomy or a related discipline. At least 5 years' experience managing irrigation systems within commercial vegetable, horticultural or large-scale farming operations. Experience managing irrigation systems within commercial vegetable, horticultural or large-scale farming operations. Strong knowledge of irrigation scheduling, crop water requirements, fertigation and water management. Experience with irrigation infrastructure, including pumps, filtration systems and automated controls. Proven leadership skills with experience managing teams and contractors. Excellent planning, organisational and problem-solving skills. Good IT skills, including Microsoft Office and irrigation management software. A proactive, continuous improvement mindset with a passion for sustainable farming. Technical Experience Irrigation scheduling, water management and fertigation systems. Drip irrigation, pump stations, reservoirs, filtration and distribution systems. Automated irrigation controls, telemetry, soil moisture sensors and weather monitoring technology. Water abstraction, environmental compliance and resource management. Preventative maintenance, asset management and performance reporting. Why work for Barfoots? Investors In People Silver Award status 24/7 On-line GP Company pension scheme Life Assurance Employee Assistance Program Benefits Platform Development opportunities Discounted leisure membership Discounted vegetable box scheme Cycle to work scheme Free onsite parking Approved training centre for Highfield qualifications Rapidly growing company Committed to Sustainability - Using our waste to power our site!
We are looking for an experienced OT Test Lead to lead testing activities across a large Operational Technology and Telemetry programme. You will be responsible for planning, coordinating, and delivering testing across OT infrastructure, industrial control systems, and supporting technologies, ensuring solutions are fully validated before deployment. Key Responsibilities Develop and manage the OT test strategy and test plans. Lead System Integration Testing (SIT), End-to-End (E2E), Factory Acceptance Testing (FAT), Operational Acceptance Testing (OAT), User Acceptance Testing (UAT), and regression testing. Coordinate testing across engineering, infrastructure, cybersecurity, operational teams, and third-party suppliers. Validate PLC, SCADA, HMI, telemetry, and enterprise system integrations. Manage defects, test reporting, and test governance. Support operational readiness, resilience, recovery, and cybersecurity testing. Essential Experience Proven experience leading testing within Operational Technology (OT) environments. Strong knowledge of PLC, SCADA, DCS, HMI, ICS, or similar industrial control systems. Experience developing test strategies, plans, and test scripts. Strong understanding of OT infrastructure, telemetry, and system integration. Experience managing defects and coordinating multidisciplinary teams. Excellent stakeholder management and communication skills.
Jul 30, 2026
Contractor
We are looking for an experienced OT Test Lead to lead testing activities across a large Operational Technology and Telemetry programme. You will be responsible for planning, coordinating, and delivering testing across OT infrastructure, industrial control systems, and supporting technologies, ensuring solutions are fully validated before deployment. Key Responsibilities Develop and manage the OT test strategy and test plans. Lead System Integration Testing (SIT), End-to-End (E2E), Factory Acceptance Testing (FAT), Operational Acceptance Testing (OAT), User Acceptance Testing (UAT), and regression testing. Coordinate testing across engineering, infrastructure, cybersecurity, operational teams, and third-party suppliers. Validate PLC, SCADA, HMI, telemetry, and enterprise system integrations. Manage defects, test reporting, and test governance. Support operational readiness, resilience, recovery, and cybersecurity testing. Essential Experience Proven experience leading testing within Operational Technology (OT) environments. Strong knowledge of PLC, SCADA, DCS, HMI, ICS, or similar industrial control systems. Experience developing test strategies, plans, and test scripts. Strong understanding of OT infrastructure, telemetry, and system integration. Experience managing defects and coordinating multidisciplinary teams. Excellent stakeholder management and communication skills.
What Are We Looking For? Due to continued growth, RSE is looking for a Senior Commissioning Engineer to join us on a permanent basis, assisting in the delivery of our projects across Scotland. You will undertake Control, Instrumentation and Electrical works including installation, testing, commissioning, and fault finding within new and existing MCC s, industrial instrumentation, analysers, motors, actuators and all other equipment used in the automated control systems associated with RSE products and those used by clients in Water industry. Please note this position will require flexibility on working location, with accommodation provided by RSE. Some of Your Key Duties Include: Provide a flexible approach to resourcing all commissioning requirements of various projects To support the project team by developing commissioning plans, site commissioning files, methodologies, programmes, method statements and risk assessments Monitor standards to ensure specification compliance and quality standards are strictly maintained Attend FAT activities and provide constructive input with regards to system controls Carry out commissioning of LV Electrical systems, Instrumentation and Control Systems including MCC s and associated field equipment Read and interpret electrical drawings, line diagrams and P&ID s Complete SAT and Asset Start-up Complete commissioning documentation Participate in HAZCOMM activities for various projects as required What Do You Need? Electrical/Instrumentation HND/HNC or equivalent Proven Instrumentation & Electrical experience full commissioning phase experience (installation, dry, wet, process testing) Previous experience holding SAP or AP position and of PTW systems Experience of iMCC s including VSD s, Softstarts and Flow, Level, Pressure and Temperature instrumentation Experience in FAT, SAT testing and fault finding Experience in telemetry end to end testing Profibus, Installation and Commissioning Certification Full UK Driving Licence. Who Are We? RSE is a trusted clean water technology company, developing market-leading products and solutions for purifying drinking water, recycling wastewater, and cleaning water in industrial processes. We are disrupting the water sector, delivering water treatment products, technologies, and services to clients across the UK. RSE provides offsite modular build solutions using a low-carbon approach compared to traditional construction methods and our unique offering to the market focuses on innovation, efficiency, and excellence. Established in 1982, RSE has grown into one of the most prominent MEICA engineering businesses in the UK water industry. We have created a complete in-house and full-service capability from project inception through to design, fabrication, and delivery by means of installation and commissioning. We additionally have one of the largest servicing and maintenance teams in the market, to ensure we re on hand for all our clients needs. Our service offering presents industry-leading innovative solutions and our dedicated staff play a key role in delivering our sustainability and wider business goals. With over 2000 staff across our group of companies, our strategic ambition will see the business continue to grow as we expand our operations and diversify our products. One of RSE s key focuses is driving servant leadership and giving our people the opportunity and responsibility to take an entrepreneurial approach in their career development. What RSE Offer To build successful teams and drive the level of quality that RSE is renowned for, we know we need the best people in the industry. Not only do we require the relevant skillsets, but we also need people with the right attitude and mentality to thrive and grow in an innovative and fast-paced environment. At RSE, you ll be given every opportunity to set the path of your own career through our Business Streams and work within dynamic teams that will require you to rise to the challenge of working for a market leader. Industry-leading salary based on your experience. Company Van. A flexible career development path, with no restrictions on where your career can go. Private Healthcare (Personal). Holiday Allowance of 31 days per year, rising to 33 days per year after 2 years service. Holiday Buy / Sell Scheme Company Pension Scheme Cycle to Work Discounted National Gym Membership Professional Fees Paid Employee Discount Platform EV/Hybrid Car Lease Scheme Access to our network of health professionals including mental health champions and Occupational Health Nurse. In a flourishing sector where there are vast career opportunities available, we believe by leading transformation in the industry our offering to the market means our people have the space to thrive. If you re interested in a career with a company that will harness your skills and provide you with the support to create your own future within the water industry, apply now.
Jun 02, 2026
Full time
What Are We Looking For? Due to continued growth, RSE is looking for a Senior Commissioning Engineer to join us on a permanent basis, assisting in the delivery of our projects across Scotland. You will undertake Control, Instrumentation and Electrical works including installation, testing, commissioning, and fault finding within new and existing MCC s, industrial instrumentation, analysers, motors, actuators and all other equipment used in the automated control systems associated with RSE products and those used by clients in Water industry. Please note this position will require flexibility on working location, with accommodation provided by RSE. Some of Your Key Duties Include: Provide a flexible approach to resourcing all commissioning requirements of various projects To support the project team by developing commissioning plans, site commissioning files, methodologies, programmes, method statements and risk assessments Monitor standards to ensure specification compliance and quality standards are strictly maintained Attend FAT activities and provide constructive input with regards to system controls Carry out commissioning of LV Electrical systems, Instrumentation and Control Systems including MCC s and associated field equipment Read and interpret electrical drawings, line diagrams and P&ID s Complete SAT and Asset Start-up Complete commissioning documentation Participate in HAZCOMM activities for various projects as required What Do You Need? Electrical/Instrumentation HND/HNC or equivalent Proven Instrumentation & Electrical experience full commissioning phase experience (installation, dry, wet, process testing) Previous experience holding SAP or AP position and of PTW systems Experience of iMCC s including VSD s, Softstarts and Flow, Level, Pressure and Temperature instrumentation Experience in FAT, SAT testing and fault finding Experience in telemetry end to end testing Profibus, Installation and Commissioning Certification Full UK Driving Licence. Who Are We? RSE is a trusted clean water technology company, developing market-leading products and solutions for purifying drinking water, recycling wastewater, and cleaning water in industrial processes. We are disrupting the water sector, delivering water treatment products, technologies, and services to clients across the UK. RSE provides offsite modular build solutions using a low-carbon approach compared to traditional construction methods and our unique offering to the market focuses on innovation, efficiency, and excellence. Established in 1982, RSE has grown into one of the most prominent MEICA engineering businesses in the UK water industry. We have created a complete in-house and full-service capability from project inception through to design, fabrication, and delivery by means of installation and commissioning. We additionally have one of the largest servicing and maintenance teams in the market, to ensure we re on hand for all our clients needs. Our service offering presents industry-leading innovative solutions and our dedicated staff play a key role in delivering our sustainability and wider business goals. With over 2000 staff across our group of companies, our strategic ambition will see the business continue to grow as we expand our operations and diversify our products. One of RSE s key focuses is driving servant leadership and giving our people the opportunity and responsibility to take an entrepreneurial approach in their career development. What RSE Offer To build successful teams and drive the level of quality that RSE is renowned for, we know we need the best people in the industry. Not only do we require the relevant skillsets, but we also need people with the right attitude and mentality to thrive and grow in an innovative and fast-paced environment. At RSE, you ll be given every opportunity to set the path of your own career through our Business Streams and work within dynamic teams that will require you to rise to the challenge of working for a market leader. Industry-leading salary based on your experience. Company Van. A flexible career development path, with no restrictions on where your career can go. Private Healthcare (Personal). Holiday Allowance of 31 days per year, rising to 33 days per year after 2 years service. Holiday Buy / Sell Scheme Company Pension Scheme Cycle to Work Discounted National Gym Membership Professional Fees Paid Employee Discount Platform EV/Hybrid Car Lease Scheme Access to our network of health professionals including mental health champions and Occupational Health Nurse. In a flourishing sector where there are vast career opportunities available, we believe by leading transformation in the industry our offering to the market means our people have the space to thrive. If you re interested in a career with a company that will harness your skills and provide you with the support to create your own future within the water industry, apply now.
What Are We Looking For? Due to continued growth, RSE is looking for a Senior Commissioning Engineer to join us on a permanent basis, assisting in the delivery of our projects across Scotland. You will undertake Control, Instrumentation and Electrical works including installation, testing, commissioning, and fault finding within new and existing MCC s, industrial instrumentation, analysers, motors, actuators and all other equipment used in the automated control systems associated with RSE products and those used by clients in Water industry. Please note this position will require flexibility on working location, with accommodation provided by RSE. Some of Your Key Duties Include: Provide a flexible approach to resourcing all commissioning requirements of various projects To support the project team by developing commissioning plans, site commissioning files, methodologies, programmes, method statements and risk assessments Monitor standards to ensure specification compliance and quality standards are strictly maintained Attend FAT activities and provide constructive input with regards to system controls Carry out commissioning of LV Electrical systems, Instrumentation and Control Systems including MCC s and associated field equipment Read and interpret electrical drawings, line diagrams and P&ID s Complete SAT and Asset Start-up Complete commissioning documentation Participate in HAZCOMM activities for various projects as required What Do You Need? Electrical/Instrumentation HND/HNC or equivalent Proven Instrumentation & Electrical experience full commissioning phase experience (installation, dry, wet, process testing) Previous experience holding SAP or AP position and of PTW systems Experience of iMCC s including VSD s, Softstarts and Flow, Level, Pressure and Temperature instrumentation Experience in FAT, SAT testing and fault finding Experience in telemetry end to end testing Profibus, Installation and Commissioning Certification Full UK Driving Licence. Who Are We? RSE is a trusted clean water technology company, developing market-leading products and solutions for purifying drinking water, recycling wastewater, and cleaning water in industrial processes. We are disrupting the water sector, delivering water treatment products, technologies, and services to clients across the UK. RSE provides offsite modular build solutions using a low-carbon approach compared to traditional construction methods and our unique offering to the market focuses on innovation, efficiency, and excellence. Established in 1982, RSE has grown into one of the most prominent MEICA engineering businesses in the UK water industry. We have created a complete in-house and full-service capability from project inception through to design, fabrication, and delivery by means of installation and commissioning. We additionally have one of the largest servicing and maintenance teams in the market, to ensure we re on hand for all our clients needs. Our service offering presents industry-leading innovative solutions and our dedicated staff play a key role in delivering our sustainability and wider business goals. With over 2000 staff across our group of companies, our strategic ambition will see the business continue to grow as we expand our operations and diversify our products. One of RSE s key focuses is driving servant leadership and giving our people the opportunity and responsibility to take an entrepreneurial approach in their career development. What RSE Offer To build successful teams and drive the level of quality that RSE is renowned for, we know we need the best people in the industry. Not only do we require the relevant skillsets, but we also need people with the right attitude and mentality to thrive and grow in an innovative and fast-paced environment. At RSE, you ll be given every opportunity to set the path of your own career through our Business Streams and work within dynamic teams that will require you to rise to the challenge of working for a market leader. Industry-leading salary based on your experience. Company Van. A flexible career development path, with no restrictions on where your career can go. Private Healthcare (Personal). Holiday Allowance of 31 days per year, rising to 33 days per year after 2 years service. Holiday Buy / Sell Scheme Company Pension Scheme Cycle to Work Discounted National Gym Membership Professional Fees Paid Employee Discount Platform EV/Hybrid Car Lease Scheme Access to our network of health professionals including mental health champions and Occupational Health Nurse. In a flourishing sector where there are vast career opportunities available, we believe by leading transformation in the industry our offering to the market means our people have the space to thrive. If you re interested in a career with a company that will harness your skills and provide you with the support to create your own future within the water industry, apply now.
Jun 01, 2026
Full time
What Are We Looking For? Due to continued growth, RSE is looking for a Senior Commissioning Engineer to join us on a permanent basis, assisting in the delivery of our projects across Scotland. You will undertake Control, Instrumentation and Electrical works including installation, testing, commissioning, and fault finding within new and existing MCC s, industrial instrumentation, analysers, motors, actuators and all other equipment used in the automated control systems associated with RSE products and those used by clients in Water industry. Please note this position will require flexibility on working location, with accommodation provided by RSE. Some of Your Key Duties Include: Provide a flexible approach to resourcing all commissioning requirements of various projects To support the project team by developing commissioning plans, site commissioning files, methodologies, programmes, method statements and risk assessments Monitor standards to ensure specification compliance and quality standards are strictly maintained Attend FAT activities and provide constructive input with regards to system controls Carry out commissioning of LV Electrical systems, Instrumentation and Control Systems including MCC s and associated field equipment Read and interpret electrical drawings, line diagrams and P&ID s Complete SAT and Asset Start-up Complete commissioning documentation Participate in HAZCOMM activities for various projects as required What Do You Need? Electrical/Instrumentation HND/HNC or equivalent Proven Instrumentation & Electrical experience full commissioning phase experience (installation, dry, wet, process testing) Previous experience holding SAP or AP position and of PTW systems Experience of iMCC s including VSD s, Softstarts and Flow, Level, Pressure and Temperature instrumentation Experience in FAT, SAT testing and fault finding Experience in telemetry end to end testing Profibus, Installation and Commissioning Certification Full UK Driving Licence. Who Are We? RSE is a trusted clean water technology company, developing market-leading products and solutions for purifying drinking water, recycling wastewater, and cleaning water in industrial processes. We are disrupting the water sector, delivering water treatment products, technologies, and services to clients across the UK. RSE provides offsite modular build solutions using a low-carbon approach compared to traditional construction methods and our unique offering to the market focuses on innovation, efficiency, and excellence. Established in 1982, RSE has grown into one of the most prominent MEICA engineering businesses in the UK water industry. We have created a complete in-house and full-service capability from project inception through to design, fabrication, and delivery by means of installation and commissioning. We additionally have one of the largest servicing and maintenance teams in the market, to ensure we re on hand for all our clients needs. Our service offering presents industry-leading innovative solutions and our dedicated staff play a key role in delivering our sustainability and wider business goals. With over 2000 staff across our group of companies, our strategic ambition will see the business continue to grow as we expand our operations and diversify our products. One of RSE s key focuses is driving servant leadership and giving our people the opportunity and responsibility to take an entrepreneurial approach in their career development. What RSE Offer To build successful teams and drive the level of quality that RSE is renowned for, we know we need the best people in the industry. Not only do we require the relevant skillsets, but we also need people with the right attitude and mentality to thrive and grow in an innovative and fast-paced environment. At RSE, you ll be given every opportunity to set the path of your own career through our Business Streams and work within dynamic teams that will require you to rise to the challenge of working for a market leader. Industry-leading salary based on your experience. Company Van. A flexible career development path, with no restrictions on where your career can go. Private Healthcare (Personal). Holiday Allowance of 31 days per year, rising to 33 days per year after 2 years service. Holiday Buy / Sell Scheme Company Pension Scheme Cycle to Work Discounted National Gym Membership Professional Fees Paid Employee Discount Platform EV/Hybrid Car Lease Scheme Access to our network of health professionals including mental health champions and Occupational Health Nurse. In a flourishing sector where there are vast career opportunities available, we believe by leading transformation in the industry our offering to the market means our people have the space to thrive. If you re interested in a career with a company that will harness your skills and provide you with the support to create your own future within the water industry, apply now.
Solus Accident Repair Centres
Birchanger, Hertfordshire
Overview Solus, part of the Aviva family, is growing our Technology capability and we're looking for a talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great opportunity to make a real impact across a large, modern technology estate. Responsibilities You'll be our subject matter expert for monitoring and performance, responsible for designing, implementing and maintaining the tools and dashboards that give us real-time visibility of our infrastructure, applications and cloud services. Your focus will include: Owning and optimising platforms such as LogicMonitor, Azure Monitor, App Insights and Log Analytics Building meaningful dashboards, alerts, telemetry pipelines and performance insights Identifying risks, trends and early indicators to prevent incidents before they happen Carrying out deep-dive investigations into performance issues and recommending improvements Working with Platform, Operations, Security and Product teams to ensure systems are reliable, available and scalable Automating responses and integrations to improve speed, accuracy and consistency Supporting major changes, deployments and post-incident reviews with data-driven evidence Qualifications Strong experience with monitoring and observability tools (LogicMonitor, Azure Monitor, App Insights, Log Analytics, Defender for Cloud) Excellent understanding of cloud performance, IaaS/PaaS, networking fundamentals, API performance and capacity modelling Skilled in dashboards, log queries (KQL), custom metrics and performance analysis Ability to diagnose complex issues across infrastructure, networks, applications or databases Confident scripting and automation skills (PowerShell, Azure Automation, Graph API) Clear communicator who can simplify technical detail for both technical and non-technical teams Desirable qualifications Microsoft certifications (AZ-900, AZ-104, AZ-305, AZ-500) or similar Experience with LogicMonitor admin, Grafana or other observability tools Familiarity with SRE concepts (SLIs, SLOs, error budgets) Understanding of ITIL processes Who are Solus? Solus, who are owned by Aviva, are one of the UK leaders in vehicle repairs, returning cars to the road in just 11 days on average and a 4.6/5 star customer rating. With an award-winning apprenticeship programme and winners of other recognised industry awards Solus are proud to be shaping the future of vehicle repair. Why Join Solus? We have so much to offer when it comes to being a Solus colleague: Competitive salary based on location, skills, experience, and qualifications. Bonus opportunity tied to your performance and the overall success of Solus. Company pension scheme with employer contributions. 33 days' holiday (including bank holidays), with the option to buy or sell up to 5 days. Save money with up to 40% discount on Aviva products and other retailer discounts. Share in Aviva's success through the Aviva Save As You Earn scheme. Supportive policies including parental and carer's leave. Wellbeing focus with tools like Group Income Protection and 24/7 GP access. At Solus, we value inclusivity and welcome all applicants. If you're excited but don't tick every box, we encourage you to apply-your unique skills might be just what we need. We guarantee an interview for disabled applicants meeting the minimum criteria-just email us after applying to let us know. Ready to join us? Apply online today, and our team will be in touch within 14 days.
May 31, 2026
Full time
Overview Solus, part of the Aviva family, is growing our Technology capability and we're looking for a talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great opportunity to make a real impact across a large, modern technology estate. Responsibilities You'll be our subject matter expert for monitoring and performance, responsible for designing, implementing and maintaining the tools and dashboards that give us real-time visibility of our infrastructure, applications and cloud services. Your focus will include: Owning and optimising platforms such as LogicMonitor, Azure Monitor, App Insights and Log Analytics Building meaningful dashboards, alerts, telemetry pipelines and performance insights Identifying risks, trends and early indicators to prevent incidents before they happen Carrying out deep-dive investigations into performance issues and recommending improvements Working with Platform, Operations, Security and Product teams to ensure systems are reliable, available and scalable Automating responses and integrations to improve speed, accuracy and consistency Supporting major changes, deployments and post-incident reviews with data-driven evidence Qualifications Strong experience with monitoring and observability tools (LogicMonitor, Azure Monitor, App Insights, Log Analytics, Defender for Cloud) Excellent understanding of cloud performance, IaaS/PaaS, networking fundamentals, API performance and capacity modelling Skilled in dashboards, log queries (KQL), custom metrics and performance analysis Ability to diagnose complex issues across infrastructure, networks, applications or databases Confident scripting and automation skills (PowerShell, Azure Automation, Graph API) Clear communicator who can simplify technical detail for both technical and non-technical teams Desirable qualifications Microsoft certifications (AZ-900, AZ-104, AZ-305, AZ-500) or similar Experience with LogicMonitor admin, Grafana or other observability tools Familiarity with SRE concepts (SLIs, SLOs, error budgets) Understanding of ITIL processes Who are Solus? Solus, who are owned by Aviva, are one of the UK leaders in vehicle repairs, returning cars to the road in just 11 days on average and a 4.6/5 star customer rating. With an award-winning apprenticeship programme and winners of other recognised industry awards Solus are proud to be shaping the future of vehicle repair. Why Join Solus? We have so much to offer when it comes to being a Solus colleague: Competitive salary based on location, skills, experience, and qualifications. Bonus opportunity tied to your performance and the overall success of Solus. Company pension scheme with employer contributions. 33 days' holiday (including bank holidays), with the option to buy or sell up to 5 days. Save money with up to 40% discount on Aviva products and other retailer discounts. Share in Aviva's success through the Aviva Save As You Earn scheme. Supportive policies including parental and carer's leave. Wellbeing focus with tools like Group Income Protection and 24/7 GP access. At Solus, we value inclusivity and welcome all applicants. If you're excited but don't tick every box, we encourage you to apply-your unique skills might be just what we need. We guarantee an interview for disabled applicants meeting the minimum criteria-just email us after applying to let us know. Ready to join us? Apply online today, and our team will be in touch within 14 days.
Electrical Estimator /E&I Estimator Location: Thornaby on Tees or Gosforth Salary : Starting at £55,000 per annum Vacancy Type : Permanent, Full Time ACEDA is a leading provider of integrated technology solutions, specialising in IT infrastructure, fire & security systems, electrical services and smart building technologies. With over 30 years experience, we deliver end-to-end solutions that help organisations operate safely, securely and efficiently. The Role ACEDA is seeking an experienced Electrical Estimator to support the continued growth of our Electrical and Integrated Technology divisions. This role will involve surveying opportunities, preparing accurate and commercially robust cost estimates, and supporting the conversion of opportunities into profitable projects across a range of sectors including utilities, infrastructure, commercial estates and public sector environments. The successful candidate will have a strong understanding of industrial electrical installations and compliance works, including the ability to review Electrical Installation Condition Reports (EICRs) and develop costed remedial solutions. Key Responsibilities and Accountabilities Surveying & Opportunity Development: Attend client sites to survey electrical installations and assess project requirements. Engage with clients to understand technical requirements, compliance obligations and operational constraints. Identify opportunities for electrical upgrades, compliance works and system improvements. Estimating & Pricing: Prepare accurate electrical cost estimates and proposals for works including: Power distribution systems Lighting and emergency lighting installations Electrical infrastructure upgrades Compliance remedial works Industrial electrical installations Responsibilities include: Producing material take-offs and labour allowances Developing detailed cost plans Obtaining and reviewing supplier and subcontractor quotations Ensuring estimates reflect labour productivity, project risks and margin expectations EICR & Compliance Remedial Works: Review Electrical Installation Condition Reports (EICRs) and identify required remedial works. Develop costed solutions for electrical compliance upgrades and improvement works. Support clients with planned compliance programmes and remedial works across their estates. Industrial Electrical & Telemetry Projects: Support preparation of tender submissions and technical proposals. Assist with scope reviews, clarifications and client queries. Project Handover: Provide structured handover information to project delivery teams following project award. Ensure clear documentation of scope, assumptions and commercial considerations. Strategic Contribution: Support the Chief Revenue Officer in identifying new opportunities and developing client relationships across ACEDA s key sectors including utilities, infrastructure, commercial states and public sector organisations. Contribute to the development of ACEDA s electrical compliance and industrial services offering, particularly around EICR programmes and remedial works. Provide technical and commercial insight to support business growth and new client opportunities. Skills & Experience Essential: Proven experience as an Electrical Estimator, Electrical Supervisor, Project Engineer or Project Manager. Strong understanding of industrial and commercial electrical installations. Ability to survey electrical works and develop scopes independently. Strong commercial awareness and understanding of project costing and margins. Experience producing detailed electrical estimates and proposals. Ability to interpret electrical drawings, schematics and specifications. Desirable: Knowledge of Electrical Installation Condition Reports (EICRs) and compliance remedial works. Familiarity with BS 7671 Wiring Regulations and inspection standards. Experience working in utilities, infrastructure or industrial environments. Exposure to telemetry, control systems or instrumentation installations. Inspection and testing qualifications such as City & Guilds 2391 / 2394 / 2395. Experience using estimating or project management systems such as Simpro What We Offer: Opportunity to work within a growing integrated technology business. Exposure to projects across electrical, fire, security and infrastructure systems. A collaborative and supportive working environment. Competitive salary and benefits package. To Apply If you feel you are a suitable candidate and would like to work for ACEDA, please do not hesitate to apply.
May 28, 2026
Full time
Electrical Estimator /E&I Estimator Location: Thornaby on Tees or Gosforth Salary : Starting at £55,000 per annum Vacancy Type : Permanent, Full Time ACEDA is a leading provider of integrated technology solutions, specialising in IT infrastructure, fire & security systems, electrical services and smart building technologies. With over 30 years experience, we deliver end-to-end solutions that help organisations operate safely, securely and efficiently. The Role ACEDA is seeking an experienced Electrical Estimator to support the continued growth of our Electrical and Integrated Technology divisions. This role will involve surveying opportunities, preparing accurate and commercially robust cost estimates, and supporting the conversion of opportunities into profitable projects across a range of sectors including utilities, infrastructure, commercial estates and public sector environments. The successful candidate will have a strong understanding of industrial electrical installations and compliance works, including the ability to review Electrical Installation Condition Reports (EICRs) and develop costed remedial solutions. Key Responsibilities and Accountabilities Surveying & Opportunity Development: Attend client sites to survey electrical installations and assess project requirements. Engage with clients to understand technical requirements, compliance obligations and operational constraints. Identify opportunities for electrical upgrades, compliance works and system improvements. Estimating & Pricing: Prepare accurate electrical cost estimates and proposals for works including: Power distribution systems Lighting and emergency lighting installations Electrical infrastructure upgrades Compliance remedial works Industrial electrical installations Responsibilities include: Producing material take-offs and labour allowances Developing detailed cost plans Obtaining and reviewing supplier and subcontractor quotations Ensuring estimates reflect labour productivity, project risks and margin expectations EICR & Compliance Remedial Works: Review Electrical Installation Condition Reports (EICRs) and identify required remedial works. Develop costed solutions for electrical compliance upgrades and improvement works. Support clients with planned compliance programmes and remedial works across their estates. Industrial Electrical & Telemetry Projects: Support preparation of tender submissions and technical proposals. Assist with scope reviews, clarifications and client queries. Project Handover: Provide structured handover information to project delivery teams following project award. Ensure clear documentation of scope, assumptions and commercial considerations. Strategic Contribution: Support the Chief Revenue Officer in identifying new opportunities and developing client relationships across ACEDA s key sectors including utilities, infrastructure, commercial states and public sector organisations. Contribute to the development of ACEDA s electrical compliance and industrial services offering, particularly around EICR programmes and remedial works. Provide technical and commercial insight to support business growth and new client opportunities. Skills & Experience Essential: Proven experience as an Electrical Estimator, Electrical Supervisor, Project Engineer or Project Manager. Strong understanding of industrial and commercial electrical installations. Ability to survey electrical works and develop scopes independently. Strong commercial awareness and understanding of project costing and margins. Experience producing detailed electrical estimates and proposals. Ability to interpret electrical drawings, schematics and specifications. Desirable: Knowledge of Electrical Installation Condition Reports (EICRs) and compliance remedial works. Familiarity with BS 7671 Wiring Regulations and inspection standards. Experience working in utilities, infrastructure or industrial environments. Exposure to telemetry, control systems or instrumentation installations. Inspection and testing qualifications such as City & Guilds 2391 / 2394 / 2395. Experience using estimating or project management systems such as Simpro What We Offer: Opportunity to work within a growing integrated technology business. Exposure to projects across electrical, fire, security and infrastructure systems. A collaborative and supportive working environment. Competitive salary and benefits package. To Apply If you feel you are a suitable candidate and would like to work for ACEDA, please do not hesitate to apply.
Join an award-winning B2B consultancy at the forefront of enterprise AI, building and owning the cloud-native platform infrastructure that powers production-grade conversational and generative AI products at scale. The role This is a platform and infrastructure engineering role - not a data science or ML engineering position. You'll own the runtime, infrastructure, and operational layers that RAG pipelines, LLM orchestration, vector search, and evaluation workflows run on, across AWS and Databricks. The focus is on building scalable, observable, secure, and cost-efficient platform infrastructure that enables AI engineering teams to ship and operate AI products reliably in production. What you'll do Design, build, and operate cloud-native AI platform infrastructure across AWS (Lambda, API Gateway, DynamoDB, S3, CloudWatch) and Databricks Deploy and operate containerised services on Kubernetes using Terraform for infrastructure-as-code Own and scale vector search infrastructure (OpenSearch, Algolia, AWS Bedrock Knowledge Bases) and embedding pipelines Build and maintain CI/CD pipelines for inference services, retrievers, ingestion workflows, and RAG components Implement observability across AI workloads using CloudWatch, MLflow, and OpenTelemetry - covering latency, throughput, cost, and system health Apply secure-by-design principles including IAM, encryption, network controls, and audit logging Work closely with AI engineers to translate prototypes and proof-of-concepts into production-ready, well-architected platform components What we're looking for Proven experience in platform, infrastructure, or software engineering roles delivering production-grade systems on AWS Strong hands-on Kubernetes experience, specifically with EKS (Elastic Kubernetes Service) and ECS (Elastic Container Service) in production environments Strong Terraform experience for infrastructure-as-code, provisioning and managing cloud infrastructure at scale Experience operating containerised services, managing CI/CD pipelines, and owning observability and reliability Familiarity with vector databases or search infrastructure (OpenSearch, Algolia) is a strong advantage Python proficiency for scripting, automation, and deploying production services Solid grasp of distributed systems, cloud-native architecture, microservices, and API design Ownership mindset - comfortable operating autonomously across reliability, performance, cost, and security Why join? You'll own the foundational platform infrastructure behind a growing suite of generative AI products, working directly with senior AI and engineering leaders. This is a deep technical ownership role with long-term architectural impact, within an organisation investing heavily in AI at scale. INDAM The Portfolio Group are acting on behalf of our client in recruiting for this position.
May 22, 2026
Full time
Join an award-winning B2B consultancy at the forefront of enterprise AI, building and owning the cloud-native platform infrastructure that powers production-grade conversational and generative AI products at scale. The role This is a platform and infrastructure engineering role - not a data science or ML engineering position. You'll own the runtime, infrastructure, and operational layers that RAG pipelines, LLM orchestration, vector search, and evaluation workflows run on, across AWS and Databricks. The focus is on building scalable, observable, secure, and cost-efficient platform infrastructure that enables AI engineering teams to ship and operate AI products reliably in production. What you'll do Design, build, and operate cloud-native AI platform infrastructure across AWS (Lambda, API Gateway, DynamoDB, S3, CloudWatch) and Databricks Deploy and operate containerised services on Kubernetes using Terraform for infrastructure-as-code Own and scale vector search infrastructure (OpenSearch, Algolia, AWS Bedrock Knowledge Bases) and embedding pipelines Build and maintain CI/CD pipelines for inference services, retrievers, ingestion workflows, and RAG components Implement observability across AI workloads using CloudWatch, MLflow, and OpenTelemetry - covering latency, throughput, cost, and system health Apply secure-by-design principles including IAM, encryption, network controls, and audit logging Work closely with AI engineers to translate prototypes and proof-of-concepts into production-ready, well-architected platform components What we're looking for Proven experience in platform, infrastructure, or software engineering roles delivering production-grade systems on AWS Strong hands-on Kubernetes experience, specifically with EKS (Elastic Kubernetes Service) and ECS (Elastic Container Service) in production environments Strong Terraform experience for infrastructure-as-code, provisioning and managing cloud infrastructure at scale Experience operating containerised services, managing CI/CD pipelines, and owning observability and reliability Familiarity with vector databases or search infrastructure (OpenSearch, Algolia) is a strong advantage Python proficiency for scripting, automation, and deploying production services Solid grasp of distributed systems, cloud-native architecture, microservices, and API design Ownership mindset - comfortable operating autonomously across reliability, performance, cost, and security Why join? You'll own the foundational platform infrastructure behind a growing suite of generative AI products, working directly with senior AI and engineering leaders. This is a deep technical ownership role with long-term architectural impact, within an organisation investing heavily in AI at scale. INDAM The Portfolio Group are acting on behalf of our client in recruiting for this position.
Who We Are Boston Consulting Group partners with leaders in business and society to tackle their most important challenges and capture their greatest opportunities. BCG was the pioneer in business strategy when it was founded in 1963. Today, we help clients with total transformation-inspiring complex change, enabling organizations to grow, building competitive advantage, and driving bottom-line impact. To succeed, organizations must blend digital and human capabilities. Our diverse, global teams bring deep industry and functional expertise and a range of perspectives to spark change. BCG delivers solutions through leading-edge management consulting along with technology and design, corporate and digital ventures-and business purpose. We work in a uniquely collaborative model across the firm and throughout all levels of the client organization, generating results that allow our clients to thrive. What You'll Do The Principal Site Reliability Engineer (SRE) is a senior technical leader responsible for shaping how reliability, automation, and operational excellence are engineered across the organisation. Operating across domains including traditional infrastructure, cloud engineering, network operations, identity, observability, security, AI-driven operations, and automated data workflows, the role focuses on designing scalable systems, reusable engineering patterns, and standardised controls that reduce operational toil, improve resilience, and embed reliability, governance, and compliance directly into delivery pipelines and operational platforms. This role will drive organisational change towards automation-first, measurable, and repeatable practices. A key part of the role is building and evolving reusable CI/CD and Terraform modules, engineering guardrails, observability patterns, and automation frameworks that can be adopted across multiple teams and domains without requiring each team to solve the same problems independently. The Principal SRE also plays an important enablement role beyond deeply technical teams, helping less technical areas of the business adopt structured, governed, and scalable ways of working. This includes translating complex engineering practices into practical standards, improving how governance is implemented through engineering controls rather than manual oversight, and driving operational maturity across a broad and diverse technology landscape. The ideal candidate is a systems thinker who understands how services, networks, identity, data flows, and operational processes fail in real-world conditions, and can apply that understanding to build automation-first, reliability-focused operating models that scale across both technical and non-technical functions. Key Responsibilities Cross-Domain Reliability Engineering Design and evolve reliability patterns across cloud, network, identity, and security domains. Identify systemic risks and failure modes across platforms and services, and define engineering solutions to mitigate them. Ensure operational activities are embedded into delivery models through automation, CI/CD integration, and event-driven workflows. Automation & Toil Reduction at Scale Lead the design of automation frameworks that eliminate manual operational tasks across multiple domains. Translate incident learnings and operational inefficiencies into scalable automation and preventative controls. Drive adoption of automation-first principles, reducing dependency on human-driven processes. Contribute to AI-driven operational use cases, including event correlation, anomaly detection, noise reduction, operational insights, and automated remediation. Ensure AIOps capabilities are grounded in reliable telemetry, clear control boundaries, and measurable operational outcomes. Observability & 24/7 Operational Excellence Define standards for telemetry, monitoring, alerting, and operational visibility across all critical systems. Ensure services are observable, measurable, and support proactive detection of issues. Improve operational readiness, incident response effectiveness, and time-to-recovery through engineering solutions. CI/CD & Platform Integration Contribute to the design of CI/CD patterns that embed reliability, security, and operational controls into pipelines. Ensure infrastructure, network, identity, and security configurations are managed through code and validated automatically. Support integration of platform services into delivery pipelines to enable consistent, repeatable deployments. Security & Identity Integration Contribute to secure-by-design patterns, including least privilege, identity-based access, and short-lived credentials. Support integration of security controls (e.g. secrets management, authentication, policy enforcement) into engineering workflows. Ensure security and compliance requirements are met through engineering controls rather than manual processes. Network & Infrastructure Reliability Support the design of resilient network architectures and segmentation aligned with Zero Trust principles. Ensure network configurations and controls are automated, validated, and observable. Contribute to infrastructure design patterns that improve availability, scalability, and fault tolerance. Design and improve operational patterns for network reliability, segmentation, visibility, and change validation. Support automation and standardisation of network controls and operational procedures to reduce manual intervention and configuration drift. Technical Leadership & Enablement Provide technical leadership across teams, influencing standards, architecture, and engineering practices. Mentor engineers on reliability engineering, automation, and systems thinking. Drive consistency through reusable patterns, frameworks, and documentation. Strategic Influence & Continuous Improvement Contribute to reliability engineering strategy and roadmap across the organisation. Communicate technical concepts, risks, and recommendations to senior stakeholders and leadership. Lead initiatives that improve reliability maturity, engineering efficiency, and operational scalability. Support less technical teams and functions in adopting structured, automated, and measurable operational practices. Act as a bridge between engineering capability and organisational change, helping scale good practice beyond core platform teams. Automated Data Workflows Design and improve automated data workflows that support operational reporting, observability, governance, and decision-making. Ensure operational data pipelines are reliable, timely, and aligned to engineering and business needs. Reusable Engineering Frameworks Build and evolve reusable modules, patterns, and frameworks for CI/CD, Terraform, and operational automation. Embed governance, validation, and reliability controls into these shared engineering assets by default. Governance by Engineering Translate governance requirements into practical engineering controls, automated checks, and repeatable standards. Help teams adopt compliant and supportable operating models without relying on manual policing or process-heavy interventions. What You'll Bring Required Qualifications 10+ years of experience in Site Reliability Engineering, Platform Engineering, or related fields. Strong hands-on experience across multiple domains, including: Cloud platforms (AWS, Azure) CI/CD and Infrastructure-as-Code (e.g. Terraform) Observability tools (e.g. Datadog, Splunk) Automation and scripting (e.g. Python) Experience designing and implementing scalable automation and reliability solutions. Deep understanding of distributed systems, failure modes, and resilience patterns. Experience integrating operational and security controls into engineering workflows. Strong stakeholder engagement and technical communication skills. Preferred Qualifications Experience with identity and access management systems (e.g. Entra ID, Vault). Experience with network architecture and security controls (e.g. firewalls, segmentation). Familiarity with Zero Trust principles and security engineering practices. Experience working in large, federated organisations with diverse technology stacks. Exposure to compliance and regulatory requirements (e.g. PCI, HIPAA, SOX). Additional info Hybrid or on-site work model. Operates as a senior individual contributor with broad cross-organisational influence. Expected to balance hands-on technical leadership with strategic direction. Occasional travel may be required for team or stakeholder engagement. Boston Consulting Group is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, age, religion, sex, sexual orientation, gender identity / expression, national origin, disability, protected veteran status, or any other characteristic protected under national, provincial, or local law, where applicable, and those with criminal histories will be considered in a manner consistent with applicable state and local laws. BCG is an E - Verify Employer. Click here for more information on E-Verify.
May 21, 2026
Full time
Who We Are Boston Consulting Group partners with leaders in business and society to tackle their most important challenges and capture their greatest opportunities. BCG was the pioneer in business strategy when it was founded in 1963. Today, we help clients with total transformation-inspiring complex change, enabling organizations to grow, building competitive advantage, and driving bottom-line impact. To succeed, organizations must blend digital and human capabilities. Our diverse, global teams bring deep industry and functional expertise and a range of perspectives to spark change. BCG delivers solutions through leading-edge management consulting along with technology and design, corporate and digital ventures-and business purpose. We work in a uniquely collaborative model across the firm and throughout all levels of the client organization, generating results that allow our clients to thrive. What You'll Do The Principal Site Reliability Engineer (SRE) is a senior technical leader responsible for shaping how reliability, automation, and operational excellence are engineered across the organisation. Operating across domains including traditional infrastructure, cloud engineering, network operations, identity, observability, security, AI-driven operations, and automated data workflows, the role focuses on designing scalable systems, reusable engineering patterns, and standardised controls that reduce operational toil, improve resilience, and embed reliability, governance, and compliance directly into delivery pipelines and operational platforms. This role will drive organisational change towards automation-first, measurable, and repeatable practices. A key part of the role is building and evolving reusable CI/CD and Terraform modules, engineering guardrails, observability patterns, and automation frameworks that can be adopted across multiple teams and domains without requiring each team to solve the same problems independently. The Principal SRE also plays an important enablement role beyond deeply technical teams, helping less technical areas of the business adopt structured, governed, and scalable ways of working. This includes translating complex engineering practices into practical standards, improving how governance is implemented through engineering controls rather than manual oversight, and driving operational maturity across a broad and diverse technology landscape. The ideal candidate is a systems thinker who understands how services, networks, identity, data flows, and operational processes fail in real-world conditions, and can apply that understanding to build automation-first, reliability-focused operating models that scale across both technical and non-technical functions. Key Responsibilities Cross-Domain Reliability Engineering Design and evolve reliability patterns across cloud, network, identity, and security domains. Identify systemic risks and failure modes across platforms and services, and define engineering solutions to mitigate them. Ensure operational activities are embedded into delivery models through automation, CI/CD integration, and event-driven workflows. Automation & Toil Reduction at Scale Lead the design of automation frameworks that eliminate manual operational tasks across multiple domains. Translate incident learnings and operational inefficiencies into scalable automation and preventative controls. Drive adoption of automation-first principles, reducing dependency on human-driven processes. Contribute to AI-driven operational use cases, including event correlation, anomaly detection, noise reduction, operational insights, and automated remediation. Ensure AIOps capabilities are grounded in reliable telemetry, clear control boundaries, and measurable operational outcomes. Observability & 24/7 Operational Excellence Define standards for telemetry, monitoring, alerting, and operational visibility across all critical systems. Ensure services are observable, measurable, and support proactive detection of issues. Improve operational readiness, incident response effectiveness, and time-to-recovery through engineering solutions. CI/CD & Platform Integration Contribute to the design of CI/CD patterns that embed reliability, security, and operational controls into pipelines. Ensure infrastructure, network, identity, and security configurations are managed through code and validated automatically. Support integration of platform services into delivery pipelines to enable consistent, repeatable deployments. Security & Identity Integration Contribute to secure-by-design patterns, including least privilege, identity-based access, and short-lived credentials. Support integration of security controls (e.g. secrets management, authentication, policy enforcement) into engineering workflows. Ensure security and compliance requirements are met through engineering controls rather than manual processes. Network & Infrastructure Reliability Support the design of resilient network architectures and segmentation aligned with Zero Trust principles. Ensure network configurations and controls are automated, validated, and observable. Contribute to infrastructure design patterns that improve availability, scalability, and fault tolerance. Design and improve operational patterns for network reliability, segmentation, visibility, and change validation. Support automation and standardisation of network controls and operational procedures to reduce manual intervention and configuration drift. Technical Leadership & Enablement Provide technical leadership across teams, influencing standards, architecture, and engineering practices. Mentor engineers on reliability engineering, automation, and systems thinking. Drive consistency through reusable patterns, frameworks, and documentation. Strategic Influence & Continuous Improvement Contribute to reliability engineering strategy and roadmap across the organisation. Communicate technical concepts, risks, and recommendations to senior stakeholders and leadership. Lead initiatives that improve reliability maturity, engineering efficiency, and operational scalability. Support less technical teams and functions in adopting structured, automated, and measurable operational practices. Act as a bridge between engineering capability and organisational change, helping scale good practice beyond core platform teams. Automated Data Workflows Design and improve automated data workflows that support operational reporting, observability, governance, and decision-making. Ensure operational data pipelines are reliable, timely, and aligned to engineering and business needs. Reusable Engineering Frameworks Build and evolve reusable modules, patterns, and frameworks for CI/CD, Terraform, and operational automation. Embed governance, validation, and reliability controls into these shared engineering assets by default. Governance by Engineering Translate governance requirements into practical engineering controls, automated checks, and repeatable standards. Help teams adopt compliant and supportable operating models without relying on manual policing or process-heavy interventions. What You'll Bring Required Qualifications 10+ years of experience in Site Reliability Engineering, Platform Engineering, or related fields. Strong hands-on experience across multiple domains, including: Cloud platforms (AWS, Azure) CI/CD and Infrastructure-as-Code (e.g. Terraform) Observability tools (e.g. Datadog, Splunk) Automation and scripting (e.g. Python) Experience designing and implementing scalable automation and reliability solutions. Deep understanding of distributed systems, failure modes, and resilience patterns. Experience integrating operational and security controls into engineering workflows. Strong stakeholder engagement and technical communication skills. Preferred Qualifications Experience with identity and access management systems (e.g. Entra ID, Vault). Experience with network architecture and security controls (e.g. firewalls, segmentation). Familiarity with Zero Trust principles and security engineering practices. Experience working in large, federated organisations with diverse technology stacks. Exposure to compliance and regulatory requirements (e.g. PCI, HIPAA, SOX). Additional info Hybrid or on-site work model. Operates as a senior individual contributor with broad cross-organisational influence. Expected to balance hands-on technical leadership with strategic direction. Occasional travel may be required for team or stakeholder engagement. Boston Consulting Group is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, age, religion, sex, sexual orientation, gender identity / expression, national origin, disability, protected veteran status, or any other characteristic protected under national, provincial, or local law, where applicable, and those with criminal histories will be considered in a manner consistent with applicable state and local laws. BCG is an E - Verify Employer. Click here for more information on E-Verify.
Who We Are Boston Consulting Group partners with leaders in business and society to tackle their most important challenges and capture their greatest opportunities. BCG was the pioneer in business strategy when it was founded in 1963. Today, we help clients with total transformation-inspiring complex change, enabling organizations to grow, building competitive advantage, and driving bottom-line impact. To succeed, organizations must blend digital and human capabilities. Our diverse, global teams bring deep industry and functional expertise and a range of perspectives to spark change. BCG delivers solutions through leading-edge management consulting along with technology and design, corporate and digital ventures-and business purpose. We work in a uniquely collaborative model across the firm and throughout all levels of the client organization, generating results that allow our clients to thrive. What You'll Do The Principal Site Reliability Engineer (SRE) is a senior technical leader responsible for shaping how reliability, automation, and operational excellence are engineered across the organisation. Operating across domains including traditional infrastructure, cloud engineering, network operations, identity, observability, security, AI-driven operations, and automated data workflows, the role focuses on designing scalable systems, reusable engineering patterns, and standardised controls that reduce operational toil, improve resilience, and embed reliability, governance, and compliance directly into delivery pipelines and operational platforms. This role will drive organisational change towards automation-first, measurable, and repeatable practices. A key part of the role is building and evolving reusable CI/CD and Terraform modules, engineering guardrails, observability patterns, and automation frameworks that can be adopted across multiple teams and domains without requiring each team to solve the same problems independently. The Principal SRE also plays an important enablement role beyond deeply technical teams, helping less technical areas of the business adopt structured, governed, and scalable ways of working. This includes translating complex engineering practices into practical standards, improving how governance is implemented through engineering controls rather than manual oversight, and driving operational maturity across a broad and diverse technology landscape. The ideal candidate is a systems thinker who understands how services, networks, identity, data flows, and operational processes fail in real-world conditions, and can apply that understanding to build automation-first, reliability-focused operating models that scale across both technical and non-technical functions. Key Responsibilities Cross-Domain Reliability Engineering Design and evolve reliability patterns across cloud, network, identity, and security domains. Identify systemic risks and failure modes across platforms and services, and define engineering solutions to mitigate them. Ensure operational activities are embedded into delivery models through automation, CI/CD integration, and event-driven workflows. Automation & Toil Reduction at Scale Lead the design of automation frameworks that eliminate manual operational tasks across multiple domains. Translate incident learnings and operational inefficiencies into scalable automation and preventative controls. Drive adoption of automation-first principles, reducing dependency on human-driven processes. Contribute to AI-driven operational use cases, including event correlation, anomaly detection, noise reduction, operational insights, and automated remediation. Ensure AIOps capabilities are grounded in reliable telemetry, clear control boundaries, and measurable operational outcomes. Observability & 24/7 Operational Excellence Define standards for telemetry, monitoring, alerting, and operational visibility across all critical systems. Ensure services are observable, measurable, and support proactive detection of issues. Improve operational readiness, incident response effectiveness, and time-to-recovery through engineering solutions. CI/CD & Platform Integration Contribute to the design of CI/CD patterns that embed reliability, security, and operational controls into pipelines. Ensure infrastructure, network, identity, and security configurations are managed through code and validated automatically. Support integration of platform services into delivery pipelines to enable consistent, repeatable deployments. Security & Identity Integration Contribute to secure-by-design patterns, including least privilege, identity-based access, and short-lived credentials. Support integration of security controls (e.g. secrets management, authentication, policy enforcement) into engineering workflows. Ensure security and compliance requirements are met through engineering controls rather than manual processes. Network & Infrastructure Reliability Support the design of resilient network architectures and segmentation aligned with Zero Trust principles. Ensure network configurations and controls are automated, validated, and observable. Contribute to infrastructure design patterns that improve availability, scalability, and fault tolerance. Design and improve operational patterns for network reliability, segmentation, visibility, and change validation. Support automation and standardisation of network controls and operational procedures to reduce manual intervention and configuration drift. Technical Leadership & Enablement Provide technical leadership across teams, influencing standards, architecture, and engineering practices. Mentor engineers on reliability engineering, automation, and systems thinking. Drive consistency through reusable patterns, frameworks, and documentation. Strategic Influence & Continuous Improvement Contribute to reliability engineering strategy and roadmap across the organisation. Communicate technical concepts, risks, and recommendations to senior stakeholders and leadership. Lead initiatives that improve reliability maturity, engineering efficiency, and operational scalability. Support less technical teams and functions in adopting structured, automated, and measurable operational practices. Act as a bridge between engineering capability and organisational change, helping scale good practice beyond core platform teams. Automated Data Workflows Design and improve automated data workflows that support operational reporting, observability, governance, and decision-making. Ensure operational data pipelines are reliable, timely, and aligned to engineering and business needs. Reusable Engineering Frameworks Build and evolve reusable modules, patterns, and frameworks for CI/CD, Terraform, and operational automation. Embed governance, validation, and reliability controls into these shared engineering assets by default. Governance by Engineering Translate governance requirements into practical engineering controls, automated checks, and repeatable standards. Help teams adopt compliant and supportable operating models without relying on manual policing or process-heavy interventions. What You'll Bring Required Qualifications 10+ years of experience in Site Reliability Engineering, Platform Engineering, or related fields. Strong hands-on experience across multiple domains, including: Cloud platforms (AWS, Azure) CI/CD and Infrastructure-as-Code (e.g. Terraform) Observability tools (e.g. Datadog, Splunk) Automation and scripting (e.g. Python) Experience designing and implementing scalable automation and reliability solutions. Deep understanding of distributed systems, failure modes, and resilience patterns. Experience integrating operational and security controls into engineering workflows. Strong stakeholder engagement and technical communication skills. Preferred Qualifications Experience with identity and access management systems (e.g. Entra ID, Vault). Experience with network architecture and security controls (e.g. firewalls, segmentation). Familiarity with Zero Trust principles and security engineering practices. Experience working in large, federated organisations with diverse technology stacks. Exposure to compliance and regulatory requirements (e.g. PCI, HIPAA, SOX). Additional info Hybrid or on-site work model. Operates as a senior individual contributor with broad cross-organisational influence. Expected to balance hands-on technical leadership with strategic direction. Occasional travel may be required for team or stakeholder engagement. Boston Consulting Group is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, age, religion, sex, sexual orientation, gender identity / expression, national origin, disability, protected veteran status, or any other characteristic protected under national, provincial, or local law, where applicable, and those with criminal histories will be considered in a manner consistent with applicable state and local laws. BCG is an E - Verify Employer. Click here for more information on E-Verify.
May 21, 2026
Full time
Who We Are Boston Consulting Group partners with leaders in business and society to tackle their most important challenges and capture their greatest opportunities. BCG was the pioneer in business strategy when it was founded in 1963. Today, we help clients with total transformation-inspiring complex change, enabling organizations to grow, building competitive advantage, and driving bottom-line impact. To succeed, organizations must blend digital and human capabilities. Our diverse, global teams bring deep industry and functional expertise and a range of perspectives to spark change. BCG delivers solutions through leading-edge management consulting along with technology and design, corporate and digital ventures-and business purpose. We work in a uniquely collaborative model across the firm and throughout all levels of the client organization, generating results that allow our clients to thrive. What You'll Do The Principal Site Reliability Engineer (SRE) is a senior technical leader responsible for shaping how reliability, automation, and operational excellence are engineered across the organisation. Operating across domains including traditional infrastructure, cloud engineering, network operations, identity, observability, security, AI-driven operations, and automated data workflows, the role focuses on designing scalable systems, reusable engineering patterns, and standardised controls that reduce operational toil, improve resilience, and embed reliability, governance, and compliance directly into delivery pipelines and operational platforms. This role will drive organisational change towards automation-first, measurable, and repeatable practices. A key part of the role is building and evolving reusable CI/CD and Terraform modules, engineering guardrails, observability patterns, and automation frameworks that can be adopted across multiple teams and domains without requiring each team to solve the same problems independently. The Principal SRE also plays an important enablement role beyond deeply technical teams, helping less technical areas of the business adopt structured, governed, and scalable ways of working. This includes translating complex engineering practices into practical standards, improving how governance is implemented through engineering controls rather than manual oversight, and driving operational maturity across a broad and diverse technology landscape. The ideal candidate is a systems thinker who understands how services, networks, identity, data flows, and operational processes fail in real-world conditions, and can apply that understanding to build automation-first, reliability-focused operating models that scale across both technical and non-technical functions. Key Responsibilities Cross-Domain Reliability Engineering Design and evolve reliability patterns across cloud, network, identity, and security domains. Identify systemic risks and failure modes across platforms and services, and define engineering solutions to mitigate them. Ensure operational activities are embedded into delivery models through automation, CI/CD integration, and event-driven workflows. Automation & Toil Reduction at Scale Lead the design of automation frameworks that eliminate manual operational tasks across multiple domains. Translate incident learnings and operational inefficiencies into scalable automation and preventative controls. Drive adoption of automation-first principles, reducing dependency on human-driven processes. Contribute to AI-driven operational use cases, including event correlation, anomaly detection, noise reduction, operational insights, and automated remediation. Ensure AIOps capabilities are grounded in reliable telemetry, clear control boundaries, and measurable operational outcomes. Observability & 24/7 Operational Excellence Define standards for telemetry, monitoring, alerting, and operational visibility across all critical systems. Ensure services are observable, measurable, and support proactive detection of issues. Improve operational readiness, incident response effectiveness, and time-to-recovery through engineering solutions. CI/CD & Platform Integration Contribute to the design of CI/CD patterns that embed reliability, security, and operational controls into pipelines. Ensure infrastructure, network, identity, and security configurations are managed through code and validated automatically. Support integration of platform services into delivery pipelines to enable consistent, repeatable deployments. Security & Identity Integration Contribute to secure-by-design patterns, including least privilege, identity-based access, and short-lived credentials. Support integration of security controls (e.g. secrets management, authentication, policy enforcement) into engineering workflows. Ensure security and compliance requirements are met through engineering controls rather than manual processes. Network & Infrastructure Reliability Support the design of resilient network architectures and segmentation aligned with Zero Trust principles. Ensure network configurations and controls are automated, validated, and observable. Contribute to infrastructure design patterns that improve availability, scalability, and fault tolerance. Design and improve operational patterns for network reliability, segmentation, visibility, and change validation. Support automation and standardisation of network controls and operational procedures to reduce manual intervention and configuration drift. Technical Leadership & Enablement Provide technical leadership across teams, influencing standards, architecture, and engineering practices. Mentor engineers on reliability engineering, automation, and systems thinking. Drive consistency through reusable patterns, frameworks, and documentation. Strategic Influence & Continuous Improvement Contribute to reliability engineering strategy and roadmap across the organisation. Communicate technical concepts, risks, and recommendations to senior stakeholders and leadership. Lead initiatives that improve reliability maturity, engineering efficiency, and operational scalability. Support less technical teams and functions in adopting structured, automated, and measurable operational practices. Act as a bridge between engineering capability and organisational change, helping scale good practice beyond core platform teams. Automated Data Workflows Design and improve automated data workflows that support operational reporting, observability, governance, and decision-making. Ensure operational data pipelines are reliable, timely, and aligned to engineering and business needs. Reusable Engineering Frameworks Build and evolve reusable modules, patterns, and frameworks for CI/CD, Terraform, and operational automation. Embed governance, validation, and reliability controls into these shared engineering assets by default. Governance by Engineering Translate governance requirements into practical engineering controls, automated checks, and repeatable standards. Help teams adopt compliant and supportable operating models without relying on manual policing or process-heavy interventions. What You'll Bring Required Qualifications 10+ years of experience in Site Reliability Engineering, Platform Engineering, or related fields. Strong hands-on experience across multiple domains, including: Cloud platforms (AWS, Azure) CI/CD and Infrastructure-as-Code (e.g. Terraform) Observability tools (e.g. Datadog, Splunk) Automation and scripting (e.g. Python) Experience designing and implementing scalable automation and reliability solutions. Deep understanding of distributed systems, failure modes, and resilience patterns. Experience integrating operational and security controls into engineering workflows. Strong stakeholder engagement and technical communication skills. Preferred Qualifications Experience with identity and access management systems (e.g. Entra ID, Vault). Experience with network architecture and security controls (e.g. firewalls, segmentation). Familiarity with Zero Trust principles and security engineering practices. Experience working in large, federated organisations with diverse technology stacks. Exposure to compliance and regulatory requirements (e.g. PCI, HIPAA, SOX). Additional info Hybrid or on-site work model. Operates as a senior individual contributor with broad cross-organisational influence. Expected to balance hands-on technical leadership with strategic direction. Occasional travel may be required for team or stakeholder engagement. Boston Consulting Group is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, age, religion, sex, sexual orientation, gender identity / expression, national origin, disability, protected veteran status, or any other characteristic protected under national, provincial, or local law, where applicable, and those with criminal histories will be considered in a manner consistent with applicable state and local laws. BCG is an E - Verify Employer. Click here for more information on E-Verify.