
Job Overview
Location
Home Working, GB
Job Type
Full-time
Category
Software Engineering
Date Posted
July 4, 2026
Full Job Description
đź“‹ Description
- • Ensure the reliability, availability, performance, and scalability of Azure-based cloud platforms and applications through proactive SRE practices
- • Monitor cloud applications and services using telemetry, metrics, logging, and observability tools to detect and resolve issues before they impact customers
- • Define, maintain, and continuously improve Service Level Objectives (SLOs), Service Level Indicators (SLIs), and Service Level Agreements (SLAs) in collaboration with engineering, product, and operations teams
- • Lead incident response activities, including outage investigations, post-incident reviews, and root cause analysis to drive continuous service improvement
- • Automate operational processes such as deployments, recovery procedures, monitoring, and documentation to reduce manual toil and operational risk
- • Embed DevOps and reliability best practices throughout the software development lifecycle by partnering with software engineering and product teams
- • Enhance platform resilience, scalability, and operational maturity using data-driven insights and SRE methodologies
- • Champion a culture of operational excellence, service ownership, and continuous learning across technical teams
- • Improve system observability by establishing and enforcing standards for monitoring, alerting, and incident management
- • Collaborate globally with engineering, product, and operations teams to deliver new features and platform enhancements without compromising stability or customer experience
- • Contribute to the evolution of cloud platform operations by identifying opportunities for automation, resilience improvements, and operational efficiency
- • Act as a technical leader for reliability initiatives, influencing organizational practices and promoting accountability for system health
- • Support the development and maintenance of runbooks, playbooks, and operational documentation to ensure consistent incident response and knowledge sharing
- • Apply error budgeting principles to balance innovation velocity with system reliability targets
- • Work within a regulated, enterprise-scale environment supporting critical scientific and industrial customer solutions
🎯 Requirements
- • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related discipline, or equivalent practical experience
- • Significant experience in Site Reliability Engineering, DevOps, Cloud Operations, or a similar role supporting enterprise-scale platforms
- • Strong hands-on experience operating and supporting production workloads on Microsoft Azure
- • Expertise in monitoring, alerting, observability platforms, and cloud-native operational practices
- • Deep understanding of SRE principles including SLOs, SLIs, SLAs, error budgets, and reliability engineering methodologies
- • Experience with automation and development using technologies such as C#, Python, PowerShell, and Infrastructure as Code (IaC)
🏖️ Benefits
- • Competitive salary and comprehensive benefits package
- • Opportunities for professional development and technical growth
- • Support for industry-recognized certification programs
- • Work alongside talented global engineering, product, and operations teams
- • Role in shaping the reliability of cloud platforms supporting critical business operations
- • Inclusive, equal opportunity culture that values diversity and inclusion
Skills & Technologies
See exactly how your profile matches this role — strengths, skill gaps, and what to do about them.
About Malvern Panalytical Ltd
Malvern Panalytical is a leading provider of analytical solutions for materials characterization. The company offers a comprehensive range of instruments and software for particle size and shape analysis, rheology, molecular weight determination, and surface area analysis. Their technologies are used across various industries, including pharmaceuticals, chemicals, materials science, and food and beverage, to help researchers and manufacturers understand and control material properties. Malvern Panalytical is part of Spectris plc, a global technology group focused on precision measurement.
Subscribe to the weekly newsletter for similar remote roles and curated hiring updates.
Newsletter
Weekly remote jobs and featured talent.
No spam. Only curated remote roles and product updates. You can unsubscribe anytime.
Similar Opportunities

Anyone AI Inc.
2 months ago

SimSpace Corporation
2 months ago
2 months ago

Compa Technologies, Inc.
2 months ago
