SRE Engineer
Skills
About this role
About Ricoh
A global leader in digital services, recognised for innovation, sustainability and a people-first culture. We feature in the Gartner Magic Quadrant , are listed in the Global 100 Most Sustainable Companies , and have been named one of Forbes’ World’s Best Employers 2025 .
At Ricoh, we believe people do their best work when they feel valued and supported. We create inclusive workplaces where you can grow, contribute, and make a positive impact while helping to build a more sustainable future.
Find your place. Transform your future
Our purpose is centred on understanding and improving how people work. By focusing on real working experiences, we support individuals to develop their skills, realise their potential and do work that feels meaningful.
People transform when they Love What They Do
This belief sits at the heart of The Ricoh Promise. It guides how we recruit, how we support our people, and how we work together every day, creating an environment where you can grow, feel valued and make a difference.
When you join us, you are encouraged to share your ideas, challenge the way things are done, and work with others to build something better. If you are looking for a place where your voice is heard, your development is supported, and your work feels meaningful, you will feel at home at Ricoh.
What you will be doing
As a Site Reliability Engineer, you will play a key role in driving reliability, performance, and operational excellence across Ricoh’s hybrid cloud and on‑prem environments. You will help shape SRE practices, support incident and problem management, embed automation, and ensure infrastructure operations meet the highest security and compliance standards, including ISO 27001.
This is a hands‑on technical role with significant influence across engineering, architecture, security, and operational teams.
Responsibilities Include:
• Delivering against SLIs, SLOs and managing error budgets for core services • Implementing standards for availability, latency, performance, capacity, and scalability • Leading and contributing to root‑cause analysis and major incident reviews • Supporting a blameless post‑mortem culture with clear action tracking • Defining and implementing SRE practices, tooling, and engineering standards • Driving infrastructure‑as‑code and automation across Azure and on‑prem • Improving image bakery pipelines for secure, repeatable server builds • Embedding observability using metrics, logs, traces, and effective alerting • Ensuring all practices align with ISO 27001 and internal security frameworks • Managing automated patching, vulnerability remediation and configuration compliance • Building dashboards and KPIs for reliability, MTTR, change failure rate, capacity and operational trends • Reducing operational toil through automation and improved tooling • Supporting the delivery and evolution of the SRE roadmap aligned to Ricoh’s transformation strategy