SaaS Monitoring Engineer CGEMJP00345837

Company: Experis – ManpowerGroup
Apply for the SaaS Monitoring Engineer CGEMJP00345837
Location: London
Job Description:

Role Title: SaaS Monitoring Engineer

Duration: Contract to run until 18/12/2026

Location: London Hybrid, 3 days onsite

Rate: Per day depending on experience (Umbrella in side IR35)

Clearance required: BPSS

Role Purpose / Summary

We are seeking a highly motivated and detail‑oriented SaaS Monitoring Engineer to join our growing cloud operations team. In this role, you will be responsible for designing, implementing, and maintaining monitoring solutions that ensure the health, performance, and reliability of our Software‑as‑a‑Service platforms. You will play a critical part in proactively identifying issues, minimizing downtime, and enabling data‑driven decision‑making through real‑time observability.

A key responsibility of this position is building and maintaining a centralised console dashboard that provides a comprehensive, real‑time view of SaaS service health, system performance, and key operational metrics.

Key Responsibilities

Console Dashboard Development

  • Design and create a centralised console dashboard that provides a real‑time overview of the health of all SaaS services.
  • Ensure the dashboard displays actionable insights, including service uptime, API performance, incident alerts, and dependency status.
  • Optimise dashboard usability by tailoring views for different stakeholders (engineering, operations, leadership).
  • Integrate data from multiple sources into a unified visualisation platform for seamless monitoring.

Incident Management & Troubleshooting

  • Set up intelligent alerting mechanisms to detect anomalies and performance degradation.
  • Investigate and troubleshoot incidents quickly to identify root causes and implement permanent fixes.
  • Collaborate with DevOps and engineering teams during incident response and post‑mortem reviews.

Automation & Optimization

  • Automate monitoring processes, alert escalation, and response workflows.
  • Continuously refine alert thresholds to reduce noise and improve signal accuracy.
  • Implement predictive monitoring techniques to anticipate potential outages.

Collaboration & Communication

  • Work closely with software engineers, DevOps, and product teams to ensure monitoring requirements are embedded early in the development lifecycle.
  • Provide insights and reporting on SaaS performance trends and operational risks.
  • Document monitoring strategies, configurations, and best practices.

Required Skills & Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or related field (or equivalent experience).
  • Proven experience in monitoring cloud‑based or SaaS environments.
  • Strong understanding of distributed systems, microservices architecture, and cloud platforms (AWS, Azure, or GCP).
  • Hands‑on experience with monitoring and visualisation tools (e.g., Grafana, Prometheus, ELK stack, Splunk, Datadog).
  • Experience building interactive dashboards and console‑based monitoring systems.
  • Proficiency in scripting or programming languages such as Python, Bash, or Go.
  • Familiarity with containerisation and orchestration tools (Docker, Kubernetes).
  • Strong analytical and problem‑solving skills with attention to detail.

Preferred Qualifications

  • Experience with Site Reliability Engineering (SRE) principles and practices.
  • Knowledge of CI/CD pipelines and DevOps methodologies.
  • Exposure to AIOps or machine learning‑based monitoring tools.
  • Certification in cloud platforms or monitoring technologies.

Key Competencies

  • Proactive mindset with a focus on preventing issues before they occur.
  • Ability to work in a fast‑paced, highly collaborative environment.
  • Strong communication skills, capable of translating technical insights into actionable business outcomes.
  • Passion for operational excellence and continuous improvement.

All profiles will be reviewed against the required skills and experience. Due to the high number of applications we will only be able to respond to successful applicants in the first instance. We thank you for your interest and the time taken to apply!

#J-18808-Ljbffr…

Posted: June 26th, 2026