HHiring Reality
← Plume

Site Reliability Engineer (SRE) Manager

Plume
Location
Ljubljana, Slovenia
Posted
9 days ago
Department
NOC L2 (Devops/SRE)
What they actually want (must-haves)
  • 4+ years of experience in network operations, infrastructure, or site reliability roles
  • At least 2+ years in a people management or team lead capacity
  • Strong understanding of networking fundamentals (TCP/IP, DNS, routing, firewalls, VPNs)
  • Proven experience owning incident response processes, including on-call rotation design and escalation management
  • Experience with monitoring, alerting, and observability tools (e.g., Grafana, Prometheus, Datadog, PagerDuty, Splunk)
  • Strong track record of driving operational improvements that measurably reduce incident volume or resolution time
Nice to have
  • Experience in telecom, ISP, networking hardware, or connected-device industries
  • Familiarity with cloud infrastructure (AWS, GCP, or Azure) and container orchestration (Kubernetes, Docker)
  • Experience with automation/scripting to reduce manual operational toil (Python, Bash, or similar)
  • Prior experience scaling a NOC or SRE team through significant growth
  • ITIL or similar operational framework certification/experience
What the job really is

As a Site Reliability Engineer (SRE) Manager at Plume, you will lead the NOC L2 team, focusing on managing escalated network and platform incidents. Your role involves mentoring engineers, driving process improvements, and ensuring effective incident response while collaborating with various teams to enhance platform reliability.

Things to weigh
  • No salary listed
  • Role involves managing a 24/7 operational environment which may require flexible hours
  • Focus on operational improvements suggests a potentially high-pressure environment during incidents
Job score2.8/5
Benefits1/5
Freshness5/5
Career value4/5
Role clarity4/5
Pay transparency0/5

Applying to Plume?

See how your résumé matches this role — and tailor it from what actually gets interviews.