This one is closed
Live roles like this one
-
B
2w ago
Senior Storage Software Engineer - DGX Cloud
NVIDIA United States $224k - $431k/yr
See every "Neocloud Service Delivery" role →
Get new “Neocloud Service Delivery” roles by email
One email a day with what is new in "Neocloud Service Delivery". Nothing new, no email.
We confirm the address first, and every mail carries an unsubscribe link. Alerts are ours, not a third party's.
Why this grade This listing scored 30/100, which is an F. It lost the most ground on freshness. See the breakdown
- Description depth 20 / 20 How much the posting actually says about the work, measured in characters of real text.
- Pay transparency 12 / 25 A published salary range, worth more than any other single factor because it is what a candidate cannot find out without applying.
- Corroboration 10 / 10 Whether more than one source carries this listing.
- Remote clarity 3 / 15 Whether "remote" means anywhere, or is quietly restricted to one country.
- Role specificity 0 / 10 Whether the listing is tagged well enough to tell what the role actually is.
- Freshness 0 / 15 How recently it was posted. Older postings are likelier to be filled or abandoned.
-15 Ghost-job penalty — Deducted for signals that this posting may not be a real, currently-open role — staleness, repeated relisting, or talent-pool language.
Every figure above is arithmetic over the posting itself — its salary field, its text, its age, its tags and how many sources carry it. How the grades work →
About the Role
Mirantis is expanding into Neocloud Infrastructure-as-a-Service — managing large-scale infrastructure to a high SLA on behalf of customers running demanding compute workloads, with the scope growing over time to cover the full stack up to the platform layer. We're looking for a Neocloud Service Delivery Manager to lead the engineering team delivering this service, own our incident and outage management process, and make sure our customers consistently get the level of service we commit to.
This is an operational leadership role. You'll manage a team of engineers responsible for infrastructure and, over time, the platform layer running on top of it. You'll own the incident lifecycle end-to-end and act as the escalation point when things go wrong — while building the reporting, process, and team rhythm that reduces how often they go wrong in the first place.
Key Responsibilities
Team Management
Manage a team of engineers responsible for large-scale infrastructure and platform operations
Own rostering and coverage across a 24x7, globally distributed operating model, ensuring continuous operational coverage
Run team performance management: 1:1s, goal-setting, skills development, and performance reviews
Identify and close skills gaps within the team as scope grows from infrastructure into platform operations
Own onboarding of new engineers into live customer environments
Service Delivery & SLA Management
Own service delivery performance against contracted SLAs across your customer portfolio
Define, track, and report on service KPIs (availability, MTTR, MTTA, ticket aging/backlog, customer satisfaction)
Chair regular service review meetings, both internal and customer-facing, backed by clear data
Maintain and continuously improve runbooks, escalation paths, and operational documentation
Manage SLA risk and escalate commercial impact to leadership where relevant
Incident & Critical Outage Management
Own the incident management process end-to-end: detection, triage, escalation, resolution, and post-incident review
Act as the primary escalation point for major and critical incidents/outages, coordinating cross-functional resources and communicating status to customers and leadership
Understand the technical detail behind incidents well enough to ask the right questions, challenge root-cause analysis, and make good calls under pressure
Run blameless post-incident reviews (RCA/PIR) and track corrective actions through to closure
Track incident trends over time to drive proactive reliability improvements rather than reactive firefighting
Ensure on-call and escalation rotas are staffed, documented, and regularly tested
Customer & Stakeholder Management
Act as a senior operational point of contact for key customers, building trust through consistent, transparent delivery
Partner with Sales and Solutions Architecture during onboarding of new customers to confirm delivery readiness
Represent service delivery performance and improvement plans in customer business reviews
Process & Continuous Improvement
Drive continuous improvement across monitoring, observability, and incident-management tooling to reduce manual effort and improve detection speed
Maintain structured onboarding/handover documentation for new customer environments
Ensure operational readiness reviews are completed before any new customer environment goes live
What Success Looks Like
SLA targets consistently met or exceeded across your customer portfolio
Incidents detected, escalated, and resolved within target timeframes, with clear customer communication throughout
A stable, well-rostered team with a visible skills growth path as scope expands from infrastructure to platform
A shrinking rate of repeat incidents, driven by disciplined post-incident follow-through
Customers who trust the team because they can see the data, not just hear reassurance
Required Skills & Experience
- Proven experience managing technical operations or managed services teams responsible for large-scale infrastructure and/or platform operations
- A technical background of some kind — you don't need to be the person fixing the issue, but you need to understand the technicalities behind incidents well enough to lead the response credibly
- Direct experience owning incident management and major incident/outage response, including customer-facing communication during live incidents
- Strong working knowledge of SLA frameworks, service reporting, and escalation management
- Experience managing distributed or shift-based teams operating on a 24x7 rotation
- Strong stakeholder management skills, able to communicate clearly to both engineers and customer executives
- Data-driven approach to service management, comfortable building and presenting operational metrics
- Must be based in the EU and eligible to work there
Preferred Skills & Experience
- Experience with observability tooling and practices (monitoring, logging, tracing, alerting) — a strong plus
- Familiarity with infrastructure and platform technologies
- ITIL or equivalent service management certification
- Experience in a neocloud, hyperscaler, colocation, or managed hosting environment
- Experience managing services for customers who own their own physical infrastructure
- Experience managing teams whose remit expanded over time from infrastructure into platform-layer operations
What does Mirantis offer you?
- Work with an established Silicon Valley leader in the cloud infrastructure industry;
- Work with exceptionally passionate, talented and engaging colleagues, helping Fortune 500 and Global 2000 customers implement next-generation cloud technologies;
- Be a part of cutting-edge, open-source innovation;
- Thrive in the high-energy environment of a young company where openness, collaboration, risk-taking, and continuous growth are valued;
- Professional development and training;
- Attend conferences and working groups;
- Company outings, happy hours, hackathons, and tech talks;
- Receive a competitive compensation package with a strong benefits plan.
It is understood that Mirantis, Inc. may use automated decision-making technology (ADMT) for specific employment-related decisions. Opting out of ADMT use is requested for decisions about evaluation and review connected with the specific employment decision for the position applied for. You also have the right to appeal any decisions made by ADMT by sending your request to
submitting your resume, you consent to the processing and storage of your personal data in accordance with applicable data protection laws, for the purposes of considering your application for current and future job opportunities.
We are a Leader for Container Management in G2 (#2 after AWS)!
We are a Leader for Container Management in G2 (#2 after AWS)!
About Mirantis
Mirantis is the Kubernetes-native AI infrastructure company, enabling organizations to build and operate scalable, secure, and sovereign infrastructure for modern AI, machine learning, and data-intensive applications. By combining open source innovation with deep expertise in Kubernetes orchestration, Mirantis empowers platform engineering teams to deliver composable, production-ready developer platforms across any environment—on-premises, in the cloud, at the edge, or in sovereign data centers. As enterprises navigate the growing complexity of AI-driven workloads, Mirantis delivers the automation, GPU orchestration, and policy-driven control needed to manage infrastructure with confidence and agility. Committed to open standards and freedom from lock-in, Mirantis ensures that customers retain full control of their infrastructure strategy.
Mirantis serves many of the world’s leading enterprises, including Adobe, DocuSign, Liberty Mutual, PayPal, Reliance Jio, Societe Generale, Splunk, and Volkswagen. Learn more at .
Originally posted on Himalayas
Apply for this role Opens jobs.smartrecruiters.com — the employer's own page, as linked by the source that listed this role
Quick question · anonymous · one tap
Would you apply to this job?
Answer to see what other job seekers said.
Your turn · no account needed
Help the next applicant
You may know something about this listing that we cannot see from here. One tap. No account needed. Signed-in reports earn points once the evidence agrees with you.
I know what it pays
Sign in with Google to earn points for reports — 100 confirmed points buy a week of Early Access.
Where this listing came from
- 21 Aug 2026 VC talent network employer ATS first sighting
- 05 Sep 2026 Himalayas also listed, 15 days later
Seen on 2 boards over 15 days.