Problem Management Specialist @Vantage Data Centers
All Others
Salary usd 85,000 - 10..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type full-time
Posted 3d ago

[Hiring] Problem Management Specialist @Vantage Data Centers

3d ago - Vantage Data Centers is hiring a remote Problem Management Specialist. πŸ’Έ Salary: usd 85,000 - 105,000 per year πŸ“Location: USA

Role Description

This position will be based remotely in the United States. Site Operations teams manage our buildings under the governance provided by Business Operations, ensuring best-in-class data center services for our customers. Our site teams comprise Critical Facility Engineers who are Vantage's front line for managing critical electrical and mechanical infrastructure operations. Providing our best-in-class services requires in-depth plant equipment knowledge and proficiency in critical facility management processes including Change, Incident, Problem, and Work Management.

Vantage is looking for a resourceful Problem Management Specialist to support the North America Site Operations organization by coordinating the end-to-end Problem Management lifecycle. The role focuses on identifying and eliminating the underlying causes of incidents, through structured root cause analysis, clear ownership, and disciplined follow-through. Serving as a central partner to Site Operations, the Specialist will help ensure that qualifying operational events are converted into complete, factual, and actionable Problem records that strengthen site reliability and reduce the likelihood and impact of recurrence.

Working across Site Operations, the Operations Management Center, Reliability Engineering, Automation Systems, Security, Customer Experience, Legal, vendors, and other technical teams, the Specialist will coordinate root cause analysis tasks, corrective actions, follow on actions, and documentation updates while maintaining traceability to the originating incident. The role will guide teams in documenting evidence-based findings, timelines, contributing factors, and support the preparation and governance of accurate internal and customer-facing RCA reports.

A successful candidate will combine strong facilitation, analysis, and operational judgment with the ability to make Problem Management practical for frontline teams. They will maintain high-quality and auditable data, challenge symptom-based conclusions, and translate technical investigations into concise, executive- and customer-ready narratives.

Qualifications

  • Demonstrated experience applying ITIL Problem Management principles, including problem identification, root cause analysis, known-error management, corrective-action tracking, and formal closure.
  • Bachelor of Science degree in Information Technology, Business Management, related field, or equivalent experience.
  • 3 years of experience in Problem Management or equivalent supporting function.
  • Hands-on experience managing Problem records, problem tasks, known errors, workarounds, and related incident and change records in ServiceNow or a comparable IT service management platform or computerized maintenance management system (CMMS).
  • Ability to facilitate structured investigations with technical and operational teams, challenge symptom-based conclusions, and guide stakeholders toward evidence-based root causes and practical corrective actions.
  • Working knowledge of root cause analysis methods such as the 5 Whys, causal-factor analysis, fault-tree analysis, fishbone diagrams, or equivalent structured techniques.
  • Strong organizational, written and verbal communication skills, with experience coordinating multiple action owners, dependencies, due dates, service-level commitments, and escalations across concurrent investigations.
  • Ability to translate complex technical findings into concise, factual, and audience-appropriate root cause analysis reports, executive updates, customer communications, and lessons learned.
  • Proven ability to work effectively with cross-functional teams, engineering, service management, vendors, and business stakeholders while maintaining objectivity, accountability, and a constructive approach during complex or high-visibility investigations.
  • Advanced skills in Microsoft Office 360 Suite – Excel, Word, Power Point, Project, and Visio.
  • Data Center, high-tech, or rapid growth industry experience is strongly preferred, but not required.

Requirements

  • Manage the end-to-end Problem Management lifecycle for qualifying Site Operations incidents, ensuring Problem records are created, assigned, progressed, and closed in accordance with established standards.
  • Facilitate structured root cause analysis with Site Operations and technical stakeholders, using incident evidence, event timelines, asset history, procedures, and operational data to distinguish root causes from symptoms and contributing factors.
  • Coordinate corrective and preventive actions across Site Operations, Reliability Engineering, the Operations Management Center, Automation Systems, Security, vendors, and other support teams, with clear owners, target dates, dependencies, and escalation paths.
  • Maintain accurate, complete, and auditable Problem, problem task, known error, workaround, outage, work order, and related change information in ServiceNow, with traceability to the originating incident.
  • Support post-incident follow-up for Site Operations P1 and P2 events by validating operational impact, outage details, investigation evidence, and required Problem Management actions.
  • Analyze recurring incidents, common failure modes, cause codes, and cross-site trends to identify systemic risks and recommend prioritized opportunities for prevention and continual improvement.
  • Publish and communicate lessons learned, known errors, workarounds, and validated corrective actions so that improvements are adopted consistently across Site Operations and applicable risks are addressed fleet-wide.
  • Identify and participate in activities focused on process improvement, automation, tools implementation, and QA testing.

Benefits

  • Salary Range: $85,000-$105,000 Base + Bonus (this range is based on Colorado market data and may vary in other locations).
  • This position is eligible for company benefits including but not limited to medical, dental, and vision coverage, life and AD&D, short and long-term disability coverage, paid time off, employee assistance, participation in a 401k program that includes company match, and many other additional voluntary benefits.
  • Compensation for the role will depend on a number of factors, including your qualifications, skills, competencies, and experience and may fall outside of the range shown.
Before You Apply
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Problem Management Specialist @Vantage Data Centers
All Others
Salary usd 85,000 - 10..
Remote Location
πŸ‡ΊπŸ‡Έ USA Only
Employment Type full-time
Posted 3d ago
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
️
πŸ‡ΊπŸ‡Έ Be aware of the location restriction for this remote position: USA Only
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 125,000+ Remote Jobs
Γ—
Apply to the best remote jobs
before everyone else

Access 125,000+ vetted remote jobs and get daily alerts.

4.9 β˜…β˜…β˜…β˜…β˜… from 500+ reviews

⚑ 129,517+ remote jobs, refreshed hourly

πŸ”” Real-time alerts: Apply first, direct to employer

πŸ›‘οΈ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later