368,611open jobs
9,439companies
50,719added this week
Browse all
Salary
$98k – $199k per year (Estimated)
Location
Remote/Hybrid (Orlando, United States)
Seniority
Senior · 12+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
NBCUniversal Careers is the official global job portal and talent acquisition hub for NBCUniversal, one of the world's leading media and entertainment enterprises. The platform connects candidates with diverse professional opportunities, internships, and early career programs across iconic brands such as Universal Pictures, NBC News, Peacock, and Universal Destinations & Experiences.

NBCUniversal is one of the world's leading media and entertainment companies. We create world-class content, which we distribute across our portfolio of film, television, and streaming, and bring to life through our global theme park destinations, consumer products, and experiences. We own and operate leading entertainment and news brands, including NBC, NBC News, NBC Sports, Telemundo, NBC Local Stations, Bravo, and Peacock, our premium ad-supported streaming service. We produce and distribute premier filmed entertainment and programming through our powerhouse film and television studios, including Universal Pictures, DreamWorks Animation, and Focus Features, and the four global television studios under the Universal Studio Group banner, and operate industry-leading theme parks and experiences around the world through Universal Destinations & Experiences, including Universal Orlando Resort, home to Universal Epic Universe, and Universal Studios Hollywood. NBCUniversal is a subsidiary of Comcast Corporation. Visit www.nbcuniversal.com for more information.

Our impact is rooted in improving the communities where our employees, customers, and audiences live and work. We have a rich tradition of giving back and ensuring our employees have the opportunity to serve their communities. We champion an inclusive culture and strive to attract and develop a talented workforce to create and deliver a wide range of content reflecting our world.

The Staff Reliability Engineer (SRE) for Workplace Engineering is responsible for the reliability, performance, security, and operational excellence of enterprise workplace collaboration & endpoint services used globally by employees and partners. This role applies an engineering mindset to operations-defining service level indicators/objectives (SLIs/SLOs), reducing toil through automation, improving observability, and strengthening incident response-to ensure a consistent, high-quality collaboration experience across messaging, meetings, voice, file sharing, knowledge sharing, device management platforms & Copilot / AI engineering.

  • Microsoft 365: Teams (chat, meetings, webinars, Teams Phone), SharePoint Online, OneDrive, Exchange Online, Microsoft Entra ID (Azure AD), Microsoft Purview, Defender for Office 365, Intune (Endpoint Management).
  • Hybrid messaging and identity integrations (as applicable): Exchange Server, directory synchronization, mail flow and routing
  • Collaboration endpoints and devices: Teams Rooms, certified headsets/cameras, conference room AV integrations
  • Ecosystem integrations: Power Platform (Power Automate/Apps), Graph API, third-party conferencing/messaging where in use (e.g., Zoom/Slack), mail hygiene/security gateways
  • Architect and optimize global Microsoft Intune and Jamf Pro environments.
  • Orchestrate Windows Updates for Business (WUfB), third-party application patching, and compliance policies to maintain a hardened security posture
  • Automated packaging and deployment of Windows applications, maintaining a rigorous cadence for third-party updates.
  • leverage PowerShell and Graph API to automate repetitive configuration tasks and self-healing remediations.
  • Partner with Security Operations to remediate vulnerabilities.
  • Develop and enforce Configuration Profiles, Compliance Policies, and Conditional Access rules
  • Own the reliability and scaling of Azure Virtual Desktop (AVD) and Windows 365 (Cloud PC), optimizing for both performance and cost-efficiency.
  • Define and operationalize SLIs/SLOs and error-budget policies for collaboration services (Teams chat/meetings/voice, SharePoint/OneDrive, Exchange) with clear customer-impact measurements.
  • Own end-to-end reliability engineering: capacity planning, performance tuning, resilience reviews, dependency mapping, and proactive risk reduction for critical collaboration journeys.
  • Demonstrated expertise in developing, operationalizing, and scaling AI engineering capabilities, including platform design, model lifecycle management, automation, reliability, and enterprise adoption.
  • Strong knowledge of AI governance frameworks, with experience establishing guardrails for responsible AI use, risk management, security, compliance, data controls, and ongoing operational oversight.
  • Build and evolve observability for collaboration platforms: health dashboards, telemetry standards, alert strategy (high signal/low noise), and synthetic monitoring aligned to user experience.
  • Lead incident response for high-severity events: establish incident roles, drive rapid triage/mitigation, coordinate cross-team communication, and produce blameless post-incident reviews with durable corrective actions.
  • Engineer automation to reduce operational toil: provisioning, policy/config drift detection, lifecycle management, reporting, and remediation using PowerShell and APIs; establish reusable runbooks and self-service patterns.
  • Strengthen change and release practices: production readiness reviews, controlled rollouts, maintenance windows, validation plans, and rollback strategies to reduce customer impact.
  • Partner with Security/Compliance to ensure collaboration services meet governance requirements (identity and access, DLP, retention, eDiscovery, information protection), while balancing usability and reliability.
  • Provide Staff-level technical leadership: set engineering standards, mentor engineers, influence roadmap priorities, and align stakeholders on reliability tradeoffs and investment.
  • Establish and lead reliability operating mechanisms (on-call standards, incident command readiness, postmortem quality, action-item governance, and quarterly reliability reviews) to improve consistency across teams.
  • Coach, mentor, and sponsor engineers across levels: provide technical guidance, review designs and postmortems, and raise the bar on documentation, runbooks, and operational readiness.
  • Drive cross-organization alignment on reliability priorities and investment by presenting trends, risks, and proposals to leadership; secure commitments and ensure delivery against measurable outcomes.
  • Serve as an escalation point for complex, cross-domain issues spanning identity, messaging, endpoints, and network dependencies; engage vendors as needed and ensure issues are driven to resolution.
  • 12+ years of experience in reliability engineering, systems engineering, DevOps, or large-scale collaboration/communications operations (enterprise or SaaS), including ownership of production services
  • Deep expertise with collaboration platforms and ecosystems: Microsoft 365 (Teams-including voice/meetings/Rooms-SharePoint Online, OneDrive, Exchange Online) and their dependencies (identity, endpoints, networking)
  • Hands-on experience defining SLIs/SLOs, building observability (metrics/logs/traces), and operating an incident management program (on-call, severity model, communications, postmortems)
  • Strong automation skills with PowerShell and APIs (Microsoft Graph preferred); ability to build tooling that improves reliability and reduces toil
  • Experience with cloud identity and access (Microsoft Entra ID/Azure AD, Conditional Access, MFA, RBAC/PIM) and collaboration governance (Purview, DLP, retention, eDiscovery) preferred
  • Bachelor’s degree in Computer Science/Engineering (or equivalent practical experience)

Desired Characteristics

  • Executive-level written and verbal communication skills; able to translate reliability data into clear decisions, tradeoffs, and action plans
  • Proven ability to influence across functions (Security, Network, End User Computing, Architecture, Product/Program) without formal authority
  • Strong systems thinking and customer satisfaction, focuses on user journeys (chat, meetings, voice, file sharing) and measurable experience outcomes
  • Demonstrated technical leadership through mentorship, sponsorship, and talent development; builds inclusive, high-performing engineering culture
  • High bar for operational excellence, insists on clear ownership, durable fixes, strong postmortems, and measurable follow-through
  • Comfort operating in ambiguity and driving large, multi-quarter improvements with measurable results

Hybrid: This position currently has a hybrid schedule, which requires contributing from the office a minimum of four days per week. The Company reserves the right to change in-office requirements at any time.

As part of our selection process, external candidates may be required to attend an in-person interview with an NBCUniversal employee at one of our locations prior to a hiring decision. NBCUniversal's policy is to provide equal employment opportunities to all applicants and employees without regard to race, color, religion, creed, gender, gender identity or expression, age, national origin or ancestry, citizenship, disability, sexual orientation, marital status, pregnancy, veteran status, membership in the uniformed services, genetic information, or any other basis protected by applicable law.

If you are a qualified individual with a disability or a disabled veteran and require support throughout the application and/or recruitment process as a result of your disability, you have the right to request a reasonable accommodation. You can submit your request to [email protected].

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,611 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Orlando
Remote • Full-Time • 3+ years exp • Cairo
C#
JavaScript
TypeScript
C#
.NET
AI/ML
Copilot
DevOps
Azure
Azure DevOps
CI/CD
Rest API
Analytics
ETL/ELT
Management
Power Apps
Power Automate
Apply
Remote • Full-Time • 3+ years exp • Cairo
C#
JavaScript
TypeScript
C#
.NET
AI/ML
Copilot
DevOps
Azure
Azure DevOps
CI/CD
Rest API
Analytics
ETL/ELT
Management
Power Apps
Power Automate
Apply
Remote • Full-Time • 1+ year exp • Cairo
C#
JavaScript
TypeScript
C#
.NET
AI/ML
Copilot
DevOps
Azure
Azure DevOps
CI/CD
Rest API
Analytics
ETL/ELT
Management
Power Apps
Power Automate
Apply
$110k – $131k per year • Remote • Full-Time • 10+ years exp • Bachelor's Degree
DevOps
AWS
Incident Management
VMWare
Apply
$217k – $304k per year • Equity • Remote • Full-Time • 8+ years exp
Go
Databases
Apache Kafka
ClickHouse
Google BigQuery
AI/ML
Flink
Recommender Systems
DevOps
Incident Management
Kubernetes
Apply
$90k – $105k per year • In office • Full-Time • 2+ years exp • New York
Management
Linear
Apply
Sr Product Manager 3 days ago
$120k – $160k per year • In office • Full-Time • 8+ years exp • Bachelor's Degree • Universal City
AI/ML
AI Agents
LLM
DevOps
OpenTelemetry
Analytics
Power BI
Management
Airtable
Confluence
Jira
Marketing
Amplitude
Apply
$65k – $80k per year • In office • Full-Time • 2+ years exp • Bachelor's Degree • Universal City
Apply
Product Manager 5 days ago
$110k – $145k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • Hollywood
AI/ML
Claude
Claude Code
Copilot
Cursor
AI Agents
DevOps
GitHub
Design
Figma
Management
Confluence
Jira
Apply
Product Manager 5 days ago
$110k – $145k per year • Remote/Hybrid • Full-Time • 5+ years exp • Bachelor's Degree • New York
AI/ML
Claude
Claude Code
Copilot
Cursor
AI Agents
DevOps
GitHub
Design
Figma
Management
Confluence
Jira
Apply
$120k – $140k per year • In office • Full-Time • Charlotte • Raleigh • Dallas • Boston • New York
SQL
Analytics
ETL/ELT
Apply
$73k – $133k per year • In office • Full-Time • 5+ years exp • Bachelor's Degree • Orlando
Bash
PowerShell
Python
DevOps
Ansible
Configuration Management
Docker
Git
Kubernetes
Kustomize
OpenShift
Puppet
Red Hat
Terraform
VMWare
Management
Confluence
Jira
Apply
$121k – $338k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • Wayne • Orlando • Washington • Atlanta • Hartford
AI/ML
Recommender Systems
Apply
$67k – $82k per year • Equity • In office • Full-Time • 3+ years exp • High School Diploma • Richmond • Orlando
DevOps
SLI/SLO/SLA
Apply
$100k – $150k per year • In office • Full-Time • 7+ years exp • Bachelor's Degree • Chicago • Charlotte • Los Angeles • Charleston • Washington
Python
SQL
Databases
Databricks
PostGIS
Snowflake
PostgreSQL
Analytics
Power BI
Tableau
Apply
See all jobs
This is one of many
368,611 more open roles from verified company boards, updated every day.