Site Reliability Engineer
Job Description
Key Skills Required
Master these to land this role
Want to know if you're a match for this job?
What You'll Do:
Assess service maturity and provide insights to development teams
Partner with development teams to implement observability best practices
Enable development teams to become autonomous with their service deployment, support, and infrastructure
Mentor developers on reliability practices, focusing on making them self-sufficient
Act as the bridge, ear, and eyes of the Platform Division teams to drive tooling and practice adoption across development teams
About You:
Deep understanding of observability practices in a distributed system environment and how it influences system design and team behavior
Practical experience with SRE concepts (SLOs, error budgets, incident management)
3–5+ years in software development, SRE, DevOps, or production development roles with experience operating production systems
Proficient in cloud-native platforms and infrastructure-as-code concepts and tools
Working knowledge of at least one programming language (TypeScript/Node.js is a plus)
Excellent communication and collaboration abilities across technical and non-technical teams
Ability to translate complex reliability concepts into actionable guidance
You enjoy enabling teams to succeed independently and measuring success by reduced dependency on you
How would you rate this job post?
See what other professionals think about this role.
Similar Opportunities
Technical Program Manager - Infrastructure Delivery
Lightning AI
United StatesSenior Product Manager, Core Product Portfolio (Agentic Marketing Platform)
Hightouch
United StatesElectrical Field Engineer
Crusoe
United StatesSenior Technical Program Manager at Pencil
Pencil
United StatesMore Openings at MaintainX
Explore Top Companies in this Space
Contentstack
SaaS / Headless CMS / Digital Experience Platform
Nectar Social
AI Social Commerce Platforms & Social CRMs / Agentic Community Management SaaS / Multimodal Social Listening Tools / E-Commerce Attribution & Marketing Automation
Abacus Insights
Healthcare Data Usability & Interoperability SaaS / Health Plan Analytics & Cloud Data Lakehouses / Payer Strategy & Healthcare IT Systems / CMS & HIPAA Compliance Management
Eastern Research Group
Environmental Consulting / Health & Safety / Regulatory Compliance / Sustainable Infrastructure
MaintainX
View Company ProfileMaintainX is a premier, enterprise-grade maintenance and work execution powerhouse engineered to orchestrate massive-scale frontline operations and intelligent, frictionless asset-management workflows. Operating as a mobile-first "digital-nervous-system" for industrial teams, the company eliminates the operational friction of traditional, legacy paper-based models—which frequently suffer from fragmented maintenance records, opaque equipment-health logic, and disconnected technician-communication pipelines—by seamlessly deploying advanced generative-AI telemetry, rigorous preventive-maintenance architectures, and cohesive cross-site integration frameworks. Moving beyond rigid legacy software paradigms, MaintainX empowers global enterprises—including leaders like Duracell, AB InBev, and McDonald’s—to dynamically synchronize their work-order management, inventory tracking, and regulatory-compliance pipelines with elite, scalable, and audit-ready execution. Under the hood, their sophisticated proprietary operational infrastructure—bolstered by AI-powered procedure generation, anomaly detection, and real-time voice-to-text transcriptions—natively manages complex global multi-site data ingestion, instantaneous automated diagnostic-routing, and compliance-hardened audit-trail workflows, providing the necessary operational foundation for the modern, AI-transformed industrial economy. What sets MaintainX apart is its uncompromising dedication to frictionless "frontline-orchestration"; by bridging the gap between highly technical, performance-intensive maintenance demands and accessible, high-velocity mobile-centric interfaces, the firm empowers modern engineering and facilities organizations to radically accelerate their uptime velocity, eliminate systemic administrative-bottlenecks, and build an unassailable foundation for continuous commercial and institutional dominance in the modern, AI-transformed global-manufacturing landscape.
Safety First
- Never pay for a job application.
- Do not share sensitive bank info.
- Verify the client before starting work.
