{"id":1405750,"url":"https://alion.io/job/mangoapps-senior-cloudops-site-reliability-engineer","title":"Senior CloudOps & Site Reliability Engineer","company":{"id":1838497,"name":"MangoApps","domain":"mangoapps.com","url":"https://alion.io/company/mangoapps","size_band":"51-200","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Schema","truth_index":{"grade":"B","score":75,"open_postings":3,"ghost_share":0,"stale_share":1,"repost_share":0,"time_to_fill_p50_days":null,"computed_at":"2026-10-01T05:45:00Z"}},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Seattle, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":123000,"max_usd":239000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":1020},"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Amazon EKS","optional":false},{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"GCP","optional":false},{"name":"IAM","optional":false},{"name":"Kubernetes","optional":false},{"name":"CI/CD","optional":true},{"name":"DNS","optional":true},{"name":"Docker","optional":true},{"name":"Linux","optional":true}],"status":"live","first_seen_at":"2026-06-29T22:19:09Z","employer_posted_date":"2026-06-29","last_verified_at":"2026-09-29T00:28:30Z","board_verified":false,"closed_at":null,"days_open":93,"trust":{"level":"stale","repost_count":0,"flags":["stale"],"days_open":93},"description":"MangoApps runs an enterprise SaaS platform that thousands of customers depend on every day. We're hiring a senior, hands-on engineer to own the reliability, availability, security, and performance of that platform in production, primarily on AWS.\nThis is a deep individual-contributor role, not a management track. You'll spend your time in the systems: tuning infrastructure, building observability, automating away toil, and leading the technical response when production is on the line. You'll influence how the rest of engineering builds and operates reliable services through your work and your judgment, not through a reporting line.\nIf you get satisfaction from understanding a production system end to end, finding the real root cause instead of the convenient one, and making the next incident less likely, this role is built for you.\nWho you are:\n5+ years of hands-on Cloud Operations and Site Reliability Engineering, operating production-scale SaaS (not pre-production or internal-only systems).\nYou operate AWS at production scale today and can speak in specifics about AWS compute, networking, IAM, EKS/Kubernetes, and the operational realities of running real workloads there. This is a hard requirement.\nA second cloud (Google Cloud or Azure) is a plus, not a substitute. We value it, but AWS depth is what the role is focused on.\nYou debug Linux at the level of \"why is this latency spike happening,\" not just \"restart the service.\"\nYou reach for automation by reflex. Manual operational work bothers you, and you've built the tooling to remove it.\nWhat will you own:\nReliability & incident response - Keep the platform available and performant. Define and continuously sharpen monitoring, alerting, and observability. Lead production troubleshooting, drive root-cause analysis, and run post-incident reviews that actually change the system afterward. Participate in the on-call rotation for the services you own.\nAWS infrastructure & operations - Design, deploy, and optimize our AWS infrastructure - compute, storage, networking, DNS, load balancing, and security services. Drive architecture improvements for reliability, scalability, performance, and cost. Own disaster recovery and business-continuity processes, and prove they work before you need them.\nContainers & orchestration - Build and operate containerized workloads on Docker with a focus on security, performance, and predictable scaling across environments.\nAutomation & Infrastructure as Code Provision - Build the scripts and tooling that make operations boring and repeatable.\nCI/CD & release engineering - Maintain CI/CD pipelines and deployment automation. Partner with engineering to make releases safer, faster, and easier to roll back.\nSecurity & compliance - Apply cloud security practices across IAM, network security, secrets management, and vulnerability remediation. Keep infrastructure aligned to our internal security standards and compliance obligations.\nWhy Join MangoApps?\nYou'll have a real impact. We have strong product-market fit, solving communication and collaboration challenges for organizations of every size around the world.\nOur work is recognized. Leading analysts like IDC, Forrester, and Gartner, and independent review sites like Capterra, consistently rate our products highly.\nYou'll keep growing. MangoApps is a genuinely collaborative place, and careers here come with steady learning and growth opportunities.\nYou'll get to solve hard problems. If complex cloud infrastructure and reliability challenges energize you, you'll feel at home.\nWe're flat. We keep the structure lean and treat everyone as an equal.","description_format":"text","description_chars":3631,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Growth opportunities"],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Artificial Intelligence","Cybersecurity","Professional Services"],"lifecycle":[{"event":"open","at":"2026-09-28T17:42:25Z"}],"liveness":{"score":11,"band":"cold","label":"Long shot","p_open":0.9,"p_active":0.454,"p_room":0.28,"age_days":93,"expected_fill_days":42,"reasons":["conf:53","urgency","win:tail","crowd:"],"computed_at":"2026-10-01T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/mangoapps-senior-cloudops-site-reliability-engineer","json_url":"https://alion.io/job/mangoapps-senior-cloudops-site-reliability-engineer.json","meta":{"generated_at":"2026-10-01T18:56:28Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1187,"day_limit":5000,"remaining_today":3813,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}