{"id":823438,"url":"https://alion.io/job/orkes-site-reliability-engineer","title":"Site Reliability Engineer","company":{"id":686350,"name":"Orkes","domain":"orkes.com","url":"https://alion.io/company/orkes-2","size_band":"201-500","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Greenhouse","truth_index":{"grade":"C","score":56,"open_postings":19,"ghost_share":0.737,"stale_share":0,"repost_share":0,"time_to_fill_p50_days":null,"computed_at":"2026-09-30T05:45:00Z"}},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"posting_text","remote_working_hours":null,"hiring_geo_confidence":"explicit","locations":[],"countries":[],"hiring_countries":["AU"],"hiring_countries_total":1,"salary":{"min":125000,"max":250000,"currency":"USD","period":"year","gross":null,"usd_annual":250000},"salary_estimate":null,"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"CI/CD","optional":false},{"name":"Datadog","optional":false},{"name":"GCP","optional":false},{"name":"Grafana","optional":false},{"name":"Incident Management","optional":false},{"name":"Kubernetes","optional":false},{"name":"OpenTelemetry","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Prometheus","optional":false},{"name":"Python","optional":false},{"name":"Terraform","optional":false},{"name":"Apache Kafka","optional":true}],"status":"live","first_seen_at":"2026-05-14T22:02:43Z","employer_posted_date":"2026-09-04","last_verified_at":"2026-09-30T10:27:43Z","board_verified":true,"closed_at":null,"days_open":139,"trust":{"level":"ghost","repost_count":0,"flags":["stale","company_stale"],"days_open":138},"description":"Who We Are\nBuilding the Backbone of Distributed Applications.\nAt Orkes, we are reimagining how developers build reliable, scalable, event-driven applications without wrangling complex infrastructure. \nBorn from Netflix's battle-tested open-source project, OSS Conductor, our platform is now powering billions of mission critical workflows across fintech, e-commerce, logistics, healthcare, and more.\nOur platform powers billions of mission critical workflows across industries like fintech, e-commerce, logistics, and healthcare helping companies run the core operations of their business with confidence. These are the workflows behind every transaction, every delivery, every patient check-in, quietly orchestrating our digital world.\nWe’re solving some of the hardest infrastructure problems: how do you make distributed systems reliable, observable, and scalable - without every team having to reinvent the wheel? We absorb the complexity of distributed systems, letting teams build with confidence.\nWe’re hiring a SRE to help us evolve and scale our platform. If you’re passionate about clean, efficient architecture, love solving tough distributed systems challenges, and thrive in an environment where your ideas actually shape the product then we’d love to talk.\nAnd the best part is we’re just getting started. \nWhat You’ll Do\nOwn reliability, availability, and performance of production systems running in cloud environments\n\nDefine and monitor SLIs/SLOs and help manage error budgets across the platform\n\nLead incident response efforts including detection, triage, mitigation, and postmortems\n\nImprove observability through logging, monitoring, alerting, and dashboards\n\nAutomate operational workflows and reduce manual toil wherever possible\n\nPartner closely with engineering teams to improve system resiliency and scalability\n\nAssist with capacity planning, infrastructure optimization, and performance tuning\n\nBuild internal tooling, runbooks, and operational best practices\n\nSupport Kubernetes-based infrastructure and distributed systems at scale\n\nAct as an escalation point for complex production and platform issues\n\nWhat We’re Looking For\n5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or related infrastructure roles\n\nStrong experience with cloud platforms such as AWS, GCP, or Azure\n\nHands-on experience with Kubernetes and containerized environments\n\nStrong understanding of distributed systems and microservices architecture\n\nExperience with observability tools such as Prometheus, Grafana, Datadog, ELK, or OpenTelemetry\n\nProficiency with infrastructure automation and scripting (Terraform, Python, Bash, etc.)\n\nExperience managing CI/CD pipelines and deployment automation\n\nStrong troubleshooting and incident management skills\n\nAbility to work cross-functionally and communicate effectively during high-pressure situations\n\nNice to Have\nExperience supporting large-scale SaaS or cloud-native platforms\n\nFamiliarity with workflow orchestration technologies such as Conductor, Temporal, or Camunda\n\nExperience with Kafka, messaging systems, or event-driven architectures\n\nKnowledge of security best practices and cloud infrastructure hardening\n\nOpen-source contributions or strong systems engineering background\n\nWhy Join Orkes?\nWork alongside a deeply technical and collaborative engineering team\n\nSolve complex distributed systems challenges at scale\n\nHigh ownership and impact in a fast-growing company\n\nRemote-friendly culture with strong engineering autonomy\n\nOpportunity to help shape the future of cloud orchestration and AI workflows\n\nMore Details:\nThe base salary for this role is $125,000-250,000 USD. However, pay is variable and when determining compensation, a number of factors will be considered: skills, years of experience, job scope, location/cost of living, and competitive compensation market data for your country/role.\nStart Date: ASAP\n\nStatus: Full-time\n\nType: Remote\n\nLocation: Australia, EMEA\n\nDepartment: Engineering\n\nReports to: Head of Engineering\n\nTravel: 15-20%\nBenefits\nComprehensive health coverage including medical, dental, and vision\n\nFlexible PTO\n\nSupport for personal development\n\nAt Orkes, we are committed to building a team that reflects a rich tapestry of perspectives, identities, and professional experiences. We believe that diversity is not just a checkbox, but a driving force behind innovation, creativity, and success. By embracing a variety of backgrounds, we cultivate an inclusive environment where every team member feels valued and empowered to bring their authentic selves to work. \nJoin us at Orkes and be a part of a team where your unique perspectives are not only welcomed but celebrated. Together we are shaping the future technology by leveraging the strength that comes from embracing diversity in all its forms. Your Journey with us is an opportunity to contribute to something greater and make a lasting impact.","description_format":"text","description_chars":4937,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"Australia","iso":"AU","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Commerce"],"lifecycle":[{"event":"open","at":"2026-09-12T13:41:32Z"}],"liveness":{"score":6,"band":"cold","label":"Long shot","p_open":1,"p_active":0.227,"p_room":0.28,"age_days":138,"expected_fill_days":42,"reasons":["conf:11","stale_co","ghost","win:tail","crowd:"],"computed_at":"2026-09-30T05:45:00Z"},"pay":{"stated_usd_annual":250000,"is_top_pay":true},"html_url":"https://alion.io/job/orkes-site-reliability-engineer","json_url":"https://alion.io/job/orkes-site-reliability-engineer.json","meta":{"generated_at":"2026-10-01T02:39:03Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":1955,"day_limit":5000,"remaining_today":3045,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}