{"id":1691576,"url":"https://alion.io/job/sysco-site-reliability-engineer","title":"Site Reliability Engineer","company":{"id":1771007,"name":"Sysco","domain":"sysco.com","url":"https://alion.io/company/sysco-com","size_band":"5000+","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"junior","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Sri Lanka"],"countries":["LK"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":null,"experience_years_min":2,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Agile","optional":false},{"name":"AIOps","optional":false},{"name":"ArgoCD","optional":false},{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"CI/CD","optional":false},{"name":"Datadog","optional":false},{"name":"GCP","optional":false},{"name":"GitHub Actions","optional":false},{"name":"Go","optional":false},{"name":"Grafana","optional":false},{"name":"Helm","optional":false},{"name":"Incident Management","optional":false},{"name":"JavaScript","optional":false},{"name":"Jenkins","optional":false},{"name":"Kubernetes","optional":false},{"name":"OpenTelemetry","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Prometheus","optional":false},{"name":"Python","optional":false},{"name":"Splunk","optional":false},{"name":"Terraform","optional":false},{"name":"TypeScript","optional":false}],"status":"live","first_seen_at":"2026-09-30T00:00:00Z","employer_posted_date":"2026-09-30","last_verified_at":"2026-10-06T01:30:41Z","board_verified":true,"closed_at":null,"days_open":6,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":6},"description":"JOB DESCRIPTION\nSoftware Site Reliability Engineer\nAbout Sysco LABS:\nSysco LABS is the Global In-House Center of Sysco Corporation (NYSE: SYY), the world’s largest foodservice company. Sysco ranks 55th in the Fortune 500 list and is the global leader in the trillion-dollar foodservice industry.\nSysco operates 333 distribution centers across 10 countries, with 75,000 colleagues serving approximately 670,000 customer locations, including restaurants, healthcare and educational facilities, lodging establishments, entertainment venues and more. For fiscal year 2026, which ended June 27, 2026, the company generated sales of more than $84 billion.\nSysco LABS Sri Lanka delivers the technology that powers Sysco’s end-to-end operations.\nSysco LABS’ enterprise technology is present in the end-to-end foodservice journey, enabling the sourcing of food products, merchandising, storage and warehouse operations, order placement and pricing algorithms, the delivery of food and supplies to Sysco’s global network and the in-restaurant dining experience of the end-customer.\nThe Opportunity\nJoin the Sysco Commercial Technology (CT) Site Reliability Engineering team as a Software Site Reliability Engineer, where you will help improve the reliability, scalability, performance, security, and operational excellence of Sysco's Commercial Technology ecosystem. Our team is responsible for delivering end-to-end reliability across the technology stack from cloud infrastructure and platform services to APIs, applications, and customer-facing digital experiences supporting enterprise API platforms, digital commerce, B2B integrations, and other business-critical systems across the CT landscape.\nThis is a Software SRE role focused on applying software engineering principles to solve reliability challenges at scale. You will design and build automation, develop internal tools and platforms, enhance observability, reduce operational toil, and deliver engineering solutions that strengthen the reliability and resilience of distributed systems.\nWorking closely with product, platform, and infrastructure engineering teams, you will contribute to building highly available, scalable, and resilient services while embedding reliability throughout the software development lifecycle.\nResponsibilities:\nOwn the reliability, scalability, performance, and operational excellence of one or more services or platform components across the Commercial Technology ecosystem.\nApply software engineering principles to design and build solutions that improve reliability, resilience, and engineering productivity.\nDesign, develop, and maintain automation, internal tools, self-service capabilities, and reliability engineering workflows.\nDefine, implement, and improve SLIs, SLOs, observability standards, dashboards, alerts, and operational metrics.\nPartner with product, platform, and infrastructure engineering teams to improve service architecture, production readiness, deployment safety, and operational excellence.\nTroubleshoot complex production issues using application code, logs, metrics, traces, APIs, databases, and infrastructure telemetry to identify root causes and implement long-term solutions.\nParticipate in incident response, postmortems, and reliability reviews, ensuring follow-up actions are implemented through engineering improvements.\nImprove deployment automation, release validation, rollback strategies, and production readiness.\nContribute to capacity planning, performance optimization, disaster recovery, and resilience testing.\nParticipate in architecture and design reviews to ensure systems are reliable, scalable, observable, secure, and operationally efficient.\nBuild and enhance shared engineering tools, libraries, and platform capabilities that improve developer productivity and service reliability.\nParticipate in an on-call rotation, using operational insights to continuously improve automation and reduce operational toil.\nRequirements:\n2+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Production Engineering, DevOps, or a related engineering role.\nStrong software engineering skills, with experience designing, developing, testing, and operating production-grade software.\nProficiency in one or more programming languages, such as Java, Go, Python, or JavaScript/TypeScript.\nAbility to read, understand, and debug application code to identify systemic issues and implement long-term engineering improvements.\nSolid understanding of distributed systems, cloud-native architectures, microservices, APIs, databases, messaging systems, caching, networking, and system design.\nExperience supporting large-scale distributed systems, enterprise SaaS solutions, digital commerce platforms, or other high-traffic environments.\nExperience applying Site Reliability Engineering principles, including SLIs, SLOs, error budgets, incident management, blameless postmortems, automation, and toil reduction.\nExperience troubleshooting production systems using application code, telemetry, and infrastructure diagnostics.\nHands-on experience with observability practices, including metrics, logs, traces, and profiling, using platforms such as Datadog, Prometheus, Grafana, OpenTelemetry, Splunk, or ELK.\nExperience with AWS, Azure, or GCP and cloud-native technologies such as containers and Kubernetes.\nExperience with CI/CD, Infrastructure as Code, and platform engineering tools such as Terraform, Helm, Argo CD, Jenkins, or GitHub Actions.\nExperience developing internal developer platforms, automation frameworks, or reliability engineering tools that improve reliability and engineering productivity.\nExperience with service meshes, API gateways, messaging systems, caching, database reliability, or resilience engineering.\nExperience with AIOps, AI-assisted operations, or modern observability platforms.\nStrong analytical, problem-solving, and systems-thinking skills.\nExcellent communication and collaboration skills, with the ability to work effectively across cross-functional engineering teams.\nDemonstrated ownership, curiosity, and a continuous-improvement mindset.\nBenefits:\nPerformance-based annual bonus \nPerformance rewards and recognition \nAgile Benefits - special allowances for Health, Wellness & Academic purposes \nPaid birthday leave \nTeam engagement allowance \nComprehensive Health & Life Insurance Cover - extendable to parents and in-laws \nHybrid work arrangement \nSysco LABS is an Equal Opportunity Employer.","description_format":"text","description_chars":6501,"description_truncated":false,"requirements":{"experience_years_min":2,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Hybrid work","Life insurance"],"hiring_locations":[{"name":"Sri Lanka","iso":"LK","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Commerce"],"lifecycle":[{"event":"open","at":"2026-10-02T10:55:42Z"}],"visa":[],"liveness":{"score":76,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.849,"p_room":0.9,"age_days":5,"expected_fill_days":13,"reasons":["conf:0","velocity","win:mid","comp:junior,brand"],"computed_at":"2026-10-05T05:45:15Z"},"pay":null,"html_url":"https://alion.io/job/sysco-site-reliability-engineer","json_url":"https://alion.io/job/sysco-site-reliability-engineer.json","meta":{"generated_at":"2026-10-06T02:22:42Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":3575,"day_limit":5000,"remaining_today":1425,"minute_limit":60,"resets_at":"2026-10-07T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":1771007},"rest":"https://alion.io/mcp/rest/get_company?id=1771007"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fsysco-site-reliability-engineer"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fsysco-site-reliability-engineer"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fsysco-site-reliability-engineer"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/sysco-site-reliability-engineer\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fsysco-site-reliability-engineer"}]}