{"id":46901,"url":"https://alion.io/job/cisco-systems-staff-site-reliability-engineer-sre-hybrid","title":"Staff Site Reliability Engineer (SRE) (Hybrid)","company":{"id":90,"name":"Cisco Systems","domain":"cisco.com","url":"https://alion.io/company/cisco-systems","size_band":"5000+","is_staffing_agency":false,"is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":{"grade":"A","score":93,"open_postings":175,"ghost_share":0,"stale_share":0.474,"repost_share":0.04,"time_to_fill_p50_days":14,"computed_at":"2026-09-24T05:45:00Z"}},"role":"DevOps","role_family":"DevOps","seniority":"staff","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["San Francisco, United States","San Jose, United States","New York, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":{"min":186900,"max":307800,"currency":"USD","period":"year","gross":null,"usd_annual":307800},"salary_estimate":null,"experience_years_min":8,"visa_sponsorship":false,"relocation_package":false,"has_equity":true,"technologies":[{"name":"AI Agents","optional":false},{"name":"AWS","optional":false},{"name":"CI/CD","optional":false},{"name":"GCP","optional":false},{"name":"Kubernetes","optional":false},{"name":"Splunk","optional":false},{"name":"Incident Management","optional":true},{"name":"Python","optional":true},{"name":"Terraform","optional":true}],"status":"live","first_seen_at":"2026-08-07T00:00:00Z","employer_posted_date":"2026-09-24","last_verified_at":"2026-09-25T03:01:22Z","board_verified":true,"closed_at":null,"days_open":49,"trust":{"level":"ok","repost_count":1,"flags":[],"days_open":49},"description":"The application window is expected to close on: 09/29/2026This is a Hybrid position requiring approximately 2 days per week on-site at Cisco offices in either San Francisco, San Jose, or New York City.\nThe Splunk Agent Resilience team is defining the future of AI resilience. Together our team provides scalable, cost-effective evaluation and guardrails that ensure AI agents behave as intended, improving reliability and reducing risks. This unified approach empowers our customers to expertly deploy and manage AI-powered applications with enhanced observability and control.\nAs a Staff Site Reliability Engineer (SRE), you will provide technical leadership for the reliability, scalability, and operational architecture of Splunk Agent Observability's platform. You will define the long-term reliability strategy, lead major infrastructure initiatives, and drive engineering excellence across deployment automation, production operations, and platform resiliency. In addition to owning complex production systems, you will influence engineering direction across teams and establish best practices for operating large-scale cloud and on-prem deployments.\nYour Impact\nDefine and drive the technical roadmap for platform reliability, scalability, and operational excellence.\n\nLead the architecture and evolution of deployment platforms supporting cloud and air-gapped customer environments.\n\nEstablish reliability engineering standards, including service level objectives (SLOs), operational readiness, capacity planning, and resiliency reviews.\n\nLead major reliability and scalability initiatives across Kubernetes, deployment infrastructure, databases, and networking.\n\nDrive automation to eliminate operational toil and improve engineering productivity.\n\nDesign and build internal platforms, frameworks, and tooling that enable reliable operations at scale.\n\nLead complex production incident response and drive systemic improvements through root cause analysis and long-term remediation.\n\nPartner with engineering leadership to influence platform architecture, deployment strategy, and production readiness.\n\nMentor engineers and raise the engineering bar through technical leadership, design reviews, and operational guidelines.\n\nCollaborate with customers and internal teams to build secure, scalable, and highly reliable deployment architectures for both cloud and on-prem environments.\n\nMinimum Qualifications\n8+ years’ experience with a Bachelor's degree or 6+ yrs with Masters or 3+ years with a PhD, or equivalent related experience; Experience to include at least 6 years in Site Reliability Engineering, Platform / Cloud / Infrastructure Engineering, or related fields.\n\n5+ years’ operating large-scale Kubernetes platforms in production.\n\nExperience designing highly available, scalable, and resilient distributed systems.\n\nExperience with AWS, GCP, or other public cloud platforms.\n\nStrong experience designing CI/CD platforms and deployment automation at scale.\n\nPreferred Qualifications\nExpertise in observability, monitoring, alerting, capacity planning, and performance engineering.\n\nStrong programming skills in Python and/or Go\n\nDeep experience with Infrastructure as Code (Terraform or similar)\n\nExperience with MLOps preferred\n\nStrong understanding of networking, distributed systems, storage, databases, and cloud architecture\n\nExperience operating both SaaS and enterprise/on-prem deployments\n\nDemonstrated technical leadership across multiple engineering teams\n\nExperience leading incident management, postmortems, and long-term reliability initiatives\n\nWhy Cisco?\nAt Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era - and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint.\nFueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere.\nWe are Cisco, and our power starts with you.\nMessage to applicants applying to work in the U.S. and/or Canada:\nThe starting salary range posted for this position is $186,900.00 to $267,700.00 and reflects the projected salary range for new hires in this position in U.S. and/or Canada locations, not including incentive compensation*, equity, or benefits.Individual pay is determined by the candidate's hiring location, market conditions, job-related skillset, experience, qualifications, education, certifications, and/or training. The full salary range for certain locations is listed below. For locations not listed below, the recruiter can share more details about compensation for the role in your location during the hiring process.\nU.S. employees are offered benefits, subject to Cisco’s plan eligibility rules, which include medical, dental and vision insurance, a 401(k) plan with a Cisco matching contribution, paid parental leave, short and long-term disability coverage, and basic life insurance. Please see the Cisco careers site to discover more benefits and perks. Employees may be eligible to receive grants of Cisco restricted stock units, which vest following continued employment with Cisco for defined periods of time.\nU.S. employees are eligible for paid time away as described below, subject to Cisco’s policies:\n10 paid holidays per full calendar year, plus 1 floating holiday for non-exempt employees\n\n1 paid day off for employee’s birthday, paid year-end holiday shutdown, and 4 paid days off for personal wellness determined by Cisco\n\nNon-exempt employees** receive 16 days of paid vacation time per full calendar year, accrued at rate of 4.92 hours per pay period for full-time employees\n\nExempt employees participate in Cisco’s flexible vacation time off program, which has no defined limit on how much vacation time eligible employees may use (subject to availability and some business limitations)\n\n80 hours of sick time off provided on hire date and each January 1st thereafter, and up to 80 hours of unused sick time carried forward from one calendar year to the next\n\nAdditional paid time away may be requested to deal with critical or emergency issues for family members\n\nOptional 10 paid days per full calendar year to volunteer\n\nFor non-sales roles, employees are also eligible to earn annual bonuses subject to Cisco’s policies.\nEmployees on sales plans earn performance-based incentive pay on top of their base salary, which is split between quota and non-quota components, subject to the applicable Cisco plan. For quota-based incentive pay, Cisco typically pays as follows:\n.75% of incentive target for each 1% of revenue attainment up to 50% of quota;\n\n1.5% of incentive target for each 1% of attainment between 50% and 75%;\n\n1% of incentive target for each 1% of attainment between 75% and 100%; and\n\nOnce performance exceeds 100% attainment, incentive rates are at or above 1% for each 1% of attainment with no cap on incentive compensation.\n\nFor non-quota-based sales performance elements such as strategic sales objectives, Cisco may pay 0% up to 125% of target. Cisco sales plans do not have a minimum threshold of performance for sales incentive compensation to be paid.\nThe applicable full salary ranges for this position, by specific state, are listed below:\nNew York City Metro Area:\n$186,900.00 - $307,800.00Non-Metro New York state & Washington state:\n$166,300.00 - $274,100.00* For quota-based sales roles on Cisco’s sales plan, the ranges provided in this posting include base pay and sales target incentive compensation combined.\n** Employees in Illinois, whether exempt or non-exempt, will participate in a unique time off program to meet local requirements.","description_format":"text","description_chars":8094,"description_truncated":false,"requirements":{"experience_years_min":8,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"bachelor","optional":false},"security_clearance":false,"languages":[]},"benefits":["Annual bonuses","Equity","Life insurance","Parental leave","Restricted stock units","Vision insurance"],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Government","AI Infrastructure","Log Analytics"],"lifecycle":[{"event":"open","at":"2026-08-07T00:00:00Z"},{"event":"close","at":"2026-08-10T13:39:49Z"},{"event":"reopen","at":"2026-09-24T15:18:47Z"}],"liveness":{"score":15,"band":"cold","label":"Long shot","p_open":1,"p_active":0.528,"p_room":0.28,"age_days":49,"expected_fill_days":14,"reasons":["conf:1","win:tail","crowd:brand"],"computed_at":"2026-09-25T04:25:31Z"},"pay":{"stated_usd_annual":307800,"is_top_pay":true},"html_url":"https://alion.io/job/cisco-systems-staff-site-reliability-engineer-sre-hybrid","json_url":"https://alion.io/job/cisco-systems-staff-site-reliability-engineer-sre-hybrid.json","meta":{"generated_at":"2026-09-25T04:25:31Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":4772,"day_limit":5000,"remaining_today":228,"minute_limit":60,"resets_at":"2026-09-26T00:00:00Z"}}}