{"id":2079770,"url":"https://alion.io/job/bp-senior-site-reliability-engineer","title":"Senior Site Reliability Engineer","company":{"id":2557,"name":"BP","domain":"bp.com","url":"https://alion.io/company/bp-com","size_band":"1001-5000","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workday","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Kuala Lumpur, Malaysia"],"countries":["MY"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":18500,"max_usd":50000,"period":"year","method":"global_role_cell_scaled_by_country","sample_n":2856},"experience_years_min":7,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"CI/CD","optional":false},{"name":"CloudFormation","optional":false},{"name":"Configuration Management","optional":false},{"name":"Grafana","optional":false},{"name":"Least Privilege","optional":false},{"name":"Linux","optional":false},{"name":"OpenTelemetry","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Prometheus","optional":false},{"name":"Python","optional":false},{"name":"Ruby","optional":false},{"name":"Terraform","optional":false},{"name":"Unix","optional":false}],"status":"live","first_seen_at":"2026-10-08T09:46:10Z","employer_posted_date":"2026-10-08","last_verified_at":"2026-10-11T15:53:49Z","board_verified":true,"closed_at":null,"days_open":3,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":3},"description":"Entity:\nTechnologyJob Family Group:\nIT&S GroupJob Description:\nRole Summary\nAs a Site Reliability Engineer, you will be responsible for improving the reliability, resilience and operational effectiveness of our technology platforms and services.\nYou will work closely with engineering and product teams to ensure systems are highly available, scalable, secure and supportable in production. You will use software engineering, automation and modern cloud practices to reduce manual effort, improve performance and strengthen production reliability.\nKey Responsibilities\nImprove the reliability, availability, performance and scalability of cloud-based applications and services.\nDesign and implement automation to reduce manual operational activities and improve engineering efficiency.\nBuild and improve monitoring, logging, alerting and observability across production systems.\nInvestigate complex production issues and drive improvements to prevent recurring failures.\nImprove system resilience, recovery and operational readiness.\nBuild and improve CI/CD pipelines to enable reliable and repeatable software delivery.\nDevelop and maintain infrastructure using Infrastructure as Code and automation.\nIdentify reliability risks, operational gaps and technical debt and drive appropriate improvements.\nImprove cloud infrastructure security and operational practices.\nDevelop reusable engineering patterns and mentor engineers across teams.\nRequired Experience and Qualifications\n7+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Engineering, Software Engineering or related technical disciplines, with strong experience operating production systems.\nStrong understanding of cloud infrastructure security, including identity and access management, least privilege, network security, secrets management and secure configuration.\nUnderstanding of security practices within CI/CD pipelines and Infrastructure as Code.\nAble to independently investigate and resolve complex technical and production problems.\nStrong communication and collaboration skills across engineering, product, security and operational teams.\nAble to influence engineering practices, drive technical improvements and mentor other engineers.\nDegree in Computer Science, Engineering or a related discipline, or equivalent professional experience.\nRelevant cloud or engineering certifications are beneficial but not essential.\nTechnical Skills\nStrong experience operating and improving production systems in cloud-based environments.\nStrong troubleshooting skills across applications, infrastructure, networking and cloud services.\nExperience managing system reliability, scalability, availability and performance.\nStrong knowledge of monitoring, logging, alerting and production diagnostics.\nExperience with incident investigation, root cause analysis and operational improvement.\nGood understanding of distributed systems, resilience and recovery practices.\nSoftware Engineering\nStrong programming and scripting skills using Python, Ruby, Go or equivalent technologies.\nStrong understanding of software engineering practices including source control, code review, automated testing and software delivery.\nStrong experience designing, building and maintaining CI/CD pipelines.\nStrong experience with deployment automation, release management and rollback or recovery practices.\nExperience building automation, tooling and reusable engineering solutions.\nCloud Infrastructure\nStrong hands-on experience with AWS, Microsoft Azure or equivalent cloud platforms.\nStrong experience with Infrastructure as Code, using technologies such as Terraform, CloudFormation or equivalent.\nStrong knowledge of Linux/Unix systems, networking and infrastructure troubleshooting.\nExperience with containers and modern cloud application infrastructure.\nExperience with observability technologies such as Prometheus, Grafana, OpenTelemetry or cloud-native equivalents.\nGood understanding of cloud services including compute, networking, storage, databases, identity and messaging.\nSkills That Set You Apart\nExperience improving reliability and operational practices across multiple services or engineering teams.\nExperience with automated recovery, resilience engineering or self-service platform capabilities.\nExperience operating large-scale or highly available distributed systems.\nStrong understanding of cloud-native engineering and modern operational practices.\nAbout bp\nAt bp, we provide the following environment and benefits to you:\nA company culture where we respect our diverse and unified teams, where we are proud of our achievements and where fun and the attitude of giving back to our environment are highly valued.\nPossibility to join our social communities and networks\nLearning opportunities and other development opportunities to craft your career path\nLife and health insurance, medical care package\nAnd many other benefits. We are an equal opportunity employer and value diversity at our company. We do not discriminate based on race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.\nWe will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, perform crucial job functions, and receive other benefits and privileges of employment.\nTravel Requirement\nNo travel is expected with this roleRelocation Assistance:\nThis role is not eligible for relocationRemote Type:\nThis position is a hybrid of office/remote working Skills:\nAgility core practices, Agility core practices, Analytics, API and platform design, Business Analysis, Cloud Platforms, Coaching, Communication, Configuration management and release, Continuous deployment and release, Data Structures and Algorithms (Inactive), Digital Project Management, Documentation and knowledge sharing, Facilitation, Information Security, iOS and Android development, Mentoring, Metrics definition and instrumentation, NoSql data modelling, Relational Data Modeling, Risk Management, Scripting, Service operations and resiliency, Software Design and Development, Source control and code management {+ 4 more}.\n\n Legal Disclaimer:\nWe are an equal opportunity employer. We do not discriminate on the basis of protected characteristics like race, religion, color, sex, national origin, sexual orientation, veteran status or disability status. Individualswith an accessibility need may request an adjustment/accommodationrelated to bp’s recruiting process (e.g., accessing the job application, completing required assessments, participating in telephone screenings or interviews, etc.). If you would like to request an adjustment/accommodationrelated to the recruitment process, please contact us.\nIf you are selected for a position and depending upon your role, your employment may be contingent upon adherence to local policy. This may include pre-placement drug screening, medical review of physical fitness for the role, and background checks.","description_format":"text","description_chars":7056,"description_truncated":false,"requirements":{"experience_years_min":7,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":{"level":"bachelor","optional":false},"security_clearance":false,"languages":[]},"benefits":["Health insurance","Relocation assistance"],"hiring_locations":[{"name":"Malaysia","iso":"MY","kind":"country"}],"hiring_excludes":[],"relocation_offered":true,"industries":["Energy & Utilities","Fossil Fuels","Oil & Gas"],"lifecycle":[{"event":"open","at":"2026-10-08T09:46:10Z"}],"visa":[],"liveness":{"score":90,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.903,"p_room":1,"age_days":1,"expected_fill_days":29,"reasons":["conf:5","velocity","win:early","comp:brand"],"computed_at":"2026-10-10T05:45:15Z"},"pay":null,"html_url":"https://alion.io/job/bp-senior-site-reliability-engineer","json_url":"https://alion.io/job/bp-senior-site-reliability-engineer.json","meta":{"generated_at":"2026-10-11T19:15:19Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler_verified","counted_by":"address","units_charged":1,"used_today":5946,"day_limit":null,"remaining_today":null,"minute_limit":300,"resets_at":"2026-10-12T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":2557},"rest":"https://alion.io/mcp/rest/get_company?id=2557"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fbp-senior-site-reliability-engineer"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fbp-senior-site-reliability-engineer"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fbp-senior-site-reliability-engineer"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/bp-senior-site-reliability-engineer\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fbp-senior-site-reliability-engineer"}]}