{"id":1234233,"url":"https://alion.io/job/duplo-senior-site-reliability-engineer","title":"Senior Site Reliability Engineer","company":{"id":177509,"name":"Duplo","domain":"tryduplo.com","url":"https://alion.io/company/duplo","size_band":"51-200","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"BambooHR","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"on_site","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Lagos, Nigeria"],"countries":["NG"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":73000,"max_usd":188000,"period":"year","method":"global_role_cell_scaled_by_country","sample_n":2013},"experience_years_min":7,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AWS","optional":false},{"name":"CI/CD","optional":false},{"name":"Datadog","optional":false},{"name":"Docker","optional":false},{"name":"Grafana","optional":false},{"name":"Kubernetes","optional":false},{"name":"PostgreSQL","optional":false},{"name":"Prometheus","optional":false},{"name":"Redis","optional":false},{"name":"Terraform","optional":false},{"name":"TypeScript","optional":false}],"status":"live","first_seen_at":"2026-07-10T00:00:00Z","employer_posted_date":"2026-07-10","last_verified_at":"2026-09-28T22:04:07Z","board_verified":true,"closed_at":null,"days_open":80,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":80},"description":"Duplo is a Lagos-based fintech startup that enables businesses in Africa to automate their spend management, simplify cross-border payments, and control business finances all on one platform.\nWe want to make B2B payments as simple as P2P payment apps. Most business payments in Africa are made offline…yikes. We are on a mission to transform this. We are backed by top investors, including Tribe Capital, Commerce Ventures, Liquid2 Ventures, My Asia VC, Soma Capital, Y Combinator, Oui Capital, and others.\nThis is a unique opportunity. You'll have the responsibility and resources to take a significant part in the creation of a paradigm-changing product that will impact millions.\nKey Responsibilities\nLead technical design and implementation for large, high-risk, or high-complexity initiatives that improve the availability, scalability, and resilience of our backend systems\nDesign robust service boundaries, define maintainable infrastructure architecture, and guide reliability standards and best practices across the engineering team\nUnderstand edge cases in money movement and design with operational safety and compliance in mind, not just uptime\nOwn observability end-to-end metrics, logging, tracing, and alerting so issues are caught before customers feel them\nBuild and refine incident response processes, lead incident diagnosis and resolution in production, and drive thorough postmortems that prevent recurrence\nWork seamlessly with cross-functional partners including product, compliance, finance operations, support, backend engineering, and external providers, communicating clearly and adapting your approach to whoever you are working with\nCollaborate closely with fellow engineers through code and infrastructure review, pairing, and technical discussion, giving and receiving feedback in a way that raises the quality of the team's work, not just your own\nMake sound, well-reasoned decisions on architectural trade-offs, deployment strategies, and rollout safety, and know when to escalate decisions with broader business impact\nDrive capacity planning, performance tuning, and cost-efficiency across infrastructure, balancing reliability against operational spend\nDefine and track SLOs/SLIs, and use them to prioritize reliability work against feature delivery\nMentor other engineers; raise engineering and operational standards and strengthen both the systems you own and the people you work with\nRequirements:\nMinimum of 7 years experience in SRE, DevOps, or backend infrastructure roles, with deep expertise in Kubernetes, Docker, and AWS in production environments\nStrong scripting/programming ability (TypeScript and/or another language used for tooling, automation, or services) to build internal tools and automate operational work\nPractical expertise with PostgreSQL operations at scale, including replication, backup/recovery, performance tuning, and patterns that protect correctness in financial systems\nPractical expertise with Redis for caching, coordination, rate-limiting, and distributed locking, including operating it reliably in production\nStrong command of observability tooling (e.g., Prometheus, Grafana, Datadog, ELK, or equivalent) and a track record of building monitoring that catches problems early\nDemonstrated experience designing and maintaining CI/CD pipelines, infrastructure as code (e.g., Terraform), and secrets management\nDemonstrated experience in fintech or other high-stakes systems, including provider integration resilience, failover design, and compliance-sensitive infrastructure\nA track record of owning complex, high-stakes incidents from detection through resolution and postmortem, and standing behind the outcomes\nStrong communication skills, with the ability to work effectively across engineering, product, compliance, and external partners\nA demonstrated habit of catching issues before they become incidents, and of treating failure as something to learn from rather than deflect\nExperience mentoring or raising the operational standard of a team, not just delivering individually","description_format":"text","description_chars":4057,"description_truncated":false,"requirements":{"experience_years_min":7,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[],"hiring_excludes":[],"relocation_offered":false,"industries":["Payment Processing & Gateways","Procurement & Spend Management"],"lifecycle":[{"event":"open","at":"2026-09-25T15:58:06Z"}],"liveness":{"score":18,"band":"cold","label":"Long shot","p_open":1,"p_active":0.489,"p_room":0.36,"age_days":80,"expected_fill_days":42,"reasons":["conf:11","win:tail","crowd:"],"computed_at":"2026-09-28T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/duplo-senior-site-reliability-engineer","json_url":"https://alion.io/job/duplo-senior-site-reliability-engineer.json","meta":{"generated_at":"2026-09-28T22:49:16Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":465,"day_limit":5000,"remaining_today":4535,"minute_limit":60,"resets_at":"2026-09-29T00:00:00Z"}}}