{"id":1216506,"url":"https://alion.io/job/wave-observability-specialist","title":"Observability Specialist","company":{"id":65,"name":"Wave","domain":"wave.com","url":"https://alion.io/company/wave","size_band":"501-1000","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Greenhouse","truth_index":{"grade":"A","score":85,"open_postings":12,"ghost_share":0,"stale_share":0.417,"repost_share":0,"time_to_fill_p50_days":64,"computed_at":"2026-09-28T05:45:00Z"}},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":null,"work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"inferred","locations":[],"countries":[],"hiring_countries":["GB","CA","ES"],"hiring_countries_total":3,"salary":null,"salary_estimate":{"min_usd":88000,"max_usd":205000,"period":"year","method":"global_role_seniority_cell","sample_n":2013},"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"CockroachDB","optional":false},{"name":"Datadog","optional":false},{"name":"Grafana","optional":false},{"name":"GraphQL","optional":false},{"name":"Honeycomb","optional":false},{"name":"Jaeger","optional":false},{"name":"Kubernetes","optional":false},{"name":"Loki","optional":false},{"name":"OpenTelemetry","optional":false},{"name":"Platform Engineering","optional":false},{"name":"PostgreSQL","optional":false},{"name":"Prometheus","optional":false},{"name":"Pyroscope","optional":false},{"name":"Python","optional":false},{"name":"Sentry","optional":false},{"name":"SLI/SLO/SLA","optional":false},{"name":"Stripe","optional":false},{"name":"Redis","optional":true}],"status":"live","first_seen_at":"2026-08-18T12:58:35Z","employer_posted_date":"2026-09-25","last_verified_at":"2026-09-28T22:36:46Z","board_verified":true,"closed_at":null,"days_open":41,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":41},"description":"Our mission\nWe're making Africa the first cashless continent.\nIn 2017, over half the population in Sub-Saharan Africa had no bank account. That's for good reason-the fees are too high, the closest branch can be miles away, and nobody takes cards. Without access to financial institutions, people are forced to keep their savings under the mattress. Small business owners rely on lenders who charge extortionate rates. Parents spend hours waiting in line to pay school fees in cash.\nWe're solving this by building financial services that just work: no account fees, instantly available, and accepted everywhere. In places where electricity, water and roads don't always work, you can still send money with Wave. In 2017, we launched a mobile app in Senegal for cash deposit, withdrawal, and peer-to-peer and business payments. Now, we have millions of users across 9 countries and are growing fast.\nOur goal is to make Africa the first cashless continent and that's where you come in...\nHow you'll help us achieve it\nWave is now the largest financial institution in Senegal and Côte d'Ivoire, with millions of users, growing rapidly year-on-year. And, we’re still in the early days of our product roadmap and potential impact on people’s everyday lives.\nAs an Observability Engineer at Wave, you will help engineers understand, operate, debug, and improve the systems that power payments for millions of users across Africa.\nYou will own and evolve Wave’s observability platform across our Python backend, GraphQL API, Postgres and CockroachDB databases, Kubernetes workloads, cloud infrastructure, and on-premises environments. Your work will make it easier for product, database, infrastructure, and security teams to detect problems and regressions early, understand system behaviour, resolve incidents quickly, and make well-informed reliability and performance decisions.\nYou'll work in the newly formed Performance & Observability team and report to the Director of Platform. You'll partner closely with infrastructure, database, and product engineering teams to improve how we instrument our services.\nIn this role you'll:\nImprove our understanding of production behaviour across application code, GraphQL APIs, databases, caches, Kubernetes workloads, cloud infrastructure, async jobs, and on-premises environments.\nPartner with product and platform teams to define meaningful service-level indicators, reduce alert fatigue, and ensure alerts are actionable, reliable, and tied to user impact.\nBuild internal tooling and self-service workflows that help engineers instrument services, investigate incidents, analyse performance, and understand dependencies.\nHelp teams identify reliability, latency, capacity, and cost issues before they become user-facing incidents.\nEstablish and implement observability standards, documentation, and training materials that scale across Wave’s engineering organisation.\nOperate and improve our observability platform (e.g., Datadog, Honeycomb, Prometheus, Grafana, OpenTelemetry) while controlling costs as data volume grows.\nYou might be a good fit if you...\nCare a lot about working on software whose mission you can believe in.\nHave a bias for action. You see a problem, you fix a problem. You get buy-in for your solutions and keep work moving.\nApproach your work with a growth mindset and use your skills and experience to mentor less-experienced engineers.\nAre excited to build world-class infrastructure that powers economic opportunity for an entire continent.\nRequirements\n5+ years of experience in observability, SRE, platform engineering, infrastructure engineering, backend engineering, or production systems engineering.\nDeep understanding of metrics, logging, tracing, profiling, alerting, dashboards, service-level indicators, and incident response workflows at scale.\nExperience building internal tools, libraries, automation, or platforms used by other engineers.\nExcellent communication and collaboration skills. This role succeeds by helping other engineers build, operate, and debug better systems.\nPragmatic judgment about when to improve tooling, when to simplify, and when to avoid unnecessary complexity.\nExperience learning, analyzing, and working on other people’s code.\nTechnical Skills\nProficiency in at least one backend language, preferably Python.\nExperience with observability tools such as Prometheus, Grafana, Datadog, OpenTelemetry, Jaeger, Tempo, Loki, Honeycomb, Sentry, or similar systems.\nExperience with at least some our stack, Postgres, CockroachDB, Redis, GraphQL, or Kubernetes.\nExperience with OpenTelemetry instrumentation and collector configuration at scale\nAbout the Performance & Observability Team\nThe Performance and Observability team was recently formed with the hiring of our first performance engineer, and is expanding with this Observability position. The team is and will continue to define its scope. The observability scope within the team currently is:\nOwnership, management and evolution of observability tools: Datadog, Honeycomb, Sentry, and Pyroscope.\nDesign of SLOs and support product teams in implementing them.\nSupport all teams in improving their alert quality.\nOwnership of observability modules and code, ensuring consistent naming conventions across all our systems.\nSome recent and potential projects, as examples of specific things the team has worked on:\nEarly detection of regressions in our code. This could be a performance, reliability, or other regression that impacts our users.\nCPU profiling in production.\nSLO support for product teams.\nScaling our Observability platform to support 4x our current volume while keeping costs in check.\nAbout engineering at Wave\nWe care about the big picture. We don’t hire engineers to just ship tickets. We hire them to solve problems. That means caring deeply about outcomes, understanding context, and jumping in wherever something’s broken, even if it’s technically “not your area.” When we see problems, inefficiencies, or opportunities to make something better, we act. We dig into operational issues, clarify fuzzy product specs, or step into unfamiliar code to help unblock teammates.\nWe move as fast as possible. Speed matters. It lets us try things quickly, get feedback early, and course-correct while it’s cheap. So we write small PRs. We aim for MVPs. We leave TODOs and file follow-ups. We don’t over-perfect v1. That said, we’re building a financial product. Some things-like money movement, correctness, or security-deserve more caution.\nWe like boring technology. We favor tools that are reliable, well-understood, and easy to debug. This keeps us focused on solving meaningful problems instead of wrestling with unpredictable infrastructure. If a new technology helps us move faster, build safer, or solve a real need, we’ll consider it. But we don’t adopt tools just because they’re new-we adopt them because they’re right.\nSimplicity is a strategy. It lets us focus our energy where it matters most: serving our users.\n#LI-DH1 #LI-REMOTE\nOur team\nWe have a rapidly growing in-country team in Senegal, Côte d'Ivoire, Mali, Burkina Faso, The Gambia, Uganda, Niger, Sierra Leone, and Cameroon plus remote team members spread across the world.\nWe're deeply passionate about our mission of bringing radically affordable financial services to the people who need them most.\nWe foster autonomy for our employees. You'll own your projects at every stage, from understanding the problem to monitoring your solution in production.\nWe raised the largest Series A in Africa in 2021. Our world-class investors, include Founders Fund, Sequoia Heritage, Stripe, Ribbit Capital, Y Combinator, and Partech Africa.\nWe are on Y Combinator's top companies by revenue.\nHow to apply\nFill out the form below, and upload a resume in English and a cover letter describing your interest in Wave and the role.\nWe review applications frequently and recommend that you apply to the role that most closely aligns with your skills, experience and career goals.\nWave is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees.","description_format":"text","description_chars":8146,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"United Kingdom","iso":"GB","kind":"country"},{"name":"Canada","iso":"CA","kind":"country"},{"name":"Spain","iso":"ES","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Payment Processing & Gateways","Money Transfer & Remittances"],"lifecycle":[{"event":"open","at":"2026-09-25T09:36:08Z"}],"liveness":{"score":77,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.86,"p_room":0.9,"age_days":40,"expected_fill_days":64,"reasons":["conf:0","win:mid"],"computed_at":"2026-09-28T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/wave-observability-specialist","json_url":"https://alion.io/job/wave-observability-specialist.json","meta":{"generated_at":"2026-09-29T00:48:19Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":735,"day_limit":5000,"remaining_today":4265,"minute_limit":60,"resets_at":"2026-09-30T00:00:00Z"}}}