{"id":1523832,"url":"https://alion.io/job/diabolocom-stateful-services-engineer","title":"Stateful Services Engineer","company":{"id":53115,"name":"Diabolocom","domain":"diabolocom.com","url":"https://alion.io/company/diabolocom","size_band":"51-200","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Lever","truth_index":{"grade":"B","score":70,"open_postings":10,"ghost_share":0.5,"stale_share":0,"repost_share":0,"time_to_fill_p50_days":null,"computed_at":"2026-10-01T05:45:00Z"}},"role":"Industrial Engineering","role_family":"Industrial Engineering","seniority":null,"employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Paris, France"],"countries":["FR"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":44000,"max_usd":78000,"period":"year","method":"role_country_seniority_unknown","sample_n":83},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Apache Kafka","optional":false},{"name":"ClickHouse","optional":false},{"name":"Incident Management","optional":false},{"name":"PostgreSQL","optional":false},{"name":"RabbitMQ","optional":false},{"name":"Redis","optional":false},{"name":"Redpanda","optional":false},{"name":"Cassandra","optional":true},{"name":"ElasticSearch","optional":true},{"name":"Linux","optional":true},{"name":"MySQL","optional":true},{"name":"OpenSearch","optional":true}],"status":"live","first_seen_at":"2026-09-30T07:59:35Z","employer_posted_date":"2026-09-30","last_verified_at":"2026-09-30T23:34:23Z","board_verified":true,"closed_at":null,"days_open":1,"trust":{"level":"ok","repost_count":0,"flags":["company_stale"],"days_open":1},"description":"Build the infrastructure behind the next generation of AI-powered customer communications\nDiabolocom is building an AI-first communication platform where AI and human agents, business workflows and communication channels work together seamlessly.\nAs we enter our next phase of growth, we are looking for a Stateful Services Engineer to help build and operate the infrastructure foundation behind the platform.\nThis is a unique opportunity to work on an infrastructure environment combining carrier-grade voice systems, on-premise data centers, distributed systems and AI workloads, while helping us build the reliability and operational maturity required to support 3x growth.\nYou will report to the newly forming Stateful Services Lead and work closely with Engineering, SRE and Infrastructure teams to improve the reliability, scalability and evolution of our stateful systems.\nYour mission\nYour goal is to build, operate and improve Diabolocom’s stateful services, with a strong focus on databases and data stores.\nYou will help ensure that our databases and stateful systems remain highly available, reliable and performant, while evolving our infrastructure so that maintenance, upgrades and scaling can be performed safely and with minimal to no downtime.\nWorking closely with the Stateful Services Lead, you will take ownership of specific technical projects and contribute to the evolution of our stateful infrastructure.\nYou will:\nOperate and maintain our stateful services, including PostgreSQL, ClickHouse, Redis, RabbitMQ and Kafka/Redpanda.\nContribute to the design, evolution and scaling of our database infrastructure, with 50+ PostgreSQL clusters and increasingly complex data workloads.\nWork on projects aimed at making database maintenance, upgrades and migrations safer and less disruptive to our services.\nImprove high availability, replication, failover, backup/restore and disaster recovery across our stateful systems.\nContribute to rebuilding our PostgreSQL clusters to support automatic failover between data centers and reduce the impact of maintenance operations.\nHelp improving our queueing infrastructure, including projects such as quorum-based RabbitMQ.\nHelp building and maintaining stable Redis infrastructure and improve the reliability of our stateful services.\nInvestigate and solve complex performance, scalability and reliability issues across databases and distributed systems.\nWork closely with development teams to understand their needs, support their use of databases and help them build reliable solutions.\nCollaborate with SRE, networking and infrastructure teams to ensure that our stateful services operate reliably within our broader infrastructure.\nContribute to monitoring, automation and operational practices that make our infrastructure easier and safer to operate.\nParticipate in our on-call and incident management rotation, helping diagnose and resolve production issues when needed.\nStay hands-on and close to the technology: you will be expected to understand how our systems work, how they fail and how to improve them.\nWho we are looking for\nWe’re looking for a technically strong and hands-on Stateful Services Engineer who enjoys solving complex infrastructure problems and taking ownership of production systems where reliability really matters.\nYou have experience operating databases or stateful services in production and understand that running these systems is about much more than simply maintaining them. You know how to troubleshoot them, how they behave under load, how they fail, and how to make them more reliable as they scale.\nYou’ll work on specific technical projects defined with the Stateful Services Lead, while having the autonomy to investigate problems, propose solutions and contribute to technical decisions.\nWe care more about your technical depth, curiosity and ability to take ownership than the exact number of years of experience you have.\nMust Have\nStrong hands-on experience as a DBA, Database Engineer, SRE or similar, with solid experience in PostgreSQL.\nExperience operating and maintaining production databases, including upgrades, migrations, replication, backups, recovery, monitoring and performance troubleshooting.\nGood understanding of high availability and reliability for stateful systems: failover, redundancy, replication, disaster recovery and minimising downtime during maintenance.\nGood understanding of distributed systems, including replication, consistency, fault tolerance and the challenges of operating stateful services at scale.\nExperience working with on-premise infrastructure or a willingness and ability to quickly adapt to an on-premise/bare-metal environment.\nGood system design skills and the ability to reason about scalability, reliability and performance.\nExperience collaborating closely with software development teams, understanding their requirements and helping them build reliable solutions around databases and stateful services.\nStrong troubleshooting and problem-solving skills, with the ability to investigate complex production issues.\nStrong hands-on mindset: you enjoy getting into the technical details and solving problems yourself.\nComfortable participating in on-call and incident management.\nFluent English, both written and spoken. French or other languages are a plus.\nNice to Have\nExperience with some of the following:\nClickHouse or another complex database/data platform.\nKafka or Redpanda and event-streaming infrastructure.\nRabbitMQ, particularly quorum-based configurations.\nRedis and distributed caching.\nOther database technologies such as Elasticsearch/OpenSearch, MySQL, Cassandra or MongoDB.\nAdvanced Linux systems administration, including processes, storage, networking and troubleshooting.\nBare-metal infrastructure and data-center environments.\nObservability and monitoring of stateful systems.\nCapacity planning, performance analysis and load testing.\nInfrastructure automation and configuration management.\nIncident management, root-cause analysis and post-mortem practices.\nAbout You\nYou take ownership. When you see a reliability or scalability problem, you investigate it and work towards a solution.\nYou are genuinely technical. You like understanding how things work under the hood and are comfortable diving deep when something breaks.\nYou think in systems. You understand that databases don't operate in isolation and enjoy working across infrastructure, networking, SRE and development teams.\nYou care about reliability. You want systems that can be maintained, upgraded and scaled without putting the business at risk.\nYou are pragmatic. You can make sensible technical trade-offs in complex environments.\nYou are curious and self-driven. You learn by doing, investigate unfamiliar technologies and are comfortable figuring things out independently.\nYou enjoy solving complex problems. You like working on systems where there isn't always an obvious answer.\nYou communicate clearly. You can explain technical topics and work effectively with the teams relying on your services.\nYou are comfortable taking responsibility. You don't need to be a manager to have an impact or take the lead on a technical problem.\nWhy join us\nOwn critical infrastructure: your work will directly impact the reliability and scalability of Diabolocom’s platform.\nSolve complex technical challenges: work with large-scale databases and stateful services, including PostgreSQL, ClickHouse, Kafka, Redis and RabbitMQ.\nWork on real infrastructure challenges: help improve PostgreSQL clusters, cross-data-center failover, RabbitMQ and Redis reliability.\nStay hands-on: spend your time solving challenging technical problems rather than managing people.\nLearn and grow: work across databases, distributed systems, messaging and infrastructure in a diverse technical environment.\nWork across teams: collaborate closely with Engineering, SRE and Infrastructure to build reliable systems at scale.\nRecruitment Process\nIntroductory call with Jade - 30 minGet to know each other and discuss your background.\n\nInterview with Sergey, CTO - 1 hourDiscuss your technical background, skills and knowledge.\n\nOn-site System Design interview with Sergey, CTO - 1 hourWork through a technical case study and explore your system design approach.\n\nTeam & cultural fit - 1 hourMeet members of the team, onsite or via video.\n\nLogistics\nParis, rue de la Paix, 75002\nHybrid model: 3 days onsite / 2 days remote\nEquipment: choose your preferred setup (Mac/Linux + monitors)\nLunch vouchers\nNavigo: 50% of the monthly pass reimbursed\nOne More Thing\nWe are not necessarily looking for someone who has already worked in this exact role or with every technology in our stack.\nThe strongest candidates are often engineers who have built and operated complex infrastructure, taken ownership of production systems and learned new technologies along the way.\nIf you enjoy solving difficult technical problems, want to work close to the infrastructure and are excited by the challenge of making stateful systems more reliable and scalable - we would love to talk.","description_format":"text","description_chars":9078,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[{"language":"French","level":"All levels","optional":false},{"language":"English","level":"Advanced (C1)","optional":false}]},"benefits":[],"hiring_locations":[{"name":"France","iso":"FR","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Unified Communications","Speech & Audio AI","Contact Center Software"],"lifecycle":[{"event":"open","at":"2026-09-30T13:23:29Z"}],"liveness":{"score":60,"band":"ok","label":"Likely open","p_open":1,"p_active":0.602,"p_room":1,"age_days":0,"expected_fill_days":36,"reasons":["conf:6","stale_co","win:early"],"computed_at":"2026-10-01T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/diabolocom-stateful-services-engineer","json_url":"https://alion.io/job/diabolocom-stateful-services-engineer.json","meta":{"generated_at":"2026-10-01T11:03:38Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":2427,"day_limit":5000,"remaining_today":2573,"minute_limit":60,"resets_at":"2026-10-02T00:00:00Z"}}}