368,530open jobs
9,432companies
50,439added this week
Browse all
Location
In office
Seniority
Senior · 4+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
Toters is a Beirut-founded on-demand e-commerce and delivery platform operating across Lebanon and Iraq. Launched in 2017, the mobile application connects customers with local restaurants, supermarkets, and retail stores for fast last-mile delivery. Beyond food and grocery logistics, the platform offers a personalized concierge service that allows users to pick up or send items locally on demand.

The Company

Toters is an on-demand e-commerce and delivery platform and operates a service that enables customers to get anything in their city at the highest level of convenience.

At Toters, technology is at the heart of everything we do. We have product teams that are working hard every day to create products that make our customers' lives easier. Our engineers are also continuously creating solutions to make our processes more efficient, all in an effort to get to our customers fast and at the best cost. If you are interested in working in a high growth startup environment, and look to be part of a team that will potentially change the way customers shop in the Middle East, apply now.

About the Role

We are looking for a Senior Site Reliability Engineerwho will play a critical role in ensuring high availability, performance, and resilience across our production systems. You will be at the heart of operational excellence, leading high-impact incident responses, building proactive monitoring systems, and engineering automation that prevents outages before they happen. If you love solving complex distributed system challenges and thrive in high-pressure environments, this role is for you.

Key Responsibilities

Incident Management & Reliability

  • Act as Incident Commanderduring major outages, leading real-time diagnosis, communication, and recovery.
  • Own and improve the end-to-end incident management lifecycle, including post-incident reviews and action plans.
  • Drive root cause analysisand proactive reliability improvements to prevent recurrence.

Monitoring & Observability

  • Design and maintain metrics, alerts, and dashboardsusing Prometheus, Grafana, and New Relic.
  • Implement SLIs/SLOsto monitor service health and drive availability targets (99.99%+ uptime).
  • Integrate log managementand distributed tracingwith tools like ELK Stackand AWS X-Ray.

Automation & Tooling

  • Develop automation scriptsand internal tooling in Python or Node.jsto reduce manual ops and accelerate recovery (MTTR improvement).
  • Build self-healing infrastructure using IaCand automation pipelines.
  • Optimize on-call workflows, escalation policies, and runbooks using PagerDuty.

Cloud Infrastructure

  • Operate and improve infrastructure hosted on AWS, ensuring reliability, cost efficiency, and scalability.
  • Collaborate with backend and platform teams to embed SRE best practicesacross engineering.

Key Qualifications

  • 4+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering.
  • Proven success managing production incidentsand participating in on-call rotations.
  • Strong hands-on experience with Prometheus, Grafana, and PagerDuty.
  • Proficient in Python or Node.jsfor automation and tooling.
  • Experience with AWS services(EC2, CloudWatch, ECS/Lambda, IAM, etc.).
  • Solid understanding of Linux systems, networking, and CI/CD pipelines.

Nice to Have

  • Experience as Incident Commanderin mission-critical environments.
  • Knowledge of New Relic, Sentry, ELK Stack, or Datadog.
  • Background implementing SLIs/SLOs/Error Budgets(Google SRE model).
  • Familiarity with Docker, Kubernetes, Terraform, or Ansible.
  • Certifications such as:
    • AWS Solutions Architect Associate/DevOps Engineer
    • ITIL Foundationor relevant reliability certifications.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,530 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$20k – $46k per year (Estimated) • Remote/Hybrid • Full-Time • 2+ years exp • Bachelor's Degree • Bengaluru
Python
SQL
AI/ML
AI Agents
Amazon SageMaker
Anomaly Detection
AWS Bedrock
Computer Vision
Embeddings
LLM
Multimodal AI
Prompt Engineering
RAG
Time Series Forecasting
DevOps
Amazon S3
AWS
AWS Lambda
CI/CD
Git
Vector
Analytics
Power BI
Tableau
Apply
DevOps Engineer 2 hours ago
$49k – $88k per year (Estimated) • Remote/Hybrid • Full-Time • Bachelor's Degree • Warsaw
Bash
Python
DevOps
AWS
Azure
Azure AKS
CI/CD
Docker
GCP
Git
GitHub
GitHub Actions
GitOps
Grafana
Kubernetes
Loki
Prometheus
Terraform
Thanos
Apply
$173k – $260k per year • In office • Full-Time • PhD • San Francisco
JavaScript
Node JS
Python
Python
Celery
Django
Flask
Databases
RabbitMQ
Redis
AI/ML
Agentforce
AI Agents
DevOps
Akamai
AWS
CI/CD
Cloudflare
CloudFormation
Helm
Jenkins
Kubernetes
Spinnaker
Terraform
Marketing
Salesforce
Apply
$91k – $182k per year (Estimated) • In office • 3+ years exp • Redwood City
Python
DevOps
AWS
CI/CD
GCP
Terraform
Apply
IT Administrator 2 hours ago
$83k – $113k per year • In office • Full-Time
Node JS
Python
JavaScript
DevOps
AWS
GCP
GitHub
Cybersecurity
Okta
SOC 2
Management
Confluence
Jira
Slack
Apply
In office • Full-Time • 5+ years exp
AI/ML
AI Agents
Apply
In office • Full-Time • 8+ years exp • Bachelor's Degree
Apply
Remote • Full-Time • 4+ years exp • United States
Python
AI/ML
LightGBM
NumPy
PyTorch
Scikit-learn
TensorFlow
XGBoost
AI Agents
Model Context Protocol
Apply
iOS Engineer II 1 month ago
In office • Full-Time • 2+ years exp • Bachelor's Degree
Swift
Swift
RxSwift
Swift Concurrency
Mobile
Bitrise
Combine
Crashlytics
Fastlane
Firebase
Reactive Programming
SwiftUI
UIKit
DevOps
CI/CD
CircleCI
Git
Apply
Android Engineer II 1 month ago
In office • Full-Time • 3+ years exp • Bachelor's Degree
Kotlin
Kotlin
RxKotlin
Mobile
Crashlytics
Dependency Injection
Firebase
Jetpack Compose
Reactive Programming
DevOps
Git
Apply
See all jobs
This is one of many
368,530 more open roles from verified company boards, updated every day.