489,618open jobs
16,287companies
72,295added this week
Browse all
Salary
$59k – $170k per year (Estimated)
Location
Remote (Saudi Arabia)
Overview
Company
Impact
Profile match
Turn every customer signal into resolution with AI. Lucidya unifies social listening, omnichannel engagement, and autonomous support into one CX platform.

About Lucidya

Lucidya is an AI-native platform for customer experience (CX) intelligence that manages entire customer lifecycles autonomously, from initial engagement through retention and growth.

Unlike platforms that only surface insights and leave the action to you, Lucidya closes the loop with proprietary NLU technology built in-house and trained on millions of multilingual conversations. This enables marketing, support, CX, and research teams to deliver personalized experiences that drive measurable improvements in customer satisfaction, retention, and lifetime value.

As we continue scaling globally, the reliability, performance, and resilience of our infrastructure become mission-critical to everything we do.

Why this role matters

At Lucidya, our platform processes massive volumes of real-time customer data. Any downtime, latency, or instability directly impacts our customers’ ability to make decisions and serve their own users.

This role exists to make sure that doesn’t happen.

As a Site Reliability Engineer, you’ll sit at the heart of our platform’s stability, owning the reliability of our cloud infrastructure and ensuring it scales seamlessly as we grow. You won’t just react to issues; you’ll anticipate them, design systems that prevent them, and build automation that removes them entirely.

If you enjoy solving complex infrastructure challenges, eliminating inefficiencies, and building systems that “just work” - this is where you’ll thrive.

What You’ll Do

You’ll be responsible for outcomes, not just tasks. Here’s what success looks like in this role:

You’ll make reliability the default

  • You’ll design and maintain infrastructure that is highly available, fault-tolerant, and scalable
  • You’ll proactively identify and eliminate single points of failure before they become incidents
  • You’ll ensure our production systems remain stable, even under increasing scale and load

You’ll own and optimize our cloud environments

  • You’ll manage and continuously improve workloads across AWS, GCP, or Azure
  • You’ll use Infrastructure as Code (Terraform) to standardize and scale infrastructure
  • You’ll optimize resource usage to balance performance and cost

You’ll run and improve Kubernetes in production

  • You’ll operate and scale Kubernetes clusters (EKS, GKE, etc.) with confidence
  • You’ll troubleshoot issues quickly and ensure smooth deployments and upgrades
  • You’ll ensure our containerized workloads perform reliably at scale

You’ll build strong observability and respond to incidents

  • You’ll implement and refine monitoring systems using tools like Prometheus, Grafana, Datadog, or ELK
  • You’ll define alerting that is meaningful, not noisy
  • You’ll respond to incidents, lead root cause analysis, and ensure we learn from every failure

You’ll automate everything that shouldn’t be manual

  • You’ll write scripts and build tooling to eliminate repetitive operational work
  • You’ll continuously improve infrastructure efficiency through automation
  • You’ll promote a culture where manual work is a temporary state, not the norm

You’ll collaborate to improve the entire system

  • You’ll work closely with DevOps and engineering teams to solve performance bottlenecks
  • You’ll contribute to CI/CD improvements and deployment reliability
  • You’ll help shape reliability best practices across the organization

What success looks like (First 90 Days)

First 30 days:

  • You’ve built a strong understanding of our infrastructure, systems, and workflows
  • You’re contributing to day-to-day operations with support from the team
  • You’ve started identifying areas for improvement in automation and reliability

By 90 days:

  • You’re independently managing infrastructure tasks and troubleshooting issues
  • You’re actively contributing to reliability and scalability improvements
  • You’ve taken ownership of parts of our infrastructure and are improving them

Requirements

Who You Are

This is what will make you successful in this role:

  • You’ve spent ~3 years working in SRE, DevOps, or infrastructure engineering, and you’ve seen what breaks at scale
  • You’re comfortable working in cloud environments like AWS, GCP, or Azure-and you understand how distributed systems behave
  • You’ve worked hands-on with Kubernetes in production and know how to troubleshoot it when things go wrong
  • You don’t just fix issues - you ask why they happened and make sure they don’t happen again

Technically, you likely:

  • Use Terraform (or similar IaC tools) to manage infrastructure
  • Work confidently with Docker and Kubernetes
  • Write scripts in Python, Bash, or similar to automate workflows
  • Understand CI/CD pipelines (Jenkins, GitHub Actions, Bitbucket, etc.)
  • Have a solid grasp of networking, load balancing, and high-availability design

When it comes to monitoring:

  • You’ve implemented tools like Prometheus, Grafana, Datadog, or ELK
  • You know the difference between useful alerts and noise
  • You focus on signals that actually drive action

What sets you apart:

  • You take ownership - you don’t wait to be told something is broken
  • You’re calm under pressure and methodical during incidents
  • You simplify complexity instead of adding to it
  • You communicate clearly, even when explaining deeply technical issues
  • You care about building systems that make other engineers more effective

Nice to Have (but not required)

  • Experience with RabbitMQ or Redis in production
  • Familiarity with Ansible or AWX
  • Exposure to multi-cloud or hybrid environments
  • Cloud certifications (AWS, GCP) or Linux certifications
  • Background from ITI (Information Technology Institute)

What the hiring process will look like

  • Screening Interview - Talent Acquisition
  • Technical Interview - SRE Lead
  • Technical Task
  • Final Interview - SRE Lead & Cloud DevOps Director
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
489,618 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Riyadh
$35k – $64k per year (Estimated) • Equity • In office • Full-Time • Toulouse
Python
SQL
Bash
DevOps
Git
Management
Jira
Apply
$94k – $130k per year • In office • Full-Time • Bachelor's Degree • Toronto
Python
Java
PHP
Ruby
Scala
AI/ML
Hadoop
Spark
Airflow
Dagster
XGBoost
Fine-tuning
Embeddings
Scikit-learn
AI Agents
NLP
TensorFlow
Keras
PyTorch
RAG
BERT
Agentforce
Knowledge Graph
Management
Slack
Apply
$197k – $314k per year • In office • Full-Time • 7+ years exp • PhD • Atlanta • Washington
Python
Go
Java
PHP
TypeScript
Ruby
Databases
MySQL
AI/ML
AI Agents
Agentforce
Management
Slack
Apply
$89k – $149k per year • In office • Full-Time • 3+ years exp • Associate's Degree • United States
Python
Apply
Electrical Engineer 2 days ago
$98k – $164k per year • In office • Full-Time • 5+ years exp • Associate's Degree • United States
Python
MATLAB
MATLAB
Simulink
Apply
$52k – $149k per year (Estimated) • Remote • Internship • Bachelor's Degree
Management
Intercom
Apply
$33k – $55k per year (Estimated) • In office • Internship • Riyadh
Apply
$77k – $177k per year (Estimated) • Remote
Python
SQL
AI/ML
AI Agents
Analytics
Tableau
Power BI
A/B Testing
Metabase
Apply
Solution Consultant 2 months ago
$79k – $188k per year (Estimated) • In office • Full-Time • Riyadh
Apply
$99k – $238k per year (Estimated) • In office • Full-Time • Riyadh
Apply
$32k – $85k per year (Estimated) • In office • Full-Time • 2+ years exp • Bachelor's Degree • Riyadh • Dammam • Jeddah
Apply
$119k – $245k per year (Estimated) • In office • Full-Time • 10+ years exp • Riyadh
Design
AutoCAD
Apply
$77k – $193k per year (Estimated) • In office • Riyadh
Python
JavaScript
TypeScript
Databases
PostgreSQL
Snowflake
Databricks
pgvector
Pinecone
Apache Kafka
Google BigQuery
BigQuery
AI/ML
Copilot
Cursor
LangGraph
LangChain
Claude Code
LlamaIndex
Dagster
dbt
AI Agents
Langfuse
LangSmith
AWS Bedrock
LLM
RAG
Braintrust
OpenAI
Anthropic
Lovable
Replit
LLM Guardrails
EU AI Act
Agentic Workflows
Frontend
Next.js
React.js
Lighthouse
DevOps
Terraform
GCP
Azure
CI/CD
AWS
Docker
Kubernetes
Vector
Amazon Kinesis
Cybersecurity
PCI DSS
GDPR
HIPAA
Threat Modeling
Design
Figma
Apply
$34k – $84k per year (Estimated) • Remote/Hybrid • Full-Time • 4+ years exp • Riyadh
Apply
Site Manager 6 hours ago
$48k – $127k per year (Estimated) • In office • Full-Time • Riyadh
Apply
See all jobs
This is one of many
489,618 more open roles from verified company boards, updated every day.