678,420open jobs
39,320companies
99,473added this week
Browse all
Location
In office
Overview
Company
Impact
Profile match
Your AI agents, assistants, and apps need unified, current and trusted business context to reason, decide, and act. Arango is the solution.

Site Reliability Engineer

About Arango:

Arango delivers a unified, natively multimodel contextual data platform that powers AI agents, assistants, and applications with the unified, current, and trusted business context needed to reason, decide, and act at scale.

The Arango Contextual Data Platform connects fragmented enterprise data with LLMs, copilots, and AI agents through a simplified architecture delivered out of the box. By combining graph, vector, document, key-value, and search capabilities in a single platform, Arango eliminates the complex stacks many organizations build to operationalize enterprise AI.

Trusted by organizations including NVIDIA, HPE, the London Stock Exchange, PSI CRO, the U.S. Air Force, NIH, Siemens,Transient.AI, Matpriskollen, and Articul8, Arango helps enterprises move from AI pilots to reliable production systems faster while lowering infrastructure complexity and total cost of ownership. Arango is a proud member of the NVIDIA Inception Program and the AWS ISV Accelerate Program. Learn more atarango.ai,LinkedIn, andG2.

Location: Europe Time zone, preferably within the Europe itself (Remote)- Spain, Portugal, etc.

Job Overview:

At ArangoDB, we are building a robust, cloud-native infrastructure to support our distributed database systems, which power mission-critical applications for a wide range of industries. We are searching for a Site Reliability Engineer (SRE) to ensure the reliability, scalability, and performance of our infrastructure and applications, with a focus on automation, monitoring, and optimizing cloud environments.

As a Site Reliability Engineer (SRE), you will be responsible for maintaining and improving the reliability of our distributed database systems running on Kubernetes and cloud environments (AWS, Google Cloud). You will design, implement, and maintain scalable infrastructure solutions, improve and expand observability into these solutions, and troubleshoot complex system issues. It is expected that you will come to work to write clean and efficient code in Golang, working closely with development teams

Your goal is to ensure high availability and performance of our cloud-based systems, automating repetitive tasks, and enhancing our CI/CD pipelines. If you're passionate about building resilient systems, managing cloud infrastructure, and using Golang to create scalable solutions (or willingness to learn Golang), we want to hear from you!

Key Responsibilities:

  • Design, implement, and maintain cloud infrastructure on AWS and Google Cloud platforms.
  • Ensure the scalability, performance, and reliability of our Kubernetes-based distributed database systems.
  • Collaborate with developers to write efficient, production-grade code in Golang to automate infrastructure management and improve system operations.
  • Optimize and automate CI/CD pipelines, deployment processes, and monitoring systems to support our production environment.
  • Develop strategies for disaster recovery, high availability, and fault tolerance.
  • Proactively identify system bottlenecks, troubleshoot, and resolve issues across the stack (network, OS, cloud infrastructure).
  • Implement monitoring, logging, and alerting systems to ensure visibility into system health and performance.
  • Participate in on-call rotations to support critical production systems and respond to incidents.
  • Collaborate with cross-functional teams to improve overall system reliability and scalability.
  • Collaborate with the Customer Success team to resolve customer issues.

Required Skills and Qualifications:

  • Experience: SRE or DevOps Engineer background in cloud-native environments. Self-organized, autonomous remote team player with strong communication skills.
  • Cloud & Infrastructure: AWS and GCP; advanced Linux internals (processes, environment variables); containerization and orchestration (Docker, Kubernetes at scale).
  • CI/CD & Observability: CI/CD pipelines (Jenkins, CircleCI); monitoring, alerting, and logging (Prometheus, Grafana, ELK stack); Git version control.
  • Networking & Security: Core networking, security best practices, and systematic troubleshooting of complex infrastructure issues.
  • Development: Programming proficiency in Golang or Python.

Nice-to-Have:

  • Experience managing distributed databases or large-scale data storage systems. •Knowledge of security best practices in cloud environments.
  • Experience with scripting languages like Python or Bash.
  • Experience with Infrastructure-as-Code (IaC) tools like Terraform is a plus. •Experience working with GitOps
  • Strong programming skills in Golang, with experience in developing automation tools, scripts, or services.

What Makes Arango Special?

At Arango, we believe that AI is only as powerful as the data foundation. Our mission is to help organizations build AI systems that can reason, decide and act based on unified, current, and trusted business context at scale. We are helping define a new category of infrastructure: the contextual data layer for AI.

Working at Arango means:

Contributing to cutting-edge AI and data infrastructure

Collaborating with experienced engineers, marketers, and product leaders

Helping shape how enterprises build AI-powered applications

If you're excited about the intersection of AI, data, and social media, we’d love to hear from you.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
678,420 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
Remote/Hybrid • 3+ years exp
Python
Go
JavaScript
TypeScript
AI/ML
LangChain
LlamaIndex
Model Context Protocol
Embeddings
AI Agents
LLM
RAG
GraphRAG
Knowledge Graph
LLM Evaluation
Edge AI
DevOps
AWS
Docker
Kubernetes
Apply
In office • 5+ years exp
Python
Bash
Databases
Neo4j
MinIO
ArangoDB
AI/ML
Embeddings
Prompt Engineering
AI Agents
RAG
Edge AI
DevOps
Rest API
GCP
OpenShift
Helm
Azure
AWS
Docker
Kubernetes
Amazon EKS
Azure AKS
Vector
Amazon S3
Apply
$65k – $149k per year (Estimated) • In office • Full-Time • 10+ years exp • Hong Kong
Python
Apply
In office
Databases
Redis
Snowflake
Neo4j
Databricks
AI/ML
AI Agents
Edge AI
DevOps
GCP
AWS
Platform Engineering
Vector
Marketing
Salesforce
LinkedIn
Apply
Remote/Hybrid • 10+ years exp
Python
SQL
Databases
ArangoDB
AI/ML
AI Agents
LLM
RAG
Edge AI
DevOps
CI/CD
AWS
Docker
Kubernetes
Apply
In office • Internship • Bachelor's Degree
Databases
ArangoDB
AI/ML
Edge AI
DevOps
AWS
Vector
Cybersecurity
Zscaler
Management
Google Workspace
Marketing
LinkedIn
Apply
In office
Databases
Redis
Snowflake
Neo4j
Databricks
AI/ML
AI Agents
Edge AI
DevOps
GCP
AWS
Platform Engineering
Vector
Marketing
Salesforce
LinkedIn
Apply
In office • 5+ years exp • Master's Degree
Databases
ArangoDB
AI/ML
AI Agents
Edge AI
DevOps
AWS
Apply
Remote/Hybrid • 10+ years exp
Python
SQL
Databases
ArangoDB
AI/ML
AI Agents
LLM
RAG
Edge AI
DevOps
CI/CD
AWS
Docker
Kubernetes
Apply
In office • 10+ years exp
Databases
ArangoDB
AI/ML
AI Agents
Edge AI
DevOps
AWS
Management
Google Workspace
Marketing
LinkedIn
Apply
See all jobs
This is one of many
678,420 more open roles from verified company boards, updated every day.