Overview
Company
Profile match
Impact
Conditions
Benefits
Hiring process
Similar jobs

Omilia

A visionary combination of technology and art that started from a small garage in 2002, Omilia is now home to solution architects, engineers, developers, linguists, and individuals who share one goal - to deliver unfeigned human experience through...

We are looking for a Senior Site Reliability Engineer with Cloud platform experience. This individual will be part of a team responsible for operating and maintaining production clusters and developing our observability solutions; they will collaborate with team members to develop automation strategies, monitoring & alerting, and ensuring overall platform reliability. Your goal will be to become an integral part of the team, making every challenge of the platform - your own challenge, and solving them accordingly.

Responsibilities

  • Ensure platform reliability and availability across production and pre-production environments through proactive monitoring, alerting, and automation.
  • First response for incidents, contribute to problem management and root cause analysis.
  • Supporting the development team's effort towards reliability, creating a solid reliability culture within the development lifecycle.
  • Develop troubleshooting documentation for production support resources.
  • Collaborate with Engineering teams to develop optimised and productive runbooks, operational documentation and automation of operational tasks.
  • Collaborate with development and cloud engineering teams to embed reliability and performance into the software delivery lifecycle.
  • Design, implement, and evolve observability solutions (metrics, logs, traces, dashboards) using tools such as Prometheus, Grafana, and ELK.
  • Participate in on-call rotations and continuously improve alert quality and response processes.
  • Champion a culture of reliability, performance, and continuous improvement across teams.

Requirements

  • Bachelor's Degree or MS in Engineering or equivalent.
  • Experience in operating at least one container orchestration cluster (Kubernetes, Docker Swarm).
  • Experience developing or maintaining software for production services at scale.
  • Experience with ELK.
  • Experience with AWS.
  • Experience with Grafana/Prometheus stack.
  • Strong scripting skills (Bash, Python or Go).
  • Excellent communication skills.
  • Thinking out of the box and anticipating challenges. It is imperative we are not simply reactive; we must expect challenges and question technologies, procedures and thinking already in place. You will be expected to constantly review and challenge at all levels.
  • Versatility. We work with agile/lean methods. We'd much rather iterate and learn than assume we know all the answers.
  • Being a team player. You don't (always) work in isolation and are excited by the thought of using your team whilst involving product, experience design, engineering, and more in the process.

Will be considered as a plus:

  • Telephony knowledge (SIP, VoIP);
  • Experience in Linux Administration (RedHat, CentOS, AL);
  • Working knowledge in Configuration Management tools (Terraform, Ansible);
  • Experience with TCP/IP and general networking concepts;
  • RDBMS knowledge (MySQL, Postgres);
  • NoSQL knowledge (Redis).

Benefits

  • Fixed compensation;
  • Long-term employment with the working days vacation;
  • Development in professional growth (courses, training, etc);
  • Being part of successful cutting-edge technology products that are making a global impact in the service industry;
  • Proficient and fun-to-work-with colleagues;
  • Apple gear.

Omilia is proud to be an equal opportunity employer and is dedicated to fostering a diverse and inclusive workplace. We believe that embracing diversity in all its forms enriches our workplace and drives our collective success. We are committed to creating an environment where everyone feels welcomed, valued, and empowered to contribute their unique perspectives without regard to factors such as race, color, religion, gender, gender identity or expression, sexual orientation, national origin, heredity, disability, age, or veteran status, all eligible candidates will be given consideration for employment.

Recommended for you based on this role

Similar stack
Same company
In your city
DevOps
Azure
SLI/SLO/SLA
Cybersecurity
Okta
Management
Google Workspace
Apply
Marousi
DevOps
Azure
SLI/SLO/SLA
Cybersecurity
Okta
Management
Google Workspace
Apply
Full-Time
DevOps
AWS
Rest API
Cybersecurity
GDPR
Marketing
Salesforce
Apply
AI/ML
Copilot
NLP
Cybersecurity
HIPAA
ISO 27001
PCI DSS
SOC 2
Marketing
Salesforce
Apply
Node JS
Python
JavaScript
AI/ML
Claude
Embeddings
Gemini
LLM
Prompt Engineering
RAG
Apply
Node JS
Python
JavaScript
AI/ML
Claude
Embeddings
Gemini
LLM
Prompt Engineering
RAG
Apply
Athens
DevOps
Ansible
AWS
Azure
Datadog
Docker
GCP
Grafana
Hyper-V
Kubernetes
Terraform
VMWare
Windows Server
Zabbix
Cybersecurity
ISO 27001
PCI DSS
SOC 2
Apply
Full-Time
Python
AI/ML
LLM
NLP
PyTorch
TensorFlow
Apply
Career impact
Discover how this job can transform your career
Get a personal career forecast for this job - salary uplift, next-level role, skill boost and a 3-year financial impact, all calculated from your profile.
Personal salary uplift vs. your current pay
Your 3-year career trajectory
Skills you will level up in this role
3-year financial impact in dollars
Create free account
Free forever • Less than a minute • No credit card

Work setup

Employment
Full-Time