368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$67k – $103k per year (Estimated)
Location
Remote (Canada, Poland, Estonia, Portugal)
Seniority
Senior
Employment
Full-Time
Overview
Company
Impact
Profile match
Glia is a customer interaction technology company headquartered in New York City and founded in 2012 under the name SaleMove. The company provides a unified platform combining voice, digital messaging, video banking, cobrowsing, and AI powered virtual assistants for regulated service industries. It serves banks, credit unions, and insurers primarily in North America and Europe, positioning its product as a single interaction layer for customer service teams.

About Glia

Glia is the #1 Banking AI platform, empowering community and regional financial institutions to create efficiencies, accelerate loan growth, drive deposits, and deliver experiences that win against megabanks and fintechs.

Glia's Banking AI Operating System is a central intelligence layer on top of existing tech stacks, activating an AI workforce of specialized agents that draw from banking data, interaction history, and integrated systems of record. These banking-trained agents automate workflows across voice and digital-from front office to back office-resulting in decreased operational costs and the Universal Banker model.

Trusted by 700+ banks and credit unions for its ironclad security and reliability, Glia delivers the industry’s first contractual no-hallucination guarantee. It’s why Glia customers quickly and confidently put Banking AI to work with measurable results from day one. More information about Glia can be found at glia.com.

The Team

You'll be joining our dedicated Observability Team, which builds the observability platform and SRE tooling used to monitor Glia’s cloud-native core infrastructure serving the conversational AI. Our team focuses on enablement and automation: we provide the standards, tooling, and platform that engineering teams use to keep their own systems available and performing optimally.

The Mission & The Work

As a Site Reliability Engineer on this team, your focus will be on building SRE and observability tooling - the platform, automation, and standards other teams use to keep their services healthy. This is a tooling and enablement role, not a production operations role: you will not be directly operating Glia’s production services. Responsibilities will include:

  • Developing standards, infrastructure and automation for dashboards, alerts, and monitors as code.

  • Partnering with development teams to establish production readiness and operational readiness.

  • Building the tooling and templates teams use to define, measure, and report on Service Level Objectives (SLOs) and Service Level Indicators (SLIs) for their services.

  • Developing tooling to automate observability and operational workflows, eliminating manual toil for engineering teams.

  • Building and improving the incident response tooling and workflows that help teams resolve outages faster and learn from them.

Our Collaboration Model

Glia Engineering is remote-first, spanning Canada, Portugal, Poland, and Estonia. The Observability team is based in Estonia, with optional offices in Tallinn and Tartu. We thrive on flexible remote collaboration, but we still bring the whole team together in Estonia twice a year for in-person innovation and connection.

Our tech stack

  • Infrastructure: AWS, Kubernetes (AWS EKS), Istio, EFK

  • Persistence: Amazon Aurora Serverless for Postgres, RabbitMQ, Amazon RDS

  • Cache: Amazon ElastiCache

  • Monitoring & Observability: DataDog with a focus on dashboards and alerts for system health

  • CI/CD: Github Actions, ArgoCD, Jenkins, Helm, with a focus on automation and pipeline optimization.

  • Infrastructure as Code: Terraform

  • Additionally, our Engineering teams use:

    • Backend: Python, Elixir, Node.js, Ruby, Go

    • Frontend: Javascript and React.js

    • Native mobile SDKs: Java and Swift

What We’re Looking for

  • Expert-level proficiency with AWS and Kubernetes (EKS), particularly in areas of observability, networking, and auto-scaling.

  • Experience with modern observability platforms (e.g., DataDog, Prometheus) and a deep understanding of metrics, logging, and tracing.

  • Deep, practical understanding of Site Reliability Engineering (SRE) principles (SLOs, error budgets, toil reduction).

  • Demonstrable experience analyzing and troubleshooting large-scale distributed systems.

  • Strong software development skills in a language like Python or Go, used to build operational tools, services, or automation.

  • Expertise in designing and operating robust CI/CD pipelines for a microservices architecture (e.g., using ArgoCD, Github Actions, Helm).

  • A systematic, data-driven approach to problem-solving and root cause analysis.

  • Proficiency in using AI tools thoughtfully, maintaining ownership of the final output while recognizing the tools' limitations.

Glia is an equal-opportunity employer. Glia does not discriminate against any employee or applicant because of race, creed, color, religion, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition (including breastfeeding), or any other basis protected by law.

The Glia Talent Acquisition team uses @glia.com and @gliatalent.com email addresses for coordinating interviews, providing updates, and sending documents.

Our hiring process involves an introduction, practical and team interviews, and a decision and offer. For more information, visit our Recruitment Privacy Notice page or contact our talent team via [email protected]

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$230k – $260k per year • Equity • Remote • Internship • Bachelor's Degree
Python
DevOps
Amazon EKS
AWS
Azure
CI/CD
GCP
Helm
Kubernetes
Terraform
Cybersecurity
FedRAMP
Orca Security
Apply
$84k – $178k per year (Estimated) • In office • Full-Time • 10+ years exp • Wellington
Java
Python
DevOps
Ansible
AWS
Azure
CI/CD
Docker
GCP
Helm
Kubernetes
Platform Engineering
Prometheus
Service Mesh
Terraform
GitLab
IAM
Apply
Platform Engineer 1 day ago
$87k – $140k per year • In office • Full-Time • 3+ years exp • Berlin
Databases
PostgreSQL
Redis
DevOps
AWS
Azure
Bicep
CI/CD
Docker
GCP
GitHub Actions
Kubernetes
OpenShift
Terraform
GitHub
Apply
Founding Engineer 1 day ago
$81k – $116k per year • In office • Full-Time • Bachelor's Degree • Munich
JavaScript
Python
TypeScript
Databases
MySQL
PostgreSQL
Frontend
Next.js
React.js
Tailwind CSS
DevOps
AWS
Azure
CI/CD
Docker
GCP
Grafana
Kubernetes
OpenTelemetry
Prometheus
Apply
$98k – $195k per year (Estimated) • In office • Full-Time • 7+ years exp • Wellington
Java
Python
SQL
Java
Spring Boot
Databases
Apache Kafka
Databricks
Neo4j
AI/ML
Flink
Spark
Frontend
GraphQL
DevOps
Azure
CI/CD
Datadog
Dynatrace
Kibana
Kubernetes
OpenShift
Platform Engineering
Splunk
Amazon ECS
Apply
$55k – $133k per year (Estimated) • In office • Full-Time • 3+ years exp
JavaScript
Node JS
AI/ML
Hallucination
DevOps
AWS
Azure
GCP
Apply
$63k – $114k per year (Estimated) • Remote • Full-Time
Elixir
Go
Python
Ruby
AI/ML
Claude
Claude Code
Hallucination
LLM
AI Agents
ISO 42001
LLM Guardrails
Model Context Protocol
DevOps
ArgoCD
AWS
CI/CD
GitHub Actions
Jenkins
Kubernetes
Terraform
GitHub
IAM
Cybersecurity
Falco
NIST CSF
Okta
PCI DSS
Snyk
SOC 2
Sysdig Secure
Apply
Equity • In office • Full-Time
Elixir
Go
Python
Ruby
AI/ML
Claude
Claude Code
Hallucination
LLM
AI Agents
ISO 42001
LLM Guardrails
Model Context Protocol
DevOps
ArgoCD
AWS
CI/CD
GitHub Actions
Jenkins
Kubernetes
Terraform
GitHub
IAM
Cybersecurity
Falco
NIST CSF
Okta
PCI DSS
Snyk
SOC 2
Sysdig Secure
Apply
In office • Full-Time
Elixir
Go
JavaScript
Node JS
Python
Ruby
Databases
Amazon Aurora
PostgreSQL
RabbitMQ
AI/ML
Hallucination
Frontend
React.js
DevOps
Amazon EKS
ArgoCD
AWS
CI/CD
Datadog
GitHub Actions
Helm
Istio
Jenkins
Kubernetes
Prometheus
Terraform
GitHub
Apply
$68k – $114k per year (Estimated) • Remote • Full-Time
Elixir
JavaScript
Ruby
Databases
Apache Kafka
DynamoDB
PostgreSQL
AI/ML
Claude
Claude Code
Copilot
Hallucination
LLM Guardrails
Frontend
React.js
DevOps
ArgoCD
AWS
CI/CD
Datadog
Docker
Jenkins
Kubernetes
Terraform
Amazon S3
GitHub
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.