1,434,312open jobs
83,785companies
217,826added this week
Browse all
Salary
≈ $113k – $234k per year (Estimated)
Location
In office (Buffalo)
Seniority
Staff

Confirmed on the employer's own hiring board on Oct 9, 2026. First seen by Alion on Oct 7, 2026.

Overview
Company
Impact
Profile match
Glint Tech Solutions is a women-owned IT staffing and recruiting firm that sources contract and direct-hire talent for enterprise clients, alongside an AI hiring product it markets as Glinto AI. Its postings cover client projects across the United States and Canada, from New York, Sunnyvale, Santa Clara and Reston to Toronto, and some defense roles require an active DoD Secret clearance. Open positions are mostly placements at client companies, including software and DevOps engineers, embedded and field application engineers, product managers, data center cooling sales staff and regulatory reporting analysts.

Job Title: Lead Site Reliability Engineer

Location: Remote within the USA, or onsite in Buffalo, NY / Wilmington, DE (client preference for candidates near these areas). New hires are required to work onsite at the client's office for the first 2-3 weeks (treated as a business trip; travel expenses covered by the company).

Company Overview

Glint Tech Solutions is a women-owned, global IT staffing and recruiting firm serving enterprise clients across the USA and Canada.

Project Description

A leading financial services client is seeking a Lead Site Reliability Engineer responsible at the expert level for ensuring the reliability, scalability, performance, and operational excellence of critical banking platforms and applications. This senior individual contributor will design, implement, and improve SRE practices across the software development lifecycle, working closely with application development, infrastructure, platform engineering, and business teams to enhance system resiliency through automation, observability, testing, and proactive operational management, while coaching and influencing others.

Key Responsibilities

  • Design, implement, and support highly available, scalable, and resilient applications and cloud infrastructure following enterprise SRE best practices
  • Define, implement, and monitor SLOs, SLIs, and error budgets for critical business services
  • Develop observability strategies using Dynatrace, OpenTelemetry (OTel), distributed tracing, metrics, logging, dashboards, and alerting
  • Analyze production telemetry to proactively identify performance bottlenecks, reliability risks, and capacity constraints
  • Lead incident response for high-severity production events and facilitate Root Cause Analysis (RCA)
  • Drive operational excellence through automation of deployments, recovery procedures, and reliability controls
  • Design and execute automated regression testing strategies to validate stability and performance
  • Create and maintain Infrastructure as Code (IaC) solutions using Terraform
  • Support and optimize Microsoft Azure environments, including App Services, scaling, and deployment automation
  • Utilize Azure Monitor, Application Insights, and Log Analytics to improve platform visibility
  • Drive performance testing, resiliency testing, and disaster recovery preparedness
  • Lead capacity planning, performance tuning, and workload optimization
  • Develop operational runbooks, incident playbooks, and standard operating procedures
  • Mentor engineers on observability, cloud engineering, automation, and SRE principles
  • Adhere to Company risk and regulatory standards, policies, and controls

Mandatory Skills

  • Strong hands-on experience with Dynatrace, OpenTelemetry (OTel), distributed tracing, metrics collection, and centralized logging
  • Proven experience designing and executing automated regression testing frameworks
  • Strong proficiency in Infrastructure as Code (IaC) using Terraform
  • Experience with CI/CD pipelines, deployment automation, and operational tooling
  • Expert knowledge of production systems monitoring, incident management, and operational troubleshooting
  • Strong understanding of application performance management, distributed systems, and cloud-native architectures
  • Strong experience with Microsoft Azure (App Services, Resource Groups, networking, scaling, deployment/release management)
  • Experience with Azure Monitor, Application Insights, Log Analytics, and Azure dashboards/alerting
  • Experience supporting cloud-native and hybrid infrastructure environments
  • Demonstrated experience implementing SRE practices - SLOs, SLIs, error budgets, incident/problem management, RCA, reliability automation
  • Ability to improve system reliability through performance tuning, capacity planning, and observability-driven insights
  • Experience developing automated recovery mechanisms and self-healing solutions
  • Knowledge of resiliency engineering patterns, disaster recovery planning, and high-availability architectures

Nice-to-Have Skills

  • Experience supporting large-scale enterprise applications in regulated environments
  • Experience working in Agile and DevOps operating models
  • Ability to work autonomously and lead complex reliability initiatives
  • Experience partnering with architecture, infrastructure, cybersecurity, and application development teams
  • Scripting/automation experience with PowerShell, Python, or Bash
  • Industry certifications in Azure, Terraform, Cloud Engineering, or Site Reliability Engineering
  • Proven experience leading major incident response and post-incident improvement efforts
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
1,434,312 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

DevOps
Similar stack
Same company
Buffalo
$156k – $234k per year • Equity • In office • Full-Time • 10+ years exp • High School Diploma • Mounds View
Databases
Apache Kafka
AI/ML
A2A
DevOps
Azure
CI/CD
AWS
Kubernetes
Cybersecurity
ISO 27001
SOC 2
HIPAA
FedRAMP
Management
Agile
ITIL
Apply
≈ $126k – $259k per year (Estimated) • In office • United States
JavaScript
Node JS
Node JS
Commander.js
Databases
ElasticSearch
DevOps
GitHub Actions
CI/CD
GitOps
ArgoCD
Kubernetes
Platform Engineering
Buildkite
Apply
$135k – $206k per year • Hybrid • 12+ years exp • Bachelor's Degree • Chicago
Apply
$118k – $180k per year • Hybrid • 10+ years exp • Bachelor's Degree • Albuquerque
Apply
$135k – $206k per year • Hybrid • 12+ years exp • Bachelor's Degree • Albuquerque
Apply
$220k – $245k per year • In office • Full-Time • Bachelor's Degree • Buffalo
Apply
Mechanical Engineer 8 hours ago
$80k – $100k per year • In office • Full-Time • Buffalo
Design
AutoCAD
Apply
$38k – $56k per year • In office • Full-Time • High School Diploma • Buffalo
Management
Microsoft Office
Apply
Billing Specialist 11 hours ago
$38k – $46k per year • In office • Full-Time • 2+ years exp • High School Diploma • Buffalo
Apply
up to $80k per year • In office • Buffalo
Apply
See all jobs
This is one of many
1,434,312 more open roles from verified company boards, updated every day.