Confirmed on the employer's own hiring board on Oct 10, 2026. First seen by Alion on Oct 10, 2026.
Join Talon.One as a Technology Operations Engineer, where you'll be responsible for maintaining the reliability, scalability, and operability of our production systems. This hands-on role involves solving production-level problems, automating repetitive tasks, and building tools to improve the efficiency of our engineering teams. You'll work closely with SRE and R&D teams, optimize incident management, enhance observability, and support core production systems. Ideal candidates have 2-4 years of experience in TechOps, DevOps, SRE, or Production Engineering, and are comfortable working with Linux, command-line tools, and monitoring.
Missions
- Identifying manual or repetitive operational friction across R&D and building clean scripts, automation, and internal tools to solve it permanently.
- Designing, building, and integrating AI agents to streamline operational workflows, ensuring proper guardrails, monitoring, and human oversight for safe execution.
- Owning and optimizing Incident.io workflows, automation, and integrations, and staying closely engaged with incident response and post-incident reviews.
Profil recherché
- Willingness to learn and build deeper expertise in production systems and reliability- 2-4 years of experience in TechOps, DevOps, SRE, Production Engineering, or a similar technical role
- Comfortable working with Linux, command-line tools, logs, and monitoring
- Proactive attitude towards improving systems, processes, and tooling
- Ability to work collaboratively with engineers across different teams
- Experience working with production systems in a SaaS or cloud environment
- Experience with scripting or automation and a mindset of "if we do it twice, can we automate it?"
- A structured approach to troubleshooting and solving operational problems
- Experience with observability platforms such as Grafana, Datadog, Prometheus, or Sentry
- Familiarity with core SRE concepts, such as Service Level Indicators/Objectives (SLIs/SLOs) and error budgets
- Experience with Kubernetes and GCP
- Experience working with APIs, integrations, CI/CD, or infrastructure automation

