We’re looking for a Site Reliability Engineer to join our Engineering Department and help build and evolve MacPaw’s internal delivery platform.
The team builds shared capabilities across the delivery lifecycle - from CI/CD, observability, and monitoring to delivery automation and AI-powered engineering tooling - helping engineers move from code to production faster and more reliably. By adopting a platform-as-a-product mindset, we enable engineers to independently plan, build, test, deploy, recover, and monitor their services.
As a Site Reliability Engineer, you’ll help improve the reliability and observability of staging environments, resolve infrastructure bottlenecks, evolve delivery capabilities, and contribute to internal tools such as AI code review and automated quality gates. You’ll work hands-on with code, building and maintaining production-grade internal tooling rather than focusing only on scripts.
If you’re excited about building reliable engineering systems and improving developer experience at scale, we’d love to hear from you!
In this role, you will:
Own and improve staging observability by reducing alert noise, closing monitoring gaps, and building a reliable early-warning system for production-like issues.
Investigate and resolve infrastructure bottlenecks affecting service stability, such as message queue backpressure and resource limits.
Establish cost and performance monitoring for AI-driven workloads running in CI/CD pipelines (e.g., AI agents, automated code reviews).
Audit, harden, and maintain CI/CD pipeline configurations to ensure seamless, secure, and predictable software delivery.
Partner with engineering teams to extend reliability tooling, deployment checks, and feature-flag mechanisms to additional services.
Requirements
Skills you’ll need to bring:
2+ years of experience in SRE, DevOps, or System Engineering roles.
Strong programming experience in Go (Golang) or a background as a developer transitioning into operations, comfortable writing production-grade internal tooling, microservices, and libraries (not just basic scripts).
Solid experience maintaining and supporting internal software delivery tooling and CI/CD pipelines.
Background in operating and maintaining existing infrastructure and mission-critical services through regular maintenance and monitoring.
Strong engineering mindset with the ability to analyze complex issues, document system workflows, and consult developers on infrastructure best practices.
Practical, hands-on use of AI tools in your own daily work
At least an Upper-Intermediate level of English and fluent Ukrainian.
As a plus:
Platform / Internal-Product mindset (X-as-a-Service, contracts, SLAs, docs/runbooks).
Hands-on experience with Python or another secondary language for automation.
Deep observability background (SLOs/Error budgets, distributed tracing, SRE-style alerting).
Familiarity with DORA metrics (Deployment Frequency, Lead Time for Changes, CFR, MTTR) or progressive delivery/feature-flag ecosystems.

