Overview
Company
Profile match
Impact
Conditions
Benefits
Hiring process
Similar jobs
We believe in a brighter future of work in schools At Arbor, we’re giving schools and their staff the tools, time and insight they need to work at their best, and drive real change. Our mission To …
We are looking for an enthusiastic and proactive Site Reliability Engineer to join our SRE team and help us ensure we provide world-class resilience and performance across the platform. The remit and focus of the role is to advise on all aspects of site reliability including availability, scalability, observability and capacity planning. It’s a broad and exciting role, so we’re looking for someone up for a challenge - if you’re an energetic and a collaborative Site Reliability Engineer, this is the role for you.
Key responsibilities
- Proactively monitor and analyse platform performance.
- Collaborate with engineering teams to address performance bottlenecks and ensure scalability.
- Assist engineering teams with implementing and reviewing SLOs
- Continually improve observability through monitoring and alerting, and dashboards, using tools such as DataDog or Prometheus for example.
- Work with other teams to ensure it is effective and provides full coverage.
- Ensure the service is highly available and resilient
- Champion best practices in design for high availability
- Devise runbooks and run game sessions to test our DR plan, H/A and backups
- Conduct assessments of capacity and plan for scaling to meet current and future business needs.
- Work closely with the Head of Platform Engineering and Head of SRE to strategize and implement scalable solutions.
- Work closely with the Platform team, feature teams and, 2nd line support and other stakeholders to ensure a good level of service is provided for our customers and embed SRE practices.
- Key player in the response and troubleshooting of incidents, ensuring rapid resolution and minimising downtime.
- Participate in blameless postmortems to identify root cause and corrective actions
- Develop and maintain playbooks and documentation
Requirements
- 7-12 years of experience
- Experience in performance monitoring and analysis
- Capacity planning experience
- Scripting and automation skills, with experience in relevant technologies.
- Experience with Infrastructure as Code, in particular, Terraform
- Understanding of relational database technologies and their cloud versions (e.g. AWS Aurora)
- Experience with messaging and distributed asynchronous workloads
- Experience with nginx or similar technologies
- Familiarity with SRE processes.
- Aware of DevOps principles like the 3 ways and 5 ideals.
Desired Skills
- Experience with other database technologies and cloud platforms.
- Past experience with Enterprise solutions running at scale
- Familiarity with Kanban and Agile development processes
- Experience with containerisation, for example Docker
- Familiarity with software best practices such as Refactoring, Clean Code, Domain-Driven Design and Test-Driven Development.
Benefits
The chance to work alongside a team of hard-working, passionate people in a role where you’ll see the impact of your work everyday. We also offer:
- Hybrid work environment
- Group Term Life Insurance paid out at 3x Annual CTC (Arbor India)
- 32 days holiday (plus Arbor Holidays). This is made up of 25 days annual leave plus 7 extra companywide days given over Easter, Summer & Christmas
- Work time: 9.30 am to 6 pm (8.5 hours only)
- Compensation - 100% fixed salary disbursement and no variable components
Recommended for you based on this role
Similar stack
Same company
In your city
Senior Product Manager
16 days ago
$92k per year • Remote • Full-Time • United Kingdom
Design
Figma
Management
Jira
Marketing
Mixpanel
Apply
SRE Tech Lead
27 days ago
$105k – $118k per year • Remote • Master's Degree
Python
DevOps
AWS
Chaos Engineering
CI/CD
Configuration Management
Datadog
Docker
Incident Management
Kubernetes
Platform Engineering
Prometheus
Terraform
Apply
Apply
Senior DevOps Engineer
1 month ago
Remote/Hybrid • Full-Time • Bachelor's Degree • Thiruvananthapuram
Databases
Amazon Aurora
DevOps
Ansible
AWS
CI/CD
CloudFormation
Datadog
Docker
Nginx
Prometheus
Terraform
Apply
Career impact
Discover how this job can transform your career
Get a personal career forecast for this job - salary uplift, next-level role, skill boost and a 3-year financial impact, all calculated from your profile.
Personal salary uplift vs. your current pay
Your 3-year career trajectory
Skills you will level up in this role
3-year financial impact in dollars
Free forever • Less than a minute • No credit card
Work setup
Location
Thiruvananthapuram
Remote work
Remote/Hybrid (India)
Employment
Full-Time
Compensation
Benefits
Annual leave, Hybrid work, Life insurance

