{"id":791579,"url":"https://alion.io/job/zerorisk-software-engineer-support-operations","title":"Software Engineer – Support & Operations","company":{"id":687703,"name":"ZeroRisk","domain":"zerorisk.io","url":"https://alion.io/company/zerorisk","size_band":"51-200","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Ashby","truth_index":null},"role":"Backend","role_family":"Backend","seniority":"middle","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"board_field","remote_working_hours":null,"hiring_geo_confidence":"inferred","locations":["Dublin, Ireland"],"countries":["IE"],"hiring_countries":["IE"],"hiring_countries_total":1,"salary":null,"salary_estimate":{"min_usd":54000,"max_usd":134000,"period":"year","method":"role_seniority_country_cell","sample_n":8},"experience_years_min":3,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Angular","optional":false},{"name":"AWS","optional":false},{"name":"Git","optional":false},{"name":"Incident Management","optional":false},{"name":"Java","optional":false},{"name":"Datadog","optional":true},{"name":"Grafana","optional":true},{"name":"JavaScript","optional":true},{"name":"PagerDuty","optional":true},{"name":"Scrum","optional":true},{"name":"TypeScript","optional":true}],"status":"closed","first_seen_at":"2026-04-23T12:10:30Z","employer_posted_date":"2026-04-23","last_verified_at":"2026-10-02T10:23:46Z","board_verified":false,"closed_at":"2026-10-02T10:23:46Z","days_open":161,"trust":{"level":"stale","repost_count":0,"flags":["stale"],"days_open":161},"description":"Role Overview\nThe Software Engineer - Support & Operations is a full stack engineer whose primary focus is the stability, reliability, and recover-ability of a production SaaS environment. This role requires genuine engineering capability - reading and reasoning about Java and Angular codebases, navigating AWS infrastructure, and tracing a production problem from a customer symptom through logs, code, and infrastructure to its root cause.\nAlongside day-to-day investigation and fix work, this engineer actively builds and maintains the internal tooling and documentation that makes the team faster and more effective over time. The right person is technically solid, methodical under pressure, and motivated by making complex problems understandable and preventable.\nThis role is well suited to an engineer with approximately three years of professional experience who is looking to deepen their full stack skills in a high-feedback, production-facing environment.\nKey Responsibilities\nProduction Investigation & Bug Fixing\nTriage and investigate production issues - querying logs, correlating events across services, and identifying root causes rather than surface symptoms.\n\nNavigate both Angular front end and Java back end codebases to trace issues end to end, from a reported UI behaviour through to service logic and data layer.\n\nImplement targeted code fixes for confirmed bugs, ensuring all changes are covered by tests, submitted via pull request, and reviewed before merging.\n\nEscalate fixes that require architectural change or touch high-risk areas of the codebase to the Senior Engineer before proceeding.\n\nProduce clear post-incident summaries covering what happened, root cause, resolution, and steps being taken to prevent recurrence.\n\nInternal Tooling\nBuild and maintain internal tools that help the team investigate and manage recurring production issues more efficiently - log query utilities, diagnostic dashboards, automation scripts, and similar.\n\nIdentify manual or repetitive investigation steps that are candidates for tooling and prioritise building solutions that save meaningful time across incidents.\n\nMaintain existing tooling to ensure it remains accurate and useful as the platform evolves.\n\nTechnology & Tooling Evaluation\nProactively research and evaluate new tools and technologies that could improve operational efficiency - observability platforms, incident management tooling, log analysis tools, and similar.\n\nProduce concise assessments of candidate tools covering capability, integration effort, cost, and recommendation, sharing findings with the Senior Engineer and Staff Engineer.\n\nStay current with relevant developments in the SaaS operations and observability space.\n\nAWS Environment\nNavigate the AWS environment to support production investigations - reviewing logs, metrics, and infrastructure state to identify environment-level contributors to issues.\n\nWork with the DevOps function where infrastructure changes are required as part of issue resolution.\n\nPlaybooks & Documentation\nBuild and maintain a library of debugging playbooks and how-to guides covering common production issues - step-by-step enough that any engineer can follow them.\n\nUpdate playbooks after every significant incident to incorporate new learnings.\n\nIdentify gaps in the playbook library and prioritise filling them based on incident frequency and impact.\n\nSkills & Experience\nEssential\nApproximately three years of professional software engineering experience with practical exposure to both Java and Angular.\n\nExperience debugging issues in a cloud-hosted or SaaS environment, comfortable working with incomplete information under time pressure.\n\nWorking knowledge of AWS and confidence navigating cloud infrastructure for investigative purposes.\n\nStrong log analysis skills - able to construct queries, correlate events across services, and draw diagnostic conclusions.\n\nClear written communication - able to explain a production issue and its resolution to both engineers and non-technical stakeholders.\n\nFamiliarity with Git-based workflows and standard code review practices.\n\nDesirable\nExperience building internal tooling or automation to support operational workflows.\n\nExposure to observability or incident management platforms such as Datadog, PagerDuty, or Grafana.\n\nUnderstanding of relational database query analysis - able to identify slow or problematic queries as part of an investigation.\n\nFamiliarity with containerised deployment environments.\n\nWays of Working\nDiagnose before fixing. Understanding the root cause before touching code - and documenting that understanding - is what reduces incidents over time rather than managing them indefinitely.\n\nEvery incident is a learning opportunity. A playbook update or post-incident note is part of the job, not an optional extra.\n\nBuild things that scale. Recurring manual investigation steps are a signal to build a tool, not a reason to repeat the same steps indefinitely.\n\nEscalate early. If an investigation is not progressing, raise it to the Senior Engineer promptly. A delayed escalation is a worse outcome than an early one.\n\nReporting Structure\nReports to the Scrum Master / Team Manager. Works closely with Senior Engineers for code review and escalation of complex fixes. Escalates systemic issues to the Staff Engineer.\nThis role offers a clear development path toward a Senior Engineer position for engineers who broaden their full stack and infrastructure skills, or toward a specialist Site Reliability Engineer (SRE) track for those with a stronger infrastructure and observability focus.","description_format":"text","description_chars":5599,"description_truncated":false,"requirements":{"experience_years_min":3,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"Ireland","iso":"IE","kind":"country"},{"name":"Dublin","iso":null,"kind":"city"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Security Compliance","RegTech, AML & Compliance"],"lifecycle":[{"event":"open","at":"2026-09-12T03:41:43Z"},{"event":"close","at":"2026-10-02T10:23:46Z"}],"visa":[],"liveness":null,"pay":null,"html_url":"https://alion.io/job/zerorisk-software-engineer-support-operations","json_url":"https://alion.io/job/zerorisk-software-engineer-support-operations.json","meta":{"generated_at":"2026-10-04T02:15:40Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":3138,"day_limit":5000,"remaining_today":1862,"minute_limit":60,"resets_at":"2026-10-05T00:00:00Z"}}}