{"id":1789908,"url":"https://alion.io/job/hcss-senior-devops-engineer","title":"Senior DevOps Engineer","company":{"id":179869,"name":"HCSS","domain":"hcss.com","url":"https://alion.io/company/hcss","size_band":"201-500","is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Gem","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"remote","remote_scope":"stated_countries","remote_scope_basis":"posting_text","remote_working_hours":null,"hiring_geo_confidence":"structured","locations":[],"countries":[],"hiring_countries":["US"],"hiring_countries_total":1,"salary":null,"salary_estimate":{"min_usd":108000,"max_usd":198000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":293},"experience_years_min":8,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"Azure DevOps","optional":false},{"name":"Azure SQL Database","optional":false},{"name":"Bicep","optional":false},{"name":"CI/CD","optional":false},{"name":"GitHub Actions","optional":false},{"name":"Grafana","optional":false},{"name":"Incident Management","optional":false},{"name":"PowerShell","optional":false},{"name":"SQL","optional":false},{"name":"Terraform","optional":false},{"name":"SLI/SLO/SLA","optional":true},{"name":"Wi-Fi","optional":true}],"status":"live","first_seen_at":"2026-09-24T20:11:29Z","employer_posted_date":"2026-10-01","last_verified_at":"2026-10-06T20:24:46Z","board_verified":true,"closed_at":null,"days_open":12,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":12},"description":"We are HCSS. For the last 40 years, we have been developing software to help construction companies streamline their operations.Based in Sugar Land, TX, our mission is helping customers achieve excellence through our proven customer-centric, end-to-end solutions and exceptionally helpful service, while providing a great life for our employees. With this mission at the core of everything we do, HCSS is a pioneer and leader in the construction software space and a consistently recognized employer. We have earned Best Companies to Work for in Texashonors for 18consecutive years and have been named a USA Today Top Workplace. HCSS has also been recognized by Built Inas a Best Place to Work in Greater Houstonand by Construction Executivefor our technology innovation, reflecting our strong culture, industry leadership, and commitment to excellence.\nWHO WE NEED:\nAs a Senior DevOps Engineer with a focus on Site Reliability Engineering (SRE), you will play a key role in driving infrastructure resilience, availability, and operational excellence across our cloud environments. A core responsibility of this role is managing and optimizing Azure SQL Elastic Pools, ensuring performance, cost-efficiency, observability, and automation are aligned with business objectives. You will lead initiatives around high availability, disaster recovery, incident response, and reliability automation. Your expertise in Azure or AWS, observability tools such as Grafana, and scalable infrastructure will be essential to ensuring our systems are robust, performant, and recoverable.\nQualifications:\n8+ years of experience in DevOps or SRE roles with a strong focus on cloud infrastructure and systems reliability\n3+ years of hands-on experience with managing Azure SQL Elastic Pools, including performance tuning, scaling, and automation\n5+ years of expertise in Azure cloud services including networking, compute, databases, and identity\n3+ years of experience applying SRE principles including SLIs, SLOs, and incident management best practices\nExtensive experience with Infrastructure as Code tools such as Terraform, Bicep, or ARM templates\nExtensive experience building and managing CI/CD pipelines with tools like Azure DevOps or GitHub Actions\nStrong scripting skills using Azure CLI and PowerShell for automation and operational tasks\nExperience with monitoring and observability platforms, ideally Grafana, or a strong foundation in similar tools\nSoft Skills:\nStrong troubleshooting and problem-solving abilities.\nExcellent communication skills and a collaborative mindset to work with cross-functional teams.\nAbility to work independently, manage multiple tasks, and prioritize efficiently.\nA proactive attitude toward continuous improvement and learning.\nPreferred Qualifications:\nManaged 10+ Elastic pools and 100+ databases in Azure\nAdvanced level certifications on cloud infrastructure like Az-400 or equivalent\nRole Responsibilities:\nAzure Elastic Pool Management:\nTake ownership of the design, scaling, and optimization of Azure SQL Elastic Pools\nMonitor and tune pool performance to ensure efficiency and SLA compliance\nEstablish observability and alerting for SQL resource consumption, errors, and performance anomalies\nAutomate provisioning, scaling, and failover using infrastructure and scripting tools\nCollaborate with database and application teams to align on resource usage strategies\nHigh Availability and Disaster Recovery:\nDesign and implement highly available and fault tolerant systems\nDevelop and maintain disaster recovery strategies across critical services\nPerform regular failover testing, documentation, and validation of recovery procedures\nWork closely with infrastructure and development teams to ensure business continuity objectives are met\nMonitoring, Observability, and Automation:\nImplement and manage observability stacks with logs, metrics, traces, and alerting\nCreate dashboards and alerts in Grafana or similar platforms to track key system indicators\nDevelop automated solutions for provisioning, monitoring, and maintenance tasks\nContinuously improve system visibility and reduce time to detect and resolve issues\nAssist in the automation of performance testing to proactively identify bottlenecks, validate scalability and ensure reliable system behavior\nIncident Management and Operational Excellence:\nEstablish and refine incident response processes, escalation workflows, and resolution protocols\nLead root cause analysis, post-incident reviews, and continuous improvement efforts, including following up to ensure identified improvements to the application or process are implemented.\nDevelop and maintain runbooks, diagnostic tools, and automated remediation solutions\nChampion a blameless culture of reliability and operational readiness across engineering teams\nCloud Infrastructure Management:\nArchitect and manage scalable and secure cloud infrastructure in Azure or AWS\nProvision and manage services including compute, networking, storage, and containerized workloads\nContinuously monitor performance, latency, and uptime to ensure system health\nApply cost optimization practices while aligning infrastructure with business goals\nInfrastructure as Code (IaC):\nDefine and implement infrastructure using tools such as Terraform\nMaintain modular, version-controlled infrastructure code that supports environment consistency\nApply automation and policy enforcement to reduce drift and improve auditability\nEnsure IaC best practices are embedded in the development lifecycle\nCollaboration and Leadership:\nMentor and support junior DevOps engineers through code reviews, knowledge sharing, and technical guidance\nPartner with cross-functional teams including development, security, and database operations to drive initiatives\nAct as a subject matter expert in SRE practices and reliability-driven engineering\nLead continuous improvement efforts across infrastructure and operations processes\nTravel Requirements:\nRemote RequirementsEmployees will be expected to come into the office on a periodic basis.\nBaseline expectations for roles are as follows but may fluctuate based on manager’s discretion:Individual Contributor - Up to 2x per year\n\nEmployees will be expected to attend HCSS sponsored events per manager discretion (ex. UGM)\nBENEFITS & PERKS:\nPart of our mission is to provide a great life for our employees. We believe that when our people are happy, they do their best work. Some of the benefits and perks we offer include:\nFlexibility to work Remotely\nMedical, dental, and vision coverage with company-paid and employee-paid options\nPaid holidays, sick days, and personal time off\nEmployee Resource Groups (ERGs) that foster connection and inclusion\nOn-site amenities including a covered basketball court, soccer field, track, pickleball/tennis courts, gym, etc.\nDog-friendly campus and WiFi-accessible courtyards\n401(k) with a 5% company match\nCoverage for employee professional development and wellness\nAnd more!","description_format":"text","description_chars":6974,"description_truncated":false,"requirements":{"experience_years_min":8,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["Professional development"],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Fleet Management","Construction Management Software"],"lifecycle":[{"event":"open","at":"2026-10-03T17:45:26Z"}],"visa":[],"liveness":{"score":80,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.803,"p_room":1,"age_days":11,"expected_fill_days":42,"reasons":["conf:1","win:early"],"computed_at":"2026-10-06T05:45:30Z"},"pay":null,"html_url":"https://alion.io/job/hcss-senior-devops-engineer","json_url":"https://alion.io/job/hcss-senior-devops-engineer.json","meta":{"generated_at":"2026-10-06T22:22:36Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler","counted_by":"address","units_charged":1,"used_today":2770,"day_limit":5000,"remaining_today":2230,"minute_limit":60,"resets_at":"2026-10-07T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":179869},"rest":"https://alion.io/mcp/rest/get_company?id=179869"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fhcss-senior-devops-engineer"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fhcss-senior-devops-engineer"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fhcss-senior-devops-engineer"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/hcss-senior-devops-engineer\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fhcss-senior-devops-engineer"}]}