{"id":880161,"url":"https://alion.io/job/worth-ai-senior-devops-engineer-infrastructure-reliability-3","title":"Senior DevOps Engineer, Infrastructure & Reliability","company":{"id":4923,"name":"Worth AI","domain":"worthai.com","url":"https://alion.io/company/worth-ai","size_band":null,"is_staffing_agency":false,"employer_type":"direct","is_intermediary":false,"listed_via":null,"ats_vendor":"Workable","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"senior","employment_type":"full_time","work_mode":"hybrid","remote_scope":null,"remote_scope_basis":null,"remote_working_hours":null,"hiring_geo_confidence":"structured","locations":["Orlando, United States"],"countries":["US"],"hiring_countries":[],"hiring_countries_total":0,"salary":null,"salary_estimate":{"min_usd":101000,"max_usd":197000,"period":"year","method":"role_seniority_country_remote_cell","sample_n":1473},"experience_years_min":null,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"Amazon EKS","optional":false},{"name":"Apache Kafka","optional":false},{"name":"AWS","optional":false},{"name":"CI/CD","optional":false},{"name":"Datadog","optional":false},{"name":"IAM","optional":false},{"name":"Kubernetes","optional":false},{"name":"PostgreSQL","optional":false},{"name":"Redis","optional":false},{"name":"SLI/SLO/SLA","optional":false},{"name":"Terraform","optional":false},{"name":"Amazon S3","optional":true},{"name":"ArgoCD","optional":true},{"name":"AWS Lambda","optional":true},{"name":"GitHub Actions","optional":true},{"name":"JavaScript","optional":true},{"name":"Python","optional":true},{"name":"Service Mesh","optional":true},{"name":"TypeScript","optional":true},{"name":"Zero Trust","optional":true}],"status":"live","first_seen_at":"2026-10-05T00:00:00Z","employer_posted_date":"2026-10-05","last_verified_at":"2026-10-11T21:16:11Z","board_verified":true,"closed_at":null,"days_open":6,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":6},"description":"Worth AI, a leader in the computer software industry, is looking for a Senior DevOps Engineer to join our Infrastructure team with a singular mission: to make our systems faster, more reliable, and more resilient while making life dramatically easier for engineers shipping software.\nThis is a hands-on build role. You will spend most of your time writing Terraform, tuning Kubernetes workloads, automating things that are currently manual, and shipping infrastructure changes to production. You'll join a small platform team with an established roadmap and existing patterns, and a strong voice in how the work gets built.\nImplement scalable Infrastructure-as-Code patterns using tools like Terraform to standardize cloud provisioning and reduce configuration drift.\nOwn and evolve our Kubernetes platform (EKS or self-managed), ensuring workloads are secure, scalable, and resilient by default.\nOptimize CI/CD pipelines to improve deployment frequency, reduce lead time, and increase confidence in releases.\nDesign and enforce secure networking, IAM, and secrets management strategies across environments.\nImprove observability by refining metrics, logs, and tracing using tools like DataDog, ensuring actionable insight into system health.\nOptimize cloud cost efficiency through rightsizing, autoscaling strategies, and architectural improvements.\nImplement disaster recovery planning, backup strategies, and multi-region resilience initiatives.\nRefactor brittle or manually managed infrastructure into automated, testable, and reproducible systems.\nIntroduce new infrastructure tooling or architectural shifts and drive adoption through documentation, workshops, and hands-on support.\nPartner with engineering teams to eliminate friction in CI/CD, deployments, and cloud environments.\nCommunicate technical trade-offs clearly across engineering and product stakeholders, balancing speed with safety.\nTechnology Stack\nCloud & Infrastructure: AWS (EKS, RDS, MSK, S3, Lambda, IAM, VPC)Containerization & Orchestration: Kubernetes, ArgoCD\nInfrastructure-as-Code: Terraform\nCI/CD: GitHub Actions\nMonitoring & Observability: DataDog\nData & Messaging: PostgreSQL, Kafka, Redis\nLanguages (as needed): Bash, Python, TypeScript, JavaScript\n\nRequirements\n8+ years in DevOps, SRE, or infrastructure engineering.\nProven experience designing and operating production Kubernetes environments at scale.\nDeep hands-on expertise with AWS infrastructure and cloud networking.\nStrong experience building and maintaining Terraform modules across large cloud environments.\nDemonstrated ownership of CI/CD systems and measurable improvement of DORA metrics.\nExperience leading incident response processes and driving meaningful postmortem outcomes.\nStrong understanding of distributed systems, event-driven architectures (Kafka), and database performance (PostgreSQL).\nProven ability to modernize legacy infrastructure and eliminate manual operational toil.\nTrack record of taking a scoped infrastructure project from an ambiguous starting point to production without needing daily direction.\nDemonstrated ability to build trust across teams while raising the reliability bar.\nSuccess Metrics\nSystem Reliability: Maintain or exceed defined SLO/SLA targets with reduced incident frequency and duration.\nInfrastructure Stability: Reduce production incidents caused by misconfiguration, manual processes, or infrastructure drift.\nOperational Efficiency: Increase the percentage of infrastructure managed through code and automation.\nCost Optimization: Improve cloud cost efficiency without sacrificing reliability or performance.\nBonus Points (Nice to Have)\nExperience coding applications\nExperience operating high-throughput Kafka clusters (MSK or self-managed).\nStrong background in database performance tuning (PostgreSQL, Redis).\nExperience implementing autoscaling strategies for high-traffic systems.\nFamiliarity with service mesh technologies.\nExperience building internal developer platforms (IDP).\nBackground in security best practices (zero-trust networking, policy-as-code).\nExperience with multi-region or globally distributed systems.\nExperience introducing platform-wide reliability frameworks (SLOs, error budgets, chaos testing).\nAll Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando.\nBenefits\nHealth Care Plan (Medical, Dental & Vision)\nRetirement Plan (401k)\nLife Insurance \nFlexible Paid Time Off \n9 paid Holidays \nFamily Leave \nRemote\nHybrid work (for Orlando Associates)\nFree Food & Snacks (Orlando)\nWellness Resources","description_format":"text","description_chars":4628,"description_truncated":false,"requirements":{"experience_years_min":null,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":["401k plan","Hybrid work","Life insurance","Parental leave","Retirement plans"],"hiring_locations":[{"name":"United States","iso":"US","kind":"country"}],"hiring_excludes":[],"relocation_offered":false,"industries":["Financial Software & Embedded Finance","Financial AI"],"lifecycle":[{"event":"open","at":"2026-09-14T00:15:34Z"}],"visa":[],"liveness":{"score":88,"band":"hot","label":"Hiring now","p_open":1,"p_active":0.877,"p_room":1,"age_days":5,"expected_fill_days":25,"reasons":["conf:2","velocity","win:early"],"computed_at":"2026-10-10T05:45:15Z"},"pay":null,"html_url":"https://alion.io/job/worth-ai-senior-devops-engineer-infrastructure-reliability-3","json_url":"https://alion.io/job/worth-ai-senior-devops-engineer-infrastructure-reliability-3.json","meta":{"generated_at":"2026-10-11T22:45:38Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers","about":"Alion is a live layer of people, companies and AI agents: who they are, whether they are real and active right now, what they do and how to work with them, readable by people and by agents and paid per call.","catalog":"https://alion.io/catalog.json","usage":{"tier":"crawler_verified","counted_by":"address","units_charged":1,"used_today":12947,"day_limit":null,"remaining_today":null,"minute_limit":300,"resets_at":"2026-10-12T00:00:00Z"}},"offers":[{"id":"company.slices","title":"One company in depth, by slice","status":"live","price":{"credits":0.02,"usd":0.002,"plus_per_slice":{"credits":0.05,"usd":0.005}},"unit":"per company, plus each slice with data","note":"the employer in depth","call":{"mcp_tool":"get_company","arguments":{"id":4923},"rest":"https://alion.io/mcp/rest/get_company?id=4923"},"human":"https://alion.io/catalog?offer=company.slices&for=job%2Fworth-ai-senior-devops-engineer-infrastructure-reliability-3"},{"id":"market.stats","title":"A market slice: pay, demand and time to fill","status":"live","price":{"credits":1,"usd":0.1},"unit":"per slice","note":"pay, demand and time to fill for this role and place","call":{"mcp_tool":"market_stats"},"human":"https://alion.io/catalog?offer=market.stats&for=job%2Fworth-ai-senior-devops-engineer-infrastructure-reliability-3"},{"id":"job.search","title":"Open jobs by role, technology, place, pay and visa","status":"live","price":{"credits":0.02,"usd":0.002},"unit":"per posting in a list","note":"similar open postings","call":{"mcp_tool":"search_jobs"},"human":"https://alion.io/catalog?offer=job.search&for=job%2Fworth-ai-senior-devops-engineer-infrastructure-reliability-3"},{"id":"company.verify","title":"Is this company real and active right now","status":"pilot","price":null,"unit":"per company","request":{"url":"https://alion.io/catalog/request","method":"POST","body":"{\"offer\": \"company.verify\", \"for\": \"job/worth-ai-senior-devops-engineer-infrastructure-reliability-3\", \"note\": \"what you need it for\"}"},"human":"https://alion.io/catalog?offer=company.verify&for=job%2Fworth-ai-senior-devops-engineer-infrastructure-reliability-3"}]}