{"id":758279,"url":"https://alion.io/job/wizdaa-senior-platform-architect","title":"Senior Platform Architect","company":{"id":689314,"name":"Wizdaa","domain":"wizdaa.com","url":"https://alion.io/company/wizdaa","size_band":null,"is_staffing_agency":true,"is_intermediary":false,"ats_vendor":"Recruitee","truth_index":null},"role":"DevOps","role_family":"DevOps","seniority":"staff","employment_type":"contractor","work_mode":"remote","remote_scope":"stated_regions","hiring_geo_confidence":"structured","locations":[],"countries":[],"hiring_countries":["AR","BR","MX","AG","AW","BS","BB","BZ","BM","BO","VG","BQ","KY","CL","CR","CW","DM","DO","EC","SV","FK","GF","GD","GP","GT","GY","HT","HN","JM","MQ","MS","NI","PA","PY","PE","PR","KN","LC","PM","SX"],"hiring_countries_total":45,"salary":null,"salary_estimate":{"min_usd":62000,"max_usd":138000,"period":"year","method":"global_role_cell_scaled_by_country","sample_n":514},"experience_years_min":5,"visa_sponsorship":false,"relocation_package":false,"has_equity":false,"technologies":[{"name":"AI Agents","optional":false},{"name":"ArgoCD","optional":false},{"name":"AWS","optional":false},{"name":"Azure","optional":false},{"name":"CI/CD","optional":false},{"name":"GCP","optional":false},{"name":"GitHub Actions","optional":false},{"name":"GitLab CI","optional":false},{"name":"GitOps","optional":false},{"name":"Google GKE","optional":false},{"name":"IAM","optional":false},{"name":"Kubernetes","optional":false},{"name":"LLM","optional":false},{"name":"LLM Guardrails","optional":false},{"name":"Platform Engineering","optional":false},{"name":"Terraform","optional":false},{"name":"Argo Workflows","optional":true},{"name":"BigQuery","optional":true},{"name":"Feast","optional":true},{"name":"Feature Store","optional":true},{"name":"FinOps","optional":true},{"name":"Gemini","optional":true},{"name":"Google BigQuery","optional":true},{"name":"Kubeflow","optional":true},{"name":"MLFlow","optional":true},{"name":"Vertex AI","optional":true}],"status":"live","first_seen_at":"2026-08-27T16:53:57Z","employer_posted_date":"2026-08-27","last_verified_at":"2026-09-24T03:13:58Z","board_verified":true,"closed_at":null,"days_open":27,"trust":{"level":"ok","repost_count":null,"flags":[],"days_open":27},"description":"We're looking for a Platform Architect who can set the standard for how we build, ship, and operate reliable cloud platforms at scale. You sit at the intersection of platform engineering and SRE. You'll own the path from infrastructure design to reliable production services, bringing DevOps rigor to complex systems.\nThis is not a ticket-processing role, and it's not a research role. You'll tackle hard problems: platform reliability, scalability, cost efficiency, deployment automation, and workload operations. You'll have the scope to solve them properly. Senior professionals here identify problems before they're asked and raise the ceiling on what the platform can do.\nWhat you will work on\nBuild and operate scalable backend and AI infrastructure, supporting real-time and batch workloads with a focus on performance, reliability, and multi-tenant architecture.\nDesign and maintain deployment workflows across services and environments, including versioning, staged rollouts, automated releases, monitoring, and safe rollback strategies.\nBuild and operate LLM and agentic systems in production, integrating model providers, APIs, gateways, tools, and external services while managing rate limits, reliability, guardrails, and graceful degradation.\nDevelop reusable services, APIs, automation, and data pipelines that support AI-powered products and internal platform capabilities.\nExtend infrastructure-as-code across the platform using Terraform and reusable patterns to provision and manage cloud services consistently across projects and environments.\nMaintain GitOps-based deployment workflows using tools such as ArgoCD, improving automation and consistency across environments and tenants.\nRun distributed workloads on Kubernetes (GKE), managing scaling, workload placement, tenant isolation, service reliability, and infrastructure capacity.\nImprove platform observability and reliability through metrics, logging, tracing, SLOs, alerting, incident response practices, and operational tooling.\nIdentify performance and infrastructure cost improvements across cloud services, compute resources, APIs, and AI workloads.\nUse agentic coding and AI development tools to accelerate engineering work, including scaffolding services, generating and reviewing infrastructure and application code, debugging, and automating repetitive workflows.\nWhat you won’t find here\nA platform team that maintains the status quo. We're actively building: new scale requirements, new architectural domains, and an ML/AI footprint that's growing fast. Senior engineers here shape how the platform evolves, and the tools available to do it are better than they've ever been.\nRequirements\nMust have\n5+ years in platform engineering, SRE, or infrastructure, with meaningful time operating production systems at scale.\nStrong SRE/DevOps foundation. You've owned reliability for production services, defined and measured SLOs, run post-mortems, and driven measurable improvements.\nDeep Terraform expertise. You actively manage complex Terraform state, reusable modules, and multi-project configurations in production, with CI-driven plan/apply workflows.\nStrong GitOps background (ArgoCD or Flux in production). You understand declarative infrastructure management at depth and have opinions on how to do it well.\nDeep Kubernetes knowledge. You've operated clusters in production, dealt with real failure modes, and understand the system at the control plane level.\nStrong cloud infrastructure background across at least one major public cloud (AWS, Azure, or GCP), including networking, compute, IAM, storage, and multi-account or multi-project design.\nHands-on experience building and operating CI/CD pipelines (GitHub Actions, Cloud Build, GitLab CI, or equivalent).\nAutomation-first thinking at a senior level. You implement systems that eliminate entire categories of manual work.\nActive user of agentic coding tools. You know how to direct them effectively, review their output critically, and use them to multiply your output.\nStrong communicator. You can articulate operational decisions, technical trade-offs, and incident summaries clearly to engineers and leadership alike.\nNice to have\nMLOps experience, including hands-on experience deploying and operating ML or AI workloads in production.\nStrong GCP experience, including VPC networking, Compute Engine, IAM, Cloud Storage, multi-project or organization design, and GKE (Standard and/or Autopilot).\nHands-on experience with GCP data services, especially BigQuery in production: partitioning and clustering, query cost and performance tuning, and dataset-level IAM. Familiarity with at least one of Dataflow, Pub/Sub, or Dataproc.\nExperience with GPU/accelerator scheduling and node lifecycle management in production (e.g., GKE node auto-provisioning, GPU time-sharing, or equivalent).\nExperience operating LLM inference at scale, managing provider quotas/throttling (TPS/TUPS), gateways, caching, and guardrails (e.g., Vertex AI, Gemini API, or equivalent).\nExperience with ML pipeline and orchestration tooling such as Argo Workflows, Kubeflow, Cloud Composer/Airflow, Vertex AI Pipelines, or equivalent.\nExperience with model registries, feature stores, and experiment tracking (e.g., MLflow, Feast, or equivalent).\nFamiliarity with model and data drift monitoring and ML-specific observability.\nBackground in FinOps: inference cost attribution, committed use discount (CUD) and reservation planning, and accelerator capacity forecasting.\nFamiliarity with data infrastructure such as object storage, CDC pipelines, or lakehouse patterns.\nExperience with multi-tenant infrastructure: isolation patterns, noisy neighbor mitigation, and tenant lifecycle management.\nPrior experience scaling ML or platform infrastructure at a startup moving toward enterprise-grade requirements.\nLocation: Remote in LATAM\nPayment in USD\nWorking hours: EST time zone","description_format":"text","description_chars":5900,"description_truncated":false,"requirements":{"experience_years_min":5,"management_years_min":null,"team_size_min":null,"manages_managers":false,"education":null,"security_clearance":false,"languages":[]},"benefits":[],"hiring_locations":[{"name":"Argentina","iso":"AR","kind":"region"},{"name":"Brazil","iso":"BR","kind":"country"},{"name":"Colombia","iso":null,"kind":"region"},{"name":"Mexico","iso":"MX","kind":"region"},{"name":"Anguilla","iso":null,"kind":"region"},{"name":"Antigua and Barbuda","iso":"AG","kind":"region"},{"name":"Aruba","iso":"AW","kind":"region"},{"name":"Bahamas","iso":"BS","kind":"region"},{"name":"Barbados","iso":"BB","kind":"region"},{"name":"Belize","iso":"BZ","kind":"region"},{"name":"Bermuda","iso":"BM","kind":"region"},{"name":"Bolivia","iso":"BO","kind":"region"},{"name":"British Virgin Islands","iso":"VG","kind":"region"},{"name":"Bonaire, Sint Eustatius and Saba","iso":"BQ","kind":"region"},{"name":"Cayman Islands","iso":"KY","kind":"region"},{"name":"Chile","iso":"CL","kind":"region"},{"name":"Costa Rica","iso":"CR","kind":"region"},{"name":"Curaçao","iso":"CW","kind":"region"},{"name":"Dominica","iso":"DM","kind":"region"},{"name":"Dominican Republic","iso":"DO","kind":"region"},{"name":"Ecuador","iso":"EC","kind":"region"},{"name":"El Salvador","iso":"SV","kind":"region"},{"name":"Falkland Islands","iso":"FK","kind":"region"},{"name":"French Guiana","iso":"GF","kind":"region"},{"name":"Grenada","iso":"GD","kind":"region"},{"name":"Guadeloupe","iso":"GP","kind":"region"},{"name":"Guatemala","iso":"GT","kind":"region"},{"name":"Guyana","iso":"GY","kind":"region"},{"name":"Haiti","iso":"HT","kind":"region"},{"name":"Honduras","iso":"HN","kind":"region"},{"name":"Jamaica","iso":"JM","kind":"region"},{"name":"Martinique","iso":"MQ","kind":"region"},{"name":"Montserrat","iso":"MS","kind":"region"},{"name":"Nicaragua","iso":"NI","kind":"region"},{"name":"Panama","iso":"PA","kind":"region"},{"name":"Paraguay","iso":"PY","kind":"region"},{"name":"Peru","iso":"PE","kind":"region"},{"name":"Puerto Rico","iso":"PR","kind":"region"},{"name":"Saint Kitts and Nevis","iso":"KN","kind":"region"},{"name":"Saint Lucia","iso":"LC","kind":"region"},{"name":"Saint Vincent and the Grenadines","iso":null,"kind":"region"},{"name":"Saint-Pierre and Miquelon","iso":"PM","kind":"region"},{"name":"Sint Maarten","iso":"SX","kind":"region"},{"name":"Suriname","iso":"SR","kind":"region"},{"name":"Trinidad and Tobago","iso":"TT","kind":"region"},{"name":"Turks and Caicos Islands","iso":"TC","kind":"region"},{"name":"Uruguay","iso":"UY","kind":"region"},{"name":"United States Virgin Islands","iso":"VI","kind":"region"}],"hiring_excludes":[],"relocation_offered":false,"industries":[],"lifecycle":[{"event":"open","at":"2026-09-11T17:41:39Z"}],"liveness":{"score":15,"band":"cold","label":"Long shot","p_open":1,"p_active":0.329,"p_room":0.45,"age_days":26,"expected_fill_days":16,"reasons":["conf:0","agency","velocity","win:tail"],"computed_at":"2026-09-23T05:45:00Z"},"pay":null,"html_url":"https://alion.io/job/wizdaa-senior-platform-architect","json_url":"https://alion.io/job/wizdaa-senior-platform-architect.json","meta":{"generated_at":"2026-09-24T05:43:06Z","cache_seconds":300,"methodology":"https://alion.io/methodology","terms":"https://alion.io/terms","contact":"https://alion.io/contact","api":"https://alion.io/developers"}}