368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$125k – $243k per year (Estimated)
Location
Remote/Hybrid (Redwood City, United States)
Seniority
Staff · 8+ years exp
Employment
Full-Time
Overview
Company
Impact
Profile match
HOAi is an AI-powered workforce automation platform for the community association management (HOA) industry, headquartered in Sunnyvale, California. Founded in 2022 by Haoyu Zha and Zhixuan Lai, the Y Combinator-backed startup was acquired in late 2024 by Vantaca, a major community management software provider backed by Cove Hill Partners.

Description

HOAi is a fast-growing startup revolutionizing the community association management industry. Our AI workforce platform integrates machine learning technology to streamline labor-heavy processes, eliminating inefficiencies and driving scalability. With rapid growth in the AI space, we are pushing boundaries to redefine industry standards.

HOAi is the leading AI solution for the community association management industry, enabling organizations to deploy AI Agents that function like experienced managers. These AI Agents go beyond traditional AI by proactively executing complex, multi-step processes with human-like reasoning-working autonomously, 24/7, across your entire operation. This transformation optimizes labor costs, enables growth without additional hires, and ensures faster, higher-quality service for residents and board members.

HOAi was acquired by Vantaca in the fall of 2024. Vantaca just achieved unicorn status with a $1.25B valuation, so it's safe to say we're past the "scrappy startup phase." We're not just building a successful company - we're building the category-defining platform that will transform how an entire industry operates.

Here's the reality of our trajectory:

  • Growing 100% year-over-year
  • Our AI product (HOAi) went from $0 to millions in months
  • Backed by Cove Hill Partners and JMI Private Equity
  • 6M+ doors on our platform, displacing legacy systems

Overview

The AI Infrastructure Engineer at HOAi is responsible for scaling and maintaining the infrastructure that powers our AI-driven products and services. This role sits at the intersection of infrastructure engineering, machine learning operations, and product development, ensuring our AI systems operate with exceptional reliability, performance, and efficiency. The ideal candidate is someone who gets excited about making AI systems fundamentally faster and more scalable. You'll work directly with our engineering and product teams to build the foundational infrastructure that enables HOAi to deliver the most advanced AI product in the community association management industry.

Accountability Key Initiatives

  • Infrastructure Ownership: Design, build, and maintain the cloud architecture, model serving infrastructure, and ML pipelines that power HOAi's products
  • Performance Optimization: Profile and optimize AI workloads to achieve sub-second inference latency while managing costs effectively
  • Scalability & Reliability: Build auto-scaling systems, implement robust failover mechanisms, and ensure 99.99% uptime for mission-critical AI services
  • MLOps Excellence: Develop and maintain CI/CD pipelines for model deployment, monitoring, and versioning across development and production environments
  • Developer Enablement: Create tooling and infrastructure that allows product engineers to deploy AI features quickly and safely
  • Security & Compliance: Implement security best practices and ensure compliance requirements are met across all AI infrastructure

Expectations for Success

  • Infrastructure uptime and reliability
  • AI inference latency (p95, p99) and throughput metrics
  • Infrastructure cost efficiency and optimization (cost per inference, GPU utilization)
  • Time to deploy new models and workflows (deployment velocity)
  • Developer satisfaction and productivity using AI infrastructure tools
  • System observability and incident response time

Responsibilities

Performance & Scalability

  • Profile and optimize database queries, API endpoints, and ML inference pipelines
  • Implement caching strategies, connection pooling, and distributed systems for scale
  • Monitor and optimize GPU utilization, memory usage, and compute costs
  • Design load balancing and auto-scaling policies for variable AI workloads
  • Build disaster recovery systems with redundancy

MLOps & Deployment

  • Build and maintain CI/CD pipelines specifically for model deployment
  • Implement model versioning, A/B testing infrastructure, and rollout mechanisms
  • Create automated testing frameworks for model quality and performance regression
  • Develop infrastructure for model monitoring, drift detection, and retraining workflows
  • Manage experiment tracking and model registry systems

Observability & Reliability

  • Implement comprehensive monitoring, logging, and alerting across the AI stack
  • Refine dashboards for real-time visibility into system health and performance
  • Conduct post-mortems and implement reliability improvements
  • Design circuit breakers, retry logic, and graceful degradation for critical services

Security & Compliance

  • Refine security best practices for AI infrastructure and data handling
  • Ensure compliance with data privacy regulations and industry standards
  • Manage credentials and access control across infrastructure
  • Support security audits and vulnerability assessments

Collaboration & Documentation

  • Work closely with Product & Engineering team to understand infrastructure needs and to enable fast, safe feature deployment
  • Document infrastructure architecture, runbooks, and operational procedures
  • Mentor team members on infrastructure best practices and tooling
  • Contribute to technical strategy and architectural decisions

Requirements

Required Experience

  • 8+ years of experience in infrastructure engineering, DevOps, or SRE
  • Strong cloud platform expertise
  • Experience building and maintaining deployment pipelines
  • Experience with PostgreSQL, Redis, or other production databases
  • Experience with APM tools, metrics, logging, and alerting
  • Familiarity with vector databases, model serving frameworks and cross-system observability and traceability
  • Managing and optimizing GPU work
  • Real-time inference with low-latency serving infrastructure
  • LLM deployment
  • Track record of achieving 10x performance improvements

Skills & Competencies

  • Able to debug complex distributed systems and find root causes
  • Obsessed with latency, throughput, and resource efficiency
  • Defaults to automating repetitive tasks and building scalable solutions
  • Understands security implications and implements best practices
  • Able to explain complex technical concepts clearly
  • Works effectively across teams and functions
  • Takes initiative to identify and solve problems before they become critical
  • Comfortable with ambiguity and changing priorities in a fast-moving startup
  • Supporting A/B deployment strategy

Mindset & Approach

  • Extreme ownership: Takes full responsibility for outcomes, not just inputs
  • Customer-focused: Understands how infrastructure decisions impact end-user experience
  • Data-driven: Makes decisions based on metrics and evidence
  • Continuous learner: Stays current with evolving technologies and best practices
  • Quality-focused: Builds systems that are reliable, maintainable, and elegant
  • Velocity-minded: Balances speed with quality; ships incrementally and iterates
  • Bar raiser: Sets and maintains high standards; elevates team performance and output quality
  • Strong delegator: Empowers others effectively; distributes work based on strengths and growth opportunities

Core Values

  • Always Growing: Likes change and enjoys finding new ways to improve their knowledge and the product. Always ready to learn quickly, helping themselves and the team grow.
  • Win as a Team: Builds trust and works together by making sure everyone communicates well. Actively involved in daily work, working closely with the team, listening to their ideas, and celebrating successes together.
  • Accountability Starts with Me: Notices problems and takes personal action to solve them.
  • Unwavering Commitment to Customer Experience: Regularly talks to customers, taking personal responsibility to understand what they need, address concerns, and make their experience better with improved Vantaca processes.
  • Innovate Boldly: We challenge the status quo and push boundaries to create meaningful change. We act with urgency and purpose, knowing that innovation drives our success.

The HOAi Way

  • Shoot for Impossible and Make it Happen: Sets audacious goals that others might see as unreachable and breaks them into actionable steps. Relentlessly perseveres through obstacles with resourcefulness and determination, turning ambitious vision into tangible results.
  • Radical Candor Leads to Humble Excellence: Gives and receives direct, honest feedback with genuine care for others' growth. Stays open-minded on the path to continuous improvement, recognizing that the best outcomes come from carrying no ego.
  • Hire and Develop the Best Talent: Actively seeks exceptional people who raise the bar and invests deeply in their growth. Creates opportunities for team members to stretch their capabilities, providing coaching and support that helps them reach their full potential.

Why You Should Join Our Team

  • Our eNPS is +68! (Google it, that is great).
  • Benefits: Medical, Dental, and Vision kick in day one.
  • Unlimited PTO (with a requirement for employees to take a minimum of one continuous week per year).
  • 401K with Company Match.
  • Remote Flexible - come to the office when needed.
  • Great parental leave benefits.
  • Named on Inc 5000 list of America’s Fastest Growing Private Companies.
  • Named on Inc 5000 Vet 100 Private Companies list multiple years in a row.
  • Winner of Coastal Entrepreneur Award, Technology Category.
  • Active employee-led Culture Committee.
  • Ongoing industry and professional development trainings available to all employees.
  • Multiple leaders on the executive committee recognized as 40 under 40 recipients for contributions to business and community.
  • We’re playing offense to win! Our product market fit and our world-class employees make us the leader in our space. We’re building something cool and people like it here.

We receive many resumes for our open positions and each one is reviewed by a human being on our recruiting team. We will compare your background with the qualifications and requirements for the position.

If you are selected for an interview you will receive an e-mail from someone on our recruiting team with an @vantaca.com email address. It may take some time for us to review all of the applications so give us some time to respond. We appreciate your interest in this role.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Redwood City
$140k – $225k per year • Remote • Full-Time • 5+ years exp • Seattle
Python
Lua
Python
FastAPI
Celery
Pydantic
SQLAlchemy
Databases
pgvector
PostgreSQL
Redis
AI/ML
LLM
RAG
Hybrid Search
Reranking
Anthropic
LLM Evaluation
LLM Guardrails
OpenAI
Model Context Protocol
DevOps
OpenTelemetry
Apply
$68k – $85k per year • In office • Full-Time • Master's Degree • San Jose
Python
AI/ML
AI Agents
DevOps
Amazon EC2
AWS
AWS Lambda
Bitbucket
CI/CD
CloudFormation
Docker
Git
Kubernetes
Terraform
Amazon S3
IAM
HPC
Cybersecurity
Least Privilege
Apply
$140k – $225k per year • Remote • Full-Time • 5+ years exp • Bachelor's Degree • Seattle
Python
TypeScript
Node JS
JavaScript
Python
FastAPI
Node JS
Fastify
Databases
PostgreSQL
Redis
Frontend
React.js
Vite
DevOps
Kubernetes
WebSockets
QA
Playwright
Vitest
Apply
$21k per year • In office • Contractor • Yekaterinburg
Python
SQL
Python
FastAPI
Flask
AI/ML
Claude
Claude Code
Embeddings
Function Calling
LLM
RAG
OpenAI
OpenAI Codex
Structured Outputs
DevOps
Docker
Git
Apply
$185k – $260k per year • Remote • Full-Time • 8+ years exp • Bachelor's Degree
DevOps
AWS
CI/CD
GCP
Kubernetes
GitHub
Cybersecurity
Clair
Dependabot
OWASP Top 10
OWASP ZAP
Snyk
Trivy
Apply
$137k – $274k per year (Estimated) • Remote • Full-Time • 7+ years exp
Databases
Apache Kafka
DevOps
Azure
Apply
$91k – $166k per year (Estimated) • Remote • Full-Time • 3+ years exp
AI/ML
ChatGPT
Perplexity
Apply
$111k – $209k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp
Go
AI/ML
Claude
Claude Code
Cursor
AI Agents
Lovable
DevOps
Platform Engineering
Apply
$149k – $269k per year (Estimated) • Remote/Hybrid • Full-Time • 5+ years exp • Redwood City
JavaScript
Databases
PostgreSQL
AI/ML
AI Agents
Anomaly Detection
dbt
LLM
Sentiment Analysis
Function Calling
Frontend
Next.js
React.js
Apply
$128k – $262k per year (Estimated) • Remote • Full-Time
Databases
LookML
AI/ML
Claude
Claude Code
Cursor
dbt
LLM
LLM Evaluation
Model Context Protocol
Management
Slack
Apply
$91k – $182k per year (Estimated) • In office • 3+ years exp • Redwood City
Python
DevOps
AWS
CI/CD
GCP
Terraform
Apply
$96k – $120k per year • In office • Internship • Bachelor's Degree • Redwood City
JavaScript
AI/ML
AI Agents
Apply
$96k – $120k per year • In office • Internship • Master's Degree • Redwood City
JavaScript
Python
AI/ML
AI Agents
Time Series Forecasting
DevOps
GitHub
Apply
$83k – $176k per year (Estimated) • In office • 12+ years exp • Bachelor's Degree • Redwood City
AI/ML
OpenAI Codex
Management
Confluence
Jira
Smartsheet
Apply
$145k – $261k per year (Estimated) • Equity • Remote/Hybrid • Full-Time • 2+ years exp • Redwood City
AI/ML
AI Agents
Robotics
Digital Twin
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.