368,634open jobs
9,437companies
50,578added this week
Browse all
Salary
$139k – $249k per year (Estimated)
Location
Remote (United States)
Seniority
Architect
Employment
Full-Time
Overview
Company
Impact
Profile match
Zencore is a leading cloud consulting and services firm. Above all else we value experience, hard work and the ability to keep things in perspective and have fun as we ensure our customers are delighted with the work we do.

As a member of our Cloud Engineering team, you will be working with fast-paced innovative companies, leveraging AI as the key driver of their transformation. Our clients will look to you as their trusted advisor, someone they can rely on and who will be there to help them along their AI journey. You will be expected to cover a large spectrum of technology topics like model optimization, high-performance training on specialized hardware (TPUs), efficient model serving (vLLM), MLOps, and complex agentic systems.

At Zencore, a Principal Architect is a key technical leader in our engineering organization and acts as an ambassador of our technical and cloud engineering expertise. Principal Architects at Zencore are able to navigate a broad technical range but are specialized in one or more domains. They are responsible for the technical oversight and end-to-end delivery of projects within our professional services business.

Whilst this is remote role, we are only considering US based candidates and do not offer visa sponsorship.

What you will do...

  • Serve as Zencore’s senior-most technical authority on the practical application of advanced artificial intelligence and machine learning.
  • Partner with the sales and business development teams in a pre-sales capacity to scope opportunities, design solutions for proposals, and act as the senior technical voice in client pitches.
  • Lead the architecture and design of sophisticated, secure, and scalable AI solutions for our clients, moving beyond standard API integrations to create genuine competitive advantages.
  • Collaborate closely with Cloud & Data Architects to guarantee the design and deployment of comprehensive client solutions.
  • Address the growing demand for private, data-sovereign AI by designing systems that meet strict GDPR and data privacy requirements. Strive for model explainability and bias mitigation, ensuring solutions adhere to ethical standards and safety guardrails.
  • Architect solutions for hosting, fine-tuning, and optimizing both proprietary (e.g., Gemini, Claude) and open-source (e.g., Llama, Mistral) models on hyperscaler platforms.
  • Lead clients in selecting optimal cloud-native technologies, prioritizing Google Cloud solutions for deploying and scaling production-grade agentic systems.
  • Guide and mentor customers and Zencore's engineering teams on advanced topics, establishing best practices for high-performance training (PyTorch, JAX, TPUs), efficient model serving (vLLM), and complex agentic systems (LangGraph, Langchain, Google ADK).
  • Devise the financial architecture of AI solutions by performing ROI analysis and implementing cost-optimization strategies to ensure large-scale deployments remain economically sustainable for customers.
  • Act as an external thought leader, contributing to the Zencore brand through blog posts, conference presentations, and community engagement.
  • Act as a "player-coach," providing hands-on leadership and fostering a culture of deep technical excellence in AI/ML.

Who we need...

  • Master’s degree in Computer Science, natural sciences, mathematics, or a related technical field, or equivalent practical experience in designing and delivering high-scale AI/ML systems.
  • Extensive experience in a senior or principal architect role with a proven track record of designing and delivering complex, production-grade machine learning systems that have created measurable business value.
  • Deep, hands-on architectural experience with at least one major cloud platform (GCP, AWS, or Azure) is required.
  • Direct, hands-on experience with Google Cloud (Vertex AI, GKE, TPUs) is a significant plus.
  • Proven expertise in LLM optimization, including techniques for quantization, pruning, efficient fine-tuning (e.g., LoRA), and high-performance serving (e.g., vLLM, TensorRT-LLM).
  • Hands-on experience with high-performance ML frameworks (e.g., JAX, PyTorch/XLA) for training or fine-tuning large-scale models.
  • Expertise in designing and deploying agentic workflows using both code-centric (e.g., LangGraph, LangChain, Google ADK) and low-code (e.g., Vertex AI Agent Builder, LangSmith Agent Builder) paradigms.
  • A strong understanding of the architectural patterns required for building secure, private, and data-sovereign AI solutions.
  • Experience with LLM observability and evaluation frameworks (e.g., LangSmith, LangFuse, Vertex AI Evaluation).
  • Exceptional communication and stakeholder management skills, with the ability to articulate complex technical concepts and their business value to both technical and non-technical audiences.
  • A passion for mentoring and a drive for continuous learning in the fast-evolving AI landscape.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$105k – $175k per year • In office • Full-Time • Raleigh
Python
AI/ML
AI Agents
Amazon SageMaker
AutoGen
AWS Bedrock
CrewAI
Edge AI
Function Calling
LangChain
LangGraph
Model Context Protocol
NLP
Prompt Engineering
PyTorch
RAG
TensorFlow
DevOps
Amazon EC2
Amazon ECS
Amazon EKS
Amazon S3
AWS
AWS Lambda
Azure
GCP
Git
Kubernetes
Apply
$200k – $270k per year • In office • Full-Time • 10+ years exp • United States
DevOps
AWS
Azure
GCP
GitLab
Jenkins
SLI/SLO/SLA
Terraform
Apply
$59k – $99k per year • In office • Full-Time • Bachelor's Degree • Boca Raton
C++
JavaScript
Python
SQL
C#
C#
.NET
DevOps
AWS
Azure
Apply
$59k – $99k per year • In office • Full-Time • Bachelor's Degree • Alpharetta
C++
JavaScript
Python
SQL
C#
C#
.NET
DevOps
AWS
Azure
Apply
$44k – $58k per year • Remote/Hybrid • Full-Time • 2+ years exp • Associate's Degree • Vancouver
Bash
Python
SQL
Databases
MS SQL
MySQL
Oracle
PostgreSQL
DevOps
AWS
Azure
GCP
VMWare
Windows Server
Apply
$135k – $299k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • New York
DevOps
AWS
GCP
Apply
Program Director 19 days ago
$117k – $260k per year (Estimated) • Remote • Full-Time • 5+ years exp • London
DevOps
GCP
Apply
$168k – $307k per year (Estimated) • Remote • Full-Time • 5+ years exp • New York
DevOps
GCP
Apply
$80k – $190k per year (Estimated) • Remote • Full-Time • Master's Degree
AI/ML
AI Agents
Claude
Fine-tuning
Gemini
Google ADK
JAX
LangChain
Langfuse
LangGraph
LangSmith
Llama
LLM
LoRA
Mistral
PyTorch
Quantization
TensorRT
TensorRT-LLM
Vertex AI
vLLM
PEFT
LLM Guardrails
TPU
DevOps
AWS
Azure
GCP
Google GKE
Kubernetes
Cybersecurity
GDPR
Apply
$107k – $208k per year (Estimated) • Remote • Contractor • 3+ years exp
Databases
MySQL
DevOps
CI/CD
GCP
Kubernetes
VMWare
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.