368,634open jobs
9,437companies
50,578added this week
Browse all
Location
In office (Shanghai)
Seniority
Staff
Employment
Full-Time
Overview
Company
Impact
Profile match
Altera is a American semiconductor company that specializes in manufacturing programmable logic devices, primary field-programmable gate arrays (FPGAs), System-on-a-Chip (SoC) FPGAs, and related design software. Originally founded in 1983, the company pioneered reconfigurable hardware before being acquired by Intel in 2015 and later re-established as an independent, standalone enterprise. Its high-performance hardware and intellectual property cores are widely used across data centers, telecommunications, automotive systems, defense, and edge AI applications.

Job Details:

Job Description:

Role Overview

We are seeking a highly experienced AI Technical Leader to drive the design and deployment of next-generation agentic AI systems optimized for FPGA/ASIC platforms.

This role sits at the convergence of:

  • Large Language Models (LLMs)

  • Agentic AI & autonomous systems

  • Retrieval-Augmented Generation (RAG)

  • Hardware-aware AI system design (FPGA / ASIC / EDA)

You will define system architecture, lead technical strategy, and deliver scalable AI solutions that tightly integrate software intelligence with hardware acceleration, with a strong focus on AI model optimization and deployment on FPGA.

Key Responsibilities

AI Architecture & System Design

  • Architect and implement end-to-end agentic AI systems (data → reasoning → action)

  • Design scalable multi-agent orchestration frameworks

  • Build enterprise-grade RAG pipelines for knowledge integration and retrieval

  • Define system-level architecture bridging cloud, edge, and hardware acceleration

Agentic AI & LLM Development

  • Develop advanced agent workflows, including:

    • Tool-using agents
    • Planning and reasoning agents
    • Multi-agent collaboration systems
  • Evaluate and integrate state-of-the-art LLMs (open-source and proprietary)

  • Optimize prompting, memory, and reasoning strategies for performance and reliability

AI Model Optimization on FPGA (Core Focus)

  • Lead AI model optimization and deployment on FPGA platforms, including:

    • Quantization (INT8 / BF16 / mixed precision)
    • Graph compilation and operator fusion
    • Model partitioning across CPU/GPU/FPGA
    • Memory and bandwidth optimization
  • Develop workflows to map AI models (e.g., CV, LLM inference) from PyTorch/Hugging Face to FPGA

  • Enable low-latency, high-efficiency inference pipelines on FPGA-based systems

  • Collaborate with hardware teams to align model architecture with FPGA constraints (DSP, BRAM, interconnect)

Hardware-Aware AI Co-Design

  • Co-design AI systems with:

    • FPGA acceleration
    • ASIC platforms
    • Heterogeneous compute architectures (CPU/GPU/FPGA)
  • Integrate AI workflows into EDA toolchains and hardware design flows

  • Drive innovations in software-hardware co-optimization

Framework & Platform Development

  • Build internal AI platforms leveraging:

    • LangChain
    • LlamaIndex
    • Custom agent orchestration frameworks
  • Develop reusable components for:

    • Memory management
    • Tool integration
    • Knowledge retrieval and indexing
  • Enable scalable deployment pipelines for AI applications on FPGA

Technical Leadership

  • Lead architecture design reviews and set technical direction

  • Mentor senior engineers and guide cross-functional teams

  • Drive innovation at the intersection of AI systems and semiconductor platforms

What Makes This Role Unique

  • Define agentic AI architecture for hardware-native AI systems

  • Work at the frontier of:

    • AI reasoning systems

    • FPGA-based acceleration

    • Next-generation computing platforms

  • High-impact role shaping AI + semiconductor convergence

Impact

You will enable a new class of intelligent systems where AI agents are tightly coupled with hardware, delivering:

  • Real-time decision-making pipelines

  • Efficient, scalable AI deployment on FPGA

  • End-to-end automation from model to hardware execution

Summary

This role is ideal for a senior technical leader who can:

  • Bridge LLM + Agentic AI + RAG with FPGA/ASIC system architecture

  • Drive AI model optimization and deployment on FPGA

  • Lead innovation in next-generation AI + hardware integrated systems

Qualifications:

Required

  • Strong experience in LLMs, agentic AI, or RAG systems

  • Proven track record in large-scale AI system architecture and deployment

  • Solid programming skills in Python/C++ and modern AI frameworks

  • Experience designing distributed or production AI systems

  • Strong system-level thinking across software and infrastructure

Preferred / Plus

  • Experience with FPGA, ASIC, or hardware acceleration platforms

  • Familiarity with FPGA development flows and toolchains (e.g., synthesis, compilation, HLS, runtime)

  • Experience with AI model optimization for hardware (quantization, compilation, deployment)

  • Exposure to EDA tools and semiconductor design workflows

  • Background in heterogeneous computing systems (CPU/GPU/FPGA)

Job Type:

Regular

Shift:

Shift 1 (China)

Primary Location:

Shanghai, China

Additional Locations:

Posting Statement:

All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance.
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
368,634 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
Shanghai
$17k – $49k per year (Estimated) • Remote/Hybrid • Kaliningrad
Node JS
Python
JavaScript
Node JS
BullMQ
Fastify
Python
FastAPI
Flask
Databases
Apache Kafka
Chroma
Pinecone
PostgreSQL
Qdrant
RabbitMQ
Redis
AI/ML
Claude
Copilot
Cursor
LangChain
LlamaIndex
LLM
Model Context Protocol
Prompt Engineering
RAG
Anthropic
OpenAI Codex
Structured Outputs
Function Calling
Frontend
Next.js
React.js
DevOps
AWS
Azure
CI/CD
Docker
GCP
Git
Gitflow
Rest API
WebSockets
GitHub
Management
Jira
Apply
Applied - AI Engineer 11 hours ago
$25k – $69k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Bengaluru • Pune
Java
Python
AI/ML
AI Agents
AWS Bedrock
Fine-tuning
Google ADK
LangGraph
LLM
LoRA
RAG
Semantic Search
LangChain
PEFT
A2A
Amazon SageMaker
AWS Strands Agents
NIST AI RMF
Semantic Search
Model Context Protocol
DevOps
Amazon EC2
Amazon EKS
AWS
Azure
CI/CD
CloudFormation
Docker
GCP
Git
GitOps
gRPC
Kubernetes
OpenTelemetry
Rest API
Terraform
Vector
Amazon S3
IAM
GitLab
Apply
Lead AI Engineer 10 hours ago
$30k – $73k per year (Estimated) • In office • Full-Time • Pune
Python
AI/ML
Fine-tuning
LLM
Reinforcement Learning
LLM Guardrails
AI Agents
DevOps
CI/CD
Docker
GitOps
Helm
Kubernetes
OpenShift
Platform Engineering
Vector
Apply
$71k – $154k per year (Estimated) • In office • Full-Time • Dublin
Java
Databases
Apache Kafka
AI/ML
Copilot
LLM
LLM Guardrails
DevOps
AWS
CI/CD
Kubernetes
GitHub
Apply
$32k – $72k per year (Estimated) • In office • Full-Time • 6+ years exp • Bengaluru
C++
Java
Python
YARA
Databases
Amazon Aurora
DevOps
Azure
GCP
Kubernetes
Cybersecurity
MITRE ATT&CK
Suricata
YARA
Zeek
Apply
$44k – $90k per year (Estimated) • In office • Full-Time • 10+ years exp • Bachelor's Degree • Bengaluru
C++
MATLAB
Python
Apply
CISO 4 days ago
$300k – $370k per year • In office • Full-Time • 15+ years exp • Bachelor's Degree • San Jose
Apply
$187k – $270k per year • In office • Full-Time • 10+ years exp • Bachelor's Degree • San Jose
Python
Apply
$30k – $72k per year (Estimated) • In office • Full-Time • 5+ years exp • Bachelor's Degree • Bengaluru
C++
Perl
Python
Verilog
VHDL
Apply
$33k – $82k per year (Estimated) • In office • Full-Time • 12+ years exp • Bengaluru
DevOps
AWS
Azure
CI/CD
GCP
IAM
Cybersecurity
Defense in Depth
GDPR
ISO 27001
Microsoft Entra ID
NIST 800-171
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Shanghai
DevOps
AWS
Cybersecurity
GDPR
PCI DSS
Analytics
Power BI
Tableau
Management
Confluence
Jira
Trello
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Shanghai
C++
Python
SQL
AI/ML
LangChain
LLM
Spark
Apply
In office • Full-Time • 15+ years exp • Bachelor's Degree • Shanghai
AI/ML
CUDA
CUDA Toolkit
LocalAI
DevOps
CI/CD
KVM
QEMU
Apply
In office • Full-Time • Shanghai
Apply
In office • Full-Time • 5+ years exp • Bachelor's Degree • Shanghai
Apply
See all jobs
This is one of many
368,634 more open roles from verified company boards, updated every day.