615,181open jobs
29,638companies
85,681added this week
Browse all
Salary
$37k – $93k per year (Estimated)
Location
Remote (India)
Seniority
Senior · 5+ years exp
Overview
Company
Impact
Profile match

Research Engineer, Applied AI  

Location: Bangalore (or throughout India remote-friendly with travel) 

About EnCharge AI:  

EnCharge AI is building the next generation AI platform. Our novel in-memory-computing architecture delivers a 10x step-function improvement in compute energy efficiency and performance for AI inference workloads. As the demands of artificial intelligence move beyond today's models, we believe fundamental underlying infrastructure must evolve. We are an experienced team of AI researchers, silicon & systems engineers, and architects backed by leading investors, poised to become the essential platform for the next wave of AI innovation. 

The Opportunity:  

Modern AI workloads-from large language models to diffusion-based generators to multimodal systems-represent some of the most compute-intensive frontiers in AI, and some of the most promising applications for our hardware’s energy efficiency advantages. We’re building a vertically integrated AI stack that will showcase the transformative potential of our silicon while delivering real value to customers today. 

We are seeking a Research Engineer to push the boundaries of AI model capability, quality, and efficiency. You’ll build fine-tuning and post training pipelines, develop rigorous benchmarking frameworks, and work at the intersection of ML research and hardware-aware optimization-ensuring our models run beautifully on our silicon. 

This is a role for someone who thrives at the boundary between research and engineering. You’ll read papers, implement techniques, and ship production-quality code-all in service of making AI inference faster, cheaper, and better. 

Key Responsibilities:  

  • Algorithmic Acceleration:  Research and implement state-of-the-art techniques to accelerate AI inference-quantization, sparsity, distillation, speculative decoding, caching strategies, and architectural modifications. Systematically characterize tradeoffs between model quality, latency, throughput, and power consumption to find optimal operating points across different use cases.  
  • Hardware Co-Design:  Partner closely with hardware, compiler, and quantization teams to ensure algorithmic improvements translate to real gains on our silicon. Identify optimizations aligned with our architecture's strengths-maximizing throughput while minimizing power. Shape the feedback loop between model development and hardware. 
  • Evaluation:  Build profiling tools and comprehensive benchmarking frameworks to understand compute bottlenecks, measure model quality across standard and domain-specific evals, and track efficiency metrics.  
  • Applied Research:  Build robust fine-tuning workflows for modern AI models, enabling rapid experimentation with LoRA, adapters, and full fine-tuning. Stay current with the rapidly evolving landscape-evaluate new architectures, implement promising techniques, and contribute insights that inform technical and go-to-market strategy. 

Qualifications:  

  • 5+ years of experience in ML research, applied ML, or ML systems 
  • Strong fundamentals in Python and PyTorch 
  • Hands-on experience with transformers, diffusion models, state space models etc. 
  • Experience fine-tuning large models and building training/evaluation pipelines 
  • Deep understanding of transformers, attention mechanisms, & optimization techniques 
  • Comfort reading and implementing techniques from research papers 

Nice to Have:  

  • Experience with efficient inference techniques (KV cache optimization, attention variants, MoE routing, flow matching) 
  • Background in hardware-aware ML optimization or quantization 
  • Familiarity with profiling tools (PyTorch Profiler, Nsight, custom instrumentation) 
  • Publications in generative modeling, efficient inference, or ML systems 
  • Contributions to open-source ML projects 
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
615,181 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

Similar stack
Same company
In your city
$12k – $30k per year (Estimated) • In office • Full-Time • 3+ years exp • India
Python
SQL
Analytics
Microsoft Excel
Apply
$125k – $165k per year • In office • Full-Time • 3+ years exp • PhD • London • Princeton • Hyderabad • New York
Python
Groovy
DevOps
Terraform
GCP
Azure DevOps
Azure
CI/CD
Jenkins
AWS
JFrog Artifactory
GitHub
Apply
Lead Data Analyst 1 day ago
$18k – $37k per year (Estimated) • In office • Full-Time • Bachelor's Degree • Hyderabad
Python
SQL
Analytics
Microsoft Excel
Apply
$19k – $40k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Gurgaon • Bengaluru
Python
SQL
Python
pySpark
AI/ML
Spark
Function Calling
AI Agents
LLM
RAG
Structured Outputs
Agentic Workflows
DevOps
Terraform
CI/CD
AWS
Docker
Vector
Apply
$122k – $227k per year (Estimated) • Equity • In office • Full-Time • 5+ years exp • Bachelor's Degree • Santa Clara • Irvine
Python
MATLAB
AI/ML
Copilot
ChatGPT
Apply
$170k – $240k per year • Remote • 10+ years exp • Bachelor's Degree
Python
Apply
$200k – $250k per year • Remote • 8+ years exp • Master's Degree
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Multimodal AI
Computer Vision
AI Agents
TensorFlow
PyTorch
LLM
Edge AI
ONNX Runtime
Agentic Workflows
Robotics
Autonomous Navigation
Apply
AI Compiler Engineer 17 days ago
$190k – $255k per year • Remote • 3+ years exp • Bachelor's Degree
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Quantization
TensorFlow
PyTorch
TPU
Edge AI
MLIR
Apply
$200k – $250k per year • Remote • 8+ years exp
DevOps
RTOS
Chips/EDA
Synopsys ZeBu
Cadence Palladium
Siemens Veloce
Apply
Lead DFT Engineer 30 days ago
$200k – $250k per year • Remote • 12+ years exp • Bachelor's Degree
Chips/EDA
Tessent
Synopsys TestMAX
Apply
See all jobs
This is one of many
615,181 more open roles from verified company boards, updated every day.