748,498open jobs
46,235companies
111,268added this week
Browse all
Salary
≈ $104k – $278k per year (Estimated)
Location
In office (Miami)
Seniority
Intern · 2+ years exp
Employment
Internship

Confirmed on the employer's own hiring board on Sep 25, 2026. First seen by Alion on Sep 24, 2026.

Overview
Company
Impact
Profile match
Brightstar.AI is a boutique AI consulting firm that partners with leading organizations to design, implement, and evolve AI-driven strategies that deliver real P&L impact.

Role Type: External Research Affiliate / Think Tank Contributor

Location: South Florida preferred

Affiliation: Local university-based Masters or PhD researcher or postdoctoral AI researcher working under faculty direction

Organization: Brightstar.AI

Role Summary

Brightstar.AI is seeking a full-time or recently graduated Masters or PhD-level Applied AI Engineer Intern - Advancing Agent Quality with a strong research mindset, affiliated with a leading (South Florida) university, to contribute to its AI Think Tank and to the hands-on build-out of Brightstar’s proprietary AI systems. The intern will pair academic rigor with applied engineering: building the evaluation environments, automated raters, and diagnostic tooling that establish - with evidence rather than impression - whether an agent is reliable enough to put in front of operators, executives, and investors.

The first concrete assignment is the AI Twin Board, Brightstar’s proprietary board-simulation product, where the intern will help build the evaluation harness end to end: realistic task suites, calibrated LLM judge, failure-mode diagnostics, and the data flywheel that feeds improvements back into the agents themselves. The AI Twin Board is the starting point, not the full scope - as Brightstar’s AI portfolio expands, the same quality discipline will be applied to other agentic systems, internal platforms, and portfolio-company deployments.

Working in coordination with university professors and research leaders, the intern will also bring up-to-date knowledge of the fast-evolving AI landscape - especially large language models, reasoning systems, emerging model capabilities, and the societal implications of advanced AI - and help translate what current and next-generation systems may enable over the next six months to three years into Brightstar’s evaluation standards and Think Tank point of view.

Key Responsibilities

· Design, build, and scale realistic agent environments and task suites that reflect how Brightstar’s agents are actually used - board-level deliberation, research synthesis, diligence workflows, and multi-step tool use.

· Develop agentic LLM judge, calibrate them against human expert evaluation, and benchmark Brightstar agents against leading frontier models.

· Build trajectory-analysis frameworks and diagnostic tooling that root-cause agent failure modes - passivity, hallucination, persona drift, brittle tool execution - and feed fixes back into agent prompts, harnesses, and architecture.

· Enable agents to generate and iterate on their own verifiers (automated test cases, checklists, ground-truth sets) to support effective exploration and iterative problem-solving.

· Harvest multi-turn interaction trajectories into high-quality datasets and reward signals that support fine-tuning and post-training of the models Brightstar relies on.

· Contribute to the AI Twin Board build-out as the initial focus, and extend the same evaluation approach to other Brightstar AI initiatives as the roadmap grows.

· Contribute frontier AI research perspectives to the Brightstar.AI Think Tank as part of weekly review sessions and select strategic discussions.

· Review relevant research papers, conference findings, and publications from leading AI institutions and labs as inputs to the Think Tank knowledge base and to Brightstar’s evaluation methodology.

· Help interpret academic and technical developments for a business and investment audience, including implications for industry transformation and applied AI opportunities.

Minimum Qualifications

· Currently enrolled in a Masters or PhD program - or serving as a postdoctoral or early-career academic researcher - in Artificial Intelligence, Machine Learning, Computer Science, Computational Linguistics, or a related field.

· 1 years of research, project, or professional experience with LLMs or Agents.

· 2 year of experience with machine learning and deep learning.

· Experience in software engineering and cloud-based development.

· Experience with agent builders, e.g. OpenSDK, OpenAI Agent Builder, etc.

· Able to communicate complex technical concepts clearly to non-technical executive audiences.

Preferred Qualifications

· In-depth knowledge of machine learning algorithms, including supervised learning and reinforcement learning.

· Hands-on experience building LLM or agent evaluation systems - autoraters, human-evaluation pipelines, benchmark design, or evaluation infrastructure.

· Experience with multi-agent orchestration, retrieval-augmented generation, and tool-use / function-calling reliability.

· Active research or publications in LLMs, reasoning systems, model behavior, AI safety, or the societal impact of advanced AI systems.

· Strong research discipline, intellectual curiosity, and the ability to separate meaningful signal from AI hype.

· Affiliated with a (South Florida) university and working under the supervision of a faculty member or professor, or recent graduate.

Why This Role Matters

Brightstar.AI ’s Think Tank is intended to provide systematic forward-looking intelligence on where AI may create value, where it may disrupt industries, and how firms can better anticipate what may work tomorrow rather than only what worked yesterday. Masters and PhD-level Applied AI Engineers are a critical part of that model, bringing academic depth, frontier perspective, and analytical rigor into the Think Tank’s quarterly insights and longer-term point of view. Moreover, this role extends beyond research to focus on the practical application of AI in terms of AI Twin Board technology, which we see as a key differentiator in our research approach and value creation methodology.

Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
748,498 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Create a free account Continue with Google
Free forever. No card. Under a minute.

Your match

How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.

Recommended for you based on this role

AI/ML
Similar stack
Same company
Miami
Research Engineer 11 days ago
$190k – $260k per year • In office • Full-Time • 2+ years exp • PhD • New York
Python
AI/ML
Fine-tuning
Multimodal AI
Speech Recognition
PyTorch
Post-training
Text-to-Speech
Machine Learning
Apply
≈ $126k – $285k per year (Estimated) • In office • 2+ years exp • Mountain View
Python
AI/ML
Spark
Reinforcement Learning
AI Agents
TensorFlow
PyTorch
Gemini
Hallucination
SFT
Post-training
Machine Learning
Apply
≈ $118k – $266k per year (Estimated) • In office • 2+ years exp • PhD • New York
Python
AI/ML
Scikit-learn
Pandas
NumPy
Apply
≈ $135k – $305k per year (Estimated) • In office • 1+ year exp • PhD • Mountain View
AI/ML
RLHF
Reinforcement Learning
Gemini
LLM
DPO
SFT
PPO
Post-training
Machine Learning
Apply
≈ $132k – $298k per year (Estimated) • In office • 1+ year exp • PhD • Mountain View
AI/ML
Reinforcement Learning
AI Agents
Gemini
LLM
SFT
Post-training
Reward Modeling
Machine Learning
Apply
$38k – $58k per year • Hybrid • Full-Time • Turin • Rome • Milan • Bologna
Python
Java
TypeScript
C#
Scala
AI/ML
LangGraph
LangChain
LlamaIndex
Prompt Engineering
Function Calling
AI Agents
Semantic Kernel
LLM
RAG
OpenAI
Anthropic
Agentic Workflows
Multi-Agent Systems
DevOps
Azure
CI/CD
Apply
$127k – $237k per year • Equity • Hybrid • Full-Time • 8+ years exp • Bachelor's Degree • Toronto
Python
C#
AI/ML
Vertex AI
AI Agents
AWS Bedrock
LLM
OpenAI
DevOps
GCP
Azure DevOps
GitHub Actions
Azure
CI/CD
AWS
Docker
Platform Engineering
GitHub
Linux
Unix
Management
Agile
Scrum
Kanban
Apply
≈ $44k – $99k per year (Estimated) • Hybrid • Full-Time • 10+ years exp • High School Diploma • Bengaluru
PowerShell
AI/ML
AI Agents
LLM
Copilot Studio
DevOps
Azure
Cybersecurity
Microsoft Entra ID
DLP
Analytics
Power BI
Management
Power Automate
Power Apps
Microsoft Teams
SharePoint
Apply
$64k – $85k per year • Hybrid • Full-Time • 4+ years exp • Bachelor's Degree • Toronto
Python
SQL
Python
pySpark
Databases
Google BigQuery
BigQuery
AI/ML
Spark
Airflow
XGBoost
Vertex AI
AI Agents
LightGBM
TensorFlow
PyTorch
LLM
Amazon SageMaker
Feature Store
Recommender Systems
Machine Learning
DevOps
Terraform
Azure
IAM
Apply
$94k – $316k per year • Hybrid • Full-Time • 12+ years exp • Associate's Degree • Arlington • Milwaukee • Dallas • Columbus • Kirkland
AI/ML
Vertex AI
AI Agents
RAG
OpenAI
Anthropic
Context Engineering
Apply
≈ $35k – $60k per year (Estimated) • In office • Internship • Bachelor's Degree • Miami
Management
Microsoft Office
Apply
$58k – $95k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Miami
Apply
$36k – $50k per year • In office • Full-Time • 2+ years exp • Miami
Apply
$81k – $108k per year • In office • 4+ years exp • Bachelor's Degree • Miami
Management
Microsoft Office
Marketing
Salesforce
Apply
Recruiter 1 day ago
$85k – $95k per year • In office • Full-Time • 4+ years exp • Bachelor's Degree • Miami
Apply
See all jobs
This is one of many
748,498 more open roles from verified company boards, updated every day.