Salary
≈ $52k – $142k per year (Estimated)
Location
In office (Tokyo)
Employment
Full-Time
Overview
Company
Impact
Profile match
Headquartered in Tokyo, Japan, Sony AI is the artificial intelligence research and development organization within the Sony Group. The company focuses on developing ethical AI technologies designed to augment human creativity across gaming, imaging, sensing, robotics, and scientific discovery. By collaborating with Sony’s global entertainment and technology divisions, it drives innovations - such as autonomous gaming agents and advanced computer vision tools - to enhance digital media and creative workflows.
Technology Field
Machine Learning
Computer Vision
Position Summary
At Sony R&D, we are pioneering the next generation of generative AI technologies that will transform how creators and audiences experience entertainment. From anime to film, music to games, we are working across the Sony Group to build AI that doesn’t just push the boundaries of research, but also empowers creators to tell their stories in new and powerful ways.
Here are two of the groundbreaking projects you could be part of:
- Video Generation AI We are also venturing into one of the most exciting frontiers of generative AI: video generation. Our goal is to develop technologies that can revolutionize how movies, music videos, games, and anime are produced across Sony’s global businesses. Current video generation models show great promise but are far from production-ready. To make them truly usable, we need to solve critical challenges such as: - Controllability: camera work, character motion, lip-sync, preventing subject fragmentation - Editability: video inpainting, style transfer, compositing - Customization: tuning for specific IPs, ensuring character consistency, adjusting motion dynamics, reinforcing models with creator feedback This project is about building video AI that creators can actually use to tell compelling stories-not just generating flashy demos. If you are excited about building the future of video creation and transforming the global entertainment industry, we’d love to have you on board.
- Anime-Focused Image Generation AI Japan is home to one of the world’s richest collections of anime IP, including those from Sony’s group companies. We are taking on the bold challenge of developing a next-generation foundation model specialized in anime image generation-an AI that can truly transform how anime is made. Existing reference-based editing AIs may deliver strong draftsmanship, but they often fail to preserve artistic style and IP consistency, making them unsuitable for real production. Our vision is to go beyond what looks technically impressive and instead create an AI that creators genuinely need and can rely on in practice. This project is not only about pushing diffusion models-it’s about crafting high-quality datasets, extending existing models, and designing the full production pipeline required to make AI actually usable in animation studios. If you’ve ever dreamed of shaping the future of anime with AI, this is your chance to lead the way.
Responsibilities
- Conduct research and development on state-of-the-art image and video generation AI models
- Design and implement datasets, pipelines, and tools tailored for creative production workflows
- Collaborate with creators, studios, and partner companies to translate research into practical solutions
- Publish findings at top-tier international conferences and contribute to Sony’s global R&D leadership
- Drive innovation that enables new experiences across anime, film, music, gaming, and beyond
Required Qualifications
- Master’s degree in Computer Science or a related field
- Strong research or engineering background in machine learning, computer vision, or generative models
- Hands-on experience with diffusion models, GANs, transformers, or related deep learning architectures
- Proficiency in Python and deep learning frameworks such as PyTorch or TensorFlow
- Demonstrated publication record in top-tier conferences or journals (e.g., CVPR, ICCV, NeurIPS, ICLR, ICML, SIGGRAPH)
- Ability to design, train, and evaluate large-scale models or pipelines
- Strong problem-solving skills and a passion for tackling open research challenges
- Ability to communicate effectively in English (both written and spoken)
Preferred Qualifications
- Ph.D. in Computer Science or a related field
- Experience in image or video generation for creative applications (e.g., anime, film, gaming, VFX)
- Experience with dataset design, annotation, and curation for generative AI
- Understanding of creative production workflows (animation pipelines, film editing, compositing, or game asset creation)
- Hands-on experience with large-scale training on distributed systems or cloud platforms
- Strong track record of collaboration with interdisciplinary teams (e.g., artists, designers, producers)
- Interest in Japanese culture, anime, or entertainment industries
- Japanese language skills are a plus, but not required
Related Product, Service
- Anime Production Support Tools
- Film and Video Production Solutions
- Game Development Assets
- Music and Entertainment Experiences
- Cross-Media IP Applications
Development Environment
- Platform: Windows, Linux, Mac, AWS, GCP, Azure
- Programming Language: Python, JavaScript, etc.
- Computing: Private GPU Cluster for Deep Learning, Multi GPU PCs
- Communication: Slack, GitHub, MS Teams, etc.
Application Requirements
Essay: Required
Coding test: Not Required
Required Skills:
Computer Vision, Machine Learning (ML)Optional Skills:
Free account
Stop reading job ads. Get the ones that fit.
One free account turns this page into a shortlist built around your stack, your level and your pay.
Match on every job. Stack, seniority, pay and location, scored against your profile.
412,258 open roles. Read straight off company career pages, refreshed every day.
Unlimited applications. Every one you send is tracked in one place, on-site or on a company board.
3 tailored CVs a month. Rewritten for the exact job you are applying to. Included free.
Free forever. No card. Under a minute.
Your match
How well do you fit this role?
Two answers are enough for a real match. No account needed.
Check my fit
Answers stay in this browser until you create an account.
Recommended for you based on this role
Similar stack
Same company
Tokyo
Remote/Hybrid • Secret • PhD
Python
AI/ML
Computer Vision
Apply
AWS Cloud Infrastructure Engineer
10 hours ago
≈ $130k – $266k per year (Estimated) • In office • Public Trust • 10+ years exp • Bachelor's Degree • Reston
Python
PowerShell
AI/ML
Fine-tuning
AWS Bedrock
LLM
Anomaly Detection
Amazon SageMaker
DevOps
Splunk
Terraform
Ansible
GitHub Actions
AWS CDK
CloudFormation
Datadog
Prometheus
CI/CD
Jenkins
AWS
Docker
Kubernetes
Configuration Management
Amazon EKS
AWS Lambda
Amazon EC2
AIOps
GitHub
Amazon S3
IAM
Amazon CloudWatch
Cybersecurity
ISO 27001
SOC 2
FedRAMP
Apply
Product Solutions Specialist
10 hours ago
$72k – $78k per year • Remote/Hybrid • 2+ years exp • Bachelor's Degree • San Diego
Python
JavaScript
SQL
C++
Apply
Research Assistant III
10 hours ago
$56k – $60k per year • In office • Bachelor's Degree • San Diego
Python
SQL
Apply
Research Assistant I
10 hours ago
$34k per year • In office • Bachelor's Degree • San Diego
Python
AI/ML
Multimodal AI
AI Agents
Anomaly Detection
Edge AI
DevOps
Docker
Robotics
Sensor Fusion
Digital Twin
Apply
Internship - Researcher / Natural Language Processing (NLP) for Multimodal Machine Learning
2 months ago
In office • Internship • 3+ years exp • Master's Degree • Tokyo
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Multimodal AI
NLP
TensorFlow
PyTorch
LLM
Pre-training
Knowledge Graph
DevOps
GitHub
Apply
Full time - Researcher / Natural Language Processing (NLP) for Multimodal Machine Learning
2 months ago
≈ $43k – $114k per year (Estimated) • In office • Full-Time • 3+ years exp • Master's Degree • Tokyo
Python
C++
C++
TensorFlow C++
PyTorch C++
AI/ML
Multimodal AI
NLP
TensorFlow
PyTorch
LLM
Pre-training
Knowledge Graph
DevOps
GitHub
Apply
Full-Time - Audio-Visual AI Research Scientist
2 months ago
≈ $50k – $135k per year (Estimated) • In office • Full-Time • PhD • Tokyo
Python
AI/ML
Multimodal AI
Computer Vision
DevOps
GitHub
Apply
Internship - Audio-Visual AI Research Scientist
2 months ago
In office • Internship • PhD • Tokyo
Python
AI/ML
Multimodal AI
Computer Vision
Apply
Research Intern for Deep Generative Modeling
2 months ago
≈ $22k – $37k per year (Estimated) • In office • Internship • Tokyo
AI/ML
Multimodal AI
Computer Vision
Apply
Apply
≈ $50k – $106k per year (Estimated) • Remote/Hybrid • Full-Time • 10+ years exp • Bachelor's Degree • Tokyo
Management
Google Docs
Apply
≈ $20k – $41k per year (Estimated) • In office • Full-Time • 3+ years exp • Bachelor's Degree • Tokyo
DevOps
Incident Management
Management
ServiceNow
Apply
Account Manager (Academic and Government)
2 days ago
≈ $29k – $72k per year (Estimated) • Remote/Hybrid • Full-Time • 3+ years exp • Tokyo
Marketing
Salesforce
Apply
Apply
This is one of many
412,258 more open roles from verified company boards, updated every day.

