Deep understanding of modern foundation models, including video models, vision-language-action models, diffusion or flow models, self-supervised learning, or world-model architectures. * Experience ...
Deep understanding of modern foundation models, including video models, vision-language-action models, diffusion or flow models, self-supervised learning, or world-model architectures. * Experience ...
You'll be responsible for designing sophisticated model architectures, fine-tuning performance ... of Video or Image processing or Computer Vision Solid programming skills for common ML frameworks ...
You'll be responsible for designing sophisticated model architectures, fine-tuning performance ... of Video or Image processing or Computer Vision Solid programming skills for common ML frameworks ...
You understand diffusion and video models viscerally -- you've debugged them, trained a LoRA, optimized inference, built a ComfyUI workflow you'd defend in public. * You can name your daily model ...
Quick apply
You understand diffusion and video models viscerally -- you've debugged them, trained a LoRA, optimized inference, built a ComfyUI workflow you'd defend in public. * You can name your daily model ...
You understand diffusion and video models viscerally - you've debugged them, trained a LoRA, optimized inference, built a ComfyUI workflow you'd defend in public. * You can name your daily model ...
You understand diffusion and video models viscerally - you've debugged them, trained a LoRA, optimized inference, built a ComfyUI workflow you'd defend in public. * You can name your daily model ...
Lead Video Producer
San Francisco, CA · On-site
$151K - $178K/yr
Your video work will tell the stories of people who keep the world's operations moving. * You are ... Champion, role model, and embed Samsara's cultural principles in every production. * Manage ...
Lead Video Producer
San Francisco, CA · On-site
$151K - $178K/yr
Your video work will tell the stories of people who keep the world's operations moving. * You are ... Champion, role model, and embed Samsara's cultural principles in every production. * Manage ...
Deep understanding of modern foundation models, including video models, vision-language-action models, diffusion or flow models, self-supervised learning, or world-model architectures. * Experience ...
Deep understanding of modern foundation models, including video models, vision-language-action models, diffusion or flow models, self-supervised learning, or world-model architectures. * Experience ...
You will work across video/image embedding models, LLM-based video understanding, and agentic pipelines that orchestrate multiple models into end-to-end workflows. This is a foundational role that ...
Quick apply
You will work across video/image embedding models, LLM-based video understanding, and agentic pipelines that orchestrate multiple models into end-to-end workflows. This is a foundational role that ...
Our system processes high volume video streams, runs YOLO based detection models, performs temporal tracking and smoothing to reduce false positives, and identifies actionable safety violations.
Our system processes high volume video streams, runs YOLO based detection models, performs temporal tracking and smoothing to reduce false positives, and identifies actionable safety violations.
Video Editor
Manhattan, NY · Hybrid
We are proud to offer a hybrid work model to support the work-life balance of our teams, and we are ... About the Role We're looking for a motivated Video Editor to help create storytelling-led, visually ...
Video Editor
Manhattan, NY · Hybrid
We are proud to offer a hybrid work model to support the work-life balance of our teams, and we are ... About the Role We're looking for a motivated Video Editor to help create storytelling-led, visually ...
Video Editor
Manhattan, NY · On-site
We are proud to offer a hybrid work model to support the work-life balance of our teams, and we are ... About the Role We're looking for a motivated Video Editor to help create storytelling-led, visually ...
Video Editor
Manhattan, NY · On-site
We are proud to offer a hybrid work model to support the work-life balance of our teams, and we are ... About the Role We're looking for a motivated Video Editor to help create storytelling-led, visually ...
AI Video Generation Architect
Washington, DC · Remote
$170K - $200K/yr
Run parallelized rendering with generative video models (e.g. Veo / Kling / Seedance). Narration ... with TTS (e.g. ElevenLabs v3 / Gemini-TTS) audio tags and SSML, pronunciation lexicons for ...
Quick apply
AI Video Generation Architect
Washington, DC · Remote
$170K - $200K/yr
Run parallelized rendering with generative video models (e.g. Veo / Kling / Seedance). Narration ... with TTS (e.g. ElevenLabs v3 / Gemini-TTS) audio tags and SSML, pronunciation lexicons for ...
AI Video Generation Architect
Washington, DC · On-site +1
$170K - $200K/yr
Run parallelized rendering with generative video models (e.g. Veo / Kling / Seedance). Narration ... with TTS (e.g. ElevenLabs v3 / Gemini-TTS) audio tags and SSML, pronunciation lexicons for ...
AI Video Generation Architect
Washington, DC · On-site +1
$170K - $200K/yr
Run parallelized rendering with generative video models (e.g. Veo / Kling / Seedance). Narration ... with TTS (e.g. ElevenLabs v3 / Gemini-TTS) audio tags and SSML, pronunciation lexicons for ...
Machine Learning Video Engineer
$150K - $277K/yr
You'll be responsible for designing sophisticated model architectures, fine-tuning performance ... video processing / computer vision. Strong fundamentals in Computer architecture Good written and ...
Machine Learning Video Engineer
$150K - $277K/yr
You'll be responsible for designing sophisticated model architectures, fine-tuning performance ... video processing / computer vision. Strong fundamentals in Computer architecture Good written and ...
Video Editor
$85K - $100K/yr
... models and forging inventive partnerships that extend our impact. About The Role WaitWhat is ... You'll edit across the full range of our video output: long-form podcast episodes, short-form ...
Video Editor
$85K - $100K/yr
... models and forging inventive partnerships that extend our impact. About The Role WaitWhat is ... You'll edit across the full range of our video output: long-form podcast episodes, short-form ...
Video Editor
OR · Remote
$85K - $100K/yr
... models and forging inventive partnerships that extend our impact. About The Role WaitWhat is ... You'll edit across the full range of our video output: long-form podcast episodes, short-form ...
Quick apply
Video Editor
OR · Remote
$85K - $100K/yr
... models and forging inventive partnerships that extend our impact. About The Role WaitWhat is ... You'll edit across the full range of our video output: long-form podcast episodes, short-form ...
You'll continuously refine how we work to support shifting growth priorities, new channels, and emerging production models. You bring a strong video production background and deep experience leading ...
You'll continuously refine how we work to support shifting growth priorities, new channels, and emerging production models. You bring a strong video production background and deep experience leading ...
Senior Research Engineer - Video Agents
San Francisco, CA · On-site +1
$123K - $169K/yr
By developing and shipping cutting-edge video generation models, you'll help redefine how hundreds of millions of people create and tell stories visually, shaping one of the most exciting product ...
Senior Research Engineer - Video Agents
San Francisco, CA · On-site +1
$123K - $169K/yr
By developing and shipping cutting-edge video generation models, you'll help redefine how hundreds of millions of people create and tell stories visually, shaping one of the most exciting product ...
Lead Video Producer
San Francisco, CA · Remote
Champion, role model, and embed Samsara's cultural principles in every production. * Manage ... Source and manage freelance Video Production support as needed for shoots and video projects.
Lead Video Producer
San Francisco, CA · Remote
Champion, role model, and embed Samsara's cultural principles in every production. * Manage ... Source and manage freelance Video Production support as needed for shoots and video projects.
Senior Vision Language Model Engineer
Santa Clara, CA · On-site
$122K - $168K/yr
Build, curate, and maintain high-quality multimodal datasets (e.g., video, sensor, language/action ... Collaborate with research, model development, performance, and product teams. * Contribute to ...
Senior Vision Language Model Engineer
Santa Clara, CA · On-site
$122K - $168K/yr
Build, curate, and maintain high-quality multimodal datasets (e.g., video, sensor, language/action ... Collaborate with research, model development, performance, and product teams. * Contribute to ...
Senior Vision Language Model Engineer
$122K - $168K/yr
Build, curate, and maintain highquality multimodal datasets (e.g., video, sensor, language/action ... Collaborate with research, model development, performance, and product teams. * Contribute to ...
Senior Vision Language Model Engineer
$122K - $168K/yr
Build, curate, and maintain highquality multimodal datasets (e.g., video, sensor, language/action ... Collaborate with research, model development, performance, and product teams. * Contribute to ...
Video Model information
See salary details
$14.18 - $18.36
5% of jobs
$18.36 - $22.53
16% of jobs
$23.37 is the 25th percentile. Wages below this are outliers.
$22.53 - $26.70
21% of jobs
The median wage is $28.46 / hr.
$26.70 - $30.88
20% of jobs
$34.07 is the 75th percentile. Wages above this are outliers.
$30.88 - $35.05
18% of jobs
$35.05 - $39.23
8% of jobs
$39.23 - $43.40
9% of jobs
$43.40 - $47.57
1% of jobs
$47.57 - $51.75
1% of jobs
$51.75 - $55.92
0% of jobs
$55.92 - $60.10
1% of jobs
$14
$31
$60
How much do video model jobs pay per hour?
Is 25 too old to start modeling?
What are the key skills and qualifications needed to thrive as a Video Model, and why are they important?
How much do video models get paid?
Is 5 ft 7 tall enough to be a model?
What are video models?
How can I get into modeling with no experience?
What is the difference between Video Model vs Video Editor?
| Aspect | Video Model | Video Editor |
|---|---|---|
| Primary Role | Showcases products or brands through modeling in videos | Creates, edits, and produces video content |
| Required Skills | Presentation, acting, understanding of branding | Editing software proficiency, storytelling, technical skills |
| Work Environment | On-camera, photoshoots, filming locations | Editing suites, post-production studios |
| Industry Usage | Fashion, advertising, influencer marketing | Media, entertainment, marketing campaigns |
While both roles involve video content, a Video Model primarily appears on camera to promote products or brands, focusing on presentation and on-screen presence. In contrast, a Video Editor works behind the scenes to craft and refine video footage, emphasizing technical editing skills. Both roles are essential in video production but serve different functions within the industry.
What are some common challenges faced by video models during shoots, and how can they be addressed?

Nvidia rating
9.3
Based on 5 frontline employees who took The Breakroom Quiz
15th of 209 rated software companies
Job description
At NVIDIA, we're not just building the future, we're generating it! Our world model team is pushing the boundaries of multimodal AI, robotics, and world foundation models for Physical AI. We are looking for a Senior Research Manager to lead world-model evaluation and benchmarking across NVIDIA's Physical AI model portfolio. This role will build the team and research agenda for evaluating world models through closed-system evaluations, where the model under test is pluggable, and open-system evaluations, where access to model internals enables deeper diagnostics, causal analysis, and mechanistic evaluation.
This is not only about leaderboards. It is about defining what makes a world model useful for Physical AI, discovering model failures, and turning those findings into better data, training recipes, model roadmaps, and downstream systems. The team will build a closed improvement loop across model evaluation, failure discovery, data generation, post-training, and re-evaluation.
What you'll be doing:
Lead a team of Research Scientists focused on world-model evaluation, benchmarking, and diagnostics for NVIDIA Physical AI models, including world foundation models, world-action models, synthetic data generation systems, robotics, simulation, and embodied AI workflows.
Define the scientific roadmap for closed-system and open-system evaluation, including open-loop and closed-loop benchmarks, metrics, failure taxonomy, model comparison, and evaluation-to-training feedback loops.
Develop benchmarks for physical plausibility, temporal consistency, scene dynamics, object permanence, spatial reasoning, action conditioning, affordances, controllability, long-horizon coherence, SDG quality, and WAM usefulness.
Develop open-system and mechanistic evaluation methods using model internals, including representation probing, causal interventions, activation analysis, ablations, sparse autoencoders, attention and feature analysis, and circuit-style diagnostics.
Drive evaluation-to-model-improvement loops with training, post-training, data curation, simulation, robotics, SDG, WAM, and applied research teams, including failure discovery, data generation, post-training priorities, model roadmap feedback, and re-evaluation.
Publish high-quality papers, technical reports, benchmarks, and open-source evaluation artifacts while establishing rigorous standards for validity, reproducibility, dataset hygiene, leakage prevention, and model comparison.
What we need to see:
Strong research background in machine learning, computer vision, multimodal AI, robotics, world models, representation learning, model evaluation, or mechanistic interpretability.
Experience leading research teams, research programs, or cross-functional technical initiatives with measurable scientific and product impact.
Deep understanding of modern foundation models, including video models, vision-language-action models, diffusion or flow models, self-supervised learning, or world-model architectures.
Experience designing serious benchmarks, evaluation datasets, metrics, diagnostic tools, or model analysis frameworks for complex ML systems.
Familiarity with world-model evaluation and open-system analysis techniques, such as physical plausibility, temporal consistency, action conditioning, counterfactual reasoning, representation probing, activation patching, causal interventions, sparse autoencoders, or feature attribution.
PhD, or equivalent experience in Computer Science, Electrical Engineering, Robotics, Machine Learning, AI, or a related field, with
12+ overall years of relevant research or engineering experience as well as 5+ years of management experience.
Ability to work onsite at NVIDIA's Santa Clara headquarters; this is not a remote position.
Ways to stand out from the crowd:
Built influential benchmarks, evaluation suites, model diagnostics, or interpretability tools used by research or production teams.
Published in areas such as world models, video generation, physical AI, embodied AI, robotics, representation learning, mechanistic interpretability, self-supervised learning, or model evaluation.
Experience evaluating generative video models, action-conditioned world models, robotics foundation models, world-action models, synthetic data generation systems, simulation systems, or vision-language-action models.
Strong point of view on what current benchmarks miss, and excitement to build the next generation of evaluation science for Physical AI.
NVIDIA is widely considered to be one of the technology world's most desirable employers. We have some of the most forward-thinking and hardworking people on the planet working for us. If you're creative, passionate and self-motivated, we want to hear from you! NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High-Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services.
You will also be eligible for equity and benefits.
This posting is for an existing vacancy.
NVIDIA uses AI tools in its recruiting processes.
NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.About Nvidia
Sourced by ZipRecruiter
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It's a unique legacy of innovation that's fueled by great technology--and amazing people. Today, we're tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what's never been done before takes vision, innovation, and the world's best talent.
Industry
Computer and electronic product manufacturing
Company size
10,000+ Employees
Headquarters location
Santa Clara, CA, US
Year founded
1993