Focus Multimodal Foundation Models · Representation Learning · Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and ...
Quick apply
Focus Multimodal Foundation Models · Representation Learning · Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and ...
Quick apply
Focus Multimodal Foundation Models · Representation Learning · Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and ...
Focus Multimodal Foundation Models Representation Learning Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and representation ...
Focus Multimodal Foundation Models Representation Learning Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and representation ...
Focus Multimodal Foundation Models • Representation Learning • Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and ...
Focus Multimodal Foundation Models • Representation Learning • Method Innovation We are looking for strong technical builders and researchers who deeply understand foundation models and ...
New York, NY · On-site
$62K - $125K/yr
Conducting original research in multimodal learning, including model design, training, and evaluation * Developing scalable methods for aligning and integrating diverse data modalities
New York, NY · On-site
$62K - $125K/yr
Conducting original research in multimodal learning, including model design, training, and evaluation * Developing scalable methods for aligning and integrating diverse data modalities
San Francisco, CA · On-site
$250K - $295K/yr
As a Machine Learning Engineer on the Applied Science team, you will design, train, and deploy Stand's flagship AI capabilities, with a central focus on the multimodal meshing of our Stand World ...
New
San Francisco, CA · On-site
$250K - $295K/yr
As a Machine Learning Engineer on the Applied Science team, you will design, train, and deploy Stand's flagship AI capabilities, with a central focus on the multimodal meshing of our Stand World ...
New
San Jose, CA · On-site
... learning. • Improve agent capabilities such as perception, memory, decision-making, and tool use ... etc. • Experience in multimodal learning, reinforcement learning, or agent systems. • ...
San Jose, CA · On-site
... learning. • Improve agent capabilities such as perception, memory, decision-making, and tool use ... etc. • Experience in multimodal learning, reinforcement learning, or agent systems. • ...
... multimodal data ... Our approach combines representation learning and generative modeling to capture structure ...
... multimodal data ... Our approach combines representation learning and generative modeling to capture structure ...
Help train and develop multimodal learning models using advanced learning techniques including RAG, self-supervised learning, semi-supervised, and transductive learning. Requirements Desired ...
Help train and develop multimodal learning models using advanced learning techniques including RAG, self-supervised learning, semi-supervised, and transductive learning. Requirements Desired ...
Help train and develop multimodal learning models using advanced learning techniques including RAG, self-supervised learning, semi-supervised, and transductive learning. Requirements Desired ...
Help train and develop multimodal learning models using advanced learning techniques including RAG, self-supervised learning, semi-supervised, and transductive learning. Requirements Desired ...
OR · On-site +1
$91K - $124K/yr
We're looking for a Senior Staff Machine Learning Scientist to help us solve challenging problems ... Design, implement, train, and optimize large-scale vision and multimodal foundation models across ...
OR · On-site +1
$91K - $124K/yr
We're looking for a Senior Staff Machine Learning Scientist to help us solve challenging problems ... Design, implement, train, and optimize large-scale vision and multimodal foundation models across ...
Help train and develop multimodal learning models using advanced learning techniques including RAG, self-supervised learning, semi-supervised, and transductive learning. Requirements Desired ...
Quick apply
Help train and develop multimodal learning models using advanced learning techniques including RAG, self-supervised learning, semi-supervised, and transductive learning. Requirements Desired ...
$180K - $450K/yr
Experience with large-scale machine learning systems and distributed training. * Strong background ... Experience with multimodal systems (vision + text, vision + audio) or real-time AI systems is a ...
$180K - $450K/yr
Experience with large-scale machine learning systems and distributed training. * Strong background ... Experience with multimodal systems (vision + text, vision + audio) or real-time AI systems is a ...
$93K - $127K/yr
We're looking for a Senior Staff Machine Learning Scientist to help us solve challenging problems ... Design, implement, train, and optimize large-scale vision and multimodal foundation models across ...
$93K - $127K/yr
We're looking for a Senior Staff Machine Learning Scientist to help us solve challenging problems ... Design, implement, train, and optimize large-scale vision and multimodal foundation models across ...
San Jose, CA · On-site
$180K - $450K/yr
Experience with large-scale machine learning systems and distributed training. * Strong background ... Experience with multimodal systems (vision + text, vision + audio) or real-time AI systems is a ...
San Jose, CA · On-site
$180K - $450K/yr
Experience with large-scale machine learning systems and distributed training. * Strong background ... Experience with multimodal systems (vision + text, vision + audio) or real-time AI systems is a ...
Santa Clara, CA · On-site
$107K - $146K/yr
The role focuses on applied research in deep learning architectures for biological data, including the development of multimodal learning systems and digital twin systems for healthcare. ...
Santa Clara, CA · On-site
$107K - $146K/yr
The role focuses on applied research in deep learning architectures for biological data, including the development of multimodal learning systems and digital twin systems for healthcare. ...
Seattle, WA · On-site
$17 - $22.75/hr
NLP/NLU, LLMs, Reinforcement Learning, Human Feedback/HITL, Deep Learning, Speech Recognition, Conversational AI, Natural Language Modeling, Multimodal Learning. In this role, you will work alongside ...
Seattle, WA · On-site
$17 - $22.75/hr
NLP/NLU, LLMs, Reinforcement Learning, Human Feedback/HITL, Deep Learning, Speech Recognition, Conversational AI, Natural Language Modeling, Multimodal Learning. In this role, you will work alongside ...
Santa Clara, CA · On-site
$102K - $125K/yr
The role involves conceptualizing and implementing deep learning architectures for biological data, developing multimodal learning systems, and collaborating with a diverse team to advance ...
Santa Clara, CA · On-site
$102K - $125K/yr
The role involves conceptualizing and implementing deep learning architectures for biological data, developing multimodal learning systems, and collaborating with a diverse team to advance ...
Description As a Machine Learning Research Engineer, you will help design and develop models and algorithms for multimodal perception and reasoning leveraging Vision-Language Models (VLMs) and ...
Description As a Machine Learning Research Engineer, you will help design and develop models and algorithms for multimodal perception and reasoning leveraging Vision-Language Models (VLMs) and ...
Research expertise in video generation/understanding, multimodal learning, or diffusion models * Demonstrated significant industry influence in the field of AI and/or recently published research in ...
Research expertise in video generation/understanding, multimodal learning, or diffusion models * Demonstrated significant industry influence in the field of AI and/or recently published research in ...
San Mateo, CA · On-site
$119K - $163K/yr
Expertise in one or more areas: computer vision, multimodal learning, deepfake detection, facial representation, adversarial machine learning, or VLM/LLM. * Strong coding skills with proficiency in ...
San Mateo, CA · On-site
$119K - $163K/yr
Expertise in one or more areas: computer vision, multimodal learning, deepfake detection, facial representation, adversarial machine learning, or VLM/LLM. * Strong coding skills with proficiency in ...
$21K - $29.5K
10% of jobs
$29.5K - $38K
14% of jobs
$39.2K is the 25th percentile. Wages below this are outliers.
$38K - $46.5K
10% of jobs
$46.5K - $55K
12% of jobs
The median wage is $57K / yr.
$55K - $63.5K
20% of jobs
$68.8K is the 75th percentile. Wages above this are outliers.
$63.5K - $72K
15% of jobs
$72K - $80.5K
4% of jobs
$80.5K - $89K
2% of jobs
$89K - $97.5K
4% of jobs
$97.5K - $106K
0% of jobs
$106K - $114.5K
9% of jobs
$21K
$61.7K
$114.5K
| Aspect | Multimodal Learning | Data Scientist |
|---|---|---|
| Required Credentials | Advanced degrees in AI, Machine Learning, or Computer Science | Bachelor's or Master's in Data Science, Statistics, or related fields |
| Work Environment | Research labs, AI development teams, academia | Business, tech companies, analytics teams |
| Industry Usage | AI research, multimedia applications, robotics | Data analysis, predictive modeling, business insights |
Multimodal Learning focuses on developing AI models that process and integrate multiple data types like images, text, and audio. Data Scientists analyze data to extract insights, build models, and support decision-making. While both roles involve data and algorithms, Multimodal Learning is specialized in AI model development for complex data integration, whereas Data Scientists work broadly across data analysis and interpretation.

Full-time
Re-posted 8 days ago
Focus
Multimodal Foundation Models · Representation Learning · Method Innovation
We are looking for strong technical builders and researchers who deeply understand foundation models and representation learning beyond simply applying existing frameworks.
Ideal candidates should have:
We value people who can bridge research and production, and who care about robustness, scalability, efficiency, and practical deployment in large-scale autonomous driving systems.
Responsibilities
1. Large-Scale Foundation Model Pretraining
2. Representation Learning & Method Innovation
3. Efficient Foundation Models & Scalable Deployment
Requirements