1

Intern Computer Vision Deep Learning Engineer Jobs in Santa Clara, CA

In this position, you will have the opportunity to be part of our extraordinary team of Computer Graphics, Computer Vision and Deep Learning researchers and engineers to discover and build solutions ...

In this position, you will be part of our extraordinary team of Computer Graphics, Computer Vision and Deep Learning researchers and engineers to discover and build solutions to previously-unsolved ...

... understanding, computer vision, deep learning, language understanding and content ranking ... Proficient in one or more programming languages such as Python, Java and C * Familiar with one or ...

... understanding, computer vision, deep learning, language understanding and content ranking ... Proficient in one or more programming languages such as Python, Java and C * Familiar with one or ...

Machine Learning Engineer

San Jose, CA · On-site

$150 - $200/hr

... understanding, computer vision, deep learning, language understanding and content ranking ... Proficient in one or more programming languages such as Python, Java and C. Familiar with one or ...

The Video Engineering group at Apple is responsible for creating the image/video core technologies ... deep learning techniques for video processing / computer vision. Strong fundamentals in Computer ...

Showing results 41-60

Intern Computer Vision Deep Learning Engineer information

See Santa Clara, CA salary details

$10

$20

$28

How much do intern computer vision deep learning engineer jobs pay per hour?

As of Sep 8, 2026, the average hourly pay for intern computer vision deep learning engineer in Santa Clara, CA is $20.01, according to ZipRecruiter salary data. Most workers in this role earn between $16.92 and $22.60 per hour, depending on experience, location, and employer.

What does an intern computer vision deep learning engineer do?

An Intern Computer Vision Deep Learning Engineer assists in developing and improving algorithms that enable computers to interpret and understand visual information from the world, such as images and videos. They often work on tasks like image classification, object detection, and facial recognition using deep learning frameworks like TensorFlow or PyTorch. Interns typically help with data collection, model training, evaluation, and sometimes deployment, all under the guidance of experienced team members. This role is a great opportunity to gain hands-on experience in machine learning and computer vision while contributing to real-world projects.

What are the key skills and qualifications needed to thrive as an intern computer vision deep learning engineer?

To thrive as an Intern Computer Vision Deep Learning Engineer, you need a solid understanding of machine learning fundamentals, computer vision concepts, and proficiency in programming languages like Python, often supported by coursework or personal projects. Familiarity with deep learning frameworks such as TensorFlow or PyTorch and experience with image processing libraries like OpenCV are typically expected. Strong problem-solving abilities, curiosity, and effective teamwork skills help interns excel in fast-paced research and development environments. These skills are essential for contributing to innovative projects and adapting to the rapidly evolving field of computer vision.

What types of projects or tasks can I expect to work on as an intern computer vision deep learning engineer?

As an Intern Computer Vision Deep Learning Engineer, you can expect to contribute to projects involving image or video analysis, such as object detection, image classification, or facial recognition. Your daily tasks might include data preprocessing, annotating datasets, training and evaluating deep learning models, and assisting with model optimization for deployment. You’ll often work closely with senior engineers and researchers, gaining hands-on experience with real-world datasets and cutting-edge frameworks. Collaboration with cross-functional teams, such as software developers and product managers, is common to ensure your models address practical business needs.

What is the difference between Intern Computer Vision Deep Learning Engineer vs Intern Machine Learning Engineer?

AspectIntern Computer Vision Deep Learning EngineerIntern Machine Learning Engineer
Required SkillsComputer vision, deep learning, CNNs, Python, TensorFlow/PyTorchMachine learning, algorithms, Python, scikit-learn, TensorFlow/PyTorch
Work EnvironmentResearch labs, tech companies, startups focusing on image/video analysisTech companies, research labs, startups working on diverse ML applications
Industry UsagePrimarily in computer vision projects like object detection, image segmentationBroader ML projects including predictive modeling, NLP, recommendation systems

Intern Computer Vision Deep Learning Engineers focus on image and video analysis using deep learning techniques, while Intern Machine Learning Engineers work on a wider range of ML applications. Both roles require strong Python skills and familiarity with deep learning frameworks, but their project focus and industry applications differ.

What are the most commonly searched types of Computer Vision Deep Learning Engineer jobs in Santa Clara, CA?

The most popular types of Computer Vision Deep Learning Engineer jobs in Santa Clara, CA are:

What are popular job titles related to Intern Computer Vision Deep Learning Engineer jobs in Santa Clara, CA?

For Intern Computer Vision Deep Learning Engineer jobs in Santa Clara, CA, the most frequently searched job titles are:

What cities near Santa Clara, CA are hiring for Intern Computer Vision Deep Learning Engineer jobs?

Cities near Santa Clara, CA with the most Intern Computer Vision Deep Learning Engineer job openings:

Infographic showing various Intern Computer Vision Deep Learning Engineer job openings in Santa Clara, CA as of July 2026, with employment types broken down into 1% As Needed, 79% Full Time, 15% Part Time, and 5% Contract. Highlights an 96% Physical, 1% Hybrid, and 3% Remote job distribution, with an average salary of $41,617 per year, or $20 per hour.

Senior Machine Learning Engineer, Perception

PlusAI

Santa Clara, CA • On-site

$150K - $250K/yr

Full-time

Retirement

Re-posted yesterday


Job description

PlusAI is a Physical AI company pioneering AI-based virtual driver software for factory-built autonomous trucks. Headquartered in Silicon Valley with operations in the United States and Europe, Plus was named by Fast Company as one of the World's Most Innovative Companies. Partners including TRATON GROUP's Scania, MAN, and International brands, Hyundai Motor Company, Iveco Group, Bosch, and DSV are working with Plus to accelerate the deployment of next-generation autonomous trucks. If you're ready to make a huge impact and drive the future of autonomy, Plus is looking for talented individuals to join its fast-growing teams.
We are seeking a highly skilled Machine Learning Engineer with deep expertise in developing Bird's Eye View (BEV) fusion models using multimodal sensor inputs, particularly LiDAR. You will play a central role in designing scalable perception algorithms that integrate data from camera, LiDAR, and radar sensors to support autonomous driving and 3D scene understanding.
Responsibilities:
  • Design, implement, and optimize BEV-based perception models that fuse camera, LiDAR, and radar inputs.
  • Benchmark perception models using large-scale datasets and well-defined quantitative metrics.
  • Collaborate cross-functionally with research, data, and deployment engineers to refine models and support real-world applications.
  • Maintain a strong focus on performance, robustness, and scalability for deployment in production systems.
  • Ensure that your work is performed in accordance with the company's Quality Management System (QMS) requirements and contribute to continuous improvement efforts.
  • Ensure team compliance with QMS, monitor quality, and drive process improvements.

Required Skills:
  • Ph.D. or Masters in AI, Computer Science, Electrical Engineering, Robotics, or a related field.
  • Ph.D. new grad or Masters + 3 years industry experience
  • Proficiency in Python and experience building deep learning pipelines.
  • Strong expertise in PyTorch, TensorFlow, or JAX.
  • Proven experience with LiDAR-based 3D perception and BEV representation models
  • Deep understanding of multimodal sensor fusion architectures and techniques.
  • Familiarity with camera, LiDAR, and radar modalities and their synchronization, calibration, and integration in perception pipelines.
  • Solid foundation in computer vision, deep learning, and 3D geometry.

Preferred Skills:
  • Industry or academic experience in autonomous vehicle perception, robotics, or related areas.
  • Hands-on experience developing deep learning models in real-world or production environments.
  • Experience with distributed training, high-performance computing, or GPU acceleration.

$150,000 - $250,000 a year
Our compensations (cash and equity) are determined based on the position, your location, qualifications, and experience.
Your opportunities joining PlusAI
Work, learn and grow in a highly future-oriented, innovative and dynamic field.
Wide range of opportunities for personal and professional development.
Catered free lunch, unlimited snacks and beverages.
Highly competitive salary and benefits package, including 401(k) plan.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.