1

Embedded Machine Learning Internship Jobs in Sunnyvale, CA

We are seeking a Machine Learning Engineer to join our team developing machine learning solutions ... Work with print software and embedded teams to integrate validated models into production code ...

Machine Learning Engineer

Fremont, CA · On-site

$150K - $220K/yr

We are seeking a Machine Learning Engineer to join our team developing machine learning solutions ... Work with print software and embedded teams to integrate validated models into production code ...

We are seeking a Machine Learning Engineer to join our team developing machine learning solutions ... Work with print software and embedded teams to integrate validated models into production code ...

next page

Showing results 1-20

Embedded Machine Learning Internship information

See Sunnyvale, CA salary details

$29.9K

$50K

$103.3K

How much do embedded machine learning internship jobs pay per year?

As of Aug 5, 2026, the average yearly pay for embedded machine learning internship in Sunnyvale, CA is $49,978.00, according to ZipRecruiter salary data. Most workers in this role earn between $38,100.00 and $54,000.00 per year, depending on experience, location, and employer.

What is an embedded machine learning internship?

An Embedded Machine Learning Internship is a temporary position designed for students or recent graduates to gain hands-on experience in developing and deploying machine learning algorithms on embedded systems. These internships typically involve working with hardware such as microcontrollers, sensors, or edge devices, and using specialized tools to optimize machine learning models for low-power and resource-constrained environments. Interns collaborate with engineers and data scientists to create efficient, real-world AI solutions that run directly on devices rather than relying on cloud computing. This role helps bridge the gap between theoretical machine learning concepts and practical implementation on embedded platforms.

What are some typical projects or tasks I might work on during an embedded machine learning internship?

During an Embedded Machine Learning Internship, you can expect to work on projects such as optimizing machine learning models to run efficiently on hardware with limited resources, integrating AI algorithms into embedded systems (like microcontrollers or IoT devices), and performing real-time data processing. You'll likely collaborate closely with software engineers and hardware designers to test models on physical devices, debug performance issues, and contribute to documentation. These experiences provide practical exposure to the challenges of deploying AI in real-world, resource-constrained environments and help build skills valuable for a future career in embedded AI.

What are the key skills and qualifications needed to thrive as an embedded machine learning intern, and why are they important?

To thrive as an Embedded Machine Learning Intern, you need a background in computer science, electrical engineering, or a related field with strong programming skills in C/C++ and Python, as well as foundational knowledge of machine learning algorithms. Experience with embedded systems development tools (such as ARM Cortex, Raspberry Pi, or Arduino), version control systems, and familiarity with ML frameworks like TensorFlow Lite or Edge Impulse is often required. Analytical thinking, problem-solving ability, and effective teamwork are vital soft skills for success in this role. These skills and qualities are crucial for efficiently developing, optimizing, and deploying machine learning solutions on resource-constrained embedded platforms.
What are popular job titles related to Embedded Machine Learning Internship jobs in Sunnyvale, CA? For Embedded Machine Learning Internship jobs in Sunnyvale, CA, the most frequently searched job titles are:
What job categories do people searching Embedded Machine Learning Internship jobs in Sunnyvale, CA look for? The top searched job categories for Embedded Machine Learning Internship jobs in Sunnyvale, CA are:
What cities near Sunnyvale, CA are hiring for Embedded Machine Learning Internship jobs? Cities near Sunnyvale, CA with the most Embedded Machine Learning Internship job openings:
Infographic showing various Embedded Machine Learning Internship job openings in Sunnyvale, CA as of July 2026, with employment types broken down into 92% Full Time, 5% Part Time, and 3% Contract. Highlights an 86% Physical, 5% Hybrid, and 9% Remote job distribution, with an average salary of $49,978 per year, or $24 per hour.

Sr. Embedded Machine Learning Engineer

Allen Control Systems

Mountain View, CA • On-site

$195K - $261K/yr

Full-time

Medical, Dental, Vision, PTO

Posted 3 days ago

New


Job description

Company Overview

Allen Control Systems (ACS) is a cutting-edge defense startup founded by two former Navy electrical engineers with a proven track record in robotics and software. We are developing an autonomous gun turret using advanced computer vision and control systems to precisely detect, track, and neutralize enemy drones.

With an engineering-first culture, ACS values technical excellence and innovation. Backed by our founders’ successful exits from two previous ventures acquired for a combined $180M in 2022, we are committed to ensuring that the groundbreaking technologies we develop will have a real-world impact.

About The Role

We are looking for a Senior Embedded Machine Learning Engineer to own the end-to-end process of taking trained ML models and deploying them efficiently onto resource-constrained edge hardware. This role sits at the intersection of machine learning, embedded systems, and hardware engineering. You will integrate, convert, and optimize models to run within strict constraints on latency, memory, power, and thermal budget, and build the supporting C++ infrastructure that hosts them on device. You will partner closely with the CVML team who build the models, the embedded and firmware teams who own the device, and the product team who define performance targets. Success means models that are not just accurate in the lab but fast, small, and dependable in the field.

What You’ll Do

  • Apply quantization, pruning, knowledge distillation, operator fusion, and graph optimization to shrink models and reduce inference cost while protecting accuracy; convert trained models into edge-deployable formats using ONNX and TensorRT.

  • Profile inference on target accelerators including GPUs, NPUs, DSPs, and FPGAs; measure latency, throughput, memory footprint, and power consumption, then drive the changes needed to hit performance targets.

  • Design, write, and maintain the C++ application code that hosts inference on device, including pre- and post-processing pipelines, data and memory management, threading, and interfaces to the rest of the embedded system; ensure the combined model and C++ stack meets real-time constraints and fits within device memory budget.

  • Build test harnesses to verify on-device accuracy against reference results and catch regressions from optimization or quantization; contribute to tooling for packaging, versioning, and delivering model updates to deployed devices.

  • Set best practices for edge deployment, review designs and code, and mentor other engineers on optimization and embedded ML techniques; work closely with research, firmware, and product teams to set realistic performance targets and feed hardware constraints back into model design.

What You’ll Need

  • 10+ years of professional software or systems engineering experience, including at least 2 years focused on deploying ML models to embedded or edge devices; Bachelor’s or Master’s degree in Computer Science, Electrical Engineering, Computer Engineering, or equivalent practical experience.

  • Very strong C++ proficiency; working knowledge of CUDA; hands-on experience with PyTorch and at least one edge inference runtime such as TensorFlow Lite, ONNX Runtime, or TensorRT.

  • Practical experience with model optimization techniques including post-training quantization, quantization-aware training, pruning, and distillation; demonstrated ability to profile and optimize for latency, memory, and power on constrained hardware.

  • Working knowledge of embedded or edge platforms such as NVIDIA Jetson, Qualcomm, ARM Cortex, or comparable NPUs and SoCs, and of Linux or an RTOS; solid grasp of computer architecture concepts relevant to inference including memory hierarchy, fixed-point arithmetic, and accelerator offload; domain experience in computer vision or sensor processing on device.

You’ll Stand Out

  • Hands-on experience deploying computer vision models for detection or tracking tasks on embedded or edge hardware.

  • Experience with NVIDIA Jetson specifically, including TensorRT optimization and deployment on Jetson platforms.

  • Background in defense, autonomous systems, or robotics where real-time reliability matters.

  • Experience building or contributing to model update and OTA delivery pipelines for deployed edge devices.

What We Offer

  • Competitive salary

  • ACS Equity Package

  • Health, Dental, Vision Insurance

  • Paid Time Off

Allen Control Systems is an Equal Opportunity Employer, providing equal employment opportunities to all employees and applicants for employment. Allen Control Systems prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws. #LI-AS1

 

Compensation Range: $195K - $261K