1

Computer Vision Engineer Intern Jobs in Atlanta, GA

Your Part in this Growth Story As a Solution Engineer Intern at Stonebranch, youll play an active ... Bachelors degree or higher in Computer Science, Information Technology, or a related field OR ...

Your Part in this Growth Story As a Solution Engineer Intern at Stonebranch, you'll play an active ... Bachelor's degree or higher in Computer Science, Information Technology, or a related field OR ...

The Field Engineer Intern will assist superintendents with daily duties including: * Assisting with ... Medical, Dental and Vision Benefits * Retirement and Savings Benefits * Flexible Spending and ...

New

Intern will work under the guidance of a senior tech lead. Internship at Hoptek could include a ... Enrolled in a BS/MS program in Industrial Engr, Computer Science or related majors * Strong ...

next page

Showing results 1-20

Computer Vision Engineer Intern information

See Atlanta, GA salary details

$12

$24

$37

How much do computer vision engineer intern jobs pay per hour?

As of Jul 30, 2026, the average hourly pay for computer vision engineer intern in Atlanta, GA is $24.44, according to ZipRecruiter salary data. Most workers in this role earn between $19.90 and $27.74 per hour, depending on experience, location, and employer.

What does a Computer Vision Engineer Intern do?

A Computer Vision Engineer Intern assists in developing and implementing algorithms that enable computers to process and interpret visual information from the world, such as images or videos. They often work on projects involving object detection, image segmentation, and other tasks related to image analysis and machine learning. Interns usually collaborate with experienced engineers, contribute to codebases, and help test and optimize models for various applications. Their work supports industries like robotics, healthcare, automotive, and more.

What types of projects and tasks can a Computer Vision Engineer Intern typically expect to work on during their internship?

As a Computer Vision Engineer Intern, you can expect to be involved in a variety of hands-on projects such as developing image recognition algorithms, annotating datasets, and testing computer vision models for accuracy and performance. Interns often collaborate closely with senior engineers and data scientists, contributing to tasks like data preprocessing, model training, and performance benchmarking. This role offers a great opportunity to gain practical experience with popular frameworks such as OpenCV and TensorFlow, and to develop skills in both research and applied development within interdisciplinary teams.

What are the key skills and qualifications needed to thrive as a Computer Vision Engineer Intern, and why are they important?

To thrive as a Computer Vision Engineer Intern, you need a solid background in computer science, mathematics, and image processing, typically supported by coursework or experience in machine learning and programming (Python, C++). Familiarity with frameworks like OpenCV, TensorFlow, or PyTorch, and experience using annotation tools or version control systems is highly valuable. Strong problem-solving, analytical thinking, and the ability to collaborate effectively help interns stand out. These skills and qualities are crucial for developing, testing, and optimizing computer vision models in a fast-paced, team-oriented environment.
What are the most commonly searched types of Computer Vision Engineer jobs in Atlanta, GA? The most popular types of Computer Vision Engineer jobs in Atlanta, GA are:
What are popular job titles related to Computer Vision Engineer Intern jobs in Atlanta, GA? For Computer Vision Engineer Intern jobs in Atlanta, GA, the most frequently searched job titles are:
Infographic showing various Computer Vision Engineer Intern job openings in Atlanta, GA as of July 2026, with employment types broken down into 1% As Needed, 81% Full Time, 14% Part Time, and 4% Contract. Highlights an 93% Physical, 2% Hybrid, and 5% Remote job distribution, with an average salary of $50,840 per year, or $24.4 per hour.

Senior Computer Vision Engineer (Egocentric), Data Foundry

Stord

Atlanta, GA

$101K - $138K/yr

Full-time

Posted 5 days ago


Stord rating

3.1

Company rating: 3.1 out of 10

Based on 7 frontline employees who took The Breakroom Quiz


Job description

hackajob is collaborating with Stord to connect them with exceptional professionals for this role.

Stord is The Consumer Experience Company, powering seamless checkout through delivery for today's leading brands. Stord is rapidly growing and is on track to double our revenue in the next 18 months. To meet and exceed this target, Stord is strategically scaling teams across the entire company, and seeking energetic experts to help us achieve our mission.

By combining comprehensive commerce-enablement technology with high-volume fulfillment services, Stord provides brands a platform to compete with retail giants. Stord manages over $10 billion of commerce annually through its fulfillment, warehousing, transportation, and operator-built software suite including OMS, Pre- and Post-Purchase, and WMS platforms. Stord is leveling the playing field for all brands to deliver the best consumer experience at scale.

With Stord, brands can increase cart conversion, improve unit economics, and drive sustained customer loyalty. Stord’s end-to-end commerce solutions combine best-in-class omnichannel fulfillment and shipping with leading technology to ensure fast shipping, reliable delivery promises, easy access to more channels, and improved margins on every order.

Hundreds of leading DTC and B2B companies like AG1, True Classic, Native, Seed Health, quip, goodr, Sundays for Dogs, and more trust Stord to deliver industry-leading consumer experiences on every order. Stord is headquartered in Atlanta with facilities across the United States, Canada, and Europe. Stord is backed by top-tier investors including Kleiner Perkins, Franklin Templeton, Founders Fund, Strike Capital, Baillie Gifford, and Salesforce Ventures.

Build the Vision Systems Powering the Future of Physical AI.

Stord operates the largest independent e-commerce fulfillment network in the U.S. — with 20+ fulfillment centers, 4,000+ warehouse associates, and nearly 100 million packages shipped annually. We are transforming this operational infrastructure into one of the most valuable sources of training data for the next generation of physical AI.

We are building a new business line at the intersection of robotics, computer vision, and AI data — and we are looking for an experienced Computer Vision Engineer to help build the technical foundation from the ground up.

This is a hands-on builder role for a technical leader who can design, prototype, and productionize perception systems that transform real-world environments into high-quality AI training data.

Why This Role:

This is a rare opportunity to build the technical foundation of a new AI business from the ground up — combining real-world operational infrastructure with cutting-edge computer vision and robotics.

You will have:

  • A structural advantage no startup can easily replicate — access to one of the largest real-world environments for collecting physical AI training data.

  • Direct exposure to the fastest-growing AI market — partnering with robotics companies, AI labs, and teams building the future of intelligent systems.

  • True technical ownership — the opportunity to define architecture, build foundational systems, and shape the future of Embodied AI data.

  • Executive partnership — working closely with Stord’s CTO and Co-Founder to define strategy, accelerate execution, and remove barriers.

What You Will Own:

You will own the early computer vision and egocentric perception stack — including data capture systems, vision pipelines, model development, and the infrastructure required to deliver high-quality datasets at scale.

Working closely with a small, highly technical team, you will help define the architecture, build the systems, and establish the technical standards for Stord’s Embodied AI data platform.

Build the Data Product & Capture Platform

  • Define and evolve Stord’s Embodied AI data products across quality tiers — from RGB egocentric video to depth-enhanced and multimodal datasets with hand pose, body pose, and rich annotations.

  • Determine the right technical investments based on customer requirements and the needs of emerging robotics and AI models.

  • Establish data quality standards and evaluation frameworks to ensure datasets meet production-level requirements.

Build the Perception Stack

  • Design and develop perception systems including:

    • Object detection, tracking, and segmentation

    • Depth estimation and 3D reconstruction

    • 6DoF pose estimation

    • Multi-view 3D hand and body pose estimation

    • Egocentric and fixed-camera perception systems

  • Build robust solutions designed for complex, real-world environments — not just benchmark datasets.

Own the Hardware + Vision Integration

  • Design and deploy camera systems and perception rigs across warehouse environments.

  • Own camera calibration, multi-camera synchronization, epipolar geometry, and 3D reconstruction workflows.

  • Develop solutions for deriving accurate spatial understanding from multimodal sensor inputs and video data.

Build Automated Labeling & Data Pipelines

  • Develop VLM-assisted and automated annotation workflows with human-in-the-loop quality systems.

  • Integrate labeling tools and processes that improve scalability while maintaining dataset accuracy.

  • Build pipelines that transform raw video into production-ready training datasets.

Train, Optimize, and Deploy Models

  • Design, fine-tune, evaluate, and optimize computer vision and multimodal models on large-scale video datasets.

  • Build reproducible training and deployment workflows that move beyond experimentation and into production.

  • Establish evaluation methodologies that measure model performance, reliability, and quality.

What You'll Need:

  • 8+ years of experience building and shipping production computer vision or perception systems (or an MS/PhD in Computer Vision, Machine Learning, Robotics, or a related field with 6+ years of hands-on industry experience).

  • Experience building perception systems that operate on real-world, imperfect data — beyond academic benchmarks.

  • Demonstrated experience developing or scaling egocentric vision, robotics perception, autonomous systems, or AI data platforms.

  • Deep expertise in computer vision fundamentals and modern tooling, including:

    • CNNs and vision transformers

    • Object detection, segmentation, and tracking

    • Depth estimation and 3D vision

    • 2D/3D pose estimation

    • Camera calibration and geometric computer vision

  • Experience owning complex perception problems end-to-end — from data collection and model development through evaluation, optimization, and deployment.

  • Strong understanding of large-scale video and multimodal datasets, including data quality measurement and evaluation methodologies.

  • Experience taking ambiguous, 0→1 technical challenges and turning them into scalable systems with limited resources.

  • Ability to influence technical direction, establish engineering standards, and mentor other engineers.

  • Expert-level Python skills and strong software engineering fundamentals; C++ experience where performance requirements demand it.


What Stord employees say

Pay

Hours and flexibility

Workplace

Get the full story on Breakroom