1

Internship Ai Infrastructure Engineer Jobs (NOW HIRING)

AI Infrastructure Engineer

Dallas, TX

$106K - $139K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Grand Rapids, MI

$103K - $135K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Charlotte, NC

$105K - $137K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Sacramento, CA

$114K - $150K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Denver, CO

$110K - $145K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

South Bend, IN

$105K - $138K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Nashville, TN

$103K - $136K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Indianapolis, IN

$102K - $134K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

New York, NY

$117K - $154K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Cleveland, OH

$104K - $136K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Chicago, IL · On-site

$110K - $145K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Houston, TX

$102K - $134K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

San Francisco, CA

$126K - $166K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Los Angeles, CA

$115K - $151K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Atlanta, GA

$103K - $135K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Miami, FL

$102K - $134K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

AI Infrastructure Engineer

Tampa, FL

$101K - $133K/yr

As an AI Infrastructure Engineer, you are a deep technical contributor with subsystem ownership and notable technical influence. Your role will be crucial in ensuring that the AI solutions you ...

next page

Showing results 1-20

Internship Ai Infrastructure Engineer information

See salary details

$11

$19

$29

How much do internship ai infrastructure engineer jobs pay per hour?

As of Jun 30, 2026, the average hourly pay for internship ai infrastructure engineer in the United States is $19.31, according to ZipRecruiter salary data. Most workers in this role earn between $16.11 and $20.91 per hour, depending on experience, location, and employer.

What is the difference between Internship Ai Infrastructure Engineer vs Data Engineer?

AspectInternship Ai Infrastructure EngineerData Engineer
Required CredentialsEnrolled in or recent graduate of Computer Science, Engineering, or related fields; some knowledge of AI and infrastructure toolsBachelor's or higher in Computer Science, Data Science, or related; experience with databases and data pipelines
Work EnvironmentInternship setting, collaborative teams, learning-focusedFull-time, technical teams managing data systems and pipelines
Employer & Industry UsageTech companies, AI startups, research labsTech firms, finance, healthcare, and other data-driven industries

The Internship Ai Infrastructure Engineer role focuses on supporting AI infrastructure projects during an internship, emphasizing learning and assisting with AI systems setup. In contrast, Data Engineers build and maintain data pipelines and infrastructure for data analysis. While both roles require knowledge of technical tools, the internship role is more entry-level and learning-oriented, whereas Data Engineers are more experienced and responsible for ongoing data management.

What does an Internship AI Infrastructure Engineer do?

An Internship AI Infrastructure Engineer assists in designing, developing, and maintaining the foundational systems that support artificial intelligence (AI) applications. They work with cloud platforms, data pipelines, and scalable computing resources to ensure that AI models can be trained and deployed efficiently. Interns may help automate workflows, optimize performance, and collaborate with data scientists and software engineers. The role provides hands-on experience with the tools and frameworks commonly used in AI engineering environments.

What are the key skills and qualifications needed to thrive as an Internship AI Infrastructure Engineer, and why are they important?

To thrive as an Internship AI Infrastructure Engineer, you need a solid understanding of computer science fundamentals, programming (especially in Python or C++), and basic knowledge of machine learning frameworks, often supported by ongoing studies in a relevant field. Familiarity with cloud platforms (like AWS, GCP, or Azure), version control systems (such as Git), and containerization tools (Docker, Kubernetes) is typically expected. Strong problem-solving abilities, curiosity, teamwork, and effective communication help interns stand out and integrate quickly into engineering teams. These skills are crucial for supporting scalable AI solutions, collaborating on complex projects, and contributing meaningfully in a fast-evolving technical environment.

What types of projects and responsibilities can an AI Infrastructure Engineer intern expect to work on?

As an AI Infrastructure Engineer intern, you can expect to be involved in projects that support the development, deployment, and scaling of AI models. Typical responsibilities may include optimizing data pipelines, maintaining and improving cloud or on-premise computing resources, and collaborating closely with data scientists to ensure efficient model training and inference. Interns often get hands-on experience with tools such as Docker, Kubernetes, and various cloud platforms, and work in cross-functional teams to troubleshoot and enhance AI workflows. This role provides a solid foundation in both software engineering and AI operations, preparing you for advanced positions in the field.
More about Internship Ai Infrastructure Engineer jobs
What cities are hiring for Internship Ai Infrastructure Engineer jobs? Cities with the most Internship Ai Infrastructure Engineer job openings:
What are the most commonly searched types of Ai Infrastructure Engineer jobs? The most popular types of Ai Infrastructure Engineer jobs are:
What states have the most Internship Ai Infrastructure Engineer jobs? States with the most job openings for Internship Ai Infrastructure Engineer jobs include:
What job categories do people searching Internship Ai Infrastructure Engineer jobs look for? The top searched job categories for Internship Ai Infrastructure Engineer jobs are:
Senior AI Infrastructure Engineer

Senior AI Infrastructure Engineer

Gatik AI

Mountain View, CA

$128K - $174K/yr

Other

Posted 13 days ago


Key responsibilities

  • Design, build, and scale high-performance AI infrastructure to support distributed training, experiment tracking, and model deployment.

  • Optimize and maintain multi-GPU and multi-node systems, including performance tuning and resource scheduling for large-scale machine learning workloads.

  • Develop, automate, and monitor AI infrastructure and data pipelines, ensuring system health, observability, and robust model lifecycle management.


Job description

Who we are

Gatik, the leader in autonomous middle-mile logistics, is revolutionizing the B2B supply chain with its autonomous transportation-as-a-service (ATaaS) solution and prioritizing safe, consistent deliveries while streamlining freight movement by reducing congestion. The company focuses on short-haul, B2B logistics for Fortune 500 retailers and in 2021 launched the world's first fully driverless commercial transportation service with Walmart. Gatik's Class 3-7 autonomous trucks are commercially deployed across major markets, including Texas, Arkansas, and Ontario, Canada, driving innovation in freight transportation. 

The company's proprietary Level 4 autonomous technology, Gatik Carrier, is custom-built to transport freight safely and efficiently between pick-up and drop-off locations on the middle mile. With robust capabilities in both highway and urban environments, Gatik Carrier serves as an all-encompassing solution that integrates advanced software and hardware powering the fleet, facilitating effortless integration into customers' logistics operations. 

About the role

We are seeking a Senior AI Infrastructure Engineer to design, build, and scale the high-performance AI platform powering our autonomous driving models. While researchers focus on developing perception, planning, and world models, you will be responsible for the underlying infrastructure that enables distributed training, experiment tracking, and seamless model deployment. You will bridge the gap between research and production, ensuring our AI stack is scalable, resilient, and highly efficient

This role is onsite 5 days a week at our Mountain View, CA office!

What you'll do
  • Distributed Training & ML Systems Support
    • Scale Research Workloads: Enable researchers to scale complex models (VLA, World Models) across multi-node setups using PyTorch Distributed, and Ray Train.
    • Performance Optimization: Architect and optimize multi-GPU setups, ensuring efficient model parallelism and data parallelism techniques across H100/A100 clusters.
    • Networking & Hardware Tuning: Optimize low-level communication (e.g., NCCL tuning, InfiniBand, or RoCE v2) to minimize latency for 3D Gaussian Splatting (3DGS) and large-scale training.
    • Intelligent Resource Scheduling: Optimize hardware utilization and cost-efficiency through Kubernetes-native GPU scheduling (NVIDIA GPU Operator, KubeFlow).
    • Inference Performance Engineering: Deploy and scale optimized model artifacts using TensorRT, ONNX Runtime, and Triton Inference Server, fine-tuning pipelines for both real-time and batch processing
  • Agentic Infrastructure & Automation
    • Self-Healing AI Infrastructure: Architect and deploy Autonomous AI Agents (LangGraph, CrewAI, or AutoGen) to monitor GPU cluster health, enabling automated real-time triage of hardware failures and NCCL timeouts.
    • Agentic DevOps & CI/CD: Develop agent-driven automation, such as Agentic PR Reviewers for infrastructure code and AI agents that proactively suggest model-specific Kubernetes resource optimizations.
    • Agentic Data Curation: Support researchers in building "Data Machines" where AI agents autonomously curate, label, and verify high-priority edge cases from raw data.
  • Model Management & Lifecycle (MLOps)
    • Automated Lifecycle Management: Design and maintain ML infrastructure leveraging MLFlow, Argo Workflows, and Kubernetes to automate the end-to-end model lifecycle.
    • Experiment & Model Tracking: Integrate feature stores and experiment tracking systems to provide a robust system of record for every model iteration.
    • Deployment Strategies: Implement robust serving mechanisms, including A/B testing, shadow deployments, and rollback mechanisms
  • Cloud-Native Foundations & Data Integration
    • Infrastructure as Code: Drive the "Everything as Code" philosophy using Terraform and Helm.
    • Data Pipelines: Collaborate with data teams to scale ETL pipelines using Apache Airflow, Kafka, and Spark for large-scale dataset management.
    • Integrated Data Factories: Collaborate with data engineering teams to scale high-bandwidth ETL pipelines using Apache Airflow, Kafka, and Spark, ensuring seamless data flow from raw sensor logs to optimized storage in S3, GCS, or Delta Lake
  • Monitoring & Observability
    • System Metrics: Define and track key ML system metrics, including training convergence, latency, throughput, and drift detection.
    • Infrastructure Health: Maintain deep visibility into platform health using Prometheus, Grafana, OpenTelemetry, and ELK Stack.
    • Deep Stack Observability: Develop comprehensive monitoring using Prometheus, Grafana, and OpenTelemetry to track low-level infrastructure health alongside high-level ML metrics like training convergence and throughput.
    • AI-Specific Metrics & Drift: Define and monitor critical ML system KPIs, including model latency, inference throughput, and feature drift detection
What we're looking for
  • Experience: 5+ years in ML infrastructure, MLOps, or DevOps supporting high-scale compute environments.
  • ML Expertise: Deep understanding of multi-GPU training strategies (FSDP, DeepSpeed, Ray Train) and high-performance networking (NCCL, InfiniBand).
  • Infrastructure Automation: Mastery of Kubernetes, Terraform, and Helm, with a focus on GPU-native orchestration.
  • AI Agent Frameworks: Proven experience building or supporting Agentic Workflows for infrastructure or data automation (e.g., using LLMs to drive DevOps tasks).
  • Platform Mastery: Expertise in MLFlow, Argo Workflows, and Kubernetes.
  • Containerization: Strong experience with Docker, Kubernetes, and Helm.
  • Data & CI/CD: Proficiency in Apache Airflow, Kafka, Spark, and GitOps automation.
  • Core Skills: Proficiency in Python and Bash; experience with Go or Rust is a plus
Bonus Qualifications
  • Advanced AI Protocols: Familiarity with the Model Context Protocol (MCP) to standardize how AI agents interact with internal databases and orchestration APIs.
  • Hybrid & Physical AI: Experience in hybrid cloud and on-prem GPU cluster management for Physical AI workloads (e.g., 3DGS, World Models).
  • Agentic Observability: Experience utilizing LLMs for semantic monitoring and log analysis to detect complex distributed system failures that traditional threshold-based alerts miss.

Salary Ranges - $180,000- $240,000

More about Gatik

Founded in 2017 by experts in autonomous vehicle technology, Gatik has rapidly expanded its presence to Mountain View, Dallas-Fort Worth, Arkansas, and Toronto. As the first and only company to achieve fully driverless middle-mile commercial deliveries, Gatik holds a unique and defensible position in the AV industry, with a clear trajectory toward sustainable growth and profitability.

We have delivered complete, proprietary AV technology - an integration of software and hardware - to enable earlier successes for our clients in constrained Level 4 autonomy.  By choosing the middle mile - with defined point-to-point delivery, we have simplified some of the more complex AV challenges, enabling us to achieve full autonomy ahead of competitors. Given extensive knowledge of Gatik's well-defined, fixed route ODDs and hybrid architecture, we are able to hyper-optimize our models with exponentially less data, establish gate-keeping mechanisms to maintain explainability, and ensure continued safety of the system for unmanned operations.

Visit us at Gatik for more company information and Careers at Gatik for more open roles.

Notable News
  • Bloomberg: Autonomous Trucking Firm Gatik Inks Contracts Worth $600 Million
  • Forbes: Hundreds' Of Gatik Robot Delivery Trucks Headed For U.S. Roads
  • Forbes:Gatik And Loblaw Announce Largest Commercial Deployment Of AV Trucks
  • Forbes: Forget robotaxis. Upstart Gatik sees middle-mile deliveries as the path to profitable AVs
  • Tech Brew: Gatik AI exec unpacks the regulations that could shape the AV industry
  • Business Wire: Gatik Paves the Way for Safe Driverless Operations ('Freight-Only') at Scale with Industry-First Third-Party Safety Assessment Framework
  • Auto Futures: Autonomous Trucking Group Gatik Secures Investment From NIPPON EXPRESS HOLDINGS
  • Automotive News: Gatik foresees hundreds of self-driving trucks on road soon, and that's just the beginning
  • Forbes: Isuzu And Gatik Go All In To Scale Up Driverless Freight Services
  • Bloomberg: Autonomous Vehicle Startup Takes Off by Picking Off Easier Routes
  • Reuters: Driverless vehicles on limited routes bump along despite US robotaxi scrutiny
Taking care of our team

At Gatik, we connect people of extraordinary talent and experience to an opportunity to create a more resilient supply chain and contribute to our environment's sustainability. We are diverse in our backgrounds and perspectives yet united by a bold vision and shared commitment to our values. Our culture emphasizes the importance of collaboration, respect and agility.

We at Gatik strive to create a diverse and inclusive environment where everyone feels they have opportunities to succeed and grow because we know that together we can do great things. We are committed to an inclusive and diverse team. We do not discriminate based on race, color, ethnicity, ancestry, national origin, religion, sex, gender, gender identity, gender expression, sexual orientation, age, disability, veteran status, genetic information, marital status or any legally protected status.