Voltai

60 jobs near Columbus, OH

System Architect

Palo Alto, CA · On-site

$285K/yr

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Hardware/software co-design and cross-layer optimization * Interface definition , memory hierarchy ...

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Collaborate with electrical, systems, and verification engineers to develop embedded software that ...

Voltai is developing innovative technologies that integrate AI with hardware and electronics ... Responsibilities : • 3+ Years of Software Engineering experience graduated from top-tier CS, EECS ...

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Collaborate with electrical, systems, and verification engineers to develop embedded software that ...

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Cross-functional collaboration between hardware, software, and controls * Building and maintaining ...

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Execute end-to-end customer demos, PoCs, and presentations spanning tools deployment, software ...

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Collaborate with electrical, systems, and verification engineers to develop embedded software that ...

About Voltai Voltai is the leading AI company building agentic systems and frontier foundation ... At Voltai, we are combining the world's best talent in the intersection of software and hardware.

Senior Forward Deployed Engineer (FDE)

Palo Alto, CA · On-site

$122K - $168K/yr

... Voltai is developing models and agents to evaluate, design, and interact with hardware and ... Experience in EDA, semiconductor, or hardware-adjacent software environments Our AI Bullshit Meter ...

System Architect

Palo Alto, CA · On-site

$285K/yr

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Hardware/software co-design and cross-layer optimization * Interface definition , memory hierarchy ...

Design Verification Engineer

Palo Alto, CA

$159K - $195K/yr

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Collaborate with ML and software teams to integrate AI models into existing DV environments.

Forward Deployed Engineer

New York, NY · On-site

$62.25 - $82.75/hr

... Voltai is developing models and agents to evaluate, design, and interact with hardware and ... Experience in EDA, semiconductor, or hardware-adjacent software environments Our AI Bullshit Meter ...

Senior Forward Deployed Engineer (FDE)

Palo Alto, CA · On-site

$122K - $168K/yr

... Voltai is developing models and agents to evaluate, design, and interact with hardware and ... Experience in EDA, semiconductor, or hardware-adjacent software environments Our AI Bullshit Meter ...

Senior Forward Deployed Engineer (FDE)

New York, NY · Remote

$107K - $146K/yr

... Voltai is developing models and agents to evaluate, design, and interact with hardware and ... Experience in EDA, semiconductor, or hardware-adjacent software environments Our AI Bullshit Meter ...

Senior Forward Deployed Engineer (FDE)

New York, NY · On-site

$114K - $157K/yr

... Voltai is developing models and agents to evaluate, design, and interact with hardware and ... Experience in EDA, semiconductor, or hardware-adjacent software environments Our AI Bullshit Meter ...

Design Verification Engineer

Palo Alto, CA · On-site

$159K - $195K/yr

About Voltai Voltai is developing world models, and agents to learn, evaluate, plan, experiment ... Collaborate with ML and software teams to integrate AI models into existing DV environments.

Showing results 41-60

Research Engineer - CUDA Kernel Engineering

Voltai

Palo Alto, CA • On-site

Full-time

Re-posted 18 days ago


Job description

About Voltai
Voltai is developing world models, and agents to learn, evaluate, plan, experiment, and interact with the physical world. We are starting out with understanding and building hardware; electronics systems and semiconductors where AI can design and create beyond human cognitive limits.

About the Team

Backed by Silicon Valley’s top investors, Stanford University, and CEOs/Presidents of Google, AMD, Broadcom, Marvell, etc. We are a team of previous Stanford professors, SAIL researchers, Olympiad medalists (IPhO, IOI, etc.), CTOs of Synopsys & GlobalFoundries, Head of Sales & CRO of Cadence, former US Secretary of Defense, National Security Advisor, and Senior Foreign-Policy Advisor to four US presidents.

About the Role

You will develop, integrate, and optimize state-of-the-art CUDA kernels to power AI models that accelerate semiconductor design and verification. Your work will enable large-scale model training, inference, and reinforcement learning systems that reason about circuit layouts, generate and validate RTL, and optimize chip architectures — running efficiently across thousands of GPUs.
You’ll build tools, performance benchmarks, and integration layers that push the limits of GPU utilization for compute-intensive workloads in AI-driven hardware design. Working closely with researchers and engineers, you’ll help make Voltai the world’s leading AI + semiconductor research organization. You’ll also release your kernels and tooling as contributions to the open-source AI and HPC ecosystems.

You might thrive in this role if you have experience with

  • Writing and optimizing CUDA kernels for large-scale AI workloads (attention, routing, graph-based operations, physics-inspired operators, etc.)

  • Profiling and optimizing GPU performance for custom compute or memory-bound workloads

  • Integrating custom kernels into cutting-edge training and inference frameworks (e.g., PyTorch, Megatron, vLLM, TorchTitan)

  • Working with the latest NVIDIA hardware and software stacks (Hopper, Blackwell, NVLink, NCCL, Triton)

  • Building GPU-accelerated primitives for graph reasoning, symbolic computation, or hardware simulation tasks

  • Collaborating with AI researchers and semiconductor experts to translate domain-specific workloads into high-performance GPU code