... manage distributed training jobs, model check-pointing, and inference serving at massive scale ... Deep practical knowledge of how large models are trained and deployed, including data/tensor ...
... manage distributed training jobs, model check-pointing, and inference serving at massive scale ... Deep practical knowledge of how large models are trained and deployed, including data/tensor ...
AI Platform Architect
Austin, TX · On-site
... manage distributed training jobs, model check-pointing, and inference serving at massive scale ... Deep practical knowledge of how large models are trained and deployed, including data/tensor ...
AI Platform Architect
Austin, TX · On-site
... manage distributed training jobs, model check-pointing, and inference serving at massive scale ... Deep practical knowledge of how large models are trained and deployed, including data/tensor ...
Principal Software Developer - GPU AI/HPC kernels
Austin, TX · On-site
$185K/yr
Aid management in planning, and delivering industry-leading software for current and future ... Exposure to Matrix/Tensor operations and numerical work * Software emulation to support FP ...
Principal Software Developer - GPU AI/HPC kernels
Austin, TX · On-site
$185K/yr
Aid management in planning, and delivering industry-leading software for current and future ... Exposure to Matrix/Tensor operations and numerical work * Software emulation to support FP ...
As an example, we manage catalog data imported from hundreds of retailers, and we build product and ... Tensor-based complementary recommendations, published at IEEE Big Data 2021 (Paper) Enhancing ...
As an example, we manage catalog data imported from hundreds of retailers, and we build product and ... Tensor-based complementary recommendations, published at IEEE Big Data 2021 (Paper) Enhancing ...
Senior Software Engineer - GPU Local AI Platforms
$121K - $160K/yr
Working knowledge of LLM inference internals: attention mechanisms, KV-cache management, continuous batching, quantization formats, and tensor parallelism * Container engineering expertise: multi ...
Senior Software Engineer - GPU Local AI Platforms
$121K - $160K/yr
Working knowledge of LLM inference internals: attention mechanisms, KV-cache management, continuous batching, quantization formats, and tensor parallelism * Container engineering expertise: multi ...
As an example, we manage catalog data imported from hundreds of retailers, and we build product and ... Tensor-based complementary recommendations, published at IEEE Big Data 2021 (Paper) Enhancing ...
As an example, we manage catalog data imported from hundreds of retailers, and we build product and ... Tensor-based complementary recommendations, published at IEEE Big Data 2021 (Paper) Enhancing ...
Senior ML Accelerator Engineer - GPU
$170K - $258K/yr
Manage relationships with internal customers to ensure our kernels and libraries meet real-world ... Experience with tensor core programming, CUTLASS and/or CuTe * Experience with ML model ...
Senior ML Accelerator Engineer - GPU
$170K - $258K/yr
Manage relationships with internal customers to ensure our kernels and libraries meet real-world ... Experience with tensor core programming, CUTLASS and/or CuTe * Experience with ML model ...
Senior Software Engineer - Local AI
Austin, TX · On-site
$121K - $160K/yr
... Tensor RT, llama.cpp and vLLM. * Strong analytical and problem-solving abilities, with the ability ... Outstanding written and oral communication skills enabling effective collaboration with management ...
Senior Software Engineer - Local AI
Austin, TX · On-site
$121K - $160K/yr
... Tensor RT, llama.cpp and vLLM. * Strong analytical and problem-solving abilities, with the ability ... Outstanding written and oral communication skills enabling effective collaboration with management ...
AI Model Architect
Austin, TX · On-site
Tensor parallelism, pipeline stages, tiling across HBM and on-chip SRAM - you'll architect the ... management at L1/L2/L3. You understand why 5G inference isn't just a data center problem in a ...
AI Model Architect
Austin, TX · On-site
Tensor parallelism, pipeline stages, tiling across HBM and on-chip SRAM - you'll architect the ... management at L1/L2/L3. You understand why 5G inference isn't just a data center problem in a ...
Manager Tensor information
See Austin, TX salary details
$33.2K - $46.5K
4% of jobs
$46.5K - $59.8K
5% of jobs
$72.9K is the 25th percentile. Wages below this are outliers.
$59.8K - $73.1K
16% of jobs
$73.1K - $86.4K
17% of jobs
The median wage is $93.5K / yr.
$86.4K - $99.7K
15% of jobs
$99.7K - $113K
12% of jobs
$122.2K is the 75th percentile. Wages above this are outliers.
$113K - $126.2K
9% of jobs
$126.2K - $139.5K
7% of jobs
$139.5K - $152.8K
3% of jobs
$152.8K - $166.1K
2% of jobs
$166.1K - $179.4K
9% of jobs
$33.2K
$105.7K
$179.4K
How much do manager tensor jobs pay per year?
What is a Manager Tensor?
What are the key skills and qualifications needed to thrive as a Manager Tensor, and why are they important?
What are some common challenges faced by a Manager Tensor when leading AI and machine learning teams?
What is the difference between Manager Tensor vs Data Scientist?
| Aspect | Manager Tensor | Data Scientist |
|---|---|---|
| Required Credentials | Bachelor's or Master's in Computer Science, Data Analytics, or related fields; certifications like TensorFlow Developer are common | Bachelor's or Master's in Data Science, Statistics, Computer Science; certifications like Certified Data Scientist are common |
| Work Environment | Leads teams, manages projects, collaborates with stakeholders in tech or AI-focused companies | Analyzes data, builds models, reports insights in tech, finance, healthcare industries |
| Employer & Industry Usage | Used in AI, machine learning, and tech companies for managing TensorFlow projects | Used across industries for data analysis, predictive modeling, and research |
The main difference is that a Manager Tensor oversees AI projects involving TensorFlow, focusing on team management and project delivery, while a Data Scientist primarily analyzes data and builds models. Both roles require technical knowledge, but the Manager Tensor role emphasizes leadership and project management within AI initiatives.
What cities near Austin, TX are hiring for Manager Tensor jobs?
Cities near Austin, TX with the most Manager Tensor job openings:
Other
This job post has expired today. Applications are no longer accepted.
Job description
About us
Graphcore is one of the world’s leading innovators in Artificial Intelligence compute.
It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry.
As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone.
Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation.
Job Summary
We are seeking for a visionary AI Platform Architect to design and oversee the comprehensive infrastructure stack that powers our most demanding distributed AI workloads. Moving beyond individual hardware components, this role acts as the unifying technical authority across hardware, software, compute, network, and storage. You will be responsible for architecting a cohesive, AI rack scale platform optimized for
trillion-parameter LLM training and high-throughput inference. By orchestrating everything from advanced clustering and distributed training frameworks down to the physical layer—spanning PCIe Gen 5/6 pathways, NVMe storage topologies, and RDMA fabrics—you will ensure our AI research and deployment teams have a flawless, frictionless, and extraordinarily powerful platform at their disposal.
Responsibilities and Duties
- End-to-End Platform Architecture: Define the holistic architecture for highly clustered AI environments, ensuring zero-bottleneck data flow between parallel storage systems, AI compute nodes, and ultra-high-bandwidth network fabrics.
- Workload Orchestration: Influence the strategy for AI workload scheduling and orchestration, utilizing tools like Kubernetes or Slurm to manage distributed training jobs, model check-pointing, and inference serving at massive scale.
- Full-Stack Optimization: Profile and eliminate system-level bottlenecks across the entire AI pipeline, tuning everything from deep learning frameworks (PyTorch, DeepSpeed, etc.) down to OS-level NUMA pinning and I/O scheduling.
- Hardware-Software Co-design: Work closely with software, firmware, and OS engineering to influence platform design, ensuring the software stack fully exploits underlying hardware capabilities, including complex ARM mesh interconnects (RNI, HNF, SNF) and advanced merchant silicon features.
- Silicon Influencing Strategy: Drive the 3-to-5-year technical vision for the AI platform. Collaborate closely with subject matter experts in processor, memory, storage, GPU, thermal, mechanical, BIOS, and Manageability disciplines to define requirements specifications to communicate and present to internal and external silicon teams to influence features, optimized board routing guidelines, power and thermal targets, and the correct feeds and speeds for a competitive AI platform. This will require a deep knowledge of the AI industry and significant market competitive analysis including TCO (OPEX / CAPEX) analysis of new technologies.
Candidate Profile
Essential:
- Experience: Demonstrated ability in systems engineering, cloud architecture, or HPC, hardware engineering with at least 4+ years functioning as a Lead or Principal Architect for large-scale AI or machine learning platforms.
- Distributed AI Frameworks: Deep practical knowledge of how large models are trained and deployed, including data/tensor/pipeline parallelism and the infrastructure requirements of modern LLM architectures.
- Systems Interconnects: Authoritative understanding of system-level bottlenecks and data pathways, including deep familiarity with PCIe Gen 5/6, NVMe namespaces, and RDMA (RoCEv2/InfiniBand) integration.
- Orchestration & Containerization: Experience with container orchestration platforms and infrastructure-as-code (IaC) tailored for GPU-heavy bare-metal and cloud environments.
- Cross-Domain Leadership: Exceptional ability to bridge the gap between AI researchers/data scientists and low-level hardware/CPU/memory/storage/GPU/network engineers, translating model requirements into strict infrastructure specifications. Ability to generate Platform engineering requirement specifications that can be used to guide and influence future silicon designs.
Desirable
- Rack scale GPU AI Platforms experience: Hands on experience with rack-as-a-system
- AI platforms that integrate all the latest networking, cooling, and GPU technologies currently present in the market.
- Software / Scripting experience: Working knowledge of scripting language such as Python/JSON to characterize workloads on bare metal AI compute systems to expose issues with current Neural engine silicon solutions.
About Cerebras Systems
Sourced by ZipRecruiter
Industry
Computer and peripheral equipment manufacturing
Company size
201 - 500 Employees
Headquarters location
Sunnyvale, CA, US
Year founded
2015