No spotlight call. Need resumes ASAP.
Role: Technical Support Specialist
Location: Onsite โ Milpitas, CA (MondayโFriday)
Duration: 3 months remaining, potential extension
Role Overview
This role is a hands-on, hardware focused technical support position supporting GPU/compute clusters in an AI lab/R&D environment. The emphasis is on hardware troubleshooting, Linux-based system support, and deep understanding of compute architecture, rather than software development.
Key Responsibilities
- Troubleshoot GPU/CPU servers, compute clusters, and networking (InfiniBand)
- Diagnose hardware issues (cabling, components, GPUs, servers)
- Rack/stack initially limited (systems already built), but may increase if extended
- Replace/install server components within racks
- Use Linux command line extensively for diagnostics and system validation
- Manage lab space and hardware inventory (re procurement access provided)
Must Have Skills (Non Negotiable)
- Strong hardware troubleshooting experience (servers, GPUs, compute systems)
- Solid understanding of computer/compute architecture
- Strong Linux skills for system bring up and troubleshooting
- Experience with GPUs and high performance compute environments
- Ability to independently diagnose and resolve hardware/system issues
Preferred / Nice to Have
- Prior data center or HPC/compute cluster experience (plus, not mandatory)
- Scripting experience (Bash, Python) โ expected if candidate has done similar roles
- Familiarity with GPU technologies (cutting edge R&D GPUs; Tesla, etc.)
- Candidates who've built systems themselves (gaming PCs, lab servers, small data centers)
Experience & Education
Minimum: 3โ4 years of relevant experience (not pure sysadmin only)
Bachelor's degree preferred, but experience matters more than degree
No travel required
Interview Process
- Initial online screening
- Likely onsite interview to assess environment fit and hands-on capability
- Suppliers encouraged to pre screen (do not submit blindly)