Lead Cloud HPC- AI Infrastructure Architect(S2S) As a Lead Cloud Integrated Infra Engineer on the ... Linux system administration in production environments * 3+ years designing or operating ...
Lead Cloud HPC- AI Infrastructure Architect(S2S) As a Lead Cloud Integrated Infra Engineer on the ... Linux system administration in production environments * 3+ years designing or operating ...
IT/EDA Engineer
Atlanta, GA · On-site
Falcomm is seeking an IT/EDA Systems Engineer who will sit at the intersection of IT infrastructure ... Experience with HPC clusters or distributed compute systems * Experience with automation tools and ...
IT/EDA Engineer
Atlanta, GA · On-site
Falcomm is seeking an IT/EDA Systems Engineer who will sit at the intersection of IT infrastructure ... Experience with HPC clusters or distributed compute systems * Experience with automation tools and ...
IT/EDA Engineer
Atlanta, GA · On-site
Falcomm is seeking an IT/EDA Systems Engineer who will sit at the intersection of IT infrastructure ... Experience with HPC clusters or distributed compute systems * Experience with automation tools and ...
IT/EDA Engineer
Atlanta, GA · On-site
Falcomm is seeking an IT/EDA Systems Engineer who will sit at the intersection of IT infrastructure ... Experience with HPC clusters or distributed compute systems * Experience with automation tools and ...
IT/EDA Engineer
Atlanta, GA · On-site
Falcomm is seeking an IT/EDA Systems Engineer who will sit at the intersection of IT infrastructure ... Experience with HPC clusters or distributed compute systems * Experience with automation tools and ...
Quick apply
IT/EDA Engineer
Atlanta, GA · On-site
Falcomm is seeking an IT/EDA Systems Engineer who will sit at the intersection of IT infrastructure ... Experience with HPC clusters or distributed compute systems * Experience with automation tools and ...
You will infuse the JPMorgan developer community with an appreciation of the impact that HPC can ... Hands-on practical experience delivering system design, application development, testing, and ...
You will infuse the JPMorgan developer community with an appreciation of the impact that HPC can ... Hands-on practical experience delivering system design, application development, testing, and ...
DevOps Engineer
Atlanta, GA · On-site
$50.75 - $69.50/hr
Collaborate with DevOps, bioinformatics, HPC, infrastructure, cybersecurity, and system administration teams. * Develop and maintain technical documentation, deployment instructions, configuration ...
DevOps Engineer
Atlanta, GA · On-site
$50.75 - $69.50/hr
Collaborate with DevOps, bioinformatics, HPC, infrastructure, cybersecurity, and system administration teams. * Develop and maintain technical documentation, deployment instructions, configuration ...
The Senior Research Applications Engineer leads application service delivery, workflow enablement ... systems, HPC services, and AI-enabled environments. BUSINESS ANALYSIS AND OPTIMIZATION : Analyze ...
The Senior Research Applications Engineer leads application service delivery, workflow enablement ... systems, HPC services, and AI-enabled environments. BUSINESS ANALYSIS AND OPTIMIZATION : Analyze ...
Sr Research Infrastructure Engineer
Augusta, GA · On-site
$102K - $138K/yr
The Senior Research Infrastructure Engineer leads infrastructure planning, capacity management ... system integration, and data movement requirements for research computing, HPC, and AI-enabled ...
Sr Research Infrastructure Engineer
Augusta, GA · On-site
$102K - $138K/yr
The Senior Research Infrastructure Engineer leads infrastructure planning, capacity management ... system integration, and data movement requirements for research computing, HPC, and AI-enabled ...
Security Engineer (OAMD)
Atlanta, GA · On-site
$125 - $150/hr
This role ensures systems comply with CDC, HHS, and federal cybersecurity requirements while ... Experience securing Linux, virtualized, HPC, and hybrid cloud environments. * Familiarity with ...
Security Engineer (OAMD)
Atlanta, GA · On-site
$125 - $150/hr
This role ensures systems comply with CDC, HHS, and federal cybersecurity requirements while ... Experience securing Linux, virtualized, HPC, and hybrid cloud environments. * Familiarity with ...
Service Engineer
Atlanta, GA · On-site
$70K - $100K/yr
... HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon ... server systems. The Service Engineer is a critical part of post-sales support and needs to ...
Service Engineer
Atlanta, GA · On-site
$70K - $100K/yr
... HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon ... server systems. The Service Engineer is a critical part of post-sales support and needs to ...
Service Engineer
Atlanta, GA · On-site
$70K - $100K/yr
... HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon ... server systems. The Service Engineer is a critical part of post-sales support and needs to ...
Service Engineer
Atlanta, GA · On-site
$70K - $100K/yr
... HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon ... server systems. The Service Engineer is a critical part of post-sales support and needs to ...
Service Engineer
Atlanta, GA · On-site
$70K - $100K/yr
... HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon ... server systems. The Service Engineer is a critical part of post-sales support and needs to ...
Service Engineer
Atlanta, GA · On-site
$70K - $100K/yr
... HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon ... server systems. The Service Engineer is a critical part of post-sales support and needs to ...
Security Engineer (OAMD)
Atlanta, GA · On-site
This role ensures systems comply with CDC, HHS, and federal cybersecurity requirements while ... Experience securing Linux, virtualized, HPC, and hybrid cloud environments. * Familiarity with ...
Security Engineer (OAMD)
Atlanta, GA · On-site
This role ensures systems comply with CDC, HHS, and federal cybersecurity requirements while ... Experience securing Linux, virtualized, HPC, and hybrid cloud environments. * Familiarity with ...
Machine Learning Engineer
Savannah, GA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Machine Learning Engineer
Savannah, GA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Machine Learning Engineer
Valdosta, GA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Machine Learning Engineer
Valdosta, GA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Service Engineer
Atlanta, GA · On-site
$70K - $100K/yr
... HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon ... server systems. The Service Engineer is a critical part of post-sales support and needs to ...
Service Engineer
Atlanta, GA · On-site
$70K - $100K/yr
... HPC and IoT/Embedded customers worldwide. We are the #5 fastest growing company among the Silicon ... server systems. The Service Engineer is a critical part of post-sales support and needs to ...
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Machine Learning Engineer
Augusta, GA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Machine Learning Engineer
Augusta, GA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Machine Learning Engineer
Kennesaw, GA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Machine Learning Engineer
Kennesaw, GA · On-site
Run and scale training experiments on cloud or HPC (AWS, GCP, SLURM, Ray), and debug throughput ... systems
Hpc Systems Engineer information
What is an HPC systems engineer?
What are the key skills and qualifications needed to thrive as an HPC systems engineer?
What are some common challenges an HPC systems engineer faces when supporting large-scale computing clusters?
What is the difference between Hpc Systems Engineer vs Hpc Network Engineer?
| Aspect | Hpc Systems Engineer | Hpc Network Engineer |
|---|---|---|
| Credentials | Typically requires a degree in computer science, engineering, or related field; certifications like Cisco CCNA or Linux certifications are common | Similar credentials; often holds networking certifications such as Cisco CCNP or CompTIA Network+ |
| Work Environment | Works on high-performance computing systems, hardware, and software integration in research or enterprise data centers | Focuses on designing, implementing, and maintaining HPC network infrastructure within data centers or research facilities |
| Industry Usage | Used in scientific research, academia, and enterprise sectors with HPC needs | Common in data centers, research institutions, and organizations requiring advanced network performance |
Hpc Systems Engineers and Hpc Network Engineers share overlapping skills in hardware, software, and certifications. However, Hpc Systems Engineers focus on overall system setup and management, while Hpc Network Engineers specialize in network infrastructure. Both roles are vital in supporting high-performance computing environments.

HPC AI Solution Architect (S2S)
Atlanta, GA • On-site
Full-time
Posted 7 days ago
Deloitte rating
8.2
Based on 93 frontline employees who took The Breakroom Quiz
Job description
As a Lead Cloud Integrated Infra Engineer on the Silicon2Service team in Deloitte's AI & Engineering practice, you will design and drive deployment of fully integrated architectures for GPU-accelerated AI factories and high-performance computing infrastructure in close partnership with Deloitte AI specialists and our ecosystem partners. You will shape end-to-end solutions-from discovery and reference architecture mapping through sizing and implementation. You will partner with Sales Executives, AI application specialists, delivery engineering, and managed services to help clients achieve measurable outcomes from private AI assets. You will lead technical solution strategy for pursuits and active opportunities and translate complex client needs into clear, complete solutions and delivery requirements.
Recruiting for this role ends on 10/3/2026.
Work you'll do
As a Lead Cloud Integrated Infra Engineer on the Silicon2Service team, you will be responsible for:
- Leading architecture for pursuits and active opportunities, including discovery, requirements, constraints, and target-state design
- Creatively defining reference architectures for on-premises, cloud, and hybrid GPU platforms across compute, network, storage, security, software and operations
- Driving architecture trade-offs and decisions across performance, scalability, reliability, locality, total cost of ownership, time-to-value, and risk
- Owning the technical solution strategy in proposals and RFPs, including architecture narrative, assumptions, dependencies, sizing guidance, and delivery approach
- Facilitating client workshops and technical reviews and translating engineering detail into executive-ready communications
- Architecting complex, innovative technology solutions with a focus on business outcomes, cost of quality, and long-term scalability and sustainability.
- Engaging with C-Suite client leadership during sales and delivery, including leading technical pre-sales discussions, shaping proposals, and supporting the closing of new business opportunities
- Supporting go-to-market strategies, including participation in industry events, conferences, and client briefings
The Silicon to Service team at Deloitte delivers end-to-end AI factories and advanced technology services that help organizations build, deploy, and operate large-scale, private AI and data platforms. We enable the next phase of enterprise AI adoption through private AI economics with cloud-like ese of use. Join this unique opportunity to work on innovative AI platforms and emerging technologies in the rapidly evolving AI market while solving complex enterprise problems for some of the world's largest organizations.
Qualifications
Required:
- 8+ years of experience in infrastructure architecture or engineering for large-scale platforms including design, implementation, operations, and optimization.
- 4+ years designing or delivering GPU-accelerated platforms for AI, ML, or high-performance computing
- 3+ years Linux system administration in production environments
- 3+ years designing or operating distributed compute clusters for AI/HPC in hybrid cloud setups, including multi-GPU topologies, partitioning, scheduler integration, and scalability for edge-to-cloud workloads.
- 2+ years with high-performance networking or storage for AI/HPC
- 2+ years building containerized platforms using Kubernetes or Red Hat OpenShift, including GPU operators/drivers, CUDA container runtime, and cluster lifecycle automation
- 2+ years automating infrastructure as code(IaC) with tools like Terraform and Ansible
- At least 2 end-to-end deployments of reference architectures in the cloud or on-prem, including variants with security controls, network segmentation, operational runbooks, and validation testing
- Experience in pre-sales or sales engineering, including discovery, solution demonstrations, and proposal/RFP contributions
- Ability to travel 50%, on average, based on the work you do and the clients and industries/sectors you serve.
- Limited immigration sponsorship may be available.
- 2+ years implementing AI/HPC cluster scheduling (Slurm and Kubernetes), including multi-tenant queues, quotas, and GPU-aware policies
- 2+ years supporting generative AI infrastructure patterns, including multi-node distributed training
- Experience with AI agents and frameworks
- Experience with high-throughput storage for AI/HPC
- Experience executing NVIDIA co-sell motions with OEMS (Dell, HPC, Lenovo), CSPs ( AWS, Azure, Google Cloud), or independent software vendors ( Run:ai, OpenShift, Weights & Biases)
You may also be eligible to participate in a discretionary annual incentive program, subject to the rules governing the program, whereby an award, if any, depends on various factors, including, without limitation, individual and organizational performance.
Deloitte is committed to providing reasonable accommodations for people with disabilities. If you require a reasonable accommodation to participate in the recruiting process, please direct your inquiries to the Global Call Center (GCC) at USTalentCICInbox@deloitte.com.
Recruiting tips
From developing a stand out resume to putting your best foot forward in the interview, we want you to feel prepared and confident as you explore opportunities at Deloitte. Check out recruiting tips from Deloitte recruiters.
Benefits
At Deloitte, we know that great people make a great organization. We value our people and offer employees a broad range of benefits. Learn more about what working at Deloitte can mean for you.
Our people and culture
Our inclusive culture empowers our people to be who they are, contribute their unique perspectives, and make a difference individually and collectively. It enables us to leverage different ways of thinking, ideas, and perspectives, and bring more creativity and innovation to help solve our clients' most complex challenges. This makes Deloitte one of the most rewarding places to work.
Our purpose
Deloitte's purpose is to make an impact that matters for our people, clients, and communities. At Deloitte, purpose is synonymous with how we work every day. It defines who we are. Our purpose comes through in our work with clients that enables impact and value in their organizations, as well as through our own investments, commitments, and actions across areas that help drive positive outcomes for our communities. Learn more.
Professional development
From entry-level employees to senior leaders, we believe there's always room to learn. We offer opportunities to build new skills, take on leadership opportunities and connect and grow through mentorship. From on-the-job learning experiences to formal development programs, our professionals have a variety of opportunities to continue to grow throughout their career.
As used in this posting, "Deloitte" means Deloitte Consulting LLP, a subsidiary of Deloitte LLP. Please see https://www.deloitte.com/us/about for a detailed description of the legal structure of Deloitte LLP and its subsidiaries.
All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability or protected veteran status, or any other legally protected basis, in accordance with applicable law.
Qualified applicants with criminal histories, including arrest or conviction records, will be considered for employment in accordance with the requirements of applicable state and local laws, including the Los Angeles County Fair Chance Ordinance for Employers, City of Los Angeles's Fair Chance Initiative for Hiring Ordinance, San Francisco Fair Chance Ordinance, and the California Fair Chance Act. See notices of various fair chance hiring and ban-the-box laws where available. Fair Chance Hiring and Ban-the-Box Notices | Deloitte US Careers
Requisition code: 365497
Job ID 365497
About Deloitte
Sourced by ZipRecruiter
Industry
Finance and insurance and business management consulting
Company size
10,000+ Employees
Headquarters location
Orlando, FL, US