CADENCE
CADENCE

60 Cadence Product Engineer Jobs Hiring Near You

... Secure DevOps pipelines across cloud platforms: Amazon AWS, Microsoft Azure, Google Cloud, IBMC cloud, Cadence software service and products • Implement infrastructure-as-code (IaC) security ...

Software Engineer II

San Jose, CA · On-site

$101K - $188K/yr

At Cadence, we hire and develop leaders and innovators who want to make an impact on the world of ... You will be a member of an expert R&D team creating technologies and products for our Liberate ...

At Cadence, we hire and develop leaders and innovators who want to make an impact on the world of ... Strong Python, C++ programming skills with production code experience * Comfortable working in ...

Experienced Toolmakers-Eyelet

Watertown, CT · On-site

$27 - $35.25/hr

Strong collaboration with engineering, production, and quality teams * A commitment to precision, safety, and high-quality manufacturing About Cadence Cadence is a leading contract manufacturer ...

About Cadence: At Cadence , we improve product performance by building solution-oriented ... We're an engineering company at heart--with over 75 engineers across our organization--and have ...

Job Summary : Cadence is a technology company focused on hiring and developing leaders and ... production stability. • Validate power and cooling readiness prior to IT deployments in ...

Quality Assurance Engineer II

Sturgeon Bay, WI · On-site

$69K - $89K/yr

... product conformity. Why should you choose Cadence? * Shape the Future of Healthcare: Join a team ... Programming CMM Software: * Develop and write CMM programs using specialized software based on ...

CNC Swiss Setup Operator - 3rd Shift

Cranston, RI · On-site

$20.25 - $27.75/hr

Why Join Us? At Cadence, you'll be part of a team that manufactures highly engineered medical ... This individual will work closely with production and quality teams to improve processes, reduce ...

Cadence is a company that hires and develops leaders and innovators in technology. They are seeking ... DevOps, and Application teams to integrate workloads across private and hybrid cloud ...

The technician will collaborate closely with production teams to identify and resolve quality ... quality engineers, and other stakeholders. * Skills in using image capture tools and analysis ...

Showing results 41-60

AI Senior Staff Systems Engineer (R51191)

Cadence

San Jose, CA • On-site

$122K - $167K/yr

Full-time

Re-posted 15 hours ago


Job description

Job Summary:
Cadence is seeking a highly skilled and experienced AI Systems Engineer to join their team. This hands-on role will lead the development, operations, and support of the entire AI infrastructure, focusing on architecting and building high-performance GPU clusters and optimizing advanced AI models.
Responsibilities:
• AI Infrastructure Architecture & Strategy: Lead the design and implementation of our next-generation AI infrastructure to support our Agentic AI initiatives. You will define the technical strategy for our on-premise GPU clusters, storage solutions, and networking to ensure optimal performance, scalability, and reliability for all our AI workloads.
• Cloud AI Service Integration: Support and secure the use of public cloud AI services, including Azure OpenAI services and Google Cloud Platform (GCP) services like Gemini . This includes managing secure access, monitoring usage, and tracking billing to ensure cost-effectiveness. You will also have hands-on experience supporting compute, GPUs, and AI services on both GCP and Azure.
• Hands-on GPU Cluster Management: Take a leadership role in the configuration, installation, and optimization of GPU server clusters. This includes advanced troubleshooting of hardware and software, performance tuning, and implementing best practices for cluster utilization and resource management. You will be an expert in administering job schedulers like LSF in a production environment, including integration with Docker for containerized job submission.
• Full-Stack AI Tech Stack Development & Operations: Architect and deploy a robust and scalable AI tech stack. You will be responsible for the end-to-end operational lifecycle, including setting up and managing deep learning frameworks ( PyTorch , TensorFlow ), containerization with Docker and Kubernetes , and implementing CI/CD pipelines for AI model development.
• Advanced LLM Deployment & Optimization: Lead the deployment, serving, and optimization of Large Language Models (LLMs). You will be an expert in techniques such as model quantization, distillation, and using high-performance serving frameworks (e.g., vLLM , TGI , TensorRT-LLM ) to maximize inference throughput and minimize latency.
• Agentic AI Workflow & Service Engineering: Architect and build production-grade Agentic AI workflows and services. You will be responsible for the technical design and implementation of systems that integrate LLMs with external tools, APIs, and databases, and will mentor other engineers on building robust and scalable AI agent applications.
• Automation & Monitoring: Develop and maintain automation scripts using languages like Python , Bash , or Perl to streamline system maintenance, deployment, and reporting. Implement and manage monitoring solutions for system health, job statuses, GPU utilization, and container performance to proactively identify and resolve issues.
• AI Systems Support & Mentorship: Act as the final escalation point for the most complex technical issues related to our AI infrastructure. You will also serve as a technical leader and mentor to other engineers, providing guidance on best practices in AI systems engineering, performance tuning, and operational excellence.
• Security and Compliance: Develop and implement security best practices for our AI systems and data, ensuring compliance with relevant regulations and protecting our intellectual property.
Qualifications:
Required:
• 10+ years of experience in a senior technical role, with at least 5 years focused on building and operating high-performance computing or AI infrastructure. Proven track record as a Principal or Senior Staff Engineer.
• Expert-level knowledge of NVIDIA GPU architecture and technologies like CUDA and cuDNN. Extensive experience with multi-GPU and multi-node training and inference.
• Proven experience with public cloud AI services, specifically managing access, usage, and billing for Azure OpenAI and Google Cloud Platform (GCP) services.
• Extensive hands-on experience with Docker: image management, container orchestration, and troubleshooting.
• Proficiency in scripting languages such as Python, Bash, or Perl.
• Deep expertise in Linux system administration (RHEL preferred), including networking, storage, and performance tuning.
• Familiarity with user authentication and integration using systems like LDAP or Active Directory.
• Strong problem-solving and communication skills with the ability to work in a multi-platform, cross-functional, and geographically distributed team.
Preferred:
• Understanding of AI job profiling and tuning (memory, GPU, I/O).
• Experience administering LSF clusters in a production or research environment. Familiarity with other job schedulers like Slurm is a plus.
• Experience with LSF Docker integration and job submission using container images.
• Experience with macOS/AppleSilicon system admin tasks and troubleshooting.
Company:
Cadence is a market leader in AI and digital twins, pioneering the application of computational software to accelerate innovation in the engineering design of silicon to systems. Founded in 1988, the company is headquartered in San Jose, USA, with a team of 10001+ employees. The company is currently Late Stage.