This is a hands-on systems engineering role at the intersection of software, firmware, hardware, and large-scale cloud infrastructure. You will work on complex GPU and server platforms, owning ...
This is a hands-on systems engineering role at the intersection of software, firmware, hardware, and large-scale cloud infrastructure. You will work on complex GPU and server platforms, owning ...
As a GPU Performance Engineer, you'll architect and implement the foundational systems that power ... Working at the intersection of hardware and software, you'll implement state-of-the-art techniques ...
As a GPU Performance Engineer, you'll architect and implement the foundational systems that power ... Working at the intersection of hardware and software, you'll implement state-of-the-art techniques ...
The role involves conducting performance characterization and analysis on large multi-GPU and multi ... programming and at least one communication runtime (MPI, NCCL, UCX, NVSHMEM) • Experience ...
The role involves conducting performance characterization and analysis on large multi-GPU and multi ... programming and at least one communication runtime (MPI, NCCL, UCX, NVSHMEM) • Experience ...
Senior System Software Engineer - GPU Performance
Santa Clara, CA · On-site
$152 - $241.50/hr
... Performance Engineer to influence the roadmap of our communication libraries. The DL and HPC ... Study the interaction of our libraries with all hardware (GPU, CPU, networking) and software ...
New
Senior System Software Engineer - GPU Performance
Santa Clara, CA · On-site
$152 - $241.50/hr
... Performance Engineer to influence the roadmap of our communication libraries. The DL and HPC ... Study the interaction of our libraries with all hardware (GPU, CPU, networking) and software ...
New
Senior Staff Software Engineer, GPU System Software
Sunnyvale, CA · On-site
$262 - $364/hr
Senior Staff Software Engineer, GPU System Software * link Copy link Google Sunnyvale, CA, USA Advanced Experience owning outcomes and decision making, solving ambiguous problems and influencing ...
Senior Staff Software Engineer, GPU System Software
Sunnyvale, CA · On-site
$262 - $364/hr
Senior Staff Software Engineer, GPU System Software * link Copy link Google Sunnyvale, CA, USA Advanced Experience owning outcomes and decision making, solving ambiguous problems and influencing ...
The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our ... We are looking for a motivated Performance engineer to influence the roadmap of our communication ...
The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our ... We are looking for a motivated Performance engineer to influence the roadmap of our communication ...
Sr. AI Software Engineer (GPU/C++)
Milpitas, CA · On-site
$166K - $283K/yr
Strong C++ programming skills and solid software engineering fundamentals. * Solid experience and hands-on knowledge with GPU programming (e.g., CUDA) and Linux development environments.
Sr. AI Software Engineer (GPU/C++)
Milpitas, CA · On-site
$166K - $283K/yr
Strong C++ programming skills and solid software engineering fundamentals. * Solid experience and hands-on knowledge with GPU programming (e.g., CUDA) and Linux development environments.
The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our ... We are looking for a motivated Performance engineer to influence the roadmap of our communication ...
The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our ... We are looking for a motivated Performance engineer to influence the roadmap of our communication ...
Senior Software Engineer - GPU Local AI Platforms
Santa Clara, CA · On-site
$143K - $189K/yr
Strong Python or C++ programming, software design, and software engineering skills. * Hands-on experience with GPU kernel development or optimization (CUDA/C++, Triton, or equivalent) - you ...
Senior Software Engineer - GPU Local AI Platforms
Santa Clara, CA · On-site
$143K - $189K/yr
Strong Python or C++ programming, software design, and software engineering skills. * Hands-on experience with GPU kernel development or optimization (CUDA/C++, Triton, or equivalent) - you ...
Senior Software Engineer - GPU Kernel Authoring & Optimization
Sunnyvale, CA · On-site
$182 - $242/hr
... GPU/accelerator software, or performance-critical systems. * Hands-on CUDA experience is required--you have written and optimized custom kernels and are fluent with the CUDA programming and memory ...
Senior Software Engineer - GPU Kernel Authoring & Optimization
Sunnyvale, CA · On-site
$182 - $242/hr
... GPU/accelerator software, or performance-critical systems. * Hands-on CUDA experience is required--you have written and optimized custom kernels and are fluent with the CUDA programming and memory ...
Sr. AI Software Engineer (GPU/C++)
Milpitas, CA · On-site
$166K - $283K/yr
Strong C++ programming skills and solid software engineering fundamentals. * Solid experience and hands-on knowledge with GPU programming (e.g., CUDA) and Linux development environments.
Sr. AI Software Engineer (GPU/C++)
Milpitas, CA · On-site
$166K - $283K/yr
Strong C++ programming skills and solid software engineering fundamentals. * Solid experience and hands-on knowledge with GPU programming (e.g., CUDA) and Linux development environments.
Sr. AI Software Engineer (GPU/C++)
Milpitas, CA · On-site
$166K - $283K/yr
Strong C++ programming skills and solid software engineering fundamentals. * Solid experience and hands-on knowledge with GPU programming (e.g., CUDA) and Linux development environments.
Sr. AI Software Engineer (GPU/C++)
Milpitas, CA · On-site
$166K - $283K/yr
Strong C++ programming skills and solid software engineering fundamentals. * Solid experience and hands-on knowledge with GPU programming (e.g., CUDA) and Linux development environments.
Senior Software Engineer - GPU Local AI Platforms
Santa Clara, CA · On-site
$143K - $189K/yr
Strong Python or C++ programming, software design, and software engineering skills. * Hands-on experience with GPU kernel development or optimization (CUDA/C++, Triton, or equivalent) - you ...
Senior Software Engineer - GPU Local AI Platforms
Santa Clara, CA · On-site
$143K - $189K/yr
Strong Python or C++ programming, software design, and software engineering skills. * Hands-on experience with GPU kernel development or optimization (CUDA/C++, Triton, or equivalent) - you ...
Senior Systems Software Engineer - GPU Performance at Scale
Santa Clara, CA · On-site
$184 - $287.50/hr
We are looking for a dedicated engineer for the Senior Systems Software Engineer role, focusing on GPU Performance at Scale. At NVIDIA, this role is uniquely positioned to drive innovation in AI and ...
New
Senior Systems Software Engineer - GPU Performance at Scale
Santa Clara, CA · On-site
$184 - $287.50/hr
We are looking for a dedicated engineer for the Senior Systems Software Engineer role, focusing on GPU Performance at Scale. At NVIDIA, this role is uniquely positioned to drive innovation in AI and ...
New
The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our ... We are looking for a motivated Performance engineer to influence the roadmap of our communication ...
The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our ... We are looking for a motivated Performance engineer to influence the roadmap of our communication ...
To excel in this role, we seek a candidate with exceptional technical expertise, who can bridge deep proficiency in high-performance C++ software engineering and low-level GPU programming with a ...
To excel in this role, we seek a candidate with exceptional technical expertise, who can bridge deep proficiency in high-performance C++ software engineering and low-level GPU programming with a ...
They are seeking a Senior Systems Software Engineer to drive innovation in AI and GPU computing, focusing on performance practices in large-scale GPU infrastructure. Responsibilities : • Lead the ...
They are seeking a Senior Systems Software Engineer to drive innovation in AI and GPU computing, focusing on performance practices in large-scale GPU infrastructure. Responsibilities : • Lead the ...
Senior Software Engineer - GPU Kernel Authoring & Optimization
Sunnyvale, CA · On-site
$143K - $189K/yr
... GPU/accelerator software, or performance-critical systems. * Hands-on CUDA experience is required--you have written and optimized custom kernels and are fluent with the CUDA programming and memory ...
Quick apply
Senior Software Engineer - GPU Kernel Authoring & Optimization
Sunnyvale, CA · On-site
$143K - $189K/yr
... GPU/accelerator software, or performance-critical systems. * Hands-on CUDA experience is required--you have written and optimized custom kernels and are fluent with the CUDA programming and memory ...
GPU Software Engineer
San Diego, CA · On-site
$98.90 - $148.30/hr
Qualcomm Graphics Software Engineers architect, design, implement, verify, and optimize the structure and performance of GPU hardware, drivers, features, applications, and tools. Qualcomm Engineers ...
GPU Software Engineer
San Diego, CA · On-site
$98.90 - $148.30/hr
Qualcomm Graphics Software Engineers architect, design, implement, verify, and optimize the structure and performance of GPU hardware, drivers, features, applications, and tools. Qualcomm Engineers ...
Sr. Staff/Principal Engineer -- GPU Driver & Systems Software
San Diego, CA · On-site
$179 - $286/hr
Sr. Staff/Principal Engineer -- GPU Driver & Systems Software Category Chip Design Location San Diego, CA or San Jose, CA Experience More than 8 Years Work Experience About the Role: Bring your GPU ...
Sr. Staff/Principal Engineer -- GPU Driver & Systems Software
San Diego, CA · On-site
$179 - $286/hr
Sr. Staff/Principal Engineer -- GPU Driver & Systems Software Category Chip Design Location San Diego, CA or San Jose, CA Experience More than 8 Years Work Experience About the Role: Bring your GPU ...
Temporary Software Engineer Gpu information
What does a temporary software engineer GPU do?
What are the typical projects a temporary software engineer GPU might work on, and how do they collaborate with permanent team members?
What are the key skills and qualifications needed to thrive as a temporary software engineer GPU, and why are they important?
What are the most commonly searched types of Software Engineer Gpu jobs in California?
The most popular types of Software Engineer Gpu jobs in California are:
What cities in California are hiring for Temporary Software Engineer Gpu jobs?
Cities in California with the most Temporary Software Engineer Gpu job openings:
Full-time
Medical, Dental, Vision, Life, Retirement, PTO
This job post has expired today. Applications are no longer accepted.
Oracle rating
8.7
Based on 152 frontline employees who took The Breakroom Quiz
55th of 245 rated software companies
Job description
Oracle Cloud Infrastructure (OCI) is seeking a Principal Systems Software Engineer to help build and evolve the low-level systems software and platform-management capabilities that power next-generation GPU infrastructure.
This is a hands-on systems engineering role at the intersection of software, firmware, hardware, and large-scale cloud infrastructure. You will work on complex GPU and server platforms, owning critical capabilities spanning BMC/service-processor software, platform management, firmware lifecycle, reliability and serviceability, telemetry, power management, hardware bring-up, and fleet operations.
You will work closely with silicon, firmware, hardware, compute, and fleet engineering teams, as well as technology and manufacturing partners, to bring new platforms from initial hardware enablement through production deployment and ongoing operation at cloud scale.
This role is well suited for an engineer with deep experience in computer systems, embedded or platform firmware, and hardware/software integration who enjoys solving difficult problems that cross traditional engineering boundaries.
Responsibilities
Key Responsibilities:
BMC & Platform-Management Systems
- Design, develop, and maintain complex BMC, service-processor, and platform-management software for GPU and server infrastructure.
- Develop capabilities for host management, baseboard management, telemetry, platform monitoring, power control, and system diagnostics.
- Design and implement secure, reliable firmware-management and update workflows across complex, multi-vendor platforms.
- Define robust interfaces and integration contracts between platform firmware, hardware components, operating systems, drivers, and higher-level infrastructure services.
Systems & Firmware Development
- Design and implement low-level systems software and firmware using technologies such as C, C++, Python, and Bash.
- Develop maintainable software for managing, monitoring, diagnosing, and provisioning server and GPU systems.
- Build automation and tooling that improves platform provisioning, onboarding, validation, diagnostics, and fleet operations.
- Develop and debug advanced platform capabilities involving areas such as RAS, telemetry, power management, high-speed I/O, and chipset/SoC services.
- Conduct design and code reviews and help establish strong engineering practices for maintainability, testing, observability, security, and reliability.
Hardware Bring-Up & Cross-Layer Debugging
- Play a leading role in initial board, device, and platform bring-up for new GPU and server systems.
- Diagnose difficult failures spanning hardware, firmware, bootloaders, operating systems, drivers, and platform services.
- Use hardware and software diagnostic techniques-including logs, schematics, JTAG, logic analyzers, emulators, and platform instrumentation-to isolate root causes.
- Partner with silicon, board, firmware, and manufacturing teams to validate end-to-end platform behavior and resolve integration issues.
- Turn complex or recurring failures into durable engineering fixes, improved diagnostics, automation, and preventive controls.
GPU Reliability, Serviceability & Operations
- Develop reliability, availability, and serviceability (RAS) capabilities for large-scale GPU and server environments.
- Improve telemetry, fault detection, logging, observability, and diagnostics used to identify and resolve platform issues.
- Develop and improve power-control and power-capping capabilities and related platform instrumentation.
- Design systems with fleet-scale reliability, fault tolerance, secure firmware lifecycle, and operational serviceability in mind.
- Support difficult platform incidents and escalations and help translate field findings into long-term product and engineering improvements.
Technical Leadership
- Own technically complex and sometimes ambiguous areas from architecture and design through implementation, validation, and deployment.
- Drive technical decisions and establish clear interfaces across teams responsible for different layers of the platform.
- Lead deep technical investigations and help teams reach evidence-based root causes for difficult system failures.
- Raise engineering standards through architecture and design reviews, code reviews, testing practices, automation, and diagnostic tooling.
- Mentor and provide technical guidance to engineers while remaining actively involved in design, coding, bring-up, and debugging.
- Collaborate effectively across software, firmware, hardware, silicon, compute, fleet, support, and external partner organizations.
In addition, you have:
- 6+ years of programming and/or scripting experience, with relevant languages such as C, C++, Python, or Bash.
- Strong systems-software, embedded-software, or firmware development experience.
- Experience integrating software or firmware with complex hardware systems.
- Strong understanding of computer hardware fundamentals and the interaction between hardware, firmware, operating systems, and software.
- Demonstrated ability to diagnose complex problems that cross hardware and software boundaries.
- Experience with software development practices including design, implementation, debugging, code review, testing, automation, and quality assurance.
- Experience working in Linux/Unix-based development or systems environments.
- Ability to independently own technically complex projects and collaborate across multiple engineering organizations.
Preferred Qualifications:
Experience in one or more of the following areas is highly desirable:
- BMC, OpenBMC, service processors, or server platform-management technologies.
- Server, GPU, accelerator, or other complex compute-platform firmware.
- GPU or server RAS, telemetry, fault management, observability, or serviceability.
- Board, device, or system bring-up.
- Hardware debugging using JTAG, logic analyzers, emulators, schematics, or related diagnostic tools.
- Firmware lifecycle management and secure firmware-update mechanisms.
- Power management, power control/capping, thermal management, or platform telemetry.
- Hardware interfaces and low-level communication protocols.
- CPU, GPU, SoC, ASIC, or FPGA-based systems.
- ARM, AMD, Intel, NVIDIA, or similarly complex compute platforms.
- Automation and diagnostics for server provisioning, validation, or fleet operations.
- Large-scale cloud or data-center infrastructure.
- Technical leadership, mentoring, architecture/design ownership, and cross-functional engineering coordination.
Location: On-Site | Santa Clara, CA
Qualifications
Disclaimer:
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $114,600 to $234,600 per annum. May be eligible for bonus, equity, and compensation deferral.
Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.
Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.
Oracle US offers a comprehensive benefits package which includes the following:
1. Medical, dental, and vision insurance, including expert medical opinion
2. Short term disability and long term disability
3. Life insurance and AD&D
4. Supplemental life insurance (Employee/Spouse/Child)
5. Health care and dependent care Flexible Spending Accounts
6. Pre-tax commuter and parking benefits
7. 401(k) Savings and Investment Plan with company match
8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
9. 11 paid holidays
10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
11. Paid parental leave
12. Adoption assistance
13. Employee Stock Purchase Plan
14. Financial planning and group legal
15. Voluntary benefits including auto, homeowner and pet insurance
The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.
Career Level - IC4
About Us
Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.
True innovation starts when everyone is empowered to contribute. That's why we're committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.
We're committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States.
Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans' status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
About Oracle
Sourced by ZipRecruiter
An Oracle career can span industries, roles, Countries and cultures, giving you the opportunity to flourish in new roles and innovate, while blending work life in. Oracle has thrived through 40+ years of change by innovating and operating with integrity while delivering for the top companies in almost every industry. In order to nurture the talent that makes this happen, we are committed to an inclusive culture that celebrates and values diverse insights and perspectives, a workforce that inspires thought leadership and innovation. Oracle offers a highly competitive suite of Employee Benefits designed on the principles of parity, consistency, and affordability. The overall package includes certain core elements such as Medical, Life Insurance, access to Retirement Planning, and much more. We also encourage our employees to engage in the culture of giving back to the communities where we live and do business. At Oracle, we believe that innovation starts with diversity and inclusion and to create the future we need talent from various backgrounds, perspectives, and abilities. We ensure that individuals with disabilities are provided reasonable accommodation to successfully participate in the job application, interview process, and in potential roles. to perform crucial job functions. That's why we're committed to creating a workforce where all individuals can do their best work. It's when everyone's voice is heard and valued that we're inspired to go beyond what's been done before.
Industry
It services
Company size
10,000+ Employees
Headquarters location
Redwood City, CA, US
Year founded
1977