Senior AI Tooling Engineer
Sunnyvale, CA · On-site
$140 - $190/hr
Nsight-compute #J-18808-Ljbffr
Sunnyvale, CA · On-site
$140 - $190/hr
Nsight-compute #J-18808-Ljbffr
Sunnyvale, CA · On-site
$140 - $190/hr
Nsight-compute #J-18808-Ljbffr
Analyze GPU execution using NVIDIA Nsight Systems and Nsight Compute. * Investigate PTX and SASS code generation to understand low-level execution behavior. * Collaborate with researchers and ...
Analyze GPU execution using NVIDIA Nsight Systems and Nsight Compute. * Investigate PTX and SASS code generation to understand low-level execution behavior. * Collaborate with researchers and ...
Analyze GPU execution using NVIDIA Nsight Systems and Nsight Compute. * Investigate PTX and SASS code generation to understand low-level execution behavior. * Collaborate with researchers and ...
Quick apply
Analyze GPU execution using NVIDIA Nsight Systems and Nsight Compute. * Investigate PTX and SASS code generation to understand low-level execution behavior. * Collaborate with researchers and ...
$143K - $189K/yr
We make NVIDIA's core developer tools - including Nsight Compute and Nsight Systems - first-class citizens for AI agents through MCP servers and Agent Skills. The space shifts every few weeks, and we ...
$143K - $189K/yr
We make NVIDIA's core developer tools - including Nsight Compute and Nsight Systems - first-class citizens for AI agents through MCP servers and Agent Skills. The space shifts every few weeks, and we ...
Santa Clara, CA · On-site
$143K - $189K/yr
We make NVIDIA's core developer tools - including Nsight Compute and Nsight Systems - first-class citizens for AI agents through MCP servers and Agent Skills. The space shifts every few weeks, and we ...
Santa Clara, CA · On-site
$143K - $189K/yr
We make NVIDIA's core developer tools - including Nsight Compute and Nsight Systems - first-class citizens for AI agents through MCP servers and Agent Skills. The space shifts every few weeks, and we ...
Santa Clara, CA · On-site
$143K - $189K/yr
We make NVIDIA's core developer tools - including Nsight Compute and Nsight Systems - first-class citizens for AI agents through MCP servers and Agent Skills. The space shifts every few weeks, and we ...
Santa Clara, CA · On-site
$143K - $189K/yr
We make NVIDIA's core developer tools - including Nsight Compute and Nsight Systems - first-class citizens for AI agents through MCP servers and Agent Skills. The space shifts every few weeks, and we ...
... Nsight for maximum performance • Work on Linux kernel internals, scheduling, memory management, and resource isolation at cluster scale • Build custom container orchestration, virtualization ...
... Nsight for maximum performance • Work on Linux kernel internals, scheduling, memory management, and resource isolation at cluster scale • Build custom container orchestration, virtualization ...
Profile workloads using NVIDIA Nsight Systems, Nsight Compute, PyTorch Profiler, and custom instrumentation. Eliminate bottlenecks in host code, CUDA kernels, memory, communication, and scheduling.
Profile workloads using NVIDIA Nsight Systems, Nsight Compute, PyTorch Profiler, and custom instrumentation. Eliminate bottlenecks in host code, CUDA kernels, memory, communication, and scheduling.
Santa Clara, CA · On-site
$143K - $189K/yr
Profile workloads using NVIDIA Nsight Systems, Nsight Compute, PyTorch Profiler, and custom instrumentation. Eliminate bottlenecks in host code, CUDA kernels, memory, communication, and scheduling.
Santa Clara, CA · On-site
$143K - $189K/yr
Profile workloads using NVIDIA Nsight Systems, Nsight Compute, PyTorch Profiler, and custom instrumentation. Eliminate bottlenecks in host code, CUDA kernels, memory, communication, and scheduling.
Palo Alto, CA · On-site
$180K - $440K/yr
Develop and tune low-level CUDA kernels (GeMM, Attention, etc.), using CUTLASS, Tensor Cores, and Nsight for maximum performance * Profile, debug, and eliminate bottlenecks across GPU memory ...
Palo Alto, CA · On-site
$180K - $440K/yr
Develop and tune low-level CUDA kernels (GeMM, Attention, etc.), using CUTLASS, Tensor Cores, and Nsight for maximum performance * Profile, debug, and eliminate bottlenecks across GPU memory ...
Santa Clara, CA · On-site
$143K - $189K/yr
Profile workloads using NVIDIA Nsight Systems, Nsight Compute, PyTorch Profiler, and custom instrumentation. Eliminate bottlenecks in host code, CUDA kernels, memory, communication, and scheduling.
Santa Clara, CA · On-site
$143K - $189K/yr
Profile workloads using NVIDIA Nsight Systems, Nsight Compute, PyTorch Profiler, and custom instrumentation. Eliminate bottlenecks in host code, CUDA kernels, memory, communication, and scheduling.
Sunnyvale, CA · On-site
$336K - $359K/yr
Profile ML workloads to identify their bottlenecks, e.g. using NVIDIA Nsight Systems * Design and implement efficiency improvements to maximize MFU and throughput, e.g. parallelism, model compilation ...
Sunnyvale, CA · On-site
$336K - $359K/yr
Profile ML workloads to identify their bottlenecks, e.g. using NVIDIA Nsight Systems * Design and implement efficiency improvements to maximize MFU and throughput, e.g. parallelism, model compilation ...
You use tools like Nsight Systems / Nsight Compute to find bottlenecks, validate hypotheses, and iterate until improvements show up in end-to-end benchmarks. * Bridges theory and practice: You can ...
You use tools like Nsight Systems / Nsight Compute to find bottlenecks, validate hypotheses, and iterate until improvements show up in end-to-end benchmarks. * Bridges theory and practice: You can ...
Profile workloads using Nsight Systems, kernel traces, and internal analysis tools. Use roofline and speed-of-light analysis to find credible headroom and drive fixes from hypothesis to measured wins.
Profile workloads using Nsight Systems, kernel traces, and internal analysis tools. Use roofline and speed-of-light analysis to find credible headroom and drive fixes from hypothesis to measured wins.
Sunnyvale, CA · On-site
$336K - $359K/yr
Profile ML workloads to identify their bottlenecks, e.g. using NVIDIA Nsight Systems * Design and implement efficiency improvements to maximize MFU and throughput, e.g. parallelism, model compilation ...
Sunnyvale, CA · On-site
$336K - $359K/yr
Profile ML workloads to identify their bottlenecks, e.g. using NVIDIA Nsight Systems * Design and implement efficiency improvements to maximize MFU and throughput, e.g. parallelism, model compilation ...
Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation * Write high-performance CUDA and Triton kernels for critical model operations * Optimize cold start ...
Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation * Write high-performance CUDA and Triton kernels for critical model operations * Optimize cold start ...
San Francisco, CA · On-site
$126K - $166K/yr
Experience with profiling and benchmarking tools (e.g., Nsight Systems, Nsight Compute) to validate performance on complex architectures. * Experience identifying and resolving compute and data flow ...
Quick apply
San Francisco, CA · On-site
$126K - $166K/yr
Experience with profiling and benchmarking tools (e.g., Nsight Systems, Nsight Compute) to validate performance on complex architectures. * Experience identifying and resolving compute and data flow ...
Palo Alto, CA · On-site
$180 - $440/hr
Develop and tune low‑level CUDA kernels (GeMM, Attention, etc.), using CUTLASS, Tensor Cores, and Nsight for maximum performance * Profile, debug, and eliminate bottlenecks across GPU memory ...
Palo Alto, CA · On-site
$180 - $440/hr
Develop and tune low‑level CUDA kernels (GeMM, Attention, etc.), using CUTLASS, Tensor Cores, and Nsight for maximum performance * Profile, debug, and eliminate bottlenecks across GPU memory ...
Palo Alto, CA · On-site
$180K - $440K/yr
Develop and tune low-level CUDA kernels (GeMM, Attention, etc.), using CUTLASS, Tensor Cores, and Nsight for maximum performance * Profile, debug, and eliminate bottlenecks across GPU memory ...
Palo Alto, CA · On-site
$180K - $440K/yr
Develop and tune low-level CUDA kernels (GeMM, Attention, etc.), using CUTLASS, Tensor Cores, and Nsight for maximum performance * Profile, debug, and eliminate bottlenecks across GPU memory ...
Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation * Write high-performance CUDA and Triton kernels for critical model operations * Optimize cold start ...
Quick apply
Profile and optimize GPU workloads using Nsight Systems, nvprof, and custom instrumentation * Write high-performance CUDA and Triton kernels for critical model operations * Optimize cold start ...
| Aspect | Nsight | Network Security Analyst |
|---|---|---|
| Required Certifications | Typically Cisco, CompTIA Security+ | CompTIA Security+, CISSP, CEH |
| Work Environment | IT consulting firms, tech companies, remote options | Corporate IT departments, security firms, government agencies |
| Industry Usage | Technology, consulting, cybersecurity | Cybersecurity, IT, finance, government |
| Common Search/Comparison | Yes | Yes |
While Nsight professionals focus on providing IT consulting and solutions, Network Security Analysts specialize in protecting networks from threats. Both roles require cybersecurity certifications and work in tech environments, but Nsight roles often involve broader IT consulting, whereas Network Security Analysts focus specifically on security measures and threat mitigation.
For Nsight jobs in California, the most frequently searched job titles are:
The top searched job categories for Nsight jobs in California are:
Cities in California with the most Nsight job openings:

$140 - $190/hr
Other
This job post has expired 1 day ago. Applications are no longer accepted.
Demonstrates expertise in AI/ML with a strong focus on improving training and inference efficiency, alongside proficiency in ML frameworks and production‑quality coding. Capable of influencing model architecture decisions and evolving toolchains to leverage advancements in AI.
Tools & Technologies