1

Cloud Performance Engineer Jobs (NOW HIRING)

The Performance Engineer will analyze and improve performance across various production deployments ... and cloud environments • Optimize latency, throughput, memory usage, batching, scheduling ...

Performance Engineer Job Location: Columbus, OH Job Type: Contract * Design and implement ... Dynatrace / AppDynamics / Splunk / Grafana / Prometheus or similar tools Cloud & Infrastructure

Info Way Solutions is seeking a Performance Engineer for their client AEO. The role involves ... • Cloud experience plus • Analyze tests results and work with Developers and Engineers to ...

Software Performance engineer

Dallas, TX · On-site

$138K/yr

Software Performance engineer Location: Anywhere is the US Onsite position Fulltime position ... Strong hands-on programming/scripting on Cloud technologies skills with Python, Java, Bash, AWS ...

Sr. Performance Engineer

Shelton, CT · On-site

$119K - $149K/yr

... cloud-native monitoring). • Experience integrating performance validation into CI/CD pipelines and partnering with SRE/DevOps practices; familiarity with chaos engineering. • Proficiency ...

Sr. Performance Engineer Franchise World Headquarters, LLC Why Join Subway? At Subway, we are not ... Proficiency scripting and automating in Java, JavaScript, or Python; comfortable working in cloud ...

Title: SAP Performance Engineer Location: Austin, TX Job Type: Full Time For this role, we are ... Experience in Cloud-based platforms like AWS and GCP. * Expertise in performance assessment and ...

Performance Engineer

O Fallon, MO · On-site

$60K - $135K/yr

Exposure to cloud platforms such as AWS, Azure, or GCP . * Experience with Performance Engineering in Kubernetes-based production environments. * Basic understanding of DevOps practices, release ...

Work extensively with Google Cloud Platform (GCP) environments and evaluate application performance ... Collaborate with development, DevOps, architecture, database, and infrastructure teams to resolve ...

Cloud Engineer

Montgomery, AL · On-site

$55.25 - $73.75/hr

Monitor cloud performance, security, and resource utilization. * Manage cloud networking, container ... Bachelor's degree. * 5+ years of cloud engineering experience-supporting IL5/IL6 cloud environments ...

Showing results 41-60

Cloud Performance Engineer information

See salary details

$23

$62

$87

How much do cloud performance engineer jobs pay per hour?

As of Sep 10, 2026, the average hourly pay for cloud performance engineer in the United States is $62.89, according to ZipRecruiter salary data. Most workers in this role earn between $53.61 and $71.63 per hour, depending on experience, location, and employer.

What cities are hiring for Cloud Performance Engineer jobs?

Cities with the most Cloud Performance Engineer job openings:

What states have the most Cloud Performance Engineer jobs?

States with the most job openings for Cloud Performance Engineer jobs include:

What are popular job titles related to Cloud Performance Engineer jobs?

For Cloud Performance Engineer jobs, the most frequently searched job titles are:

Infographic showing various Cloud Performance Engineer job openings in the United States as of September 2026, with employment types broken down into 67% Full Time, and 33% Contract. Highlights an 89% In-person, and 11% Hybrid job distribution, with an average salary of $130,802 per year, or $62.9 per hour.

Performance Engineer

Palo Alto, CA • On-site

Full-time

Re-posted 26 days ago


Job description

Job Summary:
RadixArk is an infrastructure-first company focused on building world-class open systems for inference and training in AI. The Performance Engineer will analyze and improve performance across various production deployments and benchmark LLM inference and training workloads, ensuring optimal operation of AI systems in real environments.
Responsibilities:
• Analyze and improve performance across SGLang, Miles, and RadixArk production deployments
• Benchmark LLM inference and training workloads across GPUs, TPUs, and cloud environments
• Optimize latency, throughput, memory usage, batching, scheduling, routing, and GPU utilization
• Investigate performance regressions in real customer environments
• Work closely with kernel, runtime, distributed systems, and product engineers
• Build internal tooling for profiling, tracing, benchmarking, and regression detection
• Translate customer workload characteristics into concrete performance tuning strategies
• Help define performance metrics that matter commercially, including cost-per-token and serving efficiency
• Partner with customers and cloud partners on deep technical evaluations
• Contribute performance insights back to open-source SGLang and Miles
Qualifications:
Required:
• Strong systems engineering background, especially in performance-critical software
• Experience with GPU systems, distributed systems, inference serving, ML runtimes, or high-performance computing
• Familiarity with profiling tools, performance debugging, tracing, and benchmark methodology
• Comfort working with Python and C++
• Ability to debug messy real-world performance issues across software, hardware, and infrastructure layers
• Strong communication skills — you should be able to explain performance tradeoffs to both engineers and customers
Preferred:
• Experience with CUDA, Triton, Pallas, ROCm, XLA, or kernel-level optimization is a strong plus
• Understanding of LLM inference concepts such as batching, KV cache, prefill/decode, speculative decoding, MoE, long context, and P99 latency
• Prior experience with production AI infrastructure, cloud GPU environments, or open-source ML systems is a plus
Company:
RadixArk focuses on developing infrastructure for AI inference and training systems. Founded in 2025, the company is headquartered in San Francisco, USA, with a team of 11-50 employees. The company is currently Early Stage.