1

Software Engineer Genai Jobs (NOW HIRING)

Senior AI SDE III -- GenAI & ML Platform

San Jose, CA · On-site

$143K - $189K/yr

Senior Software Development Engineer, Advertising Technology Austin Domain 11 - HomeAway onsite Jul ... Senior Software Development Engineer (GenAI, Agentic AI) Washington - Seattle Campus onsite Jun 30 ...

Required * 7+ years of professional software engineering experience, with deep, recent, multi-year Python and strong relational database and schema design skills. * Solid CS fundamentals and a ...

AI Software Engineer (GenAI / LLM)

Barnegat, NJ · On-site

$110K - $148K/yr

Descripcion del empleo Estamos buscando un/a profesional AI Software Engineer (GenAI / LLM) para incorporarse en nuestro equipo y formar parte de un proyecto internacional del sector asegurador.

WI · On-site

$150 - $210/hr

Experience in software engineering or a directly related field. * Experience in developing and/or implementing web-based or mobile applications. * Experience in a leadership role with or without ...

New

As a Visa Software Engineer, you will be an integral part of a multi-functional development team ... Proven ability to build end-to-end GenAI features, from data ingestion and backend services to APIs ...

As a Visa Software Engineer, you will be an integral part of a multi-functional development team ... Proven ability to build end-to-end GenAI features, from data ingestion and backend services to APIs ...

As a Visa Software Engineer, you will be an integral part of a multi-functional development team ... Proven ability to build end-to-end GenAI features, from data ingestion and backend services to APIs ...

As a Visa Software Engineer, you will be an integral part of a multi-functional development team ... Proven ability to build end-to-end GenAI features, from data ingestion and backend services to APIs ...

As a Visa Software Engineer, you will be an integral part of a multi-functional development team ... Proven ability to build end-to-end GenAI features, from data ingestion and backend services to APIs ...

As a Visa Software Engineer, you will be an integral part of a multi-functional development team ... Proven ability to build end-to-end GenAI features, from data ingestion and backend services to APIs ...

Required * 7+ years of professional software engineering experience, with deep, recent, multi-year Python and strong relational database and schema design skills. * Solid CS fundamentals and a ...

Showing results 41-60

Software Engineer Genai information

See salary details

$63.5K

$147.5K

$205.5K

How much do software engineer genai jobs pay per year?

As of Aug 16, 2026, the average yearly pay for software engineer genai in the United States is $147,524.00, according to ZipRecruiter salary data. Most workers in this role earn between $120,000.00 and $173,000.00 per year, depending on experience, location, and employer.

What cities are hiring for Software Engineer Genai jobs?

Cities with the most Software Engineer Genai job openings:

What states have the most Software Engineer Genai jobs?

States with the most job openings for Software Engineer Genai jobs include:

What job categories do people searching Software Engineer Genai jobs look for?

The top searched job categories for Software Engineer Genai jobs are:

Infographic showing various Software Engineer Genai job openings in the United States as of August 2026, with employment types broken down into 1% Internship, 87% Full Time, 8% Part Time, and 4% Contract. Highlights an 87% Physical, 4% Hybrid, and 9% Remote job distribution, with an average salary of $147,524 per year, or $70.9 per hour.

Staff Software Engineer - GenAI inference

Databricks

San Francisco, CA • On-site

Full-time

Re-posted 12 days ago


Job description

Job Summary:
Databricks is the data and AI company, and they are seeking a Staff Software Engineer for GenAI inference. In this role, you will lead the architecture, development, and optimization of the inference engine that powers Databricks Foundation Model API, ensuring high throughput, low latency, and robust scaling.
Responsibilities:
• Own and drive the architecture, design, and implementation of the inference engine, and collaborate on model-serving stack optimized for large-scale LLMs inference
• Partner closely with researchers to bring new model architectures or features (sparsity, activation compression, mixture-of-experts) into the engine
• Lead the end-to-end optimization for latency, throughput, memory efficiency, and hardware utilization across GPUs, and accelerators
• Define and guide standards to build and maintain instrumentation, profiling, and tracing tooling to uncover bottlenecks and guide optimizations
• Architect scalable routing, batching, scheduling, memory management, and dynamic loading mechanisms for inference workloads
• Ensure reliability, reproducibility, and fault tolerance in the inference pipelines, including A/B launches, rollback, and model versioning
• Collaborate cross-functionally on Integrating with federated, distributed inference infrastructure – orchestrate across nodes, balance load, handle communication overhead
• Drive cross-team collaboration: with platform engineers, cloud infrastructure, and security/compliance teams
• Represent the team externally through benchmarks, whitepapers, and open-source contributions
Qualifications:
Required:
• BS/MS/PhD in Computer Science, or a related field
• Strong software engineering background (6+ years or equivalent) in performance-critical systems
• Proven track record of owning complex system components and driving architectural decisions end-to-end
• Deep understanding of ML inference internals: attention, MLPs, recurrent modules, quantization, sparse operations, etc.
• Hands-on experience with CUDA, GPU programming, and key libraries (cuBLAS, cuDNN, NCCL, etc.)
• Strong background in distributed systems design, including RPC frameworks, queuing, RPC batching, sharding, memory partitioning
• Demonstrated ability to uncover and solve performance bottlenecks across layers (kernel, memory, networking, scheduler)
• Experience building instrumentation, tracing, and profiling tools for ML models
• Ability to lead through influence - work closely with ML researchers, translate novel model ideas into production systems
• Excellent communication and leadership skills, with a proactive and ownership-driven mindset
Preferred:
• Bonus: published research or open-source contributions in ML systems, inference optimization, or model serving
Company:
Databricks is a data and AI platform that unifies data engineering, analytics, and machine learning on a lakehouse architecture. Founded in 2013, the company is headquartered in San Francisco, USA, with a team of 5001-10000 employees. The company is currently Late Stage.