Go & Python Developer
Atlanta, GA · On-site
$48.25 - $66.50/hr
GPU compute (NVIDIA CUDA/DCGM, vllm, llm-d, Intel Level-Zero/ driver stack and SR-IOV)
Atlanta, GA · On-site
$48.25 - $66.50/hr
GPU compute (NVIDIA CUDA/DCGM, vllm, llm-d, Intel Level-Zero/ driver stack and SR-IOV)
Atlanta, GA · On-site
$48.25 - $66.50/hr
GPU compute (NVIDIA CUDA/DCGM, vllm, llm-d, Intel Level-Zero/ driver stack and SR-IOV)
Atlanta, GA · On-site
$126 - $262/hr
Ray Serve,vLLM/NIM/Triton, and NVIDIA Dynamo, with sandboxed execution (gVisor/Firecracker for hosted, NVIDIAOpenShell/vNodefor on-prem) for isolated, safe model execution. * Own cognitive and ...
Atlanta, GA · On-site
$126 - $262/hr
Ray Serve,vLLM/NIM/Triton, and NVIDIA Dynamo, with sandboxed execution (gVisor/Firecracker for hosted, NVIDIAOpenShell/vNodefor on-prem) for isolated, safe model execution. * Own cognitive and ...
Alpharetta, GA · On-site
$55.75 - $74/hr
Deploy, scale, and manage LLM inference servers (e.g., vLLM, Ray Serve, NVIDIA Triton) on Kubernetes across multi-cloud environments. * Implement comprehensive observability, logging, and tracing for ...
Alpharetta, GA · On-site
$55.75 - $74/hr
Deploy, scale, and manage LLM inference servers (e.g., vLLM, Ray Serve, NVIDIA Triton) on Kubernetes across multi-cloud environments. * Implement comprehensive observability, logging, and tracing for ...
Alpharetta, GA · On-site
$105K - $137K/yr
Deploy and manage containerized AI applications and model inference servers (e.g., vLLM, Ray Serve, NVIDIA Triton) on Kubernetes across multi-cloud environments (AWS, GCP, Azure). * Implement ...
Alpharetta, GA · On-site
$105K - $137K/yr
Deploy and manage containerized AI applications and model inference servers (e.g., vLLM, Ray Serve, NVIDIA Triton) on Kubernetes across multi-cloud environments (AWS, GCP, Azure). * Implement ...
Atlanta, GA · On-site
$140 - $210/hr
Experienced with LLM inference stacks (TensorRT-LLM, VLLM, llama.cpp) and developing RAG systems. * Strong problem-solving and analytical skills, with excellent communication to foster cross ...
Atlanta, GA · On-site
$140 - $210/hr
Experienced with LLM inference stacks (TensorRT-LLM, VLLM, llama.cpp) and developing RAG systems. * Strong problem-solving and analytical skills, with excellent communication to foster cross ...
Atlanta, GA · On-site
... vLLM, TensorRT-LLM, or TGI. Small language models & open-weight models • Train and optimize open-weight models such as Llama, Qwen, Mistral, or DeepSeek; build specialized small language models ...
Atlanta, GA · On-site
... vLLM, TensorRT-LLM, or TGI. Small language models & open-weight models • Train and optimize open-weight models such as Llama, Qwen, Mistral, or DeepSeek; build specialized small language models ...
Atlanta, GA · On-site
$126 - $230/hr
Deep expertise operating model-serving and inference systems (Ray, vLLM/Triton/NIM) on GPUs at production scale. * Deep observability skills: metrics, logs, traces, and OpenTelemetry. * FinOps ...
Atlanta, GA · On-site
$126 - $230/hr
Deep expertise operating model-serving and inference systems (Ray, vLLM/Triton/NIM) on GPUs at production scale. * Deep observability skills: metrics, logs, traces, and OpenTelemetry. * FinOps ...
Atlanta, GA · On-site
$107 - $177/hr
Deep expertise operating model-serving and inference systems (Ray, vLLM/Triton/NIM) on GPUs at production scale. * Deep observability skills: metrics, logs, traces, and OpenTelemetry. * FinOps ...
Atlanta, GA · On-site
$107 - $177/hr
Deep expertise operating model-serving and inference systems (Ray, vLLM/Triton/NIM) on GPUs at production scale. * Deep observability skills: metrics, logs, traces, and OpenTelemetry. * FinOps ...
Atlanta, GA · On-site
Optimize inference performance - latency, throughput, quantization, and deployment efficiency - for production, including frameworks such as vLLM, TensorRT-LLM, or TGI. Small language models & open ...
Atlanta, GA · On-site
Optimize inference performance - latency, throughput, quantization, and deployment efficiency - for production, including frameworks such as vLLM, TensorRT-LLM, or TGI. Small language models & open ...
$60.50 - $79.75/hr
Familiarity one or more deep learning frameworks (PyTorch) and modern LLM stack (VLLM, langchain / LlamaIndex) * Experience using Slurm or Kubernetes for ML job orchestration * Experience with ...
Quick apply
$60.50 - $79.75/hr
Familiarity one or more deep learning frameworks (PyTorch) and modern LLM stack (VLLM, langchain / LlamaIndex) * Experience using Slurm or Kubernetes for ML job orchestration * Experience with ...
Atlanta, GA · On-site
$60.50 - $79.75/hr
Familiarity one or more deep learning frameworks (PyTorch) and modern LLM stack (VLLM, langchain / LlamaIndex) * Experience using Slurm or Kubernetes for ML job orchestration * Experience with ...
Atlanta, GA · On-site
$60.50 - $79.75/hr
Familiarity one or more deep learning frameworks (PyTorch) and modern LLM stack (VLLM, langchain / LlamaIndex) * Experience using Slurm or Kubernetes for ML job orchestration * Experience with ...
Atlanta, GA · On-site
$182 - $242/hr
Familiarity with one or more deep learning frameworks (PyTorch) and modern LLM stack (VLLM, langchain / LlamaIndex) * Experience using Slurm or Kubernetes for ML job orchestration * Experience with ...
Atlanta, GA · On-site
$182 - $242/hr
Familiarity with one or more deep learning frameworks (PyTorch) and modern LLM stack (VLLM, langchain / LlamaIndex) * Experience using Slurm or Kubernetes for ML job orchestration * Experience with ...
$182K - $242K/yr
Familiarity one or more deep learning frameworks (PyTorch) and modern LLM stack (VLLM, langchain / LlamaIndex) * Experience using Slurm or Kubernetes for ML job orchestration * Experience with ...
Quick apply
$182K - $242K/yr
Familiarity one or more deep learning frameworks (PyTorch) and modern LLM stack (VLLM, langchain / LlamaIndex) * Experience using Slurm or Kubernetes for ML job orchestration * Experience with ...
| Aspect | Vllm | Data Analyst |
|---|---|---|
| Required Credentials | Typically requires knowledge of machine learning, AI, and programming languages like Python or R | Requires skills in statistics, Excel, SQL, and data visualization tools |
| Work Environment | Often in tech companies, research labs, or AI-focused teams | Commonly in business, finance, healthcare, and marketing sectors |
| Industry Usage | Emerging role in AI and machine learning projects | Established role in data-driven decision making |
| Common Search/Comparison | Vllm vs Data Analyst |
The main difference between Vllm and Data Analyst lies in their focus and skill set. Vllm professionals specialize in AI and machine learning models, often working in tech environments, while Data Analysts focus on interpreting data to inform business decisions. Both roles require analytical skills, but Vllm roles demand programming and AI expertise, whereas Data Analysts emphasize statistical analysis and data visualization.
For Vllm jobs in Georgia, the most frequently searched job titles are:
The top searched job categories for Vllm jobs in Georgia are:
Cities in Georgia with the most Vllm job openings:

$48.25 - $66.50/hr
Full-time
Posted 22 days ago
Sourced by ZipRecruiter
Lorven Technologies, headquartered in Plainsboro, New Jersey, United States, is a reputable company in the technology industry, specializing in providing effective IT solutions and consulting services. The company's official website, lorventech.com, offers comprehensive insights into its offerings which include but are not limited to software development, IT consulting, project management, and business analysis. Since its inception, Lorven Technologies has been committed to ensuring efficiency and reliability in delivering IT services to its global clientele, establishing itself as a trusted name in the industry.
It services
51 - 200 Employees
Plainsboro, NJ, US
2001