1

Vllm Jobs in Florida (NOW HIRING)

Senior AI Engineer

Sarasota, FL · On-site +1

$118K - $155K/yr

Experience with LLM gateways or serving infrastructure (LiteLLM, vLLM, TGI, or similar) * Vector databases (Qdrant, pgvector) and embedding pipelines * Fine-tuning open-weight models (LoRA or full ...

Be Seen First

Local/offline AI model deployment (Ollama, llama.cpp, vLLM, etc.) * Experience supporting RF equipment or specialized networking environments * CAD software support (SolidWorks or similar)

LLM inference and serving optimization * vLLM, TGI, or equivalent * Model selection trade-offs - cost, latency, capability, context window Engineering Depth * Hands-on Python - comfortable writing ...

Integrate LLM providers (OpenAI, Anthropic, Gemini, local models via Ollama / vLLM) with robust fallback, rate-limiting, and cost controls. * Deliver polished, accessible front-end experiences in ...

Integrate LLM providers (OpenAI, Anthropic, Gemini, local models via Ollama / vLLM) with robust fallback, rate-limiting, and cost controls. * Deliver polished, accessible front-end experiences in ...

Vllm information

What is a vLLM?

VLLM stands for 'Virtual Large Language Model.' In the context of AI development, VLLM professionals work with optimized inference engines for large language models, enabling faster and more efficient deployment of AI models in production environments. Their responsibilities often include integrating LLMs into applications, optimizing model performance, and ensuring scalability for real-time use cases. They may also collaborate with data scientists and engineers to manage resources and streamline AI workflows.

How does a vLLM engineer typically collaborate with data scientists and product teams during model deployment?

VLLM Engineers work closely with data scientists to understand the specific requirements and fine-tuning needs of large-scale language models. They are often responsible for integrating these models into production systems, ensuring scalability and efficiency. Collaboration with product teams is crucial to align model capabilities with user needs and to troubleshoot real-world application challenges. Frequent communication and agile workflows are common, as updates or optimizations may be needed rapidly based on feedback from both teams.

What are the key skills and qualifications needed to thrive as a machine learning engineer working with vLLM, and why are they important?

To thrive as a Machine Learning Engineer specializing in vLLM (a high-throughput LLM inference library), you need a strong understanding of machine learning principles, deep learning frameworks, and experience with Python programming. Familiarity with tools like PyTorch, CUDA, distributed computing, and cloud platforms, as well as relevant certifications in ML or data engineering, is highly valuable. Strong problem-solving, collaboration, and communication skills are essential for optimizing model performance and integrating with cross-functional teams. These capabilities ensure effective deployment and scaling of large language models, driving innovation and efficiency in AI applications.

What is the difference between Vllm vs Data Analyst?

AspectVllmData Analyst
Required CredentialsTypically requires knowledge of machine learning, AI, and programming languages like Python or RRequires skills in statistics, Excel, SQL, and data visualization tools
Work EnvironmentOften in tech companies, research labs, or AI-focused teamsCommonly in business, finance, healthcare, and marketing sectors
Industry UsageEmerging role in AI and machine learning projectsEstablished role in data-driven decision making
Common Search/ComparisonVllm vs Data Analyst

The main difference between Vllm and Data Analyst lies in their focus and skill set. Vllm professionals specialize in AI and machine learning models, often working in tech environments, while Data Analysts focus on interpreting data to inform business decisions. Both roles require analytical skills, but Vllm roles demand programming and AI expertise, whereas Data Analysts emphasize statistical analysis and data visualization.

What cities in Florida are hiring for Vllm jobs?

Cities in Florida with the most Vllm job openings:

Infographic showing various Vllm job openings in Florida as of August 2026, with employment types broken down into 95% Full Time, 4% Part Time, and 1% Contract. Highlights an 77% Physical, 6% Hybrid, and 17% Remote job distribution.

$150 - $230/hr

Other

Posted 15 days ago


Job description

Senior AI Engineer

Remote — US only · Full-time

Are you passionate about building AI products people actually use, serving millions of users? Do you want to help lead the AI engineering effort at this country's fastest-growing social media company that champions free speech? You'll be building AI-based features into a platform serving millions of users.

About the Role

You'll work directly with the platform architect and other developers to bring deep, hands-on LLM-stack expertise: you've built these systems before, you know where they break, and you know what "good" looks like in production.

The work is greenfield. You won't be maintaining someone else's pipeline — you'll be standing up the core AI infrastructure for a platform with millions of active users, and shaping the engineering practices around it as the team grows.

What We're Looking For

  • 5+ years of backend engineering experience in a statically typed language (Go, Java, Kotlin, C#)
  • You've shipped production LLM-backed features — retrieval-augmented generation, streaming responses, tool use — and lived with them after launch
  • Hands‑on experience with the LLM serving stack: routing across multiple model providers, failover, token streaming, and cost/usage metering
  • Experience building retrieval systems: vector search, embedding pipelines, context assembly, and citation-backed answers
  • You think in failure modes: hallucination, retrieval misses, provider outages, cost blowouts — and you build the instrumentation to catch them
  • Pragmatic about evaluation — you know how to measure whether AI answers are actually good (relevance, safety, source quality) with simple, repeatable tests, not just academic benchmarks
  • Strong API design instincts; comfortable owning a service end to end, from schema to deploy to on‑call
  • US-based and authorized to work in the United States

Nice to Have

  • Go (strongly preferred)
  • Python
  • Experience with LLM gateways or serving infrastructure (LiteLLM, vLLM, TGI, or similar)
  • Vector databases (Qdrant, pgvector) and embedding pipelines
  • Fine‑tuning open‑weight models (LoRA or full fine‑tunes) and the eval discipline that goes with it
  • Content moderation or safety tooling experience
  • Familiarity with Ruby on Rails or React/TypeScript (you'll integrate with both)

Our Stack

Go, Ruby on Rails, Python, React/TypeScript, PostgreSQL, Redis, RabbitMQ. Services deployed across multiple data centers.

Why TMTG?

We are a social media company dedicated to delivering users an engaging and censorship‑free experience. We believe users should be able to freely express themselves and engage with a rich diversity of viewpoints on a cancel‑proof, accessible platform. If you are interested in joining a company that is fast‑paced, rapidly growing, and committed to free speech, please get in touch.

Equal Employment Opportunity

Our Company is an equal opportunity employer that prohibits discrimination and harassment on the basis of any protected characteristic as outlined by federal, state, or local laws.

This policy applies to all employment practices within our organization, including hiring, recruiting, promotion, termination, layoff, recall, leave of absence, compensation, benefits, training, and apprenticeship. Our Company makes hiring decisions based solely on qualifications, merit, and business needs at the time.

#J-18808-Ljbffr