1

Ml Inference Jobs in Miami, FL (NOW HIRING)

Build and maintain high-throughput back-end services in Node.js, or Go serving as the inference and ... production AI / ML systems. * Expert-level proficiency in Go and TypeScript / JavaScript ...

Build and maintain high-throughput back-end services in Node.js, or Go serving as the inference and ... years building production AI / ML systems. Expert-level proficiency in Go and TypeScript ...

Build and maintain high-throughput back-end services in Node.js, or Go serving as the inference and ... production AI / ML systems. * Expert-level proficiency in Go and TypeScript / JavaScript ...

Build and maintain high-throughput back-end services in Node.js, or Go serving as the inference and ... production AI / ML systems. * Expert-level proficiency in Go and TypeScript / JavaScript ...

Red-team the enterprise's own AI-LLM-powered products, agents, RAG pipelines, and ML applications, against prompt injection, jailbreaks, model extraction and inversion, membership inference, data and ...

Senior Engineer, Offensive Security

Miami, FL · On-site

$109K - $150K/yr

... and ML applications, prompt injection, jailbreaks, model extraction and inversion, membership inference, data and supply-chain poisoning, evasion, and agent tool/sandbox abuse, validating that ...

... and ML applications, prompt injection, jailbreaks, model extraction and inversion, membership inference, data and supply-chain poisoning, evasion, and agent tool/sandbox abuse, validating that ...

... and ML applications, prompt injection, jailbreaks, model extraction and inversion, membership inference, data and supply-chain poisoning, evasion, and agent tool/sandbox abuse, validating that ...

Red-team the enterprise's own AI-LLM-powered products, agents, RAG pipelines, and ML applications, against prompt injection, jailbreaks, model extraction and inversion, membership inference, data and ...

Red-team the enterprise's own AI-LLM-powered products, agents, RAG pipelines, and ML applications, against prompt injection, jailbreaks, model extraction and inversion, membership inference, data and ...

Senior Storage Engineer

Miami, FL · On-site

$120 - $150/hr

... inference, and developer platforms demanding scalable compute capacity. Hydra Host is building the ... Deep understanding of AI/ML data pipelines: model checkpointing, data locality, and multi-tiered ...

Senior Storage Engineer

Miami, FL · On-site

$120 - $150/hr

... inference, and developer platforms demanding scalable compute capacity. Hydra Host is building the ... ML data pipelines: model checkpointing, data locality, and multi-tiered storage optimization. · ...

Showing results 21-40

Ml Inference information

See Miami, FL salary details

$35.9K

$117.4K

$187.9K

How much do ml inference jobs pay per year?

As of Aug 15, 2026, the average yearly pay for ml inference in Miami, FL is $117,392.00, according to ZipRecruiter salary data. Most workers in this role earn between $94,200.00 and $130,100.00 per year, depending on experience, location, and employer.

What is ML inference?

ML inference refers to the process of using a trained machine learning model to make predictions or decisions based on new data. After a model has been trained on historical data, inference is the phase where that model is deployed and used in real-world applications, such as recognizing speech, detecting objects in images, or recommending products. The focus in ML inference is on speed, efficiency, and scalability to ensure quick predictions, often in real time. This process is critical for practical applications like mobile apps, web services, and embedded systems. Optimizing inference involves reducing latency, memory usage, and computational requirements.

What is the difference between Ml Inference vs Data Scientist?

AspectML InferenceData Scientist
Required CredentialsKnowledge of machine learning models, programming skillsDegree in data science, statistics, or related fields
Work EnvironmentDeploying models in production, real-time data processingData analysis, model development, research
Industry UsageAI product deployment, software companiesResearch institutions, tech firms, consulting

ML Inference focuses on deploying trained models to make predictions on new data, often in real-time. Data Scientists develop and analyze models, working primarily in research and development. While both roles require understanding of machine learning, ML Inference emphasizes deployment and operationalization, whereas Data Scientists focus on model creation and analysis.

What are some common challenges faced by ML inference engineers when deploying models to production?

ML Inference Engineers often encounter challenges such as optimizing model latency and throughput to meet production requirements, ensuring compatibility with diverse hardware environments, and managing model versioning and updates without disrupting service. Additionally, balancing resource utilization and inference accuracy while monitoring real-time performance metrics is crucial. Collaboration with data scientists, DevOps, and software engineers is typically essential to streamline deployment and maintain robust, scalable inference pipelines.

What are the key skills and qualifications needed to thrive in ML inference?

To thrive in ML Inference, you need a solid background in machine learning principles, programming (Python or C++), and experience with deploying models at scale, often supported by a degree in computer science or a related field. Familiarity with frameworks and tools such as TensorFlow, PyTorch, ONNX, and cloud platforms like AWS SageMaker or Google AI Platform is typically required. Strong problem-solving skills, attention to detail, and effective communication are crucial soft skills for collaborating with multidisciplinary teams and optimizing model performance. These skills ensure efficient, scalable, and reliable deployment of machine learning solutions in real-world applications.

Is ML inference a high paying job?

ML inference roles are generally well-paying, especially for those with skills in machine learning frameworks, programming, and cloud platforms. Salaries vary based on experience, location, and industry, but they tend to be higher than average for tech-related positions.

What are popular job titles related to Ml Inference jobs in Miami, FL?

For Ml Inference jobs in Miami, FL, the most frequently searched job titles are:

What job categories do people searching Ml Inference jobs in Miami, FL look for?

The top searched job categories for Ml Inference jobs in Miami, FL are:

What cities near Miami, FL are hiring for Ml Inference jobs?

Cities near Miami, FL with the most Ml Inference job openings:

Principal Full Stack Engineer

Lennar

Miami, FL • On-site

Full-time

Medical, Dental, Vision, Retirement

Re-posted 18 days ago


Lennar rating

8.0

Company rating: 8.0 out of 10

Based on 46 frontline employees who took The Breakroom Quiz

18th of 80 rated construction


Job description

Principal Full Stack Engineer
We are Lennar
Lennar is one of the nation's leading homebuilders, dedicated to making an impact and creating an extraordinary experience for their Homeowners, Communities, and Associates by building quality homes and providing exceptional customer service, giving back to the communities in which we work and live in, and fostering a culture of opportunity and growth for our Associates throughout their career. Lennar has been recognized as a Fortune 500® company and consistently ranked among the top homebuilders in the United States.
About the Role
  • Lead the architecture and delivery of production-grade AI-powered systems across the full stack.
  • Operate at the intersection of applied AI research and software craftsmanship - designing agentic workflows, integrating large language models, and establishing engineering practices at scale.
  • Serve as the technical standard-bearer for how the organization builds, evaluates, and operates AI features.
  • Partner with Engineering and Data to translate ambiguous business problems into concrete technical roadmaps.
  • Drive AI engineering strategy across the stack, authoring RFCs and leading design reviews that raise the bar for the entire engineering org.

What You'll Do
  • Design and own end-to-end AI features: from prompt engineering and model selection through to scalable APIs and responsive front-end surfaces.
  • Architect multi-agent and agentic workflow systems using frameworks such as LangGraph, CrewAI, AutoGen, or custom orchestration layers.
  • Build retrieval-augmented generation (RAG) pipelines with production-grade chunking, embedding, indexing, and reranking strategies.
  • Integrate LLM providers (OpenAI, Anthropic, Gemini, local models via Ollama / vLLM) with robust fallback, rate-limiting, and cost controls.
  • Deliver polished, accessible front-end experiences in React that expose AI capabilities with low latency and graceful degradation.
  • Build and maintain high-throughput back-end services in Node.js, or Go serving as the inference and orchestration layer.
  • Define and implement LLM evaluation frameworks evals, red-teaming, regression suites to continuously measure model quality in production.
  • Instrument AI pipelines with tracing, latency histograms, and cost dashboards using tools such as LangSmith, Helicone, or Grafana..
  • Embed responsible-AI practices: hallucination mitigation, PII redaction, output guardrails, and bias monitoring into every feature.
  • Mentor senior engineers, conduct rigorous code and architecture reviews, and contribute to hiring by defining role expectations and screening candidates.

What We're Looking For
  • 8+ years of professional software engineering with at least 2 years building production AI / ML systems.
  • Expert-level proficiency in Go and TypeScript / JavaScript; comfortable owning the full stack from model to UI.
  • Hands-on experience designing and shipping LLM-powered features including prompt engineering, fine-tuning, and RLHF concepts.
  • Deep understanding of transformer architectures, token economics, context window management, and inference optimization.
  • Proven track record with RAG architectures, embedding models, and semantic search at scale.
  • Strong grasp of distributed systems, API design, and cloud-native infrastructure (AWS)..
  • Experience with MCP (Model Context Protocol), tool-use patterns, and multi-agent orchestration in production.
  • Demonstrated ability to lead without formal authority, influencing direction through technical credibility and clear communication.
  • Contributions to open-source AI projects or published technical writing on AI engineering topics preferred.
  • Experience in regulated or enterprise environments where auditability, security, and compliance are first-class concerns is a plus.

Life at Lennar
At Lennar, we are committed to fostering a supportive and enriching environment for our Associates, offering a comprehensive array of benefits designed to enhance their well-being and professional growth. Our Associates have access to robust health insurance plans, including Medical, Dental, and Vision coverage, ensuring their health needs are well taken care of. Our 401(k) Retirement Plan, complete with a $1 for $1 Company Match up to 5%, helps secure their financial future, while Paid Parental Leave and an Associate Assistance Plan provide essential support during life's critical moments. To further support our Associates, we provide an Education Assistance Program and up to $30,000 in Adoption Assistance, underscoring our commitment to their diverse needs and aspirations. From the moment of hire, they can enjoy up to three weeks of vacation annually, alongside generous Holiday, Sick Leave, and Personal Day policies. Additionally, we offer a New Hire Referral Bonus Program, significant Home Purchase Discounts, and unique opportunities such as the Everyone's Included Day. At Lennar, we believe in investing in our Associates, empowering them to thrive both personally and professionally. Lennar Associates will have access to these benefits as outlined by Lennar's policies and applicable plan terms. Visit Lennartotalrewards.com to view our suite of benefits.
Join the fun and follow us on social media to see what's happening at our company, and don't forget to connect with us on Lennar: Overview | LinkedIn for the latest job opportunities.
Lennar is an equal opportunity employer and complies with all applicable federal, state, and local fair employment practices laws.

What Lennar employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom


Lennar logo

About Lennar

Sourced by ZipRecruiter

Since 1954, Lennar has built over one million new homes for families across America. We build in some of the nation’s most popular cities, and our communities cater to all lifestyles and family dynamics, whether you are a first-time or move-up buyer, multigenerational family, or Active Adult.

Industry

Construction

Company size

5,001 - 10,000 Employees

Headquarters location

Miami, FL, US

Year founded

1954

Social media