1

Hugging Face Jobs (NOW HIRING)

Design and develop enterprise-grade Generative AI applications using OpenAI, AWS Bedrock, and Hugging Face. * Build and optimize Retrieval-Augmented Generation (RAG) solutions using LangChain ...

Design and develop enterprise-grade Generative AI applications using OpenAI, AWS Bedrock, and Hugging Face. * Build and optimize Retrieval-Augmented Generation (RAG) solutions using LangChain ...

AI Engineer

Little Rock, AR · On-site

$109K - $131K/yr

Rapidly test new tools, frameworks, and APIs (e.g., OpenAI, Hugging Face, Google Vertex AI) for use in forward-looking applications. * Engage in ideation sessions and support the development of MVPs ...

Java AI/LLM

Glen Lyn, VA · Remote

$52.25 - $67.50/hr

AI/LLM - hugging face model, OLAMA, LLAMA, Mistral * Agentic AI, Open AI, Gemini * Fine tuning of LLM * Lang chain, Lang flow, FAISS, vector database, Cosine similarity search. * Python programing ...

Hands-on expertise with Generative AI frameworks (e.g., Hugging Face, LangChain, OpenAI APIs). Strong background in Machine Learning, NLP, and Deep Learning. Experience with cloud platforms (Azure ...

New

Senior Engineer (ML/AI)

$107K - $146K/yr

... Hugging Face and related tooling • Translate research ideas into production-ready systems • Optimize models for low-latency, low-power, offline execution • Perform quantization, pruning, and ...

Demonstrate advanced programming expertise, particularly in Python, with deep proficiency in AI-centric libraries such as TensorFlow, PyTorch, and Hugging Face Transformers. * Architect and implement ...

Senior Machine Learning Engineer

Austin, TX · On-site

$103K - $142K/yr

... Hugging Face Transformers for model development, fine-tuning, and optimization. • Implement model optimization techniques such as quantization, pruning, distillation, and hardware specific ...

Senior Machine Learning Engineer

Austin, TX · On-site

$103K - $142K/yr

... Hugging Face Transformers for model development, fine-tuning, and optimization. • Implement model optimization techniques such as quantization, pruning, distillation, and hardware specific ...

Generative AI Engineering Intern (Graduate)

$17.25 - $22.25/hr

Hands-on experience with LLM frameworks such as LangChain, LlamaIndex, or Hugging Face Transformers. * Experience with prompt engineering techniques including few-shot learning, chain-of-thought ...

GPT 5, Claude, Gemini, Azure OpenAI Service, LangChain, LlamaIndex, Hugging Face Transformers, Python, Pinecone, ChromaDB, Weaviate, pgvector, AWS, Azure, Docker, Kubernetes, REST APIs, JSON/GeoJSON ...

Senior Machine Learning Engineer

Austin, TX · On-site

$103K - $142K/yr

... Hugging Face Transformers for model development, fine-tuning, and optimization. • Implement model optimization techniques such as quantization, pruning, distillation, and hardware specific ...

Senior Engineer (ML/AI)

$107K - $146K/yr

... Hugging Face and related tooling • Translate research ideas into production-ready systems • Optimize models for low-latency, low-power, offline execution • Perform quantization, pruning, and ...

Showing results 21-40

Hugging Face information

See salary details

$8

$15

$20

How much do hugging face jobs pay per hour?

As of Aug 6, 2026, the average hourly pay for hugging face in the United States is $15.46, according to ZipRecruiter salary data. Most workers in this role earn between $12.98 and $18.27 per hour, depending on experience, location, and employer.

What is the difference between Hugging Face vs Machine Learning Engineer?

AspectHugging FaceMachine Learning Engineer
Required CredentialsTypically requires knowledge of NLP, deep learning, and Python; certifications are optionalRequires degrees in CS or related fields; experience with ML frameworks; certifications beneficial
Work EnvironmentCollaborative, research-focused, often in tech companies or startupsDevelopment, deployment, and optimization of ML models in various industries
Employer & Industry UsageUsed by AI/ML companies, research labs, and open-source communitiesEmployed across tech, finance, healthcare, and other sectors implementing ML solutions

Hugging Face primarily focuses on NLP tools, libraries, and open-source models, serving as a platform for AI research and development. Machine Learning Engineers develop, implement, and optimize ML models across various domains. While Hugging Face offers resources and tools that ML Engineers use, the roles differ: Hugging Face is a platform, whereas Machine Learning Engineer is a job role involving hands-on model development and deployment.

More about Hugging Face jobs
What cities are hiring for Hugging Face jobs? Cities with the most Hugging Face job openings:
What states have the most Hugging Face jobs? States with the most job openings for Hugging Face jobs include:
Infographic showing various Hugging Face job openings in the United States as of July 2026, with employment types broken down into 79% Full Time, 19% Part Time, and 2% Contract. Highlights an 90% Physical, 1% Hybrid, and 9% Remote job distribution, with an average salary of $32,151 per year, or $15.5 per hour.

AI Observability Engineer

Merican Inc

Charlotte, NC • On-site

Other

Posted 12 days ago


Job description

Job Title: AI Observability Engineer

Location: Charlotte, NC / Philadelphia, PA

Job Description

We are looking for an experienced AI Observability Engineer to design, deploy, and monitor enterprise AI/ML solutions with a strong focus on Generative AI, LLMs, RAG, and AI observability. The ideal candidate will have hands-on experience building scalable AI applications, implementing monitoring frameworks, and optimizing AI model performance in production environments.

Key Responsibilities
  • Design and develop enterprise-grade Generative AI applications using OpenAI, AWS Bedrock, and Hugging Face.
  • Build and optimize Retrieval-Augmented Generation (RAG) solutions using LangChain, LangGraph, vector embeddings, and Azure AI Search.
  • Develop Agentic AI workflows and multi-agent orchestration solutions.
  • Optimize LLM inference using LoRA, QLoRA, vLLM, PagedAttention, and continuous batching.
  • Build REST APIs and AI microservices using Python and FastAPI.
  • Develop and maintain ML pipelines for model training, deployment, monitoring, and lifecycle management.
  • Implement AI observability using tools like Arize to monitor model performance, prompt quality, hallucinations, latency, and inference metrics.
  • Establish AI governance, evaluation frameworks, and guardrails for responsible AI.
  • Develop machine learning models using PyTorch, Scikit-learn, and XGBoost.
  • Build analytics dashboards and provide AI-driven insights to stakeholders.
  • Deploy containerized applications using Docker, GitHub, and CI/CD pipelines.
  • Collaborate with cross-functional teams to deliver scalable AI solutions.
Required Skills
  • Strong experience with Generative AI, LLMs, and RAG architectures.
  • Hands-on experience with OpenAI, AWS Bedrock, Hugging Face, LangChain, and LangGraph.
  • Proficiency in Python and FastAPI.
  • Experience with AI observability platforms such as Arize.
  • Knowledge of PyTorch, Scikit-learn, and XGBoost.
  • Experience with Docker, GitHub, and CI/CD pipelines.
  • Strong understanding of AI governance, model evaluation, and responsible AI practices.
  • Excellent analytical, problem-solving, and communication skills.

Merican logo

About Merican

Sourced by ZipRecruiter

Merican is a IT Service consulting firm, specialized in Digital adoption and Business automation. With our diverse collection of skilled and committed consultants, technology companies, businesses and digital experts, we provide our subject expertise and our unique client service approach, a best-in-class global model of delivery suited to the business demands of our clients. We ensure that we implement future-oriented solutions for our clients via investments in people, solutions, technologies, competencies and infrastructure.

Industry

It services

Company size

51 - 200 Employees

Headquarters location

Columbia , MD, US

Year founded

2020

Social media