1

Python Llm Jobs in Charlotte, NC (NOW HIRING)

Generative AI Engineer

Charlotte, NC · On-site

$140 - $190/hr

Deploy and optimize vLLM for high‑throughput, low‑latency LLM inference in production environments. * Build scalable APIs using Python and integrate GenAI capabilities into enterprise ...

Python Developer

Charlotte, NC · On-site

$49 - $67.75/hr

Designing, developing, and deploying Python-based applications and systems according to client ... Experience with OpenAI, Azure OpenAI, Anthropic, Google Gemini, or other LLM platforms

Python Developer

Charlotte, NC

$49 - $67.75/hr

Designing, developing, and deploying Python-based applications and systems according to client ... Experience with OpenAI, Azure OpenAI, Anthropic, Google Gemini, or other LLM platforms

... LLM-based solutions), while refining algorithms to meet business needs, and ensure smooth ... Python to create reusable customizations for non-ML, ML, and deep learning algorithms, while ...

AI Engineer

Charlotte, NC · On-site

$51.59 - $61.59/hr

The ideal candidate will have deep expertise in Python-based AI development and hands-on experience ... Experience with LLM orchestration frameworks such as LangChain, LangGraph, and LlamaIndex. Proven ...

... LLM-based solutions), while refining algorithms to meet business needs, and ensure smooth ... Python to create reusable customizations for non-ML, ML, and deep learning algorithms, while ...

Strong Python: async programming, type-driven design, and comfort working in a strict mypy codebase. * Working experience with an LLM orchestration framework such as LangChain, LangGraph, or an ...

Gen AI Python Engineer

Charlotte, NC · On-site

$90 - $120/hr

Leverage tools like SAS and R/Python to create reusable customizations for non‑ML, ML, and deep ... Contribute to thought leadership such as papers, innovative non‑ML, ML, deep learning or LLM ...

Showing results 21-40

Python Llm information

See Charlotte, NC salary details

$12

$57

$84

How much do python llm jobs pay per hour?

As of Sep 5, 2026, the average hourly pay for python llm in Charlotte, NC is $57.26, according to ZipRecruiter salary data. Most workers in this role earn between $47.21 and $65.05 per hour, depending on experience, location, and employer.

What is a Python LLM?

A Python LLM job involves working with Large Language Models (LLMs) using Python to develop, fine-tune, and deploy AI models. Responsibilities may include data preprocessing, prompt engineering, model optimization, and integration with applications. Professionals in this role often work with frameworks like TensorFlow, PyTorch, or Hugging Face Transformers. They may also contribute to improving model efficiency, reducing bias, and ensuring ethical AI usage.

What are the key skills and qualifications needed to thrive in the Python LLM position, and why are they important?

To excel as a Python LLM (Large Language Model) Engineer, you need strong skills in Python programming, machine learning, and natural language processing, typically supported by a degree in computer science or a related field. Proficiency with libraries such as TensorFlow, PyTorch, Hugging Face Transformers, and experience with model deployment platforms are often essential, alongside certifications in AI or data science. Effective communication, problem-solving abilities, and collaboration are important soft skills for working in interdisciplinary teams and delivering results in dynamic environments. These skills ensure the development, fine-tuning, and deployment of advanced language models that meet both technical and business objectives.

What are some common challenges faced by Python LLM engineers in their daily work?

Python LLM Engineers often encounter challenges related to optimizing model performance, managing large datasets, and adapting models to specific business needs. Working with large-scale language models requires balancing computational resource limitations with the need for high accuracy and efficiency. Collaboration with data scientists, product managers, and DevOps engineers is routine to ensure seamless model integration and deployment. Staying updated on the latest advancements in NLP and continuously improving models based on user feedback are also important aspects of the role.

What are the most commonly searched types of Python Llm jobs in Charlotte, NC?

The most popular types of Python Llm jobs in Charlotte, NC are:

What cities near Charlotte, NC are hiring for Python Llm jobs?

Cities near Charlotte, NC with the most Python Llm job openings:

Infographic showing various Python Llm job openings in Charlotte, NC as of August 2026, with employment types broken down into 2% Internship, 83% Full Time, 8% Part Time, 1% Temporary, and 6% Contract. Highlights an 81% Physical, 5% Hybrid, and 14% Remote job distribution, with an average salary of $119,093 per year, or $57.3 per hour.

On-Prem LLM Platform Engineer (OpenShift AI / GPU)

Infosys

Charlotte, NC • On-site

Full-time

Re-posted 25 days ago


Infosys rating

7.0

Company rating: 7.0 out of 10

Based on 62 frontline employees who took The Breakroom Quiz

153rd of 226 rated it services


Job description

Job Summary:
Infosys is a global leader in next-generation digital services and consulting, and they are seeking a Data Science Consultant 1 to join their Data and Analytics unit. The role involves participating in data preparation, model development, and collaboration with technology teams to operationalize analytics solutions.
Responsibilities:
• Participate in data extraction, transformation, and preparation.
• Resolve common data issues and ensure quality for model development.
• Participate in developing models using statistical or machine learning techniques and collaborate with technology teams to operationalize them into analytics tools or scripts.
• Participate in model testing and validation, selecting the best-performing algorithms based on statistical and business metrics.
• Participate in the development of advanced analytics and machine learning or deep learning models including LLMs using predefined processes and tools like SAS and R/ Python.
• Participate in defining analytics problems; execute visualization, analysis, and predictive modeling with senior support.
• Identify data sources and extract from RDBMS and develop UI/UX for client usage.
• Participate in model performance, while making minor adjustments, and escalate risks or compliance concerns and generate reports on deviations or schedule slippages.
• Proactively participate in detailed documentation of model development, testing, and deployment activities for reproducibility.
• Work closely with business and technology teams to translate requirements into actionable models, while communicating results effectively.
• Apply predefined quality measurement frameworks, if any, to individual project tasks.
• Participate in deploying analytics tools in test and production environment, while ensuring they meet operational requirements.
Qualifications:
Required:
• Strong hands-on experience with OpenShift (OCP) and OpenShift AI
• Deep understanding of Kubernetes cluster operations and model serving
• Experience with vLLM, Triton, TensorRT-LLM, or SGLang
• Expertise in inference optimization techniques including: Proficiency in model quantization techniques (FP8, AWQ, GPTQ) and performance tuning
• Experience with GPU orchestration and scheduling
• Working knowledge of CUDA, NCCL, and MIG concepts
• Bachelor’s degree or foreign equivalent required from an accredited institution. Will also consider three years of progressive experience in the specialty in lieu of every year of education.
• This position may require relocation and/or travel to work/project location.
• Candidates authorized to work for any employer in the United States without employer-based visa sponsorship are welcome to apply. Infosys is unable to provide immigration sponsorship for this role now or in the future.
Preferred:
• Exposure to NVIDIA H200 GPUs (highly preferred)
• Deploy and run LLMs on-prem using: OpenShift (OCP), OpenShift AI, Kubernetes ML serving patterns
• Model serving stacks: vLLM, Triton, TensorRT-LLM, SGLang
• Unfold an LLM end-to-end: Package model artifacts, containerize, configure runtime, serve endpoints
• Implement authn/authz, secrets, networking, and endpoint routing patterns
• GPU-first platform engineering: GPU orchestration, scheduling, resource quotas
• Performance tuning for throughput/latency, memory utilization
• Awareness of CUDA/NCCL, tensor parallelism, MIG concepts
Company:
Infosys is a technology company that offers consulting, outsourcing, cloud infrastructure, program management, and software services. Founded in 1981, the company is headquartered in Bangalore, IND, with a team of 10001+ employees. The company is currently Late Stage.

What Infosys employees say

Pay

Benefits

Hours and flexibility

Workplace

Get the full story on Breakroom