1

Self Decode Jobs (NOW HIRING)

The OLIX Decode Accelerator 1 (DX-1) is the first accelerator architected specifically for decode ... Improve the developer platform: build and maintain CI/CD, internal tooling, and self-service ...

CNC Programmer II - 1st Shift

Irving, TX · On-site

$25.50 - $34.75/hr

Decode blueprints, sketches, drawings, and specifications like a pro to determine dimensions ... Detail-oriented and self-motivated**-you thrive with autonomy and minimal supervision, taking ...

CNC Programmer II - 1st Shift

Irving, TX

$25.50 - $34.75/hr

Decode blueprints, sketches, drawings, and specifications like a pro to determine dimensions ... Detail-oriented and self-motivated**-you thrive with autonomy and minimal supervision, taking ...

CNC Programmer II - 1st Shift

Irving, TX · On-site

$25.50 - $34.75/hr

Decode blueprints, sketches, drawings, and specifications like a pro to determine dimensions ... Detail-oriented and self-motivated -you thrive with autonomy and minimal supervision, taking ...

AMS Engineer

Austin, TX · On-site

$264K/yr

The OLIX Decode Accelerator 1 (DX-1) is the first accelerator architected specifically for decode ... Build self-checking testbenches and reusable verification components for analog/digital ...

Electrical Automation Technician

Kinston, NC · On-site

$39K/yr

Decode complex electrical schematics and user manuals to troubleshoot and maintain cutting-edge ... Must be self-motivated and able to work with minimal supervision. * Effectively communicate the ...

Electrical Automation Technician

Kinston, NC · On-site

$38K/yr

Position Responsibilities: · Decode complex electrical schematics and user manuals to troubleshoot ... Must be self-motivated and able to work with minimal supervision. * Effectively communicate the ...

Electrical Automation Technician

Kinston, NC · On-site

$38K/yr

Position Responsibilities: • Decode complex electrical schematics and user manuals to ... Must be self-motivated and able to work with minimal supervision. * Effectively communicate the ...

Research Engineer, Multimodal Data

San Francisco, CA · On-site

$134K - $162K/yr

... decode pipelining -- so "annotate every clip in a 10K-hour corpus" stays economical. • Build the ... self-driving, robotics, or visual-data company -- or equivalent depth from a research lab. • ...

next page

Showing results 1-20

Self Decode information

See salary details

$46K

$115.4K

$172.5K

How much do self decode jobs pay per year?

As of Aug 21, 2026, the average yearly pay for self decode in the United States is $115,438.00, according to ZipRecruiter salary data. Most workers in this role earn between $80,500.00 and $134,000.00 per year, depending on experience, location, and employer.
Infographic showing various Self Decode job openings in the United States as of August 2026, with employment types broken down into 1% As Needed, 78% Full Time, 18% Part Time, and 3% Contract. Highlights an 91% Physical, 2% Hybrid, and 7% Remote job distribution, with an average salary of $115,438 per year, or $55.5 per hour.

LLM Interface GPU Consultant

Pacific Consultancy Services

Charlotte, NC • On-site

Other

Posted 15 days ago


Job description

Job Title: LLM Inference & GPU Systems Consultant

Location: Charlotte, NC (Hybrid)

Role Overview:
We are seeking an AI Infrastructure Runtime Engineer to build and maintain large-scale on-prem LLM infrastructure. This is an enterprise private GenAI environment running on NVIDIA H200 GPU clusters and an OpenShift AI deployment ecosystem. You will manage production inference internally, including self-hosting open-source LLMs like Llama. We are focused exclusively on inferencing; this role involves no model training infrastructure or fine-tuning pipelines.
Key Responsibilities
NVIDIA GPU Runtime Optimization: Drive extreme runtime efficiency and optimization for the token generation pipeline. Specifically manage prefill/decode optimization and KV cache management.
Inference Serving: Deploy and manage inference engines including vLLM and TensorRT-LLM.
Hardware Utilization: Optimize GPU throughput tuning, batching strategies, and latency optimization. Manage workload orchestration using RunAI and Kubernetes GPU orchestration.
Model Lifecycle Management: Oversee the complete Hugging Face model lifecycle, including model onboarding, deployment, and retirement.
Platform Operations: Operate and maintain the OpenShift AI ecosystem as the primary container platform for GenAI workloads.
Required Qualifications
8+ years experience working as an LLM Systems Engineer or AI Infrastructure Runtime Engineer.
8+ years hands-on experience with NVIDIA H200 clusters and runtime optimization techniques (KV Cache, prefill/decode).
Proficiency in OpenShift AI and GPU orchestration tools like RunAI.
Strong experience with modern inference frameworks, specifically vLLM and TensorRT-LLM.
Proven track record managing the Hugging Face deployment lifecycle.