This is a remote role. Candidates who live near CB offices have the option of being fully remote or ... of large language model (LLM) systems for feedback generation and annotation, and research to ...
This is a remote role. Candidates who live near CB offices have the option of being fully remote or ... of large language model (LLM) systems for feedback generation and annotation, and research to ...
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Deploy LLM solutions across cloud-based and local resources using kubernetes (llama.ccp, vllm etc)
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Deploy LLM solutions across cloud-based and local resources using kubernetes (llama.ccp, vllm etc)
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Deploy LLM solutions across cloud-based and local resources using kubernetes (llama.ccp, vllm etc)
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Deploy LLM solutions across cloud-based and local resources using kubernetes (llama.ccp, vllm etc)
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Deploy LLM solutions across cloud-based and local resources using kubernetes (llama.ccp, vllm etc)
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Deploy LLM solutions across cloud-based and local resources using kubernetes (llama.ccp, vllm etc)
AI Architect
OR · Remote
Artificial and Large Language Model Architect As an AI & LLM Architect , you will play a pivotal role in designing and implementing the technology architecture for advanced AI (including Large ...
Quick apply
AI Architect
OR · Remote
Artificial and Large Language Model Architect As an AI & LLM Architect , you will play a pivotal role in designing and implementing the technology architecture for advanced AI (including Large ...
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Deploy LLM solutions across cloud-based and local resources using kubernetes (llama.ccp, vllm etc)
Apply Large Language Models (LLMs) to a variety of applications within remote sensing such as ... Deploy LLM solutions across cloud-based and local resources using kubernetes (llama.ccp, vllm etc)
Remote AI / Machine Learning Specialist
Nashville, TN · Remote
$115K - $140K/yr
In this fully remote role, you'll work with engineering, product, and business teams to develop ... Develop and optimize large language model (LLM) applications and AI-powered workflows.
Quick apply
Remote AI / Machine Learning Specialist
Nashville, TN · Remote
$115K - $140K/yr
In this fully remote role, you'll work with engineering, product, and business teams to develop ... Develop and optimize large language model (LLM) applications and AI-powered workflows.
AI Solutions Engineer - Remote / Telecommute
Aliso Viejo, CA · Remote
$51 - $56/hr
Experience designing and implementing Large Language Model (LLM) applications utilizing platforms such as OpenAI, Anthropic Claude, Gemini, or Llama. * Proven experience building RAG (Retrieval ...
Quick apply
AI Solutions Engineer - Remote / Telecommute
Aliso Viejo, CA · Remote
$51 - $56/hr
Experience designing and implementing Large Language Model (LLM) applications utilizing platforms such as OpenAI, Anthropic Claude, Gemini, or Llama. * Proven experience building RAG (Retrieval ...
... Large Language Model (LLM) solution to check coherence and consistency. Resource will also develop and modify existing models related to customer long term engagement and retention. This role will ...
... Large Language Model (LLM) solution to check coherence and consistency. Resource will also develop and modify existing models related to customer long term engagement and retention. This role will ...
Build production-grade LLM applications and agentic systems * Deploy scalable AI solutions across ... Large Language Models and experience identifying ways to incorporate them into new domains and ...
Build production-grade LLM applications and agentic systems * Deploy scalable AI solutions across ... Large Language Models and experience identifying ways to incorporate them into new domains and ...
Build production-grade LLM applications and agentic systems * Deploy scalable AI solutions across ... Large Language Models and experience identifying ways to incorporate them into new domains and ...
Build production-grade LLM applications and agentic systems * Deploy scalable AI solutions across ... Large Language Models and experience identifying ways to incorporate them into new domains and ...
AI Engineer
San Mateo, CA · Remote
Hands-on experience with large language model (LLM) inference and/or model training using open ... Remote-friendly within the United States * Preference for candidates located near major East or ...
Quick apply
AI Engineer
San Mateo, CA · Remote
Hands-on experience with large language model (LLM) inference and/or model training using open ... Remote-friendly within the United States * Preference for candidates located near major East or ...
Staff AI/ML Engineer (Large Language Model) (TS/SCI) {S}
Aurora, CO · On-site +1
$150K - $200K/yr
Build production-grade LLM applications and agentic systems * Deploy scalable AI solutions across ... Large Language Models and experience identifying ways to incorporate them into new domains and ...
Staff AI/ML Engineer (Large Language Model) (TS/SCI) {S}
Aurora, CO · On-site +1
$150K - $200K/yr
Build production-grade LLM applications and agentic systems * Deploy scalable AI solutions across ... Large Language Models and experience identifying ways to incorporate them into new domains and ...
Staff AI/ML Engineer (Large Language Model) (TS/SCI) {S}
Aurora, CO · On-site +1
$150K - $200K/yr
Build production-grade LLM applications and agentic systems * Deploy scalable AI solutions across ... Large Language Models and experience identifying ways to incorporate them into new domains and ...
Staff AI/ML Engineer (Large Language Model) (TS/SCI) {S}
Aurora, CO · On-site +1
$150K - $200K/yr
Build production-grade LLM applications and agentic systems * Deploy scalable AI solutions across ... Large Language Models and experience identifying ways to incorporate them into new domains and ...
AI Technical Lead
Meridian, TX · On-site +1
... remote within a mutually acceptable location. #LI-Hybrid Success Looks Like: * AI systems move ... Design and implement AI-powered applications including large language model (LLM) systems and ...
AI Technical Lead
Meridian, TX · On-site +1
... remote within a mutually acceptable location. #LI-Hybrid Success Looks Like: * AI systems move ... Design and implement AI-powered applications including large language model (LLM) systems and ...
Full Stack Developer (Remote Opportunity)
$105K - $115K/yr
Integrate AI and Large Language Model (LLM) capabilities using platforms and APIs such as OpenAI ... Remote work options Note: Selected candidates will be required to complete fingerprinting at a ...
Full Stack Developer (Remote Opportunity)
$105K - $115K/yr
Integrate AI and Large Language Model (LLM) capabilities using platforms and APIs such as OpenAI ... Remote work options Note: Selected candidates will be required to complete fingerprinting at a ...
... or Large Language Model (LLM) capabilities into enterprise applications. • Experience working ... remote environment. • Strong organizational skills with exceptional attention to detail.
... or Large Language Model (LLM) capabilities into enterprise applications. • Experience working ... remote environment. • Strong organizational skills with exceptional attention to detail.
California (Remote) We are looking for devs with general cloud services / distributed services ... Experience working with Large Language Models (LLMs), particularly hosting them to run inference
Quick apply
California (Remote) We are looking for devs with general cloud services / distributed services ... Experience working with Large Language Models (LLMs), particularly hosting them to run inference
Experience integrating AI or Large Language Model (LLM) capabilities into enterprise applications ... Ability to collaborate effectively with cross-functional teams in a remote environment. * Strong ...
Experience integrating AI or Large Language Model (LLM) capabilities into enterprise applications ... Ability to collaborate effectively with cross-functional teams in a remote environment. * Strong ...
Experience integrating AI or Large Language Model (LLM) capabilities into enterprise applications ... Ability to collaborate effectively with cross-functional teams in a remote environment. * Strong ...
Experience integrating AI or Large Language Model (LLM) capabilities into enterprise applications ... Ability to collaborate effectively with cross-functional teams in a remote environment. * Strong ...
Remote Large Language Model Llm information
See salary details
$14.66 - $16.83
5% of jobs
$18.83 is the 25th percentile. Wages below this are outliers.
$16.83 - $18.99
21% of jobs
The median wage is $21.15 / hr.
$18.99 - $21.15
23% of jobs
$21.15 - $23.32
7% of jobs
$23.32 - $25.48
11% of jobs
$27.49 is the 75th percentile. Wages above this are outliers.
$25.48 - $27.64
7% of jobs
$27.64 - $29.81
6% of jobs
$29.81 - $31.97
5% of jobs
$31.97 - $34.13
5% of jobs
$34.13 - $36.30
4% of jobs
$36.30 - $38.46
3% of jobs
$14
$24
$38
How much do remote large language model llm jobs pay per hour?
What is a remote large language model LLM?
How does a remote large language model LLM engineer typically collaborate with cross-functional teams while working remotely?
What are the key skills and qualifications needed to thrive as a remote large language model LLM engineer?
What is the difference between Remote Large Language Model Llm vs Data Scientist?
| Aspect | Remote Large Language Model Llm | Data Scientist |
|---|---|---|
| Required Credentials | Advanced degrees in AI, NLP, or related fields; experience with machine learning frameworks | Degree in Data Science, Statistics, Computer Science, or related fields; strong analytical skills |
| Work Environment | Primarily remote, focused on developing and fine-tuning language models | Remote or on-site, analyzing data, building models, and generating insights |
| Employer & Industry Usage | Tech companies, AI research labs, startups working on NLP products | Tech firms, finance, healthcare, marketing, and research organizations |
While both roles involve data and machine learning, a Remote Large Language Model Llm specializes in developing and refining language models, whereas a Data Scientist focuses on analyzing data, building predictive models, and deriving insights across various domains.

Job description
College Board - Learning & Assessment
Location:
- This is a remote role. Candidates who live near CB offices have the option of being fully remote or hybrid (Tuesday and Wednesday in office). All CB employees are required to occasionally travel to meet in person for business purposes.
Role Type:
- This is a full-time position
About the Team
The Automated Scoring team provides critical insights and tools to support the design, delivery, and continuous improvement of digital assessments. We operate at the intersection of educational measurement, data science, and emerging AI technologies. Our work spans the measurement of language-based constructs, the development of large language model (LLM) systems for feedback generation and annotation, and research to ensure the validity, fairness, and reliability of our systems.
We are a collaborative, mission-driven team that values both psychometric rigor and technical expertise. We combine modern machine learning approaches, including large language models, with strong measurement principles to create scalable, trustworthy solutions that expand opportunity for students.
About the Opportunity
As a Natural Language Specialist, you will help define and advance how language-based performance is measured in high-stakes educational settings. This role sits at the intersection of natural language processing, large language models, and educational measurement, and is ideal for someone who pairs strong measurement training with a working knowledge of modern AI.
You will translate measurement constructs into LLM-based feedback and annotation systems, design and conduct the studies that establish their validity, fairness, and reliability, and ensure that what our models produce holds up to rigorous psychometric standards. You will serve as the measurement authority for cross-functional partners-psychometricians, engineers, and data scientists-shaping how language-based constructs are defined, evaluated, and applied across the team's portfolio of systems. Your primary lens will be measurement: defining what feedback and annotations mean, evidencing that they mean it, and improving them over time.
In this role, you will:
Natural Language Measurement & Psychometrics (40%)
- Define and operationalize language-based constructs for automated annotation and feedback generation
- Apply psychometric principles-reliability, validity, dimensionality, and measurement invariance-to LLM-based feedback and annotation systems
- Design and lead validity studies, including human-machine agreement, rater comparison, and fairness analyses across subgroups
- Develop and apply methods for detecting and mitigating bias in language-based scores
- Establish/Recommend annotation guidelines, feedback quality criteria, and standards for acceptable model performance
- Translate measurement requirements into specifications that guide model development and evaluation
LLM & AI Development (30%)
- Contribute to prompt design, fine-tuning, and evaluation of LLM-based feedback and annotation systems
- Develop and refine machine learning models for measuring language-based constructs
- Build evaluation frameworks that connect model behavior to measurement outcomes
- Collaborate with senior team members to translate measurement findings into production systems
- Implement high-quality, maintainable code for model development and evaluation
Research & Validation (10%)
- Lead and contribute to research studies that evaluate model performance and support assessment validity
- Apply statistical and psychometric methods to analyze results and inform model improvements
- Document methodologies and findings in a clear and rigorous manner
- Stay current with advances in educational measurement, NLP, and learning science
Data Engineering & Pipelines (10%)
- Prepare and curate datasets for measurement studies and model evaluation
- Support reproducible data processing workflows for training, evaluation, and monitoring
- Partner with engineers to integrate feedback and annotation models into scalable systems
Team Operations & Collaboration (10%)
- Collaborate closely with psychometricians, data scientists, and engineers
- Contribute to documentation, methodological standards, and team best practices
- Participate in peer reviews and knowledge sharing
- Actively raise the measurement literacy of the broader team-mentoring junior ICs, providing technical feedback on colleagues' work, and building shared standards
About You
You bring strong measurement training and a genuine interest in how modern AI can be used to measure language-based performance. You are excited about applying psychometric rigor to large language models in high-stakes educational settings.
You have:
- A Master's or PhD (or near completion) in a quantitative field such as Psychometrics, Educational Measurement, Quantitative Psychology, Statistics, Data Science, or a related discipline (measurement-focused training strongly valued)
- A solid foundation in measurement theory, including reliability, validity, and fairness; familiarity with IRT, generalizability theory, or related frameworks is highly desirable
- Experience analyzing language or text data, with exposure to NLP or large language models
- Programming skills in Python and familiarity with data science libraries (e.g., pandas, NumPy, PyTorch)
- Demonstrated ability to conduct rigorous research (e.g., thesis, publications, or applied research projects)
- Strong skills in statistical analysis and experimental design
- Familiarity with working with structured and unstructured data
- Strong attention to detail and a commitment to producing high-quality, reproducible work
- The ability to travel 5-10 times a year to College Board offices or on behalf of College Board business.
You are:
- Curious and eager to learn new tools, methods, and domains
- Thoughtful about the implications of AI systems, including fairness and validity
- Able to communicate technical and measurement concepts clearly to diverse audiences
- Comfortable working in a collaborative, cross-functional environment
- Motivated by mission-driven work in education
All roles at College Board require:
- A passion for expanding educational and career opportunities and mission-driven work
- Curiosity and enthusiasm for emerging technologies, with a willingness to experiment with and adopt new AI-driven solutions and comfort with learning and applying new digital tools independently and proactively.
- Clear and concise communication skills, written and verbal
- A learner's mindset and a commitment to growth: welcoming diverse perspectives, giving and receiving timely, respectful feedback, and continuously improving through iterative learning and user input.
- A drive for impact and excellence: solving complex problems, making data-informed decisions, prioritizing what matters most, and continuously improving through learning, user input, and external benchmarking.
- A collaborative and empathetic approach: working across differences, fostering trust, and contributing to a culture of shared success
- Authorization to work in the United States
About Our Process
- Application review will begin immediately and will continue until the position is filled. This role is expected to accept applications for a minimum of 5 business days.
- While the hiring process may vary, it generally includes: resume and application submission, recruiter phone/video screen, hiring manager interview, performance exercise such as live coding, a panel interview, a conversation with leadership and reference checks.
What We Offer
At College Board, we offer more than a paycheck- we provide a meaningful career, a supportive team, and a comprehensive package designed to help you thrive. We're a self-sustaining nonprofit that believes in fair and competitive compensation grounded in your qualifications, experience, impact, and the market.
A Thoughtful Approach to Compensation
- The hiring range for this role is $88,000-$145,000.
- Your exact salary will depend on your location, experience, and how your background compares to others in similar roles at the College Board.
- We aim to make our best offer upfront, rooted in fairness, transparency, and market data.
- We adjust salaries by location to ensure fairness, no matter where you live.
You'll have open, transparent conversations about compensation, benefits, and what it's like to work at College Board throughout your hiring process. Check out our careers page for more.
About College Board
Sourced by ZipRecruiter
Industry
Education programs administration
Company size
1,001 - 5,000 Employees
Headquarters location
New York, NY, US
Year founded
1900