1

Probabilistic Modeling Jobs in Atlanta, GA (NOW HIRING)

Senior AI/ML Engineer

Atlanta, GA ยท On-site

$100K - $138K/yr

Implement probabilistic matching techniques (e.g., Fellegi-Sunter) and ML models (gradient boosting, neural classifiers) for record linkage across the US adult population * Build candidate blocking ...

Comfort with probabilistic and statistical modeling of physical processes. * Experience with real-time systems and distributed architecture. * Experience with robotics, industrial automation, and ...

New

Comfort with probabilistic and statistical modeling of physical processes. * Experience with real-time systems and distributed architecture. * Experience with robotics, industrial automation, and ...

Genetics Tutor

Duluth, GA ยท Remote

$18 - $40/hr

Ability to explain linkage analysis, Hardy-Weinberg equilibrium, and gene regulation models while ... Emphasizes probabilistic reasoning and connects genetics to genetic counseling, forensic science ...

Senior AI Engineer

Atlanta, GA ยท On-site

$120 - $160/hr

Select and evaluate models (hosted vs open-source) based on use case constraints * Agent-Based ... Ability to debug complex issues, including probabilistic outputs * Comfort working with APIs ...

Senior AI Engineer

Atlanta, GA ยท On-site

$100K - $138K/yr

Select and evaluate models (hosted vs open-source) based on use case constraints * Agent-Based ... Ability to debug complex issues, including probabilistic outputs * Comfort working with APIs ...

Data Platform Engineer

Atlanta, GA ยท Remote

$117K - $140K/yr

Proven experience with Entity Resolution or Record Linkage (e.g., using tools like Senzing, Quantexa, or custom probabilistic matching models) * Schema Design: Ability to design flexible ontologies ...

Genetics Tutor

Johns Creek, GA ยท Remote

$18 - $40/hr

Ability to explain linkage analysis, Hardy-Weinberg equilibrium, and gene regulation models while ... Emphasizes probabilistic reasoning and connects genetics to genetic counseling, forensic science ...

Genetics Tutor

Alpharetta, GA ยท Remote

$18 - $40/hr

Ability to explain linkage analysis, Hardy-Weinberg equilibrium, and gene regulation models while ... Emphasizes probabilistic reasoning and connects genetics to genetic counseling, forensic science ...

Genetics Tutor

Marietta, GA ยท Remote

$18 - $40/hr

Ability to explain linkage analysis, Hardy-Weinberg equilibrium, and gene regulation models while ... Emphasizes probabilistic reasoning and connects genetics to genetic counseling, forensic science ...

Genetics Tutor

Woodstock, GA ยท Remote

$18 - $40/hr

Ability to explain linkage analysis, Hardy-Weinberg equilibrium, and gene regulation models while ... Emphasizes probabilistic reasoning and connects genetics to genetic counseling, forensic science ...

Genetics Tutor

Lawrenceville, GA ยท Remote

$18 - $40/hr

Ability to explain linkage analysis, Hardy-Weinberg equilibrium, and gene regulation models while ... Emphasizes probabilistic reasoning and connects genetics to genetic counseling, forensic science ...

Genetics Tutor

Atlanta, GA ยท Remote

$18 - $40/hr

Ability to explain linkage analysis, Hardy-Weinberg equilibrium, and gene regulation models while ... Emphasizes probabilistic reasoning and connects genetics to genetic counseling, forensic science ...

Showing results 41-60

Probabilistic Modeling information

What is probabilistic modeling?

Probabilistic modeling is a mathematical framework used to represent uncertain events or data by using probability distributions. Instead of giving a single outcome, it accounts for variability and randomness, allowing predictions and inferences even when information is incomplete or ambiguous. Probabilistic models are widely used in fields like statistics, machine learning, finance, and engineering to analyze data, make forecasts, and support decision-making under uncertainty.

What are the key skills and qualifications needed to thrive as a probabilistic modeler, and why are they important?

To thrive as a Probabilistic Modeler, you need a strong background in mathematics, statistics, and probability theory, often supported by a degree in applied mathematics, statistics, or a related field. Proficiency with programming languages like Python or R, and experience with statistical modeling tools and software such as TensorFlow or PyMC, are typically required. Strong analytical thinking, problem-solving abilities, and effective communication skills help translate complex models into actionable insights. These skills are vital for designing accurate models, interpreting uncertainty, and supporting data-driven decisions across various industries.

What are some common challenges faced by professionals in probabilistic modeling roles, and how can they be managed?

Professionals in probabilistic modeling often encounter challenges such as working with incomplete or noisy data, choosing the right model complexity, and ensuring model interpretability for stakeholders. Managing these challenges involves strong statistical knowledge, regular collaboration with domain experts, and effective communication to translate complex results for non-technical team members. Staying up-to-date with the latest tools and methodologies, and participating in peer reviews, can also help maintain model accuracy and reliability.

What is the difference between Probabilistic Modeling vs Data Scientist?

AspectProbabilistic ModelingData Scientist
Required CredentialsDegree in statistics, mathematics, or related fields; knowledge of probability theoryDegree in computer science, statistics, or related fields; programming skills
Work EnvironmentResearch-focused, often in analytics or data science teamsCross-functional teams, including business, engineering, and analytics
Industry UsageUsed in analytics, finance, healthcare, and research for modeling uncertaintyApplied across industries for data analysis, predictive modeling, and decision-making

Probabilistic Modeling focuses on developing models based on probability theory to understand uncertainty, while Data Scientists utilize a broader set of skills including programming, data analysis, and machine learning to extract insights from data. Both roles often overlap but serve different primary purposes within data-driven organizations.

What job categories do people searching Probabilistic Modeling jobs in Atlanta, GA look for?

The top searched job categories for Probabilistic Modeling jobs in Atlanta, GA are:

Senior AI/ML Engineer

Sumeru

Atlanta, GA โ€ข On-site

$100K - $138K/yr

Other

Re-posted 6 days ago


Job description

Role: Senior AI/ML Engineer
Location: Bellevue/Seattle, WA ; Atlanta, GA, and Frisco, TX


Need Local Candidates


Job Overview

We are seeking an AI/ML Engineer to build the intelligent systems that power identity resolution and data accessibility within our Customer Data Platform (CDP) - the authoritative source of truth for customer data across the entire US adult population.

This role focuses on developing machine learning pipelines that deduplicate, link, and resolve customer identities across disparate data sources - the core capability that transforms raw data into trusted, unified customer profiles. You will also contribute to LLM-based solutions that enable natural language querying of CDP data, making the platform accessible to business users across the organization.

You will work on both classical ML techniques and modern LLM-based approaches to ensure that every customer identity in CDP is accurately resolved, every profile is trustworthy, and every user can access the data they need.

Job Responsibilities - Identity Resolution

  • Develop and deploy entity resolution models to match and deduplicate customer records across multiple systems - directly impacting the accuracy of CDP as the source of truth
  • Implement probabilistic matching techniques (e.g., Fellegi-Sunter) and ML models (gradient boosting, neural classifiers) for record linkage across the US adult population
  • Build candidate blocking pipelines using phonetic algorithms (Soundex, Double Metaphone), token similarity, and LSH to handle billions of potential match pairs efficiently
  • Apply fuzzy matching techniques (Levenshtein, Jaro-Winkler, Jaccard) for customer attributes such as name, address, phone, and identifiers
  • Develop clustering algorithms (DBSCAN, hierarchical clustering) to create unified "golden customer profiles" that serve as the authoritative representation of each individual
  • Build embedding-based similarity systems using Sentence-BERT or transformer-based models for semantic matching
  • Implement ANN/KNN retrieval systems (FAISS, Annoy) for large-scale entity matching across population-scale datasets

Job Responsibilities - AI/LLM

  • Use LLMs (e.g., GPT, Claude) for classification and disambiguation of entity matches, improving resolution accuracy where traditional methods fall short
  • Build and support RAG pipelines to enrich customer profiles with contextual data from unstructured sources
  • Perform prompt engineering and evaluation for structured data extraction from unstructured inputs feeding into CDP
  • Contribute to NLQ-to-SQL systems, enabling business users to query CDP data using natural language - making the authoritative source of truth accessible to non-technical stakeholders
  • Support integration with vector databases (e.g., Pinecone, pgvector, Qdrant) for semantic search across customer data

Education and Work Experience

  • Bachelor's or Master's degree in Computer Science, Data Science, or related field
  • 3+ years of experience in ML/AI engineering
  • At least 1 year of experience in entity resolution, record linkage, or deduplication - ideally at scale

Technical Skills

  • Programming: Python (required)
  • Libraries: scikit-learn, HuggingFace Transformers, RapidFuzz, jellyfish
  • Experience with LLM APIs (OpenAI, Anthropic) and prompt pipelines
  • Strong SQL skills and experience with Spark or Dask for distributed processing
  • Familiarity with vector databases and embedding-based retrieval
  • Experience with ML lifecycle tools (MLflow or similar)
  • Understanding of data quality metrics and how identity resolution impacts downstream trust

Knowledge, Skills, and Abilities

  • Strong understanding of ML fundamentals and similarity matching techniques applied to customer identity
  • Ability to work with large, messy, real-world datasets spanning hundreds of millions of records
  • Understanding of precision/recall tradeoffs in identity resolution and their impact on data trust
  • Good problem-solving and analytical skills
  • Ability to collaborate with data engineering, platform, and business teams to deliver accurate customer profiles

Sumeru logo

About Sumeru

Sourced by ZipRecruiter

Industry

It services

Company size

501 - 1,000 Employees

Headquarters location

Washington, DC, US

Year founded

2002