1

Data Preprocessing Jobs in Lewisville, TX (NOW HIRING)

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Experience with data preprocessing, feature engineering, SQL, and data manipulation. * Familiarity with AI model evaluation, optimization, and hyperparameter tuning. * Strong analytical, problem ...

Showing results 21-40

Data Preprocessing information

See Lewisville, TX salary details

$43K

$154.1K

$227.4K

How much do data preprocessing jobs pay per year?

As of Aug 22, 2026, the average yearly pay for data preprocessing in Lewisville, TX is $154,126.00, according to ZipRecruiter salary data. Most workers in this role earn between $124,700.00 and $158,800.00 per year, depending on experience, location, and employer.

What is data preprocessing?

Data preprocessing is the process of cleaning, transforming, and organizing raw data into a usable format for analysis or machine learning. It involves steps such as handling missing values, removing duplicates, normalizing or scaling data, and encoding categorical variables. Proper data preprocessing helps improve the quality and performance of predictive models by ensuring the data is accurate, consistent, and suitable for analysis.

What are the key skills and qualifications needed to thrive as a data preprocessing specialist, and why are they important?

To thrive as a Data Preprocessing Specialist, you need a strong background in statistics, data cleaning, and data transformation, often supported by a degree in computer science, data science, or a related field. Proficiency with tools such as Python (pandas, NumPy), SQL, and data visualization platforms is typically essential, along with familiarity with data management systems. Attention to detail, problem-solving abilities, and effective communication are standout soft skills in this position. These skills are crucial for ensuring high-quality, reliable datasets that underpin accurate data analysis and machine learning outcomes.

What are some common challenges faced in a data preprocessing role, and how can they be effectively managed?

Professionals in Data Preprocessing often encounter challenges such as handling incomplete or inconsistent data, managing large datasets, and ensuring data quality before analysis. Addressing these issues typically involves using specialized tools to automate data cleaning, establishing clear data validation rules, and collaborating closely with data engineers and analysts. Staying updated with best practices and leveraging scripting languages like Python or R can also streamline the preprocessing workflow, making it easier to deliver reliable and accurate datasets for downstream analysis.

What is the difference between Data Preprocessing vs Data Analysis?

AspectData PreprocessingData Analysis
Primary FocusCleaning, transforming, and preparing raw data for analysisInterpreting data to extract insights and support decision-making
Skills RequiredData cleaning, scripting, understanding of data formatsStatistical analysis, data visualization, critical thinking
Work EnvironmentData engineering teams, data science projectsBusiness intelligence, research, data science teams
Tools UsedPython, R, SQL, ETL toolsExcel, Tableau, R, Python, statistical software

While data preprocessing involves preparing raw data for analysis by cleaning and transforming it, data analysis focuses on interpreting the prepared data to uncover trends and insights. Both roles are essential in the data pipeline but serve different purposes in the data lifecycle.

What are popular job titles related to Data Preprocessing jobs in Lewisville, TX?

For Data Preprocessing jobs in Lewisville, TX, the most frequently searched job titles are:

What job categories do people searching Data Preprocessing jobs in Lewisville, TX look for?

The top searched job categories for Data Preprocessing jobs in Lewisville, TX are:

Infographic showing various Data Preprocessing job openings in Lewisville, TX as of August 2026, with employment types broken down into 1% As Needed, 82% Full Time, 13% Part Time, and 4% Contract. Highlights an 87% Physical, 3% Hybrid, and 10% Remote job distribution, with an average salary of $154,126 per year, or $74.1 per hour.

Full-time

Re-posted 4 days ago


Job description

Job Title: GenAI Architect
Duration: 8+ months
Location: Dallas, TX / Tampa, FL
Job Description:
RESPONSIBILITIES
  • Design, develop, and implement Generative AI models and algorithms, using Huggingface LLM models such as Falcon, Llama2.
  • Collaborate with cross-functional teams to define AI project requirements and objectives, ensuring alignment with overall business goals.
  • Design, develop, and implement Generative AI use cases using Falcon, Llama2
  • Conduct research to stay up-to-date with the latest advancements in Generative AI, machine learning, and deep learning techniques, and identify opportunities to integrate them into customer products and services.
  • Design, develop, and implement NLP and Machine learning models
  • Productionize models - model training, model deployment, model serving, model monitoring
  • Optimize Generative AI models for improved performance, scalability, and efficiency.
  • Develop and maintain AI pipelines, including data preprocessing, feature extraction, model training, and evaluation.
  • Contribute to the establishment of best practices and standards for Generative AI development within the organization.
  • Function as a technical or project lead and while mentoring junior engineering talent

REQUIREMENTS
  • Bachelor/Master degree in Computer Science, Artificial Intelligence, Machine Learning, or a related field with 10+ years of experience.
  • 2+ years of experience in developing and implementing Generative AI models, with a strong understanding of techniques such as Falcon, Llama2.
  • Must have 5+ Telco experience
  • Proficiency in Python and have 8+ years of experience with machine learning libraries and frameworks such as TensorFlow, PyTorch, or Keras.
  • Strong knowledge of data structures, algorithms, and software engineering principles.
  • 8+ years of experience with advanced natural language processing (NLP) techniques and tools, such as SpaCy, NLTK, or Hugging Face.
  • Familiarity with data visualization tools and libraries, such as Matplotlib, Seaborn, or Plotly.
  • Azure cloud experience - deployment, MLOps
  • Significant experience architecting cutting-edge MLOps systems in enterprise environments
  • Knowledge of software development methodologies, such as Agile or Scrum.
  • Excellent problem-solving skills, with the ability to think critically and creatively to develop innovative AI solutions.
  • Strong communication skills, with the ability to effectively convey complex technical concepts to a diverse audience.
  • Proactive mindset, with the ability to work independently and collaboratively in a fast-paced, dynamic environment.