2

Remote Data Pipeline Jobs in Rochester, NY (NOW HIRING)

AI Data Architect

Rochester, NY · On-site +1

$150K - $200K/yr

... pipelines to data warehouses and RAG-powered AI systems. You'll own the full data stack across a ... This can be a remote opportunity, with 2 weeks of travel into Rochester, NY per quarter What You'll ...

AI Data Architect

Rochester, NY · Remote

$150K - $200K/yr

... pipelines to data warehouses and RAG-powered AI systems. You'll own the full data stack across a ... This can be a remote opportunity, with 2 weeks of travel into Rochester, NY per quarter What You'll ...

Engineering Lead - Remote

Rochester, NY · Remote

$105K - $155K/yr

Improve engineering practices by enhancing CI/CD pipelines, testing strategies, release processes ... Drive architectural decisions around APIs, web applications, data access, cloud technologies, and ...

Engineering Lead - Remote

Rochester, NY · Remote

$105K - $155K/yr

Improve engineering practices by enhancing CI/CD pipelines, testing strategies, release processes ... Drive architectural decisions around APIs, web applications, data access, cloud technologies, and ...

Engineering Lead - Remote

Rochester, NY · On-site +1

$105K - $155K/yr

Improve engineering practices by enhancing CI/CD pipelines, testing strategies, release processes ... Drive architectural decisions around APIs, web applications, data access, cloud technologies, and ...

... Remote Security Manger app. Fairport, NY is the headquarters for the Radionix sales & marketing ... Scope of Ownership: * CRM: HubSpot CRM * Sales pipeline, customer data, activity management ...

next page

Showing results 1-20

Remote Data Pipeline information

What is the difference between Remote Data Pipeline vs Data Engineer?

AspectRemote Data PipelineData Engineer
Required SkillsData integration, ETL tools, scriptingDatabase management, programming, data architecture
Work EnvironmentRemote, cloud-based platformsOn-site or remote, enterprise environments
CertificationsCloud certifications, data toolsData management, cloud certifications
Industry UsageTech, finance, healthcareTech, finance, healthcare

Remote Data Pipelines focus on building and maintaining automated data workflows in cloud environments, often requiring scripting and data integration skills. Data Engineers design and develop comprehensive data systems, including databases and data architecture, with broader responsibilities. Both roles are vital in data-driven industries and often overlap, but Data Engineers typically have a wider scope and require more extensive technical expertise.

What cities near Rochester, NY are hiring for Remote Data Pipeline jobs?

Cities near Rochester, NY with the most Remote Data Pipeline job openings:

Infographic showing various Remote Data Pipeline job openings in Rochester, NY as of August 2026, with employment types broken down into 1% As Needed, 82% Full Time, 13% Part Time, and 4% Contract. Highlights an 87% Physical, 3% Hybrid, and 10% Remote job distribution.

AI Data Architect

Innovative Solutions

Rochester, NY • On-site, Remote

$150K - $200K/yr

Full-time

Re-posted 12 days ago


Key responsibilities

  • Design and build data architectures on AWS, including data lakes, data warehouses, and AI systems.

  • Implement data processing pipelines, data modeling, and schema design for various data platforms.

  • Lead architecture discussions, present recommendations, and mentor team members on data architecture patterns.


Job description

As a Data/AI Architect, you'll design and build data-driven cloud architectures on AWS - from S3 data lakes and Glue ETL pipelines to data warehouses and RAG-powered AI systems. You'll own the full data stack across a variety of industries and projects: one engagement you're designing a Redshift data warehouse with medallion architecture processing 31M transactions/month, the next you're building a Bedrock Knowledge Base with OpenSearch vector search. Real ownership, real variety.
Location: This can be a remote opportunity, with 2 weeks of travel into Rochester, NY per quarter
What You'll Do:
  • Design and build S3 data lakes with multi-zone organization, partitioning strategies, lifecycle policies, and encryption
  • Implement medallion architecture (bronze/silver/gold) for data warehouses on Redshift, Snowflake, or Databricks
  • Build AWS Glue ETL pipelines (Python Shell and Spark) with incremental extraction, Data Catalog management, and optimized Parquet output
  • Design star/snowflake schemas, materialized views, and gold-layer models optimized for BI consumption (QuickSight, PowerBI)
  • Configure data warehouse platforms - Redshift with Zero-ETL from Aurora, Snowflake with Snowpipe, Databricks with Delta Lake and Auto Loader
  • Design RAG systems using Bedrock Knowledge Base with OpenSearch Serverless vector search and Titan Embeddings
  • Architect document AI pipelines using Textract, Comprehend, and Bedrock for entity extraction
  • Design SageMaker ML pipelines for training, Model Registry, and inference
  • Lead data discovery sessions with client stakeholders and present architecture recommendations to technical and business audiences
  • Mentor delivery team members on data architecture patterns and AWS data services
  • Contribute to R&D projects evaluating emerging AWS data and AI capabilities

Required Skills:
  • 5+ years professional IT experience, 2+ years professional AWS experience
  • At least one AWS Professional-level certification (Solutions Architect Professional or Data Engineer Specialty preferred)
  • Python for data pipelines (Glue jobs, Lambda, SageMaker scripts) and PySpark for Glue Spark jobs
  • SQL and NoSQL on AWS - Aurora PostgreSQL, RDS PostgreSQL, DocumentDB, DynamoDB - including schema design and query optimization
  • Data modeling - conceptual, logical, and physical models for AWS data platforms; normalized silver-layer schemas, denormalized star/snowflake gold-layer schemas, data dictionaries
  • Dimensional modeling and medallion architecture (bronze/silver/gold) on Redshift, Snowflake, or Databricks, including materialized views and incremental refresh patterns
  • AWS Glue ETL (Python Shell and Spark), Glue Data Catalog, and crawlers
  • S3 data lake architecture with partitioning, lifecycle policies, and encryptions

Preferred:
  • RAG systems with Bedrock Knowledge Base and OpenSearch Serverless vector search
  • Amazon SageMaker for ML training, Model Registry, and inference
  • AWS HealthLake, FHIR R4 transformation, and HIPAA-compliant data pipelines
  • Document AI with Amazon Textract and Comprehend
  • Amazon Athena, QuickSight, or PowerBI integration
  • Terraform or CloudFormation for data infrastructure as code
  • Step Functions, EventBridge, and Lambda for event-driven pipeline orchestration

$150,000 - $200,000 a year
The salary range provided is a general guideline. When extending an offer, Innovative considers factors including, but not limited to, the responsibilities of the specific role, market conditions, geographic location, as well as the candidate's professional experience, key skills, and education/training.
We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.