Senior Data AI Engineer
$109K - $148K/yr
Proven experience designing and implementing vector databases (e.g., Vertex AI Vector Search, Pinecone, pgvector), embedding pipelines, and knowledge graph structures that underpin RAG and semantic ...
$109K - $148K/yr
Proven experience designing and implementing vector databases (e.g., Vertex AI Vector Search, Pinecone, pgvector), embedding pipelines, and knowledge graph structures that underpin RAG and semantic ...
$109K - $148K/yr
Proven experience designing and implementing vector databases (e.g., Vertex AI Vector Search, Pinecone, pgvector), embedding pipelines, and knowledge graph structures that underpin RAG and semantic ...
Deerfield, IL · Remote
$56.25 - $69.75/hr
Database & Search: Azure SQL, CosmosDB, Vector DBs (Pinecone, FAISS, Weaviate). * DevOps & Security: Azure DevOps, Docker, Kubernetes, OAuth, JWT. * Frameworks: LangChain, LLamaIndex, Semantic Kernel ...
Quick apply
Deerfield, IL · Remote
$56.25 - $69.75/hr
Database & Search: Azure SQL, CosmosDB, Vector DBs (Pinecone, FAISS, Weaviate). * DevOps & Security: Azure DevOps, Docker, Kubernetes, OAuth, JWT. * Frameworks: LangChain, LLamaIndex, Semantic Kernel ...
Chicago, IL · On-site
$144K - $177K/yr
... of vector databases like Pinecone, FAISS, or Weaviate. · Experience with Azure SQL, CosmosDB, and scalable backend architecture. · Familiarity with LangChain, LLamaIndex, and Microsoft Semantic ...
Quick apply
Chicago, IL · On-site
$144K - $177K/yr
... of vector databases like Pinecone, FAISS, or Weaviate. · Experience with Azure SQL, CosmosDB, and scalable backend architecture. · Familiarity with LangChain, LLamaIndex, and Microsoft Semantic ...
Chicago, IL · On-site
... Vector Databases such as Pinecone, ChromaDB, FAISS, Weaviate, or Milvus · Experience integrating AI models through REST APIs · Strong understanding of embeddings, tokenization, and semantic search ...
Chicago, IL · On-site
... Vector Databases such as Pinecone, ChromaDB, FAISS, Weaviate, or Milvus · Experience integrating AI models through REST APIs · Strong understanding of embeddings, tokenization, and semantic search ...
Implement vector databases (e.g., Pinecone, FAISS, Chroma) for efficient context management and response caching. * Ensure code quality through best practices: clean code, testing, and code reviews.
Implement vector databases (e.g., Pinecone, FAISS, Chroma) for efficient context management and response caching. * Ensure code quality through best practices: clean code, testing, and code reviews.
Westmont, IL · On-site
$63.50 - $82.75/hr
... of vector databases like Pinecone, FAISS, or Weaviate. · Experience with Azure SQL, CosmosDB, and scalable backend architecture. · Familiarity with LangChain, LLamaIndex, and Microsoft Semantic ...
Quick apply
Westmont, IL · On-site
$63.50 - $82.75/hr
... of vector databases like Pinecone, FAISS, or Weaviate. · Experience with Azure SQL, CosmosDB, and scalable backend architecture. · Familiarity with LangChain, LLamaIndex, and Microsoft Semantic ...
Chicago, IL · On-site +1
$153K/yr
Architecting, optimizing relational and vector databases (PostgreSQL, SQLAlchemy, query optimization, indexes, replicas, migrations, Weaviate, Pinecone) and working with dataframes for data ...
Chicago, IL · On-site +1
$153K/yr
Architecting, optimizing relational and vector databases (PostgreSQL, SQLAlchemy, query optimization, indexes, replicas, migrations, Weaviate, Pinecone) and working with dataframes for data ...
Chicago, IL · On-site +1
$153K/yr
Architecting, optimizing relational and vector databases (PostgreSQL, SQLAlchemy, query optimization, indexes, replicas, migrations, Weaviate, Pinecone) and working with dataframes for data ...
Chicago, IL · On-site +1
$153K/yr
Architecting, optimizing relational and vector databases (PostgreSQL, SQLAlchemy, query optimization, indexes, replicas, migrations, Weaviate, Pinecone) and working with dataframes for data ...
Chicago, IL · On-site
... or vector databases for knowledge grounding. Demonstrated 25% reduction in hallucination or ... Vector DB (Vertex Matching Engine, Pinecone, or Weaviate) Ideal Profile * 3-5 years hands-on GCP ...
Chicago, IL · On-site
... or vector databases for knowledge grounding. Demonstrated 25% reduction in hallucination or ... Vector DB (Vertex Matching Engine, Pinecone, or Weaviate) Ideal Profile * 3-5 years hands-on GCP ...
Chicago, IL · On-site
$65 - $85.50/hr
... Vector Databases Design, Model Routing, LLM Orchestration, Developing Agents and building Agentic Mesh, technologies such as LangChain, LlamaIndex, PineCone, Milvus, PyTorch, Tensor Flow and ...
Chicago, IL · On-site
$65 - $85.50/hr
... Vector Databases Design, Model Routing, LLM Orchestration, Developing Agents and building Agentic Mesh, technologies such as LangChain, LlamaIndex, PineCone, Milvus, PyTorch, Tensor Flow and ...
RetrievalAugmented Generation (RAG) using vector databases (e.g., Pinecone, FAISS, Milvus, Weaviate) * Implement CI/CD pipelines for agent code, prompts, and policies. * Integrate GenAI agents with ...
RetrievalAugmented Generation (RAG) using vector databases (e.g., Pinecone, FAISS, Milvus, Weaviate) * Implement CI/CD pipelines for agent code, prompts, and policies. * Integrate GenAI agents with ...
Retrieval‑Augmented Generation (RAG) using vector databases (e.g., Pinecone, FAISS, Milvus, Weaviate) * Implement CI/CD pipelines for agent code, prompts, and policies. * Integrate GenAI agents ...
Quick apply
Retrieval‑Augmented Generation (RAG) using vector databases (e.g., Pinecone, FAISS, Milvus, Weaviate) * Implement CI/CD pipelines for agent code, prompts, and policies. * Integrate GenAI agents ...
Understanding of prompt engineering, RAG (retrieval-augmented generation), and vector databases (e.g., Pinecone, Weaviate, Chroma). * Solid understanding of Agile/Scrum practices.
Quick apply
Understanding of prompt engineering, RAG (retrieval-augmented generation), and vector databases (e.g., Pinecone, Weaviate, Chroma). * Solid understanding of Agile/Scrum practices.
Chicago, IL · On-site
... with vector databases (e.g., Pinecone), and building AI agents and multi-service AI architectures-particularly within Azure environments. • Cloud & DevOps Proficiency - Familiarity with Azure ...
Chicago, IL · On-site
... with vector databases (e.g., Pinecone), and building AI agents and multi-service AI architectures-particularly within Azure environments. • Cloud & DevOps Proficiency - Familiarity with Azure ...
Chicago, IL · On-site
$124K - $161K/yr
RAG & Vector Databases: Experience parsing, cleaning, and feeding structured and unstructured data into LLMs, alongside strong knowledge of vector stores (e.g., Pinecone). Model Context Protocols:
Chicago, IL · On-site
$124K - $161K/yr
RAG & Vector Databases: Experience parsing, cleaning, and feeding structured and unstructured data into LLMs, alongside strong knowledge of vector stores (e.g., Pinecone). Model Context Protocols:
Chicago, IL · On-site +1
$65 - $85.50/hr
Vector Databases: Azure Cosmos DB, Pinecone, or Weaviate * DevOps & MLOps: Azure DevOps, GitHub Actions, Docker, Kubernetes About EisnerAmper: EisnerAmper is one of the largest accounting, tax, and ...
Chicago, IL · On-site +1
$65 - $85.50/hr
Vector Databases: Azure Cosmos DB, Pinecone, or Weaviate * DevOps & MLOps: Azure DevOps, GitHub Actions, Docker, Kubernetes About EisnerAmper: EisnerAmper is one of the largest accounting, tax, and ...
Chicago, IL · On-site +1
$65 - $85.50/hr
Vector Databases: Azure Cosmos DB, Pinecone, or Weaviate * DevOps & MLOps: Azure DevOps, GitHub Actions, Docker, Kubernetes About EisnerAmper: EisnerAmper is one of the largest accounting, tax, and ...
Chicago, IL · On-site +1
$65 - $85.50/hr
Vector Databases: Azure Cosmos DB, Pinecone, or Weaviate * DevOps & MLOps: Azure DevOps, GitHub Actions, Docker, Kubernetes About EisnerAmper: EisnerAmper is one of the largest accounting, tax, and ...
Chicago, IL · On-site
$165K - $216K/yr
Exposure to vector databases or semantic search tooling (e.g., Pinecone, Weaviate, pgvector) * Background in a technical presales or solutions engineering role at an AI/ML or data platform company ...
Chicago, IL · On-site
$165K - $216K/yr
Exposure to vector databases or semantic search tooling (e.g., Pinecone, Weaviate, pgvector) * Background in a technical presales or solutions engineering role at an AI/ML or data platform company ...
Chicago, IL · On-site
$180K - $190K/yr
Vector Databases (OpenSearch, Pinecone, Weaviate, Chroma, FAISS) * AI Governance, Responsible AI, Model Monitoring, and Security * Machine Learning & MLOps practices, including model lifecycle ...
Chicago, IL · On-site
$180K - $190K/yr
Vector Databases (OpenSearch, Pinecone, Weaviate, Chroma, FAISS) * AI Governance, Responsible AI, Model Monitoring, and Security * Machine Learning & MLOps practices, including model lifecycle ...
$85K - $162K/yr
Experience with vector databases (FAISS, Pinecone, Weaviate) and embeddingbased retrieval analysis. * Solid grasp of software architecture basics, distributed system concepts, and integration ...
$85K - $162K/yr
Experience with vector databases (FAISS, Pinecone, Weaviate) and embeddingbased retrieval analysis. * Solid grasp of software architecture basics, distributed system concepts, and integration ...
| Aspect | Pinecone Vector Databases | Data Engineers |
|---|---|---|
| Primary Role | Managing and deploying vector database solutions for AI/ML applications | Designing, building, and maintaining data pipelines and infrastructure |
| Skills & Certifications | Knowledge of vector databases, cloud platforms, programming (Python, SQL) | Data modeling, ETL processes, cloud services, programming (Python, Java) |
| Work Environment | Tech companies, AI startups, cloud providers | Data-driven organizations, tech firms, finance, healthcare |
While Pinecone Vector Databases specialists focus on deploying and managing vector database solutions for AI applications, Data Engineers build and maintain the data infrastructure that supports these systems. Both roles require programming skills and familiarity with cloud platforms, but their core responsibilities differ: one centers on database management, the other on data pipeline development.
You have a clear vision of where your career can go. And we have the leadership to help you get there.At CNA, we strive to create a culture in which people know they matter and are part of something important, ensuring the abilities of all employees are used to their fullest potential.
A senior individual contributor role responsible for designing, building, and operationalizing end-to-end AI and machine learning solutions that accelerate CNA's migration to a modern cloud data lakehouse. The engineer works across structured and unstructured data domains - including documents, images, audio, and transactional records - to unlock analytical value through scalable pipelines, RAG architectures, vector databases, and knowledge graphs. This role may also provide guidance to others to support the building of complex technical capabilities.JOB DESCRIPTION:
Essential Duties & Responsibilities
Performs a combination of duties in accordance with departmental guidelines:
Design and build AI solutions that accelerate data migration from legacy systems to the cloud, ensuring scalability, reliability, and governance compliance.
Design and implement scalable ingestion and transformation pipelines across structured (SQL, relational) and unstructured (documents, images, audio, email, call transcripts) data sources, applying OCR, NLP preprocessing, and document chunking strategiesoptimizedfor LLM consumption.
Implement modernlakehousepatterns on Google Cloud Platform (GCP) - including data governance, cataloging, and lineage tracking - to ensure data is reliably discoverable, auditable, and fit for AI/ML workloads at scale.
Design and implement vector databases, embedding pipelines, and knowledge graph structures that serve as the foundational retrieval layer for RAG and other AI applications.
Productionize and operationalize AI solutions and advanced analytics in a DevOps/MLOpsenvironment, including automated testing, monitoring, and rollback capabilities.
Cultivate innovation by proactively proposingnew ideasandidentifyingthe right combination of tools and frameworks to turn business problems into analytics solutions.
Researches,identifiesand implements process improvements that address complex technology gaps. Builds strong knowledge of technology enablers.
May perform additional duties as assigned.
Reporting Relationship
Typically Director or above
Skills, Knowledge & Abilities
Deep expertise building scalable ingestion and transformation pipelines across structured and unstructured data sources; strong background migrating workloads from legacy systems to modern cloud platforms.
Skilled in parsing and normalizing diverse content types - PDFs, emails, images, and call transcripts - using OCR, NLP preprocessing (tokenization, entity extraction, summarization), and document chunking strategies optimized for LLM consumption.
Proven experience designing and implementing vector databases (e.g., Vertex AI Vector Search, Pinecone, pgvector), embedding pipelines, and knowledge graph structures that underpin RAG and semantic search applications
Strong SQL and data analytical skills; experience building data marts and feature datasets for data science and ML applications.
Strong coding fluency in Python; hands-on experience with BigQuery, Claude Code, RAG architectures, LLMs, ADK, and prompt engineering techniques
Expertise in building ML platforms and data pipelines at scale; familiarity with major ML algorithms, deep learning, NLP, information retrieval, and data mining techniques
Experience with GCP services (Vertex AI, Dataflow, BigQuery, Cloud Run, Pub/Sub); comfort with distributed computing frameworks (Apache Spark, Dataproc) for large-scale data processing.
Solid experience managing diverse data sources including preprocessing, cleansing, and verifying data integrity to meet data science and ML requirements
Demonstrated experience with machine learning, deep learning, information retrieval, NLP, or data mining - particularly applied to unstructured or semi-structured data
Hands-on experience with vector databases, embedding models (e.g., text-embedding-gecko, OpenAI Ada, Cohere), and end-to-end RAG pipeline design
Experience using Agile methods preferred.
Strong communication and interpersonal skills and the ability to work effectively with peers and team members in a highly matrixed environment.
Preferred experience with the insurance industry, its products and services.
Experience in implementing big data processing technology. Apache Spark preferred.
Education & Experience
Bachelor's Degree in Computer Science, Engineering, Mathematics, Computational Statistics, Data Science, or a related technical field (or equivalent experience);Master's Degreepreferred.
Typically7+ years of experience in data engineering, ArtificialIntelligenceor Machine Learning.
2+ years of codingproficiencyin at least one programming language (Python, Java, SQL).
Applicable certifications preferred (GCP, Data Engineering).
#LI-KJ1 #LI-HYBRID
In certain jurisdictions, CNA is legally required to include a reasonable estimate of the compensation for this role. In District of Columbia, California, Colorado, Connecticut, Illinois, Maryland, Massachusetts, New York and Washington, the national base pay range for this job level is $72,000 to $141,000 annually.Salary determinations are based on various factors, including but not limited to, relevant work experience, skills, certifications and location. CNA offers a comprehensive and competitive benefits package to help our employees - and their family members - achieve their physical, financial, emotional and social wellbeing goals. For a detailed look at CNA's benefits, please visitcnabenefits.com.
CNAutilizesAI-enabled technology during the recruiting process. For more information, please visitourcareers page.
CNA is committed to providing reasonable accommodations to qualified individuals with disabilities in the recruitment process. To request an accommodation, please contactleaveadministration@cna.com