1

Data Io Jobs in California (NOW HIRING)

Senior Product Manager, Frame.io

San Jose, CA · On-site

$148K - $195K/yr

About Frame.io Frame.io is transforming the way creative work is carried out. Creative ... Translate user feedback, data, and market trends into actionable roadmaps. * Prototype and validate ...

About Frame.io Frame.io is transforming the way creative work is carried out. Creative ... Translate user feedback, data, and market trends into actionable roadmaps. * Prototype and validate ...

Senior Product Manager, Frame.io

San Francisco, CA · On-site

$149K - $196K/yr

About Frame.io Frame.io is transforming the way creative work is carried out. Creative ... Translate user feedback, data, and market trends into actionable roadmaps. * Prototype and validate ...

About Frame.io Frame.io is redefining how creative work happens. Creative professionals around the ... Translate user feedback, data, and market trends into actionable roadmaps. * Prototype and validate ...

... world's data-centric applications-from cloud and communications infrastructure to automotive ... About the Role As an FPGA DDR and IO Subsystem Architec t, you will define and drive the ...

next page

Showing results 1-20

Data Io information

See California salary details

$10

$30

$67

How much do data io jobs pay per hour?

As of Aug 24, 2026, the average hourly pay for data io in California is $30.76, according to ZipRecruiter salary data. Most workers in this role earn between $17.65 and $39.78 per hour, depending on experience, location, and employer.

What is a Data Io?

A Data IO (Input/Output) job typically involves managing the flow of data between systems, databases, and storage units. Professionals in this role ensure efficient data transfer, optimize data pipelines, and maintain data integrity. They work with technologies like ETL tools, cloud storage, and APIs to facilitate seamless data movement. Strong programming skills in Python, SQL, or other languages are often required.

What are the typical daily responsibilities of a Data Io?

A Data IO Specialist is responsible for managing the seamless transfer, validation, and processing of data between systems or databases on a daily basis. Tasks often include designing and maintaining ETL pipelines, monitoring data quality, troubleshooting transfer issues, and collaborating with data engineers, analysts, and business stakeholders to meet project requirements. This role typically involves working within a larger data team and requires clear documentation and communication to ensure data flows support business needs. Daily work may also involve responding to data incidents or implementing improvements to optimize data integration processes.

What are the key skills and qualifications needed to thrive in the Data Io position?

To thrive as a Data IO (Input/Output) Specialist, you need strong analytical skills, proficiency in data management, and a background in computer science or information technology. Familiarity with data integration tools, scripting languages like Python or SQL, and systems such as ETL platforms or data warehouses is frequently required. Excellent attention to detail, problem-solving abilities, and effective communication are crucial soft skills for this position. These competencies ensure the accurate and efficient transfer, transformation, and management of data across various systems and teams.

What are the most commonly searched types of Data Io jobs in California?

The most popular types of Data Io jobs in California are:

What cities in California are hiring for Data Io jobs?

Cities in California with the most Data Io job openings:

Infographic showing various Data Io job openings in California as of August 2026, with employment types broken down into 1% As Needed, 88% Full Time, 9% Part Time, and 2% Contract. Highlights an 85% Physical, 4% Hybrid, and 11% Remote job distribution, with an average salary of $63,986 per year, or $30.8 per hour.

AI Engineer -- Research Agents (Full-Stack)

Socket.dev

San Francisco, CA • On-site

$180 - $260/hr

Other

Posted 6 days ago


Job description

What you’ll do
  • Design and ship agentic systems (tool calling, multi-agent workflows, structured outputs) that reliably fetch, extract, and normalize data across the web and APIs.
  • Build and operate search/indexing pipelines on OpenSearch/Elasticsearch (schema design, analyzers, reindex/data migration strategies, relevance tuning).
  • Own robust web scraping: directory crawling, CAPTCHA handling, headless browsers, rotating proxies, anti-bot evasion, and backoff/retry policies.
  • Develop backend services in Python + FastAPI with clean contracts and strong observability.
  • Scale workloads on AWS + Temporal (batch/queue workers, autoscaling, fault tolerance, cost control).
  • Parallelize external API requests safely (rate limits, idempotency, circuit breakers, retries, dedupe).
  • Integrate third‑party APIs for enrichment and search; model and cache responses; manage schema evolution.
  • Transform and analyze data using Pandas (or similar) for normalization, QA, and reporting.
  • Pitch in across the stack: billing (Stripe), and occasional front‑end changes to ship end‑to‑end features.
Minimum requirements
  • Hands-on experience with agentic architectures (tool calling, structured outputs/JSON, planning/execution loops) and prompt engineering.
  • Deep knowledge of OpenSearch/Elasticsearch: index design, analyzers, ingestion pipelines, snapshots, rolling upgrades, and zero-downtime reindexing/data migrations.
  • Proven web scraping expertise: solving CAPTCHAs, session/auth flows, proxy rotation, stealth techniques, and legal/ethical constraints.
  • AWS + Temporal in production (at least two of: ECS/EKS, Lambda, SQS/SNS, Batch, Step Functions, CloudWatch).
  • Building high-throughput data/IO pipelines with concurrency (asyncio/multiprocessing), resilient retries, and rate‑limit aware scheduling.
  • Integrating diverse external APIs (auth patterns, pagination, webhooks); designing stable interfaces and backfills.
  • Strong data wrangling with Pandas or equivalent; comfort with large CSV/Parquet workflows and memory/perf tuning.
  • Familiarity with Stripe (subscriptions, metered billing, webhooks) and basic front‑end changes (React/TypeScript or similar).
  • Excellent ownership, product sense, and pragmatic debugging.
Nice to have
  • Entity resolution/record linkage at scale (probabilistic matching, blocking, deduping).
  • Experience with Langfuse, OpenTelemetry, or similar for tracing/evals; task queues (Celery/RQ), Redis, Postgres.
  • Search relevance (BM25/vector/hybrid), embeddings, and retrieval pipelines.
  • Playwright/Selenium, stealth browsers, anti‑bot frameworks, CAPTCHA providers.
  • CI/CD, infrastructure as code (Terraform), and cost/perf observability.
  • Security & compliance basics for data handling and PII.
#J-18808-Ljbffr