We believe leaders stay hands-on: even as we grow, everyone (including managers) continues to ship ... Web scraping and API integrations at scale. * Startup founder or early-stage experience is a huge ...
Quick apply
We believe leaders stay hands-on: even as we grow, everyone (including managers) continues to ship ... Web scraping and API integrations at scale. * Startup founder or early-stage experience is a huge ...
Quick apply
We believe leaders stay hands-on: even as we grow, everyone (including managers) continues to ship ... Web scraping and API integrations at scale. * Startup founder or early-stage experience is a huge ...
New York, NY ยท On-site
$57.75 - $79/hr
About Tavily We're building the infrastructure layer for agentic web interaction at scale. Our API ... Managing Kubernetes clusters across multiple environments and regions * Owning infrastructure as ...
New York, NY ยท On-site
$57.75 - $79/hr
About Tavily We're building the infrastructure layer for agentic web interaction at scale. Our API ... Managing Kubernetes clusters across multiple environments and regions * Owning infrastructure as ...
$62.25 - $82.75/hr
About Tavily We're building the infrastructure layer for agentic web interaction at scale. Our API ... Managing Kubernetes clusters across multiple environments and regions * Owning infrastructure as ...
$62.25 - $82.75/hr
About Tavily We're building the infrastructure layer for agentic web interaction at scale. Our API ... Managing Kubernetes clusters across multiple environments and regions * Owning infrastructure as ...
New York, NY ยท On-site
$147K - $224K/yr
Managing Kubernetes clusters across multiple environments and regions * Owning infrastructure as ... Real scaling challenges - bursty scraping workloads, cache invalidation, multi-region, millions of ...
New York, NY ยท On-site
$147K - $224K/yr
Managing Kubernetes clusters across multiple environments and regions * Owning infrastructure as ... Real scaling challenges - bursty scraping workloads, cache invalidation, multi-region, millions of ...
New York, NY ยท On-site
$147K - $224K/yr
Managing Kubernetes clusters across multiple environments and regions * Owning infrastructure as ... Real scaling challenges - bursty scraping workloads, cache invalidation, multi-region, millions of ...
New York, NY ยท On-site
$147K - $224K/yr
Managing Kubernetes clusters across multiple environments and regions * Owning infrastructure as ... Real scaling challenges - bursty scraping workloads, cache invalidation, multi-region, millions of ...
New York, NY ยท On-site +1
$134K - $176K/yr
By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that ... Design and operate prepaid credit/wallet systems with auto top-up and balance management * Support ...
New York, NY ยท On-site +1
$134K - $176K/yr
By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that ... Design and operate prepaid credit/wallet systems with auto top-up and balance management * Support ...
New York, NY ยท On-site
$134K - $176K/yr
By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that ... Design and operate prepaid credit/wallet systems with auto top-up and balance management * Support ...
New York, NY ยท On-site
$134K - $176K/yr
By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that ... Design and operate prepaid credit/wallet systems with auto top-up and balance management * Support ...
New York, NY ยท On-site
$147K - $224K/yr
By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that ... Design and operate prepaid credit/wallet systems with auto top-up and balance management * Support ...
New York, NY ยท On-site
$147K - $224K/yr
By connecting LLMs to high-quality, trustworthy web content, we help developers build agents that ... Design and operate prepaid credit/wallet systems with auto top-up and balance management * Support ...
... content, scraping, and other forms of fraud and abuse, and some time on product infrastructure ... Our stack is React (web app), React Native (mobile app), Node.js, and Typescript, running on GCP ...
... content, scraping, and other forms of fraud and abuse, and some time on product infrastructure ... Our stack is React (web app), React Native (mobile app), Node.js, and Typescript, running on GCP ...
$40.4K - $50.3K
13% of jobs
$57.4K is the 25th percentile. Wages below this are outliers.
$50.3K - $60.2K
17% of jobs
$60.2K - $70K
10% of jobs
The median wage is $76.6K / yr.
$70K - $79.9K
16% of jobs
$79.9K - $89.8K
17% of jobs
$92.6K is the 75th percentile. Wages above this are outliers.
$89.8K - $99.7K
10% of jobs
$99.7K - $109.6K
5% of jobs
$109.6K - $119.5K
3% of jobs
$119.5K - $129.3K
1% of jobs
$129.3K - $139.2K
3% of jobs
$139.2K - $149.1K
5% of jobs
$40.4K
$83.7K
$149.1K
| Aspect | Manager Web Scraping | Data Analyst |
|---|---|---|
| Primary Focus | Overseeing web scraping projects, managing scraping teams, ensuring data quality | Analyzing data sets, generating reports, interpreting data insights |
| Skills & Certifications | Web scraping tools, programming (Python, SQL), project management | Statistical analysis, data visualization, Excel, SQL |
| Work Environment | Tech teams, data engineering, project management | Business units, data teams, reporting environments |
| Industry Usage | Web data collection, market research, e-commerce | Business intelligence, marketing, finance |
While both roles involve working with data, the Manager Web Scraping primarily oversees the technical process of extracting data from websites, managing teams, and ensuring data quality. In contrast, a Data Analyst focuses on interpreting data, creating reports, and providing insights to support business decisions. Both roles require technical skills, but their core responsibilities and work environments differ.
Cities near Commack, NY with the most Manager Web Scraping job openings:

Full-time
Medical, Dental, Vision, Retirement, PTO
Re-posted 11 days ago
Clipbook builds AI-powered media monitoring and analytics tools for communications, PR, and public affairs teams. We help them track, search, and make sense of coverage across text, audio, and video.
We launched in 2023 and have grown to 200+ clients including BCG, Weber Shandwick, and dozens of government agencies. We bootstrapped to seven-figures in ARR before raising a $3.3M seed round (co-led by Mark Cuban). We plan to raise our Series A this year and we are aiming for eight-figures revenue by end of year.
Our founding team has backgrounds at BCG, Bain, Harvard, Stanford, Oxford, the White House, and Congress, and has previously built startups backed by Sequoia, Tiger Global, Insight Partners, Coatue, and NFX.
The RoleYou'd be one of the first engineers at Clipbook — joining a small but mighty engineering team with engineers from Meta, Stripe, and AWS. You'll have real influence over architecture and technical direction, and what you build will ship to 200+ customers quickly.
What You'll DoArchitect & build core backend systems. Drive architectural decisions across our backend stack (Python, Node.js, PostgreSQL). Own features from concept → deployment → observing users rely on what you built. We're still laying critical foundations, so we value a pragmatic, "strong opinions, weakly held" mindset — especially when decisions are expensive to unwind later.
Design scalable data infrastructure. Build and maintain pipelines for ingesting, normalizing, deduplicating, indexing, and querying massive, multi-modal datasets (text, audio, video) across news, social media, policy, and more. A lot of this is normalization, deduplication, and edge case handling.
Integrate AI into real-world workflows. Put LLMs and ML models into production for real workflows: RAG pipelines, embeddings, prompt execution, agentic systems, fine-tuned models. Production means reliable, monitored, and cost-conscious.
Design systems that scale. Build systems that will hold up as we grow 10x, while being practical about what to invest in now vs. what can wait.
Develop performant APIs and services. Create robust interfaces and internal services that power our product end-to-end, ensuring reliability, security, and a seamless experience for customers.
Collaborate closely with users. Join customer calls, hear what's working and what isn't, and build in response. Rapidly iterate to deliver solutions that genuinely move the needle for comms/public affairs teams.
Shape Clipbook's engineering culture. As one of the first engineers, you'll influence everything — code quality, system design principles, documentation standards, and how we build as a team. We believe leaders stay hands-on: even as we grow, everyone (including managers) continues to ship.
2–10+ years building and scaling production backend systems. You've taken systems from zero to production and owned the full lifecycle. Strong across backend fundamentals (services, data models, distributed systems), with deeper expertise in one or more areas — and a desire to keep expanding your breadth over time.
Genuine excitement to build quickly & ship fast. We care about getting things into users' hands and iterating from there.
You take ownership. When something is broken or unclear, you fix it or flag it without waiting to be asked.
Comfortable with ambiguity. You make reasonable calls with incomplete information, communicate them clearly, and adjust as you learn more.
Strong experience with cloud infrastructure, containerization, and CI/CD. We're building systems that will one day serve Fortune 500 executives in real-time — so reliability matters. Familiarity with AWS/GCP, containerization (Docker/K8s), and CI/CD pipelines ensures we can deliver high uptime and iterate quickly.
A future leader. You'll help shape our engineering culture and can quickly grow into leadership as the company scales.
Data Engineering: Spark, Kafka, Flink, BigQuery, streaming pipelines, ETL at scale.
Distributed Systems: High-volume scaling, fault tolerance, eventual consistency.
AI/ML: LLMs in production, RAG, embeddings, fine-tuning, inference optimization.
Backend: Python, Node.js, PostgreSQL, high-concurrency systems.
Computer vision or multimodal model experience (audio, video, image). This is central to our product, so it's a meaningful differentiator.
Semantic search, vector databases, or meaning-aware retrieval.
LLM fine-tuning, RLHF, or eval pipelines.
Web scraping and API integrations at scale.
Startup founder or early-stage experience is a huge plus.
We're hiring for the application and data layer. This probably isn't the right fit if your background is primarily in firmware, embedded systems, hardware engineering, networking, or infrastructure/DevOps.
A Few Things Worth KnowingWe care more about what you ship than when or where you work, but this is an early-stage company growing fast. The pace reflects that. Our culture is intense and driven by H&H (hunger and hustle).
As one of the first engineers, you'll wear a lot of hats. There's no platform team to hand things off to yet.
Salary: $150K–$220K
Equity: Early-stage grant with significant upside
Benefits: Medical, dental, vision, 401(k), unlimited PTO
Growth: A clear path to engineering leadership as we scale