1

Video Labelling Jobs in Edison, NJ (NOW HIRING)

Revenue Operations Manager

New York, NY ยท On-site

$100K - $160K/yr

In 18 months, we've built the combined technology of FrameIO (acquired by Adobe for $1.275B) and LucidLink ($40M ARR), layered with proprietary AI for visual search and video labeling. We process ...

Product Designer

New York, NY ยท On-site

$120K - $160K/yr

In 18 months, we've built the combined technology of FrameIO (acquired by Adobe for $1.275B) and LucidLink ($40M ARR), layered with proprietary AI for visual search and video labeling. We process ...

Graphic Designer

New York, NY ยท On-site

$95K - $135K/yr

In 18 months, we've built the combined technology of FrameIO (acquired by Adobe for $1.275B) and LucidLink ($40M ARR), layered with proprietary AI for visual search and video labeling. We process ...

Senior Software Engineer, Backend

New York, NY ยท On-site

$180K - $250K/yr

What You'll Work On Video + Data Platform Backend * Design and build high-throughput services for ... Build APIs and internal primitives that make large datasets usable for research, labeling, QA, and ...

AI Engineer

New York, NY

$125K - $150K/yr

Work with messy, multimodal sports data from tracking systems, video and computer vision outputs, audio, commentary, text, and structured feeds, including imperfect labels and ambiguous real-world ...

AI Engineer

New York, NY ยท On-site

$125K - $150K/yr

Work with messy, multimodal sports data from tracking systems, video and computer vision outputs, audio, commentary, text, and structured feeds, including imperfect labels and ambiguous real-world ...

Assistant, Marketing

New York, NY ยท On-site

$30K/yr

Organize department, label and artist meetings (catering, booking conference rooms, conference lines, video conferences, run A/V, etc) * Liaise with other departments, including Publicity, Digital ...

This role will lead negotiations with record labels and artist management in support of our original content video productions and exclusive artist merchandising offerings. The Business Affairs Lead ...

next page

Showing results 1-20

Video Labelling information

See Edison, NJ salary details

$15

$26

$42

How much do video labelling jobs pay per hour?

As of Aug 5, 2026, the average hourly pay for video labelling in Edison, NJ is $26.32, according to ZipRecruiter salary data. Most workers in this role earn between $19.90 and $30.10 per hour, depending on experience, location, and employer.

What does a video labelling do?

A typical day in Video Labelling involves reviewing video footage, identifying and annotating specific objects or events according to project guidelines, and entering this data into specialized software tools. Team members often collaborate with data scientists, engineers, or quality assurance leads to ensure accuracy and consistency in the annotations. Depending on the project and employer, you may work independently or as part of a larger team, sometimes with set quotas or deadlines. This work is crucial for developing and refining AI and machine learning models, making attention to detail and adherence to standards especially important. Over time, experienced video labelling professionals may progress to quality assurance roles or team leads overseeing larger annotation projects.

What are the key skills and qualifications needed to thrive in video labelling, and why are they important?

To thrive as a Video Labelling professional, you should have excellent attention to detail, basic computer proficiency, and familiarity with visual content analysis. Knowledge of annotation platforms, video editing software, or AI training tools is often required, and experience with data labelling systems can be beneficial. Strong communication, reliability, and the ability to follow detailed guidelines are important soft skills for this role. These abilities ensure high-quality, consistent data annotation that directly supports machine learning and computer vision projects.

What is a video labelling?

A Video Labelling job involves annotating or tagging objects, actions, or events in video footage to train machine learning models. This process helps AI systems recognize and interpret visual data accurately. Tasks may include drawing bounding boxes, classifying scenes, or adding timestamps for specific events. Video labelling is commonly used in industries like autonomous driving, security surveillance, and content moderation.

What are the most commonly searched types of Video Labelling jobs in Edison, NJ? The most popular types of Video Labelling jobs in Edison, NJ are:
What are popular job titles related to Video Labelling jobs in Edison, NJ? For Video Labelling jobs in Edison, NJ, the most frequently searched job titles are:
What cities near Edison, NJ are hiring for Video Labelling jobs? Cities near Edison, NJ with the most Video Labelling job openings:
Infographic showing various Video Labelling job openings in Edison, NJ as of July 2026, with employment types broken down into 78% Full Time, 16% Part Time, 2% Temporary, 3% Contract, and 1% Summer. Highlights an 95% Physical, 1% Hybrid, and 4% Remote job distribution, with an average salary of $54,751 per year, or $26.3 per hour.

Research Scientist, Video Understanding & World Models

Mecka AI

New York, NY โ€ข On-site

$200K - $250K/yr

Full-time

Re-posted 4 days ago


Job description

About Mecka AI
Mecka AI is building the data infrastructure layer for robotics and embodied AI.
We partner with leading AI labs and robotics companies to deliver high-quality, real-world datasets used to train, evaluate, and deploy robotic systems. Our work sits directly between research, data, and real-world execution - where model performance is dictated by data quality.
Our Mission
Robotics will become the largest industry in human history - larger than anything that has come before it. As intelligent machines move into the physical world, they will dramatically expand global GDP, raise the material standard of living for everyone, and ultimately help make humanity a multiplanetary civilization. None of that happens without one thing: enormous amounts of high-quality, real-world data.
Mecka AI builds that foundation. We are the data infrastructure layer for robotics and embodied AI - the substrate that teaches machines to perceive, reason, and act in reality. Get this right, and we accelerate the most important technological transition of our time.
Our Culture
  • Excellence as the baseline. We hold an extremely high bar and expect the best work of your career. Mediocrity isn't interesting to us.
  • Highly technical. We reason from first principles, not by analogy. The best argument wins - regardless of title or tenure.
  • Truth-seeking. We are relentlessly honest with ourselves and each other. We chase reality - measured, not assumed - and kill our own bad ideas fast.
  • Maniacal urgency. The work matters and the clock is real. We move fast, ship, measure, and iterate.
  • Extreme ownership. You own outcomes end-to-end - no hand-offs, no excuses, no waiting for permission.
  • Hardcore. This is a high-intensity environment for people who want to do the defining work of their lives.
The Role
We are looking for a Research Scientist, Video Understanding to own Mecka's video understanding agenda end-to-end: train large-scale video representation and video-language models on our egocentric + stereo corpus, and turn the resulting checkpoints into production signals the rest of the stack ships on.
This role is focused on large model training, video encoders, video-language models, VLMs/VLAs, and temporal representation learning on real-world robotics data.
What You'll Work On
Large-Scale Training & Architecture
  • Own model architecture and training strategy across Mecka's task families (manipulation, locomotion, daily activity, long-horizon behavior).
  • Run self-supervised and multimodal pretraining (VideoMAE / VJEPA / VideoPrism / InternVideo-class) with rigorous evals and clean ablations.
Video-Language & Multimodal Modeling
  • Train and fine-tune video encoders and video-language models (temporal transformers, joint-embedding models, contrastive objectives, masked modeling, instruction/video alignment).
  • Incorporate useful priors (pose, depth, camera motion, optical flow) when it improves representation quality.
Research โ†’ Production Signals
  • Turn checkpoints into usable artifacts: embeddings and model outputs that downstream systems can reliably consume (retrieval, labeling, QA, analytics).
  • Build a disciplined training + eval workflow with regression tracking and reproducible runs.
Who You Are
Required Background
  • Deep experience training large models in PyTorch (or equivalent), including multi-GPU or distributed training.
  • Strong understanding of modern video representation learning and/or multimodal modeling.
  • Ability to run rigorous experiments and communicate results clearly.
  • Warning: Research Scientist positions require hyper-specific expertise. Please limit your applications to one research role. Applying to multiple Research Scientist positions suggests a lack of focus and may result in the rejection of all submissions. You may, however, apply to other non-research roles alongside your research application.

Strong Signals:
  • Experience with video VLMs / VLA-adjacent systems (VideoCLIP, InstructBLIP-Video, LLaVA-Video-class).
  • Experience with egocentric / embodied datasets (Ego4D, EgoExo4D, EPIC-Kitchens, Something-Something).
  • Strong software engineering discipline: you write research code that can be shipped.
Why This Role
  • Work on a domain - egocentric embodied video - where data is scarce everywhere except here.
  • Own a research agenda that directly feeds production systems and product outcomes.