... annotation workflows to support continuous model improvement. • Explore model compression, inference acceleration, and serving optimizations to enable efficient real-time video processing at scale.
... annotation workflows to support continuous model improvement. • Explore model compression, inference acceleration, and serving optimizations to enable efficient real-time video processing at scale.
... annotation workflows to support continuous model improvement. • Explore model compression, inference acceleration, and serving optimizations to enable efficient real-time video processing at scale.
... annotation workflows to support continuous model improvement. • Explore model compression, inference acceleration, and serving optimizations to enable efficient real-time video processing at scale.
In this role, you will work on transcription, annotation, and evaluation tasks that help train and ... Video - Listen to, analyse, and transcribe audio and video content in Hindi, following detailed ...
In this role, you will work on transcription, annotation, and evaluation tasks that help train and ... Video - Listen to, analyse, and transcribe audio and video content in Hindi, following detailed ...
Data Operations Engineer
Mountain View, CA · On-site
$136K - $163K/yr
Preferred : • Experience with multimodal datasets (text, image, video, audio, or 3D). • Familiarity with data annotation, labeling workflows, or dataset preparation for machine learning. • ...
Data Operations Engineer
Mountain View, CA · On-site
$136K - $163K/yr
Preferred : • Experience with multimodal datasets (text, image, video, audio, or 3D). • Familiarity with data annotation, labeling workflows, or dataset preparation for machine learning. • ...
... video experiences). • Improve data quality through annotation, filtering, augmentation, synthetic generation, captioning, and in-depth data studies, particularly for visual and audio data. • ...
... video experiences). • Improve data quality through annotation, filtering, augmentation, synthetic generation, captioning, and in-depth data studies, particularly for visual and audio data. • ...
Experience with video understanding (temporal consistency, tracking, video segmentation) * Experience with foundation models for data annotation * Experience with MLOps tooling (Weights & Biases ...
Experience with video understanding (temporal consistency, tracking, video segmentation) * Experience with foundation models for data annotation * Experience with MLOps tooling (Weights & Biases ...
Experience with video understanding (temporal consistency, tracking, video segmentation) * Experience with foundation models for data annotation * Experience with MLOps tooling (Weights & Biases ...
Quick apply
Experience with video understanding (temporal consistency, tracking, video segmentation) * Experience with foundation models for data annotation * Experience with MLOps tooling (Weights & Biases ...
Experience with video understanding (temporal consistency, tracking, video segmentation) * Experience with foundation models for data annotation * Experience with MLOps tooling (Weights & Biases ...
Experience with video understanding (temporal consistency, tracking, video segmentation) * Experience with foundation models for data annotation * Experience with MLOps tooling (Weights & Biases ...
Experience with video understanding (temporal consistency, tracking, video segmentation) * Experience with foundation models for data annotation * Experience with MLOps tooling (Weights & Biases ...
Experience with video understanding (temporal consistency, tracking, video segmentation) * Experience with foundation models for data annotation * Experience with MLOps tooling (Weights & Biases ...
Computational Linguist, AI Evaluation
San Francisco, CA · On-site
$140K - $160K/yr
Video is 90% of the world's data. Most of it is invisible to machines. TwelveLabs builds the ... You will also be responsible for automating as much of the repetitive partnership and annotation ...
Computational Linguist, AI Evaluation
San Francisco, CA · On-site
$140K - $160K/yr
Video is 90% of the world's data. Most of it is invisible to machines. TwelveLabs builds the ... You will also be responsible for automating as much of the repetitive partnership and annotation ...
Member of Technical Staff - Imagine Model
$180K - $440K/yr
... video experiences). * Improve data quality through annotation, filtering, augmentation, synthetic generation, captioning, and in-depth data studies, particularly for visual and audio data. * Design ...
Member of Technical Staff - Imagine Model
$180K - $440K/yr
... video experiences). * Improve data quality through annotation, filtering, augmentation, synthetic generation, captioning, and in-depth data studies, particularly for visual and audio data. * Design ...
Office-based Video Data Contributor (Part-time)
Menlo Park, CA · On-site
$20/hr
... video datasets that help train advanced AI models ... Our Robotics team develops scalable data collection and annotation pipelines that enable the ...
Office-based Video Data Contributor (Part-time)
Menlo Park, CA · On-site
$20/hr
... video datasets that help train advanced AI models ... Our Robotics team develops scalable data collection and annotation pipelines that enable the ...
AI Data Strategist
Redwood City, CA · On-site
$148K - $192K/yr
Experience with embodied AI, video, or time-series data. * Familiarity with evaluation pipelines, active learning, or data-centric AI. * Exposure to annotation tooling such as Labelbox, Scale, CVAT ...
AI Data Strategist
Redwood City, CA · On-site
$148K - $192K/yr
Experience with embodied AI, video, or time-series data. * Familiarity with evaluation pipelines, active learning, or data-centric AI. * Exposure to annotation tooling such as Labelbox, Scale, CVAT ...
Product Operations Lead
$150K - $250K/yr
... video, audio, and multimodal data for AI training and evaluation. The team works on highly specific ... Source, onboard, and manage a distributed human workforce for data annotation, curation, and ...
New
Quick apply
Product Operations Lead
$150K - $250K/yr
... video, audio, and multimodal data for AI training and evaluation. The team works on highly specific ... Source, onboard, and manage a distributed human workforce for data annotation, curation, and ...
New
Product Operations Lead
San Francisco, CA · On-site
Video alone makes up 80% of internet traffic, and across modalities, data has become the enabling ... Source, onboard, and manage a distributed human workforce for data annotation, curation, and ...
Product Operations Lead
San Francisco, CA · On-site
Video alone makes up 80% of internet traffic, and across modalities, data has become the enabling ... Source, onboard, and manage a distributed human workforce for data annotation, curation, and ...
Technical Program Manager, Dataset Operations
Santa Clara, CA · On-site
$151K - $196K/yr
Computer vision / video datasets * Robotics / embodied AI data * Autonomous driving datasets * Understanding of: * Annotation workflows * Data quality evaluation * Dataset biases and coverage ...
Technical Program Manager, Dataset Operations
Santa Clara, CA · On-site
$151K - $196K/yr
Computer vision / video datasets * Robotics / embodied AI data * Autonomous driving datasets * Understanding of: * Annotation workflows * Data quality evaluation * Dataset biases and coverage ...
Research Engineer, Multimodal
Redwood City, CA · On-site
$225K - $400K/yr
... annotation workflows to support continuous model improvement. * Explore model compression, inference acceleration, and serving optimizations to enable efficient real-time video processing at scale.
Research Engineer, Multimodal
Redwood City, CA · On-site
$225K - $400K/yr
... annotation workflows to support continuous model improvement. * Explore model compression, inference acceleration, and serving optimizations to enable efficient real-time video processing at scale.
Experience working with human feedback data collection and annotation pipelines * Strong aesthetic sense and understanding of video quality assessment * Familiarity with alignment techniques such as ...
Quick apply
Experience working with human feedback data collection and annotation pipelines * Strong aesthetic sense and understanding of video quality assessment * Familiarity with alignment techniques such as ...
Experience working with human feedback data collection and annotation pipelines * Strong aesthetic sense and understanding of video quality assessment * Familiarity with alignment techniques such as ...
Experience working with human feedback data collection and annotation pipelines * Strong aesthetic sense and understanding of video quality assessment * Familiarity with alignment techniques such as ...
Member of Technical Staff - Imagine Model
Palo Alto, CA · On-site
$180K - $440K/yr
... video experiences). * Improve data quality through annotation, filtering, augmentation, synthetic generation, captioning, and in-depth data studies, particularly for visual and audio data. * Design ...
Member of Technical Staff - Imagine Model
Palo Alto, CA · On-site
$180K - $440K/yr
... video experiences). * Improve data quality through annotation, filtering, augmentation, synthetic generation, captioning, and in-depth data studies, particularly for visual and audio data. * Design ...
Video Annotation information
What is a Video Annotation job?
A Video Annotation job involves labeling objects, activities, or events within videos to help train machine learning models. Annotators use specialized tools to draw bounding boxes, segment frames, or classify scenes to improve AI's ability to recognize visuals. This work is essential for applications like autonomous vehicles, facial recognition, and action recognition in AI systems.
What are the typical daily responsibilities of a Video Annotation specialist?
As a Video Annotation specialist, your daily tasks will generally involve watching video footage, identifying relevant objects or actions, and accurately labeling or tagging frames according to specific project guidelines. You may also review and validate annotations to ensure quality and consistency, collaborate with team members or project managers to clarify labeling instructions, and document any ambiguities or challenges encountered during annotation. Most roles are structured with clear targets or quotas for completed work, and you may work independently or as part of a larger team supporting AI development projects. The position requires strong concentration and the ability to handle repetitive tasks efficiently while maintaining high standards of accuracy.
What are the key skills and qualifications needed to thrive in the Video Annotation position, and why are they important?
To excel in Video Annotation, you need strong attention to detail, visual analysis skills, and familiarity with video processing concepts, often supported by a diploma or coursework in computer science or a related field. Knowledge of annotation tools such as CVAT, Labelbox, or VGG Image Annotator, and, in some cases, experience with basic scripting or data management platforms, is highly valued. Excellent focus, time management, and the ability to follow precise instructions help individuals stand out in this position. These abilities are crucial for ensuring the accuracy and quality of annotated video data, which directly impacts AI and machine learning model performance.

Full-time
Posted 19 days ago
Job description
Character.AI empowers people to connect, learn and tell stories through interactive entertainment. As a Research Engineer on the Multimodal team, you will build and advance video and image generation models that enhance user interactions with AI characters.
Responsibilities:
• Lead fine-tuning and continued training of video generation models, including image-to-video and joint audio-visual generation.
• Design and experiment with novel model architectures for multimodal generation, including multimodal conditioning (voice, structured text, reference images).
• Leverage techniques such as LoRA, RLHF, and full-parameter fine-tuning to improve model quality across diverse visual scenarios.
• Design and build large-scale data pipelines and automated annotation workflows to support continuous model improvement.
• Explore model compression, inference acceleration, and serving optimizations to enable efficient real-time video processing at scale.
Qualifications:
Required:
• Strong passion for pushing the boundaries of visual AI, with a self-driven, hands-on approach to solving complex technical problems
• Proficient in PyTorch with end-to-end experience across data processing, model training, and deployment
• Solid understanding of video and image generation architectures, including diffusion models, DiT, ControlNet, and SOTA video generation models
• Experience with multimodal model training, including working with audio, vision, and language modalities together
• Experience with distributed training tools (FSDP, DeepSpeed, etc.)
• Experience with large-scale data processing, dataset construction, and automated data cleaning
Preferred:
• Experience with joint audio-visual or speech-conditioned generation models
• Experience with AIGC, video effects, character animation, or asset generation products
• Familiarity with ML deployment and orchestration (Kubernetes, Slurm, Docker, cloud platforms)
• Publications in relevant venues (NeurIPS, ICLR, CVPR, ECCV, ICCV, or similar)
Company:
Character.ai provides open-ended conversational applications in which users create characters and converse with them. Founded in 2021, the company is headquartered in Menlo Park, USA, with a team of 51-200 employees. The company is currently Growth Stage.