Video Annotator: Building Video Classifiers using Vision-Language
Description
Netflix Tech Blog on a framework for training ML models for video understanding. Describes an internal system to create video classifiers using combined video and text (subtitle/metadata) analysis.
Resource details
- Category path
- Intro & Learning›Tutorials & Case Studies
- Classification
- Not yet classified
Video Annotator: Building Video Classifiers using Vision-Language is cataloged in Intro & Learning and Tutorials & Case Studies. Its topics include machine learning, video classification, Netflix, vision-language model, and ml framework.
URL
Tags
Related resources
- Building Service Topology at Scale (Netflix Tech Blog)Real-time service dependency-mapping system for streaming microservices at scale, architecture and lessons learned.
- High Quality Video Encoding at Scale (Netflix Tech Blog)Deep-dive into Netflix's EC2-based elastic encode pipeline, per-title encoding, and automated quality scoring.
- EME, CDM, AES, CENC, and Keys — Understanding DRM Building BlocksOTTVerse explainer (Mile-High Video blog-award winner) tying together the core DRM protocol/technology building blocks: EME, CDM, AES, CENC…
- Tech @ WBD — Warner Bros Discovery Engineering BlogWarner Bros Discovery's engineering blog covering streaming, video, and product tech posts from internal engineers.
- MirrorFly IM SoftwareMirrorFly is a self-hosted instant messaging software that enables businesses to create a white-label team chat and meeting app. The soluti…
- A Hitchhiker's Guide to Structural SimilarityNetflix-affiliated deep-dive paper on SSIM quality metrics and their application to streaming QoE evaluation.
- Building Your Own Scalable Video Streaming Server (Part 1)Code-along tutorial series building a microservices-based video streaming server with FFmpeg HLS segmenting.
- Building Your First Video Pipeline: FFmpeg & MediaMTX Basics (Part 1)Technical tutorial walking through demux→decode→filter→encode→mux pipeline stages toward building a low-latency streaming setup with FFmpeg…
- Video Encoding to AV1 Guide (WIP)Long-running community-maintained guide to AV1 software encoding (libaom/aomenc) on the Level1Techs forum.
- Fora Soft: Video Streaming Deep-Dive GuideFree 9-chapter technical guide covering HLS, DASH, WebRTC, CMAF, CDN, and DRM fundamentals.
- Unreeling Netflix: Understanding and Improving Multi-CDN Movie DeliveryAcademic paper (Princeton/Bell Labs) measuring Netflix's multi-CDN delivery architecture, DNS-based CDN selection, and AWS backend.
- Deep Dive into HLS Protocol ArchitectureTechnical explainer covering TS-vs-fMP4 tradeoffs, MSE API, and byte-level container overhead analysis.
- ExoPlayer Architecture & DRM (System Design Lesson)Structured system-design lesson covering ExoPlayer architecture and DRM integration.
- Streaming Tech Sweden 2025 RecordingsSession recordings from the 2025 Streaming Tech Sweden conference, covering Nordic streaming architecture and client-side encryption talks.
- Ronald S. Bultje — Random thoughtsPersonal blog by ex-Google VP9 lead dev, dav1d/Eve encoder co-creator, Two Orioles founder. Deep VP9/AV1 codec internals writeups.