SceneSeg (Local-to-Global Scene Segmentation)
Description
CVPR 2020 multimodal movie scene segmentation using place and audio features on the MovieNet dataset; reference ML baseline with code.
Resource details
- Category path
- Media Tools›Scene Detection & Segmentation
- Provider
- Not yet classified
- Format
- Not yet classified
- Skill level
- Not yet classified
SceneSeg (Local-to-Global Scene Segmentation) is cataloged in Media Tools and Scene Detection & Segmentation. Provider: Not yet classified. Format: Not yet classified. Skill level: Not yet classified.
URL
Related resources
- SceneSegmentation-SCRLOfficial PyTorch implementation of CVPR 2022 Scene Consistency Representation Learning for movie-level scene segmentation (MovieNet/SceneSe…
- MDLSeg: Parameter-Free Video Scene Segmentation via Minimum Description LengthAcademic paper proposing a parameter-free scene-segmentation method claimed to outperform PySceneDetect.
- KVQ: NTIRE 2024 Short-form UGC Video Quality Assessment ChallengeCVPR NTIRE 2024 challenge dataset and code for short-form UGC video quality assessment, 600 videos/3600 processed clips.
- GHVQ: Benchmark and Metric for AI-Generated Human-Activity Video QualityAcademic paper introducing a benchmark dataset and objective quality metric for AI-generated human-activity video, with companion code.
- SparseSync: Audio-Visual Synchronisation with Trainable SelectorsBMVC 2022 Spotlight paper and code for sparse-in-space-and-time audio-visual sync detection.
- 2BiVQA: Double Bi-LSTM No-Reference Video Quality AssessmentNo-reference VQA model for UGC video using double Bi-LSTM, companion code to arXiv:2208.14774.
- Evaluating Video Quality Metrics for Neural and Traditional Codecs (4K/UHD-1)Academic comparison of VMAF, AVQBits, and FasterVQA no-reference quality metrics on 4K content.
- CompressedVQA: Full/No-Reference VQA for Compressed UGC VideoICME 2021 grand challenge-winning full-reference and no-reference video quality models for compressed user-generated content.
- nr-vqa-consumervideoNo-reference video quality metric tuned for consumer video capture artifacts (sensor noise, motion blur, shake) rather than compression art…
- aws-batch-with-FFmpegReference architecture running FFmpeg containers (ARM64/x86-64/NVIDIA/Xilinx) on AWS Batch with Spot compute environments and SDK/REST job…