TimeSide
TimeSide is a Python framework for scalable audio processing, analysis, imaging, transcoding, streaming, and labeling. It provides a high-level API designed to enable complex processing on large datasets of audio or video assets with a plug-in architecture and an extensible web f
Link
Related resources
- audioprintPython MFCC-based perceptual-hash audio fingerprinting library.
- Eyevinn auto-subtitlesWhisper-based automatic subtitle generation tool that chunks large audio and outputs VTT/SRT/JSON while preserving sync.
- python-video-converterPython library wrapping ffmpeg/ffprobe with a declarative Converter().convert() API for scripting format/audio/video conversion pipelines.
- Normalize-AudioBatch EBU R128 loudness normalization tool for MKV files via ffmpeg, preserves video stream and caches loudness analysis.
- BackgroundMattingV2Real-time high-resolution background matting requiring a captured background plate, achieving 4K 30fps on RTX 2080 Ti, companion to arXiv:2…
- aws-batch-with-FFmpegReference architecture running FFmpeg containers (ARM64/x86-64/NVIDIA/Xilinx) on AWS Batch with Spot compute environments and SDK/REST job…
- Evaluating Video Quality Metrics for Neural and Traditional Codecs (4K/UHD-1)Academic comparison of VMAF, AVQBits, and FasterVQA no-reference quality metrics on 4K content.
- Bitrate Ladder Construction via Transfer Learning and Spatio-Temporal FeaturesML-based bitrate ladder optimization achieving 94.1% complexity reduction at only 1.71% BD-rate cost using transfer learning.
- VMAF-torchPyTorch reimplementation of VMAF metric, GPU/gradient-friendly for use in learned codec optimization loops.
- CompressedVQA: Full/No-Reference VQA for Compressed UGC VideoICME 2021 grand challenge-winning full-reference and no-reference video quality models for compressed user-generated content.