Media Tools
Page 15 of 30: Explore 704 curated media processing and video editing tools for transcoding, analysis, and manipulation on Awesome Video.
About this collection
Media Tools brings together 704 curated resources from across the video technology landscape. Notable topics include AI & Machine Learning Tools, Ads & QoE, Audio & Subtitles, and Audio Analysis & Processing.
Subcategories
- AI & Machine Learning Tools
- Ads & QoE
- Audio & Subtitles
- Audio Analysis & Processing
- Batch Processing & Automation
- Color Grading & Correction Tools
- Color Science & Histogram Analysis
- Conversion & Format Tools
- Effects & Compositing Tools
- Metadata Extraction & Management
- Non-linear Editing Suites
- Quality Analysis & Metrics
- Scene Detection & Segmentation
- Subtitle & Caption Tools
- VMAF PSNR SSIM Tools
- Video Analytics & Benchmarking
Resources
- imscJS: IMSC/TTML/SMPTE-TT/EBU-TT-D rendererRenders IMSC/TTML/SMPTE-TT/EBU-TT-D subtitles and captions to HTML
- pyvideotrans — video translation/dubbing pipelineEnd-to-end pipeline: ASR → subtitle translation → multi-role dubbing → audio-video re-sync, offline or via API models.
- stable-tsWhisper-based transcription with forced alignment and audio indexing, including a mode that loads audio in 30s chunks for near-streaming su…
- subconv — SCC (CEA-608) to WebVTT converterRuby library converting SCC (EIA-608/CEA-608) caption files to WebVTT; niche 608 conversion path.
- ttconv: subtitle/caption format converterConverts EBU STL, IMSC/TTML/SMPTE-TT/EBU-TT-D and 608/SCC into IMSC, WebVTT and SRT
- whisper-vtt2srt — AI transcript WebVTT→SRT cleanerZero-dependency CLI/Python tool converting Whisper WebVTT to clean SRT, fixing karaoke-accumulation and filtering glitches.
- Audio Offset FinderFind the offset of an audio file within another audio file
- AudioAlignGUI research tool for aligning/synchronizing audio and video recordings of the same event captured from multiple sources.
- Chromaprint — audio fingerprinting libraryC library and fpcalc CLI generating acoustic fingerprints from PCM audio, used for content identification and dedup in media pipelines.
- EBU ADM ToolboxConfig-driven eat-process CLI for validating and transforming ADM (ITU-R BS.2076) next-generation-audio files, building on libadm, libbw64,…
- EBU ADM Toolbox DocumentationOfficial documentation for eat-process configuration, ADM validation profiles and processing blocks in the EBU ADM Toolbox.
- FFmpeg NormalizeAudio Normalization for Python/ffmpeg.
- MPEG-H Authoring SuiteOfficial toolset to author, monitor, analyze and encode MPEG-H Audio, with MPF and BWF/ADM import/export for immersive and personalized aud…
- Olaf (Overly Lightweight Acoustic Fingerprinting)Portable C library for acoustic fingerprinting, runs on embedded devices and WASM.
- PanakoContent-based audio search and fingerprinting system robust to noise and distortion, with stream-monitoring subapps, academic-backed.
- R128 (NDR)GUI loudness measurement app wrapping ffmpeg/ffprobe for EBU R128, built by German broadcaster NDR.
- SparseSync: Audio-Visual Synchronisation with Trainable SelectorsBMVC 2022 Spotlight paper and code for sparse-in-space-and-time audio-visual sync detection.
- SpectroMap: Peak Detection Algorithm for Audio FingerprintingAcademic paper on a peak-detection algorithm for audio fingerprinting with implementation.
- SyncSinkSynchronizes media files sharing common audio by computing sub-second offsets to align multi-camera/multi-mic recordings.
- ViSQOL v3 — Objective Audio and Speech Quality MetricGoogle's full-reference perceptual audio/speech quality metric producing MOS-LQO scores, with C++ CLI/API, Python bindings and a retrainabl…
- Virinext Bitstream AnalyzerBitstream syntax analyser and validator for AAC, AC-3/E-AC-3, MPEG-2 Audio, FLAC, Opus and PCM elementary streams.
- ac3esbrowserGUI analyzer that displays syntax elements of AC-3 / E-AC-3 (Dolby Digital / DD+) elementary streams.
- audalignPython package that aligns audio files via fingerprinting/cross-correlation/spectrograms and inserts silence to auto-sync multi-source capt…
- audfprintLandmark-based audio fingerprinting in Python, Shazam-style, supports fingerprint databases of ~1M entries.