video-SALMONN-2
video-SALMONN 2 is a powerful audio-visual large language model (LLM) that generates high-quality audio-visual video captions, which is developed by the Department of Electronic Engineering at Tsinghua University and ByteDance.
Category
Audio engineering
Pricing
Free / open source
GitHub repo →
Top video-SALMONN-2 alternatives
mediapipeCross-platform, customizable ML solutions for live and streaming media.spleeterDeezer source separation library including pretrained models.eqMacmacOS System-wide Audio Equalizer & Volume Mixer 🎧pedalboard🎛 🔊 A Python library for audio.DALIA GPU-accelerated library containing highly optimized building blocks and an execution engine for data processing to accelerate deep learninseek-tuneAn implementation of Shazam's song recognition algorithm.
See the full list of video-SALMONN-2 alternatives →