
Mistral 7B
A powerful large-scale language model with about 7.3 billion parameters, developed by Mistral.AI, demonstrates excellent multilingual processing power and reasoning performance.
Meta introduces the world's first unified multimodal audio separation model that supports text, visual, and time cues to accurately separate target sounds from complex audio and video.
SAM Audio is the world's first unified multimodalaudio separationThe model realizes intelligent parsing and interactive extraction of complex audio scenes by fusing textual, visual and temporal cues. The core goal is to enable users to accurately isolate specific target sounds from mixed audio or video as if they were “listening with their eyes”, for example, by clicking on a musical instrument on the screen, typing text to describe the sound source, or marking a time segment, all of which can be accomplished with a single click.







