
Canopy MultimediaLarge ModelIt is developed by Wanxing Technology, which is China's first large model of audio and video multimedia creation pendant class, with powerful audio and video generation and processing capabilities.
I. Composition and function
Canopy Multimedia Grand Model consists of Video Grand Model, Audio Grand Model, Picture Grand Model, and Language Grand Model with the following core capabilities:
- One-click movie: can quickly generate high-quality audio and video content.
- AI Art Design: provides advanced image and graphic design capabilities.
- Text-generated music: Generate music that matches the scene based on the text description.
- Audio Enhancement: Optimize audio quality to enhance the listening experience.
- Audio Analysis: Deeply analyze the audio to extract key information.
- Multi-language dialog: supports natural dialog interaction in multi-language environments.
II. Technical characteristics
- Multimedia Fusion: The Canopy Multimedia Grand Model integrates multimedia elements such as video, audio, pictures and language to realize the fusion generation of multimedia content.
- Vertical solutions: Provide specialized solutions for digital creative pendant creation scenarios to meet the needs of creation in specific fields.
- Arithmetic data and application localization: Based on 1.5 billion user behaviors and 10 billion localized high-quality audio and video data deposits, training is conducted on the basis of domestic arithmetic and servers to ensure the efficiency and applicability of the model.
III. Application and iteration
- Application Scenario: Canopy multimedia big model has been applied in video production, movie production, advertising industry and other markets, and shows strong driving force and innovation.
- Capability Iteration: Currently, the Canopy Multimedia Big Model has iterated nearly 100 audio and video atomic capabilities, including Vincent theme video, Vincent 3D video, AI singers, and video AI soundtracks,digital personbroadcasting, etc., and continue to iterate on key multimodal capabilities such as vision and hearing.
IV. Cooperation and strategy
Wanxing Technology has reached a tripartite arithmetic cooperation with Matou Arithmetic and Huawei Cloud, and has reached a strategic cooperation on large-model arithmetic with CAGT, to jointly promote the application and development of high-quality arithmetic in the era of large models.
Canopy Multimedia Big Model is a multimedia authoring pendant big model based on audio/video generative AI technology, which has powerful audio/video generating and processing capabilities to provide professional solutions for digital creative pendant authoring scenarios. Through continuous technology iteration and cooperation strategy implementation, Canopy Multimedia Model is leading the new wave of audio and video creation.
data statistics
Relevant Navigation

AliQianwen, the first native language world model released on June 24, 2026, uses plain text to uniformly simulate seven major digital environments, allowing agents to practice in a "virtual world."

DeepSeek-V3
Hangzhou Depth Seeker has launched an efficient open source language model with 67.1 billion parameters, using a hybrid expert architecture that excels at handling math, coding and multilingual tasks.

DeepSeek-Math-V2
The world's first large model of mathematical reasoning in open source form to reach the gold medal level of the International Mathematical Olympiad (IMO), realizing the rigor of reasoning and the ability to solve difficult mathematical problems through a self-verification framework.

Tongyi LM
Launched by AliCloud, the ultra-large-scale pre-trained language model has powerful natural language processing and comprehension capabilities, and is able to simulate human thinking for tasks such as multi-round conversations and copywriting, and serves a number of industries and scenarios to provide users with intelligent solutions.

InternLM
Shanghai AI Lab leads the launch of a comprehensive big model research and development platform, providing an efficient tool chain and rich application scenarios to support multimodal data processing and analysis.

ZhiPu AI BM
The series of large models jointly developed by Tsinghua University and Smart Spectrum AI have powerful multimodal understanding and generation capabilities, and are widely used in natural language processing, code generation and other scenarios.

ERNIE X1 Turbo
Baidu has launched a new generation of high-level AI assistants to disassemble complex tasks and automate the entire process with autonomous deep thinking, multimodal toolchain invocation and extreme cost advantages.

SenseNova
Shangtang Technology has launched a comprehensive big model system with powerful natural language processing, text-born diagrams and other multimodal capabilities, aiming to provide efficient AI solutions for enterprises.
No comments...
