
SenseNovaLarge ModelIt is a comprehensive big modeling system launched by ShangTech.
Background & Release:
- Day by day new SenseNova is a big modeling system announced by ShangTech in April 2023 by Chairman and CEO Xu Li.
- The system was approved to go live in August 2023, marking its official availability to the public.
- On May 29th, 2024, ShangTech announced that its "DayDayNew" Big Model will soon undergo a major upgrade, and officially released the DayDayNew Big Model 5.0 Cantonese version to the public.
Main Functions and Features::
- natural language processing (NLP): Automatically transforms data into meaningful analytics and visualization results through a combination of code generation and automated execution through capabilities such as natural language generation, intent recognition, logic understanding and code interpreters.
- Vincentian graphic ability (geology):: Includesdigital personThe video generation platform "SenseAvatar" (SenseAvatar) and other features can provide users with rich visual content generation services.
- Model Development Functions: Support users to develop models according to their needs and provide customized AI solutions.
Technical characteristics::
- Hybrid Expert Architecture (MoE): Nisshin SenseNova 5.0 utilizes the MOE Hybrid Expert Architecture, an architecture that allows the model to complete inference with a small number of parameters activated, improving the model's processing efficiency and responsiveness.
- Volume of training data: Based on more than 10TB tokens of training data, it ensures that the model has a robust knowledge base and a wide range of applications.
- inference context window: Reaching around 200K allows the model to handle longer text sequences and more complex contextual relationships.
Performance benchmarking::
Rizhixin SenseNova 5.0 fully benchmarks the GPT-4 Turbo in terms of comprehensive performance and meets or exceeds the GPT-4 Turbo in mainstream objective reviews, especially in terms of natural language capability, text-to-graph capability, multimodal and data analysis capability.
application scenario::
Risen SenseNova big models have been widely used in many fields such as finance, healthcare, education, etc., such as intelligent customer service, intelligent marketing, investment research analysis, research report writing, medical and healthcare language big models, etc., which provide powerful AI support for the industry.
language version::
In addition to the standard version, ShangTech has also released the Day Day New Big Model 5.0 Cantonese version, which further extends the model's language support capabilities.
Rizhixin SenseNova Big Model is a comprehensive big model system launched by ShangTech, with powerful natural language processing, text-born graph capabilities and model development functions. Through advanced hybrid expert architecture and a large amount of training data, it realizes performance comparable to GPT-4 Turbo, and is widely used in a variety of fields, providing users with efficient and flexible AI services.
data statistics
Relevant Navigation

OpenAI introduces small AI models with inference capabilities and cost-effective pricing, designed for developers and users to optimize application performance and efficiency.

Bunshin Big Model X1
Baidu launched an advanced large language model with deep thinking, multi-modal support and multi-tool invocation capabilities to meet the needs of multiple domains with excellent performance, affordable price and rich functionality.

InternLM
Shanghai AI Lab leads the launch of a comprehensive big model research and development platform, providing an efficient tool chain and rich application scenarios to support multimodal data processing and analysis.

Evo 2
The world's largest biology AI model, jointly developed by multiple top organizations, is trained based on massive genetic data and can accurately predict genetic variants and generated sequences to help breakthroughs in life sciences.

Doubao
ByteDance launched a self-developed big model. Through byte jumping internal 50 + business scene practice verification, daily 100 billion tokens large use of continuous polishing, to provide multi-modal capabilities, with high quality model effect for the enterprise to create a rich business experience

Ovis2
Alibaba's open source multimodal large language model with powerful visual understanding, OCR, video processing and reasoning capabilities, supporting multiple scale versions.

TranslateGemma
Google's open source lightweight multimodal translation model supports 55 languages and image translations, with performance that exceeds larger models, taking into account both mobile and cloud deployments, and facilitating efficient globalized communication.

SKYMEDIA
Wanxing Technology has developed China's first audio and video multimedia creation pendant big model, which integrates video, audio, picture and language processing capabilities to provide powerful AI creation support for the digital creative field.
No comments...
