AI Sharing Circle

Daily sharing of the latest AI products, projects, frameworks, paper interpretations, etc.~
Ling-V2 - 蚂蚁百灵开源的MoE架构语言模型系列

Ling-V2 - The MoE Architecture Language Model Series of Ant Centurion Open Source

Ling-V2 is a family of large-scale language models based on the MoE architecture introduced by the Ant-Belling team. The first version, Ling-mini-2.0, has 16 billion total parameters, with only 1.4 billion parameters activated per input token.
10mos ago
051K
Xiaomi-MiMo-Audio - 小米开源的首个原生端到端语音大模型

Xiaomi-MiMo-Audio - Xiaomi Open Source's First Native End-to-End Speech Big Model

Xiaomi-MiMo-Audio is Xiaomi's open source 7-billion-parameter end-to-end speech macromodel with powerful features such as multi-language dialog, speech continuation, less-sample generalization, and audio understanding, which is able to reach the SOTA level in speech intelligence and audio understanding benchmarks, surpassing Google Gemi...
10mos ago
058.3K
WebWeaver - 阿里通义开源的新型双智能体框架

WebWeaver - Ali Tongyi open source new dual-intelligence body framework

WebWeaver is a new dual-intelligence body framework introduced by Alibaba Tongyi team, which is mainly used in open deep research, and can simulate the human research process, which is divided into two intelligences: planning and writing.
10mos ago
057K
MCP Registry - GitHub推出的官方MCP服务器管理平台

MCP Registry - The official MCP server management platform from GitHub.

The MCP Registry is a centralized platform from GitHub that helps developers discover and install MCP servers more easily.The MCP Registry is here to help developers quickly find the AI tools they need in one place, greatly simplifying...
10mos ago
054.5K
通义DeepResearch - 阿里通义开源的深度研究智能体

Tongyi DeepResearch - Ali Tongyi Open Source Deep Research Intelligence Body

Tongyi DeepResearch (Tongyi DeepResearch) is an open source intelligent body launched by Alibaba, designed for deep information retrieval and complex task reasoning, with 30 billion parameters, supporting multiple reasoning modes, including ReAct mode and deep mode...
10mos ago
060.4K
OpenAI《在AI时代保持领先》PDF指南 - 附下载链接

OpenAI's PDF Guide to Staying Ahead in the Age of AI - with Download Links

Staying ahead in the age of AI is an AI leadership guide from OpenAI that helps business leaders maintain a competitive edge in the age of AI. The guide points to the rapid growth of AI, with faster model releases, lower costs, and faster enterprise adoption...
10mos ago
060.9K
浙江大学免费PDF资料《大模型基础》 - 附下载链接

Free PDF of Fundamentals of Large Models from Zhejiang University - with download link

Fundamentals of Large Models provides an in-depth analysis of the core technologies and practical paths of Large Language Models (LLMs). Starting from the fundamental theory of language modeling, it systematically explains the principles of model design based on statistics, recurrent neural networks (RNN), and Transformer architecture, focusing on the three major big language model...
10mos ago
066.1K
LLaSO - 逻辑智能推出的业界首个全面开源的语音模型

LLaSO - The Industry's First Fully Open Source Speech Model from Logic Intelligence

LLaSO is an open source speech model launched by Beijing Depth Logic Intelligence Technology Co. Ltd, which solves the problems of data dispersion and insufficient task coverage in the field of large-scale speech language modeling by integrating speech and text data and providing alignment datasets, command fine-tuning datasets and evaluation benchmarks.
10mos ago
048.6K
混元3D 3.0 - 腾讯推出的3D生成模型,支持超高清建模

Hybrid 3D 3.0 - Tencent's 3D generated models with UHD modeling support

Hybrid 3D 3.0 is an advanced 3D generation model launched by Tencent, based on 3D-DiT hierarchical sculpting technology, with a geometric resolution of up to 1536³, capable of generating ultra-high-definition, detail-rich 3D models, and excelling in character modeling, with the ability to accurately shape the five senses and body shape.
10mos ago
067K
Mini-o3 - 字节、港大联合开源的视觉推理模型

Mini-o3 - Bytes, HKU Joint Open Source Visual Reasoning Model

Mini-o3 is an open source model jointly launched by ByteDance and the University of Hong Kong, focusing on solving complex visual search problems. The model has a powerful multi-round interactive reasoning capability, and can locate the target through deep exploration and trial-and-error.
10mos ago
052.4K