AI Sharing Circle

Daily sharing of the latest AI products, projects, frameworks, paper interpretations, etc.~
LLaSO - 逻辑智能推出的业界首个全面开源的语音模型

LLaSO - The Industry's First Fully Open Source Speech Model from Logic Intelligence

LLaSO is an open source speech model launched by Beijing Depth Logic Intelligence Technology Co. Ltd, which solves the problems of data dispersion and insufficient task coverage in the field of large-scale speech language modeling by integrating speech and text data and providing alignment datasets, command fine-tuning datasets and evaluation benchmarks.
12mos ago
057.1K
混元3D 3.0 - 腾讯推出的3D生成模型,支持超高清建模

Hybrid 3D 3.0 - Tencent's 3D generated models with UHD modeling support

Hybrid 3D 3.0 is an advanced 3D generation model launched by Tencent, based on 3D-DiT hierarchical sculpting technology, with a geometric resolution of up to 1536³, capable of generating ultra-high-definition, detail-rich 3D models, and excelling in character modeling, with the ability to accurately shape the five senses and body shape.
1yrs ago
077.8K
Mini-o3 - 字节、港大联合开源的视觉推理模型

Mini-o3 - Bytes, HKU Joint Open Source Visual Reasoning Model

Mini-o3 is an open source model jointly launched by ByteDance and the University of Hong Kong, focusing on solving complex visual search problems. The model has a powerful multi-round interactive reasoning capability, and can locate the target through deep exploration and trial-and-error.
1yrs ago
063.1K
GPT-5-Codex - OpenAI推出的最强编程模型

GPT-5-Codex - The Most Powerful Programming Model Introduced by OpenAI

GPT-5-Codex is a powerful programming optimization model from OpenAI, further enhanced by GPT-5 and designed for software engineers. The model generates high-quality code quickly, supports multiple programming languages, and optimizes existing code to improve performance.
1yrs ago
062.5K
MiniMax Music 1.5 - MiniMax最新推出的AI音乐生成模型

MiniMax Music 1.5 - MiniMax's latest AI music generation model

MiniMax Music 1.5 is an advanced AI music generation tool that supports generating up to 4 minutes of music based on users' natural language descriptions. The model supports a variety of music styles and mood customization, generating a natural and full vocal color, smooth transitions, richly layered arrangements...
1yrs ago
073K
AnyI2V - 复旦联合阿里达摩院等开源的智能图像动画生成框架

AnyI2V - Fudan, Ali Dharma Institute and other open source framework for intelligent image animation generation

AnyI2V is an image animation generation framework jointly launched by Fudan University, Alibaba Dharma Institute and others, which supports the conversion of static conditional images (e.g., grids, point clouds, etc.) into dynamic videos without the need for complex training processes and large amounts of data.
1yrs ago
065.5K
SRPO - 腾讯混元推出的文本到图像生成模型

SRPO - Text-to-Image Generation Model launched by Tencent Mixed Meta

SRPO (Semantic Relative Preference Optimization) is a text-to-image generation model introduced by Tencent Hybrid, which optimizes the reward mechanism through text conditioned signals to achieve online adjustment of rewards and reduce offline fine-tuning dependency.
1yrs ago
082.3K
Qwen3-Next - 阿里通义推出的最新基础模型

Qwen3-Next - the latest base model from Ali Tongyi

Qwen3-Next is a new generation of hybrid architecture big model open source by Ali Tongyi, combining Gated DeltaNet and Gated Attention technology, good at dealing with long text, fast inference and saving computing resources.
1yrs ago
061.9K
文心大模型X1.1 - 百度推出的深度思考模型,理解能力更强

Wenshin Big Model X1.1 - Baidu's Deep Thinking Model for Better Understanding

Wenxin Big Model X1.1 is a deep thinking model launched by Baidu, based on a hybrid reinforcement learning framework that focuses on improving language understanding and generation. The model excels in handling complex questions, following instructions and simulating the behavior of intelligences, and can accurately provide knowledgeable answers and high-quality text content.
1yrs ago
067.4K
混元图像2.1 - 腾讯推出的开源文生图模型

Hybrid Image 2.1 - Tencent's Open Source Vendor Graph Model

HunyuanImage 2.1 is Tencent's open source graphic model, designed for high-quality image generation. The model supports native 2K resolution, can accurately render complex scenes and details, so that the character's expression and movement can be vividly reproduced.
1yrs ago
064.2K