AI Sharing Circle

AI is changing the world!
Nano Banana - 谷歌推出的AI图像编辑模型

Nano Banana - AI image editing model launched by Google

Nano Banana is the Gemini 2.5 Flash Image codename for Gemini, an AI image generation and editing model from Google that generates detailed, photorealistic images based on simple text prompts to make high-quality modifications to existing images.
11mos ago
087.6K
Genie Envisioner - 智元联合北航等开源的通用机器人操作平台

Genie Envisioner - Jiyuan's open-source general-purpose robotics platform with Beihang and others

Genie Envisioner (GE) is a unified platform for robot operation developed by the Genie Robotics team in collaboration with the National University of Singapore, Beijing University of Aeronautics and Astronautics and other organizations. It allows robots to better understand and perform tasks by "imagining first, then acting".
11mos ago
061.3K
DINOv3 - Meta AI推出的新一代自监督视觉基础模型

DINOv3 - Next Generation Self-Supervised Vision Base Model from Meta AI

DINOv3 is a next-generation self-supervised vision base model from Meta AI, which adopts a self-supervised learning paradigm to learn image features without labeling data. It solves the feature degradation problem by improving data preparation and introducing Gram anchoring, and improves the generalization...
11mos ago
073.8K
Matrix-Game 2.0 - 昆仑万维开源自研的交互式世界模型

Matrix-Game 2.0 - Interactive World Model developed by KunlunWanwei

Matrix-Game 2.0 is a self-developed interactive world model released by Kunlun SkyWork AI. Matrix-Game 2.0 is the industry's first open-source, real-time, long-sequence interactive generation model for general-purpose scenarios. The model is able to run at 25 FPS through a visually-driven interaction scheme in multiple...
12mos ago
067.3K
Baichuan-M2 - 百川智能推出开源的医疗增强大模型

Baichuan-M2 - Baichuan Intelligence Launches Open Source Healthcare Enhanced Big Model

Baichuan-M2 is an open source medical augmented large model launched by Baichuan Intelligence. It performs well in the medical field, especially in the HealthBench review with a score of 60.1, surpassing OpenAI's gpt-oss120b and many other open source models, becoming a global...
12mos ago
066.2K
Qwen-Flash - 通义千问推出的高性能、低成本语言模型

Qwen-Flash - A high-performance, low-cost language model from Tongyi Chien-quan

Qwen-Flash is a high-performance, low-cost language model introduced in the Alibaba Tongyi Thousand Questions series, designed for fast response and efficient processing of simple tasks. Based on the advanced Mixture-of-Experts (MoE) architecture, it is realized by sparse expert network...
12mos ago
064.8K
SkyReels-A3 - 昆仑万维推出的音频驱动数字人创作工具

SkyReels-A3 - Audio-Driven Digital Human Creation Tool from KunlunWangwei

SkyReels-A3 is an audio-driven digital human creation tool from Kunlun World Wide Group. SkyReels-A3 is an audio-driven digital human creation tool, which can generate high-quality dynamic video content through simple inputs (e.g., portrait images and voice), make static photos "come alive", and replace lines for existing videos with new lip-syncs that the characters will automatically...
12mos ago
060.8K
MiniMax Speech 2.5 - MiniMax推出的语音生成模型

MiniMax Speech 2.5 - Speech Generation Model from MiniMax

MiniMax Speech 2.5 is an advanced speech generation model developed by MiniMax team. It has made significant progress in the field of speech synthesis, especially in multilingual expressiveness, timbre reproduction accuracy and language coverage. The model supports 40 languages...
12mos ago
068K
GPT-5 - OpenAI推出的最强语言模型,统一智能系统

GPT-5 - The Strongest Language Model Introduced by OpenAI, Unified Intelligence System

GPT-5 is the latest language model released by OpenAI with several upgrades. It is a unified intelligence system with a built-in real-time router that automatically switches between efficient and deep thinking modes according to the complexity of the problem, realizing fast response and accurate answers.GPT-5 has several versions, including the one for general...
12mos ago
066.4K
dots.vlm1 - 小红书hi lab开源的多模态大模型

dots.vlm1 - Small red book hi lab open source multimodal big model

dots.vlm1 is the first multimodal big model open-sourced by Little Red Book hi lab. Based on NaViT, a 1.2 billion parameter visual encoder trained from scratch, and DeepSeek V3 Large Language Model (LLM), it has powerful visual perception and text inference...
12mos ago
066.1K