Latest AI Resources

Total 3185 articles posts
Tizzy.ai - 百度推出的AI搜索应用

Tizzy.ai - AI search app launched by Baidu

Tizzy.ai is an AI intelligent search application launched by Baidu.Tizzy.ai is based on Baidu's big model technology, with powerful intelligent search functions, can quickly answer questions, deep thinking and assist in decision-making.Tizzy.ai has a simple interface, no ads and pop-ups, and the bottom of the guide...
1yrs ago
081.7K
MagicTryOn - 浙大和vivo等机构推出的视频虚拟试穿框架

MagicTryOn - Video Virtual Try-On Framework from ZJU and Vivo and others

MagicTryOn is an advanced video virtual try-on framework launched by the School of Computer Science and Technology of Zhejiang University in collaboration with vivo and other organizations. The framework replaces the traditional U-Net architecture with an innovative Diffusion Transformer (DiT) architecture, combined with a fully self-attentive machine...
1yrs ago
081.7K
袋鼠参谋 – 美团推出的商家AI智能决策应用

Kangaroo Staff - AI Intelligent Decision Making App for Merchants by Meituan

Kangaroo Staff is a merchant-oriented AI intelligent decision-making application launched by Meituan to help merchants solve problems in store opening and operation. Based on Meituan's massive catering data and more than 10 years of online operation experience, through conversational interaction, it provides merchants with precise information on track selection, store opening location, dish development, store operation and other scenarios, such as...
1yrs ago
081.7K
BAGEL - 字节跳动推出的开源多模态基础模型

BAGEL - Open source multimodal base model launched by Wordpress

BAGEL is a multimodal base model open-sourced by ByteDance with 14 billion parameters, of which 7 billion are active. The model base with the Mixed Transformer Expert Architecture (MoT) captures pixel-level and semantic-level features of an image with two independent encoders, respectively, to support efficient processing of images, text, video...
1yrs ago
081.6K
AI抖音 - 抖音推出的智能深度思考与搜索应用

AI Jitterbug - Intelligent Deep Thinking and Search App by Jitterbug

AI Jitterbug is an intelligent deep thinking and search application launched by Jitterbug to provide users with a more efficient and intelligent content acquisition experience. Based on Jitterbug's powerful content ecosystem and AI technology, it provides users with more comprehensive and detailed answers through connected search and reasoning capabilities.
1yrs ago
081.5K
冷饭热炒:流畅的创作儿童故事视频

Cold food, hot food: smooth creation of children's storytelling videos

The hottest money-making idea in the early days of AI tutorials was to teach you to create children's story illustrated books, which later spawned children's story animations, these videos are as easy to watch as they are to operate, and then they get stuck. The reason for this is still a slightly high threshold, and the quality of the content and generated images is worrying. Until ... gpt and cloude generation...
2yrs ago
081.5K
Sparkify - 谷歌推出的AI动画视频生成平台

Sparkify - AI animated video generation platform from Google

Sparkify is an AI animated video generation platform launched by Google. The platform is based on Gemini 2.5 and Veo2 models, users input questions or complex concepts, Sparkify in 2 minutes to generate intuitive animated short videos to explain the relevant knowledge points. The platform supports text, image and...
1yrs ago
081.5K
D-Human:克隆数字人营销短视频生产专家

D-Human: Cloning Digital Human Marketing Short Video Production Experts

Comprehensive introduction D-Human is a digital human video production platform, invested by Xiaomi, led by a doctor of Chinese Academy of Sciences research and development. It supports SaaS, API, OEM multiple cooperation methods, and provides 1:1 real life restoration technology, 8 minutes of video material to clone yourself or others. The platform greatly reduces the creation...
2yrs ago
081.4K
GenFlow超能搭子 – 百度文库推出的通用AI Agent

GenFlow Super Hitchhiker - Generalized AI Agent from Baidu Literature Library

GenFlow Super Hitchhiker is a general-purpose AI Agent launched by Baidu Literature Library, which allows users to autonomously disassemble tasks, call up Baidu Literature Library's 1.4 billion document libraries and online resources, and generate PPTs, reports, charts, posters, and other full-modal content in an extremely fast manner by simply typing in the natural language commands.
1yrs ago
081.1K
Make - AI无代码自动化工作流搭建平台

Make - AI's no-code automated workflow building platform

Make is an AI-driven no-code automation platform that helps organizations improve efficiency and innovation based on automated processes. The platform offers more than 2,000 pre-built apps that support a variety of business scenarios, such as marketing, sales, finance, etc. Make's core features include no-code visual process creation, AI...
1yrs ago
081K
Comate AI IDE - 文心快码推出多模态、多智能体协同的AI IDE

Comate AI IDE - Wenshin Quickcode Launches Multimodal, Multi-Intelligence Body Collaboration AI IDE

Comate AI IDE is the industry's first multi-modal, multi-intelligence body collaborative AI native IDE launched by Baidu Wenshin Express Code, with powerful multi-modal capabilities, support for the design draft of a key to the code (F2C), image to the code, and the natural language to the code, the front-end development scenarios in the performance of the outstanding...
1yrs ago
080.8K
琴乐大模型 - 腾讯推出的AI音乐创作模型

Piano Music Big Model - AI Music Composition Model by Tencent

Qin Music Grand Model is an advanced AI music creation grand model jointly launched by Tencent AI Lab and Tencent TME Tianqin Lab. The model intelligently generates high-quality stereo audio or multi-track sheet music based on user-inputted keywords, descriptive statements or audio clips in English and Chinese.
1yrs ago
080.5K
Mapify - XMind推出的AI思维导图生成工具

Mapify - AI Mind Map Generator from XMind

Mapify is an AI mind map generator tool launched by XMind team. It can quickly convert text, PDF, web pages, video, audio and other formats into structured mind maps, helping users efficiently extract and organize key information.
1yrs ago
080.5K
MineContext - 字节开源的主动式上下文感知AI伙伴

MineContext - Bytes Open Source Active Context-Aware AI Partner

MineContext is an active context-aware AI partner open-sourced by the ByteDance Viking team to help users efficiently manage massive amounts of information and improve the efficiency of knowledge work. Over the screenshot and content understanding technology, automatically record the user's daily operations (such as browsing the web, editing documents, etc.), support...
11mos ago
080.5K
奇妙元:数字人视频制作与直播服务平台|声音克隆|形象克隆

WonderWon: Digital Human Video Production and Live Streaming Service Platform|Voice Cloning|Image Cloning

Comprehensive Introduction Wonderful Dollar is a platform for digital persona video production and live streaming services, providing the function of generating videos from photos and PPTs, as well as the service of translating videos into different languages. Users can customize digital characters for news reporting, educational content, corporate promotion and many other areas. The platform also provides interactive digital staff...
2yrs ago
080.2K
Paper2Slides - 香港大学开源的学术论文转为幻灯片AI工具

Paper2Slides - HKU open source academic papers into slides AI tool

Paper2Slides is an open source AI tool from the Data Intelligence Laboratory of the University of Hong Kong that converts academic papers into professional slides or posters in one click. Using RAG (Retrieval Augmented Generation) technology, directly parsing the document content rather than relying on network information, to ensure that the generated PPT is highly consistent with the original...
9mos ago
080K
Gemini CLI - 谷歌开源的编程Agent

Gemini CLI - Google Open Source Programming Agent

Gemini CLI is Google's open source AI programming tool based on incorporating the Gemini Big Model into the developer's endpoint to provide developers with powerful AI capabilities. The tool understands code, manipulates files, executes commands, and dynamically troubleshoots problems to help developers efficiently write generation...
1yrs ago
079.6K
有道数字人:虚拟形象播报与实时交互平台|免费制作克隆数字人

Arigatou Digital Human: Virtual Image Broadcasting and Real-time Interaction Platform|Free Clone Digital Human Creation

Comprehensive introduction Wealth Digital People is a platform that integrates advanced AI technology, focusing on providing virtual image broadcasting and real-time interactive services. The platform utilizes self-developed speech recognition, speech synthesis, multimodal perception and document Q&A technologies to create realistic digital human doppelgangers for users, supporting video production, translation, teaching...
2yrs ago
079.1K
有道小P - 网易有道推出的新一代AI全科学习助手

Youdao Xiao P - A new generation of AI general learning assistant launched by NetEase Youdao

Youdao Little P is an AI all-subject learning assistant launched by NetEase Youdao, designed for K12 students, equipped with the Youdao Ziyi education big model, covering elementary school, middle school and high school all-subject Q&A, providing personalized learning advice. With AI word search and AI translation functions, Youdao Little P helps students quickly solve language problems...
1yrs ago
078.9K
Seed-OSS - 字节跳动团队开源的全新AI模型

Seed-OSS - A new AI model open-sourced by the Wordpress team

Seed-OSS is a large family of language models open-sourced by the Byte Jump Seed team, focusing on long text and reasoning tasks. The model performs well in complex logical reasoning and multi-step reasoning, with high accuracy and efficient problem solving.Seed-OSS supports long text contexts up to 512K...
1yrs ago
078.8K
VoxCPM - 面壁智能联合清华开源的端到端TTS模型

VoxCPM - Faceted Intelligence and Tsinghua Open Source End-to-End TTS Model

VoxCPM is a speech generation model jointly open-sourced by Facade Intelligence and Shenzhen International Graduate School of Tsinghua University.VoxCPM adopts an end-to-end diffusion autoregressive architecture to generate continuous speech representations directly from text, breaking through the limitations of traditional discrete disambiguation. Through hierarchical language modeling and finite state quantization...
1yrs ago
078.3K
Qwen-Image-Edit - 阿里通义开源的图像编辑模型

Qwen-Image-Edit - Ali Tongyi open source image editing model

Qwen-Image-Edit is an all-purpose image editing model introduced by Ali Tongyi, built on the Qwen-Image architecture with 20 billion parameters. The model combines both semantic and appearance editing capabilities, and can perform low-level visual appearance editing on images (e.g., adding, deleting...
1yrs ago
077.8K
RoboOS 2.0 - 智谱开源的跨本体具身大小脑协作框架

RoboOS 2.0 - Wisdom Spectrum's Open Source Cross-Ontology Embodied Brain-Size Collaboration Framework

RoboOS 2.0 is an open-source framework for cross ontology brain collaboration, promoting the transformation of robots from single intelligence to collaborative group intelligence. The framework realizes efficient division of labor with a "big brain" architecture, where the cloud brain is responsible for complex decision-making and collaboration, and the small brain module focuses on executing specific skills.
1yrs ago
077.6K
MuseSteamer - 百度推出的视频生成大模型

MuseSteamer - Baidu Launches Big Model for Video Generation

MuseSteamer is a large model for multimodal video generation launched by Baidu. The model can quickly generate high-quality dynamic video content based on text descriptions or images provided by the user, supporting a variety of clarity and functionality versions to meet the needs of different scenarios of creation.
1yrs ago
077.5K
优雅YOYA - 中科闻歌推出的AI音视频内容创作平台

Elegant YOYA - AI Audio/Video Content Creation Platform Launched by Sinotech Winkler

Elegant YOYA is a multimodal literate video platform launched by Zhongke Wenge, the platform is based on AI multimodal technology to empower the whole chain of video content creation. Users only need to input the theme requirements, the platform can quickly generate scripts, images, videos, and can complete intelligent editing, voice synthesis and character mouth drive and other operations, the output...
1yrs ago
077.4K