Latest AI Resources

Total 3143 articles posts
AsrTools:语音转字幕工具,内置剪映、快手、必剪接口的轻量客户端

AsrTools: speech-to-subtitle tool, lightweight client with built-in interfaces to Cutscene, Racer, and Must-Cut

Comprehensive Introduction AsrTools is an intelligent speech-to-text tool with built-in interfaces from big players such as Cutscene, Racer, Must Cut, etc. It does not require GPU or cumbersome configuration, and supports efficient multi-threaded batch processing. It is based on PyQt5 development, beautiful and user-friendly interface, able to output SRT and TXT format words...
2yrs ago
0103.8K
通义万相:AI创意作画|文生图|图生图|虚拟模特|个人写真|涂鸦作画

Tongyi Wanxiang: AI Creative Painting|Text-to-Picture|To-Picture|Virtual Modeling|Personal Portrait|Doodle Painting

Comprehensive Introduction Tongyi Wanxiang is an AI creative painting platform under Aliyun, providing a variety of AI art creation functions. Users can create in a variety of ways such as text to generate images, image to generate images, graffiti painting, virtual modeling and personal portraits. The platform is based on the self-developed Composer combination of generating...
2yrs ago
0103.1K
Sonic:音频驱动肖像图片生成面部表情生动的数字人口播视频

Sonic: Audio-driven portrait images generate digital demo videos with vivid facial expressions

General Introduction Sonic is an innovative platform focusing on global audio perception designed to generate vivid portrait animations driven by audio. Developed by a team of researchers from Tencent and Zhejiang University, the platform utilizes audio information to control facial expressions and head movements to generate natural and smooth animated videos.S...
1yrs ago
0102.9K
紫东太初:多模态大模型平台,支持文本创作、图像生成、3D理解、信号分析等任务

Zidong Taichu: Multi-modal large model platform supporting text creation, image generation, 3D understanding, signal analysis and other tasks

Comprehensive Introduction Zidong Taichu is a new-generation multimodal big model platform launched by the Institute of Automation of the Chinese Academy of Sciences and the Wuhan Institute of Artificial Intelligence. The platform supports multiple tasks such as multi-round question and answer, text creation, image generation, 3D understanding and signal analysis, with powerful cognitive, understanding and creation capabilities. Zidong ...
2yrs ago
0102.7K
RD-Agent:自动化数据驱动研发工具,通过AI技术推动以数据为导向的研发过程

RD-Agent: an automated data-driven R&D tool to drive data-driven R&D processes through AI technology

Comprehensive Introduction RD-Agent is an open source tool from Microsoft designed to automate and optimize the research and development (R&D) process. The tool focuses on data-driven scenarios to improve the efficiency of model and data development through artificial intelligence techniques.RD-Agent integrates research...
1yrs ago
0102.6K
Llasa 1~8B:高品质语音生成和克隆的开源文本转语音模型

Llasa 1~8B: an open source text-to-speech model for high quality speech generation and cloning

General Introduction Llasa-3B is an open source text-to-speech (TTS) model developed by the Audio Lab of the Hong Kong University of Science and Technology (HKUST Audio). The model is based on the Llama 3.2B architecture, which has been carefully tuned to provide high-quality speech generation that not only supports multiple...
1yrs ago
0102.6K
AI投资系统:自动化A股投资决策系统,利用多智能体系统分析市场数据

AI investment system: automated A-share investment decision-making system that utilizes a multi-intelligence system to analyze market data

Comprehensive Introduction A_Share_investment_Agent is an A-share investment decision aid based on a multi-intelligence system. The system is designed to analyze market data, calculate the intrinsic value of stocks, analyze market sentiment, and fundamental data through multiple collaborative intelligences to...
2yrs ago
0102.5K
Qwen-Agent:基于Qwen的智能代理应用框架,包括工具调用、代码解释器、RAG和Chrome扩展。

Qwen-Agent: Qwen-based framework for intelligent agent applications, including tool calls, code interpreters, RAGs and Chrome extensions.

Comprehensive Introduction Qwen-Agent is an intelligent agent application framework developed based on Qwen 2.0 and above, with capabilities such as command following, tool usage, planning and memorization. The framework provides a variety of sample applications such as browser assistants, code interpreters and custom assistants...
2yrs ago
0102.4K
MagicQuill:智能交互式图像涂鸦编辑系统,精准局部涂鸦编辑

MagicQuill: Intelligent Interactive Image Graffiti Editing System, Precise Localized Graffiti Editing

General Introduction MagicQuill is an open-source AI interactive image editing tool jointly launched by Hong Kong University of Science and Technology (HKUST), Ant Group, Zhejiang University and University of Hong Kong. The tool aims to achieve accurate localized editing of images in an intelligent and interactive way.MagicQuill...
2yrs ago
0102.1K
混元文生视频:生成写实镜头感的高质量视频,腾讯开源视频生成大模型

Hybrid Vincennes video: generating realistic footage sense of high-quality video, Tencent open source video generation large model

Comprehensive Introduction Tencent Mixed Yuan Text Generation Video (available in Yuanbao APP) is a video generation platform based on AI technology launched by Tencent. The platform utilizes the Tencent Mixed Yuan Big Model with powerful cross-domain knowledge and natural language understanding to generate high-quality videos based on users' text descriptions...
2yrs ago
0101.9K
Agent TARS:使用视觉和命令操作电脑的开源智能体

Agent TARS: An Open Source Intelligence Using Vision and Commands to Operate Computers

Comprehensive Introduction Agent TARS is a multimodal AI intelligence open-sourced by ByteDance.The core feature is to visually understand web content and combine command line and file system operations to help users complete complex computer tasks. Instead of requiring manual operations like traditional tools, it can self...
1yrs ago
0101.9K
InvSR:开源图像超分辨率项目,提升图像分辨率质量

InvSR: Open source image super-resolution project to improve the quality of image resolution

General Introduction InvSR is an innovative open-source image super-resolution project based on diffusion inversion techniques capable of converting low-resolution images into high-quality, high-resolution images. The project utilizes the rich a priori knowledge of images embedded in pre-trained large-scale diffusion models to support, through a flexible sampling mechanism, the...
2yrs ago
0101.8K
Gauth(Gauthmath):使用AI解决作业问题,提供详细解答,字节旗下海外作业辅导APP

Gauth (Gauthmath): uses AI to solve homework problems and provide detailed answers, Byte's overseas homework tutoring app

General Introduction Gauth (formerly known as Gauthmath) is an AI homework helper website designed for students. It utilizes advanced AI technology and a team of professional tutors to provide homework answering services in a variety of subjects from math to chemistry. Users can upload an image or type in a question to quickly get...
1yrs ago
0101.8K
SadTalker:让照片说话|嘴型同步音频|合成口型同步视频|免费数字人

SadTalker: Make Photos Talk | Mouth Synchronized Audio | Synthesized Mouth Synchronized Video | Free Digital People

General Introduction SadTalker is an open source tool that combines a single still portrait photo with an audio file to create realistic talking avatar videos for a variety of scenarios such as personalized messages, educational content, and more. The revolutionary use of 3D modeling technologies such as ExpNet and PoseVA...
1yrs ago
0101.6K
ChatWiki:轻量级开源企业知识库AI问答系统

ChatWiki: lightweight open source enterprise knowledge base AI Q&A system

Comprehensive Introduction ChatWiki is an open source knowledge base AI Q&A system officially launched by Sesame Small Customer Service, built on Large Language Modeling (LLM) and Retrieval Augmented Generation (RAG) technology. It provides out-of-the-box data processing and model calling capabilities to help companies quickly build their own knowledge...
2yrs ago
0100.9K
SciSpace:一站式学术研究与论文写作平台,为学生和研究人员提供一体化 AI 工具

SciSpace: A one-stop academic research and paper writing platform with integrated AI tools for students and researchers

General Introduction SciSpace (formerly Typeset.io) is an AI-powered platform designed for academic research and writing. It provides a wealth of tools and resources to help researchers and students find, understand and write about literature more efficiently. The platform integrates literature management, automatic gr...
2yrs ago
0100.8K
Dessix.io:灵感捕捉、思维组织和AI协作于一体的笔记工具

Dessix.io: inspiration capture, thought organization and AI collaboration in one note-taking tool

General Introduction Dessix.io is an all-in-one note-taking tool with integrated AI collaboration features designed to help users capture inspiration, organize their thoughts and create efficiently. With Dessix, users can easily collect web content or text snippets, utilize AI to automatically generate summaries and keywords, and simplify letter...
2yrs ago
0100.7K
InstantIR:受损图像修复与图像高清放大开源项目,最低16G显存

InstantIR: damaged image repair and image high-definition zoom open source project, minimum 16G video memory

General Description InstantIR is an innovative single-image restoration model developed by the InstantX team, designed to resurrect your damaged images with extremely high-quality and realistic details, capable of high-quality restoration of damaged images. The tool not only restores the details of the image...
2yrs ago
0100.6K
TRELLIS:Microsoft开发的3D资产生成模型,支持多种格式和灵活编辑

TRELLIS: Microsoft-developed 3D asset generation model with multiple format support and flexible editing

General Introduction TRELLIS is a large-scale 3D asset generation model developed by Microsoft. It is capable of receiving text or image prompts and generating high-quality 3D assets in a variety of formats, such as radial fields, 3D Gaussians, and meshes.At the heart of TRELLIS is a unified structured latent...
2yrs ago
0100.5K