Latest AI Resources

Total 3143 articles posts
Fish Agent:端到端AI语音克隆助手,实时语音对话助理,Fish Speech衍生项目

Fish Agent: end-to-end AI voice cloning assistant, real-time voice conversation assistant, Fish Speech spin-off project

Comprehensive Introduction Fish Speech Derivative Project Fish Agent is a revolutionary end-to-end AI speech cloning system developed based on the V0.1 3B model architecture. As a fully end-to-end speech clone processing system, its most important feature is the use of innovative speechless...
2yrs ago
094.7K
法行宝:AI法律顾问,人工智能法律咨询,百度AI法律平台

Fa Xing Bao: AI Legal Advisor, Artificial Intelligence Legal Consultation, Baidu AI Legal Platform

Comprehensive Introduction LawXinbao is an intelligent legal service platform launched by Baidu, which integrates advanced artificial intelligence technology with a professional legal knowledge base. The platform is dedicated to providing users with convenient and professional legal intelligent services, including intelligent legal Q&A, case analysis, contract review and other functions. Through deep learning...
2yrs ago
094.7K
AutoAgent:通过自然语言快速创建并部署AI智能体的框架

AutoAgent: a framework for rapid creation and deployment of AI intelligences through natural language

General Introduction AutoAgent is an open source AI intelligences framework developed by the Data Intelligence Laboratory of the University of Hong Kong (HKUDS) and hosted on GitHub.It allows users to rapidly create and deploy customized AI intelligences by describing their requirements in purely natural language, without any programming base...
1yrs ago
094.7K
VideoRAG:理解超长视频的RAG框架,支持多模态检索和知识图谱构建

VideoRAG: A RAG framework for understanding ultra-long videos with support for multimodal retrieval and knowledge graph construction

Comprehensive Introduction VideoRAG is a retrieval-enhanced generative framework designed for processing and understanding very long contextual videos. The tool combines a graph-driven textual knowledge base with hierarchical multimodal context encoding to efficiently process on a single NVIDIA RTX 3090 GPU...
1yrs ago
094.7K
通义听悟:阿里通义音视频内容转录AI助手

Tongyi Listening and Understanding: Ali Tongyi Audio and Video Content Transcription AI Assistant

Comprehensive Introduction Tongyi Listening and Understanding is a work-study AI assistant launched by Aliyun, focusing on transcribing and analyzing audio and video content. It relies on AliCloud's powerful AI models to transcribe audio and video content into text in real time, and provides translation, summarization, positioning and other functions. Tongyi Listening Woo supports multiple languages and scenarios...
2yrs ago
094.6K
RunPod:专为AI设计的GPU云服务,快速冷启动SD且按秒付费

RunPod: GPU Cloud Service Designed for AI with Fast Cold Start SD and Pay Per Second

Comprehensive Introduction RunPod is a cloud computing platform designed for AI, aiming to provide developers, researchers and enterprises with a one-stop solution for AI model development, training and scaling. The platform integrates on-demand GPU resources, serverless reasoning, and automatic scaling for AI projects across...
2yrs ago
094.3K
Genesis:开源生成式物理引擎,实现基于真实物理的4D动态世界模拟

Genesis: open source generative physics engine for real physics-based 4D dynamic world simulation

General Introduction Genesis is a generative physics world designed for general purpose robotics and embodied AI learning. It provides a unified simulation platform that supports the simulation of a wide range of materials and physical phenomena.Genesis aims to unlock generative AI and physics simulation by combining...
2yrs ago
094.3K
OpenAOE:大模型群聊框架:同时与多个大语言模型聊天

OpenAOE: Large Model Group Chat Framework: Chatting with Multiple Large Language Models Simultaneously

Comprehensive Introduction OpenAOE is an open source large model group chat framework, aiming to solve the problem of the lack of chat frameworks in the current market with multiple models responding in parallel. With OpenAOE, users can talk to multiple Large Language Models (LLMs) at the same time and get parallel output. The framework supports ...
1yrs ago
094.3K
AICamp:适合团队使用的大模型集成聊天平台,接入自有API或免费使用GPT-4o-mini

AICamp: an integrated chat platform for teams with large models, access to its own API or free use of GPT-4o-mini

Comprehensive Introduction AICamp is a comprehensive AI platform designed to simplify the use of various AI tools and models. It provides a shared workspace for teams, facilitating team members to collaborate and improve productivity.AICamp offers a wide range of advanced AI features to help organizations bring...
2yrs ago
094.2K
Dream API:oneapi/newapi中转API,针对个人用户提供免费公益API

Dream API: oneapi/newapi transit API, providing free public service API for individual users.

Introduction After recommending many free large model API services in the Chief AI Sharing Circle, I suddenly found an important issue: there are official free small size scales; there are reverse API models; but there has been no free "official conversion" API. The reason why it has not been recommended is that the free "official conversion" API is not available. The reason why I haven't recommended it is that free "official conversion" is not available for large models.
2yrs ago
094.1K
MegaParse:解析各类型文档为LLM可用数据,完整保留文档中的表格、图片等所有信息

MegaParse: parses all types of documents into LLM-available data, preserving all information in the document such as tables, pictures, etc. in its entirety

Comprehensive Introduction MegaParse is a powerful and versatile document parsing tool designed to optimize data processing for the Large Language Model (LLM). Whether you are working with text, PDF, PowerPoint presentations or Word documents, MegaParse...
2yrs ago
094.1K
老师帮 - AI教师工作助手,支持课件一键转PPT

Teacher's Help - AI teacher's work assistant, support courseware to PPT in one click

Teacher Help is an AI intelligent tool platform designed for teachers to improve their work efficiency and teaching quality based on AI technology. The platform provides a variety of functions, including lesson plan generation, one-click conversion of courseware to PPT, homework and test question design, student comment generation, and teaching plan writing. The platform supports text translation...
1yrs ago
093.9K
DeepPiano - 智曲科技推出的AI钢琴应用

DeepPiano - AI Piano App by Smartquote Technology

DeepPiano is an intelligent piano application with a big model as its core launched by Zhiqu Technology. Through advanced artificial intelligence technology, it provides a variety of convenient functions for piano players and learners.DeepPiano can realize intelligent sheet music page turning, automatic recognition of playing progress, no need to manually operate...
1yrs ago
093.9K
VideoLingo:视频转录单词级时间轴字幕,视频字幕翻译和本地化配音开源工具

VideoLingo: video transcription word-level timeline subtitles, video subtitle translation and localized dubbing open source tools

General Description VideoLingo is a one-stop video translation and localization dubbing tool designed to generate Netflix-grade, high-quality subtitles, eliminating raw machine translation and multi-line subtitles, and adding high-quality voiceovers that enable global knowledge to be shared across language barriers. By...
2yrs ago
093.8K
Edraw.AI(亿图):在线协作白板工具,AI生成流程图和多种图表

Edraw.AI: Online collaborative whiteboard tool, AI-generated flowcharts and multiple diagrams

Comprehensive Introduction Edraw.AI is a revolutionary AI-powered online visualization whiteboard collaboration platform that integrates more than 40 intelligent tools and a library of carefully designed templates. The platform uses advanced AI technology to quickly transform users' textual thoughts into professional visual diagrams. The platform supports...
2yrs ago
093.7K
Akool:生成图像和视频营销素材|视频换脸|视频翻译|人像说话

Akool: Generate images and video marketing materials | Video Face Swap | Video Translation | Portrait Speak

General Introduction Akool is a focus on personalized visual marketing and advertising. Through advanced AI technology, AKOOL can help users easily create high-quality, personalized video content for a wide range of fields such as advertising, online education, art creation and e-commerce. It provides face transposition...
2yrs ago
093.7K
WeClone:用微信聊天记录和语音训练数字分身

WeClone: training digital doppelgangers with WeChat chats and voices

Comprehensive introduction WeClone is an open source project that uses WeChat chat logs and voice messages, combined with large language models and speech synthesis technology, to allow users to create personalized digital doppelgangers. The project can analyze the user's chat habits to train the model , but also a small number of voice samples to generate realistic sound...
1yrs ago
093.7K
Media.io:多功能在线媒体处理工具,在线视频、音频、图像编辑器

Media.io: Multi-functional online media processing tool, online video, audio, image editor

General Introduction Media.io is a powerful online AI video editing and media file processing platform. It helps users to enhance, convert, compress, etc. videos, audios and pictures. In addition to the basic editing functions, there are also features like video cartoonization, AI song cover generation, audio desc...
1yrs ago
093.6K
Midjourney:创造你想象中的图像|Midjourney中文官网介绍|官网开放免费测试

Midjourney: Create the images of your imagination|Midjourney Chinese website introduction|Official website open for free testing

Midjourney Introduction Midjourney is an independent research lab exploring new mediums of thought and expanding the imagination of the human species. It provides an AI service that generates images based on textual descriptions, allowing users to create a variety of art forms, from realistic to abstract wind...
2yrs ago
093.6K
Kolors:生成高质量图像的文本到图像模型,支持生成中文海报

Kolors: text-to-image model for generating high-quality images, support for generating Chinese posters

Comprehensive Introduction Kolors is a large-scale text-to-image generation model developed by the Racer team, based on potential diffusion techniques. The model is trained on billions of text-image data pairs, and is capable of generating high-quality, complex semantically accurate images with support for both Chinese and English input.Kolors in visual quality...
2yrs ago
093.5K
Leon:用语音交流的个人助理,创建自定义技能完成各类任务,保护隐私,离线运行

Leon: a personal assistant that communicates by voice, creates customized skills to accomplish all kinds of tasks, protects privacy, runs offline

General Introduction Leon is an open source personal assistant that runs on your server. It performs the tasks you ask for and can interact with the user via voice or text.Leon emphasizes privacy protection, supports offline operation, and allows users to create and share custom skills to extend its functionality...
1yrs ago
093.5K
GeekAI:自部署商业化多功能AI助手,完整接入多模型API运营后台

GeekAI: Self-deployed commercialized multi-functional AI assistant with complete access to multi-model API operation backend

Comprehensive introduction GeekAI is a full set of open source solutions for AI assistants based on AI big language model API implementation. The project comes with an operations management backend , out of the box , integrated with ChatGPT, Azure, ChatGLM, Xunfei Starfire, Wenxin Yiyin and many other p...
2yrs ago
093.5K
MegaTTS3:合成中英文语音的轻量模型

MegaTTS3: A Lightweight Model for Synthesizing Chinese and English Speech

Comprehensive Introduction MegaTTS3 is an open source speech synthesis tool developed by ByteDance in cooperation with Zhejiang University, focusing on generating high-quality Chinese and English speech. Its core model is only 0.45B parameters , lightweight and efficient , support for mixed Chinese and English speech generation and speech cloning . The project is hosted on ...
1yrs ago
093.4K
Devika:开源的AI软件工程师智能体,能够理解、拆分指令为子任务并编写代码

Devika: open-source AI software engineer intelligence that understands, splits instructions into subtasks and writes code

General Introduction Devika is an advanced AI software engineer that understands high-level human instructions, breaks them down into steps, studies the relevant information, and writes code to achieve a given goal. It intelligently develops software using large-scale language models, planning and reasoning algorithms, and web browsing capabilities.D...
1yrs ago
093.4K
AI2SRT:利用 Gemini模型,一键为长视频创建解说短视频或视频总结

AI2SRT: Create short narrated videos or video summaries for long videos with one click using Gemini models

Comprehensive Introduction AI2SRT is an open source project that utilizes the GeminiAI Big Model to generate short narrated videos and video summaries for long videos with one click, while supporting audio and video transcription subtitles. The project aims to simplify the video content creation process and provide efficient subtitle generation and translation functions. Users can pass...
2yrs ago
093.3K
WeaveFox:前端智能研发平台,能够根据设计图直接生成源代码

WeaveFox: a front-end intelligence development platform that generates source code directly from design drawings

Comprehensive Introduction WeaveFox is an AI front-end intelligent R&D platform launched by Ant Group, aiming to improve the efficiency and quality of front-end development through AI technology. The platform is based on Ant's self-developed Bailing multimodal large model, which is able to generate front-end source code directly based on design drawings, and supports multiple clients and technology stacks...
2yrs ago
093.2K
PantoMatrix(EMAGE):全身手势生成框架,从音频生成全身手势的3D动画框架

PantoMatrix (EMAGE): full-body gesture generation framework, 3D animation framework for generating full-body gestures from audio

Comprehensive Introduction PantoMatrix is an advanced full-body gesture generation framework capable of generating complete human movements from audio and partial gestures, including face, partial body, hand and full-body movements. The framework utilizes the latest multimodal datasets and deep learning techniques to provide high-quality 3D...
2yrs ago
093.2K