Latest AI Resources

Total 3143 articles posts
BISHENG(文擎毕昇):构建企业级AI应用的开源LLM DevOps平台

BISHENG: Open Source LLM DevOps Platform for Building Enterprise AI Applications

Comprehensive Introduction BISHENG is an open source LLM (Large Language Model) DevOps platform designed for next generation enterprise AI applications. The platform provides powerful and comprehensive features including generative AI workflows, RAG (Retrieval Augmented Generation), intelligent agents, unified model management...
2yrs ago
0121K
Granola:AI会议助手,转录会议讨论内容并用AI增强会议记录

Granola: AI meeting assistant that transcribes meeting discussions and enhances meeting notes with AI

General Introduction Granola is a tool that utilizes artificial intelligence technology to improve meeting efficiency and record accuracy. It transcribes meetings in real-time, automatically generates smart notes, and provides detailed meeting analysis.Granola aims to help users better manage meeting records and improve work...
2yrs ago
0120.7K
TRV:将幻灯片/PPT和讲解备注快速生成演讲视频

TRV: Rapidly Generate Presentation Videos from Slides/PPTs and Explanatory Notes

General Introduction TRV is an open source tool, hosted on GitHub, designed to help users quickly convert slides and presentation notes into videos with narration. It automatically generates audio and video content from incoming presentation files through simple command line operations, suitable for those who need to quickly create presentations...
1yrs ago
0120.4K
Linly-Talker:数字人智能对话系统,结合大语言模型与视觉模型,实现互动新体验

Linly-Talker: An Intelligent Dialogue System for Digital People, Combining Big Language Modeling and Visual Modeling for a New Interactive Experience

Comprehensive Introduction Linly-Talker is an innovative digital human dialog system that combines Large Language Models (LLMs) with visual models to create a novel approach to human-computer interaction. The system integrates a variety of technologies such as Whisper, Linly, Micros...
1yrs ago
0120.2K
VITA:开源视觉与语音实时交互的多模态大语言模型

VITA: Open Source Multimodal Large Language Model for Real-Time Interaction between Vision and Speech

General Introduction VITA is a leading open source interactive multimodal large language modeling project, pioneering the ability to achieve true full multimodal interaction. The project launched VITA-1.0 in August 2024, pioneering the first open source interactive fully-modal large language model.2024...
2yrs ago
0120.1K
WrenAI:对话式数据分析AI助手,直接获取答案、SQL查询与分析报表

WrenAI: Conversational Data Analytics AI Assistant with Direct Access to Answers, SQL Queries & Analytics Reports

General Introduction WrenAI is an open source SQL AI assistant specifically designed to help data teams, product teams and business teams gain data insights through natural language conversations. It is capable of converting natural language into SQL queries, generating charts, spreadsheets and reports, supporting multilingual...
2yrs ago
0119.7K
RealtimeVoiceChat:低延迟与AI进行自然口语对话

RealtimeVoiceChat: low-latency natural spoken conversation with AI

General Introduction RealtimeVoiceChat is an open source project focused on real-time, natural conversations with artificial intelligence via voice. Users use a microphone to input their voice, and the system captures the audio through a browser, quickly converts it to text, and a large-scale language model (LLM) generates back...
1yrs ago
0119.4K
小红书AI运营助手:自动生成和发布小红书文章

Xiaohongshu AI operation assistant: automatically generate and publish Xiaohongshu articles

Comprehensive Introduction Xiaohongshu AI Operation Assistant (xhsaipublisher) is an automation tool designed for publishing articles on the Xiaohongshu platform. The program combines a graphical user interface with automation scripts that utilize big model technology to generate content and automatically log in and publish via browser...
2yrs ago
0118K
R2R:多模态内容解析并结合知识图谱与混合搜索的先进AI检索(RAG)系统

R2R: An Advanced AI Retrieval (RAG) System for Multimodal Content Parsing and Combining Knowledge Graph with Hybrid Search

Comprehensive Introduction R2R (RAG to Riches) is an advanced AI retrieval system supporting Retrieval Augmented Generation (RAG) functionality with production-ready features. Built on a containerized RESTful API, the system provides multimodal content parsing, hybrid search functionality...
2yrs ago
0117.4K
Linly-Dubbing:智能视频多语言AI配音/翻译工具

Linly-Dubbing: Intelligent Video Multilingual AI Dubbing/Translation Tool

Comprehensive Introduction Linly-Dubbing is an intelligent multilingual AI dubbing and translation tool designed to provide users with high-quality multilingual video dubbing and subtitle translation services by integrating advanced AI technology. The tool is especially suitable for international education, global content localization and other scenarios, helping...
2yrs ago
0117.1K
百聆 (Bailing):低延时的开源语音对话助手,轻松实现自然对话交流

Bailing: a low-latency open source voice dialog assistant that easily realizes natural conversational exchanges

Comprehensive Introduction Bailing (Bailing) is an open source voice conversation assistant designed to engage in natural conversations with users through speech. The project combines speech recognition (ASR), voice activity detection (VAD), large language modeling (LLM) and speech synthesis (TTS) technologies to achieve...
2yrs ago
0116.4K
番茄创作工具:将授权小说和短剧文稿转视频,生成短视频用于推广引流

Tomato Creation Tool: Convert licensed novels and short play scripts to video, generating short videos for promotion and traffic generation

Comprehensive Introduction Tomato Darling Center's Copy to Video Creation Tool is a powerful AIGC (Artificial Intelligence Generated Content) tool designed to help content creators quickly convert written copy to video. The tool simplifies the production of copy to video through semantic analysis, illustration generation and video export...
2yrs ago
0116.4K
Qwen Chat:使用Qwen系列所有模型,图像生成、文档处理和网络搜索

Qwen Chat: image generation, document processing and web search using all models of the Qwen family

Comprehensive Introduction Qwen Chat (Tongyi Qianqian Overseas Edition) is a multi-functional AI assistant platform developed by Aliyun, aiming to provide users with comprehensive AI services. The platform integrates chatbot, image and video understanding, image generation, document processing, web search integration, tool li...
1yrs ago
0116.3K
Deep Live Cam:开源的实时AI换脸工具,一张照片就能实现实时换脸直播

Deep Live Cam: open source real-time AI face-swapping tool, a photo can realize real-time face-swapping live

General Introduction Deep Live Cam is an open source artificial intelligence tool designed to enable real-time face replacement and deep fake video generation from a single photo. The tool utilizes advanced deep learning algorithms to enable real-time face replacement in live streams or video calls, protecting user privacy and adding fun...
2yrs ago
0116.2K
蝉镜:数字人视频创作平台,拥有数百款数字人模板以及克隆专属数字人形象(付费)

Cicada Mirror: digital human video creation platform with hundreds of digital human templates and cloning of exclusive digital human images (paid)

General Introduction Cicada is a platform focusing on digital human video creation, utilizing AI technology to simplify the video production process. Users can choose different digital human images, input copy and generate videos with multi-language voiceovers. The platform provides a rich library of templates and materials, which are suitable for a variety of fields such as advertising and marketing, education and training...
2yrs ago
0116.2K
LibreChat:模仿ChatGPT界面交互的AI对话开源项目

LibreChat: mimic ChatGPT interface interaction AI dialog open source project

General Introduction LibreChat is a free, open source AI chat platform with extensive customization options and support for multiple AI providers, services and integrations. It brings together all AI conversations in one place with a familiar interface and innovative features, supporting multiple AI models, plugins and multiple languages. By...
2yrs ago
0116.1K
Smolagents: open source project for rapid development of AI intelligences and lightweight construction of intelligences

Smolagents: open source project for rapid development of AI intelligences and lightweight construction of intelligences

Comprehensive Introduction Smolagents is a lightweight intelligent agent library developed by HuggingFace that focuses on simplifying the development process of AI agent systems. The project is known for its clean design philosophy, with only about 1000 lines of core code, yet provides powerful feature integration capabilities. It is most ...
2yrs ago
0116K
LocalAI:开源的本地AI部署方案,支持多种模型架构,WebUI统一管理模型和API

LocalAI: open source local AI deployment solutions, support for multiple model architectures, WebUI unified management of models and APIs

General Introduction LocalAI is an open source local AI alternative designed to provide an API interface compatible with OpenAI, Claude, and others. It supports running on consumer-grade hardware, does not require a GPU, and is capable of text, audio, video, image generation, and speech cloning for multiple...
2yrs ago
0115.9K
RMBG-2-Studio:批量移除图像和视频背景的开源程序,基于RMBG 2.0优化

RMBG-2-Studio: open source program for batch removal of image and video backgrounds, optimized for RMBG 2.0

General Introduction RMBG-2-Studio is an enhanced background removal and replacement application developed based on the BRIA-RMBG-2.0 model. The application is designed to provide users with efficient and accurate image background processing capabilities for a variety of image types, including e-commerce, gaming and...
2yrs ago
0115.2K
Repo Prompt:依赖本地文件夹上下文进行写作、对话与优化代码

Repo Prompt: Relying on Local Folder Context for Writing, Conversing, and Optimizing Code

General Introduction Repo Prompt is a native application built for the macOS platform, dedicated to simplifying the process for developers working with native code using advanced AI language models. The tool helps developers manage and modify code files in an intelligent way, significantly improving development efficiency...
2yrs ago
0115.2K
NeoAI:让AI接管电脑远程操作,使用自然语言控制电脑的开源项目

NeoAI: Open source project that lets AI take over remote operation of computers and control them using natural language

General Introduction NeoAI is an innovative open source AI assistant tool that allows users to easily control and manage their computers through natural language conversations. Without writing any code, users can simply use everyday conversations to find files, automate tasks, manage devices, etc.NeoAI...
2yrs ago
0115K
轻竹PPT:AI一键生成PPT,在线PPT制作,Word、PDF文档转PPT

Light Bamboo PPT: AI one-key generation of PPT, online PPT production, Word, PDF document to PPT

Comprehensive Introduction Light Bamboo PPT (QZOffice) is an online service platform that utilizes artificial intelligence technology to help users quickly create professional-grade presentations. Users can automatically generate PPT templates by entering themes or key points, and edit and share them online to enjoy a convenient PPT making experience...
2yrs ago
0115K
Copilot:Microsoft Copilo智能AI助手,生产力工具| 微软Copilo国内访问

Copilot: Microsoft Copilo Intelligent AI Assistant, Productivity Tools | Microsoft Copilo Domestic Access

Copilot General Description Copilot is introduced by Microsoft as an artificial intelligence aid that can be integrated into Microsoft 365. It understands users' natural language to help them get information faster and be more productive.Copilot integrates with a wide range of applications including...
2yrs ago
0114.8K
智谱清言:GLM模型驱动的智能对话工具,支持创建智能体、长文档解读、AI数据分析

Smart Spectrum Clear Speech: a GLM model-driven intelligent dialog tool that supports creation of intelligences, long document interpretation, and AI data analysis

Comprehensive Introduction The website of 智谱清言(chatglm.cn), relying on GLM (Generative Language Model) technology, provides an intelligent communication platform. The platform supports multi-round conversations, content creation and message summarization, and aims to provide advanced...
1yrs ago
0114.3K
Goose:开源可扩展的编程智能体,自动化执行编程全流程任务

Goose: open source scalable programming intelligences that automate the full range of programming tasks

General Introduction Goose is an open source AI agent tool developed by Block, Inc. designed to help developers automate everyday development tasks. It supports a wide range of Large Language Models (LLMs) and interacts with users via the command line or desktop application interfaces.Goose can perform a wide range of tasks from agent...
2yrs ago
0114.3K
ViMax - 香港大学开源的多智能体视频生成框架

ViMax - Open Source Multi-intelligent Body Video Generation Framework at the University of Hong Kong

ViMax is an open source multi-intelligence body video generation framework from the Data Science Laboratory of the University of Hong Kong, which can automate the whole process from creative input to video output. Integration of script generation , scene design , shot planning and video rendering and other functions , to support users to generate coherent film and television grade video through natural language description ...
8mos ago
0114.3K
MatAnyone: 提取视频指定目标人像的开源工具,生成目标人像视频

MatAnyone: Extract video to specify the target portrait of the open-source tool to generate the target portrait video

General Introduction MatAnyone is an open source project focusing on video keying, developed and released on GitHub by a research team at S-Lab, Nanyang Technological University, Singapore. It provides users with stable and efficient video processing capabilities through coherent memory propagation techniques, especially...
1yrs ago
0114.2K
AnimateAI:使用AI生成角色一致的动画视频,儿童动画视频生成工具

AnimateAI: Generate character-consistent animated videos using AI, animated video generation tool for kids

Comprehensive Introduction AnimateAI is an all-encompassing AI video generation tool designed for the creation of animated video series. With advanced AI technology, users can quickly generate high quality video series, saving time and cost. Whether you are creating animated stories, movie trailers, inspirational short...
2yrs ago
0113.8K
Dify-WebUI:基于Dify API的桌面智能对话客户端,提供企业级AI对话能力

Dify-WebUI: Desktop Intelligent Conversation Client based on Dify API, providing enterprise-grade AI conversation capabilities

Comprehensive Introduction Dify-WebUI is a modern desktop smart conversation app based on the Dify API, designed to provide enterprises with powerful AI conversation capabilities. The application supports a variety of preset theme colors to meet the personalized needs of enterprises, and has a knowledge base management function to support...
2yrs ago
0113.7K
SP-MangaEditer:专业四格漫画插图创作工具,生成图像、编辑漫画页面

SP-MangaEditer: Professional four-panel manga illustration creation tool, generating images, editing manga pages

General Introduction SP-MangaEditer is an independent manga editing platform designed for manga creators. The platform supports image generation, layer editing, image adjustment, filter application and many other functions to help users easily create high-quality manga illustrations. Users can simply manipulate...
2yrs ago
0113.6K
Wasitai:检查图像是否由AI生成的简单工具,提供图像检测API

Wasitai: a simple tool to check if an image is generated by AI, providing an image detection API

General Introduction Wasitai is a powerful and handy tool that helps users easily detect whether an image is generated by AI or not. With the advancement of AI in the field of image generation, many tools and platforms are available to create realistic, high-quality images from text, sketches, or other images. However, not all...
2yrs ago
0113.3K
bilive:B站无人监守直播录制与自动切片、上传工具

bilive: Unsupervised live recording and automatic slicing and uploading tools for B station

Comprehensive Introduction bilive is a tool designed for B station live recording, providing extremely fast live recording, auto-slicing, pop-up rendering and subtitle generation. The tool is compatible with ultra-low configuration machines, supports 7x24 hours unattended recording, automatically recognizes and renders pop-ups and subtitles, automatically slices and...
2yrs ago
0113.1K