Latest AI Resources

Total 3185 articles posts
元真数字人:数字人直播、口播短视频,商业化AI虚拟人直播工具

Yuanzhen digital people: digital people live, oral short video, commercialization AI avatar live tool

Comprehensive Introduction Yuanzhen Digital People is a leading AIGC (Artificial Intelligence Generated Content) platform dedicated to providing users with one-stop services such as digital people live broadcasting, short video production and AI assistant. The platform integrates AI algorithm synthesis and GPT-style big models, supports users to create exclusive Q&A models, provides real...
2yrs ago
0109.3K
火山方舟:大模型训练与云计算服务,注册送150元等额算力

Volcano Ark: Big Model Training and Cloud Computing Service, Sign Up for $150 Equivalent Arithmetic

Comprehensive Introduction Volcano Ark is a cloud computing platform launched by Volcano Engine that focuses on big model services, aiming to provide enterprises with a complete solution from model selection, training to application. Relying on ByteDance's deep accumulation in the field of AI, Volcano Ark integrates the big model resources of several top AI companies...
2yrs ago
0109.2K
通义听悟:阿里通义音视频内容转录AI助手

Tongyi Listening and Understanding: Ali Tongyi Audio and Video Content Transcription AI Assistant

Comprehensive Introduction Tongyi Listening and Understanding is a work-study AI assistant launched by Aliyun, focusing on transcribing and analyzing audio and video content. It relies on AliCloud's powerful AI models to transcribe audio and video content into text in real time, and provides translation, summarization, positioning and other functions. Tongyi Listening Woo supports multiple languages and scenarios...
2yrs ago
0109.1K
opensource_notebooklm:基于Deepseek-V3和PlayHT TTS的NotebookLM开源实现

opensource_notebooklm: open source implementation of NotebookLM based on Deepseek-V3 and PlayHT TTS

General Introduction Open Source NotebookLM is an innovative artificial intelligence project that combines Deepseek-V3's language understanding capabilities with PlayHT's speech synthesis technology, aiming to create an intelligent note-taking conversation system. The project was developed by Build Fast w...
2yrs ago
0109.1K
OpenAOE:大模型群聊框架:同时与多个大语言模型聊天

OpenAOE: Large Model Group Chat Framework: Chatting with Multiple Large Language Models Simultaneously

Comprehensive Introduction OpenAOE is an open source large model group chat framework, aiming to solve the problem of the lack of chat frameworks in the current market with multiple models responding in parallel. With OpenAOE, users can talk to multiple Large Language Models (LLMs) at the same time and get parallel output. The framework supports ...
2yrs ago
0109K
RunPod:专为AI设计的GPU云服务,快速冷启动SD且按秒付费

RunPod: GPU Cloud Service Designed for AI with Fast Cold Start SD and Pay Per Second

Comprehensive Introduction RunPod is a cloud computing platform designed for AI, aiming to provide developers, researchers and enterprises with a one-stop solution for AI model development, training and scaling. The platform integrates on-demand GPU resources, serverless reasoning, and automatic scaling for AI projects across...
2yrs ago
0109K
AnyText:生成和编辑多语言图像文本,高可控在图像中生成多行中文

AnyText: Generate and edit multi-language image text, highly controllable to generate multiple lines of Chinese in the image

Comprehensive Introduction AnyText is a revolutionary multilingual visual text generation and editing tool developed based on the diffusion model. It generates natural, high-quality multilingual text in images and supports flexible text editing features. It was developed by a team of researchers and presented at ICLR 2024...
2yrs ago
0108.9K
GLM-PC(智谱牛牛)正式发布内测下载,真正可以控制电脑的AI

GLM-PC (Smart Spectrum Bull) officially released for internal download, the real AI that can control the computer

GLM-PC (Bull) Introduction GLM-PC is a desktop application based on the CogAgent model, which is able to perform complex tasks quickly through natural language commands. It has the ability of task planning and interface understanding, and can autonomously complete various computer operations according to user instructions. Notes for use...
2yrs ago
0108.8K
AI2SRT:利用 Gemini模型,一键为长视频创建解说短视频或视频总结

AI2SRT: Create short narrated videos or video summaries for long videos with one click using Gemini models

Comprehensive Introduction AI2SRT is an open source project that utilizes the GeminiAI Big Model to generate short narrated videos and video summaries for long videos with one click, while supporting audio and video transcription subtitles. The project aims to simplify the video content creation process and provide efficient subtitle generation and translation functions. Users can pass...
2yrs ago
0108.8K
Flow(Laminar):构建智能体的轻量级任务引擎,简化并灵活管理任务

Flow (Laminar): a lightweight task engine for building intelligences that simplifies and flexibly manages tasks

Comprehensive Introduction Flow is a lightweight task engine designed for building AI agents, emphasizing simplicity and flexibility. Unlike traditional node- and edge-based workflows, Flow uses a dynamic task queuing system that supports parallel execution, dynamic scheduling, and intelligent dependency management. Its core concept is ...
2yrs ago
0108.7K
shadcn/ui:组件库构建平台

shadcn/ui: component library building platform

General Introduction shadcn/ui is an open source component library building platform that provides beautiful and customizable UI components that users can copy and paste into their applications. The platform supports a variety of front-end frameworks and provides detailed installation and usage guidelines to help developers quickly get started...
2yrs ago
0108.7K
Orchestra: Building Smart AI Teams for Easier and More Efficient Multi-Intelligence Collaborative Development

Orchestra: Building Smart AI Teams for Easier and More Efficient Multi-Intelligence Collaborative Development

Comprehensive Introduction Orchestra is an innovative lightweight Python framework that focuses on building multi-intelligence collaborative systems based on the Large Language Model (LLM). It employs a unique method of arranging intelligences so that multiple AI intelligences can work together harmoniously like a symphony orchestra. By modeling ...
2yrs ago
0108.5K
DeepPiano - 智曲科技推出的AI钢琴应用

DeepPiano - AI Piano App by Smartquote Technology

DeepPiano is an intelligent piano application with a big model as its core launched by Zhiqu Technology. Through advanced artificial intelligence technology, it provides a variety of convenient functions for piano players and learners.DeepPiano can realize intelligent sheet music page turning, automatic recognition of playing progress, no need to manually operate...
1yrs ago
0108.5K
反谱 - AI音乐转谱平台,支持音频文件转五线谱和简谱

AntiScore - AI music transcription platform, supports audio files to pentatonic and simple music.

AntiSpectrum is an innovative online AI music conversion platform, based on advanced AI technology, to convert audio files (such as MP3, FLAC, etc.) into pentatonic and simple scores. AntiSpectrum has a vocal separation function, which separates the vocals from the accompaniment in the music, making it easy for music production and mixing. AntiSpectrum supports converting MIDI files...
1yrs ago
0108.4K
Index-AniSora - B站推出的开源动漫视频生成模型

Index-AniSora - Open Source Anime Video Generation Model by B Station

Index-AniSora is an advanced anime video generation model open source by Beili Beili. The model can generate a coherent animation video based on a single picture , support a variety of styles , such as drama , national animation , VTuber content and so on. The model is based on the diffusion model architecture , combined with the space-time mask module , 3D...
1yrs ago
0108.3K
DomoAI:智能视频艺术风格转换|图像转视频|文本转视频

DomoAI: Intelligent Video Art Style Conversion|Image to Video|Text to Video

General Description DomoAI has recently launched its Video to Video feature, which converts existing videos into a completely different art style with amazing results. It allows users to easily create unique styles of visual art. Other features included in the platform can convert still images to motion video, text to picture...
2yrs ago
0108.3K
AgentIQ:灵活连接和管理AI智能体的开源工具

AgentIQ: An open source tool for flexible connection and management of AI intelligences

General Introduction AgentIQ is an open source tool from NVIDIA designed to help developers efficiently connect and manage AI intelligences. It enables intelligences from different frameworks to seamlessly collaborate, connect enterprise data and tools, and build workflows like calling functions. The tool's biggest...
1yrs ago
0108.3K
Sketch-Gen:生成高质量线稿和草图,反推图像提示词,一键安装包

Sketch-Gen: Generate high-quality line drawings and sketches, backpropagate image cue words, one-click package installation

General Introduction Sketch-Gen is an AI technology-based line drawing and sketch generation tool designed to help artists and designers quickly generate high-quality line drawings and sketches. The tool is derived from the Paints-UNDO project and utilizes advanced machine learning models that can...
2yrs ago
0108.3K
Rows:数据驱动的电子表格工具

Rows: a data-driven spreadsheet tool

Comprehensive Introduction Rows is an innovative spreadsheet tool designed to simplify importing, converting and sharing data. It not only provides the functionality of a traditional spreadsheet, but also integrates real-time data reporting and automated data analysis. With a simple interface and without writing code, users can create strong...
2yrs ago
0108.2K
AnimeGamer:用语言指令生成动漫视频和角色互动的开源工具

AnimeGamer: An Open Source Tool for Generating Anime Videos and Character Interactions with Language Commands

AnimeGamer is an open source tool launched by Tencent ARC Lab. Users can generate anime videos with simple language commands, such as "Sousuke drive around in a purple car", as well as allow different anime characters to interact with each other, such as Kiki from The Witch's House, and Sky City...
1yrs ago
0108.2K
天工AI:全能AI助手,助力高效工作与生活

Tiangong AI: All-around AI assistant for efficient work and life

Comprehensive Introduction Tiangong AI is the first all-round AI assistant in China, which integrates various functions such as search, dialog, writing, document analysis, drawing, PPT production and so on. It is able to understand the user's intention, search for information from all over the internet, and summarize, generalize and integrate through advanced AI technology to output high-quality, no...
2yrs ago
0108.2K
COSINE:智能理解代码库,让开发者轻松理解和编写代码的AI工具(内测)

COSINE: Intelligent Understanding Codebase, an AI tool that makes it easy for developers to understand and write code (in beta)

General Introduction Cosine is a revolutionary AI-driven code understanding platform that provides deep codebase understanding and analysis services for modern software developers. Supporting over 50 programming languages, the platform utilizes a unique technical architecture that combines a specialized search engine, vector database, and ...
2yrs ago
0108.1K
Midjourney:创造你想象中的图像|Midjourney中文官网介绍|官网开放免费测试

Midjourney: Create the images of your imagination|Midjourney Chinese website introduction|Official website open for free testing

Midjourney Introduction Midjourney is an independent research lab exploring new mediums of thought and expanding the imagination of the human species. It provides an AI service that generates images based on textual descriptions, allowing users to create a variety of art forms, from realistic to abstract wind...
2yrs ago
0108.1K
老师帮 - AI教师工作助手,支持课件一键转PPT

Teacher's Help - AI teacher's work assistant, support courseware to PPT in one click

Teacher Help is an AI intelligent tool platform designed for teachers to improve their work efficiency and teaching quality based on AI technology. The platform provides a variety of functions, including lesson plan generation, one-click conversion of courseware to PPT, homework and test question design, student comment generation, and teaching plan writing. The platform supports text translation...
1yrs ago
0108K
Leon:用语音交流的个人助理,创建自定义技能完成各类任务,保护隐私,离线运行

Leon: a personal assistant that communicates by voice, creates customized skills to accomplish all kinds of tasks, protects privacy, runs offline

General Introduction Leon is an open source personal assistant that runs on your server. It performs the tasks you ask for and can interact with the user via voice or text.Leon emphasizes privacy protection, supports offline operation, and allows users to create and share custom skills to extend its functionality...
2yrs ago
0108K
Perplexica:1比1复刻 Perplexity AI 功能和界面的开源AI搜索引擎

Perplexica: an open source AI search engine that replicates Perplexity AI's features and interface 1 to 1

Comprehensive Introduction Perplexica is an open source AI-driven search engine designed to provide answers that delve deep into the Internet. It uses advanced machine learning algorithms, such as similarity search and embedding techniques, to optimize search results and provide clear answers with cited sources.Perple...
2yrs ago
0107.9K
通义千问:阿里推出的多模态大模型,拥有文本回答、图片理解、视频解析能力

Tongyi Thousand Questions: a large multimodal model launched by Ali with text answering, image understanding, and video parsing capabilities

Comprehensive Introduction Tongyi Thousand Questions is an intelligent big model developed by Aliyun, aiming to provide a human-like interaction experience through deep learning and natural language processing technology. It can quickly generate creative copy to add fun to life, and serve as a learning assistant to help users easily learn all kinds of knowledge. With cutting-edge technology and evolving...
2yrs ago
0107.9K
法行宝:AI法律顾问,人工智能法律咨询,百度AI法律平台

Fa Xing Bao: AI Legal Advisor, Artificial Intelligence Legal Consultation, Baidu AI Legal Platform

Comprehensive Introduction LawXinbao is an intelligent legal service platform launched by Baidu, which integrates advanced artificial intelligence technology with a professional legal knowledge base. The platform is dedicated to providing users with convenient and professional legal intelligent services, including intelligent legal Q&A, case analysis, contract review and other functions. Through deep learning...
2yrs ago
0107.9K
BuffGPT:企业级生成式AI应用低代码开发平台

BuffGPT: A Low-Code Development Platform for Enterprise-Grade Generative AI Applications

Comprehensive Introduction BuffGPT is an open source AI application development platform based on the Large Language Model (LLM), providing out-of-the-box features such as data processing, model invocation, RAG retrieval, and visual workflow orchestration to help users easily build and operate generative AI applications. The platform supports privatization...
2yrs ago
0107.7K
VideoLingo:视频转录单词级时间轴字幕,视频字幕翻译和本地化配音开源工具

VideoLingo: video transcription word-level timeline subtitles, video subtitle translation and localized dubbing open source tools

General Description VideoLingo is a one-stop video translation and localization dubbing tool designed to generate Netflix-grade, high-quality subtitles, eliminating raw machine translation and multi-line subtitles, and adding high-quality voiceovers that enable global knowledge to be shared across language barriers. By...
2yrs ago
0107.7K
Motionvid.ai:用文字或草图快速生成演示动画视频

Motionvid.ai: Quickly generate animated demo videos with text or sketches

General Introduction Motionvid.ai is an online tool that utilizes artificial intelligence to help users quickly create professional animated videos. Its best feature is to generate animations with smooth dynamics and high-quality visual effects in seconds through text descriptions or hand-drawn sketches. Users don't need to master complex...
1yrs ago
0107.6K
Agentic Security:开源的LLM漏洞扫描工具,提供全面的模糊测试和攻击技术

Agentic Security: open source LLM vulnerability scanning tool that provides comprehensive fuzz testing and attack techniques

General Introduction Agentic Security is an open source LLM (Large Language Model) vulnerability scanning tool designed to provide developers and security professionals with comprehensive fuzz testing and attack techniques. The tool supports customized rule sets or agent-based attacks and is able to integrate LLM AP...
2yrs ago
0107.5K
SegAnyMo:从视频中自动分割任意运动物体的开源工具

SegAnyMo: open source tool to automatically segment arbitrary moving objects from video

General Introduction SegAnyMo is an open source project developed by a team of researchers at UC Berkeley and Peking University, including members such as Nan Huang. This tool focuses on video processing and can automatically recognize and segment arbitrary moving objects in a video, such as people, animals or...
1yrs ago
0107.5K
MegaParse:解析各类型文档为LLM可用数据,完整保留文档中的表格、图片等所有信息

MegaParse: parses all types of documents into LLM-available data, preserving all information in the document such as tables, pictures, etc. in its entirety

Comprehensive Introduction MegaParse is a powerful and versatile document parsing tool designed to optimize data processing for the Large Language Model (LLM). Whether you are working with text, PDF, PowerPoint presentations or Word documents, MegaParse...
2yrs ago
0107.5K
VideoRAG:理解超长视频的RAG框架,支持多模态检索和知识图谱构建

VideoRAG: A RAG framework for understanding ultra-long videos with support for multimodal retrieval and knowledge graph construction

Comprehensive Introduction VideoRAG is a retrieval-enhanced generative framework designed for processing and understanding very long contextual videos. The tool combines a graph-driven textual knowledge base with hierarchical multimodal context encoding to efficiently process on a single NVIDIA RTX 3090 GPU...
2yrs ago
0107.5K
Edraw.AI(亿图):在线协作白板工具,AI生成流程图和多种图表

Edraw.AI: Online collaborative whiteboard tool, AI-generated flowcharts and multiple diagrams

Comprehensive Introduction Edraw.AI is a revolutionary AI-powered online visualization whiteboard collaboration platform that integrates more than 40 intelligent tools and a library of carefully designed templates. The platform uses advanced AI technology to quickly transform users' textual thoughts into professional visual diagrams. The platform supports...
2yrs ago
0107.4K
CogVLM2:开源多模态模型,支持视频理解与多轮对话

CogVLM2: Open Source Multimodal Modeling with Support for Video Comprehension and Multi-Round Dialogue

Comprehensive Introduction CogVLM2 is an open source multimodal model developed by the Tsinghua University Data Mining Research Group (THUDM), based on the Llama3-8B architecture, and designed to provide performance comparable to or even better than GPT-4V. The model supports image understanding, multi-round dialogs, and visual ...
2yrs ago
0107.4K