Latest AI Resources

Total 3143 articles posts
Maxun:开源无代码平台,自动抓取网页数据并转换为API或电子表格

Maxun: open source no-code platform that automatically crawls web data and converts it to APIs or spreadsheets

Comprehensive Introduction Maxun is an open source no-code web data extraction platform that allows users to train robots in minutes to automatically crawl web data and convert it into APIs or spreadsheets. The platform supports paging and scrolling, can adapt to changes in website layout, provides powerful data crawling...
2yrs ago
088.2K
百度秒哒:一句话AI生成完整前后端企业网站(开放内测)

Baidu second da: a sentence of AI to generate a complete front and back end of the enterprise website (open internal test)

Comprehensive Introduction Seconda is an innovative no-code development tool launched by Baidu Intelligent Cloud, designed to help users quickly build apps through natural language without the need to master complex programming skills. Relying on Baidu's powerful AI technology and big model support, the platform is able to transform users' creative ideas into practical usable...
1yrs ago
088.1K
YuE:将歌词转化为完整歌曲的基础模型,支持多种音乐风格

YuE: Transforms lyrics into a base model of a complete song, supporting a wide range of musical styles

General Introduction YuE is an open source full song generation base model that focuses on transforming lyrics into full songs. Unlike other models that can only generate short snippets of non-vocal music, YuE is capable of generating full songs with lead and backing vocals up to several minutes in length. The model addresses music generation in...
2yrs ago
088.1K
VoAPI:高颜值的AI模型转发接口管理系统,官网每日提供免费API额度

VoAPI: High-value AI model forwarding interface management system, the official website provides free API quota on a daily basis

Comprehensive Introduction VoAPI is a new high-color and high-performance AI model interface management and distribution system, which is mainly used for personal or enterprise internal management and distribution channels. Developed based on NewAPI, the system provides rich functional modules and optimized user interface, aiming to enhance...
2yrs ago
088.1K
AI Test Kitchen:Google创意生成与AI技术实验平台

AI Test Kitchen: Google's Experimental Platform for Idea Generation and AI Technology

Comprehensive Introduction AI Test Kitchen is an experimentation platform launched by Google Labs to explore the combination of artificial intelligence and creativity. The platform allows users to experience and give feedback on emerging AI technologies such as LaMDA.The platform provides a variety of tools to help users transform ideas into real...
2yrs ago
088.1K
VideoMind:视频按时间戳定位内容与问答的开源项目

VideoMind: video by timestamp positioning content and Q&A open source project

General Introduction VideoMind is an open source multimodal AI tool focused on inference, Q&A and summary generation for long videos. It was developed by Ye Liu of the Hong Kong Polytechnic University and a team from Show Lab at the National University of Singapore. The tool mimics human understanding of video...
1yrs ago
088.1K
海绵音乐:智能AI音乐创作平台,文字和图片生成音乐

Sponge Music: Intelligent AI music creation platform, text and image generated music

General Introduction SpongeBob Music is a music creation platform based on artificial intelligence technology. Users only need to enter a sentence of inspiration or upload a picture to generate an exclusive piece of music. The platform provides a variety of music styles and creation tools to help users easily create high-quality music. Whether you are a professional musician or...
2yrs ago
088.1K
Handy - 开源免费的本地AI语音转文字工具

Handy - Open Source Free Native AI Speech to Text Tool

Handy is open source and free local speech to text tool, supporting Windows, MacOS and Linux systems, developed by Rust and React. It is suitable for quick transcription and text input by processing voice data locally without uploading it to the cloud to ensure privacy and security.
9mos ago
088K
Story-Adapter:根据长篇故事生成连续且风格一致的图像插画

Story-Adapter: generating continuous and consistent graphic illustrations based on a long story

General Introduction Story-Adapter is an innovative story visualization framework that converts textual stories into coherent image sequences. Developed by researchers, this project employs an iterative approach that requires no training to generate high-quality story illustrations. The framework is characterized by its ability to handle long...
2yrs ago
088K
快标书:AI协助生成不同行业竞标方案(投标标书)

Fast bidding: AI assistance in generating bidding programs (tender bids) for different industries

Comprehensive Introduction Fast Bidding is a bidding production platform based on AI big model technology, aiming to provide efficient and intelligent bidding document solutions for enterprises and individuals. By integrating the bidding AI knowledge base, intelligent catalog generation and online editing functions, the platform helps users quickly generate high...
1yrs ago
088K
GenXD:生成任意3D和4D场景视频的开源框架

GenXD: open source framework for generating videos of arbitrary 3D and 4D scenes

General Introduction GenXD is an open source project, developed by the National University of Singapore (NUS) and Microsoft team. It focuses on generating arbitrary 3D and 4D scenes , to solve the real-world 3D and 4D generation due to insufficient data and model design complexity brought about by the problem . The project was developed by ...
1yrs ago
088K
Fun-ASR - 钉钉、通义联合推出的新一代语音识别模型

Fun-ASR - A New Generation of Speech Recognition Models Jointly Launched by Nail and Tongyi

Fun-ASR is a big model of speech recognition jointly launched by Nail and Tongyi Labs. The model has been trained with massive audio data and can accurately recognize multi-industry terminology, such as Internet, technology, home decoration, etc., significantly improving the recognition accuracy. The model combines with Nail enterprise information for inference optimization to reduce the illusion problem...
11mos ago
087.9K
Agenta:集成到AI应用的提示词与模型效果评估工具

Agenta: a tool for evaluating the effectiveness of cue words and models integrated into AI applications

Comprehensive Introduction Agenta is an open source AI model management tool specialized in helping users easily experiment with cue words, test model effects and monitor runs. It is suitable for people who want to develop AI applications quickly, providing a platform that is simple to operate. You can use it to try the effect of different cue words on...
1yrs ago
087.9K
MMAudio:为视频画面生成同步音效与配乐,视频到音频的多模态联合训练工具

MMAudio: generating synchronized sound effects and soundtracks for video footage, video-to-audio multimodal co-training tool

General Introduction MMAudio is an open-source project aiming to generate high-quality synchronized audio through joint multimodal training. Developed by Ho Kei Cheng et al. at the Chinese University of Hong Kong, the project's main function is to generate synchronized audio based on video and/or text input.MM...
2yrs ago
087.9K
MedRAX: 利用多模态大模型进行胸部X光片分析的智能体

MedRAX: A Smart Body for Chest X-ray Analysis Using Multimodal Large Models

Comprehensive Introduction MedRAX is a state-of-the-art AI intelligence designed for chest radiograph (CXR) analysis. It integrates state-of-the-art CXR analysis tools and multimodal large language models to dynamically process complex medical queries without additional training.MedRAX, through its modular design...
1yrs ago
087.7K
鬼手剪辑:视频去重|短剧解说|视频翻译|去除字幕

Ghost Hand Clips: video de-emphasis|skit commentary|video translation|subtitle removal

Comprehensive Introduction The official website of Ghost Hand Clips is designed to provide efficient video translation and subtitle removal tools for video creators, merchants and MCN organizations. Using powerful AI technology, Ghost Hand Clips is able to achieve intelligent translation of video content, subtitle removal and video personalization, helping users break through the language barrier and easily play...
2yrs ago
087.6K
笔格设计:在线图片编辑器,免费使用图像生成工具,轻松制作精美图片

Pen Grid Design: online photo editor, free to use the image generation tool, easy to create beautiful pictures

General Introduction Pen Grid Design is a website that provides online image editing and design services. Users can easily create and edit all kinds of images, including posters, PPT, GIF, etc. through this platform. Pen Grid Design provides a wealth of design materials and templates, and supports AI smart tools, such as AI image generation, A...
2yrs ago
087.6K
SmartRead:自动标注技术PDF文档并提供相关引用源

SmartRead: Automatically annotate technical PDF documents and provide relevant citation sources

Comprehensive Introduction SmartRead is an AI-based open source tool designed for technical documents. It can automatically analyze PDF files, mark key content, such as important terms, titles or core ideas to help users quickly understand complex documents. At the same time, it can also provide with the main document...
1yrs ago
087.6K
Director:智能视频代理框架,用自然语言描述执行视频搜索、编辑和生成工作流

Director: Intelligent Video Agent Framework for Performing Video Search, Editing, and Generation Workflows with Natural Language Descriptions

General Introduction Director is an open source framework designed to simplify and optimize video interactions and workflows by building intelligent video agents. The framework is based on VideoDB's "video-as-data" infrastructure and is capable of handling complex video tasks such as searching, editing, compiling and generating...
2yrs ago
087.6K
Agentic Security:开源的LLM漏洞扫描工具,提供全面的模糊测试和攻击技术

Agentic Security: open source LLM vulnerability scanning tool that provides comprehensive fuzz testing and attack techniques

General Introduction Agentic Security is an open source LLM (Large Language Model) vulnerability scanning tool designed to provide developers and security professionals with comprehensive fuzz testing and attack techniques. The tool supports customized rule sets or agent-based attacks and is able to integrate LLM AP...
1yrs ago
087.6K
Moondream:批量反推图像提示词的开源轻量级视觉语言模型

Moondream: an open source lightweight visual language model for batch backpropagation of image cue words

Comprehensive Introduction Moondream is an open source lightweight visual language model designed to enable image description capabilities through deep learning and computer vision techniques. The model is able to run efficiently on a variety of platforms and is particularly suitable for edge devices.Moondream uses advanced techniques and...
2yrs ago
087.4K
Voiceflow:写作构建AI智能体,部署客户服务对话工具|编排客户服务流程

Voiceflow: Writing to Build AI Intelligence, Deploying Customer Service Conversation Tools | Orchestrating Customer Service Processes

General Introduction Voiceflow is a collaboration platform designed for teams to design, develop and publish chat and voice assistants. Since its inception in 2018, Voiceflow is dedicated to providing developers and designers with a powerful collaborative platform for conversational AI agents. Through this platform...
2yrs ago
087.4K
Orchestra: Building Smart AI Teams for Easier and More Efficient Multi-Intelligence Collaborative Development

Orchestra: Building Smart AI Teams for Easier and More Efficient Multi-Intelligence Collaborative Development

Comprehensive Introduction Orchestra is an innovative lightweight Python framework that focuses on building multi-intelligence collaborative systems based on the Large Language Model (LLM). It employs a unique method of arranging intelligences so that multiple AI intelligences can work together harmoniously like a symphony orchestra. By modeling ...
2yrs ago
087.3K
Fey: 金融市场研究工具,提升投资决策的智能助手

Fey: Financial market research tools, intelligent assistants to enhance investment decisions

General Introduction Fey is an intelligent assistant designed for modern investors, providing real-time market data and personalized investment advice. With a simple and intuitive interface, users can easily access important financial information and market trends.Fey's core features include stock tracking, financial analysis, personalized new...
2yrs ago
087.2K
Tough Tongue AI:与AI对话练习面试与职场沟通技巧

Tough Tongue AI: Practice Interview and Workplace Communication Skills by Talking to an AI

General Introduction Tough Tongue AI is an artificial intelligence platform designed for practicing tough conversations. Users can simulate a variety of complex conversational situations, such as job interviews, salary negotiations, sales presentations, etc. by selecting preset scenarios or creating custom scenarios. The platform provides video and...
2yrs ago
087.1K
Artflow:创作人物一致性的动画故事和虚拟数字人口播视频

Artflow: Creating character-consistent animated stories and virtual digital pop-up videos

General Description Artflow is an online platform that enables users to upload photos, train exclusive AI characters, and create character-consistent videos and animated stories. Offering free training for the first time, users can customize their identity to create unique images and videos for a variety of scenarios. Monthly ...
2yrs ago
087.1K