Open-source AI tools target users who value controllability, secondary development, and data boundaries, and can be used for on-premises deployment, model research, or internal enterprise integration. When selecting a code, you need to check the license, maintenance activity, hardware requirements, installation difficulty, and community ecosystem together, rather than just whether the code repository is public.
AstronClaw
Lobster assistant
AstronClaw is an AI office assistant launched by iFLYTEK, focusing on the rapid deployment of exclusive AI assistants and 24×7 hours online services. Based on OpenClaw's core capabilities, AstronClaw supports the creation of deeply customizable personal AI assistants and can integrate efficient skills to enable multi-channel information interaction. As an AI office product for efficiency improvement and task collaboration scenarios, AstronClaw is suitable for information processing, content generation, office communication, and daily workflow collaboration, helping individuals and teams achieve a continuous online and scalable smart office experience with a lower threshold.
PicoClaw
Lobster assistant
PicoClaw is an open-source AI chat assistant that focuses on lightweight, cross-platform, and efficient operation. Built on Go, PicoClaw features fast startup, low resource usage, and low deployment threshold, making it suitable for developers and teams to quickly build intelligent assistants on local devices, edge hardware, or multiple system environments. As an AI chat product for practical use cases, PicoClaw supports connecting to multiple communication channels such as Telegram, Discord, Slack, Feishu, DingTalk, Enterprise WeChat, LINE, and QQ, and can be used for message interaction, intelligent Q&A, collaborative communication, and automated assistant scenarios. For users looking for low-cost deployment and a stable experience, PicoClaw is an AI chat tool that balances performance, flexibility, and scalability.
Ugly - AI Video Open Source Community
AI video generation
Xianchou AI is an open source community focusing on AI video creation, creating content and communication platforms around "creating, sharing, learning, and reproducing popular AI videos". Users can browse popular AI video cases in Presenting Ugly AI, get creative ideas and prompt inspiration that can be referenced, and quickly build their own AI video creation process; You can also publish your work and experiences, and discuss model selection, camera pacing, and short video storytelling with other creators. For individual creators and teams who want to improve the efficiency of AI video, Ugly AI makes AI video tutorials, prompts, and replication methods easier to search and reuse through community precipitation, helping to continuously produce more stable AI video content.
AudioCraft
AI music creation
AudioCraft is a library of generative audio and AI music tools and online demos launched by Meta AI, integrating capabilities such as MusicGen text-to-music, AudioGen text-to-sound effects, and EnCodec neural audio compression. You can quickly generate different styles of soundtracks, ambient sounds, and realistic sound effects with a single prompt, and control the rhythm, atmosphere, and duration through prompts, which is suitable for short video soundtracks, game sound prototypes, advertising ambient sounds, and creative inspiration verification. AudioCraft also provides open-source code and models for developers to deploy on-premises, integrate into workflows, and develop reactively.
XiaoZhi AI
MCP Toolset
XiaoZhi AI is an open-source AI hardware project and console service for developers and hardware vendors, focusing on low-latency voice conversations and scalable agent capabilities. Based on the Model Context Protocol (MCP), Xiaozhi AI connects speech recognition, intent understanding, memory and tool calling in series, and adapts to a variety of common chips and development boards, making it convenient and quick to make products such as emotional companionship, smart home control, desktop voice assistants, and in-vehicle robots. Through the Xiaozhi AI console, users can create agents, configure roles and behaviors, and complete device binding and management, making Xiaozhi AI more worry-free from prototyping to mass production access.
LTX-2
AI video generation
LTX-2 is an AI video generation engine for creative production, focusing on 4K 48fps synchronization with audio and video, allowing AI videos to be integrated from draft to finished film. The LTX-2 supports text-to-video and image-to-video, with multi-keyframe and depth control for stable motion and coherent footage for VFX, advertising, storyboarding, and image restoration. Open source and API design for studios, developers, and enterprises to run on consumer-grade GPUs. Through LoRA fine-tuning and process tools, LTX-2 achieves high fidelity, high efficiency, and controllability in AI video generation, AI video editing, and large-scale production scenarios, helping teams quickly build branded visual styles and stable delivery capabilities.
Deta Surf
AI browser
Deta Surf is an AI browser and smart notebook that blends browsing and note-taking, allowing web pages, documents, and ideas to flow naturally in one place. Deta Surf supports web search, PDF and YouTube embedded reading and questioning, automatically generates key point summaries and in-depth analysis, and keeps traceable citation links in your notes. Collect sites, images, and files with Notebooks, organize information with Vertical Tabs, and create interactive apps and visualizations with Surflets. Deta Surf is open-source and local-first, with data stored on the device and importable for import and export; Support for large models or local LLMs of your choice, using your own keys to meet the needs of privacy, security, and efficient research.
Helium Browser
AI browser
Helium Browser is a privacy-first, open-source browser that is safe to use with unbiased ad blocking and tracking protection enabled by default, blocking third-party cookies, and minimizing fingerprinting. Based on Chromium, it is compatible with all extensions and anonymously access the extension store; The interface is light, fast to start, and supports Split View split screen and native !bangs to quickly and directly access websites, and can also be used offline. With no built-in cloud sync and password management, Helium Browser offers self-hosted options and fast security updates for users and teams that value efficiency, control, and data sovereignty.
Browser Operator
AI browser
Browser Operator is an open-source, privacy-friendly AI browser for research, analysis, and automation scenarios. Browser Operator is based on the Chromium architecture, supports local AI inference and cloud large model collaboration, and provides AI Agent Studio to build AI agent processes, automatically capture web page information, extract key points, and generate summaries, and can complete forms, price comparisons, and lead sorting across multiple sites. Browser Operator has built-in research assistants, smart shopping and talent search templates, lowering the threshold for construction; The whole process is transparent and auditable, and the data is locally controllable. Currently in alpha with priority support for macOS, it's suitable for individuals and teams that value privacy, security, and efficiency, elevating their daily browsing into orchestrated AI browser workflows.
BrowserOS
AI browser
BrowserOS is an open-source, privacy-first AI browser that focuses on "turning words into actions". BrowserOS uses local AI agents to understand natural language commands and automatically complete web page operations such as clicking, typing, and jumping, making it suitable for high-frequency research and form process automation. It supports Split View to put models such as ChatGPT, Claude, and Gemini in sidebar conversations, making it more efficient to ask questions while reading; Compatible with the Chromium ecosystem, commonly used Chrome extensions and data can be seamlessly migrated. BrowserOS provides MCP server integration, which can connect to Gmail, Calendar, Docs, Sheets, and Notion with one click, upgrading the browser to a work hub. It also supports local models such as Ollama and LM Studio or comes with its own API key to ensure that the data is locally controllable. BrowserOS covers macOS, Windows, and Linux across platforms, and has built-in semantic retrieval and highlighting to comprehensively improve information retrieval and web office efficiency.
Nook Browser
AI browser
Nook Browser is an open-source, privacy-focused, and modern browser. Nook Browser takes local priority as the core, and the data is saved on your device, and you can export and delete it with one click, reducing the risk of synchronization and leakage. It organizes work and interests through Rooms/Spaces, avoids overwhelming tabs, and offers Peek Quick Preview and Split View split screen for efficient multitasking. Nook Browser pursues a simple, less intrusive online experience, with an open roadmap and community co-creation mechanism, making it easy to scale and customize. For users and teams that value privacy, efficiency, and sustainable evolution, Nook Browser is the ideal browser choice for lightweight and scalability.
Zen Browser
AI browser
Zen Browser is an open-source browser that emphasizes focus and efficiency, advocating for a "quieter internet." Zen Browser prioritizes privacy and reduces distractions and tracking, making it suitable for daily online and web work that requires stability and security. It offers features like Workspaces, Compact Mode, Glance Quick Switching, and Split View for clearer multitasking management. Zen Browser is continuously iterated by the community, with a beautiful interface, fast performance, and strong scalability, which can meet the high-frequency usage scenarios of developers and content creators. If you value privacy, efficiency, and comfort, Zen Browser is the browser choice that balances aesthetics and functionality.
Ladybird
AI browser
Ladybird is a new non-profit-led, independent web browser and web engine that emphasizes web standards first and built from the ground up. Ladybird aims to be cross-platform, focusing on Linux and macOS at this stage, pursuing performance, stability, and security, without introducing other browser kernel code, nor commercializing it through default search shares or tokens. The project is open source and welcoming developers, collaborating with the community through GitHub to advance progress, with an alpha version planned for early users in 2026, suitable for users and development teams who value openness, transparency, and control.
Scira AI
AI search engine
Scira AI is an open-source AI search engine and Perplexity alternative that focuses on "real-time, verifiable, and self-hostable." It combines RAG (Retrieval-Augmented Generation) with Search Grounding to return answers with source citations, making it suitable for research, learning, and knowledge retrieval scenarios. Scira AI supports multi-model switching, covering Grok, Claude, DeepSeek, Qwen, OpenAI, etc., providing PDF analysis, history and voice reading capabilities, and realizes topic monitoring and timing tracking through Scira Lookout. Q: What is Scira AI? A: Scira AI is an open-source AI search tool that provides real-time search with trusted citations. As an alternative to Perplexity, Scira AI combines speed, transparency, and control, and can be used directly on the official website or self-built to meet the AI search needs of individuals and teams.
Cloudflare VibeSDK
AI programming tools
Cloudflare VibeSDK is an open-source reference implementation for AI programming that helps teams quickly build and deploy AI applications on the Cloudflare Developer Platform. It takes the core path of "from natural language to runnable applications", combining Workers, Pages, KV, Durable Objects, Queues, R2, Vectorize, and AI Gateway to cover full-link scenarios of generation, storage, inference, and observation. Brief definition: Cloudflare VibeSDK is an all-in-one sample platform for AI application development and delivery on the Cloudflare edge network. Common use cases include intelligent customer service, retrieval-augmented generation, and full-stack chat applications for teams looking to validate and scale their AI products with low operational costs.
Genmo Mochi 1
AI video generation
Genmo Mochi 1 is an open-source AI video generation tool for creators and teams, supporting text-to-video, image-to-video, and per-shot control, highlighting physical consistency, character coherence, and high-prompt alignment. Genmo Mochi 1 offers both cloud playground and private deployment, making it easy to quickly produce in scenarios such as advertising storyboards, product demonstrations, educational animations, and social media videos. As an "open-source AI video generation model", Genmo Mochi 1 is easy to use and scalable, making it suitable for integrating AI video generation into existing creative processes and production systems.
Kilo Code
AI programming tools
Kilo Code is an open-source AI programming agent extension for VS Code, focusing on the integration of "architecture-coding-debugging". Built-in Orchestrator, Architect, Code, and Debug modes can automatically disassemble tasks, generate implementations, locate and fix errors. It supports on-premises and cloud large models, OpenRouter and its own API Key, and is compatible with MCP Marketplace and common development processes. Definition: Kilo Code is an end-to-end AI programming tool that completes the process from solution to delivery within VS Code, suitable for individuals and teams to efficiently implement complex projects, reduce repetitive labor and hallucinatory code.
21st.dev
MCP Toolset
21st.dev is a UI component discovery and management platform built for front-end design and development teams, integrating AI intelligent agents to retrieve and import high-quality components in mainstream IDEs with one click. Users can browse, share, replicate, and customize open-source components with real-time preview and version control to help unify team visual style and development specifications. The platform provides comprehensive API interfaces and Magic MCP functions, supporting batch generation, parametric customization, and rapid prototype iteration, seamlessly connecting design and code processes. The user-friendly interface and community-driven ecosystem allow design engineers and "vibe coders" to unleash their creativity in collaboration, improving project development efficiency and consistency. With no on-premises deployment required, cloud collaboration and exportable seed projects allow teams to easily expand and maintain component libraries and shorten project time-to-live.
Waveformer
AI music creation
Waveformer is an open-source AI music generation platform developed by Replicate, combined with Meta's MusicGen model, allowing users to quickly generate original music from text prompts, which can be saved as audio files or videos with waveform animations. Users can simply input descriptive text, such as "light electronic dance music" or "melodious piano melody," to generate musical compositions that meet their needs, suitable for various scenarios such as video production, podcasting, and gaming. The platform offers an intuitive user interface that supports a wide range of musical styles and moods, catering to different creative needs. Waveformer's source code is open-source on GitHub, enabling developers to engage in both local and secondary development, driving the development of AI music technology.
Melodisco
AI music creation
Melodisco is an innovative platform that combines an AI music player and generator that allows users to explore, listen, and create AI-generated music. Users can browse through trending, newly released, or random AI songs through the platform, or use the built-in music generation tools to create unique melodies based on their preferences. Melodisco offers an intuitive user interface that supports a wide range of musical styles and moods, making it accessible to music lovers, content creators, and developers. Additionally, the platform's open-source project has been released on GitHub, supporting both local deployment and secondary development, encouraging the community to jointly promote the development of AI music technology.
Piano Genie
AI music creation
Piano Genie is an AI music creation tool developed by the Google Magenta team designed to help users without a musical background play the piano with ease. Users can simply use the 1 to 8 number keys on the keyboard or tap the colored buttons on the screen, and Piano Genie converts these simple inputs into a full 88-key piano performance in real-time. The system is based on a recurrent neural network autoencoder trained on 1,400 performances in international piano competitions, which is able to understand the structure and rhythm of music and generate smooth and expressive melodies. Piano Genie offers an intuitive interface with support for MIDI inputs and outputs, making it suitable for various scenarios such as education, entertainment, and music creation. Its source code is open-sourced on GitHub, encouraging developers to engage in local deployment and secondary development, driving the development of AI music technology.
Open Voice OS
AI audio processing
OpenVoiceOS (OVOS) is a community-driven, open-source voice AI platform designed to create custom voice-controlled interfaces for various devices. The platform emphasizes privacy and security, allowing users to process voice data locally and avoid sending sensitive information to the cloud, enhancing data protection. OVOS supports a wide range of hardware platforms, including Raspberry Pi, Mycroft devices, and Linux desktops and laptops, for embedded systems and low-profile devices. Its modular architecture includes components such as ovos-core, ovos-listener, and ovos-messagebus, and supports plug-in speech recognition (STT) and text-to-speech (TTS) engines, allowing users to choose the appropriate plug-in according to their needs. OVOS also provides a wealth of developer tools and documentation to facilitate developers to create and deploy custom voice applications. As a continuation of the Mycroft project, OpenVoiceOS is committed to providing a voice assistant solution that is transparent, customizable, and respects user privacy.
n8n
AI office assistant
n8n is an open-source workflow automation platform that combines visual build with code flexibility and supports over 500 application integrations. Users can create multi-step automated processes through drag-and-drop interfaces or JavaScript/Python scripts, suitable for IT operations, security response, sales analysis, and other scenarios. n8n has built-in AI capabilities that support the construction of LLM-based agent systems for natural language processing, data extraction, and automated decision-making. The platform offers self-hosted and cloud deployment options, ensuring data security and compliance. With its strong scalability and active community support, n8n is an ideal choice for technology teams to achieve efficient automation.
DorkGPT
AI search engine
DorkGPT is an AI-powered Google Dork query generator tool designed to help users efficiently construct advanced search statements to reveal hidden web page data and potential security vulnerabilities. Through natural language processing technology, users only need to input simple query requirements, and DorkGPT can automatically generate accurate Google Dork statements, suitable for various scenarios such as cybersecurity audits, open-source intelligence (OSINT) research, and academic data retrieval. With an emphasis on responsible use, the platform encourages users to explore information within a legal and ethical framework, making it an ideal aid for security researchers, journalists, and data analysts.