LobeHub screenshot thumbnail

LobeHub

With over 82,000 GitHub stars and 700,000+ downloads, LobeHub has evolved from its origins as LobeChat into a comprehensive multi-agent AI collaboration platform where humans and autonomous agent teams co-evolve. The platform's Agent Harness architecture functions as an operating system between AI models and applications, handling prompt presets, tool orchestration, lifecycle hooks, planning, filesystem access, and sub-agent management across 25+ model providers including OpenAI, Anthropic Claude, Google Gemini, DeepSeek, Mistral, Groq, AWS Bedrock, Azure OpenAI, and local models through Ollama. Agent Groups enable sophisticated collaboration with sequential, parallel, iterative, and debate orchestration modes, allowing multiple specialized agents to tackle complex workflows simultaneously. The Agent Builder creates production-ready agents from natural language descriptions with auto-configuration, drawing from a marketplace of 505+ pre-built agents and 10,000+ MCP-compatible skills and plugins. Pages provide collaborative document editing with multi-agent co-authoring, while Schedules automate agent runs around the clock without human supervision. The knowledge base leverages PostgreSQL with pgvector for RAG-powered retrieval, and Personal Memory gives agents transparent, editable context that evolves through continual learning. The full self-hosted stack deploys via Docker Compose with PostgreSQL, Redis, RustFS for S3-compatible storage, and SearXNG for private web search, all configurable through environment variables. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. LobeHub Community licensed.

Deploy
Octop screenshot thumbnail

Octop

Modern engineering teams and busy households deploy Octop to run private, autonomous AI agents equipped with persistent memory workspaces, scheduled cron jobs, and direct browser automation. Users can orchestrate specialist agents tailored for software development, IT operations, content generation, and system diagnostics through an interactive React dashboard. The platform connects directly to Discord, Feishu, DingTalk, and WeCom, allowing team members to delegate complex tasks without leaving their everyday messaging apps. An integrated remote desktop and browser control engine lets agents navigate websites, capture screenshots, fill forms, and operate graphical software autonomously. Administrators can assign distinct MBTI personality profiles to agents, establish granular role-based permissions, configure scheduled cron workflows, and integrate custom Model Context Protocol servers for external tool access. Long-term memory persists across conversations through dedicated workspace files, ensuring contextual continuity whenever switching between underlying language models or team collaborators. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
QwenPaw screenshot thumbnail

QwenPaw

Designed as a unified personal AI workstation, QwenPaw operates as an autonomous digital copilot that coordinates scheduled automations, interactive coding sessions, and multi-channel team communication across everyday messaging apps. Users can dispatch long-running research tasks, process office documents including PDF, Excel, and Word files, and execute browser-based data collection without manual intervention. The built-in web console and terminal interface provide full visibility into agent reasoning, allowing operators to inspect intermediate thinking steps, review proposed file edits, and approve sensitive tool calls. Through direct connectors for Discord, Telegram, DingTalk, Lark, and WeChat, teams can trigger specialized skills or query shared workspaces directly from their existing group channels. A self-evolving personal memory system continuously indexes chat interactions and local resources into editable Markdown files, ensuring knowledge persists across restarts and task delegations. Built-in guardrails including a sandboxed execution runtime, tool access policies, and automated skill scanners protect underlying host files from unauthorized modifications during autonomous scripting runs. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. Apache 2.0 licensed.

Deploy
World Monitor screenshot thumbnail

World Monitor

World Monitor replaces twenty or more browser tabs by fusing geopolitical, military, financial, and infrastructure signals onto a single interactive canvas built with TypeScript, Vite, globe.gl with Three.js for the 3D globe, and deck.gl with MapLibre GL for the flat map. World Monitor ingests live data from 530 upstream sources including ACLED and UCDP for conflict events, OpenSky Network for military and civilian aircraft, AISStream for vessel positions, NASA FIRMS for satellite fire detection, USGS for earthquakes, and FRED, IMF, BIS, and Finnhub for macroeconomic and market data. The Country Instability Index v8 computes real-time stress scores across 31 Tier-1 nations, while the finance radar tracks 29 stock exchanges, commodities, and cryptocurrency with a 7-signal market composite. Six specialized dashboard variants — World, Tech, Finance, Commodity, Energy, and Happy — deploy from a single codebase. AI summarization runs locally through Ollama and LM Studio integration or optionally via Groq and OpenRouter cloud providers, with Transformers.js powering browser-side inference. The platform supports 26 languages with native-language feeds and RTL layout, and provides MCP server integration for AI agent connectivity. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. AGPL-3.0 licensed.

Deploy
TencentDB Agent Memory screenshot thumbnail

TencentDB Agent Memory

TencentDB Agent Memory provides a team-level memory hub that transforms AI agent conversations, documents, and codebases into four governed, shareable memory assets: Chat Memory for conversation history, Skills extracted from completed tasks, LLM-Wiki built from document ingestion, and Code-Graph generated from codebase analysis. The four-tier semantic pyramid structures long-term memory from L0 raw conversation capture through L1 episodic extraction and L2 scenario aggregation to L3 persona synthesis, enabling hierarchical drill-down via node and result references instead of flat vector recall. The Node.js Gateway sidecar handles capture, extraction, storage, recall, and pipeline scheduling through RESTful HTTP v2 endpoints on port 8420, while the Memory Proxy intercepts Anthropic-format API calls to inject team memory context into Claude Code, CodeBuddy, and other coding agents transparently. Local SQLite with the sqlite-vec extension provides the default storage backend with hybrid BM25 keyword plus vector embedding plus reciprocal rank fusion retrieval requiring zero external API dependencies. Teams manage ownership, versions, status, visibility, usage counts, and agent bindings through the Memory Hub dashboard with role-based access control separating System Admin and team-level Admin and Member permissions. Official TypeScript and Python SDKs provide programmatic access for custom framework integration beyond the built-in OpenClaw plugin and Hermes Agent adapter. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
RAGFlow screenshot thumbnail

RAGFlow

RAGFlow has established itself as one of the most widely adopted open-source RAG engines available, powering production AI systems that demand traceable, hallucination-free answers from complex enterprise data. The platform processes PDF, DOCX, Excel, and PPT files through vision-based deep document understanding with layout analysis and OCR, extracting structured knowledge from tables, charts, and images that simpler parsers miss entirely. RAGFlow's hybrid retrieval pipeline combines vector search with BM25 keyword matching and multi-stage reranking across configurable document stores including Elasticsearch, InfiniFlow's Infinity engine, OpenSearch, and OceanBase. Developers connect any combination of LLM providers — OpenAI, DeepSeek, Anthropic Claude, Google Gemini, and locally-hosted models via Ollama — through a unified configuration layer. The visual agent workflow system enables multi-step reasoning chains with persistent memory, tool calling, and pre-built templates for common enterprise scenarios. RAGFlow synchronizes data from Confluence, S3, Notion, and Google Drive, and delivers answers through chat integrations with Feishu, Discord, Telegram, and Line. The Python SDK and RESTful API on port 9380 provide programmatic access to knowledge base management, document parsing, and conversational retrieval. The full stack deploys via Docker Compose with MySQL for metadata, Redis for task orchestration, and MinIO for object storage. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. Apache 2.0 licensed.

Deploy
OmniRoute screenshot thumbnail

OmniRoute

OmniRoute is an AI gateway, aggregating 338 LLM providers including OpenAI, Anthropic Claude, Google Gemini, DeepSeek, Kimi, MiniMax, and GLM into a single OpenAI-compatible endpoint at localhost:20128. The gateway catalogs over 1,200 models across 90 free-tier providers and 40 free-forever providers, automatically rotating through tier-1, tier-2, and tier-3 fallback chains when any provider exhausts its quota or returns errors. RTK plus Caveman stacked token compression reduces eligible context by 15 to 95 percent before forwarding requests, cutting API costs dramatically without degrading output quality. OmniRoute exposes its full routing engine through a built-in MCP server with 104 tools across 31 scopes over stdio, HTTP, and SSE transports, plus an A2A protocol server with six autonomous agent skills and JSON-RPC 2.0 streaming. The gateway integrates directly with Claude Code, Cursor, GitHub Copilot, Codex CLI, OpenCode, and Cline through standard base-URL configuration. Seventeen routing strategies include latency-optimized, cost-minimized, and auto-scoring modes that evaluate candidates on success rate, context fit, model fitness, quota state, and circuit-breaker health. The Next.js dashboard provides real-time provider status, usage analytics, combo chain configuration, and model catalog browsing via a responsive PWA. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
Agent Zero screenshot thumbnail

Agent Zero

Agent Zero equips large language models with a persistent virtual Linux desktop, letting autonomous AI agents write code, research the web, and execute multistep tasks without supervision. Operating within an isolated container, the system opens an interactive XFCE graphical desktop in your browser canvas where the agent interacts with GUI applications like Blender and terminal sessions directly. Its embedded browser features interactive annotation tools, letting you click web elements to inspect stylesheet hierarchies, extract component markup, or submit automated testing instructions. You can cowork on spreadsheets, presentations, and Markdown specifications in real time alongside the agent, preserving document revisions through snapshot time travel. A multi-tier memory architecture backed by FAISS vector indexing and SearXNG web discovery retains problem-solving context across projects, while hierarchical delegation spins up specialized subordinate agents to parallelize complex engineering audits. The platform connects seamlessly to commercial LLMs via OpenAI, Anthropic, or OpenRouter, runs private local models through Ollama, and integrates custom toolkits via the Model Context Protocol and community Plugin Hub. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
MindsHub screenshot thumbnail

MindsHub

Backed by $50M+ from Benchmark, Y Combinator, and NVIDIA with 800+ contributors and 39,000+ GitHub stars, MindsHub Cowork is the unified AI workspace where open-source models handle entire projects — research, reporting, internal tools, scheduled operations — and return finished, shareable deliverables. The platform runs two interchangeable open-source agent harnesses, Anton and Hermes, swappable from a dropdown without losing context. A built-in Model Router pre-wires 25+ models spanning Anthropic Claude, OpenAI GPT, Google Gemini, DeepSeek, Qwen, Kimi, Grok, and MindsHub Air with automatic failover — no per-provider API keys required. A secure credentials vault connects BigQuery, PostgreSQL, Salesforce, HubSpot, Zendesk, Gong, Gmail, Google Drive, Notion, Linear, Stripe, and Slack, keeping secrets scoped per connection so agents never see raw keys. Agent output becomes publishable artifacts — documents, dashboards, apps, and code — each deployable to a live shareable URL. Cross-session persistent memory, a reusable skill library, and a background scheduler supporting hourly, daily, and weekly cadences enable autonomous recurring workflows. The architecture separates a React/Vite frontend (shipping as both Electron desktop app and web SPA) from a FastAPI backend with a versioned REST API at /api/v1 covering conversations, projects, artifacts, schedules, and connectors. Self-host via Docker Compose with nginx on port 3000 and the API on port 26866, or deploy on-prem, in a VPC, or air-gapped. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
Local Deep Research screenshot thumbnail

Local Deep Research

Execute comprehensive multi-angle investigations across scientific journals and web repositories using Local Deep Research, an autonomous intelligence engine designed for researchers, analysts, and knowledge workers. Users can submit complex analytical questions through a self-hosted web console to initiate agentic research cycles powered by LangGraph workflows. The application dynamically queries search providers including SearXNG, arXiv, PubMed, Wikipedia, and Google, iteratively evaluating discovered content to formulate targeted follow-up queries. Discovered academic papers and web sources can be indexed into an encrypted personal library to combine internal documentation with external search intelligence. Reports are synthesized with structured outlines, section headings, and verifiable academic citations that link directly to source references. Individual research sessions, system logs, and user credentials remain isolated in per-user SQLCipher databases protected by AES-256 encryption. Automated subscriptions deliver scheduled news digests and topic monitoring summaries directly to connected notification channels. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
big-AGI screenshot thumbnail

big-AGI

big-AGI is an open-source generative AI workspace that provides a unified, local-first interface for orchestrating multi-model reasoning, automated code execution, and custom persona workflows across private infrastructure. Users query multiple large language models simultaneously through the Beam scatter-gather engine, which prompts independent AI systems in parallel, compares candidate completions side by side, and merges optimal passages into a single refined response. Knowledge workers assemble tailored AI personas equipped with specialized system instructions, custom temperature settings, and predefined document context to handle domain-specific tasks ranging from architectural design reviews to legal contract analysis. The application renders rich multimedia outputs including interactive Mermaid sequence diagrams, LaTeX mathematical formulas, syntax-highlighted code blocks with live execution previews, and AI-generated image generation canvases. Teams integrate local inference servers like Ollama and LocalAI alongside commercial API endpoints to route confidential datasets strictly through internal networks while monitoring per-prompt token usage and operational latency. Users attach complex PDF documents, spreadsheets, and source code repositories for automatic parsing and semantic retrieval, while local-first storage engines ensure private chat transcripts and custom presets remain encrypted on host drives. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
Blinko screenshot thumbnail

Blinko

Capture fleeting thoughts as interactive flash cards and transform unorganized notes into an AI-searchable personal brain with Blinko. Users can jot down rapid notes, format rich markdown documents with code snippets and task lists, and attach multimedia files through a responsive web interface. An integrated retrieval-augmented generation engine indexes every note into vector embeddings, allowing users to query their entire personal archive using conversational natural language. You can connect local Ollama models or remote providers such as OpenAI, DeepSeek, Anthropic, and Grok to summarize lengthy entries and suggest hierarchical classification tags automatically. The interactive daily review interface presents random cards to help review, organize, or archive older insights before they are forgotten. Built-in tools include an integrated global music player for focused writing sessions, custom RSS feed ingestion, and password-protected note sharing with configurable expiration dates. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. GPL-3.0 licensed.

Deploy
FastGPT screenshot thumbnail

FastGPT

FastGPT lets you build production AI agents and knowledge base chatbots through a visual drag-and-drop workflow editor, connecting any LLM provider to your documents with retrieval-augmented generation that cites sources and reduces hallucination. The workflow canvas chains LLM calls, conditional branching, HTTP requests, code sandbox execution, and plugin nodes into complex conversation flows and agent skill pipelines without writing backend code. The knowledge base engine ingests documents in ten formats (TXT, Markdown, HTML, PDF, DOCX, PPTX, CSV, XLSX, URL scraping, and CSV batch import) then applies automatic chunking, hybrid vector retrieval with semantic reranking, and QA-pair splitting to deliver accurate, citation-backed answers. FastGPT connects to virtually any LLM provider through its AI Proxy aggregation layer: OpenAI GPT-4o, Anthropic Claude, Google Gemini, DeepSeek, Qwen, ERNIE Bot, and models hosted via Ollama all work through a unified OpenAI-compatible API. Bidirectional MCP support enables agents to call external tools and expose their own capabilities to other systems. Completed applications can be shared via login-free links, embedded as iframe widgets, or integrated with WeCom, Lark, DingTalk, and WeChat Official Accounts through the published REST API. Application operation logs, conversation annotation, and per-model usage analytics provide full lifecycle governance for compliance-sensitive deployments. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. FastGPT Open Source License (Apache 2.0 based) licensed.

Deploy
Eclaire screenshot thumbnail

Eclaire

Built to consolidate fragmented digital life into a private personal cloud, Eclaire centralizes tasks, notes, documents, photos, and web bookmarks into a unified workspace powered by local artificial intelligence. Users can save web articles to generate clean Markdown copies and searchable PDF archives, automatically bypassing clutter and dead links. The integrated document ingestion engine extracts tabular data from spreadsheets, performs optical character recognition across uploaded receipts and handwritten images, and indexes multi-page PDF manuals for rapid full-text search. Conversational assistant panels allow operators to interrogate their accumulated library using natural language, asking for cross-document summaries, research timelines, or action items extracted from meeting transcripts. Background workers prioritize asynchronous indexing queues, categorizing image libraries with visual object tags and managing recurring task deadlines with automated Telegram alert notifications. External automation flows can interact directly with the unified storage layer via an OpenAI-compatible REST API, allowing mobile shortcuts and custom scripts to ingest bookmarks and dictate notes remotely. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
Doccano screenshot thumbnail

Doccano

Doccano is a text annotation platforms for building machine learning training datasets. The web-based interface supports text classification for sentiment analysis and document categorization, sequence labeling for named entity recognition with overlapping entity support and relation extraction between labeled spans, and sequence-to-sequence annotation for text summarization and machine translation pairs. Collaborative annotation enables multiple annotators to work on the same project simultaneously with per-user progress tracking, annotation guidelines, example assignment to specific members, and filtering by assignee. Auto-labeling integrates with external machine learning model APIs through configurable request and response mapping templates, allowing pre-annotation that annotators can review and correct. Data import accepts plain text, JSONL, CoNLL, and Excel formats, while export produces JSONL and CoNLL datasets compatible with spaCy, Hugging Face Transformers, PaddleNLP, and other training frameworks through the doccano-transformer library. The Django backend with Django REST Framework exposes a complete RESTful API for programmatic project creation, dataset management, and annotation retrieval via the official doccano-client Python library. Celery handles background tasks including dataset import and export processing with Flower providing task monitoring. One-click deployment supports AWS CloudFormation and Heroku alongside Docker Compose for self-hosted environments. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. MIT licensed.

Deploy
EDDI screenshot thumbnail

EDDI

Deploy autonomous conversational AI agents and coordinate complex multi-agent workflows through EDDI, an open-source orchestration middleware that turns declarative configuration files into secure, production-grade enterprise assistants. Engineering teams connect AI models from twelve different commercial and local providers, using bilateral Model Context Protocol tools to let external desktop clients and coding assistants interact directly with running conversational services. Autonomous agents collaborate through structured group interaction patterns including round table discussions, peer reviews, Delphi consensus rounds, and devil's advocate debates to refine generated solutions before delivery. Declarative workflow extensions automate external REST API calls with dynamic request templating, extracting structured properties into persistent conversation memory and chaining multi-step service interactions without custom script glue. Built-in secret vaults safely manage sensitive API credentials, while the centralized web dashboard provides live conversation inspection, token quota enforcement, and comprehensive audit logging for regulatory compliance. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. Apache 2.0 licensed.

Deploy