Dify
Dify turns the notoriously complex process of building production-grade AI applications into a visual drag-and-drop experience that teams can actually ship and maintain. With over 87,000 GitHub stars and backing from prominent investors, the platform has become the go-to open-source LLMOps solution for organizations that refuse to be locked into proprietary AI stacks. The visual workflow canvas lets developers wire together LLM calls, conditional logic, iteration loops, tool invocations, and human-in-the-loop checkpoints without writing boilerplate integration code. Its RAG pipeline engine handles the full document lifecycle from ingestion of PDFs, Word documents, and HTML through configurable chunking strategies, embedding with models from OpenAI or open-source alternatives, vector storage in Weaviate, Qdrant, Pinecone, or pgvector, and hybrid semantic-plus-keyword retrieval with citation tracking. Dify integrates with hundreds of model providers including OpenAI GPT-4o, Anthropic Claude, Google Gemini, Mistral, Llama, and any OpenAI-compatible endpoint like Ollama for fully local inference. The agent framework supports both ReAct and function-calling strategies with 50-plus built-in tools spanning Google Search, DALL-E, Stable Diffusion, WolframAlpha, and custom API definitions. Published apps can be deployed as hosted web interfaces, embedded chat widgets, REST API endpoints, or MCP-compatible tools. Enterprise features include role-based access control, SSO integration, and audit logging. A built-in marketplace enables teams to share and reuse model providers, tools, and workflow templates across projects. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. Apache 2.0 licensed with an open-source community edition.
Weaviate
Embark on a semantic adventure with Weaviate, the open-source vector database that's like a GPS for your data! Imagine a world where your data points are like stars in the cosmos, and Weaviate is your trusty telescope, helping you navigate the galaxy of information based on the 'semantic vibes' they emit. With the ability to stand alone or play nice with a smorgasbord of vectorizing modules, this database is your ticket to a lightning-fast vector similarity search that zips through raw vectors or data objects with the speed of a comet, even when you throw in a few filters. It's like having a search party that combines the Sherlock Holmes of keyword sleuthing with the psychic predictions of vector search techniques for mind-blowingly accurate results. Want to play quiz master with your data? Pair it with any generative model and host a Q&A session right in your dataset! Hosting Weaviate on RepoCloud is like getting first-class seats for your apps on a spaceship, at economy prices. Buckle up for a next-gen vector database experience that will power your apps to infinity and beyond!
Immich
With over 110,000 GitHub stars and one of the fastest-growing open-source communities in the self-hosted space, Immich delivers a Google Photos-grade experience entirely on your own hardware. The platform handles automatic background backup from Android and iOS devices, deduplication, and support for RAW formats, LivePhotos, and MotionPhotos. Its machine learning pipeline runs facial recognition and clustering locally on your server, enabling you to group photos by person without sending a single image to the cloud. CLIP-based semantic search lets you find images by describing their content in natural language, while metadata-driven search covers EXIF data, dates, and locations. The web interface built with SvelteKit provides a responsive timeline view, albums, shared albums with configurable permissions, public sharing links with optional passwords and expiry dates, partner sharing for family libraries, and a global map plotting photos by GPS coordinates. Administrative features include multi-user support with per-user storage quotas, OAuth integration, API key management, and a user-defined storage structure for organizing files on disk. The architecture uses PostgreSQL for metadata, Redis with BullMQ for background job queues handling thumbnail generation, video transcoding, and smart search indexing, and exposes over 400 REST API endpoints documented via OpenAPI with auto-generated SDKs for web, mobile, and CLI clients. Running on a dedicated VPS on RepoCloud with guaranteed CPU, RAM, and SSD, full root SSH access, and a browser serial console. AGPL-3.0 licensed.
Perplexica
Meet Perplexica, the brainy, open-source search engine that doesn't just search the web—it practically reads your mind! Powered by the wizardry of AI, Perplexica digs through the digital universe to fetch not just any answers, but the ones you actually need. Thanks to its advanced machine learning chops, it's like having a personal research assistant who's always on the ball. Built on the robust SearxNG framework, it not only keeps you in the loop with the latest info but also guards your privacy like a digital fortress. Host it on RepoCloud, and you'll wonder how you ever browsed the net without it!
Farfalle
Live web search plus an LLM of your choice: Farfalle is an open-source, self-hosted answer engine in the Perplexity mold. Queries route through one of several search providers - self-hosted SearXNG for a fully independent stack, or Tavily, Serper, and Bing APIs - and the model composes a cited answer from the retrieved results. Model flexibility is the core design: run llama3, mistral, gemma, or phi3 locally through Ollama for zero per-query cost and full privacy, use cloud models like GPT-4o or Groq-hosted Llama 3 for speed, or route to any provider via LiteLLM. An Expert Search mode uses an agent that plans a multi-step search strategy and executes it for harder questions, and chat history keeps prior research sessions available. The stack is a Next.js and shadcn/ui frontend over a FastAPI backend with Redis rate limiting, shipped as a pre-built Docker image. A browser search-engine entry pointing at your instance makes it the default search from the address bar. Paired with SearXNG and Ollama, the whole pipeline runs with no external API at all.
Perplexica3
Meet Perplexica, your new AI-driven search engine sidekick that doesn't just search the web—it practically reads your mind! Harnessing the power of advanced machine learning, Perplexica dives deep into the ocean of the internet to fetch not just any answers, but the most relevant, crystal-clear responses, all while keeping your privacy under lock and key. Built on the robust SearxNG framework, it's like having a super-smart librarian who also respects your personal space. Hosted on RepoCloud, it's not just smart—it's also a wallet-friendly genius!