Lemonade 11.5.0: Mastering Intelligent Model Routing and AI Workflows

Discover Lemonade 11.5.0: New intelligent model routing policies, server-side job engines, ModelScope support, and MCP client capabilities for local AI.

Discover Lemonade 11.5.0: New intelligent model routing policies, server-side job engines, ModelScope support, and MCP client capabilities for local AI.

Master real-time speech-to-text with WhisperLive. Learn how to deploy Faster Whisper, TensorRT, and OpenVino backends for high-performance AI streaming.

Secure your meetings with Meetily. A 100% local AI assistant for real-time transcription and summaries using Whisper and Ollama. No cloud, no bots.

Discover Colibri: a tiny C engine that streams experts from disk to run massive 744B parameter models like GLM-5.2 on machines with only 25GB of RAM.

Explore LongCat-2.0, a massive 1.6 trillion parameter MoE model featuring 1M context window and superior performance on SWE-bench Pro.

Upgrade your AI with advanced RAG techniques. Explore semantic chunking, reranking, and Agentic RAG architectures to eliminate hallucinations.

Discover Strix, the open-source AI toolkit that performs autonomous penetration testing, validates vulnerabilities with PoCs, and generates security patches.

Discover how Baidu's Unlimited OCR uses R-SWA to solve memory explosion in end-to-end, long-horizon document parsing tasks.

A week-long experiment with 100+ agents optimizing Gemma 4 revealed surprising social emergence, self-policing, and a 5x speed improvement in vLLM.

Learn how to install and configure Nanobot, the open-source personal AI agent you can truly own. Easy setup for local and cloud LLM providers.