High-performance LLM inference engine — drop-in replacement for Ollama with faster multi-turn inference, lower TTFT, and higher throughput through prefix caching and continuous batching.
An open-source, code-first Go toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.
#1 Persistent memory for AI coding agents based on real-world benchmarks
Official repo for spec & SDK of MCP Apps protocol - standard for UIs embedded AI chatbots, served by MCP servers
✨ Self hosted, always-on AI agent platform run in containers. Create multiple bots with long memory, and connect them to Telegram, Discord, Feishu(Lark), Matrix, etc.
wasmcloud 是一个编写可移植业务逻辑的平台,可以在从边缘到云的任何地方运行,它拥有一个安全的默认、无锅炉板的开发者体验和快速反馈回路