vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Browse the latest quality-filtered tools added to YunoDB from verified public project data.
1130 tools
A high-throughput and memory-efficient inference and serving engine for LLMs
Native LLM inference server for Apple Silicon. OpenAI + Anthropic API compatible. No Python. Includes MLX Core macOS app with chat, agent mode, and tool calling
Fastest enterprise AI gateway (50x faster than LiteLLM) with adaptive load balancer, cluster mode, guardrails, 1000+ models support & <100 µs overhead at 5k RPS
An open-source Agent-first Identity and Access Management (IAM) /LLM MCP & agent gateway and auth server with web UI supporting OpenClaw, MCP, OAuth, OIDC, SAML
DeepSeek-native AI coding agent for your terminal. Engineered around prefix-cache stability — leave it running.
Universal provider proxy for OpenAI Codex & Claude Code — use any LLM (Claude, Gemini, Grok, DeepSeek, Ollama…) with Codex CLI, App, SDK, and Claude Code
Code Editor for the AI Agents Era - Run an army of Claude Code, Codex, etc. on your machine
A simple, performant, and scalable Jax LLM!
SGLang is a high-performance serving framework for large language models and multimodal models.
Zero-config entity resolution that feeds a durable identity layer: resolve messy records from any source into stable golden entities — a Customer 360 with whole
Consider it done. The open-source AI agent that works out of the box · 想到,就能做到。开源、开箱即用的 AI Agent。
Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
Gradient descent, but the parameters are agents — a parallel, asynchronous framework for self-evolving agents (skills, prompts, harnesses). Diffs are the gradie
A persistent workspace for development work that self-improves and continues beyond one session.
The Best AI Agent Framework for Agent Collaboration.
🐢 Open-Source Evaluation & Testing library for LLM Agents
[EMNLP2025] "LightRAG: Simple and Fast Retrieval-Augmented Generation"
Agentic AI platform that harnesses Visual LLM Chaining to build proactive digital assistants
LLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Web dashboard for Hermes Agent — multi-platform AI chat, session management, scheduled jobs, usage analytics
Open-source desktop AI agent for tools, files, knowledge, workflows, and real deliverables.
Open Science is an open-source, local-first, model-agnostic AI research workbench for scientific discovery.
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separatio
SeaTunnel is a multimodal, high-performance, distributed, massive data integration tool.