zilliztech/claude-context · GitHub
BlockRunAI/ClawRouter: The agent-native LLM router for OpenClaw. 41+ models, 1ms routing, USDC payments on Base & Solana via x402
Portkey - Production Stack for Gen AI Builders
wilpel/caveman-compression: Caveman Compression is a semantic compression method for LLM contexts
Together AI – Fast Inference, Fine-Tuning & Training
AIMLAPI.com - Access 400+ AI Models with a Single AI API
AgenticOS - The Operating System for Autonomous AI Agents
NVIDIA NIM APIs
Syllabi - Open Source AI Chatbot Platform with RAG
Source-code: https://github.com/Achu-shankar/Syllabi
LLMLingua Series | Effectively Deliver Information to LLMs via Prompt Compression
Source-code: https://github.com/microsoft/LLMLingua
Amazon Q - AWS
Stability-AI/StableSwarmUI: StableSwarmUI, A Modular Stable Diffusion Web-User-Interface, with an emphasis on making powertools easily accessible, high performance, and extensibility.
cumulo-autumn/StreamDiffusion · GitHub
alibaba/MNN · GitHub
AMD-AIG-AIMA/Instella: Fully Open Language Models with Stellar Performance
microsoft/semantic-kernel: Integrate cutting-edge LLM technology quickly and easily into your apps
BentoML: Build, Ship, Scale AI Applications
Source-code: https://github.com/bentoml/OpenLLM
Memori – The memory fabric for enterprise AI
OpenRouter
sgl-project/sglang: SGLang is a fast serving framework for large language models and vision language models.
vllm-project/vllm: A high-throughput and memory-efficient inference and serving engine for LLMs