Fast Inference | Fast hosted models, one API
DevPass by LLM Gateway - All-Access Dev Plans for AI Coding
Experiential Labs · The open source AI gateway
Source-code: https://github.com/experientiallabs/experiential
Modal: High-performance AI infrastructure
E2B | The Enterprise AI Agent Cloud
Daytona - Secure Infrastructure for Running AI-Generated Code
drumih/turbo-fieldfare · GitHub
OmniRoute — Free AI Gateway for Multi-Provider LLMs
Source-code: https://github.com/diegosouzapw/OmniRoute
9Router - Free AI Router | Smart Fallback for Claude, Codex & More
Source-code: https://github.com/decolua/9router
lyogavin/airllm: AirLLM 70B inference with single 4GB GPU · GitHub
Andyyyy64/whichllm: Find the local LLM that actually runs and performs best on your hardware
Sim — The AI Workspace | Build, Deploy & Manage AI Agents
Source-code: https://github.com/simstudioai/sim
Caveman — the token-efficient stack for agent-native builders
Source-code: https://github.com/juliusbrussee/caveman
zilliztech/claude-context · GitHub
Portkey - Production Stack for Gen AI Builders
BlockRunAI/ClawRouter: The agent-native LLM router for OpenClaw. 41+ models, 1ms routing, USDC payments on Base & Solana via x402
wilpel/caveman-compression: Caveman Compression is a semantic compression method for LLM contexts
AIMLAPI.com - Access 400+ AI Models with a Single AI API
NVIDIA NIM APIs
AgenticOS - The Operating System for Autonomous AI Agents
Memori – The memory fabric for enterprise AI
Syllabi - Open Source AI Chatbot Platform with RAG
Source-code: https://github.com/Achu-shankar/Syllabi
Lambda | The Superintelligence Cloud