Fast Inference | Fast hosted models, one API
DevPass by LLM Gateway - All-Access Dev Plans for AI Coding
Desert Ant Labs: On-device AI models and SDKs
Experiential Labs · The open source AI gateway
Source-code: https://github.com/experientiallabs/experiential
Waku — one native app for all your coding agents
Source-code: https://github.com/egoist/waku
mlx-optiq · Run LLMs locally on your Mac
Open-source: https://pypi.org/project/mlx-optiq/
Modal: High-performance AI infrastructure
E2B | The Enterprise AI Agent Cloud
Daytona - Secure Infrastructure for Running AI-Generated Code
drumih/turbo-fieldfare · GitHub
OmniRoute — Free AI Gateway for Multi-Provider LLMs
Source-code: https://github.com/diegosouzapw/OmniRoute
9Router - Free AI Router | Smart Fallback for Claude, Codex & More
Source-code: https://github.com/decolua/9router
lyogavin/airllm: AirLLM 70B inference with single 4GB GPU · GitHub
ayushh0110/ScreenMind: AI-powered screen memory — captures, analyzes, and lets you search/chat your screen history. Powered by Gemma 4 . 100% local, 100% private.
Local Waifu | A self-learning AI companion for macOS
Andyyyy64/whichllm: Find the local LLM that actually runs and performs best on your hardware
Locally AI - Run AI models locally on your iPhone, iPad, and Mac
Odysseus — A Self-Hosted AI Workspace
Source-code: https://github.com/odysseus-dev/odysseus
Cotabby - local AI autocomplete for macOS
Source-code: https://github.com/FuJacob/cotabby
Sim — The AI Workspace | Build, Deploy & Manage AI Agents
Source-code: https://github.com/simstudioai/sim
Caveman — the token-efficient stack for agent-native builders
Source-code: https://github.com/juliusbrussee/caveman
mozilla-ai/llamafile: Distribute and run LLMs with a single file. · GitHub
LocalAI – Offline AI Chat LLM - Apps on Google Play
LLM Hub – 15+ Private AI Models on Android, Free
Source-code: https://github.com/timmyy123/LLM-Hub