Browse: Infrastructure
LLM7.io offers a single API gateway that connects you to a wide array of leading AI models from various providers.
This repository contains comprehensive pricing and configuration data for LLMs. It powers cost attribution for 200+ enterprises running 400B+ tokens through Portkey AI Gateway every day.
The worldโs fastest AI model gateway (450x less overhead than LiteLLM). Unified access to LLMs across endpoints (openAI, self-hosted, etc.) behind a single authentication layer - with API key generati
A simple, yet handy, LLM gateway.
Plano is an AI-native proxy and data plane for agentic apps โ with built-in orchestration, safety, observability, and smart LLM routing so you stay focused on your agents core logic.
An open-source, cloud-native, high-performance gateway unifying multiple LLM providers, from local solutions like Ollama to major cloud providers such as OpenAI, Groq, Cohere, Anthropic, Cloudflare an
gproxy is a Rust-based multi-channel LLM proxy that exposes OpenAI / Claude / Gemini-style APIs through a unified gateway, with a built-in admin console, user/key management, and request/usage auditin
Privacy-first LLM proxy and AI gateway โ load balancing, multi-provider routing, API key management, usage tracking, rate limiting. Self-hosted. Zero knowledge of your prompts.
Open-source AI gateway written in Rust, with token compression for Claude Code, Codex... and any other LLM client.
FastMCP Server for USPTO data
Fast, small, and fully autonomous AI personal assistant infrastructure, ANY OS, ANY PLATFORM โ deploy anywhere, swap anything ๐ฆ
This Guidance demonstrates how to streamline access to numerous large language models (LLMs) through a unified, industry-standard API gateway based on OpenAI API standards
High-scale LLM gateway, written in Rust. OpenTelemetry-based observability included
OmniRoute is an AI gateway for multi-provider LLMs: an OpenAI-compatible endpoint with smart routing, load balancing, retries, and fallbacks. Add policies, rate limits, caching, and observability for
A unified AI model hub for aggregation & distribution. It supports cross-converting various LLMs into OpenAI-compatible, Claude-compatible, or Gemini-compatible formats. A centralized gateway for pers
โก๏ธ Open-source AI Gateway โ Use any SDK to call 100+ LLMs. Built-in failover, load balancing, cost control & end-to-end tracing.
The only fully local production-grade Super SDK that provides a simple, unified, and powerful interface for calling more than 200+ LLMs.
Model Context Protocol Server for Swift
Controller generation for Javalin, Helidon SE.
Build and run autonomous AI agents with OpenClaw, Hermes, multiple model providers, orchestration, delegation, memory, skills, schedules, and chat connectors.
Zero trust LLM gateway. OpenAI-compatible proxy with semantic routing and load balancing across OpenAI, Anthropic, Ollama, vLLM, and any compatible backend. Identity-based access, virtual A
OpenAI-compatible HTTP LLM proxy / gateway for multi-provider inference (Google, Anthropic, OpenAI, PyTorch). Lightweight, extensible Python/FastAPIโuse as library or standalone service.
TensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation, optimization, and experimentation.
LLM API load-balancing gateway. LLM API ่ด่ฝฝๅ่กก็ฝๅ ณ.
Universal LLM Gateway: One API, every LLM. OpenAI/Anthropic-compatible endpoints with multi-provider translation and intelligent load-balancing.
Self-hosted orchestration layer for autonomous AI agent teams. Shared memory, heartbeat scheduling, vault-first secrets, and cross-model peer review โ one command to deploy.
Open-source multi-tenant AI agent platform โ 14 specialized agents, 195+ tools, 37+ AI models. Self-hosted. Fork and deploy your own AI operations team.
Zero-code LLM security & observability proxy. Real-time prompt injection detection, PII scanning, and cost control for OpenAI-compatible APIs. Built in Rust.
The world's first Autonomous Product Engine (APE): AI agents research your market, generate features, and ship code as PRs. Convoy mode, crash recovery, cost tracking, 80+ API endpoints. Self-hosted v
โก๏ธ Blazing fast LLMs API Gateway written in Go
SmarterRouter: An intelligent LLM gateway and VRAM-aware router for Ollama, llama.cpp, and OpenAI. Features semantic caching, model profiling, and automatic failover for local AI labs.
OllamaFreeAPI: Free Distributed API for Ollama LLMs Public gateway to our managed Ollama servers with: - Zero-configuration access to 50+ models - Auto load-balanced across global nodes - Free tier w
A High-Availability, Transparent, and Smart Multi-Vendor Proxy for Claude Code. Support Claude Plans, GitHub Copilot, Google Antigravity, ZAI/GLM, MiniMax, Qwen, Xiaomi, Kimi, Doubao...
KawaiiGPT โ Open-source LLM gateway accessing DeepSeek, Gemini, and Kimi-K2 through reverse-engineered Pollinations API with no API keys required, built-in prompt injection capabilities for security r
Python LLM-RAG deep agent using LangChain, LangGraph and LangSmith built on Quart web microframework and served using Hypercorn ASGI and WSGI web server.
Type-safe code generator for GraphQL schemas โ produces clients and server interfaces for Dart, Flutter, Java, and Spring Boot. Features built-in caching with TTL/tag-based invalidation, JSO
๐ฆ [2026.03.10] Hardened OpenClaw deployment on a single VPS: one command, production-ready.
๐ Use Claude Code CLI for free with NVIDIA's unlimited API. This proxy converts requests to NIM format and integrates with a Telegram bot for remote control.
๐ Transform Google Antigravity API into an OpenAI-compatible gateway, featuring multi-account support, token management, and real-time monitoring.
LLM proxy to observe and debug what your AI agents are doing.
Cloud native, ultra-high performance AI&API gateway, LLM API management, distribution system, open platform, supporting all AI APIs.๐ฆไบๅ็ใ่ถ ้ซๆง่ฝ AI&API็ฝๅ ณ๏ผLLM API ็ฎก็ใๅๅ็ณป็ปใๅผๆพๅนณๅฐ๏ผๆฏๆๆๆAI API๏ผไธ้ไบOpenAIใAzureใ
๐ Next Generation Multi-tenant AI One-Stop Solution. Builtin Admin & Billing System. Enterprise-Grade Unified LLM Gateway Support for 200+ Models And 35+ Providers, Load Balacing w/ Priority-base Rou
โ๏ธ Simplify your projects with MoLi, a fast and flexible Molang interpreter in Java, designed for easy integration and high performance.
๐ Streamline your AI CLI interactions with AIO Coding Hub, a unified gateway for Claude, Codex, and Gemini requests. Simplify setup and enhance stability.
๐ ๏ธ Streamline your development with Kiro, an agentic IDE that transforms prototypes into production using spec-driven methods and AI-powered coding support.
Enable peer-to-peer collaboration between AI agents with human supervision for complex task coordination and decision-making.
๐ Access the reverse-engineered GitHub Copilot API through this proxy, enabling streamlined integration for your development needs.
๐ Connect your phone directly to AI agents with OpenClaw Gateway, an open-source WebSocket solution free from third-party oversight.
AI-powered web app builder โ describe it, build it, ship it. 2-agent LangGraph system (Sonnet 4.5 + o4-mini) generates React apps from natural language with live preview and one-click deploy.
๐ ๏ธ Manage Minecraft Bedrock addons easily with Bedrock-Addon-Wrangler. Simplify formats, resolve UUID issues, and streamline your server experience.
Lightweight coordination server for autonomous AI coding agents โ task claiming, file locks, message passing, and health monitoring over REST
The open-source hub to build & deploy GPT/LLM Agents โก๏ธ
LSP server leveraging LLMs for code completion (and more?)
Complete open-source AI collaboration suite and multi-agent platform featuring LLM orchestration, automation, and virtual assistants. Scales seamlessly from small deployments to large enterprise envir
