LiteLLM is an open-source Large Language Model (LLM) gateway designed to streamline access, management, and monitoring of over 100 LLM providers through a unified, OpenAI-compatible API. By offering a consistent interface, LiteLLM simplifies the integration of diverse LLMs, enabling developers to focus on building applications without the complexities of handling multiple APIs. Key Features: - Unified API Integration: Access and manage over 100 LLM providers, including OpenAI, Azure, Anthropic
Run BiOS is serverless inference for teams running large language models in production. It exposes an OpenAI-compatible API, so you can point the OpenAI SDK at Run BiOS and keep your existing code, authenticating with a standard Authorization: Bearer header. Six model families run behind one API — Claude, DeepSeek, GLM, Kimi, MiniMax and Qwen — with context windows up to one million tokens. BiOS Adaptive routes each request for quality, speed and budget against a published price ceiling. Chat,
nRouter is an LLM gateway that unifies access to multiple model providers behind a single OpenAI-compatible API key. Model tokens pass through at flat list price with 0% token markup; a visible platform fee (4% pay-as-you-go, 2% Pro, 0% Max) is added on top of credit purchases. Plans: pay-as-you-go credits from $5, Pro $50/month with $100/month managed-router allowance, Max $200/month with $400/month allowance, and custom Enterprise. Full comparison: https://nrouter.ai/pricing
This easy to use online form builder will get you ready in minutes.
Cloaked AI is an encryption-in-use solution that protects vector embeddings without compromising usability or hampering AI use cases like anomaly detection, biometric identification, semantic search, and so on. Cloaked AI works with all known vector databases, including those from Pinecone, Weaviate, Qdrant, Elastic, and AWS OpenSearch.
Opsmeter is an AI Cost & Inference Control platform for teams building with LLMs. It shows what caused your AI bill by attributing spend and latency to endpoint tags, users/tenants, models, and prompt versions. Teams can set budgets, receive overrun alerts, and export analytics for finance and operations workflows. Integration is provider-agnostic and lightweight, with privacy-first telemetry focused on metadata rather than prompt content.
Glama.ai is a comprehensive AI workspace and integration platform that offers a unified interface to leading LLM providers, including OpenAI, Anthropic, and others. It supports the Model Context Protocol (MCP) ecosystem, enabling developers and enterprises to easily build, manage, and connect MCP-compatible services with AI agents such as Claude and GPT-4.
Bud AI Foundry is a control panel for your GenAI deployments, designed to maximize infrastructure performance and give you full control over every aspect, from deployment administration to compliance and security management—All in one place