Source Project
hugalafutro/model-hotel
Multi-Provider AI Gateway - No personal logs by design. Model autodiscovery, Failover groups, High availability, Android companion app, and more. "Because we have LiteLLM at home"
Alternatives Engine
Compare open-source alternatives to hugalafutro/model-hotel by fit, deployment, maintenance, quality, and agent readiness.
Decision Summary
Source Project
Multi-Provider AI Gateway - No personal logs by design. Model autodiscovery, Failover groups, High availability, Android companion app, and more. "Because we have LiteLLM at home"
Best For
Not Best For
Comparison Table
comet-ml/opik has the strongest combined agent score and maintenance profile in this comparison.
Alternative Match
Similar llm eval with docker/local deployment overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI and Anthropic.
Alternative Match
Similar llm eval with docker/local deployment overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
Debug, evaluate, and monitor your LLM applications, RAG systems, and agentic workflows with comprehensive tracing, automated evaluations, and production-ready dashboards.
Alternative Match
Same llm eval intent with llm_gateway overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
This repository contains comprehensive pricing and configuration data for LLMs. It powers cost attribution for 200+ enterprises running 400B+ tokens through Portkey AI Gateway every day.
Alternative Match
Similar llm eval with docker/local deployment overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
AI Observability & Evaluation
Alternative Match
Similar llm eval with local/cloud deployment overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
🐢 Open-Source Evaluation & Testing library for LLM Agents
Alternative Match
Similar llm eval with local/cloud deployment overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
Evaluation and Tracking for LLM Experiments and AI Agents
Alternative Match
Similar llm eval with docker/local deployment overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
The platform for LLM evaluations and AI agent testing
Alternative Match
Same llm eval intent with llm_gateway overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
Modular, open source LLMOps stack that separates concerns: LiteLLM unifies LLM APIs, manages routing and cost controls, and ensures high-availability, while Langfuse focuses on detailed observability, prompt versioning, and performance evaluations.
Alternative Match
Similar llm eval with docker/local deployment overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.
Alternative Match
Same llm eval intent with llm_gateway overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
Make every model a CodeX engine through an OpenAI-compatible Responses API gateway
Alternative Match
Same llm eval intent with llm_gateway overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
RouterArena: An open framework for evaluating LLM routers with standardized datasets, metrics, an automated framework, and a live leaderboard.
Alternative Match
Similar llm eval with docker/local deployment overlap.
Fit: Strong replacement candidate with overlapping indexed use cases.
Fast, flexible LLM inference
Data Source
1136 loaded projects. Generated at 2026-08-16T15:43:58.120Z.