One-line proxy — cost tracking, 95% cache savings, 100+ model providers. Here are 10 similar tools you can use — including open-source options.
Real-time LLM pricing across all providers — MCP server for Claude Code
Enterprise vector semantic caching with context-enabled search
Privacy-first search API — 2K free/month
GitHub of ML — 500K+ models, free Spaces, ZeroGPU