We check every product on this page once a day and record whether it still answers.
Routes prompts to the cheapest capable model tier and offers a terminal coding agent claiming lower cost than Claude Code on a benchmark.
Free
What to know▼
How it works
Routes every prompt to the cheapest model tier that can handle it, escalating only when needed; its terminal agent Klaat Code claims to solve the same 33-task benchmark as Claude Code at 5.4x lower cost ($0.027 vs $0.146/task).
Pricing
Free beta: 75 requests/day, no card; tool calls are free, only messages count against quota.
We check every product on this page once a day and record whether it still answers.
Reverse proxy that enforces hard spending caps on LLM API calls across providers by swapping the base URL.
Watch out · AGPL-3.0, no telemetry.
FreeNo trackingSelf-hosted
What to know▼
How it works
A transparent reverse proxy enforcing hard hourly/daily spending caps per LLM API client via a one-line base_url change, returning 429 once the cap is hit rather than calling the upstream.
What's different
Works with OpenAI, Anthropic, Mistral, Groq, DeepSeek, and 5 more providers.
Pricing
Hosted beta at proxai.eu is free, no credit card; self-hostable as a ~10 MB Docker image.
We check every product on this page once a day and record whether it still answers.
A local gateway that lets every AI coding client share one set of MCP server connections instead of configuring each tool separately, cutting how many tools your agent has to load.
Watch out · The '91% overhead cut' figure is the product's own claim; we haven't independently benchmarked it.
FreeNo sign-upWorks offline
What to know▼
How it works
A local-first, open-source gateway. You set up and authenticate each MCP server once, and every connected AI client — Claude, Cursor, VS Code, Codex — shares that same connection instead of you configuring it per tool.
What's different
Claims up to 91% less tool-token overhead by giving an agent a handful of meta-tools instead of loading hundreds of individual tool definitions; API keys stay in the OS keychain.
Pricing
Free, open source, no sign-up.
Best for
Developers running multiple AI coding clients against the same set of MCP servers who don't want to re-authenticate each one separately.
Watch out
The '91% overhead cut' figure is the product's own claim; we haven't independently benchmarked it.
We check every product on this page once a day and record whether it still answers.
A synced multi-model AI chat app across iOS, Android, and web using your own API keys.
FreeWorks offline
What to know▼
How it works
Lets you chat with ChatGPT, Claude, Gemini, and 10+ other AI models in one app across iOS, Android, and web, synced in real time, using your own API keys.
What's different
Includes real-time cost tracking, local-first storage, saved Notes, and cross-checking replies with a second model.
Pricing
Try free, no key needed; no markup or extra subscription with your own keys.
We check every product on this page once a day and record whether it still answers.
Counts your AI usage locally across ChatGPT, Claude, Gemini, Claude Code and Cursor in one dashboard, showing your real total and how close you are to each limit.
FreeNo sign-up
What to know▼
What's different
Counts locally rather than relying on each provider's own scattered usage view.
We check every product on this page once a day and record whether it still answers.
A tool that generates a themed sci-fi/fantasy AI universe with tradable in-game artifacts and a leaderboard.
FreeNo sign-up
What to know▼
How it works
Generates a themed AI universe (sci-fi, fantasy, cyberpunk or steampunk) with tradable artifacts in about 7 seconds, with a live marketplace, TC token staking and an AI assistant called JARVIS.
Pricing
3 free generations, no signup needed.
Best for
People wanting to generate and trade AI-created world assets.
We check every product on this page once a day and record whether it still answers.
A Windows taskbar utility that tracks remaining usage limits and reset times across Claude, Codex, Cursor, and other AI tools.
Watch out · Windows only; works with Claude, Codex, Cursor, Gemini, and Copilot.
FreeWorks offline
What to know▼
How it works
Keeps your remaining AI capacity and reset times visible above the Windows taskbar, alerting you when a limit comes back early or partially restores; hovering shows per-model and per-project breakdowns and an equivalent API cost.
Pricing
Free and local-first.
Watch out
Windows only; works with Claude, Codex, Cursor, Gemini, and Copilot.
We check every product on this page once a day and record whether it still answers.
Benchmark database showing which open-weight AI models run on which consumer GPUs, with install recipes and a GPU advisor.
FreeNo sign-up
What to know▼
How it works
Provides honest benchmarks for 84 open-weight AI models across 27 consumer GPUs from RTX 3060 to Apple Silicon, each pair rated as runs, tight fit, or won't fit with real speed and VRAM numbers, plus 800+ step-by-step install recipes for llama.cpp, Ollama, ComfyUI, and MLX, and a GPU Advisor.
We check every product on this page once a day and record whether it still answers.
A self-hostable AI chat platform that routes queries to the cheapest LLM and adds live market data, positioned as an alternative to ChatGPT or Claude you run yourself. For developers and cost-conscious teams.
FreeOpen sourceSelf-hosted
What to know▼
How it works
A self-hostable, open-source AI chat platform that runs on your own infrastructure, integrates live financial market data into chat, and routes each query to the best/cheapest LLM automatically across OpenAI, Anthropic, SiliconFlow and others.
What's different
Self-hosted with full data ownership and no vendor lock-in, plus intelligent model routing claimed to cut API costs by up to 60% versus using a single provider.
Pricing
Free forever to self-host, with Free / Pro / MAX tiers and per-model quotas.
Best for
Developers and cost-conscious teams who want a ChatGPT/Claude alternative running on their own infrastructure.
We check every product on this page once a day and record whether it still answers.
A growing collection of free, no-signup online calculators covering AI tokens, fees, solar savings, and health tracking.
FreeNo sign-up
What to know▼
How it works
A collection of free calculators covering AI token estimation, PayPal/Etsy fees, solar savings, water intake, pregnancy due dates, and period tracking.
We check every product on this page once a day and record whether it still answers.
A free color palette generator and design-token tool for designers and developers, extracting colors from images and exporting to CSS, Tailwind, or design tokens.
FreeCan export my data
What to know▼
How it works
Generates palettes, extracts colors from images, checks contrast, and exports to CSS, Tailwind, and design tokens.
We check every product on this page once a day and record whether it still answers.
A free local VS Code/Cursor extension that turns your Claude Code session files into a browsable dashboard with full transcripts, tool calls, sub-agents and per-response token and cost tracking.
FreeWorks offline
What to know▼
How it works
A VS Code/Cursor/VSCodium extension that turns local Claude Code session files into a browsable workspace with full transcripts, tool calls, sub-agents, per-response token and cost tracking, an analytics dashboard, and a CLAUDE.md memory view.
What's different
100% local-first — nothing leaves your machine.
Pricing
Free.
Best for
Developers using Claude Code who want to review sessions, costs and tool calls without a CLI-only cost counter.
We check every product on this page once a day and record whether it still answers.
Converts PDFs, Word, Excel, PowerPoint, CSV and JSON files into clean Markdown formatted for feeding into ChatGPT or Claude, free to start with no account required.
FreeNo sign-up
What to know▼
How it works
Converts PDFs, Word, Excel, PowerPoint, CSV, and JSON files into clean Markdown formatted to work well as input for AI models like ChatGPT and Claude.
What's different
Optimizes the output specifically to use fewer tokens and produce better AI answers, rather than just generic format conversion.
Pricing
Free to start, no account needed.
Best for
People who need to feed messy documents into ChatGPT or Claude and want cleaner, more token-efficient input.
We check every product on this page once a day and record whether it still answers.
Generates AI color palettes as design.md token files for tools like Claude, Cursor, and v0 — a quick color-scale generator for developers and designers doing AI-assisted work.
FreeNo sign-up
What to know▼
How it works
Generates AI color palettes as design.md token files — the format used by tools like Claude, Cursor and v0 — including 12-step scales, light/dark modes and CSS variables.
What's different
Outputs directly in the design.md token format that AI coding tools expect, rather than a generic palette export.
Pricing
Free, no signup.
Best for
Developers and designers building AI-assisted workflows who need design tokens fast.
We check every product on this page once a day and record whether it still answers.
Honey is a free, open-source (MIT) skill that reduces token usage for AI coding agents like Claude Code by enforcing concise code and prose conventions and cheaper bulk reads.
Free
What to know▼
How it works
enforces YAGNI/stdlib-first code, answer-first prose, compact JSON/ESON handoffs between agents, and renders bulk reads as images to cut token cost
Pricing
free, MIT license
Best for
developers running AI coding agents (Claude Code, GPT, etc.) who want lower token usage
We check every product on this page once a day and record whether it still answers.
Routes each API prompt to the cheapest model that can handle it, cutting AI API costs — a developer tool for teams with real LLM spend, with a free tier.
Free
What to know▼
How it works
Routes each API prompt through a single endpoint to the cheapest model capable of handling it, cutting AI API costs by 30-80%.
Pricing
Free tier included.
Best for
Developer teams with real LLM API spend looking to cut costs without changing their integration.
We check every product on this page once a day and record whether it still answers.
A free, local-first CLI that reads logs already on your machine to estimate spend on AI coding tools like Claude Code and Cursor.
FreeWorks offline
What to know▼
How it works
A local-first CLI that reads logs already on the machine (no proxy, no traffic in the middle) to estimate spend on AI coding tools like Claude Code, Cursor, and Codex from published model pricing, printing monthly reports and shareable cards.
Pricing
Free; install via 'npx lmspend'. Optional dashboard adds history, budgets, alerts, and team roll-ups.
Best for
Individuals and teams wanting to track AI coding tool costs before the invoice arrives.