AI Cost Monitor

Zero-dependency CLI that benchmarks latency (TTFT) and pricing across OpenAI-compatible AI providers — before you wire them into production.

Dependencies: 0 Python: 3.10+ License: MIT Probe cost: ~fractions of a cent
View on GitHub Quick start
⚡

Latency probe

Measures real time-to-first-token with streaming requests. Best, median and worst across N runs.

💰

Pricing catalog

Fetches a provider's /v1/models catalog with per-million token prices and discounts.

📦

Zero dependencies

Pure Python standard library. Clone and run — nothing to install.

🔒

Wallet-safe

Probes send max_tokens=1 and warn before spending anything.

Quick start

# clone and configure
git clone https://github.com/frost13ru-svg/ai-cost-monitor.git
cd ai-cost-monitor && cp .env.example .env

# benchmark a model (near-zero cost)
./run.sh probe --model gpt-6-luna --runs 3