SCANNING● LIVE DATAUTC 2026-08-13 00:00:00
VENDORS 23RECEIPTS 35CRITICAL 57D 0 NEWv0.9.4-rc1
gotnerfedgotnerfed
PUBLIC WATCHDOGEST. 2025·MIT LICENSE·NO VENDOR MONEY

Did you get
nerfed?

The public watchdog for AI tools. We catch every quiet price hike, limit cut, and silent model swap your vendors make - written up, dated, and sourced the same day. 35+ receipts on 23 vendors and counting.

Beyond the receipts, we ship the tools to do something about it - the Plan Optimizer finds what you're overpaying for, the Email Decoder translates vendor PR back into plain English, and NerfBench catches model swaps the day they happen.

23vendors tracked35+receipts filed100%source-citedMITopen dataset
LIVE FEED
scanner.log · tail -fPREVIEWONLINE · scanning
14:32:07SCANfetched cursor.com/pricing · 200 OK · 14.3kb
14:32:06DIFFanthropic.com/claude-code · 3 nodes changed · scoring...
14:32:04OKopenai.com/codex/pricing · no change · last seen 11h ago
14:32:01JURYAI Jury convened on receipt #midjourney-2026-02-v7-tier-restructure · 3 models
14:31:58SCANfetched github.com/features/copilot · 200 OK
14:31:55NERFRECEIPT FILED · midjourney · tier removed · score 73
14:31:51OKperplexity.ai/pricing · checksum match
14:31:48DIFFv0.app/pricing · footer footnote added · ignoring
14:31:42SCANstarting daily NerfBench run · 50 prompts × 4 models
14:31:40OKwindsurf.com/pricing · checksum match
14:31:36JURYAI Jury convened on receipt #v0-2025-10-idle-burn · 3 models
14:31:33SCANfetched lovable.dev/pricing · 200 OK · 8.1kb
14:31:29OKmidjourney.com/pricing · checksum match
14:31:24SCANfetched suno.com/pricing · 200 OK · 6.2kb
14:31:20DIFFgemini.google.com/pricing · 2 nodes changed · scoring...
14:31:15NERFRECEIPT FILED · vercel v0 · billing model shift · score 68
14:31:11OKdeepseek.com/pricing · checksum match
14:31:06SCANfetched replit.com/pricing · 200 OK · 11.7kb
14:31:02JURYAI Jury convened on receipt #cursor-2025-06-credit-rebill · 3 models
THE RECEIPTS

Every change we've caught, newest first.

Each row is a documented vendor move with a primary source, a plain-English summary, and a severity score. Click any row for the full receipt.

Browse all 35
DATEVENDORWHAT CHANGEDKINDSEV
2026-08-0310d agoGemini (Google)Forecast watch - Gemini (Google) on deckQuiet week receipt-side, but the forecast flags Gemini (Google) as the vendor most likely to nerf next. 94% chance of a change in the next 30 days based on cadence.OTHERMINOR2026-07-2024d agoGemini (Google)Forecast watch - Gemini (Google) on deckQuiet week receipt-side, but the forecast flags Gemini (Google) as the vendor most likely to nerf next. 94% chance of a change in the next 30 days based on cadence.OTHERMINOR2026-07-061mo agoGemini (Google)Forecast watch - Gemini (Google) on deckQuiet week receipt-side, but the forecast flags Gemini (Google) as the vendor most likely to nerf next. 95% chance of a change in the next 30 days based on cadence.OTHERMINOR2026-06-221mo agoGemini (Google)Forecast watch - Gemini (Google) on deckQuiet week receipt-side, but the forecast flags Gemini (Google) as the vendor most likely to nerf next. 82% chance of a change in the next 30 days based on cadence.OTHERMINOR2026-06-092mo agoClaude (Anthropic)Claude Fable 5 hard-refuses cybersecurity, bio/chem, and 'distillation' prompts - silently routing them to the older Opus 4.8Anthropic launched Claude Fable 5 on June 9, 2026 - a new 'Mythos-class' flagship 'made safe for general use' at $10/M input and $50/M output. The catch for security and science users: when Fable 5 detects a request in three domains - offensive/exploitation cybersecurity, most biology and chemistry, and prompts that try to 'distill' Claude's capabilities - it does not answer; the request is automatically handled by the older Claude Opus 4.8 instead. Anthropic says >95% of sessions hit no fallback and users are told when it happens. The unrestricted version, Claude Mythos 5, is gated to authorized cybersecurity partners and selected biology researchers only.OTHERMINOR2026-06-012mo agoGitHub CopilotGitHub Copilot moves to usage-based billingOn June 1, 2026 Copilot shifts from request-based to usage-based billing. Devs widely framed it as shrinkflation -- same price, less effective output.BILLINGMINOR2026-05-192mo agoGeminiGemini 3.5 Flash added to the developer API; Google's migration doc warns it costs more than Gemini 3 FlashNew model SKU added to the Gemini developer API. Gemini 3 Flash Preview remains available at its prior pricing. Google's own migration guide at ai.google.dev/gemini-api/docs/interactions/whats-new-gemini-3.5 instructs developers to switch to the new model and explicitly notes it is more expensive. Default reasoning effort changed from high to medium. Computer Use feature not supported in 3.5 Flash.TIER REMOVEDMINOR2026-05-192mo agoGeminiGemini app + AI Mode default switched from Gemini 3 Flash to 3.5 Flash with no in-app noticeAt Google I/O 2026, Google made Gemini 3.5 Flash the new default model behind the Gemini app (900M MAU) and AI Mode in Google Search (1B+ MAU) globally. The prior default was Gemini 3 Flash, which had held that spot since December 17, 2025. No proactive user-facing notice accompanied the swap. Google's developer migration documentation explicitly states the new model is 'more expensive than Gemini 3 Flash Preview.'MODEL SWAPMINOR
Showing 8 of 35 receipts · 1 criticalBrowse all 35 receipts
BIGGEST CATCH

The single worst vendor move this week.

Most receipts are minor. Some aren't. This one cost users real money - silently, overnight, with no announcement. This is what the watchdog exists for.

Full receipt
LIVE SCOREBOARD · DAY 224

Four models, fifty prompts, scored every day.

Cells show today's score out of 10 with a 7-day trend. Cheap-tier models on purpose - the prices most teams actually pay.

Open scoreboard
judged by Llama 3.3 70B (Groq) · temp 0.1 · daily 14:00 UTC
4 MODELS REPORTING · OVERALL = MEAN OF 4 CATEGORIESNEXT RUN · 14:00 UTC
Δ-pills compare to yesterday · sparklines show the 7-day overall trendFull history & per-prompt detail
PLAN OPTIMIZER

Find what you're overpaying for.

Most paid AI stacks have 2-3 subscriptions with a free or open-source equivalent that does the same job. Tell us what you pay for; we tell you what to drop, what to keep, and what to swap in. Free. No signup. Math runs on your machine.

Open optimizer
PLAN OPTIMIZER · LIVE · CLIENT-SIDE

Most paid AI stacks carry 2–3 subscriptions with a free or open-source equivalent that does the same job. Drop in what you pay for and we'll surface the cheapest stack that still covers your usage - then ship a step-by-step migration playbook to switch over.

Math runs in your browser. Nothing leaves your machine. Median Personal-plan saves $31/month on the first run; teams average $190/month across 5 seats. Free, no signup.

FREE · no signup · share via URL · math runs on your machine

SAMPLE STACK · 4 PAID SUBSSAVES $40/MO · $480/YR
CURRENTRECOMMENDEDΔ /MO
DROPCursor Pro$20/moaider OSS$0 · BYOK−$20
DROPClaude Code$20/moOpenHands OSS$0 · self-host−$20
KEEPGitHub Copilot$10/mo=keep - good $/value$10/mo$0
KEEPChatGPT Plus$20/mo=keep - voice + img gen$20/mo$0
NET MONTHLY$70 $30 · same coverage · free playbook
−$40/MO
THE TOOLS RACK

Receipts are the surface. These are the levers.

Every tool we ship is free to read and run. They share one job: turn the data we collect into a thing you can actually do on a Tuesday afternoon - cancel a sub, decode a vendor email, switch IDEs without losing your config, or catch a model swap before your prod errors do.

All tools
Plan OptimizerFREE

Tell us your paid AI subs; we find the cheapest free / OSS stack that still covers your usage. Math runs in your browser.

FREE STACKCLIENT-SIDENO SIGNUP
LIVE · OPTIMIZE · SAMPLE STACK
Cursor Pro · $20−$20
Claude Code · $20−$20
Copilot · $10$0
NET / MO−$40
FREECLIENT-SIDE · NO SIGNUP
Email DecoderFREE

Paste the "we're improving our pricing" email. We translate corporate-speak into what they actually did to you, with citations.

RECEIPT-GROUNDEDCITEDPASTE-IN
DECODER · IN → OUT
VENDOR: we're improving pricing transparency
VENDOR: to better serve our users
→ Pro price +$10/mo. Fast-credit cap −50%.
3 receipts cited.
SOURCEDEVERY CLAIM CITED
Migration PlaybooksFREE

Step-by-step instructions to move off the platform that just nerfed you - config, keys, prompts, history - in under an hour.

51 PLAYBOOKS~45 MIN AVGVERIFIED
MIGRATE · CURSOR → AIDER
Export Cursor settings.json
Install aider via uv pip
Map keymap.json → .aiderrc
Done · 18 min total
51PLAYBOOKS LIVE
NerfBenchFREE · MIT

The first public, daily behavioral benchmark for frontier models. When a vendor silently swaps a model, our score moves before the screenshots do.

DAY 22414:00 UTCMIT
NERFBENCH · DAY 224 · CURSOR
91
CURSOR · ▲ +12 · 7D
Day 224RUNS 14:00 UTC DAILY
Vendor CompareFREE

Side-by-side scorecards: limits, model defaults, comms history, Nerf Index, response time. Pick the next one without buyer's remorse.

23 VENDORSCOMMS HXSOURCED
COMPARE · CURSOR vs OPENAI
CURSOR
91
NERF · ▲ +12
OPENAI
38
NERF · ▼ −2
23VENDORS TRACKED
AskFREE

Plain-English Q&A grounded in our full receipts archive. "Did Replit just change billing again?" - get a sourced answer in 2 seconds.

FULL ARCHIVE~2sSOURCED
ASK · GROUNDED IN THE ARCHIVE
Did Replit change billing
Yes - Replit moved the Core plan from a flat agent quota to effort-based metered pricing (rolled to existing subscribers July 1, 2025).
↗ 3 SOURCES
SOURCEDEVERY ANSWER CITED
PRICING · SELF-HOSTABLE · MIT

Free forever for the public site. Pay when you want receipts that page you.

Reading every receipt and using every tool is and always will be free. Paid plans add private API access, alert webhooks, the per-vendor risk index for procurement, and quarterly NerfBench / Compare reports. No vendor money. No ads.

Full feature matrix
PUBLIC

Free

$0/forever
1 SEAT1 ALERT

Read every receipt, run every tool, and watch one vendor. No signup, no card, no tracking.

Includes
  • Every receipt · RSS + Bluesky bot
  • Full Decoder, Ask & Compare
  • NerfBench public scoreboard + daily seed
  • 1 vendor email alert
  • 30-day diff history
  • Sunday digest email
  • Embeddable badge (with attribution)
Start reading
FOR TEAMS

Teams

$29/mo
5 SEATSHOURLY

A procurement layer for your squad. Shared watchlists, hourly scans, contract-grade scorecards across 5 seats.

Everything in Personal, plus
  • 5 seats · shared watchlists · CSV export
  • Compare Pro · contract-grade scorecards
  • Cost Analyzer across your team's stack
  • Custom severity rules · hourly scans
  • 3-year diff history
  • Audit log of who acked what
  • Shared webhooks + team Slack channel
  • Custom watch URLs
FOR BUILDERS

API

$99/mo
10K EVENTSWEBHOOKS

Build on our data. Firehose API, jury verdicts, NerfBench scores, embeddable widgets - programmatic and remixable.

Everything in Teams, plus
  • Firehose API · 10k events/mo · webhooks
  • AI Jury raw verdicts + per-receipt rubric
  • NerfBench API · daily prompts & scores
  • Embeddable widgets · RBAC · multiple keys
  • Forever retention
  • White-label badges (no attribution)
  • Webhook replay + failure queue
  • SLA monitoring + status webhook
◆ FOR PROCUREMENTEnterprise$499 /mo · custom contract
Everything in API + SSO (SAML/OIDC), procurement risk dashboard, unlimited seats & events, dedicated Slack, SOC2/GDPR reporting, and a custom contract with SLA. For procurement orgs that need a paper trail their CISO will sign off on.
Talk to us →