Back to Langfuse

Model Price Audit Memory

.agents/skills/add-model-price/references/model-audit-memory.md

4.7.024.6 KB
Original Source

Model Price Audit Memory

This file is an optional snapshot of the latest automated audit whose per-model results add useful context for a future run. It is orientation only; reconfirm every price and tier against official provider sources before making a change or reporting a row as confirmed.

The audit agent may replace the snapshot below with its complete current table. Keep only one snapshot, never append an unbounded run history, never persist a partial set of checked models, and do not update this file only to refresh the audit date.

Latest useful snapshot

Audit date: 2026-08-07

All prices listed as $X / MTok (per million tokens). Per-token JSON values: divide by 1,000,000.

ProviderModel / pricing entryPricing checkedPrice confirmedTiering checkedTiering correctChangeOfficial source(s)Comments
Anthropicclaude-fable-5Input $10/MTok, Output $50/MTok, 5m $12.50/MTok, 1h $20/MTok, read $1/MTokYesFlat 1M context at standard pricingYesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed via full pricing table fetch, unchanged since Aug 4.
Anthropicclaude-mythos-5Same as Fable 5YesFlat 1M contextYesNonehttps://platform.claude.com/docs/en/about-claude/pricingLimited availability (Project Glasswing). Re-confirmed unchanged.
Anthropicclaude-opus-5Input $5/MTok, Output $25/MTok, 5m $6.25/MTok, 1h $10/MTok, read $0.50/MTokYesFlat 1M contextYesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed. Fast mode $10/$50 MTok unchanged.
Anthropicclaude-opus-4-8Same as Opus 5YesFlat 1M contextYesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed unchanged.
Anthropicclaude-opus-4-7Same as Opus 5YesFlat 1M contextYesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed unchanged.
Anthropicclaude-opus-4-6Same as Opus 5YesFlat 1M contextYesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed. inference_geo: "us" still adds 1.1x.
Anthropicclaude-opus-4-5-20251101Same as Opus 5YesFlat 1M contextYesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed unchanged.
Anthropicclaude-opus-4-1-20250805Input $15/MTok, Output $75/MTok, 5m $18.75/MTok, 1h $30/MTok, read $1.50/MTokYesDeprecated — no tieringNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingStill listed as "retired, except on Bedrock and Google Cloud" on the current page. Entry retained, unchanged.
Anthropicclaude-opus-4-20250514Input $15/MTok, Output $75/MTok, 5m $18.75/MTok, 1h $30/MTok, read $1.50/MTokYesRetired except Google Cloud — no tieringNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed present on current page's main table, unchanged.
Anthropicclaude-sonnet-5Input $2/MTok, Output $10/MTok (through Aug 31, 2026); 5m $2.50/MTok, 1h $4/MTok, read $0.20/MTokYesFlat 1M context; introductory pricing through Aug 31 2026YesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed. Standard pricing $3/$15 (cache $3.75/$6/$0.30) still scheduled for Sep 1, 2026 — update the file then.
Anthropicclaude-sonnet-4-6Input $3/MTok, Output $15/MTok, 5m $3.75/MTok, 1h $6/MTok, read $0.30/MTokYesFlat 1M contextYesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed unchanged.
Anthropicclaude-sonnet-4-5-20250929Input $3/MTok, Output $15/MTok, 5m $3.75/MTok, 1h $6/MTok, read $0.30/MTokYesNo large-context tier (200k hard context-window cap)YesNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed unchanged; the Large Context tier removed on Aug 4 2026 was not reintroduced.
Anthropicclaude-sonnet-4-20250514Input $3/MTok, Output $15/MTok, 5m $3.75/MTok, 1h $6/MTok, read $0.30/MTokYesRetired except Bedrock/Google Cloud — no tieringNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed present on current page's main table, unchanged.
Anthropicclaude-haiku-4-5-20251001Input $1/MTok, Output $5/MTok, 5m $1.25/MTok, 1h $2/MTok, read $0.10/MTokYesNo large-context tier (200k context window, not on flat 1M list)Not applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed unchanged.
Anthropicclaude-3-5-haiku-20241022Input $0.80/MTok, Output $4/MTok, 5m $1/MTok, 1h $1.60/MTok, read $0.08/MTokYesRetired except Bedrock/Google CloudNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingRe-confirmed present on current page's main table ("Claude Haiku 3.5").
Anthropicclaude-3.7-sonnet-20250219Input $3/MTok, Output $15/MTok, cache $3.75/$6/$0.30NoNot on current pageNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingNot on current page this run either. Legacy prices retained, not re-verified.
Anthropicclaude-3.5-sonnet-20241022Input $3/MTok, Output $15/MTok, cache $3.75/$6/$0.30NoNot on current pageNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingNot re-verified this run. Legacy prices retained.
Anthropicclaude-3-5-sonnet-20240620Input $3/MTok, Output $15/MTok, cache $3.75/$6/$0.30NoNot on current pageNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingNot re-verified this run. Legacy prices retained.
Anthropicclaude-3-opus-20240229Input $15/MTok, Output $75/MTokNoNot on current pageNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingNot re-verified this run. Legacy.
Anthropicclaude-3-sonnet-20240229Input $3/MTok, Output $15/MTokNoNot on current pageNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingNot re-verified this run. Legacy.
Anthropicclaude-3-haiku-20240307Input $0.25/MTok, Output $1.25/MTokNoNot on current pageNot applicableNonehttps://platform.claude.com/docs/en/about-claude/pricingNot re-verified this run. Legacy.
AWS Bedrockclaude-3-5-sonnet-20240620 / claude-3.5-sonnet-20241022 (Public Extended Access SKU)$6.00/MTok input, $30.00/MTok output, $7.50/MTok cache write, $0.60/MTok cache readYes (SKU confirmed real, Aug 4 2026)Distinct dated SKU, not a context-length tierNot applicableUnresolvedhttps://aws.amazon.com/bedrock/pricing/Not re-verified this run; permanent documented limitation (model-ID string match cannot distinguish billing SKU) — see provider-sources-and-price-keys.md.
OpenAIgpt-5.6-solInput $5/MTok, Cached $0.50/MTok, Cache write $6.25/MTok, Output $30/MTokYesLarge Context (>272K): $10/$1.00/$12.50/$45YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed via full standard-pricing-table dump this run.
OpenAIgpt-5.6-terraInput $2/MTok, Cached $0.20/MTok, Cache write $2.50/MTok, Output $12/MTokYesLarge Context (>272K): $4/$0.40/$5.00/$18YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed stable.
OpenAIgpt-5.6-lunaInput $0.20/MTok, Cached $0.02/MTok, Cache write $0.25/MTok, Output $1.20/MTokYesLarge Context (>272K): $0.40/$0.04/$0.50/$1.80YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed stable.
OpenAIgpt-5.5-2026-04-23 (alias gpt-5.5)Input $5/MTok, Cached $0.50/MTok, Output $30/MTokYesLarge Context (>272K): $10/$1.00/$45YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged. No cache-write pricing for this model.
OpenAIgpt-5.5-pro-2026-04-23 (alias gpt-5.5-pro)Input $30/MTok, Output $180/MTok; no cacheYesLarge Context (>272K): $60/$270YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5.4Input $2.50/MTok, Cached $0.25/MTok, Output $15/MTokYesLarge Context (>272K): $5.00/$0.50/$22.50YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5.4-2026-03-05Same as gpt-5.4YesLarge Context (>272K): $5.00/$0.50/$22.50YesNonehttps://developers.openai.com/api/docs/pricingDated snapshot sibling; re-confirmed.
OpenAIgpt-5.4-proInput $30/MTok, Output $180/MTok; no cacheYesLarge Context (>272K): $60/$270YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5.4-pro-2026-03-05Same as gpt-5.4-proYesLarge Context (>272K): $60/$270YesNonehttps://developers.openai.com/api/docs/pricingDated snapshot sibling; re-confirmed.
OpenAIgpt-5.4-miniInput $0.75/MTok, Cached $0.075/MTok, Output $4.50/MTokYesNo large-context tier (dashes confirmed)YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5.4-mini-2026-03-17Same as gpt-5.4-miniYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingDated snapshot sibling; re-confirmed.
OpenAIgpt-5.4-nanoInput $0.20/MTok, Cached $0.02/MTok, Output $1.25/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5.4-nano-2026-03-17Same as gpt-5.4-nanoYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingDated snapshot sibling; re-confirmed.
OpenAIgpt-5.3-codexInput $1.75/MTok, Cached $0.175/MTok, Output $14.00/MTokYesNo large-context tier (400k context window, single tier)YesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5.2-2025-12-11Input $1.75/MTok, Cached $0.175/MTok, Output $14.00/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged (row: "gpt-5.2").
OpenAIgpt-5.1-2025-11-13Input $1.25/MTok, Cached $0.125/MTok, Output $10/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5-2025-08-07Input $1.25/MTok, Cached $0.125/MTok, Output $10/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5-mini-2025-08-07Input $0.25/MTok, Cached $0.025/MTok, Output $2/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5-nano-2025-08-07Input $0.05/MTok, Cached $0.005/MTok, Output $0.40/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-5-pro-2025-10-06Input $15/MTok, Output $120/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged via full standard-pricing-table dump (previously only confirmed via dedicated model page).
OpenAIgpt-5.2-pro-2025-12-11Input $21/MTok, Output $168/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged via full standard-pricing-table dump.
OpenAIgpt-4.1-2025-04-14Input $2/MTok, Cached $0.50/MTok, Output $8/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-4.1-mini-2025-04-14Input $0.40/MTok, Cached $0.10/MTok, Output $1.60/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-4.1-nano-2025-04-14Input $0.10/MTok, Cached $0.025/MTok, Output $0.40/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-4o-2024-08-06Input $2.50/MTok, Cached $1.25/MTok, Output $10/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-4o-2024-05-13Input $5/MTok, Output $15/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingNewly cross-checked this run via full table dump; matches existing file value exactly.
OpenAIgpt-4o-mini-2024-07-18Input $0.15/MTok, Cached $0.075/MTok, Output $0.60/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIo1Input $15/MTok, Cached $7.50/MTok, Output $60/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingNewly cross-checked this run via full table dump; matches existing file value exactly.
OpenAIo1-proInput $150/MTok, Output $600/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingNewly cross-checked this run via full table dump; matches existing file value exactly.
OpenAIo3-proInput $20/MTok, Output $80/MTok; no cacheYesNo large-context tier (200k context window)YesNonehttps://developers.openai.com/api/docs/pricing https://developers.openai.com/api/docs/models/o3-proNewly cross-checked this run (not in prior snapshot); matches existing file value exactly.
OpenAIo3-2025-04-16Input $2/MTok, Cached $0.50/MTok, Output $8/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIo3-mini-2025-01-31Input $1.10/MTok, Cached $0.55/MTok, Output $4.40/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIo4-mini-2025-04-16Input $1.10/MTok, Cached $0.275/MTok, Output $4.40/MTokYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-4-turbo-2024-04-09Input $10/MTok, Output $30/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingNewly cross-checked this run via full table dump; matches existing file value exactly.
OpenAIgpt-4-0613Input $30/MTok, Output $60/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingNewly cross-checked this run via full table dump; matches existing file value exactly.
OpenAIgpt-3.5-turboInput $0.50/MTok, Output $1.50/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingRe-confirmed unchanged.
OpenAIgpt-3.5-turbo-0125Input $0.50/MTok, Output $1.50/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingNewly cross-checked this run via full table dump; matches existing file value exactly.
OpenAIgpt-3.5-turbo-1106Input $1.00/MTok, Output $2.00/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingNewly cross-checked this run via full table dump; matches existing file value exactly.
OpenAIgpt-3.5-turbo-instructInput $1.50/MTok, Output $2.00/MTok; no cacheYesNo large-context tierYesNonehttps://developers.openai.com/api/docs/pricingNewly cross-checked this run via full table dump; matches existing file value exactly.
OpenAIdavinci-002Input $2.00/MTok, Output $2.00/MTok (base/non-fine-tuned inference); no cacheYesNo large-context tierYesUpdatedhttps://developers.openai.com/api/docs/pricingFixed a long-standing bug (created Jan 2024, never updated). The plain davinci-002 entry (matches bare model ID, not the ft:davinci-002:... fine-tuned form) had been priced at $6/$12, which is actually the Fine-tuning (Legacy) table's Training/Input/Output rate for davinci-002, not the base-model inference rate. Three independent targeted fetches of the official pricing page confirmed the base "Standard/Specialized models" table shows davinci-002 at $2.00/$2.00, distinct from the Fine-tuning table's $6.00 training + $12.00/$12.00 input/output. The separate ft:davinci-002 entry (id cls08rv9g000508jq5p4z4nlr) already correctly holds the $12/$12 fine-tuning inference rate and was left unchanged.
OpenAIbabbage-002Input $0.40/MTok, Output $0.40/MTok (base/non-fine-tuned inference); no cacheYesNo large-context tierYesUpdatedhttps://developers.openai.com/api/docs/pricingFixed the same class of bug as davinci-002. The plain babbage-002 entry's output price was $1.60 (the Fine-tuning Legacy table's output rate), corrected to $0.40 (the Standard/base-model table rate); input was already correct at $0.40. The separate ft:babbage-002 entry (id cls08s2bw000608jq57wj4un2) already correctly holds the $1.60/$1.60 fine-tuning inference rate and was left unchanged.
OpenAIgpt-5-chat-latestInput $1.25/MTok, Cached $0.125/MTok, Output $10/MTokNoNo provider tieringNot applicableNonehttps://developers.openai.com/api/docs/models/gpt-5-chat-latestNot re-verified this run (not present in the full standard-table dump, which does not include this alias); retained from July 2026 audit.
Googlegemini-2.5-flashInput $0.30/MTok, Audio $1/MTok, Output $2.50/MTok, Cache read $0.03/MTok (audio $0.10/MTok)YesNo large-context tierYesNonehttps://ai.google.dev/pricingRe-confirmed unchanged.
Googlegemini-2.5-flash-liteInput $0.10/MTok, Audio $0.30/MTok, Output $0.40/MTok, Cache read $0.01/MTok (audio $0.03/MTok)YesNo large-context tierYesNonehttps://ai.google.dev/pricingRe-confirmed unchanged.
Googlegemini-2.5-proInput $1.25/$2.50 MTok (≤200K/>200K), Output $10/$15, Cache read $0.125/$0.25YesLarge Context (>200K) confirmedYesNonehttps://ai.google.dev/pricingRe-confirmed unchanged.
Googlegemini-3.5-flashInput $1.50/MTok, Output $9.00/MTok, Cache read $0.15/MTokYesNo large-context tierYesNonehttps://ai.google.dev/pricingRe-confirmed unchanged.
Googlegemini-3.5-flash-liteInput $0.30/MTok, Output $2.50/MTok, Cache read $0.03/MTokYesNo large-context tierYesNonehttps://ai.google.dev/pricing https://ai.google.dev/gemini-api/docs/pricingRe-confirmed via two fresh targeted verbatim fetches explicitly separating Free/Paid tier columns for the "Context caching price" cell: Free tier = "Not available", Paid tier = "$0.03" (text/image/video) plus a non-representable $1.00/MTok/hour storage price. An initial broad (non-targeted) fetch this run incorrectly reported "Not available" for this model's caching — a repeat of the exact free/paid column-collapse artifact documented for July 2026; always use a targeted verbatim-quote fetch for this specific cell, never trust a broad table-dump summary for it.
Googlegemini-3.1-flash-liteInput $0.25/$0.50 (text/audio), Output $1.50, Cache read $0.025/$0.05YesNo large-context tierYesNonehttps://ai.google.dev/pricingRe-confirmed unchanged.
Googlegemini-3.1-flash-lite-previewSame as GA gemini-3.1-flash-liteNoNo large-context tierNot applicableNonehttps://ai.google.dev/pricingNot separately listed on official page this run either; not re-verified.
Googlegemini-3.1-pro-previewInput $2/$4 MTok (≤200K/>200K), Output $12/$18YesLarge Context (>200K) confirmedYesNonehttps://ai.google.dev/pricingRe-confirmed unchanged.
Googlegemini-3-flash-previewInput $0.50/$1.00 (text/audio), Output $3.00, Cache read $0.05/$0.10YesNo large-context tierYesNonehttps://ai.google.dev/pricingRe-confirmed unchanged.
Googlegemini-3-pro-previewInput $2/$4 MTok (≤200K/>200K), Output $12/$18NoLarge Context (>200K) set in fileNot applicableNonehttps://ai.google.dev/pricingStill not listed on official AI Studio page this run either; existing prices retained, not re-verified.
Googlegemini-3.6-flashInput $1.50/MTok, Output $7.50/MTok, Cache read $0.15/MTokYesNo large-context tierYesNonehttps://ai.google.dev/pricingRe-confirmed unchanged.
Googlegemini-2.0-flashInput $0.10/MTok, Output $0.40/MTokNoDeprecated (shut down June 1, 2026)Not applicableNonehttps://ai.google.dev/pricingNot re-verified this run; retained for backward compatibility.
Googlegemini-2.0-flash-001Same as gemini-2.0-flashNoDeprecated (shut down June 1, 2026)Not applicableNonehttps://ai.google.dev/pricingNot re-verified this run; retained for backward compatibility.

Unresolved findings (updated 2026-08-07)

  1. claude-sonnet-5 introductory pricing — Introductory pricing ($2/$10/MTok) expires August 31, 2026. Standard pricing ($3/$15/MTok, cache $3.75/$6/$0.30) takes effect September 1, 2026. The pricing file must be updated before or on that date.

  2. claude-opus-4-1-20250805 retirement — Deprecated, past its originally stated Aug 5 2026 retirement date but still listed on the official pricing page as "retired, except on Bedrock and Google Cloud" this run. File entry retained; no action required, but check whether it is fully removed from the official page in the next audit.

  3. AWS Bedrock "Claude 3.5 Sonnet (Public Extended Access)" pricing — Confirmed real in the Aug 4 2026 audit (see provider-sources-and-price-keys.md) but not representable in Langfuse's schema because it matches by model-ID string only. No file change should be made for this; treat it as a permanent, documented limitation rather than something to re-investigate each run.

  4. Legacy Claude 3.x / 3.5 / 3.7 models not on the current pricing page — Not re-verified this run (claude-3.7-sonnet-20250219, claude-3.5-sonnet-20241022, claude-3-5-sonnet-20240620, claude-3-opus-20240229, claude-3-sonnet-20240229, claude-3-haiku-20240307). Existing prices retained. Low priority since these are retired/legacy.

  5. gemini-3.1-flash-lite-preview and gemini-3-pro-preview — Still not separately listed on the official AI Studio pricing page. Existing prices retained without fresh confirmation. Re-verify if these move from preview to GA or gain their own pricing row.

  6. OpenAI "cache writes" is still gpt-5.6-family-only — As of this audit (re-confirmed via the full standard-pricing-table dump), the distinct 1.25x-of-input cache-write billing dimension applies only to gpt-5.6-sol, gpt-5.6-terra, and gpt-5.6-luna. Every other checked OpenAI model still shows "—" for cache writes. Future audits should re-check this column whenever a new OpenAI reasoning model is added.

  7. Base vs. fine-tuning legacy pricing confusion is a real historical bug class — The davinci-002/babbage-002 fix in this audit (Aug 7 2026) revealed that OpenAI's pricing page lists the same base model name in two different tables: the "Standard" table (bare inference pricing) and a "Fine-tuning" table (which additionally shows a Training cost plus a different, higher Input/Output inference rate for legacy fine-tuned models). A pricing-file entry whose matchPattern matches only the bare model ID (no ft: prefix) must use the Standard table's price, never the Fine-tuning table's price, even though both rows share the exact same model name in the source page. Future audits touching any OpenAI base model that also has a legacy fine-tuning tier (currently: gpt-3.5-turbo, davinci-002, babbage-002, and the fine-tunable snapshots gpt-4.1-2025-04-14, gpt-4.1-mini-2025-04-14, gpt-4.1-nano-2025-04-14, gpt-4o-2024-08-06, gpt-4o-mini-2024-07-18, o4-mini-2025-04-16) should double-check which table a fetched number came from before applying it to the bare (non-ft:) entry.

  8. Legacy/embedding/base-completion catalog tail not covered this run — Entries such as text-ada-001, text-babbage-001, text-curie-001, text-davinci-00{1,2,3}, text-embedding-*, the Vertex *-bison*/*-gecko* PaLM family, claude-1.x/claude-2.x, and gemini-1.0-*/gemini-pro were not re-fetched this run (consistent with prior audits) since they are long-retired and out of the "flagship text/chat/reasoning model" scope. If a future task explicitly asks to audit embeddings or PaLM-era models, treat this as unverified starting ground, not confirmed.