Close Menu
  • Home
  • AI News
  • AI Startups
  • Deep Learning
  • Interviews
  • Machine-Learning
  • Robotics

Subscribe to Updates

Get the latest creative news from FooBar about art, design and business.

What's Hot

Constructing Protected Structure for Enterprise AI – Unite.AI

July 31, 2026

A complete-company AI work platform for regulated industries

July 31, 2026

ValidSoft Unveils Full AI Belief Intelligence Stack to Safe People & AI Brokers from Identification Verification to Execution

July 31, 2026
Facebook X (Twitter) Instagram
Smart Homez™
Facebook X (Twitter) Instagram Pinterest YouTube LinkedIn TikTok
SUBSCRIBE
  • Home
  • AI News
  • AI Startups
  • Deep Learning
  • Interviews
  • Machine-Learning
  • Robotics
Smart Homez™
Home»Robotics»OpenAI Cuts API Costs on Its Two Cheaper GPT-5.6 Tiers – Unite.AI
Robotics

OpenAI Cuts API Costs on Its Two Cheaper GPT-5.6 Tiers – Unite.AI

Editorial TeamBy Editorial TeamJuly 30, 2026Updated:July 31, 2026No Comments5 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr Reddit WhatsApp Email
OpenAI Cuts API Costs on Its Two Cheaper GPT-5.6 Tiers – Unite.AI
Share
Facebook Twitter LinkedIn Pinterest WhatsApp Email



OpenAI lowered the API value of its two lower-cost GPT-5.6 fashions on July 30, 2026, chopping the most affordable tier by 80% and the mid-tier by 20% whereas leaving its flagship untouched. The change is logged within the firm’s personal API changelog and is already stay on the printed price card.

Per million enter and output tokens, the usual charges now learn:

  • GPT-5.6 Luna: 20 cents and $1.20, down from $1 and $6
  • GPT-5.6 Terra: $2 and $12, down from $2.50 and $15
  • GPT-5.6 Sol: $5 and $30, unchanged, matching the speed its predecessor GPT-5.5 nonetheless carries

All three tiers reached basic availability on July 9, 2026 on the increased costs, which places the repricing three weeks into the household’s industrial life.

The cuts run by means of each service tier on the sheet, not simply the headline price. Batch and Flex processing, each half the usual value, now put Luna at 10 cents enter and 60 cents output. Cached enter reads, discounted 90%, drop to 2 cents per million tokens on Luna and 20 cents on Terra. Lengthy-context requests, which invoice at double the enter price and 1.5 instances output, land at 40 cents and $1.80 for Luna. Patrons going by means of Amazon (AMZN ) Bedrock are billed by AWS, and OpenAI notes these charges can differ from its personal.

The cheaper tiers are the place high-volume manufacturing site visitors lives: classification, extraction, request routing, first-pass drafting, and the lengthy agent loops the place one consumer instruction can set off dozens of mannequin calls earlier than it returns a solution. A five-fold lower on the tier absorbing that quantity adjustments the arithmetic on which workloads are price automating in any respect.

The place the cheaper tiers sit in opposition to Claude

At 20 cents in and $1.20 out, Luna undercuts Anthropic’s least expensive printed mannequin, Haiku 4.5, by an element of 5 on enter and roughly 4 on output, in response to Anthropic’s pricing web page. Terra’s new price sits under the $3 and $15 that Claude Sonnet 5 is scheduled to cost as soon as its introductory price of $2 and $10 lapses on August 31, 2026. On the prime of each lineups, Sol nonetheless prices extra on output than Opus 5, which Anthropic costs at $5 and $25.

Precedence processing turns into Quick mode

The identical changelog entry retires Precedence Processing and replaces it with Quick mode. For Sol, OpenAI says Quick mode runs as much as 2.5 instances customary velocity at twice the value, and the swap is backward suitable: requests already tagged for precedence path to Quick mode and not using a code change. Quick-mode charges are $10 and $60 for Sol, $4 and $24 for Terra, and 40 cents and $2.40 for Luna.

Anthropic sells the identical product underneath the identical title and the identical phrases — quick mode for Opus 5, as much as 2.5 instances quicker at twice customary pricing. The 2 price playing cards now converge on region-pinned inference as effectively. OpenAI expenses a ten% uplift on fashions launched on or after March 5, 2026 when a buyer requires knowledge residency; Anthropic payments US-only inference at 1.1 instances its customary price.

What made the cheaper tiers cheaper

OpenAI printed its accounting a day earlier than the lower. In a July 29, 2026 engineering submit, 5 members of its technical employees described optimizations throughout inference and the agent harness behind Codex and ChatGPT Work. Sol, working inside Codex, rewrote the corporate’s manufacturing GPU kernels; mixed with broader kernel work, OpenAI says that lower end-to-end serving prices by 20%. Sol additionally redesigned its personal speculative-decoding draft mannequin throughout a whole lot of experiments, which the corporate credit with elevating token-generation effectivity by greater than 15%.

These are OpenAI’s personal figures for its personal stack. The submit closed on a dedication to cross “under-the-hood enhancements again to our customers and prospects within the type of extra broadly accessible, cost-efficient intelligence.” The speed card adopted a day later.

The harness work factors on the identical price middle the value lower does. OpenAI caps instrument output at 10,000 tokens by default and retains model-visible historical past append-only, so an agent loop resending its directions and gear definitions at each step hits the immediate cache as an alternative of paying full enter charges. On a job that takes 30 mannequin requests, that’s 30 probabilities to keep away from recomputing the identical prefix.

The cuts land whereas patrons are auditing inference spend and widening which groups contact the fashions: OpenAI’s personal analysis discovered employees utilizing ChatGPT effectively past their job titles. Amazon’s engineering group moved to cap AI spending after price overruns the identical day, and OpenAI shipped exhausting spend limits for API organizations and tasks on July 22, 2026, letting directors set a month-to-month ceiling that begins returning errors as soon as tracked spend reaches it.

For a workforce already routing bulk work to Luna, the identical calls now price a fifth of what they did, with a ceiling they’ll set within the dashboard. With Sol’s value unchanged, the stay choice stays the place OpenAI has put it for the reason that household shipped: which tier every request requires.



Supply hyperlink

Editorial Team
  • Website

Related Posts

Constructing Protected Structure for Enterprise AI – Unite.AI

July 31, 2026

Arun Hiremath, Chief Enterprise Officer and Co-Founding father of EvoluteIQ – Interview Collection – Unite.AI

July 30, 2026

Xsight Labs Raises $300M for Programmable AI Community Silicon – Unite.AI

July 30, 2026
Misa
Trending
Robotics

Constructing Protected Structure for Enterprise AI – Unite.AI

By Editorial TeamJuly 31, 20260

Half one ended with a declare: enterprise AI will succeed when establishments learn to construct…

A complete-company AI work platform for regulated industries

July 31, 2026

ValidSoft Unveils Full AI Belief Intelligence Stack to Safe People & AI Brokers from Identification Verification to Execution

July 31, 2026

The Full AI Agent Engineer Expertise Stack You Want in 2026

July 31, 2026
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo
Our Picks

Constructing Protected Structure for Enterprise AI – Unite.AI

July 31, 2026

A complete-company AI work platform for regulated industries

July 31, 2026

ValidSoft Unveils Full AI Belief Intelligence Stack to Safe People & AI Brokers from Identification Verification to Execution

July 31, 2026

The Full AI Agent Engineer Expertise Stack You Want in 2026

July 31, 2026

Subscribe to Updates

Get the latest creative news from SmartMag about art & design.

The Ai Today™ Magazine is the first in the middle east that gives the latest developments and innovations in the field of AI. We provide in-depth articles and analysis on the latest research and technologies in AI, as well as interviews with experts and thought leaders in the field. In addition, The Ai Today™ Magazine provides a platform for researchers and practitioners to share their work and ideas with a wider audience, help readers stay informed and engaged with the latest developments in the field, and provide valuable insights and perspectives on the future of AI.

Our Picks

Constructing Protected Structure for Enterprise AI – Unite.AI

July 31, 2026

A complete-company AI work platform for regulated industries

July 31, 2026

ValidSoft Unveils Full AI Belief Intelligence Stack to Safe People & AI Brokers from Identification Verification to Execution

July 31, 2026
Trending

The Full AI Agent Engineer Expertise Stack You Want in 2026

July 31, 2026

OpenAI Cuts API Costs on Its Two Cheaper GPT-5.6 Tiers – Unite.AI

July 30, 2026

Arun Hiremath, Chief Enterprise Officer and Co-Founding father of EvoluteIQ – Interview Collection – Unite.AI

July 30, 2026
Facebook X (Twitter) Instagram YouTube LinkedIn TikTok
  • About Us
  • Advertising Solutions
  • Privacy Policy
  • Terms
  • Podcast
Copyright © The Ai Today™ , All right reserved.

Type above and press Enter to search. Press Esc to cancel.