02-347-7730  |  Saeree ERP - Complete ERP System for Thai Businesses Contact Us

AI News Roundup — July 2026: GPT-5.6, Claude Opus 5, Gemini 3.6 and Thailand's TH-AI Passport

AI News Roundup — July 2026: GPT-5.6, Claude Opus 5, Gemini 3.6 and Thailand's TH-AI Passport
  • 19
  • August

"AI News Roundup — July 2026" — the short answer is this was the month all three frontier labs shipped in the same two weeks: OpenAI made GPT-5.6 generally available with ChatGPT Work on 9 July, Google launched Gemini 3.6 Flash on the 21st, and Anthropic closed the month with Claude Opus 5 on the 24th. Thailand also switched on its TH-AI Passport programme on 1 July. This roundup collects every announcement that mattered — models, pricing, deals, capital markets and Thai policy — and ends with what an enterprise should actually do about it.

In one line: July 2026 saw every frontier lab replace its flagship inside a fifteen-day window, mid-tier API pricing fall sharply, and Thailand open access to leading models for the first time through TH-AI Passport.

Timeline: what actually happened in July 2026

If June was the month of disruption (see our June 2026 AI roundup), July was the month of shipping — three flagship models replaced inside two and a half weeks, followed almost immediately by aggressive mid-tier price cuts.

DatePlayerEvent
1 JulAnthropicClaude Fable 5 restored after the U.S. Commerce Department export-control directive was lifted on 30 June
1 JulThailand (MDES)TH-AI Passport reaches its go-live date — access to 31 models from 14 providers
8 JulxAIGrok 4.5 ships with a 500,000-token context window — the first release since the company went public
9 JulOpenAIGPT-5.6 goes generally available (Sol / Terra / Luna) alongside ChatGPT Work
10 JulSK HynixRaises $26.5bn in a U.S. IPO — the largest listing ever by a foreign company on a U.S. exchange
10 JulAppleSues OpenAI, io Products and two former employees in the Northern District of California over hardware trade secrets
21 JulGoogleGemini 3.6 Flash launches with 3.5 Flash-Lite and 3.5 Flash Cyber
23 JulAnthropicClaude voice mode updated so users can pick Opus, Sonnet or Haiku
24 JulAnthropicClaude Opus 5 launches — same price as Opus 4.8, materially stronger
24 JulMetaMuse Spark 1.2 ships as Meta's new frontier model, a shift away from the Llama line
27 JulCognizant + AnthropicExpanded partnership to bring Claude to Cognizant's enterprise clients
30 JulOpenAICuts GPT-5.6 Luna pricing by 80% and Terra by 20% — three weeks after GA

1. OpenAI — three GPT-5.6 tiers and an agent that finishes the job

GPT-5.6 became generally available on 9 July after a limited preview with a small group of approved partners that began on 26 June. The structural change from previous releases is that OpenAI split the family into three named tiers instead of shipping one model with adjustable effort — full detail in our piece on GPT-5.6: Sol, Terra and Luna.

TierPrice, input / output per 1M tokensChange on 30 July
Sol (flagship)$5 / $30unchanged
Terra (balanced)$2 / $12down 20%
Luna (cost-efficient)$0.20 / $1.20down 80%

Arguably the bigger story was ChatGPT Work, launched the same day — an agentic platform built to complete whole jobs rather than answer questions. In the launch demo, OpenAI's finance team issued a single instruction and watched the system pull data from Slack, run a variance analysis, update an Excel forecast model, generate a slide deck, and publish an interactive dashboard as a shareable site.

Worth noting: cutting Luna by 80% within three weeks of GA says something about where the competition is heading. The battleground is shifting to cost-per-job on high-volume, low-complexity work — not the benchmark score of the top-end model.

2. Anthropic — Fable 5 returns, Opus 5 closes the month

Anthropic's July opened with good news: Claude Fable 5 came back on 1 July after the export-control directive that forced its suspension was lifted on 30 June. Initially Fable 5 counted toward up to 50% of weekly usage limits, with that window extended repeatedly until permanent availability rules were announced on 20 July.

Then on 24 July, Claude Opus 5 replaced Opus 4.8 at the Opus tier. The detail enterprises noticed was that pricing did not move — still $5 / $25 per million tokens — while capability rose across the board, with a 1M-token context window, up to 128K output tokens, and a five-level effort setting. The mid-tier Claude Sonnet 5 continues to carry high-volume work at a lower price point.

Two smaller items round out the month and point at the direction of travel: on 23 July Claude's voice mode gained model selection (Opus, Sonnet, Haiku), and on 27 July Cognizant expanded its partnership to bring Claude into its own enterprise client base. The contest is moving from "who scores higher" to "who gets embedded in real business systems first".

3. Google — a cheaper Flash tier and a full push into agents

On 21 July Google shipped Gemini 3.6 Flash alongside 3.5 Flash-Lite and 3.5 Flash Cyber, positioning the Flash tier explicitly for speed, cost and high-volume agentic work rather than maximum reasoning depth.

AttributeGemini 3.6 Flash
Price (per 1M tokens)$1.50 input / $7.50 output
Context / output1,000,000 / 64,000 tokens
Accepted inputsText, image, video, audio, PDF
Throughput~280 tokens/second
Token efficiency~17% fewer output tokens than its predecessor — fewer reasoning steps and tool calls per workflow
Knowledge cutoffMarch 2026

Beyond the models, Google launched Gemini Robotics ER 2 for embodied reasoning and made AlphaEvolve — an agent that optimises code — generally available to Google Cloud customers.

4. The rest of the field

  • Grok 4.5 (8 July) — 500,000-token context, roughly $2 per million input tokens, around 80 tokens/second; xAI's first release after going public.
  • Meta Muse Spark 1.2 (24 July) — Meta's new frontier model, a deliberate move away from continuing the Llama line.
  • Mistral Leanstral 1.5 — goes beyond code generation to produce mathematical proof, via Lean 4, that software behaves as specified. Worth watching for safety-critical systems.

Capital markets and courts — the month the returns question surfaced

On 10 July SK Hynix raised $26.5bn by selling 177.9 million American depositary shares at $149 each, making it the largest listing by a foreign company in U.S. history — past Alibaba's 2014 record. The demand driver was memory for AI workloads, plainly.

In the same month, though, the market started asking harder questions. Investors began probing the financing structures behind the AI infrastructure build-out after reports of very large financing guarantees for data-centre projects. And on 10 July Apple filed suit against OpenAI, io Products and two former Apple employees over hardware trade secrets.

Risk to plan for: record capital inflows and rising doubt about returns arrived in the same month. If you are budgeting AI spend for next year, design for provider portability — do not wire your whole architecture to one vendor's API with no exit path.

Thailand — TH-AI Passport goes live, and an AI bill is coming

1 July 2026 was the go-live date for TH-AI Passport, which opens access to 31 AI models from 14 providers, including OpenAI, Google, Claude, DeepSeek and Meta AI.

The rationale is an access gap. Figures cited for the programme put Thailand's AI technology access rate at roughly 10.67%, against about 23.5% in Vietnam — more than double.

On the legal side, the government is targeting September 2026 to move a draft AI act into the legislative process, focused specifically on high-risk uses rather than blanket regulation, with an AI Governance Practice Center acting as a sandbox where public agencies, private firms and academia can apply international AI governance principles in practice. That direction lines up with what we covered in Digital Government 2026, and it matters for any organisation selling into Thai public-sector procurement.

So what should an enterprise actually do?

A busy news month does not mean everything needs changing. This table separates what deserves action now from what can wait.

TopicDo nowDo not rush
Falling pricesReview high-volume tasks (summarising, classification, validation) for a move to a cheaper tierDowngrade accuracy-critical work just because the price dropped
New flagships everywhereBuild your own eval set — 20-30 real cases — and re-run it on every releaseSwap production models on launch-day headlines without measuring
Agent platformsPick one narrow, well-scoped job, run it end to end, measure the hours actually savedLet an agent write to financial systems before a human approval gate exists
TH-AI PassportCheck whether your organisation is eligible and on what termsCancel existing contracts to wait on this channel alone
The draft AI actInventory where AI is already used, who owns it, and what data leaves your perimeterBuild a full compliance programme before the actual text is published

The view from an ERP team

What gets clearer every month is that model capability is no longer the bottleneck in enterprise AI projects — data and permissions are. The best model on the market still cannot answer "how much budget is left in this category" if the numbers live across several versions of a spreadsheet and nothing can say who is allowed to see what.

To be straight about our own product: the Saeree ERP AI Assistant is still in training and is not something we ship to customers today. What the system does do now is the groundwork underneath it — budget, procurement, inventory and accounting data in one database, role-based access control, and a complete approval trail. Those are preconditions for connecting AI safely, not optional extras. For the broader picture of how agents enter ERP work, see What is Agentic AI and our field notes on AI agents as coworkers.

Conclusion

July 2026 settled two open questions. First, the model race has not slowed — all three flagships turned over inside a fortnight. Second, the fight is migrating away from benchmark scores toward cost-per-job and distribution into real business systems, visible in the 80% Luna price cut, the launch of ChatGPT Work, and Anthropic's choice to partner with a firm that already owns enterprise relationships.

For organisations operating in Thailand there is an extra layer this month that previous ones lacked: policy became concrete. TH-AI Passport reached its go-live date and a draft AI act is targeted for the legislature by September. The question for the rest of this year is no longer only "which model do we use" but "who can we answer to about how we use it".

A month of AI news takes ten minutes to read. The harder question — is our own data ready for an AI to work with — is not answered anywhere in it.

- The Saeree ERP team

References

Connecting AI to real operations starts with data that is ready

Talk to the Saeree ERP team about consolidating budget, procurement, inventory and accounting data into one database with role-based access — the groundwork AI needs before it can be useful.

Request a Free Consultation

Tel 02-347-7730 | sale@grandlinux.com

Saeree ERP Author

About the Author

Sureeraya Limpaibul

Managing Director, Grand Linux Solution Co., Ltd. & Founder of Saeree ERP — providing end-to-end ERP advisory and services.