Weekly roundup: GPT-5.6 general release, Gemini 3.5 Pro launch, Grok 4.5, Apple-Alibaba Qwen approval, Ollama’s $65M raise
Weekly roundup: GPT-5.6 general release, Gemini 3.5 Pro launch, Grok 4.5, Apple-Alibaba Qwen approval, Ollama’s $65M raise

Weekly roundup: GPT-5.6 general release, Gemini 3.5 Pro launch, Grok 4.5, Apple-Alibaba Qwen approval, Ollama’s $65M raise

Busy nine days on the frontier, so here's a consolidated summary with sources.

Model launches: OpenAI released GPT-5.6 broadly on July 9 — three variants (Sol $5/$30, Terra $2.50/$15, Luna $1/$6 per 1M tokens), 1.05M context (Axios, OpenAI). xAI shipped Grok 4.5 on July 8 at $2/$6, positioned as a workhorse for routine knowledge work (TechCrunch, Axios). Google's Gemini 3.5 Pro reached GA today, July 17, after a full architectural rebuild — 2M token context, Deep Think reasoning layer, ~$1.25/$10 (TechTimes, BigGo).

Distribution news: Apple's iOS 27 public beta opened Siri AI to the general public (TechCrunch, 9to5Mac), and China's CAC approved Apple Intelligence for launch there, running on Alibaba's Qwen (TechCrunch, CNBC). Microsoft added Claude as a model option in Copilot Chat as part of 40+ July updates.

Open source: Ollama raised a $65M Series B — 8.9M monthly developers, claims presence in 85% of the Fortune 500 (TechCrunch). Hugging Face reported Chinese open-weight models took 41% of downloads this spring, surpassing US models. Mistral put a new open-weight frontier model into early access.

Less rosy: Grok Build was caught uploading users' repos to xAI-controlled cloud storage; xAI open-sourced the CLI and promised data deletion (The Register). Hugging Face disclosed a production intrusion executed end-to-end by an autonomous AI agent. And China's new anthropomorphic AI rules took effect July 15 — ByteDance and Alibaba shut down user-created agents entirely.

My take as someone building on top of these APIs: the pricing collapse is the real story. Luna at $1/$6 and Grok at $2/$6 means capability that cost 15-30x more two years ago is now nearly free at the margin. In practice that changes architecture decisions — pipelines I run through cheap fast models today would've needed careful cost engineering last year. But the Grok Build incident is a reminder that when tokens get this cheap, the vendor's data practices become the actual differentiator. I'd also flag the China agent shutdown as underrated: millions of users lost working agents overnight because of a regulatory change. If you're building agentic products, jurisdictional platform risk deserves a line in your risk register.

What's everyone else seeing on the cost side — has anyone re-benchmarked their pipelines against Luna or Grok 4.5 yet?

submitted by /u/ksraj1001
[link] [comments]