Google DeepMind: Their new model outperforms 3.6 Flash on key coding tasks such as debugging, designs more functional web layouts with fewer prompts, and improves reasoning and accuracy for real‑world business workflows. Google AI: Announces Gemini 3.7 Flash as their most intelligent coding and agent workhorse, now powering Gemini Spark in the Gemini app to help users tackle long to‑do lists. Google AI Developers: Shows Gemini 3.7 Flash in Antigravity turning a single architecture spec into native, production‑ready code for Flutter, SwiftUI, Jetpack Compose, React Native and NativeScript. DeepSeek: Launches DeepSeek‑V4‑Pro with major agent upgrades, flexible reasoning effort for simple to complex tasks, and native OpenAI‑compatible API support; also rolls out new API pricing with off‑peak rates 50 % lower than peak starting Aug 16, 2026.
GLM‑5.2 is Z.ai’s newest open‑source model designed for long‑horizon coding tasks. It supports a solid 1 million‑token context, introduces flexible “effort” levels for balancing speed and capability, and uses the IndexShare architecture to cut per‑token FLOPs by 2.9×. Benchmarks show it outperforms its predecessor (GLM‑5.1) and ranks as the strongest open‑source coding model, closing the gap to leading closed‑source systems.
GitHub announces that Agentic Workflows are now in public preview, offering intelligent automations with guardrails, observability, and cost controls. GitHub Changelog notes the workflows can now use the built‑in GITHUB_TOKEN instead of personal access tokens, improving security and simplicity. OpenCode reports that DeepSeek V4 Pro, Fable 5, and the North Mini Code model (256K context, fully open source) are now available on its platform. OpenRouter launches an Activity explorer that shows real‑time spending, token usage, cache hit rates, agents, and trends for models like Fable. RyanLee shares that his high‑performance MSA kernel library is open‑source and that the M3 weights are expected to be released on Friday, with a link to the GitHub paper.
OpenCode reports that DeepSeek V4 Flash is now available in OpenCode Zen. The new model can be accessed through the platform’s environment. This addition expands the AI tools developers, students, and enthusiasts can use.
OpenRouter now supports "apply_patch," a server tool that lets models propose file edits using V4A diffs through the Responses API. The model generates a patch, and OpenRouter validates the diff syntax server-side. This feature allows for more efficient and accurate file editing. xAI has released grok-build-0.1 in public beta via the xAI API. This model powers the Grok Build CLI and excels at agentic coding, priced at $1/m input and $2/m output. Google AI has released an episode of Release Notes featuring the architects of Gemini, including @JeffDean, @koraykv, @OriolVinyalsML, and @NoamShazeer. They discuss their journey and the people behind the model. LangChain has released LangSmith LLM Gateway, which enforces spend limits and redacts PII before requests reach the model. They also announced Deep Agents v0.6, which makes harness profiles a first-class abstraction, allowing for production-grade performance at lower costs. NVIDIA has announced a new era of PC, but the details are unclear. OpenAI has launched Rosalind Biodefense to help trusted builders develop new biodefense and pandemic preparedness capabilities. They are also expanding trusted access to GPT-Rosalind for select U.S. government and allied partners.