Z.ai: announced GLM‑5.2, a frontier‑intelligence model with open weights, a 1 M‑token context window and two reasoning effort levels for coding and agentic tasks. OpenCode: reported that GLM‑5.2 has risen to 6th place on their leaderboard within three days of release. ollama: highlighted GLM‑5.2 as the strongest open‑source coding model yet, now available on Ollama’s US cloud powered by NVIDIA AI Blackwell GPUs. GitHub Changelog: warned that Opus 4.6 (fast) will be deprecated in all Copilot experiences on June 29 2026 and urged migration to Opus 4.8 (fast). Cursor: introduced a new /automate skill that lets agents set up automations from plain‑language task descriptions, configuring triggers, instructions and tools automatically.
OpenCode: Kimi 2.7 Code is now available in Go with image support, optimized for coding and priced similarly to 2.6. Kimi.ai: Builders using the Kimi K2.7 Code API can earn 20%–30% extra quota by topping up $100+ before July 2, with one bonus per account. ollama: The Kimi‑K2.7‑Code model is now hosted on Ollama’s US cloud on NVIDIA B300 GPUs, keeping data private and never used for training; try it with “ollama launch claude --model”. Z.ai: GLM‑5.2, the new flagship model, is now accessible to all GLM Coding Plan users—including Lite, Pro, Max, and Team tiers.
GitHub announces that Agentic Workflows are now in public preview, offering intelligent automations with guardrails, observability, and cost controls. GitHub Changelog notes the workflows can now use the built‑in GITHUB_TOKEN instead of personal access tokens, improving security and simplicity. OpenCode reports that DeepSeek V4 Pro, Fable 5, and the North Mini Code model (256K context, fully open source) are now available on its platform. OpenRouter launches an Activity explorer that shows real‑time spending, token usage, cache hit rates, agents, and trends for models like Fable. RyanLee shares that his high‑performance MSA kernel library is open‑source and that the M3 weights are expected to be released on Friday, with a link to the GitHub paper.
NVIDIA AI announced the release of Nemotron 3 Ultra, a 550 B MoE model that speeds inference fivefold, lowers agentic task costs up to 30 % and excels at coding, deep research, and long‑horizon planning. OpenCode noted that Nemotron 3 Ultra is now free with 1 M context and fully open source. Ollama said the model is available on its cloud platform, offering launch commands for Claude, Hermes and OpenClaw. OpenAI introduced a new memory system for ChatGPT that automatically tracks important details, doubles memory capacity for Plus and Pro users in the US, and lets users review and steer remembered content via a summary. Cursor added an interactive context‑usage report in its canvas, breaking down token distribution across prompts, tools, rules and skills.
OpenCode reports that DeepSeek V4 Flash is now available in OpenCode Zen. The new model can be accessed through the platform’s environment. This addition expands the AI tools developers, students, and enthusiasts can use.