Model Signal logo Model Signal Fast, verified AI updates

Tag

Claude

Published stories tagged with Claude.

AI Models 3 min read

Latest from X - 2026-06-30 to 2026-07-02

Google Research: Introduced TabFM, a foundation model for tabular data classification and regression that can generate high‑quality predictions on unseen tables in a single forward pass. ClaudeDevs: Raised Claude Platform API rate limits for all users, removed spend‑based tiers, and gave the latest Sonnet and Haiku models five‑times higher limits at the top tier. Google AI: Launched two workflow updates - one model for ultra‑fast image generation and another to instantly animate those images, both at a fraction of the usual cost, highlighted by Nano Banana 2 Lite. GitHub Changelog: Announced that Gemini 2.5 Pro and Gemini 3 Flash will be deprecated in all Copilot experiences on July 31 2026, urging migration to Gemini 3.1 Pro or Gemini 3.5 Flash beforehand. Z.ai: Rolled out ZCode, the official development environment for GLM‑5.2, offering 1.5× usage quota for GLM Coding Plan subscribers, BYOK support, and cross‑platform availability on macOS, Windows, and Linux. GitHub: Made Kimi Moonshot’s Kimi K2.7 Code generally available in GitHub Copilot as the first open‑weight model selectable in the model picker, noting lower cost with performance comparable to top‑tier models.

AI Models 4 min read

Previewing GPT‑5.6 Sol, Terra, and Luna

OpenAI has begun a limited preview of the GPT‑5.6 family: **Sol** (flagship), **Terra** (balanced, 2× cheaper than GPT‑5.5), and **Luna** (fast, low‑cost). Sol introduces a new “max” reasoning mode and an “ultra” mode that uses sub‑agents. Early benchmarks show state‑of‑the‑art performance on coding (Terminal‑Bench 2.1), biology (GeneBench v1), and cybersecurity (ExploitBench, ExploitGym). The models ship with OpenAI’s most robust safety stack to date, but the preview may block or delay some requests.

AI Models 4 min read

GLM‑5.2 Brings 1M‑Token Context to Coding Agents

GLM‑5.2 is Z.ai’s newest open‑source model designed for long‑horizon coding tasks. It supports a solid 1 million‑token context, introduces flexible “effort” levels for balancing speed and capability, and uses the IndexShare architecture to cut per‑token FLOPs by 2.9×. Benchmarks show it outperforms its predecessor (GLM‑5.1) and ranks as the strongest open‑source coding model, closing the gap to leading closed‑source systems.

Coding 2 min read

Latest from X - 2026-06-12 to 2026-06-13

OpenCode: Kimi 2.7 Code is now available in Go with image support, optimized for coding and priced similarly to 2.6. Kimi.ai: Builders using the Kimi K2.7 Code API can earn 20%–30% extra quota by topping up $100+ before July 2, with one bonus per account. ollama: The Kimi‑K2.7‑Code model is now hosted on Ollama’s US cloud on NVIDIA B300 GPUs, keeping data private and never used for training; try it with “ollama launch claude --model”. Z.ai: GLM‑5.2, the new flagship model, is now accessible to all GLM Coding Plan users—including Lite, Pro, Max, and Team tiers.

AI Models 3 min read

Latest from X - 2026-06-10 to 2026-06-12

Kimi.ai: Released the open‑source Kimi‑K2.7‑Code model, reporting +21.8% on Kimi Code Bench v2, +11.0% on Program Bench, +31.5% on MLS Bench Lite, and 30% lower reasoning overthinking. Google Research: Launched Gemini‑SQL2, a text‑to‑SQL system built on Gemini 3.1 Pro that hits state‑of‑the‑art scores on the BIRD benchmark. OpenAI: Added saved Codex rate‑limit resets for Go, Plus, Pro, and Business tiers (one free reset) and a two‑week invite program letting Plus/Pro users earn extra resets by inviting friends. Claude: Made dynamic workflows in Claude Code generally available, letting the model orchestrate parallel sub‑agents for complex tasks like codebase‑wide bug hunts and verify work before returning results.

AI Models 3 min read

Claude Fable 5: A New Leap for Autonomous Coding

Claude Fable 5 is Anthropic’s latest Mythos‑class model, released with safety classifiers that route risky queries to Claude Opus 4.8. It claims state‑of‑the‑art performance on coding benchmarks, dramatically faster codebase migrations, and higher token efficiency than previous Claude models.

Coding 3 min read

Latest from X - 2026-06-08 to 2026-06-09

Claude announced Claude Fable 5, a Mythos‑class model deemed safe for general use and now available everywhere, while Claude Mythos 5 stays limited to Glasswing partners. OpenRouter, Cursor, Visual Studio Code and Devin Desktop all reported that Claude Fable 5 is now live on their platforms, with Cursor noting a 72.9% score on CursorBench, 8 points above the previous best. Google DeepMind highlighted its 3.5 Live Translate, which streams speech into over 70 languages while preserving tone, pace and pitch for natural conversation. Kimi.ai introduced Kimi Work, a local desktop AI agent that can run up to 300 parallel agents and, via a WebBridge extension, navigate browsers to search, scroll and interact.

Coding 3 min read

Latest from X - 2026-06-04

NVIDIA AI announced the release of Nemotron 3 Ultra, a 550 B MoE model that speeds inference fivefold, lowers agentic task costs up to 30 % and excels at coding, deep research, and long‑horizon planning. OpenCode noted that Nemotron 3 Ultra is now free with 1 M context and fully open source. Ollama said the model is available on its cloud platform, offering launch commands for Claude, Hermes and OpenClaw. OpenAI introduced a new memory system for ChatGPT that automatically tracks important details, doubles memory capacity for Plus and Pro users in the US, and lets users review and steer remembered content via a summary. Cursor added an interactive context‑usage report in its canvas, breaking down token distribution across prompts, tools, rules and skills.

AI Models 6 min read

Latest from X - 2026-06-01 to 2026-06-02

Qwen: introduces Qwen3.7-Plus, a multimodal agent model that unifies vision and language with both GUI and CLI operation and serves as a coding and productivity assistant. OpenAI: frontier models and Codex are now generally available on AWS via Amazon Bedrock, extending enterprise security, compliance, and governance workflows. xAI: Composer 2.5 is now inside Grok Build, described as a fast, highly intelligent model for long‑running tasks and complex instructions. LangChain: highlights Fleet for secure agent access to private resources and adds LangSmith LLM Gateway spend limits that return a 402 error when caps are hit. Google Antigravity: is becoming a scientific workbench with a Science Skills bundle that runs complex workflows like protein analysis using Alpha* models and dozens of databases; Google Gemma: releases the first gemma‑skills iteration, enabling agents to build with Gemma, use MTP for speed, pick model size, and locate up‑to‑date resources. ClaudeDevs: resets 5‑hour and weekly rate limits for Pro/Max plans and fixes excessive parallel subagents; Cursor: raises usage limits for Teams and adds a Premium seat with 5× usage at 3× cost; Visual Studio Code: demos orchestrating agents via the VS Code Agents window; NVIDIA: adds real‑time AI media tools including Synthetic Video Detector (up to 92% accuracy, 22 ms latency), RTX Video Super Resolution and Frame Generation; Vercel: enables remote execution of Conductor’s parallel coding agents on fast Sandboxes; Perplexity: launches Search as Code, a new architecture that writes Python to call its search stack directly, now default in the Perplexity Agent API.

AI Models 2 min read

Claude Code Releases

Claude Code has released version 2.1.152, which includes several new features and bug fixes. The update applies review findings to the working tree after a review, simplifies code, and improves the user experience.