Model Signal logo Model Signal Fast, verified AI updates

Tag

Kimi

Published stories tagged with Kimi.

AI Models 3 min read

Latest from X - 2026-06-30 to 2026-07-02

Google Research: Introduced TabFM, a foundation model for tabular data classification and regression that can generate high‑quality predictions on unseen tables in a single forward pass. ClaudeDevs: Raised Claude Platform API rate limits for all users, removed spend‑based tiers, and gave the latest Sonnet and Haiku models five‑times higher limits at the top tier. Google AI: Launched two workflow updates - one model for ultra‑fast image generation and another to instantly animate those images, both at a fraction of the usual cost, highlighted by Nano Banana 2 Lite. GitHub Changelog: Announced that Gemini 2.5 Pro and Gemini 3 Flash will be deprecated in all Copilot experiences on July 31 2026, urging migration to Gemini 3.1 Pro or Gemini 3.5 Flash beforehand. Z.ai: Rolled out ZCode, the official development environment for GLM‑5.2, offering 1.5× usage quota for GLM Coding Plan subscribers, BYOK support, and cross‑platform availability on macOS, Windows, and Linux. GitHub: Made Kimi Moonshot’s Kimi K2.7 Code generally available in GitHub Copilot as the first open‑weight model selectable in the model picker, noting lower cost with performance comparable to top‑tier models.

Coding 2 min read

Latest from X - 2026-06-12 to 2026-06-13

OpenCode: Kimi 2.7 Code is now available in Go with image support, optimized for coding and priced similarly to 2.6. Kimi.ai: Builders using the Kimi K2.7 Code API can earn 20%–30% extra quota by topping up $100+ before July 2, with one bonus per account. ollama: The Kimi‑K2.7‑Code model is now hosted on Ollama’s US cloud on NVIDIA B300 GPUs, keeping data private and never used for training; try it with “ollama launch claude --model”. Z.ai: GLM‑5.2, the new flagship model, is now accessible to all GLM Coding Plan users—including Lite, Pro, Max, and Team tiers.

AI Models 3 min read

Latest from X - 2026-06-10 to 2026-06-12

Kimi.ai: Released the open‑source Kimi‑K2.7‑Code model, reporting +21.8% on Kimi Code Bench v2, +11.0% on Program Bench, +31.5% on MLS Bench Lite, and 30% lower reasoning overthinking. Google Research: Launched Gemini‑SQL2, a text‑to‑SQL system built on Gemini 3.1 Pro that hits state‑of‑the‑art scores on the BIRD benchmark. OpenAI: Added saved Codex rate‑limit resets for Go, Plus, Pro, and Business tiers (one free reset) and a two‑week invite program letting Plus/Pro users earn extra resets by inviting friends. Claude: Made dynamic workflows in Claude Code generally available, letting the model orchestrate parallel sub‑agents for complex tasks like codebase‑wide bug hunts and verify work before returning results.

Coding 3 min read

Latest from X - 2026-06-08 to 2026-06-09

Claude announced Claude Fable 5, a Mythos‑class model deemed safe for general use and now available everywhere, while Claude Mythos 5 stays limited to Glasswing partners. OpenRouter, Cursor, Visual Studio Code and Devin Desktop all reported that Claude Fable 5 is now live on their platforms, with Cursor noting a 72.9% score on CursorBench, 8 points above the previous best. Google DeepMind highlighted its 3.5 Live Translate, which streams speech into over 70 languages while preserving tone, pace and pitch for natural conversation. Kimi.ai introduced Kimi Work, a local desktop AI agent that can run up to 300 parallel agents and, via a WebBridge extension, navigate browsers to search, scroll and interact.

AI Models 4 min read

Latest from X - 2026-05-29

OpenRouter now supports "apply_patch," a server tool that lets models propose file edits using V4A diffs through the Responses API. The model generates a patch, and OpenRouter validates the diff syntax server-side. This feature allows for more efficient and accurate file editing. xAI has released grok-build-0.1 in public beta via the xAI API. This model powers the Grok Build CLI and excels at agentic coding, priced at $1/m input and $2/m output. Google AI has released an episode of Release Notes featuring the architects of Gemini, including @JeffDean, @koraykv, @OriolVinyalsML, and @NoamShazeer. They discuss their journey and the people behind the model. LangChain has released LangSmith LLM Gateway, which enforces spend limits and redacts PII before requests reach the model. They also announced Deep Agents v0.6, which makes harness profiles a first-class abstraction, allowing for production-grade performance at lower costs. NVIDIA has announced a new era of PC, but the details are unclear. OpenAI has launched Rosalind Biodefense to help trusted builders develop new biodefense and pandemic preparedness capabilities. They are also expanding trusted access to GPT-Rosalind for select U.S. government and allied partners.