TabFM is Google Research’s new foundation model that predicts on tabular classification and regression tasks without any per‑dataset training, hyperparameter tuning, or manual feature engineering. It leverages in‑context learning (ICL) with a hybrid attention architecture and is pretrained on hundreds of millions of synthetic tables. Benchmarks on the TabArena suite show TabFM (both default and ensemble variants) achieving higher Elo scores than heavily tuned traditional models.
Google Research: Introduced TabFM, a foundation model for tabular data classification and regression that can generate high‑quality predictions on unseen tables in a single forward pass. ClaudeDevs: Raised Claude Platform API rate limits for all users, removed spend‑based tiers, and gave the latest Sonnet and Haiku models five‑times higher limits at the top tier. Google AI: Launched two workflow updates - one model for ultra‑fast image generation and another to instantly animate those images, both at a fraction of the usual cost, highlighted by Nano Banana 2 Lite. GitHub Changelog: Announced that Gemini 2.5 Pro and Gemini 3 Flash will be deprecated in all Copilot experiences on July 31 2026, urging migration to Gemini 3.1 Pro or Gemini 3.5 Flash beforehand. Z.ai: Rolled out ZCode, the official development environment for GLM‑5.2, offering 1.5× usage quota for GLM Coding Plan subscribers, BYOK support, and cross‑platform availability on macOS, Windows, and Linux. GitHub: Made Kimi Moonshot’s Kimi K2.7 Code generally available in GitHub Copilot as the first open‑weight model selectable in the model picker, noting lower cost with performance comparable to top‑tier models.
OpenAI has begun a limited preview of the GPT‑5.6 family: **Sol** (flagship), **Terra** (balanced, 2× cheaper than GPT‑5.5), and **Luna** (fast, low‑cost). Sol introduces a new “max” reasoning mode and an “ultra” mode that uses sub‑agents. Early benchmarks show state‑of‑the‑art performance on coding (Terminal‑Bench 2.1), biology (GeneBench v1), and cybersecurity (ExploitBench, ExploitGym). The models ship with OpenAI’s most robust safety stack to date, but the preview may block or delay some requests.
Google DeepMind announces Gemini 3.5 Flash now includes a native computer‑use tool, enabling developers to create agents that can see and act across browsers, mobile and desktop interfaces. OpenAI says its new GPT‑5.5 Instant version is more conversational, better at grasping intent, handling complex constraints, and improving shopping‑related interactions. Mistral AI introduces Mistral OCR 4, which provides structured output with bounding boxes, block classification and inline confidence scores for 170 languages. Visual Studio Code’s latest release adds a unified model customization picker for tuning and shows total chat‑session costs, simplifying cost tracking and usage insight.
GLM‑5.2 is Z.ai’s newest open‑source model designed for long‑horizon coding tasks. It supports a solid 1 million‑token context, introduces flexible “effort” levels for balancing speed and capability, and uses the IndexShare architecture to cut per‑token FLOPs by 2.9×. Benchmarks show it outperforms its predecessor (GLM‑5.1) and ranks as the strongest open‑source coding model, closing the gap to leading closed‑source systems.
Kimi.ai: Released the open‑source Kimi‑K2.7‑Code model, reporting +21.8% on Kimi Code Bench v2, +11.0% on Program Bench, +31.5% on MLS Bench Lite, and 30% lower reasoning overthinking. Google Research: Launched Gemini‑SQL2, a text‑to‑SQL system built on Gemini 3.1 Pro that hits state‑of‑the‑art scores on the BIRD benchmark. OpenAI: Added saved Codex rate‑limit resets for Go, Plus, Pro, and Business tiers (one free reset) and a two‑week invite program letting Plus/Pro users earn extra resets by inviting friends. Claude: Made dynamic workflows in Claude Code generally available, letting the model orchestrate parallel sub‑agents for complex tasks like codebase‑wide bug hunts and verify work before returning results.
Google announced Gemini 3.5 Live Translate, an audio model that streams speech‑to‑speech translation in real time for more than 70 languages. It is available now in public preview via the Gemini Live API, in private preview for Google Meet, and globally in the Google Translate mobile app.
OpenRouter (@OpenRouter) has integrated its models into ComfyUI workflows, allowing users to leverage OpenRouter models directly within ComfyUI. GitHub (@github) highlights the 2026 Partner Pack, offering exclusive discounts and perks for maintainers. The GitHub Innovation Graph provides economic data on trends in GDP, inequality, and emissions, which researchers find valuable. Google AI Developers (@googleaidevs) showcases successful implementations of Managed Agents in the Gemini API, including Eigent_AI's root cause analysis and llama_index's document processing template. NVIDIA (@nvidia) discusses the Dell AI Factory with NVIDIA, which enables companies to build, run, and scale AI, with NemoClaw powering agentic AI on-prem. OpenAI (@OpenAI) shares Terence Tao's experience with AI, which gives researchers more freedom to experiment and pursue unconventional ideas. LangChain (@LangChain) emphasizes the efficiency of LangSmith Sandboxes, which pause automatically when idle, and encourages users to create agents using everyday language with LangSmith Fleet.
OpenRouter now supports "apply_patch," a server tool that lets models propose file edits using V4A diffs through the Responses API. The model generates a patch, and OpenRouter validates the diff syntax server-side. This feature allows for more efficient and accurate file editing. xAI has released grok-build-0.1 in public beta via the xAI API. This model powers the Grok Build CLI and excels at agentic coding, priced at $1/m input and $2/m output. Google AI has released an episode of Release Notes featuring the architects of Gemini, including @JeffDean, @koraykv, @OriolVinyalsML, and @NoamShazeer. They discuss their journey and the people behind the model. LangChain has released LangSmith LLM Gateway, which enforces spend limits and redacts PII before requests reach the model. They also announced Deep Agents v0.6, which makes harness profiles a first-class abstraction, allowing for production-grade performance at lower costs. NVIDIA has announced a new era of PC, but the details are unclear. OpenAI has launched Rosalind Biodefense to help trusted builders develop new biodefense and pandemic preparedness capabilities. They are also expanding trusted access to GPT-Rosalind for select U.S. government and allied partners.
LangChain (@LangChain) LangChain has released a new course on LangSmith Fleet Essentials, allowing users to build, use, and manage agent fleets for complex tasks without coding. The course is a quickstart guide to building and improving agents. LangSmith Engine is also mentioned, which optimizes self-improving loops. A keynote from @hwchase17 highlighted the future of agents. Google DeepMind (@GoogleDeepMind) Google DeepMind's Gemini for Science tools aim to help scientists achieve their next breakthrough.