Codex Drops Summary Compaction for Fresh Context Windows
Three merged openai/codex pull requests swap lossy summary compression for a fresh context window, plus history and notes tools models can query on demand.
119 verified stories covering AI Coding, product updates and industry developments. Page 1 of 2
Three merged openai/codex pull requests swap lossy summary compression for a fresh context window, plus history and notes tools models can query on demand.
ByteDance delays Doubao 2.2 past August to strengthen coding, tool-calling and agent skills, days after splitting its Seed AI unit into four divisions.
OpenAI is ending Cursor's model access by November 12, citing distrust of SpaceX after its $60 billion buyout of Cursor, even as Anthropic doubles down on Claude.
DeepSeek retook OpenCode's daily usage lead the moment GLM-5.3-Flash started charging, ending its 41-trillion-token free run.
OpenAI is testing a "Persistent mode" for Codex CLI that keeps the coding agent running until a user pauses it, with no launch date set.
Anthropic's Warp case study shows agents rewriting their own skill files from human feedback, with every change reviewed like code — no retraining involved.
Tobi Lütke said Shopify may drop Claude Code unless it reads AGENTS.md; Anthropic replied within hours, promising more flexible configuration support.
OpenAI's own staff use Codex almost universally, but paying subscribers barely touch agent tools — a gap that reveals what's really blocking AI agents today.
OpenAI is reinstating the five-hour usage limit for Plus accounts on ChatGPT Work and Codex, while a routing bug quietly downgraded 3% of Pro and Thinking requests to GPT-5.5-mini.
GPT-5.6 Sol built an RV32IM CPU gate by gate inside a puzzle-game sandbox, named it Codex-R32 itself, and got 1993's DOOM running on it — a serious test of cross-layer design consistency.
ByteDance folds AI coding tool TRAE and no-code agent platform Coze into Doubao, betting office is the entry point neither could carry alone.
An unnamed model called Ox Alpha appeared on OpenRouter with a free 1.05-million-token context window, and fingerprint clues point to Zhipu AI's GLM lineup.
Nvidia is paying AI coding startup Poolside $6 billion to license its model-training system and hiring 109 of its engineers, while Poolside's founders and company stay independent instead of being acquired outright.
Anthropic publishes a six-stage AI-native SDLC playbook that chains intent.md, spec.md and plan.md through Skills, Hooks, CLAUDE.md and evals.
OpenAI says the latest Codex rate-limit spikes trace back to sub2api-style subscription reselling, and warns the practice will keep getting flagged.
OpenAI open-sourced Codex's execution layer via three integration tiers, with Cisco and a tax-prep firm already using it to build agents.
Anthropic has Claude handle on-call for CI/CD incidents, delivering evidence-backed analysis in a median 14 minutes as engineer code output rose eightfold.
Zhipu's GLM-5.3 API prices at $1.40/$4.40 per million tokens and scores 60 on Artificial Analysis's Intelligence Index, ranking 8th of 182 models — but it burns tokens fast.
Cursor pushed its Origin code-hosting beta to all paid users Monday morning; hours later GitHub suffered a global outage lasting almost seven hours.
Qualcomm has open-sourced Mojo's compiler under Apache 2.0 with the LLVM exception, a month after acquiring Modular — though outside code contributions aren't accepted yet.
OpenCode's $10-a-month Go plan lifts DeepSeek Flash's monthly quota from $15 to $30, while Pro holds at $15 and adds peak and off-peak pricing.
OpenAI says AI now triages almost all initial security alerts, leaving high-impact calls to humans, according to an August 17 post from Greg Brockman.
DeepSeek's V4 Pro API now points to the 0813 version, pairing a 1M-token context and 384K output ceiling with a clear price warning.
Qwen3.8-2.4T-A95B has moved from a promise to an open-weight ModelScope repository, giving developers a 4.89TB Max-class model card to test.
InclusionAI open-sourced Ling-3.0-tiny, a 7.9B-parameter MoE model that activates 1.3B parameters per token and targets local agent workloads.
Thinking Machines released Inkling-Small with open weights, 276B total parameters, 12B active parameters, multimodal input and benchmark scores close to the 975B Inkling.
DeepSeek put V4-Flash 0731 into API beta, added Responses API support for Codex, and reported large agent-benchmark gains from post-training alone.
Unity China released Tuanjie Engine 2.0 with Codely, positioning game agents as auditable engineering workflow rather than one-line game generation.
Laguna S 2.1 gives enterprises a self-hosted coding-agent base, trading frontier scores for local control and clearer deployment economics.
Qwen3.8-Max-Preview arrives with 2.4T parameters, Qoder access and steep credit discounts before a full model card or independent benchmarks.
Google is reportedly months behind on Gemini 3.5 Pro as coding performance becomes the main test of its flagship model.
Moonshot AI has launched Kimi K3 with 2.8 trillion parameters, a 1M-token context window and published API pricing, while the open-weight proof still depends on weights and model cards.
Agnes-2.5-Flash uses a free coding model to pressure agentic developer tools. The piece reviews the verified facts and why the signal matters beyond one announcement.
Grok Build is accused of uploading repository bundles beyond opened files. The piece reviews the verified facts and why the signal matters beyond one announcement.
Claude-powered Bun rewrite sparks Zig backlash. A concise localization of the verified facts and the industry signal behind the story.
Grok 4.5 targets coding seats with low prices. A concise localization of the verified facts and the industry signal behind the story.
GhostApproval exposes symlink risks in AI coding tools. A concise localization of the verified facts and the industry signal behind the story.
Meta prices Muse Spark 1.1 for agent work is reframed for global technology readers, with the key numbers, product claims and open questions kept intact.
MoonBit bets on an AI-native programming language is reframed for global technology readers, with the key numbers, product claims and open questions kept intact.
Claude Code faces Chinese NVDB risk warning is reframed for global technology readers, with the key numbers, product claims and open questions kept intact.
Software factories, enterprise agents and investor-operators.
Model choice, Azure hosting and coding-model price pressure.
AI coding agents, ACP protocol support and editor strategy.
Meta told some applied-AI engineers to seek approval before using Claude Code or Codex, citing model distillation and contract risks.
Anthropic positioned Claude Sonnet 5 as a mid-tier model for autonomous agent work, with promotional pricing below the Opus tier.
Cursor Mobile lets users start a coding agent from a phone or take over an agent already running on desktop.
Chamath returns as CEO after 8090 Labs raises $135M
Google forms a coding push as AI-written code stalls near half
Wix’s Base44 ships its own model as ARR reaches $150M
Tenet Security describes Agentjacking, where attackers inject fake error reports that coding agents treat as trusted repair instructions, exposing secrets and deployment credentials.
Z.ai’s open-weight GLM-5.2 scored 74.4 on FrontierSWE, near Claude Opus and above GPT-5.5 in some coding benchmarks, while keeping token prices far lower.
Unconfirmed OpenAI signals suggest GPT-5.6 may focus on longer context, slower but stronger reasoning, and agentic coding.
Tenet's Agentjacking research shows how trusted telemetry can become executable instructions for coding assistants.
The load from always-on coding agents is testing GitHub's human-era infrastructure and Microsoft's Azure migration plan.
Days after its IPO, SpaceX exercised an all-stock option to acquire Anysphere, the company behind Cursor.
Moonshot AI released Kimi K2.7-Code with a HighSpeed mode, strong self-reported tool-use scores and a much lower output price than Claude Opus 4.8.
xAI added an Agent Dashboard to Grok Build, lowering access from the $300 SuperGrok Heavy tier to cheaper plans and joining the race to manage many coding agents at once.
A Shanghai Jiao Tong University-led benchmark shows agents often find the right files but identify the exact bug-relevant lines only about 14% to 19% of the time.
Zhipu's GLM-5.2 pairs a 744B-parameter MoE design with a 1M-token window and a coming MIT-style open-weight release for coding agents.
Moonshot's open MoE model offers far cheaper input and output tokens while improving agent-style coding benchmarks over K2.6.
Cursor's classifier lets low-risk agent actions proceed while pausing higher-stakes steps, reducing approval fatigue for developers.
Anthropic's new command copies a live coding session so developers can try risky approaches without losing the original context.
Moonshot open-sources Kimi K2.7-Code for AI coding.
The acquisition adds cloud development infrastructure to Codex, tightening OpenAI’s push from coding assistant toward full software production workflows.
OpenAI’s purchase of the Gitpod-linked cloud sandbox company points Codex toward managed development environments, not just code suggestions.
Cohere’s 30B-parameter MoE model activates about 3B parameters, supports 256K context and is released under Apache 2.0.
Cursor plans a London European headquarters and wants to grow its regional team to about 200 people by year-end.
The Postgres-based backend platform says its user base has more than doubled since the last round and agents now deploy most databases on the platform.
Copilot’s AI Credits make agentic sessions and frontier models usage-based, exposing developers to bills far above the old flat subscription.
New role plugins, Sites and annotations turn Codex from a coding assistant into a broader workflow tool for non-developers.
A reported Google Play pilot asks developers to sell source code, including dormant projects, while the email itself avoids saying AI.
At Build 2026, Microsoft presented seven self-trained models, with MAI-Code-1-Flash beating Claude Haiku 4.5 on its cited coding benchmarks.
GitHub is replacing Copilot premium requests with AI Credits, making heavier agentic coding sessions pay closer to their actual token cost.
At Build 2026, Microsoft positioned Windows as an agent platform and said Project Polaris will become the default GitHub Copilot reasoning engine.
Sekai has raised a $20 million Series A after users generated 15 million mini apps and spent more than an hour a day on the mobile platform.
Copilot is moving from premium request counts to AI Credits, exposing the real token cost of agent-style coding.
AI coding startup Cognition raises over $1 billion at a $26 billion valuation, claiming 90% of its internal code is now generated by its own AI agent Devin.
At Anthropic's first European developer event, a show of hands revealed that most attendees had merged pull requests written entirely by Claude without reading the code.
Cursor's 3.5 update moves its Automations feature from a standalone web page into the IDE's Agents Window, enabling multi-repo support and agent-driven workflows beyond coding.
Google launched Antigravity 2.0 at I/O 2026, a standalone agent platform that can orchestrate 93 sub-agents to build an operating system in 12 hours.
Cursor's third-generation coding model Composer 2.5 matches Opus 4.7 and GPT-5.5 on key benchmarks while costing only a tenth per task.
Google's Gemini 3.5 Flash outperforms GPT-5.5 on agentic benchmarks while being dramatically cheaper, and becomes the default model across Gemini app and Search.
OpenAI added Codex controls to ChatGPT on iPhone, iPad and Android, letting users monitor and approve coding tasks away from the desktop.
Cursor’s cloud agents now support multi-repository environments, faster builds and enterprise controls for teams with microservice codebases.
Coder Agents gives enterprises a self-hosted, model-agnostic way to run AI coding agents without sending code or prompts outside their network.
Cognition AI, the maker of the AI software engineer Devin, is reportedly in funding talks at a $25 billion valuation, a sixfold increase in 13 months.
Sundar Pichai revealed during a quarterly earnings call that 75% of Google's new code is now AI-generated and human-reviewed, up from 25% 18 months ago. The shift is reshaping developer roles and highlights internal use of rival tools like Claude Code at DeepMind.
DeepSeek released V4-Pro and V4-Flash under MIT license, achieving top competitive programming scores while pricing output at $3.48/M — a fraction of GPT-5.5 and Claude Opus 4.7.
A new JetBrains survey of over 10,000 professional developers reveals that 90% now use AI tools, with Claude Code surging from 3% to 18% adoption in six months while GitHub Copilot's growth stalls.
Anthropic redesigned the Claude Code desktop app with multi-session management, built-in dev tools, and Routines for automated, event-triggered coding tasks.
AI coding tools have been widely adopted for two years, but a new phenomenon called "Tokenmaxxing" is emerging where developers treat AI token budgets as a status symbol. Data shows code churn is up to 9.4 times higher, and only 10-30% of AI-generated code survives after a few weeks.
Alibaba's Qwen team has released Qwen3.6-35B-A3B under Apache 2.0, a 35B-parameter MoE model that activates only 3B per token, scoring 73.4% on SWE-bench Verified and outperforming Google's Gemma 4-31B by 21 points.
AI coding startup Cursor is negotiating a $2 billion funding round at a pre-money valuation exceeding $50 billion, nearly doubling its valuation in six months.
Agoda's engineering team found that while AI coding tools significantly increased individual developer output, project delivery speed remained unchanged, with code review becoming the new bottleneck.
In the first week of April 2026, Cursor v3, OpenAI's Codex plugin for Claude Code, and an emergent three-layer stack show that AI coding tools are diverging into specialized layers rather than converging into one.
Apple is sending Siri team engineers to an AI coding bootcamp focused on using tools like Claude Code, as the company scrambles to overhaul Siri ahead of WWDC in June.
Factory has closed a $150 million Series B led by Khosla Ventures, reaching a $1.5 billion valuation, with Morgan Stanley, EY, and Palo Alto Networks already paying customers for its enterprise-grade AI coding agent platform.
Z.ai (formerly Zhipu AI) released GLM-5.1, a 754B-parameter open-source MoE model that scored 58.4% on SWE-Bench Pro, surpassing GPT-5.4 and Claude Opus 4.6. The model was trained entirely on Huawei Ascend 910B chips with MindSpore, using zero NVIDIA hardware.
OpenAI has deployed its GPT-5.3-Codex-Spark model on Cerebras WSE-3 chips instead of Nvidia GPUs, achieving over 1,000 tokens per second — roughly 15 times faster than the standard version. The move marks OpenAI's first production deployment on non-Nvidia hardware.
OpenAI has updated Codex with background computer use, allowing it to operate other macOS apps autonomously while users continue working. The update also includes a built-in browser, image generation, over 90 new plugins, SSH access, and self-scheduling capabilities.