AI Coding Tools Articles
53 articles across AI Catchup's news, guides, tutorials, and comparisons.
All AI Coding Tools articles
Codex Composer Predictions Beta Suggests the Next Message in Desktop Threads
OpenAI added composer predictions to Codex in beta. Eligible personal ChatGPT Pro users aged 18 and older can see complete next-message suggestions in the Codex desktop app for supported local and SSH threads, accept them with Tab, edit before sending, and turn the feature off in Settings.
Claude Code Subagents vs Agent Teams vs Dynamic Workflows: Which Parallel Approach to Use
Claude Code runs parallel work five ways. Use subagents for side tasks inside one session, agent view to dispatch background sessions, agent teams when workers must talk to each other, dynamic workflows for dozens to hundreds of cross-checked agents, and Projects for cloud work spanning days.
Cursor Adds GLM 5.3 and GLM 5.3 Flash to Its Model Picker
Cursor lists Z.ai's GLM 5.3 and GLM 5.3 Flash in its model picker. GLM 5.3 is for higher-quality work at $1.4 per million input tokens and $4.4 per million output tokens; Flash is smaller and faster at $0.15 and $0.5. Both support Cursor's agent tools and draw from the third-party Other Models usage pool.
Claude Code Mods Let Developers Customize the CLI and Desktop App
Anthropic introduced Claude Code mods as a plugin-based way to customize how Claude Code behaves and looks. Small JavaScript or TypeScript modules can add interface elements, change tool calls, or replace built-in features. Mods run with Claude Code's access to your machine, so review their source before installing.
Codex Cloud Reusable Environments Keep Coding Tasks Moving
OpenAI's Codex Cloud uses reusable, published environments that package a repository, tools, dependencies, and access settings. Create the setup on desktop or web, then run isolated coding tasks from desktop, web, or mobile. Tasks can continue while your computer sleeps, subject to plan eligibility and workspace access.
Claude Code Can Now Create Design, Slide, and Doc Artifacts
Claude Design, Claude Slides, and Claude Docs work inside Claude Code, so developers can point Claude at files and RFCs, ask for a design review deck or UI mockup, edit it, and share it from desktop or terminal. As of October 8, 2026 the three templates are out of beta and on every plan.
Claude Code Desktop Adds Pop-Out Panes for Diff and Terminal Work
ClaudeDevs says Claude Code Desktop can now pop out any pane into its own window. You can move a diff or terminal to another screen while Claude keeps working, dock it back later, and run sessions side by side or stacked.
Claude Fable 5.1 and Mythos 5.1: Same Price, Cheaper Cache Reads, Fewer Safeguard Interventions
Anthropic released Claude Fable 5.1 on September 1, 2026 as claude-fable-5-1 at $10 per million input and $50 per million output tokens, with cache reads cut 75% to $0.25 per million. Anthropic estimates typical workloads cost about 25% less than Fable 5. Claude Mythos 5.1 is the same model with fewer cyber and life-science safeguards, limited to vetted US organizations.
OpenAI Plans to End Direct Cursor Model Access on November 12
OpenAI plans to end Cursor's direct access to OpenAI models. November 12, 2026 is proposed, and as of October 9 neither company has confirmed it. Only OpenAI models picked inside Cursor are affected, about 5% of its traffic. Local Chat and Agent can keep OpenAI through your own API key or the Codex extension; Tab, Auto and Cloud Agents cannot.
Ox Alpha Was Z.ai's GLM-5.3-Flash, and the Free Week Is Over
Ox Alpha was Z.ai's GLM-5.3-Flash. OpenRouter's model page now names the developer, the stealth listing serves no providers, and OpenCode has dropped Ox Alpha Free. The named model lists at $0.15 per 1M input tokens and $0.50 output; the 50% launch discount ended September 9, 2026, and as of September 10 Z.ai's pricing page carries list prices only.
Codex Opens GPT-5.6 Sol's 1M-Token Context to ChatGPT Accounts
OpenAI's Tibo Sottiaux said on August 17, 2026 that GPT-5.6 Sol's 1M-token context option in Codex now works through ChatGPT accounts, not only API keys. OpenAI's docs list a 1,050,000-token API window for the model and a model_context_window setting, but no ChatGPT-account rule. GPT-6.1 Sol is now Codex's recommended model.
OpenAI's GPT-5.6 Builder's Guide Shows How to Make Agents More Efficient
OpenAI's GPT-5.6 builder's guide lays out a Responses API architecture for longer-running agents: persist reasoning, compact context, delegate work across agents, and move deterministic tool processing into code. OpenAI reports that these patterns raised GPT-5.6 Sol's ARC-AGI-3 score from 13.3% to 38.3% while using roughly 6x fewer output tokens.
OpenAI Expands Daybreak With GPT-5.6-Cyber
OpenAI expanded its Daybreak cybersecurity program with Blue and Red access tiers and introduced GPT-5.6-Cyber, a purpose-trained model for authorized vulnerability research, exploit validation, and security testing. Access is limited to approved defenders and organizations.
Claude Code Artifacts Expand to Pro and Max Plans (Private Until You Share)
Claude Code Artifacts, which launched for Team and Enterprise, reached Pro and Max plans in July 2026. On Pro and Max an artifact is private until you share it, and a public link is the way to share it; sharing with specific people or a whole organization is a Team and Enterprise feature.
Claude Desktop for Linux Beta: Ubuntu and Debian, With Caveats
Anthropic shipped a beta of the Claude desktop app for Linux on Ubuntu 22.04+ and Debian 12+ (x86_64 or arm64), with Chat, Cowork, and Claude Code on all paid plans. The app does not self-update; updates arrive through apt, and as of August 2026 a directly installed .deb registers the repo itself. Computer Use and Dictation are missing.
Claude Sonnet 5 Launch: Anthropic's Most Agentic Sonnet, Since Replaced by Sonnet 5.5
Anthropic launched Claude Sonnet 5 on June 30, 2026 as its most agentic Sonnet yet, and made its $2 input and $10 output per million token pricing permanent. Claude Sonnet 5.5 replaced it on September 28; Anthropic now lists Sonnet 5 as a legacy model, still available on the API as claude-sonnet-5.
Cursor for iOS: Cloud Agents Go Mobile-First in Public Beta
Cursor's native iOS app lets developers launch and manage coding agents from a phone. Its October 6, 2026 Remote Control update documents how to pair with agents running on a computer, including the Enterprise admin toggle and the requirement to keep that computer on and online. The app also supports cloud agents, mobile review, and PR merges.
Codex Permission Profiles: Least-Privilege Controls for Local Agent Work
OpenAI shipped Codex permission profiles in beta -- reusable, inheritable policies that replace the coarse sandbox_mode/sandbox_workspace_write combo. A profile binds OS-enforced filesystem read/write/deny rules (down to **/*.env) to per-domain network and Unix-socket rules. Enterprise admins get fail-closed allowlists via requirements.toml. Profiles govern local sandboxed command execution only, not MCP servers, app connectors, browser, or cloud.
Claude Code Adds Artifacts: Live, Shareable Pages for PR Walkthroughs and Dashboards
Artifacts in Claude Code turn an in-progress session into a live web page that updates as the session works. Launched in June 2026 as a beta for Team and Enterprise with org-only sharing, they are now on Pro, Max, Team, and Enterprise, and an artifact can be shared to a public link anyone can open.
OpenAI Previews GPT-5.6: Sol, Terra, and Luna in Limited Preview
OpenAI previewed the GPT-5.6 family in June 2026 with trusted partners: Sol, a frontier flagship it called a step function better than GPT-5.5; Terra, competitive with GPT-5.5 at 2x lower cost; and Luna, its most cost-efficient model. All three became generally available on July 9, 2026, and GPT-6 has since superseded them.
Anthropic Suspended Fable 5 and Mythos 5 on June 12, 2026, and Restored Them on July 1: What Happened
A US export control directive forced Anthropic to disable Claude Fable 5 and Mythos 5 for all customers on June 12, 2026. The controls were lifted June 30 and Fable 5 returned globally July 1 with a new safety classifier. As of September 2026 both models are active, and Fable 5.1 is the current Fable model.
Claude Fable 5 and Mythos 5: Mythos-Class Capability Goes General, With Caveats
Anthropic launched Claude Fable 5 on June 9, 2026 as a Mythos-class model made safe for general use, plus Claude Mythos 5 for vetted cyberdefenders via Project Glasswing, both at $10 per million input and $50 per million output tokens. Fable 5.1 and Mythos 5.1 replaced them on September 1; Anthropic now lists Fable 5 as a legacy model.
Codex for Every Role: Role-Specific Plugins, Codex Sites, and Annotations Beyond Code
OpenAI announced six role-specific Codex plugins (62 apps, 110 skills), Sites, and annotations for documents, spreadsheets, and slides on June 2, 2026. As of September 2026, Sites is a public beta on Plus, Pro, Business, Enterprise, and Edu plans, and the plugin directory is shared by ChatGPT and Codex across the desktop app, the web, and Codex CLI.
Codex Build iOS Apps Plugin: Mirror the Simulator in the Browser and Hot-Reload SwiftUI Previews
OpenAI's Build iOS Apps plugin for Codex bundles nine iOS and Swift skills. The June 2026 addition mirrors the iOS Simulator into the Codex in-app browser and hot-reloads Swift Package-backed SwiftUI previews. XcodeBuildMCP handles simulator build, run, and debug. It installs from the /plugins browser in Codex CLI or the Plugins tab in the ChatGPT desktop app.
Cursor Shared Canvases: Publish an Agent Canvas and Share It With Your Team via URL
Cursor added shared canvases -- you can now share a canvas from Cursor with your team by generating a link to a live snapshot that teammates open in the browser. Recipients view it read-only in the Cursor Dashboard, so you distribute a working dashboard or report instead of a full chat thread. Shared canvases are available on Pro, Teams, and Enterprise plans.
Claude Opus 4.8 Fast Mode: 2.5x Faster Output Tokens; Opus 5.5 Is Now the Default
Anthropic launched fast mode for Claude Opus 4.8 on May 28, 2026: the same model at 2.5x the output speed, three times cheaper than before. Opus 5.5 is now the fast mode default in Claude Code v2.1.280 and later, at $8/$40 per MTok; Opus 4.8 still supports it at $10/$50, billed to usage credits on subscription plans.
Claude Code ships a security-guidance plugin for in-session vulnerability checks
Anthropic shipped an official security-guidance plugin for Claude Code. It runs automatic vulnerability checks while Claude edits files, at the end of each turn, and when Claude runs commits or pushes through its Bash tool.
Claude Agent SDK Monthly Credit on Paid Plans: Paused in June, Now Covered by Max and Team API Credits
The Claude Agent SDK monthly credit never started: Anthropic paused it on June 15, 2026. Since October 7, 2026, Max and Team plans instead include monthly API credits that cover the Agent SDK, `claude -p`, the Claude API, and Managed Agents. Pro and Enterprise get no credit; plan sign-in usage still draws from subscription limits.
Claude Code 2.1.142: `claude agents` Gains Session Flags, Fast Mode Defaults to Opus 4.7, MCP Tool Timeout Honored
Claude Code 2.1.142 (May 14, 2026) added eight dispatch flags to `claude agents`, moved fast mode's default from Opus 4.6 to Opus 4.7, and made `MCP_TOOL_TIMEOUT` lift the 60-second cap on remote MCP servers. The flags and timeout fix stand; fast mode now defaults to Opus 5.5 (v2.1.280 and later) at $8/$40 per MTok.
Claude Code +50% Weekly Limits Promo Ended September 13, 2026; Limits Are Now 25% Higher Permanently
Anthropic's +50% Claude Code weekly-limits promotion ended September 13, 2026 at 11:59 PM PT. On September 14 Anthropic rewrote its support article: starting September 14, 2026, weekly limits in Claude Code are 25% higher than they were before the promotion for Pro, Max, Team, and seat-based Enterprise plans. Five-hour limits are unchanged.
Codex in the ChatGPT Mobile App: Run, Review, and Steer Codex from Your Phone
Yes, you can drive Codex from your phone. The ChatGPT mobile app on iOS and Android connects to a host running the ChatGPT desktop app on macOS or Windows, and from the phone you start work, review diffs and terminal output, and approve actions. Codex still executes on the host, which must stay awake and online.
Codex Hooks and Programmatic Access Tokens: Setup, Trust Model, and What Actually Runs Today
Codex access tokens are ChatGPT Business and Enterprise workspace credentials for non-interactive Codex CLI runs. Create one at chatgpt.com/admin/access-tokens, then authenticate with CODEX_ACCESS_TOKEN or codex login --with-access-token. Hooks are the in-session extensibility framework: eleven lifecycle events. As of August 2026, command and mcp_tool handlers execute; at launch only command handlers did.
Codex CLI 0.130.0 Adds `remote-control`, Richer Plugin Sharing Metadata, and Better App-Server Thread Paging
OpenAI shipped Codex CLI 0.130.0 in May 2026. The release adds a new `codex remote-control` command for starting a headless, remotely controllable app-server, improves app-server clients with paging options for large threads (unloaded/summary/full turn items), expands plugin sharing with link metadata and discoverability controls, adds Bedrock auth support for AWS console-login credentials from `aws login` profiles, and fixes several app-server/thread reliability issues.
Claude Code 2.1.133: `worktree.baseRef` Default Returns to `origin/<default>`, MCP OAuth Proxy Honored Across the Whole Flow
Anthropic shipped Claude Code 2.1.133 on May 7, 2026. The headline is a worktree-base behavior change: a new `worktree.baseRef` setting (`fresh` | `head`) defaults to `fresh`, which moves `EnterWorktree`'s base back to `origin/<default>` after three days of branching from local `HEAD` (since 2.1.128 on May 4). The release also routes `HTTP(S)_PROXY` / `NO_PROXY` / mTLS through the entire MCP OAuth flow (discovery, dynamic client registration, token exchange, refresh), exposes effort level to hooks via `$CLAUDE_EFFORT`, adds Linux sandbox path overrides, and fixes a refresh-token race that was 401-ing parallel sessions.
Codex CLI 0.129.0 Adds Modal Vim Composer, Redesigned Resume/Fork Picker, and a `/hooks` Browser
OpenAI shipped Codex CLI 0.129.0 on May 7, 2026. The release brings modal Vim editing to the TUI composer via `/vim`, a redesigned resume/fork picker, a raw scrollback mode, workspace-aware `/diff`, a new `/hooks` browser with before/after compaction support, expanded plugin management with workspace sharing and share access controls, theme-aware status lines, and Codex Apps auth surfaced through Guardian. Plus a long bug-fix list across Linux/Windows sandboxes, MCP, and TUI input handling.
Cursor Adds Enterprise Model Controls, Soft Spend Limits, and Richer Usage Analytics
Cursor's May 4, 2026 update gave Enterprise admins provider- and model-level allow and block lists, soft spend limits with alerts at 50%, 80%, and 100%, and usage analytics by user and product surface. The June 1, 2026 deadline to migrate old blocklists has passed; model access now lives under Team Settings, Models.
Warp Goes Open Source: AGPL Client, MIT UI Framework, and a New `settings.toml`
On April 27, 2026 (changelog v0.2026.04.27.15.32) Warp open-sourced its client at github.com/warpdotdev/warp under AGPL v3, with the `warpui` UI framework crates released under MIT. The same release adds a TOML settings file editable from the settings page or by asking Warp's agent. The server stays closed-source. OpenAI is the founding sponsor.
Cursor SDK: Programmatic Cursor Agents in TypeScript and Python, Local or Cloud
Cursor launched its SDK in public beta on April 29, 2026 as a TypeScript package. As of October 9, 2026 it ships for TypeScript (`@cursor/sdk`) and Python (`cursor-sdk`), released together at version 1.0.35, and the docs no longer call it beta. Agents run locally, on Cursor-hosted VMs, or on self-hosted workers.
Anthropic's Claude Code Post-Mortem: Three Engineering Missteps Behind the Spring 2026 Quality Decline
Anthropic published a post-mortem on April 23, 2026 explaining the Claude Code quality regression that ran from early March through mid-April: a March 4 default-effort downgrade from high to medium, a March 26 caching change that wiped reasoning history every turn, and an April 16 verbosity prompt that capped responses at 25 words between tool calls. All three were resolved by April 20, the API was unaffected, and Anthropic reset usage limits for all subscribers.
Cursor 3.2 Adds /multitask Async Subagents, Worktrees Polish, and Multi-Root Workspaces
Cursor 3.2 shipped on April 24, 2026 with three changes: a `/multitask` command that fans a request out to async subagents instead of queueing, an improved worktrees experience for background branch work, and multi-root workspaces for cross-repo sessions. As of September 2026 the last two do not combine -- Cursor documents worktrees as disabled inside multi-root workspaces -- and `/multitask` no longer appears anywhere in Cursor's documentation.
GPT-5.5 Launch: State-of-the-Art Agentic Coding, 1M Context, and a New Pro Tier
OpenAI launched GPT-5.5 on April 23, 2026, with state-of-the-art scores on Terminal-Bench 2.0 (82.7%) and GDPval (84.9%) and a GPT-5.5 Pro tier. GPT-6.1 Sol and GPT-6 Astra have since replaced it, and it retires from ChatGPT and Codex on October 14, 2026. The API keeps it at $5/M input and $30/M output.
Claude Code Subagent Patterns: 10 Reusable Agent Definitions
A Claude Code subagent is a delegated worker with its own context window, defined as a markdown file with YAML frontmatter in .claude/agents/. Subagents can edit files when you grant Edit or Write, nest three layers deep by default, and run 20 at a time. These 10 definitions cover the highest-value delegations.
Claude Opus 4.7 Best Practices: How to Actually Get the Most Out of the Upgrade
Claude Opus 4.7 follows instructions more literally than 4.6, runs longer agentic tasks more reliably, and ships a new xhigh effort level. Anthropic's launch-day guidance is to specify the task up front, batch your interactions, use auto mode, and default to xhigh; this guide adds the verification, recap, and scoping habits that make those gains show up in your sessions.
Cursor Self-Documentation: Cursor Spawns a Subagent to Read Its Own Docs Before Answering
On April 17, 2026, Cursor's Eric Zakariasson announced that Cursor can more accurately answer questions about itself: it spawns a subagent to fetch the latest Cursor docs and updates before answering. The announcement was an X post; Cursor's subagent docs list only Explore, Bash, and Browser as built-in subagents and do not document it separately.
Codex CLI vs Claude Code vs Cursor: 2026 Architecture Deep-Dive (Sandboxing, Context, Plugins, Scheduling)
Codex CLI, Claude Code and Cursor now all confine local shell commands in an OS-enforced sandbox and all run hook scripts, so the real split is defaults and composition. Codex sandboxes by default, Claude Code's sandbox is opt-in, and Cursor sandboxes under Auto-review. Choose by default posture, Windows support and how unattended work runs.
Claude Opus 4.7 Launch: Stronger Coding, xhigh Effort, and a New Cyber Safeguards Tier
Anthropic launched Claude Opus 4.7 on April 16, 2026, with gains over Opus 4.6 in advanced software engineering at the same $5/$25 pricing, a new xhigh effort level, /ultrareview, higher-resolution vision and the first cyber safeguards. It is now a legacy model, succeeded by Opus 4.8, Opus 5 and Opus 5.5.
Cursor Canvases: When to Ask the Agent for a UI Instead of Text
Cursor canvases are agent-generated interactive artifacts -- dashboards, reports, audits -- that render beside the chat. Cursor saves each one to your workspace canvas list, so you can reopen, edit, and rerun it later with fresh data. Publishing a canvas as a browser link for teammates needs a paid plan and a team.
Warp's Universal Agent Support: The Terminal as an Agentic Development Environment
Warp's universal agent support runs Claude Code, Codex, OpenCode, Gemini CLI and a dozen other CLI agents inside its open-source terminal, with vertical tabs, code review hand-off and Remote Control. Those agents bill to their own accounts, not Warp's. Warp's paid plans (Build $20, Max $200, Business $50 a month) cover its own agent.
Master Claude Code's 1M Context Window: Rewind, Compact, Clear, and Subagents
Claude Code's 1M token context window opens longer autonomous sessions but introduces 'context rot' -- degraded performance as the window fills. Master four turn-end tools: /rewind to drop bad branches, /compact to summarize and continue, /clear to start fresh with a distilled brief, and subagents to wall off noisy work in their own context.
Inside Claude Code's Rebuilt Desktop: Parallel Agents, Drag-Drop Panes, Side Chat
Anthropic rebuilt the Claude Code desktop app on April 14, 2026 around parallel agent workflows: a multi-session sidebar, drag-and-drop panes, an in-app file editor, integrated terminal, side chat, three view modes, and SSH on macOS. Four months on the app also creates routines and local scheduled tasks, browses external sites, and controls your computer.
Claude Code Routines: Schedule, API, and GitHub-Trigger Your AI Agents
Claude Code Routines is Anthropic's new way to run saved Claude Code configurations automatically -- by schedule, API call, or GitHub event. Routines run on Anthropic's cloud infrastructure with a prompt, repo, and MCP connectors. Available in research preview on Pro, Max, Team, and Enterprise plans.
AI Tools Landscape: What Changed in Early 2026
Three shifts defined AI tooling in early 2026: MCP settled as the cross-tool standard, coding assistants grew past autocomplete into multi-file workflow partners, and narrowly autonomous agents reached production on supervised tasks. MCP's tipping point came in June 2025, when Visual Studio Code made MCP support generally available.
Cursor vs Claude Code in 2026: Which AI Coding Tool Is Better?
Neither tool is terminal-only or editor-only any more: Cursor ships a CLI and Claude Code ships a desktop app, IDE extensions and a browser surface. Choose Cursor for inline tab autocomplete and cross-lab model choice, though OpenAI plans to end Cursor's direct model access on a proposed November 12; choose Claude Code for one agent across surfaces.