What's rising in the AI coding ecosystem?
GPT-5.6 is released with three tiers: Sol for agentic coding, Terra for balanced performance, and Luna for speed and low cost 1056,1310. Sol is designed for long-horizon coding and agentic work 1057, and Sol on Medium is the preferred setting for coding with GPT-5.6 1255,1312. ChatGPT Work, built on GPT-5.6 and Codex, transforms ChatGPT into a general work agent that completes projects across apps 1069. GPT-5.6 Sol outperforms Claude and Gemini for solo builders due to benchmark leadership and the integrated Work/Codex interface 1420. Sol is also favored over Fable 5 for autonomous code building, reviewing, and screen testing because of lower token usage and better autonomy 1792. ChatGPT Work is preferred over Claude for automating the middle steps of office work 1417. GPT-5.6 Sol demonstrates real-time performance in Blender at 750 tokens/s without acceleration (1972), and is preferred over Claude Code for 3D modeling because it can generate textures directly (1978).[11]
Claude Code's new desktop launch signals a shift from copilot to autonomous software engineer, with the IDE becoming the operating system for development 1270. Claude Code originated from Anthropic's internal 'clide' tool 855. Version 2.1.203 adds background session auto-recovery and blocks unauthorized destructive actions 797. Version 2.1.201 reduces interruptions by removing mid-conversation system role 218. The Artifacts feature turns sessions into live, interactive web pages with version history 223. Claude Cowork enables cross-device task continuation with background processing when the laptop is closed 794. Claude Code skills are evolving from static instructions to self-improving, benchmarked software (1976). A hidden TeamMateTool in v2.1.19 reveals a leader-worker system, pointing to built-in multi-agent capabilities (1987). The new Home tab unifies Chat and Cowork and allows scheduling Skills to run autonomously 1009. Anthropic is turning internal tools into products, similar to Google's evolution 1003.[10]
Multi-agent collaboration is on the rise: codex-plugin-cc lets Claude Code delegate tasks to Codex in the same terminal 141. A new tool enables monitoring and managing multiple coding agents from mobile TUI and web 484. Self-improving agentic routines where Claude prompts Claude are gaining traction 1172. AI programming shifts from individual code completion to multi-agent collaboration, where Claude in Slack can fix bugs, write PRs, and send daily reports, making task decomposition and context management the new core skills 1108. Running multiple Claude Code sessions in TMux with agent teams and Git worktree isolation enables adversarial code reviews and parallel feature testing 376. Hermes Agent 0.18.0 Runtime introduces /moa and /learn commands 520, and Hermes Agent can learn and autonomously execute repetitive workflows after a single demonstration 321.[7]
Cursor remains the best all-rounder for AI coding, while Claude Code excels at design and research, Codex offers useful plugins, and Opencode provides the best value 1265. The Cursor $20 plan is the top value pick 1322. Interest in Cursor alternatives is rising following the SpaceX acquisition 1030. Cursor's decision to restrict Sonnet 5 and GPT 5.5 to Max-only plans signals a shift toward higher-tier monetization, causing user backlash 860. Grok 4.5 is now available in Cursor at $2/M input and $6/M output tokens 930. OpenAI's Sites feature allows turning an idea into a live site that can be published and shared 1171. The Browser Company's new desktop app introduces a fully functional in-app browser, a cloud browser for agents, and side chat integrating ChatGPT/Codex 1170. Cursor's iOS app enables building from anywhere with always-on cloud agents 379. Continue v1.2.24 shifts away from Hub slugs toward explicit model definitions in config templates 622.[9]
Cost efficiency is driving model choices: GLM-5.2 and MiniMax-M3 are preferred over Opus 4.8 for cost-sensitive tasks, offering similar intelligence at 1/5 the price 373. Running a 671B model locally on a Mac Studio is now feasible, signaling a shift toward on-device inference for large reasoning models 1555. The NVIDIA DGX Spark, with 128GB unified memory and 1 petaflop AI compute, enables running 200B parameter models locally without quantization, paying for itself in 10 months vs cloud subscriptions 229. To use Claude Fable 5 cost-effectively in Claude Code, delegate roles: Fable for overall design, Opus for heavy reasoning, Sonnet for execution and task management 125. Claude Sonnet 5 is supported across the Anthropic, Bedrock, Vertex, Claude Code, SAP AI Core, OpenRouter, and Vercel AI Gateway providers 367. Anthropic's free workshops and official courses provide a structured path to building autonomous agents with Claude 727,573.[7]
Anthropic's session on self-learning agents with memory and dreaming shows significant gains: Rakuten cut first-pass mistakes by 90%, Harvey saw 6x gain on legal benchmark 2327. Fable 5 is the best model for creative thinking, but GPT-5.6 Sol is almost as good and much cheaper, making it the better value 2315,2149,2059,1969. Sol on Medium is preferred over GPT-5.5 xhigh for coding work due to being an upgrade 2312,2140,1959,1636,1474,1701,1549,1408,1312. GPT 5.6 Sol is preferred over Claude Fable 5 for autonomous building, reviewing, and testing after initial decisions are made 2255,2219,2421. HyperFrames is the best free video editing tool, outperforming Claude Code-based tools 2620,2144. A 4-hour course on mastering Claude Code helped a non-MIT/Stanford applicant land a $750k offer from Anthropic 2520.[20]
OpenAI releases GPT-5.6 and integrates Codex into ChatGPT, with a developer AMA scheduled 1634,1472,1354,1310. GPT-5.6 Sol is preferred over Fable 5 for faster one-shot reasoning in Devin 1173. DeepSWE's GPT-5.6 models are preferred over Claude Code for coding agent benchmarks 1169. GPT-5.6's Computer Use update enables faster, more token-efficient multi-step tasks with batching and parallel operations 1063. GPT-Realtime-2.1-mini is now available in the API, bringing reasoning and tool use to the Realtime mini lineup at the same cost as GPT-Realtime-mini 694.[8]
Claude Code 2.1.19 contains a hidden TeamMateTool with leader-worker system, revealed by running strings on the binary, indicating a trend toward built-in multi-agent capabilities (1987). Claude Code is evolving skills from static instructions to self-improving software that is benchmarked, evaluated, and refined automatically (1976). Claude Tag, a new feature allowing Claude to be invoked in team conversations like Slack, signals a shift from developer tool to enterprise-wide AI assistant 514. Claude Code web's egress policy blocking GitHub is a new issue affecting workflows that rely on cloning repos 113.[4]
Vibe coded projects fail for the same reasons as hand-coded ones: lack of value and marketing, not scaling issues 1635,1550. Today's LLMs are less likely to default to building everything in React compared to last year, reducing the need to explicitly request no React 109. The 'understand to participate' framing is a useful way to describe the cognitive debt problem when working with coding agents 260. As AI agents take on longer-running work, engineering focus moves to setting direction, reviewing work, and designing better systems 104. Anthropic's model naming scheme (Haiku < Sonnet < Opus < Fable < Mythos) is ambiguous, with Fable's position relative to Opus uncertain 2142.[6]
Added new claims: self-learning agents with memory/dreaming (2327), Fable 5 vs GPT-5.6 Sol comparisons (2315, 2149, 2059, 1969), Sol on Medium preference (2312, 2140, 1959, 1636, 1474, 1701, 1549, 1408, 1312), GPT 5.6 Sol preferred over Fable 5 for autonomous work (2255, 2219, 2421), HyperFrames as best free video editing tool (2620, 2144), Claude Code course leading to $750k offer (2520), GPT-5.6 release and Codex integration (1634, 1472, 1354, 1310), GPT-5.6 Sol preferred in Devin (1173) and b