Tag: coding-agents
All the articles with the tag "coding-agents".
- AI Signals
Kimi K3 SWE Marathon: 5 Tests Before You Switch
Updated:Kimi K3 is released, but SWE Marathon scores are only one signal. Test repo quality, tool calls, context compression, subagents, and privacy before migrating.
- AI & Games
Phaser Game Agent Review: Build a Browser Game From One Prompt
Updated:A developer-focused review of Phaser Game Agent, its MCP setup, generated code and assets, sandbox architecture, pricing, and practical limitations.
- AI Signals
Kimi K3 Is Here: What the 2.8T Open Model Actually Changes
Moonshot's Kimi K3 combines 2.8T parameters, sparse MoE routing, native vision, and a 1M context window. Agent reliability and serving cost remain open.
- AI Tools & Benchmarks
How to Evaluate Coding Agents: Productivity, Quality, Cost, and Risk
A field-tested scorecard for comparing coding agents with real repository tasks, durable quality metrics, total cost, and controlled rollout evidence.
- AI Engineering
GhostApproval and Friendly Fire Expose the Trust Problem in AI Coding Agents
Two new attack patterns show how malicious repositories can turn coding agents, approval dialogs, and automated security reviews against developers.
- AI Tools & Benchmarks
Microsoft's CLI Coding-Agent Study Found 24% More Merged PRs. Here Is the Catch
A Microsoft field study links Claude Code and Copilot CLI adoption to more merged pull requests, but the result is narrower than a productivity claim.
- AI Signals
AI Week in Review: GPT-5.6, J-Space, Coding Agents, and Cheaper Intelligence
A developer-focused briefing on the week's most consequential AI releases, research findings, and production engineering signals.
- AI Tools & Benchmarks
GPT-5.6 Has 72 Configurations: The Cheapest Model May Be the Better Default
Model size alone no longer predicts value. Here is how to evaluate GPT-5.6 tier, reasoning effort, speed, and agent cost together.