Coding agents
-
Research
ExecCritic shows bad generated tests can make coding agents worse
ExecCritic finds generated tests can help or hurt repository repair, supporting a design where tests are independently qualified and frozen before they guide coding agents.
-
Agents
Arm AI Portal makes hardware optimization callable from coding agents
Arm AI Portal exposes optimized models, performance data and deployment guidance through MCP, making Arm-specific AI optimization directly accessible to coding agents.
-
Agents
VS Code Agent Merge turns PR blockers into a continuous agent loop
VS Code 1.136 lets agents repeatedly fix PR reviews, CI failures and conflicts, raising a new need for independent merge controls.
-
Agents
GitHub extends Copilot content exclusions to app and CLI, but gaps remain
Copilot app and CLI now respect content exclusions, while GitHub still documents gaps in IDE agent modes, symlinks and remote filesystems.
-
Agents
OpenAI sets a November cutoff for its models inside Cursor
OpenAI plans to end model supply to Cursor after SpaceX's acquisition, creating a November 12 migration deadline for developer workflows.
-
Research
openJiuwen turns the coding-agent harness into a measurable systems layer
openJiuwen proposes a composable, runtime-adaptive coding-agent harness and shows why model-matched benchmark audits matter for agent engineering.