For Developers
-
Agents
Mastra Factory beta makes coding-agent autonomy a stage-by-stage decision
Mastra Factory enters beta with configurable human and agent gates across issue intake, planning, coding and review, plus self-hosting options and caveats.
-
Agents
GitHub makes Copilot agent permissions enforceable above local settings
GitHub adds enterprise-managed deny, ask and allow rules for Copilot agent operations that local auto-approval and saved grants cannot weaken.
-
Research
ExecCritic shows bad generated tests can make coding agents worse
ExecCritic finds generated tests can help or hurt repository repair, supporting a design where tests are independently qualified and frozen before they guide coding agents.
-
Agents
Arm AI Portal makes hardware optimization callable from coding agents
Arm AI Portal exposes optimized models, performance data and deployment guidance through MCP, making Arm-specific AI optimization directly accessible to coding agents.
-
Agents
VS Code Agent Merge turns PR blockers into a continuous agent loop
VS Code 1.136 lets agents repeatedly fix PR reviews, CI failures and conflicts, raising a new need for independent merge controls.
-
Research
Context privilege escalation turns agent memory into a security boundary
A study across 12 agent harnesses shows how context can gain authority across roles and scopes, making provenance-aware non-escalation controls a runtime requirement.
-
Research
CordisBench shows where agent harnesses should replace reasoning with verification
CordisBench finds that models struggle with growing harness lifecycle interactions even when software can compute the tested state consequences exactly.
-
Agents
GitHub lets Copilot approvals count toward merge requirements
Copilot code review can now count as a required pull-request approval, with enterprise, repository and path controls around where AI approval applies.
-
Agents
GitHub extends Copilot content exclusions to app and CLI, but gaps remain
Copilot app and CLI now respect content exclusions, while GitHub still documents gaps in IDE agent modes, symlinks and remote filesystems.
-
Research
HarnessDev shows why self-improving agents need release gates
HarnessDev shows why evolving AI agent harnesses need hidden-task validation, executor checks, runtime evidence and rollback before promotion.