New Features & Tech Innovations
-
Agents
Mastra Factory beta makes coding-agent autonomy a stage-by-stage decision
Mastra Factory enters beta with configurable human and agent gates across issue intake, planning, coding and review, plus self-hosting options and caveats.
-
Agents
GitHub makes Copilot agent permissions enforceable above local settings
GitHub adds enterprise-managed deny, ask and allow rules for Copilot agent operations that local auto-approval and saved grants cannot weaken.
-
Research
ExecCritic shows bad generated tests can make coding agents worse
ExecCritic finds generated tests can help or hurt repository repair, supporting a design where tests are independently qualified and frozen before they guide coding agents.
-
Agents
Zscaler puts AI agents on the path to automatic threat containment
Zscaler's Agentic SOC links AI investigation to automated security controls, making authorization, audit trails and rollback central to deployment.
-
New Products
ChatGPT Images 2.5 turns generation into a multi-turn production workflow
OpenAI's Images 2.5 adds Sketch, comments, templates and Flare/Sunburst API models, pushing image generation toward a stateful, versioned production workflow.
-
Research
OpenAI’s Navier–Stokes claim comes with a Lean proof, but acceptance is a separate gate
OpenAI has published a proposed Navier–Stokes Millennium Problem solution with a formal Lean proof; the result remains subject to independent mathematical scrutiny.
-
Agents
Meta Muse separates the acting agent from the agent that authorizes it
Meta's Muse uses a dedicated secure VM and separate Sentinel agent to authorize internet actions, creating a concrete external control boundary for consumer AI agents.
-
Agents
Arm AI Portal makes hardware optimization callable from coding agents
Arm AI Portal exposes optimized models, performance data and deployment guidance through MCP, making Arm-specific AI optimization directly accessible to coding agents.
-
Open Source
SwarmLLM splits a 27B model across browser tabs
SwarmLLM v0.2.0 splits a 27B model across browser tabs with WebGPU and WebRTC, trading single-device memory limits for network and peer-trust constraints.
-
Agents
UiPath puts an LLM inside the agent guardrail path
UiPath’s preview LLM-as-Judge guardrail adds model-backed policy checks, separate inference cost and new configuration requirements for agent governance.