Agents
-
Agents
DeepSeek Harness 0.1.7 turns agent configuration into a runtime plugin surface
DeepSeek Harness 0.1.7 makes agent capabilities more dynamically composable. Aipolix examines the benefits and the new operational risks.
-
New Models
GPT-6 Sol and Luna make model routing a first-order agent economics decision
GPT-6 Sol and Luna introduce a 20x price spread, large context and different cost boundaries. Aipolix examines agent routing economics.
-
New Models
Claude Opus 5.5 cuts agent costs, but safeguard routing changes how teams should evaluate it
Claude Opus 5.5 cuts token and cache costs while sensitive requests may use fallback models. Aipolix examines the impact on agent evaluation.
-
New Models
Grok 4.7 keeps token prices flat, but agent economics depend on the full task
Grok 4.7 keeps basic API token rates unchanged, but independent tests show higher output consumption and provider-specific long-context pricing.
-
Agents
Plugin4Shell exposes a gap in coding-agent plugin checks; patched clients still need an integrity audit
What Plugin4Shell changes about Git-pinned plugins in Claude Code, Codex, Copilot and Gemini CLI, with qualified patch status and an audit plan.
-
Agents
Qwen Code 0.24.1 gives workflow agents narrower tool access, while browser support remains unfinished
Qwen Code 0.24.1 ships per-subagent tool allowlists and a browser SDK. What is usable now, what remains unfinished, and how developers can verify the boundaries.
-
Agents
VS Code 1.138 brings coding agents into Dev Containers, but permissions remain separate
Microsoft adds Dev Container agent sessions and cross-app Codex continuity in VS Code 1.138. Aipolix examines rollout limits and the separate security boundaries.
-
AI Governance
OpenAI releases six misalignment reports, but disclosure is not a safety gate
OpenAI discloses six model misalignment cases involving handoff instructions, external data transfers and agent coordination, while warning that the sample is not representative.
-
New Models
Gemini 3.8 Live separates a spoken turn from the work behind it
Gemini 3.8 Live Extended Thinking adds background reasoning and async tools, requiring voice-agent clients to track interaction state beyond spoken turns.
-
🧾 Traceability: Who (or What) Wrote This Line of Code? (7/10)
How to audit AI-generated code: intent-to-effect provenance, signed prompts and models, structured authorship, correlation IDs and evidence across system boundaries.