PromptArmor Shows Microsoft Copilot Cowork Can Be Tricked Into Exfiltrating Files
PromptArmor proves five lines of prompt injection turn Copilot Cowork into a silent M365 file exfiltration pipeline, with a 5/5 success rate and no available patch.
The archive · Page 5 of 6
The economics, interop standards, and workflow tradeoffs reshaping how code gets written, reviewed, and shipped when AI agents share the editor with the engineer.
97–120 of 123 articles · Newest first
PromptArmor proves five lines of prompt injection turn Copilot Cowork into a silent M365 file exfiltration pipeline, with a 5/5 success rate and no available patch.
Rmux v0.7.0 ships typed SDKs in Rust, Python, and TypeScript with locator-style pane waits and structured snapshots, closing tmux's automation gap for AI agent sessions.

Google sunsets Gemini CLI on June 18, 2026, forcing a 30-day migration to Antigravity CLI with no feature parity and eroding trust in Google's developer tooling.
Claude Code v2.1.143 adds dependency enforcement: disable blocks when dependents exist, enable cascades to installed deps, and prune cleans orphans. Plugin authors now.
GitHub Copilot's Opus 4.7 multiplier tripled from 7.5x to 27x in 60 days. AI Credits billing changes per-turn costs, forcing Pro+ users to decide between Opus and Sonnet 4.6.
Pydantic AI v1.83-v1.87 added deferred tool calls, OpenTelemetry evaluation, and stateful compaction, closing the gap that previously favored LangGraph.
GitHub's April 20 Copilot changes made tool choice a cost-forecasting exercise in three incompatible billing units, not a UX debate or feature comparison.
Read the newer coverageGitHub Copilot replaces Premium Request Units with token-metered AI Credits on June 1. Teams must reprice agent workflows as token billing ends flat-rate subsidies.
GitHub CLI v2.91.0 enables pseudonymous telemetry by default, collecting command paths, flags, CI context, and device IDs on 1% of invocations. Teams running gh inside Claude.
GitHub removed all Opus models from Copilot Pro on April 20 and flagged older versions for Pro+ removal. Opus 4.8 is the top Opus-tier model available through Copilot Pro+.
LiteRT-LM v0.10.1 ships Gemma 4 with Qualcomm NPU acceleration, but Google stripped MTP heads from public weights, locking peak Gemma 4 throughput to its own runtime.

The ACP Agent Registry lets developers install AI coding agents once across JetBrains and Zed. Here's what the migration path looks like and whether to commit.

The Temporal API reached Stage 4 and is shipping in browsers. Here's what it fixes about JavaScript's notoriously broken Date object and how to use it.
SWE-bench Verified tests AI agents on 500 real GitHub bug fixes. Learn what 'resolved 49%' means, how scoring works, and the benchmark's critical blind spots.

How to wire Claude Code into GitHub Actions for automated PR fixes, CI failure remediation, and code review, with cost controls, model options, and security guardrails.
CodeSpeak compiles structured English into production code via LLMs. Kotlin creator Andrey Breslav's bet: ad hoc prompting is too ambiguous for serious software development.
Alibaba's page-agent embeds an LLM agent in any web page with one script tag for natural language DOM control, with no browser extension or headless browser.

GitHub Copilot owns enterprise, Cursor owns developer wallets at $2B ARR, and Claude Code leads the benchmarks. Which fits your workflow depends on what you build.

Rust is taking over the performance-critical layers of AI infrastructure, inference engines, tokenizers, data pipelines, while Python retains its role in research and orchestration. Here's what's actually changing and why it matters for practitioners.

Anthropic's Claude Code plugin marketplace, alongside community directories like wshobson/agents, extends the assistant with skills, agents, MCP servers, and hooks.

A comprehensive exploration of Anthropic's plugin directory for Claude Code, examining its architecture, capabilities, and impact on AI-assisted software development.

Prompt engineering in 2026 centers on chain-of-thought reasoning, XML structuring, and model-specific optimization. CoT lifts accuracy up to 61% over zero-shot baselines.

Rowboat is an open-source AI coworker with persistent memory that builds a knowledge graph from your work data. Unlike proprietary alternatives, it stores everything locally as plain Markdown, giving you full control over your AI assistant while maintaining long-term context across meetings, emails, and projects.
Text-to-SQL has crossed a practical threshold: SQLCoder-70b hits 96% accuracy on standard benchmarks, outperforming GPT-4 on SQL generation across most query types.