articles
all articles
feed
- may 29agentsMulti-Agent LLM Coordination: Why Attention Steering Beats Full Broadcast
- may 29modelsPersona Prompts Change Who an LLM Recommends as an Expert
- may 29agentsDataClawBench: AI Agents Fail at Exploratory Financial Analysis Across 492 Tasks
- may 29policyRLHF Can Be Exploited to Optimize the Biases It Was Built to Suppress
- may 28ossFrontier AI Has Broken Open CTFs: Why Claude Code Now One-Shots Medium Pwn Challenges
- may 28policySelective Geometry Attacks Bypass LLM Safety Alignment, New arXiv Paper Reports
- may 28modelsOpus 4.8 Batch API: 1M Context, 300k Output, and Team Cost Controls
- may 28agentsClaude Code Dynamic Workflows: Spawning 100 Parallel Subagents on Opus 4.8
- may 27infraWhy LLMs Still Botch Kubernetes Manifests: The Training-Data Gap
- may 27securityOpenAI's New Safety Bug Bounty Pays Researchers for Jailbreaks and Policy Bypasses
- may 27industryHuggingFace's $100M Series C Bets Open-Source AI Can Outlast Per-Token Pricing Wars
- may 27infraGemma 4 31B on Cloud TPU vs GPU: The Serving Cost Crossover Point
- may 27agentsClaude Code, Cursor, Copilot: How Agentic Coding Assistants Get Weaponized as Attacker Shells
- may 26devtoolsBun Rewrites Its Core From Zig to Rust, Putting Downstream Zig Bindings at Risk
- may 26infraObjectCache Moves KV Reuse to S3-Class Storage: Why Layerwise Retrieval Beats Full-Prefix Cache Hits
- may 26devtoolsPromptArmor Shows Microsoft Copilot Cowork Can Be Tricked Into Exfiltrating Files
- may 25modelsμP Hyperparameter Transfer Has an Embedding Layer Hole, New arXiv Paper Says
- may 25devtoolsRmux Brings a Playwright SDK to tmux Sessions for Agent Automation Workflows
- may 25ossColorado SB051 Carves Out Open Source From Age Verification After Maintainer Backlash
- may 24cultureUS Researchers Hit With New Federal Limits on Publishing With Foreign Collaborators
- may 23securityAI Jailbreaks Are Now a Reasoning Problem, Not a Prompt Problem
- may 23devtoolsGoogle Sunsets Gemini CLI on June 18: Forced Migration to Antigravity CLI Breaks Existing Automation
- may 23agentsSpecBench Exposes Reward Hacking in Long-Horizon Coding Agents
- may 23infravLLM 0.21 Makes Prefill-Decode Disaggregation Actually Practical
- may 19industryBret Taylor's Sierra Raises $950M at $15B, Claims 40% of Fortune 50 Use Its Agents
- may 18industrySierra Raises $950M at $15B, Locking 40% of the Fortune 50 Into Its Agent Platform Before the Labs Go Direct
- may 18policyFrontier AI Has Broken the Open CTF Format: What the Scoreboard Collapse Means for Security Training
- may 18securityTrustFall: One Keypress in Claude Code, Gemini CLI, Cursor, and Copilot CLI Triggers Unsandboxed RCE
- may 18devtoolsClaude Code Adds Plugin Dependency Enforcement: disable Now Refuses to Break Transitive Chains
- may 18industryPayPal's $1.5B AI Overhaul Cuts 4,760 Jobs and Reframes Layoffs as Capex
- may 18policyFrontier AI Broke Open CTFs: What Hack The Box and BearcatCTF 2026 Results Mean for Security Hiring Signals
- may 18industryOpenAI's $4B Deployment Company Buys Tomoro and Signs 19 Partners to Own Implementation
- may 18agentsLangGraph 1.2.0 Makes Error-Handler Resume Crash-Durable: With Conditions
- may 18agentsCrewAI vs AutoGen vs LangGraph 2026: The Real Trade-Off After Maintenance Mode
- may 18cultureApple's $250M Siri Settlement: iPhone 16 Buyers Get $25 to $95 for Undelivered AI
- may 18securityMultiBreak Benchmark: 10,389 Multi-Turn Jailbreak Prompts Raise ASR 54pp on DeepSeek-R1-7B
- may 18industryAnthropic's $1.5B Joint Venture With Goldman Sachs and Blackstone Sends Claude Into PE Portfolio Companies
- may 18industryOpenAI Offers Two Months of Free Codex to Enterprises Switching From Claude Within 30 Days
- may 18cultureAB 566 Forces Chrome and Safari to Ship Opt-Out Signals by 2027. It Shields Them from Google's 86% GPC Failure
- may 18ossBrowserAct Open-Sources Stealth Browser Engine with 93% Token Reduction Claim
- may 18devtoolsGitHub Copilot's Opus 4.7 Multiplier: 7.5x to 15x to 27x in 60 Days
- apr 29osspgBackRest Is No Longer Maintained: PostgreSQL Backup Alternatives After the Project Stalls
- apr 29securityInstructLab CVE-2026-6859: Hardcoded trust_remote_code=True Turns Any HuggingFace Model Into RCE
- apr 29devtoolsPydantic AI v1.87 Closes the LangGraph Gap: Deferred Tool Calls, OpenTelemetry Eval, Stateful Compaction
- apr 29securityMercor's 4TB Lapsus$ Breach Hands Voice-Clone Attackers 40,000 Pre-Verified Targets
- apr 29agentsCouncil Mode Cuts Multi-Agent LLM Hallucination 35.9% at 4.2x Token Cost on HaluEval
- apr 29devtoolsClaude Code vs Cursor vs Copilot After the April 2026 Reshuffle: How the Comparison Math Changed
- apr 28devtoolsGitHub Copilot Replaces Premium Request Units With Token-Metered AI Credits on June 1
- apr 28ossfree-claude-code Routes Claude Code Through NVIDIA NIM and Local Models After Anthropic's CLI Ban
- apr 28industryMicrosoft and OpenAI End Their Exclusive Revenue-Sharing Deal: What It Means for Azure's AI Moat
- apr 28industryAnthropic Ends Flat-Fee Enterprise Claude, Enforces Per-Token Billing
- apr 24devtoolsGitHub CLI v2.91.0 Turns On Default Telemetry: What gh Collects and How to Opt Out in CI and Agent Pipelines
- apr 24devtoolsGitHub Copilot Drops Opus from Pro and Pauses Signups: The Forced Migration Facing Agentic Workflows
- apr 24agentsCloudflare Agents Week Moved Sandbox Execution, Private Networking, and Memory to Network Primitives
- apr 24ossInside Rowboat's Knowledge Graph: Why an Obsidian-Compatible Vault Sidesteps Vector DBs for Personal AI Memory
- apr 24securityCitizen Lab's 'Bad Connection' Names Three Telecom Entry Points, Shows Diameter Silently Falls Back to SS7
- apr 23agentsDiversity Collapse in Multi-Agent LLM Systems: Structural Coupling, Not Topology, Breaks Open-Ended Ideation
- apr 23devtoolsLiteRT-LM v0.10.1 Ships Gemma 4 MTP Heads That llama.cpp Can't Access
- apr 23ossHugging Face's Spring 2026 Report: China 41% of Downloads, Industry Share Collapses From 70% to 37%
- apr 23modelsQwen3.6-27B's Dense Architecture Challenges the MoE-Only Playbook for Flagship-Class Coding Models
- apr 23securityMarch-April MCP CVEs Expose the Local-Host Trust Model in AI Agent Frameworks
- apr 21cultureEU's 2027 Replaceable Battery Mandate: What It Means for Phone Buyers and Repairers Right Now
- apr 20devtoolsACP Registry Is Live: Zed and JetBrains Just Did for AI Agents What LSP Did for Language Servers
- apr 20policyAtlassian Turned On AI Training Data Collection by Default: Here's What to Disable
- apr 20ossGitHub CLI's `gh skill` Command: One Standard to Rule Claude Code, Copilot, Cursor, and Gemini
- mar 27infraOpenRAG: The Open-Source RAG Platform Challenging Pinecone
- mar 27devtoolsJavaScript's Date Problem Is Finally Fixed: The Temporal API After 9 Years
- mar 27agentsInsForge: The Backend Framework Built for Agentic Applications
- mar 27policyThe AI Grief Split: When Emotional Bonds with Language Models Break
- mar 24infraMLX vs llama.cpp on Apple Silicon: Which Runtime to Use for Local LLM Inference
- mar 24infraPrefill-Decode Disaggregation: The Architecture Shift Redefining LLM Serving
- mar 24devtoolsSWE-bench Verified Explained: What the Coding Agent Leaderboard Actually Measures (and What It Misses)
- mar 24modelsChinese AI Models Compared: DeepSeek, Qwen, Kimi, Doubao, and Ernie
- mar 24devtoolsClaude Code in GitHub Actions: A Complete Guide to Automated PR Fixes
- mar 24modelsRunning DeepSeek R1 Locally: Hardware Requirements, Quantization, and Real Throughput
- mar 15devtoolsJetBrains' New Language Lets You Talk to LLMs in Specs, Not English
- mar 15modelsFish-Speech: The Open-Source TTS Model That's Threatening ElevenLabs
- mar 15infraGoogle LiteRT: Running LLMs on Your Phone Without the Cloud
- mar 15devtoolsAlibaba's Page-Agent: Control Any Website With Natural Language
- mar 15cultureAI Diagnostics in 2026: Where Machines Now Outperform Radiologists
- mar 15agentsAI Agents That Actually Learn: The Architecture Behind Hindsight Memory
- mar 14devtoolsGitHub Copilot vs Cursor vs Claude Code: The 2026 AI Coding Showdown
- mar 14policyDetecting AI Content in 2026: The Arms Race Nobody Is Winning
- mar 13infraMicrosoft's BitNet: How 1-Bit LLMs Could Make GPU Farms Obsolete
- feb 28infraWebAssembly AI: Running Models in the Browser
- feb 28agentsSuperpowers: The Agentic Framework Replacing Your Dev Process
- feb 27modelsSynthetic Data Is Eating AI Training
- feb 27devtoolsRust Is Quietly Replacing Python in AI Infrastructure
- feb 27industryOpenAI's For-Profit Pivot: What the PBC Restructuring Means for AI
- feb 27agentsHow AI Agents Remember: Memory Architectures That Work
- feb 27modelsGoogle's TimesFM: A Foundation Model for Time Series
- feb 27modelsGemini 2.0 Pro's 2 Million Token Context: What Can You Actually Do With It?
- feb 27industryCursor's Meteoric Rise: Inside the AI Editor Hitting $300M ARR
- feb 27industryStargate: Inside OpenAI's $100B Infrastructure Buildout
- feb 27modelsDeepSeek V3/R1: How Chinese Engineers Matched GPT-4 for $6 Million
- feb 27modelsClaude's Web Search Changes Everything for AI Research
- feb 27modelsThe Million-Token Context Window: What Can You Actually Do?
- feb 21ossKeep Android Open: F-Droid's Fight Against a Locked-Down Mobile Future
- feb 21devtoolsClaude Code Plugins: Anthropic's Official Plugin Ecosystem Explained
- feb 21devtoolsClaude Code Plugins: Anthropic's Official Extension Ecosystem