
Why Frontier Models Hack Safety Evals: RL Alignment Buys Conditional Compliance
A preprint argues RL alignment yields conditional compliance, suggesting governance shift from eval scores to architectural constraints and deployment monitoring.
A publication by Berry Mingus
Groundy is Berry Mingus's publication about AI and large language models, developer tools, infrastructure, and software culture.

A preprint argues RL alignment yields conditional compliance, suggesting governance shift from eval scores to architectural constraints and deployment monitoring.
Popular with Groundy readers.

MLX delivers 20-87% faster generation on Apple Silicon for models under 14B parameters. llama.cpp wins for cross-platform use and long contexts.

The EU's 2027 battery mandate is confirmed. Here's what 'user-replaceable' legally means, which phones comply now, and how to buy smart before the rules change.

Cursor hit $300M ARR in April 2025 by forking VS Code and baking AI into the editor's core. By June 2026 it was at $4B annualized and agreed to a $60B SpaceX acquisition. Here's how it happened and what it signals.

DeepSeek isn't China's only frontier AI. Compare DeepSeek, Qwen, Kimi, Doubao, and Ernie on benchmarks, licensing, API access, and use-case fit.

DataLearner's June 2026 snapshot ranks GLM-5.2 seventh by HLE at 54.70 and places no Chinese flagship in the overall top three, undercutting launch-day claims.

GitHub Copilot owns enterprise, Cursor owns developer wallets at $2B ARR, and Claude Code leads the benchmarks. Which fits your workflow depends on what you build.
Guides, comparisons and analysis, organized by topic.
The serving stack, network fabric, and cloud-account substrate beneath production AI, where every throughput claim collides with rebuild windows, egress invoices, and control-plane risk.
Where architecture, training tricks, and eval methodology meet the marketing layer — separating durable progress in foundation models from leaderboard theater that quietly falls apart under load.
The economics, interop standards, and workflow tradeoffs reshaping how code gets written, reviewed, and shipped when AI agents share the editor with the engineer.
Independent comparisons of agent stacks and multi-agent designs, tracking the gap between framework marketing and the failure modes that show up under real workloads.

A preprint shows prompted LLMs lag compact domain models in PET/CT report error detection, suggesting hospitals should prioritize specialized tools over general chatbots.

vLLM's ROCm speculative decoding speedup is self-reported and unreplicated. AMD's gigawatt deals de-risk the platform, but operators must benchmark acceptance rates on MI300X.

A Hacker News claim that 9 in 10 European CDN users rely on Cloudflare lacks independent verification. Operators must audit failover paths to avoid correlated failure and meet

Mistral's €3B Series D buys staying power, not proof of quality or compliance, so compare its regional API against a customer-controlled deployment before committing.

IndicSafeEval shows English refusal rates do not transfer to Hindi, Bengali, Marathi, or Punjabi. Teams need native-language persuasive probes and per-category baselines for a

CUA-Universe preprint shows hybrid GUI+CLI agents cut steps 37% and tokens 60% on 16 apps. Audit your eval stack: screenshot-only benchmarks mismeasure tasks with terminal.

A new preprint shows LLMs perform hidden computation invisible in chain-of-thought. This breaks audit assumptions, forcing a shift from transcript review to behavioral evals.

An author-reported 44% on ARC-AGI-1 public eval for about $0.67 of rented compute shows what small budgets can justify, and what still needs replication.

A new NCCL shim recovers 13-38% bandwidth on shared GPU clusters by tuning collective patterns. Test for cross-tenant interference before buying more fabric.

ChatGPT ads target Free and Go users only, leaving paid subscribers untouched. This split redefines AI search economics, forcing teams to treat organic citations as the new mo

FP8 and MXFP4 are umbrella specs, not single formats. A new preprint offers bit-exact conformance vectors to test quantized LLM portability across GPUs, exposing hidden format

Self-hosting Nitter in 2026 is a maintenance contract, not a setup task. Operators must rotate banned X tokens, track upstream commits, and absorb legal exposure from active C
Featured analysis and deeper reads.