groundy
articlessearch

Groundy — independent coverage of developer tools, infrastructure, and platforms

devtools
Two graphite-shaded folded paper sculptures stand on warm ivory paper with matching copper pointed tips. The left has many narrow folds; the right has broad angular planes.

4-Bit vs 8-Bit Quants: Why Accuracy Benchmarks Miss Distribution Drift

A preprint argues zero-shot accuracy misses distribution drift in quantized LLMs, recommending divergence metrics like JSD and TV against BF16 bases for safer deployment.

9 min
agents

Do More Tools Make LLM Agents Worse? When to Gate External Evidence

infra

Running MoE LLMs on a Single GPU: What Expert Offloading Actually Costs

policy

Can General LLMs Catch Radiology Report Errors Before Clinicians Sign Off?

  1. infraCloudflare Runs Nearly 9 in 10 European CDNs: The Concentration Problem
  2. industryMistral's €3B Raise and the Real Cost of Sovereign AI for European Enterprises
  3. policyWhy English-Only LLM Red Teaming Misses Indic-Language Jailbreaks
  4. policyWhy Chain-of-Thought Monitoring Misses Hidden LLM Reasoning
  5. modelsmdlARC's 44% ARC-AGI-1 Claim: What the Small Budget Leaves Out
  6. infraNoisy Neighbors at the Fabric: Why Shared GPU Clusters Throttle Your Jobs
  7. industryChatGPT Ads Change the Economics of AI Search Visibility
  8. modelsFP8 vs MXFP4 vs BF16: Why Your Quantized LLM Disagrees Across GPUs
  9. infraRunning Your Own Nitter Instance: What the Post-Takedown Comeback Requires
  10. devtoolsClaude Code vs TERMy: When Terminal Help Needs No Model
  11. infraRunning Local LLMs on a $60 Used GPU: What AMD's BC-250 Can and Can't Do
  12. infraPrivate Vector Search vs TEEs: Can RAG Retrieval Be Outsourced Safely?
  13. devtoolsDo LLMs Spread Reasoning Like a Virus? Code Review in a Monoculture
  14. devtoolsGPT-5.6 Is Now Microsoft 365 Copilot's Default: Seat Budgets Can't Assume a Stable Model
  15. modelsRunning a 104GB LLM on a 48GB Mac: What Expert Streaming Costs
  16. infraRunning LLMs in the Browser: Can Privacy Be Verified Instead of Promised?
  17. infraDoes Cloudflare's Adaptive Intelligence Change the Bot Defense Build-vs-Buy Math?
  18. policySafety RL Can Backfire: Why the Training Environment Decides the Direction
  19. devtoolsDo LLM Users Get Better With Practice? What Longitudinal Chat Logs Show
  20. industryOpenAI Caps Microsoft Revenue Share: The Azure Buyer's Renegotiation Guide
  21. modelsGPT-6 Astra on ARC-AGI-3: What the Agentic Score Actually Measures
  22. devtoolsRunning the React Compiler in Vite: Memoization Without useMemo
  23. devtoolsTcl/Tk vs Electron for Internal Tools: GUIs Without a Browser Engine
  24. infraCloudflare Compresses Its Cache With Zstandard: The Storage-vs-CPU Trade
browse all 778 articles →