groundy

Groundy — independent coverage of developer tools, infrastructure, and platforms

models

Detecting AI-Generated Audio: Why Decay Tails Betray Voice Clones

A new preprint shows AI audio leaks in decay tails via group delay. It offers a cheap, watermark-free filter for fraud screening, though hold-out accuracy is only 66.7%.

6 min
devtools

How Cursor Uses GPT-5: Where Your Prompts Actually Go

industry

OpenAI's Cursor Decision Is a Vendor Exit Drill for AI Coding Tools

models

Reward Hacking Starts in the Verifier: Rule Checks vs LLM Judges for Math RL

  1. policyHugging Face's AI Action Plan Reply: Open Weights vs the Frontier Lab Lobby
  2. devtoolsDiagnosing LLM Prompt Injection Detectors Before You Gate an Agent on Them
  3. infraAlibaba's ScaleSense: When Learned Autoscaling Beats Provisioning Rules
  4. policyWhy Temperature 0 Won't Save Your Financial AI Audit
  5. infraAWS Cognito Postmortem: The Real Cost of Free Managed Auth
  6. modelsCan You Serve LLMs on 2-Bit Weights? What Ultra-Low-Bit Quantization Costs
  7. devtoolsLocal Coding LLMs Hallucinate Packages: Slopsquatting Defenses Compared
  8. policyPayPal Blocks GrapheneOS: A Payment Contingency Guide for Open Source
  9. infraHow Cloudflare Saved 100 TB in 1.1.1.1's DNS Cache and What Operators Can Copy
  10. devtoolsTesting Cloud APIs Without a Cloud Account: LocalStack vs Emulator Synthesis
  11. modelsDo LLMs Still Need BPE? What RL-Trained Tokenizers Change
  12. policyDoes the EU AI Act Exempt Open-Weight Models Like Qwen and Llama?
  13. agentsQwen-Agent Stretches 8k to 1M Context: Do You Need a Long-Context Model?
  14. industryLLM Uncertainty Methods Compared: What Actually Catches Hallucinations
  15. infraWhere Simple RAG Breaks: Multi-Document QA Needs Hierarchy, Not More Chunks
  16. agentsCan AI Agents Do Root Cause Analysis? What Cloud-OpsBench Measures
  17. agentsWhich AI Agents Behave Badly: Cloudflare's Agentic Internet Data
  18. policyAI Agents Outgrow OAuth: What Task-Scoped Authorization Requires
  19. modelsWhy Prompt Caching Can Change Model Outputs: A Prefix Invariance Audit
  20. industryAI Video Generators as World Simulators: What VGI-Bench Actually Measures
  21. industryLLMs Reading Earnings Filings: Where KPI Extraction Still Fails
  22. devtoolsPDF Parsers vs Math Formulas: Building RAG Over Scientific Papers
  23. modelsGDPR Deletion Requests vs LLM Weights: What Machine Unlearning Actually Removes
  24. policyAI Bias Audits vs Ethics Audits: What Each Actually Catches
browse all 735 articles →