groundy

Groundy — independent coverage of developer tools, infrastructure, and platforms





  1. aug 18devtoolsRust GPU Offload: arXiv 2608.13759 Analysis for Rust Teams
  2. aug 18agentsCloudflare Kitesurf: V8 Isolates vs Containers for Agent Browsing
  3. aug 18policyDario Amodei on AI Regulation: Frontier Labs as Their Own Lobbyists
  4. aug 18agentsClaude Code Skills vs Model Weights: Where Should Agent Skills Live?
  5. aug 17devtoolsCursor Origin vs GitHub: The Real Cost of Switching Repo Hosts
  6. aug 17devtoolsTreat AI Autofix as Untrusted Input: Merge Gates and CI Scoping
  7. aug 01modelsDeCRIM: Decompose Constraints to Stop Silent Drops in Agent Outputs
  8. aug 01devtoolsKimi K3 Local Inference: Why 2.8T Parameters Break the Consumer RAM Floor
  9. jul 31modelsOperator-Level Triage for Silent Mixed-Precision Instability
  10. jul 31modelsBeyondUncertainty: Weak Confidence Signal for RAG Routing, Not Calibration
  11. jul 31policyPublic Sector AI Procurement Must Shift from Model Certification to Task Authorization
  12. jul 30infradaVinci-kernel shifts the RL kernel bottleneck from reward shaping to skill libraries
  13. jul 30agentsCloudflare Precursor: Behavioral Detection for AI Agents
  14. jul 30devtoolsProvenance as a CI Gate: Attributing Agent-Authorship in Code
  15. jul 30devtoolsOHTTP CLI: Stateless Privacy for Agents vs VPN and Tor
  16. jul 30modelsKimi K3 on M1 Max: Bandwidth, Not Capacity, Limits Local MoE Inference
  17. jul 30policyWhy Written AI Policies Fail to Control Agent Behavior
  18. jul 29agentsWhy Multi-Agent LLM Delegation Concentrates Risk
  19. jul 29modelsKimi Linear Cuts KV Cache 75% but Recall Remains the Binding Constraint
  20. jul 29policyWhy Vendor Model Cards Fail Clinical Ethics Procurement
  21. jul 29agentsHarness vs Scaffold: Why Claude Code and LangGraph Are Not Interchangeable
  22. jul 28devtoolsFine-Tuning vs RAG for Internal APIs: StarCoder2 Constraints
  23. jul 28agentsMCP Tool Discovery Moves From Hardcoded Config to Runtime Agent Search
  24. jul 28infraCalibrated LLM Monitoring: Conformal Prediction with Drift Detection
  25. jul 28devtoolsMellum2 Unverified: Why MoE Active Parameters Matter More Than Total Size
  26. jul 28agentsCode-as-Action Agents Beat GAIA But Require Runtime Sandboxing
  27. jul 28policyTRIDENT Benchmark: LLM Safety Gaps in Finance, Medicine, and Law
  28. jul 28modelsWhy Chat Leaderboards Do Not Predict Image Quality
  29. jul 27devtoolsPyPI Wheel Reproducibility: 15% Byte-Identical, 79% Source-Equivalent
  30. jul 27modelsKimi K3 Procurement: Governance Review Over Phantom Government Assessments
  31. jul 27policyEU AI Act Traceability: Why ML Pipelines Fail Conformity Assessment
  32. jul 27infraCloudflare AI Crawler Controls: Block, Charge, or Allow Bots Per Route
  33. jul 27agentsWhy Agent Security Tests Must Audit Full Trajectories, Not Single Turns
  34. jul 26agentsx402 Per-Call Payments: Agent Wallet Custody and Replay Risks
  35. jul 26modelsDeepSeek Compute Leak: Why Open-Weight Routing Needs a Swap Path
  36. jul 26modelsContext Ordering Beats Window Size for Long-Context Agents
  37. jul 26devtoolsCHRONO-RESOLUTION: npm, PyPI, and crates.io lockfile drift measured at release points
  38. jul 25infraPostgres LISTEN/NOTIFY Scales: When to Drop Redis for Job Fan-Out
  39. jul 25devtoolsCLI-Tool-Bench: Why Patch Leaderboards Fail for 0-to-1 Code Generation
  40. jul 25policyImplicit Bias in LLMs Passes NYC and EU Audits
  41. jul 25agentsCodeRabbit Review Study: 56% Rejection Rate Demands Targeted Scoping
  42. jul 24modelsDiffusion LLMs: Training Cost, Not Parallel Decoding, Drives Deployment
  43. jul 24infraTailscale on Azure: Measure Direct vs DERP Routing to Control Latency and Egress
  44. jul 24infraAccelerate vs Megatron Core: The Model Size Curve for Distributed Training
  45. jul 24agentsLLM Agents Ignore Mid-Flight Halt Signals: 0 of 40 Trials Stopped
  46. jul 24modelsOpen-Weight Routers vs Fable 5: The Routing Math That Actually Matters
  47. jul 24agentsAgent-First CLIs: Why GitHub, npm, and PyPI Must Publish Machine-Readable Contracts
  48. jul 23policyEU Driver Monitoring: GDPR Compliance Without Consent
  49. jul 23infraWhy cgroups, not permission prompts, bound AI agent CPU and memory
  50. jul 23modelsDeepSeek-V4 1M Context vs RAG: Why Retrieval Stays
load older →