groundy

infrastructure & runtime

  1. jun 26infraVercel CDN Cache Tags vs Path Purging: When Tag Invalidation Wins
  2. jun 26infraPrisma Joins the Vercel Marketplace: The ORM Becomes the Database Vendor
  3. jun 26infraOpenAI on AWS Bedrock: Routing Math to Run Before You Move Traffic
  4. jun 25infraVercel's Function Observability: What Native Metrics Replace and What They Don't
  5. jun 25infraAWS Databases on the Vercel Marketplace: The Cross-Cloud Latency Tax
  6. jun 25infraTurso on the Vercel Marketplace: Edge SQLite vs the Serverless Connection Pool
  7. jun 24infraVercel on the AWS Marketplace: What the Listing Does to Procurement and Lock-In
  8. jun 24infraServing Cold MoE Models: CrossPool Disaggregates KV Cache and Weights
  9. jun 24infraVercel's In-Function Concurrency: What It Does to Cold Starts and Billing
  10. jun 24infraPoisoning a RAG Retriever: How Conflict-Aware Edits Inject False Knowledge
  11. jun 24infraVercel Raised Its CDN Origin Timeout to Two Minutes: What Breaks First
  12. jun 24infraGradio-Lite Runs Model Inference in the Browser via Pyodide, No Server
  13. jun 24infraCloudflare AI Gateway Adds Spend Limits to Cap the Runaway Inference Bill
  14. jun 24infraVercel Now Honors stale-if-error: Serving Stale Cache When the Origin Dies
  15. jun 23infraVercel's Manual CDN Purge API: Cache Control Without a Redeploy
  16. jun 23infraCloudflare Now Routes Public Traffic to Private Apps via DNS, No VPN
  17. jun 23infraGitHub's AI Capacity Crunch Pushes Microsoft to Rent AWS Compute
  18. jun 21infraCloudflare's Temporary Accounts Give AI Agents Disposable Credentials
  19. jun 21infraRunning Long-Context Agents on a 4-Bit KV Cache: Where Accuracy Breaks
  20. jun 20infraWhen LLM-Generated CUDA Kernels Pass Tests but Get the Math Wrong
  21. jun 19infraRunning GLM-5.2 at Home: SGLang, vLLM, Transformers, and KTransformers Setup Guide
  22. jun 16infraAWS Bedrock Now Requires Data Sharing for Mythos: The Self-Hosting Calculus
  23. jun 15infravLLM Cold Start Latency: Why Scale-to-Zero LLM Serving Stalls
  24. jun 15infraThe Vercel-AWS Deal Reveals Where AI Inference Runs
  25. jun 11infraRunning RAG on a Snapdragon NPU: The On-Device Retrieval Tradeoff
  26. jun 10infraGraphRAG vs VectorRAG: Does the Graph Index Earn Its Cost?
  27. jun 10infraMiniMax M3 Ships 1M Context and Desktop Control as Open Weights
  28. jun 10infraDeepSeek-V4 FlashMemory: Sparse Attention for Million-Token Context
  29. jun 09infraIs Cloudflare's Bot Traffic Surge Real? The Measurement Dispute
  30. jun 07infraIndexing Images for RAG: kapa.ai's Approach to Multimodal Retrieval
  31. jun 06infraThe RTX Spark Bet on Unified Memory for Local LLMs: Where Bandwidth Caps It
  32. jun 06infraPod-Level Remote Attestation in Kubernetes: Confidential Workloads on dstack
  33. jun 05infraGenerating GPU Kernels for Moore Threads Silicon: Can LLMs Break CUDA Lock-In?
  34. jun 05infraMicrosoft's Azure Linux Goes General-Purpose: The Container Base-Image Play
  35. jun 05infraCloudflare Acquires VoidZero, the Company Behind Vite's Rust Toolchain
  36. jun 05infraPutting a Datacenter V100 in a Gaming PC: The Local LLM Math
  37. may 27infraWhy LLMs Still Botch Kubernetes Manifests: The Training-Data Gap
  38. may 27infraGemma 4 31B on Cloud TPU vs GPU: The Serving Cost Crossover Point
  39. may 26infraObjectCache Moves KV Reuse to S3-Class Storage: Why Layerwise Retrieval Beats Full-Prefix Cache Hits
  40. may 23infravLLM 0.21 Makes Prefill-Decode Disaggregation Actually Practical
  41. mar 27infraOpenRAG: The Open-Source RAG Platform Challenging Pinecone
  42. mar 24infraMLX vs llama.cpp on Apple Silicon: Which Runtime to Use for Local LLM Inference
  43. mar 24infraPrefill-Decode Disaggregation: The Architecture Shift Redefining LLM Serving
  44. mar 15infraGoogle LiteRT: Running LLMs on Your Phone Without the Cloud
  45. mar 13infraMicrosoft's BitNet: How 1-Bit LLMs Could Make GPU Farms Obsolete
  46. feb 28infraWebAssembly AI: Running Models in the Browser
  47. feb 19infraTailscale Peer Relays: The Missing Piece for True P2P Networking
  48. feb 19infraDNS-Persist-01 Validation: Let's Encrypt's Model for Permanent ACME Certificate Authorization
  49. feb 12infraThe Complete Guide to Local LLMs