groundy
articlessearch

infrastructure & runtime

  1. Vercel In-Function Concurrency: What It Changes for Stateful Node.js
  2. Running LLMs on AMD GPUs With ROCm: What Actually Works
  3. Cloudflare Meerkat: What Globally Distributed Consensus Costs at the Edge
  4. Vercel CDN Now Honors External Origin Cache-Control: Audit Your Headers
  5. AI Found Real Bugs in Cloudflare's Circl Crypto Library
  6. Pruning RAG Context: What to Cut Before the LLM Sees It
  7. Cloudflare's x402 Gateway: What Per-Request API Billing Actually Needs
  8. Doubao 2.1 Pro: What 180 Trillion Daily Tokens Means for Inference Infrastructure
  9. Every CUDA Kernel Pays a Launch Tax: The Host-to-Device Walkthrough
  10. Vercel Montreal Region: Audit Residency Before You Migrate
  11. GLM-5.2 on vLLM and Ascend: Open Weights Beyond NVIDIA
  12. How Vercel Runs Its Own CDN in Front of Discourse: A Self-Dogfooding Case Study
  13. Vercel Runtime Logs Surface CDN Cache Hits, Not the Eviction Cause
  14. Multimodal Knowledge Graph RAG vs Vector RAG: What MKG-RAG-Bench Shows
  15. Vercel Observability Now Tracks Redirects and Rewrites Beside Function Errors
  16. Cloudflare Workflows Saga Rollbacks: Compensating Actions in Serverless Orchestration
  17. Static Corpus RAG: The Bible Case for Separating Churn from Algorithm Complexity
  18. Vercel's KIKO Milano Black Friday Case Study: What the Scaling Claims Skip
  19. Vercel Postgres vs Neon vs Supabase: When the Bundled DB Wins
  20. Fine-Tuning a 20B LLM With RLHF on a 24GB GPU: What Fits
  21. Vercel Flat Rate CDN Beta: Break-Even Math for Spiky Workloads, Tax for the Rest
  22. Where DeepSeek Weights Actually Run on Vercel's AI Gateway
  23. Vercel's Anti-Lock-In Pitch: What the Open-Source Bet Still Locks In
  24. Vercel Adds Tag-Based CDN Cache Invalidation: Surrogate Keys at the Edge
  25. GLM 5.2 Fast on Vercel AI Gateway: What Routing Through Wafer Actually Buys
  26. Vercel CDN Cache Tags vs Path Purging: When Tag Invalidation Wins
  27. Prisma Joins the Vercel Marketplace: The ORM Becomes the Database Vendor
  28. OpenAI on AWS Bedrock: Routing Math to Run Before You Move Traffic
  29. Vercel's Function Observability: What Native Metrics Replace and What They Don't
  30. AWS Databases on the Vercel Marketplace: The Cross-Cloud Latency Tax
  31. Turso on the Vercel Marketplace: Edge SQLite vs the Serverless Connection Pool
  32. Vercel on the AWS Marketplace: What the Listing Does to Procurement and Lock-In
  33. Serving Cold MoE Models: CrossPool Disaggregates KV Cache and Weights
  34. Vercel's In-Function Concurrency: What It Does to Cold Starts and Billing
  35. Poisoning a RAG Retriever: How Conflict-Aware Edits Inject False Knowledge
  36. Vercel Raised Its CDN Origin Timeout to Two Minutes: What Breaks First
  37. Gradio-Lite Runs Model Inference in the Browser via Pyodide, No Server
  38. Cloudflare AI Gateway Adds Spend Limits to Cap the Runaway Inference Bill
  39. Vercel Now Honors stale-if-error: Serving Stale Cache When the Origin Dies
  40. Vercel's Manual CDN Purge API: Cache Control Without a Redeploy
  41. Cloudflare Now Routes Public Traffic to Private Apps via DNS, No VPN
  42. GitHub's AI Capacity Crunch Pushes Microsoft to Rent AWS Compute
  43. Cloudflare's Temporary Accounts Give AI Agents Disposable Credentials
  44. Running Long-Context Agents on a 4-Bit KV Cache: Where Accuracy Breaks
  45. When LLM-Generated CUDA Kernels Pass Tests but Get the Math Wrong
  46. Running GLM-5.2 at Home: SGLang, vLLM, Transformers, and KTransformers Setup Guide
  47. AWS Bedrock Now Requires Data Sharing for Mythos: The Self-Hosting Calculus
  48. vLLM Cold Start Latency: Why Scale-to-Zero LLM Serving Stalls
  49. The Vercel-AWS Deal Reveals Where AI Inference Runs
  50. Running RAG on a Snapdragon NPU: The On-Device Retrieval Tradeoff