<?xml version="1.0" encoding="UTF-8"?><rss version="2.0" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:media="http://search.yahoo.com/mrss/"><channel><title>Groundy · Most read</title><description>The articles Groundy readers are spending time with.</description><link>https://groundy.com/</link><language>en-us</language><ttl>60</ttl><lastBuildDate>Mon, 14 Sep 2026 01:01:18 GMT</lastBuildDate><atom:link href="https://groundy.com/popular.xml" rel="self" type="application/rss+xml"/><atom:link href="https://groundy.com/feeds/" rel="alternate" type="text/html"/><image><url>https://groundy.com/rss-icon.png</url><title>Groundy · Most read</title><link>https://groundy.com/</link><width>144</width><height>144</height></image><item><title>MLX vs llama.cpp on Apple Silicon: Which Runtime to Use for Local LLM Inference</title><link>https://groundy.com/articles/mlx-vs-llamacpp-on-apple-silicon-which-runtime-to-use-for-local-llm-inference/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/mlx-vs-llamacpp-on-apple-silicon-which-runtime-to-use-for-local-llm-inference/</guid><description>MLX delivers 20-87% faster generation on Apple Silicon for models under 14B parameters. llama.cpp wins for cross-platform use and long contexts.</description><pubDate>Tue, 24 Mar 2026 16:48:05 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/mlx-vs-llamacpp-on-apple-silicon-which-runtime-to-use-for-local-llm-inference/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/mlx-vs-llamacpp-on-apple-silicon-which-runtime-to-use-for-local-llm-inference/image.jpg?v=53aae127609cb8d8&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Two stacks of dark model-weight slabs rest in contrasting runtime cradles: a seamless copper enclosure beside a modular warm-gray frame with green joints.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;MLX delivers 20-87% faster generation on Apple Silicon for models under 14B parameters. llama.cpp wins for cross-platform use and long contexts.&lt;/p&gt;&lt;p&gt;Berry Mingus · Infrastructure &amp;amp; Runtime · 9 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/mlx-vs-llamacpp-on-apple-silicon-which-runtime-to-use-for-local-llm-inference/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-08-29T00:00:00.000Z</atom:updated><media:content url="https://groundy.com/feed-images/mlx-vs-llamacpp-on-apple-silicon-which-runtime-to-use-for-local-llm-inference/image.jpg?v=53aae127609cb8d8" type="image/jpeg" medium="image" width="960" height="540" fileSize="78909"><media:description type="plain">Two stacks of dark model-weight slabs rest in contrasting runtime cradles: a seamless copper enclosure beside a modular warm-gray frame with green joints.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/mlx-vs-llamacpp-on-apple-silicon-which-runtime-to-use-for-local-llm-inference/thumb.jpg?v=53aae127609cb8d8" width="640" height="360"/><category>Infrastructure &amp; Runtime</category><category>mlx</category><category>llama-cpp</category><category>apple-silicon</category><category>on-device-inference</category><category>macos</category><category>local-llm</category><enclosure url="https://groundy.com/feed-images/mlx-vs-llamacpp-on-apple-silicon-which-runtime-to-use-for-local-llm-inference/image.jpg?v=53aae127609cb8d8" length="78909" type="image/jpeg"/></item><item><title>EU&apos;s 2027 Replaceable Battery Mandate: What It Means for Phone Buyers and Repairers Right Now</title><link>https://groundy.com/articles/eus-2027-replaceable-battery-mandate-what-it-means-for-phone-buyers-and/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/eus-2027-replaceable-battery-mandate-what-it-means-for-phone-buyers-and/</guid><description>The EU&apos;s 2027 battery mandate is confirmed. Here&apos;s what &apos;user-replaceable&apos; legally means, which phones comply now, and how to buy smart before the rules change.</description><pubDate>Tue, 21 Apr 2026 16:00:00 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/eus-2027-replaceable-battery-mandate-what-it-means-for-phone-buyers-and/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/eus-2027-replaceable-battery-mandate-what-it-means-for-phone-buyers-and/image.jpg?v=f216be807ee534b1&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;A phone battery lifts from a gasketed chassis beside an appliance with a hand-removable battery.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;The EU&amp;apos;s 2027 battery mandate is confirmed. Here&amp;apos;s what &amp;apos;user-replaceable&amp;apos; legally means, which phones comply now, and how to buy smart before the rules change.&lt;/p&gt;&lt;p&gt;Berry Mingus · Culture &amp;amp; Society · 6 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/eus-2027-replaceable-battery-mandate-what-it-means-for-phone-buyers-and/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-07-03T00:00:00.000Z</atom:updated><media:content url="https://groundy.com/feed-images/eus-2027-replaceable-battery-mandate-what-it-means-for-phone-buyers-and/image.jpg?v=f216be807ee534b1" type="image/jpeg" medium="image" width="960" height="540" fileSize="118316"><media:description type="plain">A phone battery lifts from a gasketed chassis beside an appliance with a hand-removable battery.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/eus-2027-replaceable-battery-mandate-what-it-means-for-phone-buyers-and/thumb.jpg?v=f216be807ee534b1" width="640" height="360"/><category>Culture &amp; Society</category><category>eu-regulation</category><category>right-to-repair</category><category>consumer-electronics</category><category>sustainability</category><category>batteries</category><enclosure url="https://groundy.com/feed-images/eus-2027-replaceable-battery-mandate-what-it-means-for-phone-buyers-and/image.jpg?v=f216be807ee534b1" length="118316" type="image/jpeg"/></item><item><title>Open-Weight LLM Leaderboards 2026: Where DeepSeek, Qwen, and GLM Rank</title><link>https://groundy.com/articles/open-weight-llm-leaderboards-2026-where-deepseek-qwen-and-glm-rank/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/open-weight-llm-leaderboards-2026-where-deepseek-qwen-and-glm-rank/</guid><description>DataLearner&apos;s June 2026 snapshot ranks GLM-5.2 seventh by HLE at 54.70 and places no Chinese flagship in the overall top three, undercutting launch-day claims.</description><pubDate>Fri, 26 Jun 2026 22:09:15 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/open-weight-llm-leaderboards-2026-where-deepseek-qwen-and-glm-rank/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/open-weight-llm-leaderboards-2026-where-deepseek-qwen-and-glm-rank/image.jpg?v=c1ab6cf17d5cb6c7&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Three identical model modules sit before different independent physical test fixtures.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;DataLearner&amp;apos;s June 2026 snapshot ranks GLM-5.2 seventh by HLE at 54.70 and places no Chinese flagship in the overall top three, undercutting launch-day claims.&lt;/p&gt;&lt;p&gt;Berry Mingus · Models &amp;amp; Research · 8 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/open-weight-llm-leaderboards-2026-where-deepseek-qwen-and-glm-rank/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/open-weight-llm-leaderboards-2026-where-deepseek-qwen-and-glm-rank/image.jpg?v=c1ab6cf17d5cb6c7" type="image/jpeg" medium="image" width="960" height="540" fileSize="112557"><media:description type="plain">Three identical model modules sit before different independent physical test fixtures.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/open-weight-llm-leaderboards-2026-where-deepseek-qwen-and-glm-rank/thumb.jpg?v=c1ab6cf17d5cb6c7" width="640" height="360"/><category>Models &amp; Research</category><category>llm-leaderboards</category><category>open-weight-models</category><category>ai-benchmarks</category><category>chinese-ai-models</category><category>model-evaluation</category><category>model-selection</category><enclosure url="https://groundy.com/feed-images/open-weight-llm-leaderboards-2026-where-deepseek-qwen-and-glm-rank/image.jpg?v=c1ab6cf17d5cb6c7" length="112557" type="image/jpeg"/></item><item><title>Chinese AI Models Compared: DeepSeek, Qwen, Kimi, Doubao, and Ernie</title><link>https://groundy.com/articles/the-chinese-ai-model-ecosystem-deepseek-qwen-kimi-doubao-and-ernie-compared/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/the-chinese-ai-model-ecosystem-deepseek-qwen-kimi-doubao-and-ernie-compared/</guid><description>DeepSeek isn&apos;t China&apos;s only frontier AI. Compare DeepSeek, Qwen, Kimi, Doubao, and Ernie on benchmarks, licensing, API access, and use-case fit.</description><pubDate>Tue, 24 Mar 2026 14:57:11 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/the-chinese-ai-model-ecosystem-deepseek-qwen-kimi-doubao-and-ernie-compared/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/the-chinese-ai-model-ecosystem-deepseek-qwen-kimi-doubao-and-ernie-compared/image.jpg?v=b938e0a1826c5a3c&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Six distinct model cartridges surround a shared copper manifold, with different housings and connectors representing varied deployment choices.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;DeepSeek isn&amp;apos;t China&amp;apos;s only frontier AI. Compare DeepSeek, Qwen, Kimi, Doubao, and Ernie on benchmarks, licensing, API access, and use-case fit.&lt;/p&gt;&lt;p&gt;Berry Mingus · Models &amp;amp; Research · 10 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/the-chinese-ai-model-ecosystem-deepseek-qwen-kimi-doubao-and-ernie-compared/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-06-19T00:00:00.000Z</atom:updated><media:content url="https://groundy.com/feed-images/the-chinese-ai-model-ecosystem-deepseek-qwen-kimi-doubao-and-ernie-compared/image.jpg?v=b938e0a1826c5a3c" type="image/jpeg" medium="image" width="960" height="540" fileSize="139615"><media:description type="plain">Six distinct model cartridges surround a shared copper manifold, with different housings and connectors representing varied deployment choices.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/the-chinese-ai-model-ecosystem-deepseek-qwen-kimi-doubao-and-ernie-compared/thumb.jpg?v=b938e0a1826c5a3c" width="640" height="360"/><category>Models &amp; Research</category><category>deepseek</category><category>qwen</category><category>kimi</category><category>doubao</category><category>ernie</category><category>chinese-ai</category><category>alibaba</category><category>baidu</category><category>bytedance</category><enclosure url="https://groundy.com/feed-images/the-chinese-ai-model-ecosystem-deepseek-qwen-kimi-doubao-and-ernie-compared/image.jpg?v=b938e0a1826c5a3c" length="139615" type="image/jpeg"/></item><item><title>4-Bit vs 8-Bit Quants: Why Accuracy Benchmarks Miss Distribution Drift</title><link>https://groundy.com/articles/4-bit-vs-8-bit-quants-why-accuracy-benchmarks-miss-distribution-drift/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/4-bit-vs-8-bit-quants-why-accuracy-benchmarks-miss-distribution-drift/</guid><description>A preprint argues zero-shot accuracy misses distribution drift in quantized LLMs, recommending divergence metrics like JSD and TV against BF16 bases for safer deployment.</description><pubDate>Sat, 12 Sep 2026 04:11:48 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/4-bit-vs-8-bit-quants-why-accuracy-benchmarks-miss-distribution-drift/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/4-bit-vs-8-bit-quants-why-accuracy-benchmarks-miss-distribution-drift/image.jpg?v=c64017a7814bcafb&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Two graphite-shaded folded paper sculptures stand on warm ivory paper with matching copper pointed tips. The left has many narrow folds; the right has broad angular planes.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;A preprint argues zero-shot accuracy misses distribution drift in quantized LLMs, recommending divergence metrics like JSD and TV against BF16 bases for safer deployment.&lt;/p&gt;&lt;p&gt;Berry Mingus · Developer Tools · 9 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/4-bit-vs-8-bit-quants-why-accuracy-benchmarks-miss-distribution-drift/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/4-bit-vs-8-bit-quants-why-accuracy-benchmarks-miss-distribution-drift/image.jpg?v=c64017a7814bcafb" type="image/jpeg" medium="image" width="960" height="540" fileSize="47269"><media:description type="plain">Two graphite-shaded folded paper sculptures stand on warm ivory paper with matching copper pointed tips. The left has many narrow folds; the right has broad angular planes.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/4-bit-vs-8-bit-quants-why-accuracy-benchmarks-miss-distribution-drift/thumb.jpg?v=c64017a7814bcafb" width="640" height="360"/><category>Developer Tools</category><category>llm-quantization</category><category>model-evaluation</category><category>distribution-drift</category><category>gguf</category><category>divergence-metrics</category><category>developer-tools</category><enclosure url="https://groundy.com/feed-images/4-bit-vs-8-bit-quants-why-accuracy-benchmarks-miss-distribution-drift/image.jpg?v=c64017a7814bcafb" length="47269" type="image/jpeg"/></item><item><title>FP8 vs MXFP4 vs BF16: Why Your Quantized LLM Disagrees Across GPUs</title><link>https://groundy.com/articles/fp8-vs-mxfp4-vs-bf16-why-your-quantized-llm-disagrees-across-gpus/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/fp8-vs-mxfp4-vs-bf16-why-your-quantized-llm-disagrees-across-gpus/</guid><description>FP8 and MXFP4 are umbrella specs, not single formats. A new preprint offers bit-exact conformance vectors to test quantized LLM portability across GPUs, exposing hidden format</description><pubDate>Mon, 07 Sep 2026 07:15:12 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/fp8-vs-mxfp4-vs-bf16-why-your-quantized-llm-disagrees-across-gpus/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/fp8-vs-mxfp4-vs-bf16-why-your-quantized-llm-disagrees-across-gpus/image.jpg?v=9ac8aca899d82bb1&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Three presses shape identical copper blanks into subtly different gears beside one shared brass reference gauge.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;FP8 and MXFP4 are umbrella specs, not single formats. A new preprint offers bit-exact conformance vectors to test quantized LLM portability across GPUs, exposing hidden format&lt;/p&gt;&lt;p&gt;Berry Mingus · Models &amp;amp; Research · 13 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/fp8-vs-mxfp4-vs-bf16-why-your-quantized-llm-disagrees-across-gpus/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/fp8-vs-mxfp4-vs-bf16-why-your-quantized-llm-disagrees-across-gpus/image.jpg?v=9ac8aca899d82bb1" type="image/jpeg" medium="image" width="960" height="540" fileSize="141208"><media:description type="plain">Three presses shape identical copper blanks into subtly different gears beside one shared brass reference gauge.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/fp8-vs-mxfp4-vs-bf16-why-your-quantized-llm-disagrees-across-gpus/thumb.jpg?v=9ac8aca899d82bb1" width="640" height="360"/><category>Models &amp; Research</category><category>quantization</category><category>fp8</category><category>mxfp4</category><category>gpu-compatibility</category><category>llm-inference</category><category>conformance-testing</category><enclosure url="https://groundy.com/feed-images/fp8-vs-mxfp4-vs-bf16-why-your-quantized-llm-disagrees-across-gpus/image.jpg?v=9ac8aca899d82bb1" length="141208" type="image/jpeg"/></item><item><title>GitHub Copilot vs Cursor vs Claude Code: The 2026 AI Coding Showdown</title><link>https://groundy.com/articles/github-copilot-vs-cursor-vs-claude-code-2026-ai-coding/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/github-copilot-vs-cursor-vs-claude-code-2026-ai-coding/</guid><description>GitHub Copilot owns enterprise, Cursor owns developer wallets at $2B ARR, and Claude Code leads the benchmarks. Which fits your workflow depends on what you build.</description><pubDate>Sat, 14 Mar 2026 16:29:11 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/github-copilot-vs-cursor-vs-claude-code-2026-ai-coding/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/github-copilot-vs-cursor-vs-claude-code-2026-ai-coding/image.jpg?v=48b3cff2cc3273ee&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Three distinct precision tools work on different parts of one folded paper construction.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;GitHub Copilot owns enterprise, Cursor owns developer wallets at $2B ARR, and Claude Code leads the benchmarks. Which fits your workflow depends on what you build.&lt;/p&gt;&lt;p&gt;Berry Mingus · Developer Tools · 11 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/github-copilot-vs-cursor-vs-claude-code-2026-ai-coding/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-08-22T00:00:00.000Z</atom:updated><media:content url="https://groundy.com/feed-images/github-copilot-vs-cursor-vs-claude-code-2026-ai-coding/image.jpg?v=48b3cff2cc3273ee" type="image/jpeg" medium="image" width="960" height="540" fileSize="98594"><media:description type="plain">Three distinct precision tools work on different parts of one folded paper construction.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/github-copilot-vs-cursor-vs-claude-code-2026-ai-coding/thumb.jpg?v=48b3cff2cc3273ee" width="640" height="360"/><category>Developer Tools</category><category>ai-tools</category><category>developer-tools</category><category>coding-assistants</category><category>github-copilot</category><category>cursor</category><category>claude-code</category><enclosure url="https://groundy.com/feed-images/github-copilot-vs-cursor-vs-claude-code-2026-ai-coding/image.jpg?v=48b3cff2cc3273ee" length="98594" type="image/jpeg"/></item><item><title>Running MoE LLMs on a Single GPU: What Expert Offloading Actually Costs</title><link>https://groundy.com/articles/running-moe-llms-on-a-single-gpu-what-expert-offloading-actually-costs/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/running-moe-llms-on-a-single-gpu-what-expert-offloading-actually-costs/</guid><description>FluxMoE streams MoE weights from host DRAM to save VRAM, trading capacity for bandwidth. Author-reported gains on multi-GPU setups lack independent replication.</description><pubDate>Fri, 11 Sep 2026 20:14:36 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/running-moe-llms-on-a-single-gpu-what-expert-offloading-actually-costs/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/running-moe-llms-on-a-single-gpu-what-expert-offloading-actually-costs/image.jpg?v=6fdab584d9e10514&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Graphite illustration on warm ivory paper: a rack of upright slabs connects by a narrow copper bridge carrying three blocks to a press-like station holding three more blocks.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;FluxMoE streams MoE weights from host DRAM to save VRAM, trading capacity for bandwidth. Author-reported gains on multi-GPU setups lack independent replication.&lt;/p&gt;&lt;p&gt;Berry Mingus · Infrastructure &amp;amp; Runtime · 9 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/running-moe-llms-on-a-single-gpu-what-expert-offloading-actually-costs/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/running-moe-llms-on-a-single-gpu-what-expert-offloading-actually-costs/image.jpg?v=6fdab584d9e10514" type="image/jpeg" medium="image" width="960" height="540" fileSize="69372"><media:description type="plain">Graphite illustration on warm ivory paper: a rack of upright slabs connects by a narrow copper bridge carrying three blocks to a press-like station holding three more blocks.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/running-moe-llms-on-a-single-gpu-what-expert-offloading-actually-costs/thumb.jpg?v=6fdab584d9e10514" width="640" height="360"/><category>Infrastructure &amp; Runtime</category><category>mixture-of-experts</category><category>expert-offloading</category><category>gpu-inference</category><category>vllm</category><category>memory-bandwidth</category><category>llm-serving</category><enclosure url="https://groundy.com/feed-images/running-moe-llms-on-a-single-gpu-what-expert-offloading-actually-costs/image.jpg?v=6fdab584d9e10514" length="69372" type="image/jpeg"/></item><item><title>Private Vector Search vs TEEs: Can RAG Retrieval Be Outsourced Safely?</title><link>https://groundy.com/articles/private-vector-search-vs-tees-can-rag-retrieval-be-outsourced-safely/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/private-vector-search-vs-tees-can-rag-retrieval-be-outsourced-safely/</guid><description>Spruce claims 0.21-2.97s private retrieval via MPC, challenging TEE defaults. Compare self-hosted Qdrant, TEEs, and cryptographic outsourcing for privacy-sensitive RAG.</description><pubDate>Sun, 06 Sep 2026 08:06:35 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/private-vector-search-vs-tees-can-rag-retrieval-be-outsourced-safely/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/private-vector-search-vs-tees-can-rag-retrieval-be-outsourced-safely/image.jpg?v=f14c585eb61729eb&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;One sealed capsule enters an opaque enclave vault while another splits across two distant server towers and rejoins only at a locked corpus cabinet.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;Spruce claims 0.21-2.97s private retrieval via MPC, challenging TEE defaults. Compare self-hosted Qdrant, TEEs, and cryptographic outsourcing for privacy-sensitive RAG.&lt;/p&gt;&lt;p&gt;Berry Mingus · Infrastructure &amp;amp; Runtime · 13 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/private-vector-search-vs-tees-can-rag-retrieval-be-outsourced-safely/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/private-vector-search-vs-tees-can-rag-retrieval-be-outsourced-safely/image.jpg?v=f14c585eb61729eb" type="image/jpeg" medium="image" width="960" height="540" fileSize="134727"><media:description type="plain">One sealed capsule enters an opaque enclave vault while another splits across two distant server towers and rejoins only at a locked corpus cabinet.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/private-vector-search-vs-tees-can-rag-retrieval-be-outsourced-safely/thumb.jpg?v=f14c585eb61729eb" width="640" height="360"/><category>Infrastructure &amp; Runtime</category><category>rag</category><category>vector-search</category><category>privacy</category><category>mpc</category><category>tee</category><category>qdrant</category><category>infrastructure</category><enclosure url="https://groundy.com/feed-images/private-vector-search-vs-tees-can-rag-retrieval-be-outsourced-safely/image.jpg?v=f14c585eb61729eb" length="134727" type="image/jpeg"/></item><item><title>Cursor&apos;s Meteoric Rise: Inside the AI Editor Hitting $300M ARR</title><link>https://groundy.com/articles/cursor-s-meteoric-rise-inside-ai-editor-hitting-300m/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/cursor-s-meteoric-rise-inside-ai-editor-hitting-300m/</guid><description>Cursor hit $300M ARR in April 2025 by forking VS Code and baking AI into the editor&apos;s core. By June 2026 it was at $4B annualized and agreed to a $60B SpaceX acquisition. Here&apos;s how it happened and what it signals.</description><pubDate>Fri, 27 Feb 2026 14:15:59 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/cursor-s-meteoric-rise-inside-ai-editor-hitting-300m/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/cursor-s-meteoric-rise-inside-ai-editor-hitting-300m/image.jpg?v=ae395d6ebc57c361&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;A precision editing tool and folded paper construction sit on a modular work surface that extends into additional bays.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;Cursor hit $300M ARR in April 2025 by forking VS Code and baking AI into the editor&amp;apos;s core. By June 2026 it was at $4B annualized and agreed to a $60B SpaceX acquisition. Here&amp;apos;s how it happened and what it signals.&lt;/p&gt;&lt;p&gt;Berry Mingus · Industry &amp;amp; Business · 8 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/cursor-s-meteoric-rise-inside-ai-editor-hitting-300m/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-06-19T00:00:00.000Z</atom:updated><media:content url="https://groundy.com/feed-images/cursor-s-meteoric-rise-inside-ai-editor-hitting-300m/image.jpg?v=ae395d6ebc57c361" type="image/jpeg" medium="image" width="960" height="540" fileSize="114459"><media:description type="plain">A precision editing tool and folded paper construction sit on a modular work surface that extends into additional bays.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/cursor-s-meteoric-rise-inside-ai-editor-hitting-300m/thumb.jpg?v=ae395d6ebc57c361" width="640" height="360"/><category>Industry &amp; Business</category><category>ai-industry</category><category>cursor</category><category>startups</category><category>developer-tools</category><category>enterprise-software</category><category>venture-capital</category><enclosure url="https://groundy.com/feed-images/cursor-s-meteoric-rise-inside-ai-editor-hitting-300m/image.jpg?v=ae395d6ebc57c361" length="114459" type="image/jpeg"/></item><item><title>AI Code Generation Benchmarks 2026: Which Model Actually Writes Better Code?</title><link>https://groundy.com/articles/ai-code-generation-benchmarks-2026-which-model-actually/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/ai-code-generation-benchmarks-2026-which-model-actually/</guid><description>Frontier and open-weight coding models now post similar benchmark scores, but real-world software engineering exposes gaps between leaderboard results and practical utility.</description><pubDate>Sun, 15 Feb 2026 11:01:06 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/ai-code-generation-benchmarks-2026-which-model-actually/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/ai-code-generation-benchmarks-2026-which-model-actually/image.jpg?v=dd6924deccb996b7&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Different tool-steel keys face a precision test fixture and a second partly concealed lock, with no declared winner.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;Frontier and open-weight coding models now post similar benchmark scores, but real-world software engineering exposes gaps between leaderboard results and practical utility.&lt;/p&gt;&lt;p&gt;Berry Mingus · Models &amp;amp; Research · 9 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/ai-code-generation-benchmarks-2026-which-model-actually/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-06-19T00:00:00.000Z</atom:updated><media:content url="https://groundy.com/feed-images/ai-code-generation-benchmarks-2026-which-model-actually/image.jpg?v=dd6924deccb996b7" type="image/jpeg" medium="image" width="960" height="540" fileSize="120212"><media:description type="plain">Different tool-steel keys face a precision test fixture and a second partly concealed lock, with no declared winner.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/ai-code-generation-benchmarks-2026-which-model-actually/thumb.jpg?v=dd6924deccb996b7" width="640" height="360"/><category>Models &amp; Research</category><category>ai-research</category><category>benchmarks</category><category>code-generation</category><category>comparison</category><category>llm</category><category>open-source</category><enclosure url="https://groundy.com/feed-images/ai-code-generation-benchmarks-2026-which-model-actually/image.jpg?v=dd6924deccb996b7" length="120212" type="image/jpeg"/></item><item><title>Constitutional AI Self-Amendment Hits the Metacognition Wall</title><link>https://groundy.com/articles/constitutional-ai-self-amendment-hits-the-metacognition-wall/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/constitutional-ai-self-amendment-hits-the-metacognition-wall/</guid><description>An arXiv preprint argues LLM metacognition is coarse and context-dependent, suggesting self-amendment requires external calibration and human approval gates.</description><pubDate>Sun, 13 Sep 2026 08:57:23 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/constitutional-ai-self-amendment-hits-the-metacognition-wall/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/constitutional-ai-self-amendment-hits-the-metacognition-wall/image.jpg?v=176b1244a631c388&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Graphite illustration of a closed ivory book secured by a forest-green strap and brass clasp. A blank loose sheet overlaps its edge beside a copper-framed mirror showing a blurred reflection, on warm ivory paper.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;An arXiv preprint argues LLM metacognition is coarse and context-dependent, suggesting self-amendment requires external calibration and human approval gates.&lt;/p&gt;&lt;p&gt;Berry Mingus · Ethics, Policy &amp;amp; Safety · 8 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/constitutional-ai-self-amendment-hits-the-metacognition-wall/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/constitutional-ai-self-amendment-hits-the-metacognition-wall/image.jpg?v=176b1244a631c388" type="image/jpeg" medium="image" width="960" height="540" fileSize="51868"><media:description type="plain">Graphite illustration of a closed ivory book secured by a forest-green strap and brass clasp. A blank loose sheet overlaps its edge beside a copper-framed mirror showing a blurred reflection, on warm ivory paper.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/constitutional-ai-self-amendment-hits-the-metacognition-wall/thumb.jpg?v=176b1244a631c388" width="640" height="360"/><category>Ethics, Policy &amp; Safety</category><category>llm-metacognition</category><category>constitutional-ai</category><category>ai-governance</category><category>self-amendment</category><category>confidence-calibration</category><category>agentic-systems</category><enclosure url="https://groundy.com/feed-images/constitutional-ai-self-amendment-hits-the-metacognition-wall/image.jpg?v=176b1244a631c388" length="51868" type="image/jpeg"/></item><item><title>Do More Tools Make LLM Agents Worse? When to Gate External Evidence</title><link>https://groundy.com/articles/do-more-tools-make-llm-agents-worse-when-to-gate-external-evidence/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/do-more-tools-make-llm-agents-worse-when-to-gate-external-evidence/</guid><description>Preprints show forced evidence gathering raised clinical agent error from 28.3% to 34.3%, while curated tool menus lifted ToolBench success to 0.898, suggesting gating is key.</description><pubDate>Fri, 11 Sep 2026 20:58:31 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/do-more-tools-make-llm-agents-worse-when-to-gate-external-evidence/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/do-more-tools-make-llm-agents-worse-when-to-gate-external-evidence/image.jpg?v=e01df841c4181024&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;A partly open copper gate admits an ivory puzzle piece toward an interlocking ivory, gray and green sequence. Loose ceramic pieces wait beside it on textured ivory paper.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;Preprints show forced evidence gathering raised clinical agent error from 28.3% to 34.3%, while curated tool menus lifted ToolBench success to 0.898, suggesting gating is key.&lt;/p&gt;&lt;p&gt;Berry Mingus · Agents &amp;amp; Frameworks · 9 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/do-more-tools-make-llm-agents-worse-when-to-gate-external-evidence/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/do-more-tools-make-llm-agents-worse-when-to-gate-external-evidence/image.jpg?v=e01df841c4181024" type="image/jpeg" medium="image" width="960" height="540" fileSize="62130"><media:description type="plain">A partly open copper gate admits an ivory puzzle piece toward an interlocking ivory, gray and green sequence. Loose ceramic pieces wait beside it on textured ivory paper.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/do-more-tools-make-llm-agents-worse-when-to-gate-external-evidence/thumb.jpg?v=e01df841c4181024" width="640" height="360"/><category>Agents &amp; Frameworks</category><category>llm-agents</category><category>tool-gating</category><category>mcp</category><category>evidence-handling</category><category>clinical-ai</category><category>tool-selection</category><enclosure url="https://groundy.com/feed-images/do-more-tools-make-llm-agents-worse-when-to-gate-external-evidence/image.jpg?v=e01df841c4181024" length="62130" type="image/jpeg"/></item><item><title>Terminal vs Browser for AI Agents: What a Hybrid GUI+CLI Benchmark Shows</title><link>https://groundy.com/articles/terminal-vs-browser-for-ai-agents-what-a-hybrid-gui-cli-benchmark-shows/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/terminal-vs-browser-for-ai-agents-what-a-hybrid-gui-cli-benchmark-shows/</guid><description>CUA-Universe preprint shows hybrid GUI+CLI agents cut steps 37% and tokens 60% on 16 apps. Audit your eval stack: screenshot-only benchmarks mismeasure tasks with terminal.</description><pubDate>Mon, 07 Sep 2026 19:36:21 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/terminal-vs-browser-for-ai-agents-what-a-hybrid-gui-cli-benchmark-shows/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/terminal-vs-browser-for-ai-agents-what-a-hybrid-gui-cli-benchmark-shows/image.jpg?v=c3e9957918173b6d&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;A glazed visual workspace and a bank of precise punched-strip levers connect through one copper clutch to a shared task carriage.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;CUA-Universe preprint shows hybrid GUI+CLI agents cut steps 37% and tokens 60% on 16 apps. Audit your eval stack: screenshot-only benchmarks mismeasure tasks with terminal.&lt;/p&gt;&lt;p&gt;Berry Mingus · Developer Tools · 14 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/terminal-vs-browser-for-ai-agents-what-a-hybrid-gui-cli-benchmark-shows/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/terminal-vs-browser-for-ai-agents-what-a-hybrid-gui-cli-benchmark-shows/image.jpg?v=c3e9957918173b6d" type="image/jpeg" medium="image" width="960" height="540" fileSize="111988"><media:description type="plain">A glazed visual workspace and a bank of precise punched-strip levers connect through one copper clutch to a shared task carriage.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/terminal-vs-browser-for-ai-agents-what-a-hybrid-gui-cli-benchmark-shows/thumb.jpg?v=c3e9957918173b6d" width="640" height="360"/><category>Developer Tools</category><category>computer-use-agents</category><category>gui-cli-routing</category><category>agent-evaluation</category><category>cua-universe</category><category>developer-tools</category><category>security-scoping</category><enclosure url="https://groundy.com/feed-images/terminal-vs-browser-for-ai-agents-what-a-hybrid-gui-cli-benchmark-shows/image.jpg?v=c3e9957918173b6d" length="111988" type="image/jpeg"/></item><item><title>How Cloudflare Detects MCP Traffic and What Enterprises Should Filter</title><link>https://groundy.com/articles/how-cloudflare-detects-mcp-traffic-and-what-enterprises-should-filter/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/how-cloudflare-detects-mcp-traffic-and-what-enterprises-should-filter/</guid><description>Cloudflare claims MCP traffic classification, but stdio servers remain invisible to edge proxies. Build a filter policy that gates remote endpoints and treats local tool on as</description><pubDate>Mon, 07 Sep 2026 08:56:55 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/how-cloudflare-detects-mcp-traffic-and-what-enterprises-should-filter/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/how-cloudflare-detects-mcp-traffic-and-what-enterprises-should-filter/image.jpg?v=99c3a30763217335&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Remote parcels pass through a green edge inspection gate while a short underground local conduit bypasses it.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;Cloudflare claims MCP traffic classification, but stdio servers remain invisible to edge proxies. Build a filter policy that gates remote endpoints and treats local tool on as&lt;/p&gt;&lt;p&gt;Berry Mingus · Agents &amp;amp; Frameworks · 11 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/how-cloudflare-detects-mcp-traffic-and-what-enterprises-should-filter/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/how-cloudflare-detects-mcp-traffic-and-what-enterprises-should-filter/image.jpg?v=99c3a30763217335" type="image/jpeg" medium="image" width="960" height="540" fileSize="147836"><media:description type="plain">Remote parcels pass through a green edge inspection gate while a short underground local conduit bypasses it.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/how-cloudflare-detects-mcp-traffic-and-what-enterprises-should-filter/thumb.jpg?v=99c3a30763217335" width="640" height="360"/><category>Agents &amp; Frameworks</category><category>mcp</category><category>cloudflare</category><category>edge-security</category><category>agent-security</category><category>json-rpc</category><category>egress-control</category><enclosure url="https://groundy.com/feed-images/how-cloudflare-detects-mcp-traffic-and-what-enterprises-should-filter/image.jpg?v=99c3a30763217335" length="147836" type="image/jpeg"/></item><item><title>Noisy Neighbors at the Fabric: Why Shared GPU Clusters Throttle Your Jobs</title><link>https://groundy.com/articles/noisy-neighbors-at-the-fabric-why-shared-gpu-clusters-throttle-your-jobs/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/noisy-neighbors-at-the-fabric-why-shared-gpu-clusters-throttle-your-jobs/</guid><description>A new NCCL shim recovers 13-38% bandwidth on shared GPU clusters by tuning collective patterns. Test for cross-tenant interference before buying more fabric.</description><pubDate>Mon, 07 Sep 2026 07:59:55 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/noisy-neighbors-at-the-fabric-why-shared-gpu-clusters-throttle-your-jobs/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/noisy-neighbors-at-the-fabric-why-shared-gpu-clusters-throttle-your-jobs/image.jpg?v=f8ce0220d5a63d7f&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Six weaving engines share one congested copper-thread junction where a green adaptive shuttle works through a knot.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;A new NCCL shim recovers 13-38% bandwidth on shared GPU clusters by tuning collective patterns. Test for cross-tenant interference before buying more fabric.&lt;/p&gt;&lt;p&gt;Berry Mingus · Infrastructure &amp;amp; Runtime · 12 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/noisy-neighbors-at-the-fabric-why-shared-gpu-clusters-throttle-your-jobs/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><media:content url="https://groundy.com/feed-images/noisy-neighbors-at-the-fabric-why-shared-gpu-clusters-throttle-your-jobs/image.jpg?v=f8ce0220d5a63d7f" type="image/jpeg" medium="image" width="960" height="540" fileSize="151759"><media:description type="plain">Six weaving engines share one congested copper-thread junction where a green adaptive shuttle works through a knot.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/noisy-neighbors-at-the-fabric-why-shared-gpu-clusters-throttle-your-jobs/thumb.jpg?v=f8ce0220d5a63d7f" width="640" height="360"/><category>Infrastructure &amp; Runtime</category><category>gpu-clusters</category><category>nccl</category><category>infiniband</category><category>distributed-training</category><category>network-congestion</category><category>multi-tenant</category><enclosure url="https://groundy.com/feed-images/noisy-neighbors-at-the-fabric-why-shared-gpu-clusters-throttle-your-jobs/image.jpg?v=f8ce0220d5a63d7f" length="151759" type="image/jpeg"/></item><item><title>How Spotify Cut Claude Code Token Usage 90%: What It Takes to Replicate</title><link>https://groundy.com/articles/how-spotify-cut-claude-code-token-usage-90-what-it-takes-to-replicate/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/how-spotify-cut-claude-code-token-usage-90-what-it-takes-to-replicate/</guid><description>Spotify&apos;s claimed 90% token cut via Portal is unverified. Benchmark token proxies against Claude Code&apos;s native compaction and subscription ceilings before building.</description><pubDate>Sat, 05 Sep 2026 20:04:28 GMT</pubDate><content:encoded>&lt;p&gt;Spotify&amp;apos;s claimed 90% token cut via Portal is unverified. Benchmark token proxies against Claude Code&amp;apos;s native compaction and subscription ceilings before building.&lt;/p&gt;&lt;p&gt;Berry Mingus · Agents &amp;amp; Frameworks · 12 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/how-spotify-cut-claude-code-token-usage-90-what-it-takes-to-replicate/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><category>Agents &amp; Frameworks</category><category>claude-code</category><category>token-optimization</category><category>llm-infrastructure</category><category>agent-fleets</category><category>cost-engineering</category><category>prompt-caching</category></item><item><title>Claude Code in GitHub Actions: A Complete Guide to Automated PR Fixes</title><link>https://groundy.com/articles/how-to-run-claude-code-as-a-github-actions-agent-for-automated-pr-fixes/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/how-to-run-claude-code-as-a-github-actions-agent-for-automated-pr-fixes/</guid><description>How to wire Claude Code into GitHub Actions for automated PR fixes, CI failure remediation, and code review, with cost controls, model options, and security guardrails.</description><pubDate>Tue, 24 Mar 2026 14:36:19 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/how-to-run-claude-code-as-a-github-actions-agent-for-automated-pr-fixes/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/how-to-run-claude-code-as-a-github-actions-agent-for-automated-pr-fixes/image.jpg?v=0b884f43f19cf341&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;A precision repair fixture works on a branching paper strip inside a transparent enclosure with a testing tray and guarded return latch.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;How to wire Claude Code into GitHub Actions for automated PR fixes, CI failure remediation, and code review, with cost controls, model options, and security guardrails.&lt;/p&gt;&lt;p&gt;Berry Mingus · Developer Tools · 10 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/how-to-run-claude-code-as-a-github-actions-agent-for-automated-pr-fixes/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-06-19T00:00:00.000Z</atom:updated><media:content url="https://groundy.com/feed-images/how-to-run-claude-code-as-a-github-actions-agent-for-automated-pr-fixes/image.jpg?v=0b884f43f19cf341" type="image/jpeg" medium="image" width="960" height="540" fileSize="135317"><media:description type="plain">A precision repair fixture works on a branching paper strip inside a transparent enclosure with a testing tray and guarded return latch.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/how-to-run-claude-code-as-a-github-actions-agent-for-automated-pr-fixes/thumb.jpg?v=0b884f43f19cf341" width="640" height="360"/><category>Developer Tools</category><category>claude-code</category><category>github-actions</category><category>ci-cd</category><category>coding-agents</category><category>automation</category><enclosure url="https://groundy.com/feed-images/how-to-run-claude-code-as-a-github-actions-agent-for-automated-pr-fixes/image.jpg?v=0b884f43f19cf341" length="135317" type="image/jpeg"/></item><item><title>Google&apos;s TimesFM: A Foundation Model for Time Series</title><link>https://groundy.com/articles/google-s-timesfm-foundation-model-time/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/google-s-timesfm-foundation-model-time/</guid><description>TimesFM is Google&apos;s decoder-only transformer for zero-shot time-series forecasting, pretrained on ~100 billion real-world points across domains without retraining.</description><pubDate>Fri, 27 Feb 2026 16:43:07 GMT</pubDate><content:encoded>&lt;p&gt;TimesFM is Google&amp;apos;s decoder-only transformer for zero-shot time-series forecasting, pretrained on ~100 billion real-world points across domains without retraining.&lt;/p&gt;&lt;p&gt;Berry Mingus · Models &amp;amp; Research · 9 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/google-s-timesfm-foundation-model-time/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-08-19T00:00:00.000Z</atom:updated><category>Models &amp; Research</category><category>machine-learning</category><category>forecasting</category><category>time-series</category><category>foundation-models</category><category>google-research</category><category>zero-shot</category><category>quantile-forecasting</category></item><item><title>WiFi DensePose: Full-Body Tracking Through Walls Using Your Router</title><link>https://groundy.com/articles/wifi-densepose-full-body-tracking-through-walls-using-your/?utm_source=rss&amp;utm_medium=feed&amp;utm_campaign=popular</link><guid isPermaLink="true">https://groundy.com/articles/wifi-densepose-full-body-tracking-through-walls-using-your/</guid><description>WiFi routers can perform full-body pose estimation through walls using Channel State Information, turning everyday network infrastructure into a covert tracking system.</description><pubDate>Wed, 18 Feb 2026 12:20:50 GMT</pubDate><content:encoded>&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/wifi-densepose-full-body-tracking-through-walls-using-your/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;&lt;img src=&quot;https://groundy.com/feed-images/wifi-densepose-full-body-tracking-through-walls-using-your/image.jpg?v=462830a7660b5b0c&quot; width=&quot;960&quot; height=&quot;540&quot; alt=&quot;Wireless transceivers and faint copper wave contours surround a partition that partly conceals a wooden pose-study mannequin.&quot; style=&quot;max-width:100%;height:auto&quot; /&gt;&lt;/a&gt;&lt;/p&gt;&lt;p&gt;WiFi routers can perform full-body pose estimation through walls using Channel State Information, turning everyday network infrastructure into a covert tracking system.&lt;/p&gt;&lt;p&gt;Berry Mingus · Models &amp;amp; Research · 6 min read&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://groundy.com/articles/wifi-densepose-full-body-tracking-through-walls-using-your/?utm_source=rss&amp;amp;utm_medium=feed&amp;amp;utm_campaign=popular&quot;&gt;Read the full article on Groundy →&lt;/a&gt;&lt;/p&gt;</content:encoded><dc:creator>Berry Mingus</dc:creator><atom:updated>2026-08-19T00:00:00.000Z</atom:updated><media:content url="https://groundy.com/feed-images/wifi-densepose-full-body-tracking-through-walls-using-your/image.jpg?v=462830a7660b5b0c" type="image/jpeg" medium="image" width="960" height="540" fileSize="101679"><media:description type="plain">Wireless transceivers and faint copper wave contours surround a partition that partly conceals a wooden pose-study mannequin.</media:description></media:content><media:thumbnail url="https://groundy.com/feed-images/wifi-densepose-full-body-tracking-through-walls-using-your/thumb.jpg?v=462830a7660b5b0c" width="640" height="360"/><category>Models &amp; Research</category><category>ai-research</category><category>privacy</category><category>surveillance</category><category>computer-vision</category><category>wifi</category><category>rf-sensing</category><enclosure url="https://groundy.com/feed-images/wifi-densepose-full-body-tracking-through-walls-using-your/image.jpg?v=462830a7660b5b0c" length="101679" type="image/jpeg"/></item></channel></rss>