Detecting AI-Generated Audio: Why Decay Tails Betray Voice Clones
A new preprint shows AI audio leaks in decay tails via group delay. It offers a cheap, watermark-free filter for fraud screening, though hold-out accuracy is only 66.7%.
The Groundy archive · Page 4 of 34
Browse Groundy's complete archive of 797 articles on AI, developer tools and infrastructure. Page 4 of 34.
73–96 of 797 articles · Newest first
A new preprint shows AI audio leaks in decay tails via group delay. It offers a cheap, watermark-free filter for fraud screening, though hold-out accuracy is only 66.7%.
OpenAI is cutting off Cursor's model supply by Nov 12, 2026. This exposes the uncontracted data hop between your IDE and the lab, forcing a governance audit and fallback plan.
SpaceX's acquisition of Cursor exposes a structural risk: AI editors are thin layers over third-party model APIs. Export agent histories and warm a second CLI to mitigate exit
Rule checkers miss format variants while LLM judges get hacked during RL. This failure map from arXiv 2505.22203 shows why hybrid verifiers and reward audits are now essential
LLMs4OL 2026 results show post-hoc vocabulary filtering achieves 0.92 F1 but bans non-taxonomic relations. Learn how to manage the term-governance burden this creates.

GLM-5.2 on vLLM v0.28.0: enable MTP, fix k_norm warnings, and verify FP8 precision. Measured gains from research on other models, not GLM-5.2.
Hugging Face's RFI response contrasts open-weight contract terms with proprietary API restrictions. Teams should choose models based on fees, usage limits, and entity bans, as
TutorTrace data shows AI adoption metrics fail to measure junior skill. Use behavioral windows and guided help-seeking states to judge code review evidence instead of dash.
J-Zero claims zero-data self-play beats baselines by 8.0 points on unverifiable tasks, but judge verdict flips of 5.3 to 48.4 percent suggest the gains may be self-validated.
A new preprint shows 77% of prompt injection detector decisions flip on single-token removal. Probe your own traffic slices to separate calibration drift from exploitable gaps
KubeCap automates Kubernetes capability minimization via LLM-inferred rules, cutting permissions by 55% in Go workloads. Validate inferred drops against runtime behavior to CI
Alibaba's ScaleSense uses learned query-level estimation to cut AnalyticDB costs by up to 5.22x. This audit covers the risks of opaque autoscalers and the limits of vendor-own
A new preprint argues agent benchmark scores track the evaluation harness as much as the model. Teams should pin harness versions and ablate scaffold changes to avoid misat.
A new arXiv survey shows temperature 0 does not guarantee reproducible financial AI outputs. Hardware, batching, and parallelism cause divergence that sampling controls miss,
A founder's Cognito postmortem highlights hidden auth costs. This guide maps decision axes for startups, separating verified AWS capabilities from unverified customization and
DeepPlanner trains planning into research agents via RL, but single-source preprint data demands caution. Compare prompt, graph, and weights layers to decide when fine-tuning.
arXiv 2508.06753 reports 2-bit LLM serving with up to 7x speedups on Intel Xe2. For self-hosters, the binding cost is not memory but task-specific eval coverage to verify QAT-
Local coding LLMs hallucinate package names at rates up to 73% on adversarial prompts. A new preprint compares defense layers, showing that dependency safety must shift from a
Provider APIs, client libraries, and MCP standardize different parts of the agent tool stack. Route by fleet shape to avoid schema drift and keep one authoritative source for
PayPal's non-bank status and lack of FDIC insurance expose open source projects to sudden funding freezes. This guide details how to engineer payment redundancy and stage a no
Cloudflare claims 100 TB saved in 1.1.1.1 cache, but the mechanism is unverified. Operators can cut memory now by tuning Redis eviction policies and admission controls, not by
Compare LocalStack, Azurite, and CloudEmu for CI cloud testing. LocalStack now requires an auth token; Azurite covers only storage. Synthesis shifts fidelity risk from code to
An ICML 2026 paper shows RL can learn token boundaries end-to-end, beating straight-through baselines at 100M parameters. For serving or fine-tuning, keep your frozen BPE or S

GLM-5.3-Flash and Qwen3.8-Flash-Next lack confirmed pricing and independent evals. Hold your router, verify per-token costs, and run a 50-case domain holdout before switching.