<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/"><channel><title>Debugging on Bits-Entangled</title><link>https://stondo.github.io/tags/debugging/</link><description>Recent content in Debugging on Bits-Entangled</description><generator>Hugo -- 0.162.1</generator><language>en-us</language><lastBuildDate>Sat, 12 Sep 2026 13:00:00 +0000</lastBuildDate><atom:link href="https://stondo.github.io/tags/debugging/index.xml" rel="self" type="application/rss+xml"/><item><title>The Word Salad Only opencode Could See: Hunting a Phantom Garbage Bug on DeepSeek-V4.1-Flash</title><link>https://stondo.github.io/posts/dsv41-flash-word-salad-top-p-opencode/</link><pubDate>Sat, 12 Sep 2026 13:00:00 +0000</pubDate><guid>https://stondo.github.io/posts/dsv41-flash-word-salad-top-p-opencode/</guid><description>The new V4.1-Flash lane benchmarked clean for hours, then a user report: it starts producing garbage after a while, and only in some CLIs. The hunt ran through four test batteries that found nothing, a tcpdump that captured nothing, a reverse proxy that caught the actual payloads — and ended at a two-line JSON file fixing a sampling default nobody had set.</description></item><item><title>From locklocklock to 262K Context: GLM-5.3-Flash on Two RTX PRO 6000</title><link>https://stondo.github.io/posts/glm-5.3-flash-two-rtx-pro-6000-from-garbage-to-verified/</link><pubDate>Sun, 30 Aug 2026 09:00:00 +0000</pubDate><guid>https://stondo.github.io/posts/glm-5.3-flash-two-rtx-pro-6000-from-garbage-to-verified/</guid><description>GLM-5.3-Flash booted on my two RTX PRO 6000 Blackwell and answered every prompt with deterministic garbage. The investigation ran from a fake kernel bug I wrote myself, through a full exoneration of the linear-attention stack, to a live per-layer bisect that cornered the real culprit in the sparse-MLA prefill path. The fix came from a completely different quantization and kernel stack, and the model now passes 261,900-token needle retrieval on my desk.</description></item></channel></rss>