<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>The Thinking Layer</title><description>Interactive, numbers-first explainers on how AI actually works: models, infrastructure, research and real-world AI.</description><link>https://thethinkinglayer.com/</link><item><title>One token through a 2.4-trillion-parameter model</title><link>https://thethinkinglayer.com/posts/one-token-through-qwen/</link><guid isPermaLink="true">https://thethinkinglayer.com/posts/one-token-through-qwen/</guid><description>Every word an LLM writes costs a full trip through its weights. I followed one prompt through Qwen3.8-Max to see what that trip actually looks like, in bytes, GPUs and memory.</description><pubDate>Tue, 29 Sep 2026 00:00:00 GMT</pubDate></item></channel></rss>