<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>lilting channel (English)</title><description>Notes on tech and daily life</description><link>https://lilting.ch/</link><language>en-us</language><item><title>Qwen3-Embedding Local CPU vs API: +35% TTS Latency, API Adds None</title><link>https://lilting.ch/en/articles/qwen3-embedding-modelscope-api-tts-contention/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen3-embedding-modelscope-api-tts-contention/</guid><description>Qwen3-Embedding-0.6B on CPU vs ModelScope&apos;s API on a Ryzen 7 5800HS voice server. Local embedding adds up to +2.6s per voice turn by competing for CPU; the API adds none.</description><pubDate>Tue, 08 Sep 2026 16:55:00 GMT</pubDate><category>Qwen</category><category>ModelScope</category><category>RAG</category><category>Vector Search</category><category>Embedding</category><category>SQLite</category><category>Python</category><category>M5Stack</category><category>TTS</category><category>STT</category><category>API</category><category>Experiment</category></item><item><title>Qwen3-Embedding + Qdrant on CPU: 1.1s/query, 1.5-1.7GB RAM</title><link>https://lilting.ch/en/articles/qwen3-embedding-qdrant-cpu-benchmark/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen3-embedding-qdrant-cpu-benchmark/</guid><description>Tested Qwen3-Embedding-0.6B on CPU and Qdrant&apos;s local mode, then checked the 1.5-1.7GB RAM footprint against a Ryzen 7 5800HS voice server with only 4GB VRAM.</description><pubDate>Tue, 08 Sep 2026 15:44:47 GMT</pubDate><category>Qwen</category><category>RAG</category><category>Vector Search</category><category>Embedding</category><category>Python</category><category>M5Stack</category><category>SQLite</category><category>Experiment</category></item><item><title>Anima DiT Residual Steering on M1 Max Flips Flat Color, Not Hair Color Alone</title><link>https://lilting.ch/en/articles/anima-dit-color-binding-probe/</link><guid isPermaLink="true">https://lilting.ch/en/articles/anima-dit-color-binding-probe/</guid><description>Tested on M1 Max 64GB with plain Anima-Base v1.0. A red-to-blue residual direction added to one of Blocks 24-27 flips a flat image, but on a hair mask it paints a blue slab. No hair-only direction found.</description><pubDate>Mon, 07 Sep 2026 16:33:00 GMT</pubDate><category>AI</category><category>Image Generation</category><category>Anima</category><category>Anima-Base</category><category>ComfyUI</category><category>Qwen</category><category>Experiment</category><category>Multi-character</category></item><item><title>Codex on Astra ate my weekly quota in half a day, so I rebuilt the orchestration</title><link>https://lilting.ch/en/articles/codex-orchestration-usage/</link><guid isPermaLink="true">https://lilting.ch/en/articles/codex-orchestration-usage/</guid><description>Codex on Astra hit the weekly usage cap in under 12 hours. Fixes: 120s minimum wait in multi_agent_v2, an 8-field JSON receipt between agents, Astra demoted to advisor, Sol high on top.</description><pubDate>Mon, 07 Sep 2026 08:22:33 GMT</pubDate><category>Codex</category><category>OpenAI</category><category>AI Agents</category><category>Developer Productivity</category><category>Automation</category></item><item><title>How I designed long-term StackChan memory with Qwen3-Embedding and Qdrant</title><link>https://lilting.ch/en/articles/qwen-qdrant-stackchan-memory-design/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen-qdrant-stackchan-memory-design/</guid><description>A pre-implementation design for adding Qwen3-Embedding and Qdrant memory to an RTX 3050 Ti and CoreS3 voice-chat stack without depending on StackChanWorld&apos;s API. It keeps raw logs, searchable memory, persona, and body separate.</description><pubDate>Sun, 06 Sep 2026 11:17:25 GMT</pubDate><category>Qwen</category><category>AI Agents</category><category>M5Stack</category><category>RAG</category><category>TTS</category><category>STT</category><category>Python</category></item><item><title>Qwen-Image and FLUX ran on ModelScope API-Inference without being in the model list</title><link>https://lilting.ch/en/articles/modelscope-api-inference-magicube-probe/</link><guid isPermaLink="true">https://lilting.ch/en/articles/modelscope-api-inference-magicube-probe/</guid><description>My ModelScope account switched from a monthly API-Inference quota to Magicube coins, so I measured what each call costs. Qwen-Image, Z-Image, Krea-2-Turbo and FLUX worked even though none of them appeared in the /v1/models list I got; of the embedding IDs I tried only Qwen3-Embedding went through; audio paths were 404 and Wan produced no video. In these tests 400s and 404s cost nothing, while a 200 with an empty body was charged.</description><pubDate>Fri, 04 Sep 2026 18:05:00 GMT</pubDate><category>Qwen</category><category>ModelScope</category><category>Image Generation</category><category>API</category><category>Experiment</category></item><item><title>What is known about how Eric Lu factored the 862-bit RSA-260 after 35 years</title><link>https://lilting.ch/en/articles/rsa-260-factored-how-computed/</link><guid isPermaLink="true">https://lilting.ch/en/articles/rsa-260-factored-how-computed/</guid><description>862-bit RSA-260 fell on September 3, 2026 after 35 years. I verified the posted factor in Python on an M4 Mac mini, traced the viral &apos;seven months by hand&apos; story to a joke, and sized the compute from RSA-250&apos;s 2,700 core-years with the GNFS formula.</description><pubDate>Fri, 04 Sep 2026 05:49:24 GMT</pubDate><category>Math</category><category>Cryptography</category><category>Security</category></item><item><title>LZ&apos;s 248 keV dark matter candidate event drew five theory papers overnight</title><link>https://lilting.ch/en/articles/lz-248kev-event-five-followup-papers/</link><guid isPermaLink="true">https://lilting.ch/en/articles/lz-248kev-event-five-followup-papers/</guid><description>One 248 keV nuclear recoil in 2.84 tonne-years, global 2.6σ (0.5% background probability), and five arXiv papers within two hours, three landing on a roughly 1 to 1.1 TeV higgsino. What LZ ruled in and out, from the preprint.</description><pubDate>Thu, 03 Sep 2026 05:50:12 GMT</pubDate><category>Research</category><category>Physics</category></item><item><title>Feynman&apos;s two postulates tested with 1.4M photon paths, in high school physics</title><link>https://lilting.ch/en/articles/feynman-path-integral-postulates-photon-test/</link><guid isPermaLink="true">https://lilting.ch/en/articles/feynman-path-integral-postulates-photon-test/</guid><description>Science Advances 2026: 1,419,857 single-photon paths reconstructed to test Feynman&apos;s 1948 postulates directly. Explained with arrows from energy conservation and the double slit, down to what 94% fidelity means.</description><pubDate>Thu, 03 Sep 2026 04:15:07 GMT</pubDate><category>Research</category><category>Physics</category></item><item><title>GPT-OSS keeps its answers after its hidden states are swapped for a TPR formula</title><link>https://lilting.ch/en/articles/emergent-symbolic-structure-tpr-discover/</link><guid isPermaLink="true">https://lilting.ch/en/articles/emergent-symbolic-structure-tpr-discover/</guid><description>McCoy, Soulos, Linzen and Smolensky swap every GPT-OSS input-token hidden state for a closed-form tensor product formula; accuracy drops at most 2.36 points and 31 causal interventions average 0.903.</description><pubDate>Wed, 02 Sep 2026 04:20:00 GMT</pubDate><category>AI</category><category>LLM</category><category>OpenAI</category><category>Research</category><category>Interpretability</category></item><item><title>LEWM predicts the on-screen person&apos;s emotion, not the model&apos;s own</title><link>https://lilting.ch/en/articles/large-emotional-world-model-lewm/</link><guid isPermaLink="true">https://lilting.ch/en/articles/large-emotional-world-model-lewm/</guid><description>LEWM predicts a 7-class emotion label for the human on screen, not a state of its own. Its 45.72% is a cosine-similarity delta vs WorldGPT, and &apos;self-aware&apos; never appears in the paper.</description><pubDate>Tue, 01 Sep 2026 13:14:36 GMT</pubDate><category>AI</category><category>LLM</category><category>Research</category><category>Multimodal</category></item><item><title>Probes arrived 10 minutes after the cohttp path traversal fix PR went public</title><link>https://lilting.ch/en/articles/cohttp-patch-probe-oss-embargo/</link><guid isPermaLink="true">https://lilting.ch/en/articles/cohttp-patch-probe-oss-embargo/</guid><description>Probes for percent-encoded traversal hit the cohttp maintainer&apos;s logs 10 minutes after the fix PR opened. Mandiant now puts mean time-to-exploit at -7 days, VulnCheck&apos;s data disagrees.</description><pubDate>Sun, 30 Aug 2026 16:57:00 GMT</pubDate><category>Security</category><category>Vulnerability</category><category>OSS</category><category>AI Agents</category><category>CVE</category><category>Project Glasswing</category><category>Anthropic</category></item><item><title>Why the Moon gains 56 or 58.7 µs a day, and what CGPM votes on in October</title><link>https://lilting.ch/en/articles/lunar-reference-time-tcl-cgpm/</link><guid isPermaLink="true">https://lilting.ch/en/articles/lunar-reference-time-tcl-cgpm/</guid><description>Draft Resolution D reaches the 28th CGPM in October alongside the leap-second vote. One lunar time scale, unscaled TCL, and why 56 and 58.7 µs/day are both correct.</description><pubDate>Sun, 30 Aug 2026 14:49:55 GMT</pubDate><category>Space</category><category>NASA</category></item><item><title>Anthropic&apos;s Model Hardware Standard took a laser relock from 58% to 99.3%</title><link>https://lilting.ch/en/articles/anthropic-model-hardware-standard-mhs/</link><guid isPermaLink="true">https://lilting.ch/en/articles/anthropic-model-hardware-standard-mhs/</guid><description>MHS is Anthropic&apos;s driver standard for agents running lab hardware. What it standardizes, where MCP sits underneath, and the numbers QuEra, CMU and Genentech reported.</description><pubDate>Sat, 29 Aug 2026 13:26:00 GMT</pubDate><category>MCP</category><category>Anthropic</category><category>Claude</category><category>AI</category><category>AI Agents</category></item><item><title>Qwen3.8-Flash-Next 125B on EVO-X2 gfx1151: tg128 23.26, 128K context ceiling</title><link>https://lilting.ch/en/articles/qwen38-flash-next-evo-x2-rocm-test/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen38-flash-next-evo-x2-rocm-test/</guid><description>Tested on EVO-X2 (gfx1151) under Windows + ROCm: the official b10666 binary runs it at pp512 159.25 / tg128 23.26, and the earlier crash was my own workaround flag.</description><pubDate>Fri, 28 Aug 2026 13:40:00 GMT</pubDate><category>LLM</category><category>Local LLM</category><category>Qwen</category><category>llama.cpp</category><category>AMD</category><category>ROCm</category><category>Experiment</category></item><item><title>Qwen&apos;s capybara mascot: traced to Oct 11, 2024, and the prompts still say bear</title><link>https://lilting.ch/en/articles/qwen-capybara-mascot-timeline/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen-capybara-mascot-timeline/</guid><description>Traced across 31 blog posts and 464 Hugging Face repos: Qwen&apos;s capybara debuted Oct 11, 2024, hit blog banners in Jan 2025, and its own edit prompts all say &apos;the bear&apos;.</description><pubDate>Thu, 27 Aug 2026 10:27:00 GMT</pubDate><category>Qwen</category><category>AI</category><category>HuggingFace</category><category>GitHub</category><category>Anthropic</category></item><item><title>Codex &apos;suspended (tty input)&apos; on macOS: stdio MCP servers keep the tty</title><link>https://lilting.ch/en/articles/codex-suspended-tty-input-mcp/</link><guid isPermaLink="true">https://lilting.ch/en/articles/codex-suspended-tty-input-mcp/</guid><description>On an M4 Mac mini (Codex 0.150.1), stdio MCP servers keep the controlling terminal, so a child reading the tty SIGTTIN-stops the CLI. Shell tools got setsid; MCP didn&apos;t.</description><pubDate>Thu, 27 Aug 2026 06:58:00 GMT</pubDate><category>Codex</category><category>OpenAI</category><category>MCP</category><category>macOS</category><category>Docker</category><category>Terminal</category><category>Experiment</category></item><item><title>125B Qwen3.8-Flash-Next hits 18 tok/s on M1 Max 64GB with N-gram table on SSD</title><link>https://lilting.ch/en/articles/qwen38-flash-next-llamacpp-m1max-test/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen38-flash-next-llamacpp-m1max-test/</guid><description>Tested on M1 Max 64GB: AtomicChat&apos;s M64 GGUF keeps the 51B N-gram table in its own shard, so a llama.cpp PR #27742 build leaves it on SSD and runs the 125B MoE at 17.6 tok/s.</description><pubDate>Thu, 27 Aug 2026 03:05:00 GMT</pubDate><category>LLM</category><category>Local LLM</category><category>Qwen</category><category>llama.cpp</category><category>Apple Silicon</category><category>Experiment</category></item><item><title>Redrawing a B&amp;W icon with a WAI-Anima LoRA, prompts beat i2i for the pose</title><link>https://lilting.ch/en/articles/kana-icon-pose-anima-lora/</link><guid isPermaLink="true">https://lilting.ch/en/articles/kana-icon-pose-anima-lora/</guid><description>Tested on M1 Max 64GB, ComfyUI v0.30.1: i2i and Qwen-Image-Edit lost the character, so I rebuilt the pose one prompt at a time, plus why negatives do nothing at cfg 1.0.</description><pubDate>Wed, 26 Aug 2026 10:26:26 GMT</pubDate><category>LoRA</category><category>AI</category><category>Image Generation</category><category>Anima</category><category>ComfyUI</category><category>Experiment</category></item><item><title>Leap seconds are set to end May 20, 2027, before the first negative leap second</title><link>https://lilting.ch/en/articles/leap-second-abolition/</link><guid isPermaLink="true">https://lilting.ch/en/articles/leap-second-abolition/</guid><description>CGPM votes this October on Draft Resolution C, which stops leap seconds on May 20, 2027, driven by a 30% risk of a negative leap second by 2035. From the 2012 Reddit and 2017 Cloudflare outages to the leap second list still shipping in macOS tzdata.</description><pubDate>Tue, 25 Aug 2026 09:40:00 GMT</pubDate><category>Linux</category><category>Cloudflare</category><category>Go</category></item><item><title>Anima 3.8B with Qwen3.5-4B passes from-behind, gets worse at out-of-frame crops</title><link>https://lilting.ch/en/articles/anima-3-8b-qwen35-prompt-adherence/</link><guid isPermaLink="true">https://lilting.ch/en/articles/anima-3-8b-qwen35-prompt-adherence/</guid><description>Tested on M1 Max 64GB, ComfyUI v0.33.3: from-behind and a back-row drummer now pass, out-of-frame crops get worse, and Anima-Base LoRAs need a 52-block key remap.</description><pubDate>Sun, 23 Aug 2026 20:09:00 GMT</pubDate><category>LoRA</category><category>AI</category><category>Image Generation</category><category>Anima</category><category>Anima-Base</category><category>ComfyUI</category><category>Qwen</category><category>Experiment</category></item><item><title>LLaMA Pro block expansion on Qwen3-0.6B learns my blog but forgets Wikipedia</title><link>https://lilting.ch/en/articles/llama-pro-block-expansion-qwen3-blog-corpus/</link><guid isPermaLink="true">https://lilting.ch/en/articles/llama-pro-block-expansion-qwen3-blog-corpus/</guid><description>Tested on M1 Max 64GB with MLX: LLaMA Pro-style block expansion on Qwen3-0.6B-Base vs LoRA vs full FT, trained on 722 blog posts. Held-out PPL drops 14.7 to 10.5, but Wikipedia-ja rises 13.5 to 17.9, more forgetting than full FT. A 20% Wikipedia mix nearly removes it, and renumbering 28-layer LoRA keys reproduces the Anima-2.9B result in numbers.</description><pubDate>Sun, 23 Aug 2026 08:06:00 GMT</pubDate><category>LLM</category><category>MLX</category><category>Qwen</category><category>Anima</category><category>Apple Silicon</category><category>LoRA</category><category>AI</category><category>Experiment</category></item><item><title>Qwen-MM-Plugins core in Claude Code and Codex, PDF/video/STL with no API key</title><link>https://lilting.ch/en/articles/qwen-mm-plugins-core-claude-code-codex/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen-mm-plugins-core-claude-code-codex/</guid><description>16GB M4 Mac mini test of QwenLM&apos;s Qwen-MM-Plugins core. Claude Code adds a plugin_ prefix that breaks --allowedTools, codex exec rejects MCP calls by default. PDF, STL and video read with no API key.</description><pubDate>Sat, 22 Aug 2026 05:20:00 GMT</pubDate><category>Qwen</category><category>MCP</category><category>Claude Code</category><category>Codex</category><category>AI Agents</category><category>Multimodal</category><category>Experiment</category></item><item><title>How Z3 found the leap-year magic number 1073750999</title><link>https://lilting.ch/en/articles/leap-year-magic-number-z3/</link><guid isPermaLink="true">https://lilting.ch/en/articles/leap-year-magic-number-z3/</guid><description>A three-instruction leap-year test hides the constants 1073750999, 3221352463, and 126976. I rebuilt an 18-bit version with Z3, then traced how the multiply, mask, and threshold encode divisibility by 4, 100, and 400.</description><pubDate>Sat, 22 Aug 2026 05:11:50 GMT</pubDate><category>C</category><category>Algorithms</category><category>Math</category><category>Experiment</category></item><item><title>Renumbered keys bring Anima-Base LoRAs back on 40-layer Anima-2.9B, 0/3 to 3/3</title><link>https://lilting.ch/en/articles/anima-2-9b-layer-expansion-lora-compat/</link><guid isPermaLink="true">https://lilting.ch/en/articles/anima-2-9b-layer-expansion-lora-compat/</guid><description>Tested on M1 Max 64GB, ComfyUI v0.33.3. Anima-2.9B inserts 12 DiT blocks, so Anima-Base LoRA keys hit the wrong layers with zero warnings, 0/3. Renumbering the keys gives 3/3.</description><pubDate>Sat, 22 Aug 2026 04:50:00 GMT</pubDate><category>LoRA</category><category>AI</category><category>Image Generation</category><category>Anima</category><category>Anima-Base</category><category>ComfyUI</category><category>Qwen</category><category>Experiment</category></item><item><title>Disabling the FUJITSU logo fixed USB boot freezes on my ESPRIMO G5010/E</title><link>https://lilting.ch/en/articles/esprimo-g5010-fujitsu-logo-usb-freeze/</link><guid isPermaLink="true">https://lilting.ch/en/articles/esprimo-g5010-fujitsu-logo-usb-freeze/</guid><description>My Fujitsu ESPRIMO G5010/E froze at the FUJITSU logo with any bootable USB inserted. Disabling the boot logo in BIOS fixed it. D3804-A1x BIOS changelog included.</description><pubDate>Thu, 20 Aug 2026 17:15:00 GMT</pubDate><category>Linux</category><category>Ubuntu</category><category>UEFI</category></item><item><title>Ben Joffe&apos;s Mersenne weekday trick vs plain %7 in clang, V8, and CPython on M4</title><link>https://lilting.ch/en/articles/fast-day-of-week-mersenne-benchmark/</link><guid isPermaLink="true">https://lilting.ch/en/articles/fast-day-of-week-mersenne-benchmark/</guid><description>Benchmarked on an M4 Mac mini: Ben Joffe&apos;s 2-instruction weekday hack beats plain %7 by 1.6-6.4x in clang, Rust, and V8, loses 3x in CPython. Plus a 25x V8 -0 deopt trap.</description><pubDate>Thu, 20 Aug 2026 09:10:00 GMT</pubDate><category>C</category><category>Rust</category><category>JavaScript</category><category>Python</category><category>Algorithms</category><category>Benchmark</category><category>Apple Silicon</category><category>Math</category><category>Experiment</category></item><item><title>Qwen3.8-27B on Strix Halo (EVO-X2): ROCm llama.cpp + MTP at 22 tok/s in Q8_0</title><link>https://lilting.ch/en/articles/qwen38-27b-evo-x2-rocm-mtp-test/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen38-27b-evo-x2-rocm-mtp-test/</guid><description>Tested on GMKtec EVO-X2 (Ryzen AI Max+ 395): Q8_0 + MTP beats a 4-bit M1 Max at 22 tok/s, thinking burns 32,712 chars before any HTML, and the NSFW refusal line moves.</description><pubDate>Wed, 19 Aug 2026 07:41:00 GMT</pubDate><category>LLM</category><category>Local LLM</category><category>Qwen</category><category>llama.cpp</category><category>AMD</category><category>ROCm</category><category>Experiment</category></item><item><title>Qwen3.8-27B on M1 Max 64GB: MLX vs Ollama, and a 50k-char Thinking Runaway</title><link>https://lilting.ch/en/articles/qwen38-27b-mlx-ollama-m1max-test/</link><guid isPermaLink="true">https://lilting.ch/en/articles/qwen38-27b-mlx-ollama-m1max-test/</guid><description>Tested on M1 Max 64GB: Qwen3.8-27B hits ~19 tok/s on both MLX and Ollama, but the default reasoning_effort=xhigh blew thinking up to 50,373 chars. Why Ollama dodges it.</description><pubDate>Tue, 18 Aug 2026 07:00:04 GMT</pubDate><category>LLM</category><category>Local LLM</category><category>Qwen</category><category>Ollama</category><category>MLX</category><category>Apple Silicon</category><category>Experiment</category></item><item><title>GitHub Advisory Database Still Shows the Rejected Fake SQLite CVEs as Critical</title><link>https://lilting.ch/en/articles/sqlite-fake-cve-advisory-database-unreviewed/</link><guid isPermaLink="true">https://lilting.ch/en/articles/sqlite-fake-cve-advisory-database-unreviewed/</guid><description>Checked August 17, 2026 via the GitHub Advisory Database API: all six fake SQLite CVEs rejected by MITRE still show CVSS up to 9.8 as type unreviewed, withdrawn_at null.</description><pubDate>Mon, 17 Aug 2026 04:28:34 GMT</pubDate><category>Security</category><category>CVE</category><category>Vulnerability</category><category>SQLite</category><category>AI</category><category>OSS</category><category>GitHub</category></item></channel></rss>