tag: #llm
(10 posts)
Show all tags
-
2026-08-21
:: MTP speculative decoding costs code correctness on two Qwen-lineage models
::
#agentic-ai,
#ai,
#llm,
#self-hosted
-
2026-08-21
:: Ornith 1.5 35B A3B on Strix Halo: MTP pays off at n=1, not at the default n=3
::
#agentic-ai,
#ai,
#hardware,
#llm,
#self-hosted
-
2026-08-12
:: The small model didn't fabricate. It stopped citing.
::
#agentic-ai,
#ai,
#llm,
#self-hosted
-
2026-08-07
:: AMD buys Taalas: etched models, and what it means for local inference
::
#agentic-ai,
#ai,
#hardware,
#llm,
#self-hosted
-
2026-07-27
:: Poolside says Laguna needs an H200. We ran it on a single H100.
::
#agentic-ai,
#ai,
#hardware,
#llm,
#self-hosted
-
2026-07-26
:: DeepSeek-V4-Flash on Strix Halo: it runs, and now we know how fast
::
#ai,
#hardware,
#llm,
#self-hosted
-
2026-07-26
:: Laguna-S-2.1 on a mini-PC: the honest numbers
::
#agentic-ai,
#ai,
#hardware,
#llm,
#self-hosted
-
2026-06-15
:: Running an H100 at Trail Openers: what it actually costs in money, energy, and CO₂
::
#agentic-ai,
#ai,
#hardware,
#llm,
#self-hosted,
#sustainability
-
2026-06-05
:: H100 vs Strix Halo: the gap is bigger than the first benchmark suggested
::
#agentic-ai,
#ai,
#hardware,
#llm,
#self-hosted,
#sustainability
-
2026-04-30
:: New Strix Halo? Five things that will cost you hours.
::
#ai,
#hardware,
#llm,
#self-hosted