~/ai-muninn
search
⌘K
blog
model anatomy
github
中
~ / blog
/
tag / dflash2
❯
grep -r "#dflash2" ~/blog
2 matches
date
read
title
2026-09-21
23m
Two Modded RTX 2080 Tis Hit 153.8 tok/s on Qwen3.8-27B With FastLLM + DFlash2
#fastllm
#dflash2
#speculative-decoding
#2080-ti
2026-08-23
20m
[Benchmark] Qwen3.8-27B hits 65 tok/s on one DGX Spark — 12.3 without speculative decoding
#dgx-spark
#gb10
#sm121
#nvfp4
← back to all posts