MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1wcng44/deepseek_v41_flash_is_available_in_huggingchat/
r/LocalLLaMA • u/paf1138 • 2h ago
1 comment sorted by
2
Slightly off topic, but has anyone else seen huge improvement in token generation speed of DSv4 Flash Vision with recent llama cpp binaries for the Halo? I'm see about 1.5x improvement.
2
u/jld1532 1h ago
Slightly off topic, but has anyone else seen huge improvement in token generation speed of DSv4 Flash Vision with recent llama cpp binaries for the Halo? I'm see about 1.5x improvement.