Signal
Advertise
Signal
Advertise
Sign in
Submit
Discover trends that matter
Trending repositories
Daily
Weekly
Monthly
Yearly
Live mentions
Topics
GitHub trending
Repositories
Developers
Insights
Stats
Log in
MohammadHaishemKhawaja/hashyy — GitHub trending stats & insights | Trendshift
Kane CLI
Bifrost
Busbar
MohammadHaishemKhawaja/hashyy
#
LLM inference
A full, unpruned 177B MoE model at 11.5 tok/s on one 12 GB GPU. Expert streaming for llama.cpp, with the measurement harness and every dead end.
Visit GitHub
Like MohammadHaishemKhawaja/hashyy, 0 likes
0
Bookmark MohammadHaishemKhawaja/hashyy, 0 bookmarks
0
Python
9
3
2 contributors
Custom license
Social mentions
Recent discussions about this repository across the web
Qwen3.8-Flash-Next 177B at 11–15 tok/s on a single RTX 5070 12GB with 32GB DDR4 RAM [P]
r/LLM
Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4
r/OpenSourceeAI
Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4
r/LLM
Qwen3.8-Flash-Next 177B at 11–15 tok/s on a single RTX 5070 12GB with 32GB DDR4 RAM
r/ArtificialInteligence
Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4
r/Qwen_AI
Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4
r/u_Tommey_DZ
Qwen3.8-Flash-Next 177B running at 11–15 tok/s on a single RTX 5070 12GB + 32GB RAM DDR4
r/LocalLLaMA
RTX 5070 12GB + 32GB RAM Qwen3.8-Flash-Next 177B running at ~ 11-15 tok/s
r/nvidia
Qwen3.8-Flash-Next 177B running at ~11–15 tok/s on a single RTX 5070 12GB + 32GB RAM
r/LocalLLM
Repository activities
repository's daily and monthly activities across stars, forks, merged PRs, issues, and closed issues