Signal
Advertise
Signal
Advertise
Sign in
Submit
Discover trends that matter
Trending repositories
Daily
Weekly
Monthly
Yearly
Live mentions
Topics
GitHub trending
Repositories
Developers
Insights
Stats
Log in
alesha-pro/llama.cpp-mirai-s — GitHub trending stats & insights | Trendshift
Bifrost
Omnigraph
alesha-pro/llama.cpp-mirai-s
#
Local LLM
llama.cpp with Mirai S (2.4-bit Qwen3.8-27B) support: 128K context on a 12 GB GPU
Visit GitHub
Like alesha-pro/llama.cpp-mirai-s, 0 likes
0
Bookmark alesha-pro/llama.cpp-mirai-s, 0 bookmarks
0
C++
4
1
1 contributors
MIT License
Social mentions
Recent discussions about this repository across the web
Then I went after long context, because my agent runs go past 50K tokens and it got slow there. With q8 or q4 KV on Ampere, llama.cpp runs decode attention in a vector kernel that reads every KV head…
@superalesha · x.com
Repository activities
repository's daily and monthly activities across stars, forks, merged PRs, issues, and closed issues