Signal
Advertise
Signal
Advertise
Sign in
Submit
Discover trends that matter
Trending repositories
Daily
Weekly
Monthly
Yearly
Live mentions
Topics
GitHub trending
Repositories
Developers
Insights
Stats
Log in
perkel666/MegaCapybara — GitHub trending stats & insights | Trendshift
Busbar
Bifrost
Kane CLI
perkel666/MegaCapybara
#
AI infrastructure
#
LLM inference
The fastest inference engine for Qwen3.8-27B on the NVIDIA RTX 5090: up to 500 tokens/s for one agent and up to 2,000 tokens/s for many, contexts up to 1M tokens, and a launcher that shows what every setting costs. Windows and Linux.
Visit GitHub
Like perkel666/MegaCapybara, 0 likes
0
Bookmark perkel666/MegaCapybara, 0 bookmarks
0
19
2
2 contributors
MIT License
website
Social mentions
Recent discussions about this repository across the web
The fastest interference engine for RTX5090 and Qwen3.8 27B. Twice as fast as ninfer. 500+ t/s single coding, 2000+t/s up to 12 agents at the same time with 800k context. Smart VRAM-RAM-DISC Cache…
r/LocalLLaMA
Repository activities
repository's daily and monthly activities across stars, forks, merged PRs, issues, and closed issues