Signal
Advertise
Signal
Advertise
Sign in
Submit
Discover trends that matter
Trending repositories
Daily
Weekly
Monthly
Yearly
Live mentions
Topics
GitHub trending
Repositories
Developers
Insights
Stats
Log in
halt95/qwen38-flash-next-3090s — GitHub trending stats & insights | Trendshift
Kane CLI
Bifrost
Busbar
halt95/qwen38-flash-next-3090s
Qwen3.8-Flash-Next at its full 262K context on 4x RTX 3090 with vLLM: 806,792-token KV pool, three 262K sessions resident, MTP, host-mapped PLE, pinned build and container recipes.
Visit GitHub
Like halt95/qwen38-flash-next-3090s, 0 likes
0
Bookmark halt95/qwen38-flash-next-3090s, 0 bookmarks
0
Python
12
2
1 contributors
Apache License 2.0
website
Social mentions
Recent discussions about this repository across the web
Flash-Next v2.5.1 is out: Qwen3.8-Flash-Next (125B MoE) on 4× RTX 3090, no NVLink. • Agent follow-ups resume from cache: 0.96 s to first token • 8 agents at once: 92.5 % cache reuse • 925K-token KV…
@Funball2 · x.com
Further improved Running Qwen3.8 Flash-Next: Updated to vllm 0.30 Opt-in Repeatable greedy output Still 3 x 262k sessions running at once in the kv pool Opt-in GPU/NUMA auto-select for other box…
@Funball2 · x.com
Usually tweet once in a blue moon to complain at a company. Deciding to shift gears. Finally managed to get Qwen3.8-Flash-Next running on 4x RTX 3090s: 807K KV pool, three full 262,144-token sessions…
@Funball2 · x.com
Repository activities
repository's daily and monthly activities across stars, forks, merged PRs, issues, and closed issues