Signal
Advertise
Signal
Advertise
Sign in
Submit
Discover trends that matter
Trending repositories
Daily
Weekly
Monthly
Yearly
Live mentions
Topics
GitHub trending
Repositories
Developers
Insights
Stats
Log in
kwindla/aiewf-eval — GitHub trending stats & insights | Trendshift
Featured
open-connector
kwindla/aiewf-eval
#
NLP
A long-context eval
Visit GitHub
Python
154
19
3 contributors
Social mentions
Recent discussions about this repository across the web
On my 30-turn voice conversation benchmark, Muse Glimmer doesn't do well. This benchmark tests long context recall and tricky tool calling judgment in a back-and-forth conversation. The "agentic"…
@kwindla · x.com
Cerebras inference is very fast. So fast that it changes how we think about configuring our LLMs for voice agent use cases. Kimi K2.6 is a 1T parameter reasoning model that @cerebras serves at 650 -…
@kwindla · x.com
Gemini 3.5 Flash is out today. Here are numbers from my main voice and task agent benchmarks. Some notes: All the Gemini 3 models so far are too slow to work well for voice agents. Gemini 2.5 Flash…
@kwindla · x.com
Repository activities
repository's daily and monthly activities across stars, forks, merged PRs, issues, and closed issues