Advertisement
Advertisement
⚡ Community Insights
Discussion Sentiment
57% Positive
Analyzed from 401 words in the discussion.
Trending Topics
#between#https#tiny#token#local#prompt#lean#measured#results#harnesses
Discussion Sentiment
Analyzed from 401 words in the discussion.
Trending Topics
Discussion (15 Comments)Read Original on HackerNews
I'd love to see a tiny, reproducible benchmark repo that anyone can drop on their own hardware and then run against all harnesses at once to compare the per turn prefix token count, time to the first token, experienced tokens/sec (and prefill), cache reuse % and a pass rate on a deterministic set of small tasks. I think it could also be useful to have some way to share results and hardware for others to compare.
it grew out of annoyance of dependencies on js runtimes, probably similar to you. mine additionally works on solaris and esp32.
could be interesting to collaborate!
"it spreads up to 50% between nights, so nothing between the lean arms is a finding."
https://m.youtube.com/watch?v=c_fQoDkULl0 (see around 8:00)
“Chad” initially looked interesting but the minute I saw the ai-written markdown and giant commit I just left. I just can’t bring myself to read someone elses’ slop, regardless of performance.
If all a developer hand writes is a truthy and readable markdown document, I really don’t care if the rest of the project is vibe coded, but I struggle to get interested in AI generated summaries and docs.