Back to News
Advertisement
Advertisement

⚑ Community Insights

Discussion Sentiment

50% Positive

Analyzed from 117 words in the discussion.

Trending Topics

#harness#test#models#claude#code#more#model#interesting#paired#same

Discussion (2 Comments)Read Original on HackerNews

HarHarVeryFunnyβ€’about 4 hours ago
This seems a strange thing to test given that Claude Code is optimized for Anthropic models.

A test of how different models perform in more of a model-agnostic harness like OpenCode would be much more interesting, perhaps paired with a control experiment of how those same models performed on the same task when using their respective native harnesses.

rpdillonβ€’about 3 hours ago
Yeah, it's an interesting test because he gave Kimi credit for running inside the harness he liked, which I presume is Claude Code. But he also said it was noted for its harness sensitivity, and then said it didn't do as well as Opus.

A lot of folks don't seem to understand that it's the harness and its behavior paired with a particular model that produces the result, and you have to focus on both aspects. He barely acknowledges the importance of the harness in his write-up, beyond the two notes above.