Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

If I am reading this right, the 7b model performs as well as qwen3.6-35b-a3b at coding?

K2 horizon 7b scores 70.6 on swe-bench-verified.

Qwen3.6-35b-a3b scores a 70.0 on swe-bench-verified.

That’s pretty interesting. I assume the benchmark and reality don’t line up, but i’m downloading it now to find out.

If it’s anywhere near true, it unlocks local llm coding on a whole new class of machines (anything with 8gb vram).

 help



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: