K2 horizon 7b scores 70.6 on swe-bench-verified.
Qwen3.6-35b-a3b scores a 70.0 on swe-bench-verified.
That’s pretty interesting. I assume the benchmark and reality don’t line up, but i’m downloading it now to find out.
If it’s anywhere near true, it unlocks local llm coding on a whole new class of machines (anything with 8gb vram).
K2 horizon 7b scores 70.6 on swe-bench-verified.
Qwen3.6-35b-a3b scores a 70.0 on swe-bench-verified.
That’s pretty interesting. I assume the benchmark and reality don’t line up, but i’m downloading it now to find out.
If it’s anywhere near true, it unlocks local llm coding on a whole new class of machines (anything with 8gb vram).