<span class="vcard">/u/Creative-Fig522</span>
/u/Creative-Fig522

1.7B model leading strict-7 formal reasoning above Qwen3-8B and Gemma-4-26B – specialists eating generalist territory?

Most of the reasoning gains coming out of the big labs are still tied to scale. More params, more compute, better reasoning. That's been the play for a while. Ran into TwIL-LM2 which flips the script for narrow tasks. PEFT LoRA adapter on SmolLM2-1…

1.7B model leading strict-7 formal reasoning above Qwen3-8B and Gemma-4-26B – specialists eating generalist territory?

Most of the reasoning gains coming out of the big labs are still tied to scale. More params, more compute, better reasoning. That's been the play for a while. Ran into TwIL-LM2 which flips the script for narrow tasks. PEFT LoRA adapter on SmolLM2-1…