Santiago Valdarrama
@svpino
This open-source model scores 95.4% on BIG-Bench Hard’s logic subset with just 3.66B parameters.
TwIL-LM3-Pro is a model small enough to run locally on a laptop.
Specialized post-training is super cool. You can get a ton of small models when you focus on specific tasks.