Santiago Valdarrama

@svpino

A fine-tuned model can outperform frontier models and be cheaper and faster to run. Literally, every company I've met wants this. I want you to see these results from fine-tuning Qwen3 4B on AWS. It smokes both the out-of-the-box model and Claude Sonnet 4.6.
打开原帖#511482
  1. Research

    François Chollet: The critical distinction between base LLMs (2024 and earlier) and mod…
  2. Industry

    Michael Truell: Grok Bot gets useful work done, without you needing to ask.
  3. Industry

    Aravind Srinivas: We're open-sourcing our multimodal Decision model and offering it thr…