Rohan Paul

@rohanpaul_ai

China's best small reasoning model (VibeThinker-3B) just got beaten by a 3.6B model from Austin. > webAI's TwIL-LM3-Pro, a 3.66B local model, lifts IBM Granite's formal-logic score by 28% > the recommended Q4 build is a 2.09GiB file that runs through llama.cpp on CPU or local GPU, so private data can stay on the device. > TwIL-LM3-Pro now leads every small model webAI compared on formal logic, scoring roughly 35% above Weibo's VibeThinker-3B, 24% above Qwen3.5-4B and 47% above Liquid AI's LFM2.5-8B-A1B.
打开原帖#511482
  1. Industry

    Databricks: Agentic apps are putting new pressure on the data stack
  2. Industry

    Cursor: Innate is a team of 7 building capable robots for everyday life
  3. Industry

    Runway: Runway Head of Robotics Andy Chen laid out Runway's unique approach t…