Charles Frye

@charles_irl

New on the @modal blog: a write-up by me and Shreya on our work to speed up AI-SQL queries - why AI-SQL? why high-throughput inference? - how can we possibly be >10x faster than vLLM? - what does the future of inference engineering look like? Read here: https://modal.com/blog/quail-billion-tpm
打开原帖#511482
  1. Research

    Andrew Ng: The loudest voices stoking fears about AI dangers have made tremendou…
  2. Research

    Jeff Dean: Proud to have collaborated with many others on quite a few of these t…
  3. Research

    Jeff Dean: The safety data for Waymo gets better and better