Perplexity

@perplexity_ai

Inside Photon, p99 response time dropped from about 800 ms to about 65 ms. It uses about 20% fewer serving machines and stores 2.5x as much data per document. How we built Photon: https://t.co/7c64kR5OB5 Get started with Fast Search: https://t.co/Qn4OPZfqnx
打开原帖#511482
  1. Industry

    Alexandr Wang: “hey muse, can you get the gulfstream in position for me to go to my…
  2. Industry

    Aravind Srinivas: $1 for a 1000 requests at 200 ms
  3. Industry

    Alexandr Wang: elon: “and meta AI, the only reason i’m here is because you are a fri…