Rohan Paul
@rohanpaul_ai
Together, that's what separates infrastructure from a demo. Recognition that knows your domain, one stack for every latency target, and a design that treats context as a first-class input rather than an afterthought.
Open source from NetEase Youdao.
The main question is whether stable, domain-aware streaming recognition becomes normal for voice systems.