Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

They significantly lowered latency compared to EPYC/Xeon, which is critical for streaming agents (e.g. text/audio/video agents).


What latency? How much is it compared to LLM inference speed?


See the Redpanda comment/link here.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: