Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

From the paper:

"Our CXL solution achieves substantial gains for diverse workloads, including up to a 25% reduction in server count for disaggregated ML inference"

How does using worse RAM result in 25% reduction of server count for given workloads?



Because it's used in addition to, not in place of, the better RAM.


CXL is adding "slow" RAM over PCIe, basically. Not replacing.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: