Andrew Lin is Deputy Director at Phison Electronics, where he has spent his career on NAND flash translation layer (FTL) design. He currently focuses on enterprise QLC SSDs and system architecture — including work like aiDAPTIV+, which uses flash as an intelligent extension of GPU memory for large-scale AI workloads.
The AI ecosystem has successfully scaled past what can be supported with DRAM. We will show that inference performance can be greatly improved and the DRAM footprint can be reduced with the introduction of a large Flash tire. The presentation will highlight the benefits and explain the experiments used to validate this claim. The advantages apply to client solutions all the way up to full enterprise deployments.