BASIC-Prefetcher: Bin-based Address and Size-Informed Caching for AI-Driven SSD Workloads
Read latency is a critical bottleneck for NAND-based SSDs in AI-driven datacenter workloads, where model parameters and key-value data are frequently swapped between main memory and storage. Existing prefetching schemes operate on block-level address sequences that have been stripped of application context by the files...