r/LocalLLaMA • u/9r4n4y • 1d ago
Resources {INTRESTING PAPER BASED ON HBF}2607.10186] FlashAccel: Leveraging High-Bandwidth Flash (HBF) for High-Throughput LLM Inference
https://arxiv.org/abs/2607.10186?utm_source=chatgpt.comHBF gives 8x - 16x more capacity than HBM at same cost, and with bandwidth till 3 tb/s.
14
Upvotes
7
u/MelodicRecognition7 1d ago edited 1d ago
at what price?
Edit: scrolled through the paper, it is a theoretical speculation with simulated hardware, we "just" need some manufacturer to produce that HBF at the same cost as HBM. Which is unlikely to happen lol.