#
HBF# is emerging as a complementary technology to #
HBM# rather than a competing one, with the potential to address HBM's high cost and limited capacity. While HBM delivers unmatched memory bandwidth, it is difficult for it alone to meet the massive capacity requirements of #
LLM# inference. HBF, on the other hand, leverages stacked Flash technology to provide sufficient bandwidth while offering several times the memory capacity of HBM at a significantly lower cost.
Although #
NVIDIA# has yet to announce any plans to adopt the technology, industry sentiment remains broadly positive. The interest shown by industry leaders such as #
Google# and #
Tenstorrent# have shown a willingness to consider adopting HBF, though specific use cases have not yet been determined. #
TrendForce# believes the long-term architecture will likely combine HBM for ultra-high-speed computation with HBF for high-density data storage, creating a hybrid memory hierarchy that could become a key enabler of large-scale AI commercialization.