Pesquisa | Portal Regional da BVS (teste)

Deep random forest with ferroelectric analog content addressable memory.

Yin, Xunzhao; Müller, Franz; Laguna, Ann Franchesca; Li, Chao; Huang, Qingrong; Shi, Zhiguo; Lederer, Maximilian; Laleni, Nellie; Deng, Shan; Zhao, Zijian; Imani, Mohsen; Shi, Yiyu; Niemier, Michael; Hu, Xiaobo Sharon; Zhuo, Cheng; Kämpfe, Thomas; Ni, Kai.

Sci Adv ; 10(23): eadk8471, 2024 Jun 07.

Artigo em Inglês | MEDLINE | ID: mdl-38838137

RESUMO

Deep random forest (DRF), which combines deep learning and random forest, exhibits comparable accuracy, interpretability, low memory and computational overhead to deep neural networks (DNNs) in edge intelligence tasks. However, efficient DRF accelerator is lagging behind its DNN counterparts. The key to DRF acceleration lies in realizing the branch-split operation at decision nodes. In this work, we propose implementing DRF through associative searches realized with ferroelectric analog content addressable memory (ACAM). Utilizing only two ferroelectric field effect transistors (FeFETs), the ultra-compact ACAM cell performs energy-efficient branch-split operations by storing decision boundaries as analog polarization states in FeFETs. The DRF accelerator architecture and its model mapping to ACAM arrays are presented. The functionality, characteristics, and scalability of the FeFET ACAM DRF and its robustness against FeFET device non-idealities are validated in experiments and simulations. Evaluations show that the FeFET ACAM DRF accelerator achieves â¼106×/10× and â¼106×/2.5× improvements in energy and latency, respectively, compared to other DRF hardware implementations on state-of-the-art CPU/ReRAM.

Ferroelectric compute-in-memory annealer for combinatorial optimization problems.

Yin, Xunzhao; Qian, Yu; Vardar, Alptekin; Günther, Marcel; Müller, Franz; Laleni, Nellie; Zhao, Zijian; Jiang, Zhouhang; Shi, Zhiguo; Shi, Yiyu; Gong, Xiao; Zhuo, Cheng; Kämpfe, Thomas; Ni, Kai.

Nat Commun ; 15(1): 2419, 2024 Mar 18.

Artigo em Inglês | MEDLINE | ID: mdl-38499524

RESUMO

Computationally hard combinatorial optimization problems (COPs) are ubiquitous in many applications. Various digital annealers, dynamical Ising machines, and quantum/photonic systems have been developed for solving COPs, but they still suffer from the memory access issue, scalability, restricted applicability to certain types of COPs, and VLSI-incompatibility, respectively. Here we report a ferroelectric field effect transistor (FeFET) based compute-in-memory (CiM) annealer for solving larger-scale COPs efficiently. Our CiM annealer converts COPs into quadratic unconstrained binary optimization (QUBO) formulations, and uniquely accelerates in-situ the core vector-matrix-vector (VMV) multiplication operations of QUBO formulations in a single step. Specifically, the three-terminal FeFET structure allows for lossless compression of the stored QUBO matrix, achieving a remarkably 75% chip size saving when solving Max-Cut problems. A multi-epoch simulated annealing (MESA) algorithm is proposed for efficient annealing, achieving up to 27% better solution and ~ 2X speedup than conventional simulated annealing. Experimental validation is performed using the first integrated FeFET chip on 28nm HKMG CMOS technology, indicating great promise of FeFET CiM array in solving general COPs.

RESUMO

RESUMO

ENVIAR RESULTADO:

SELEÇÃO DE REFERÊNCIAS

DETALHE DA PESQUISA