כתבה
arXiv cs.LG ·
NeuroFlex: Lossless Element-Level ANN-SNN Co-Execution for Efficient Sparse Inference
תקציר מקורי באנגליתarXiv:2609.14092v1 Announce Type: cross Abstract: Sparse DNN accelerators specialize in ANN or SNN execution, leaving energy or latency on the table when workload characteristics vary within a layer. Hybrid accelerator designs that switch modes at layer or tile granularity suffer from low PE utilization since one core type idles whenever the other is active. NeuroFlex is the first accelerator to assign every output element independently to ANN or SNN execution mode with zero accuracy loss. We extend integer-exact ANN-SNN equivalence from layers to individual output elements, thereby enabling mode switching with no conversion error. An offline cost-guided scheduler scores each element by its marginal energy-delay trade-off and packs work across PEs, achieving 97-99% PE utilization compared
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית