יום שלישי, 15 בספטמבר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

דחיסה סמנטית דינמית

Dynamic Semantic Compression for Efficient Latent-Space Inference in Large Language Models
חוקרים הציגו שיטה חדשה לדחיסה סמנטית דינמית, המאפשרת קיצור זמן ומידע במודלים שפה גדולים. השיטה משתמשת באוטואנקודר סמנטי דינמי ומקצרת את הקלט והפלט.
תקציר מקורי באנגליתarXiv:2609.15338v1 Announce Type: cross Abstract: Large Language Models (LLMs) primarily perform inference at the token level, resulting in substantial memory overhead and compromised computational efficiency. In this paper, we propose a Dynamic Semantic Extraction and Inference (DSEI) framework, which achieves segment-level inference within the latent space through a two-stage training strategy. First, we construct a Dynamic Semantic Autoencoder (DSAE) via self-supervised learning. DSAE dynamically extracts segment-level semantics and compresses them into compact latent representations via adaptive semantic weighting and gated fusion. Subsequently, we integrate the DSAE into the LLM architecture and train the model to infer over dense latent space. DSEI substantially reduces both input an
קרא במקור המקורי