יום שישי, 9 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

PMTRM: יחידת תצוגה זיכרון-פסאודו-זמני ללמידת מדיניות גוף

PMTRM: Pseudo-Memory Temporal Re-encoding Module for Embodied Policy Learning
נוסחאות חדשות ללמידת מדיניות גוף: יחידת תצוגה זיכרון-פסאודו-זמני שמציגה היסטוריה מוגבלת של מצבים ופעולות שנעשו. זה עוזר להבחין בין פזות שונות במהלך תפעול של רובוט.
תקציר מקורי באנגליתarXiv:2610.11168v1 Announce Type: cross Abstract: Robotic manipulation often contains repeated motions whose local observations look similar at different phases. When these phases require different actions, a policy that relies mainly on the current observation may repeat completed motions or switch phases at the wrong time. To address this phase ambiguity, we present the Pseudo-Memory Temporal Re-encoding Module (PMTRM), a lightweight plug-in module with only 7.61M parameters that encodes a bounded history of executed states and actions into a latent sequence for existing policies. To help distinguish phases, a temporal heterogeneity objective penalizes positive similarity between distant positions in this sequence, while anchor and reconstruction losses preserve information needed for ac
קרא במקור המקורי