יום שני, 5 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

מרכזים של מושגים קליניים במודלים גדולים של שפה

Clinical Concept Centers in LLMs
חוקרים גילו מרכזים של מושגים קליניים במרחב הלטנטי של מודלים גדולים של שפה. המרכזים האלה משמשים כמנגנונים פנימיים המניעים את התנהגות המודל. המחקר מראה כי המרכזים האלה יכולים לשמש לשיפור הביצועים הקליניים של המודלים.
תקציר מקורי באנגליתarXiv:2610.02829v1 Announce Type: cross Abstract: Large language models are increasingly used in clinical settings. However, research into the reliability and performance of these models has focused almost entirely on the language substrate, scoring what the model says. Mechanistic interpretability has found that the latent space carries a higher fidelity of representation than the text: internal representations not only encode substantially more than the output verbalizes, but the stated reasoning also systematically omits features that causally drive the answer. An evaluation of model behavior in terms of mechanistic interpretability has not been explored in clinical decision support. In this work, we extend behavioral evaluation into the latent space and ask whether clinical concepts ex
קרא במקור המקורי