יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.AI ·

בקרת פרסונה מדורגת

Persona Dosing: Calibrated Activation Steering for Graded Trait Control
חוקרים פיתחו שיטה לבקרת פרסונה מדורגת עבור מודלים שפה. השיטה, PersonaDose, מאפשרת שליטה על ביטוי תכונות אישיות במודלים כמו Llama-3.1-8B, Qwen3-8B ו-Gemma-3-4B.
תקציר מקורי באנגליתarXiv:2609.36388v1 Announce Type: new Abstract: An activation-steering coefficient sets intervention strength, but requesting a particular degree of persona expression requires a behavioral scale. We study persona dosing: controlling a language model through a trait description and a requested mean intensity. PersonaDose specializes a shared, description-conditioned FLAS controller on persona responses, then calibrates its flow time against measured trait expression. Training responses are not paired with requested target intensities. Across Llama-3.1-8B, Qwen3-8B, and Gemma-3-4B, PersonaDose raises core-trait expression at the Persona Vectors coherence floor of 75 by 33.2, 18.3, and 17.8 points over contrastive activation addition. Calibration-selected settings retain an expression advant
קרא במקור המקורי