כתבה
arXiv cs.CL ·
When a Kindergartener Solves Calculus: Measuring Capability Leakage in Role-Prompted Reasoning Models
תקציר מקורי באנגליתarXiv:2609.39846v1 Announce Type: new Abstract: We investigate the problem of role-capability leakage (RCL), in which a role-prompted reasoning model generates convincing in-role text while continuing to exhibit capabilities on benchmarks that exceed those implied by the assigned role. For example, when a model is prompted to assume the role of a kindergarten student, one might expect its performance on a mathematics benchmark to reflect kindergarten-level ability rather than expert-level proficiency in solving calculus problems. We introduce RoleCapBench, a curriculum-grounded benchmark for evaluating RCL across six educational roles and four assessment levels spanning elementary school through A-level, and use it to evaluate three open-weight reasoning models. We find that although the m
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית