יום ראשון, 4 באוקטובר 2026 LIVE
AI־INFO

כתבה arXiv cs.LG ·

בדיקת תקן: דגמי System One נגד מודלי מחלקות מאומנים ומודלי שפה לשערי החלטה אוטומטיים

Benchmarking System One decision models against trained classifiers and language models for automated decision gates
בדיקת תקן של דגמי System One נגד מודלי מחלקות מאומנים ומודלי שפה לשערי החלטה אוטומטיים. המחקר משווה את תוצאות דגמי System One עם אלה של מודלי מחלקות מאומנים ומודלי שפה. התוצאות מציעות כללים תפעוליים לשערי החלטה אוטומטיים.
תקציר מקורי באנגליתarXiv:2610.00346v1 Announce Type: new Abstract: Software that hands branching decisions to a model needs a declared option and a probability it can threshold. Typed decision models, also called System One models, return such probabilities without generating text, while supervised classifiers and generative language models are the established alternatives. Under matched conditions, one harness sends eight decision-model checkpoints from six families, including the hosted model Jev, and two generative comparators the same semantic requests, and scores supervised and zero-shot classifiers on the same workflow, intent and social-science items. The ranking of the model classes depends on the conditions. With the task's own labels, small trained classifiers are the most accurate on intents and n
קרא במקור המקורי