כתבה
arXiv cs.AI ·
רישום ואימות של רצפי פיד: קידוד תמונה-לשון עם רשתות רישום מחודשות
Grounded and Faithful P&ID Reasoning: Constraining Vision-Language Models with Recovered Evidence Graphs
מודלי תמונה-לשון מתארים רצפי פיד באופן נכון עם רשתות רישום מחודשות. ניתן לשאול את המודלים לשאלות על רצפי פיד ולקבל תשובות נכונות. זה חשוב לביטחון ולבטיחות בתעשיית התעשיות הכימיות.
תקציר מקורי באנגליתarXiv:2609.05880v1 Announce Type: cross Abstract: Piping and Instrumentation Diagrams (P&IDs) are the authoritative maps of process plants: isolation, maintenance, and HAZOP decisions depend on what connects to what. Vision-language models describe these sheets fluently, yet they often invent or miss process connections---and an invented or missed link can reverse an isolation or reachability call, so a plant decision cannot trust a fluent answer that was never checked against the linework. We instead recover an explicit graph of the drawing---its symbols, the process connections between them, and the tags that name them---and then require the model to answer only by querying that graph through seven read-only operators, so a topology claim is returned only when it cites the query results
קרא במקור המקורי
arxiv.org
פתח כתבה מקורית