تدقيق قائم على الأدلة لوكلاء التداول المعياريين
الملخص
تقترح هذه الورقة تدقيقا على مستوى الادعاء للوكلاء المعياريين الذين يخططون وينفذون ويتحققون وينقحون، انطلاقا من محدودية الاعتماد على درجة إجمالية واحدة للمهمة. ويسجل التدقيق الأدلة لكل استنتاج، ويعين حكما من أربعة: مدعوم أو غير مدعوم أو غير محسوم أو لم يُقيّم، ويحدد نطاق انطباق الاستنتاج. والغرض منه توضيح ما يثبته التقييم بشأن الوكيل ومكوناته.
تسهم ثلاث طرق في جمع الأدلة: تقيس سياسات أوراكل المثالية التحسن الممكن بلوغه لمجموعة أفعال محددة؛ ويساعد الاستبدال بمكون مثالي واحدا تلو الآخر على تحديد القيمة المفقودة، مع مراعاة إمكان حجبها بسبب مراحل لاحقة؛ كما يختبر فحص مستقل ما إذا كانت درجة أداة التحقق تدعم فعلا الحد المنسوب إليها. وفي سوق اصطناعي ذي أنظمة خفية، يجد التدقيق أن القيمة المقاسة لمعلومات النظام المثالية تتغير مع مجموعة الأفعال، وأن مولد السيناريوهات يفقد كثيرا من إشارة النظام، وأنه يمكن تجاوز أداة تحقق وقت التشغيل من دون تغيير واضح في النتائج. وتتعلق هذه النتائج بوكيل وبيئة واحدين؛ أما المساهمة الأوسع فهي البروتوكول.
الأفكار الرئيسية
- قد لا تحدد درجات المهمة الإجمالية وحدها المكون الذي تسبب في نتيجة أو ما الذي تصادق عليه أداة التحقق.
- يرفق التدقيق بكل ادعاء دليلا وحكما من أربعة أنواع وحدود نطاقه.
- تقدر سياسات أوراكل المثالية المكاسب الممكنة مقارنة بمجموعة أفعال محددة صراحة.
- قد يحدد استبدال المكونات موضع فقدان القيمة، لكن الآثار اللاحقة قد تترك النتائج غير محسومة.
- تقتصر نتائج السوق الاصطناعي على الوكيل والبيئة المدروسين.
الوسوم
النص الكامل
# Verify Claims, Not Scores: Evidence-Based Verification of Modular Agents # Verify Claims, Not Scores: Evidence-Based Verification of Modular Agents When developers change one component of an agent, such as its controller, a learned model or its verifier, they usually judge the change by an aggregate task score. That score cannot tell whether improvement was attainable, which component lost value, or what the agent's own checks certify. We introduce a claim-specific verification audit for modular agents that plan, act, check and refine. Instead of scoring the agent, the audit scores the evidence: each conclusion is recorded with the evidence behind it, one of four verdicts (supported, unsupported, unresolved or not evaluated) and the boundary within which it holds. Three tools supply that evidence. Oracle policies measure attainable improvement under an explicitly stated action set, so that a low value can be traced to the evaluation rather than to the environment. Replacing one component at a time with a perfect counterpart locates lost value, with null results read as unresolved whenever a downstream component could mask them. A separate test asks whether the verifier's score identifies the quantity it is read as bounding. Applied to a constrained portfolio-allocation agent in a synthetic market with known hidden regimes, the audit shows that the value of perfect regime information depends on the action set used to measure it, that the scenario generator discards most of the regime signal while better local fidelity does not improve decisions, and that the runtime verifier can be bypassed with no visible change in outcomes. The contribution is the protocol and the evidential distinctions it enforces; the empirical findings are specific to the agent and environment studied.
يُعرض النص كاملًا مع نسبه إلى مصدره وفقًا لترخيصه. الترخيص: abstract CC0
أعدّ وكيل الأبحاث في Stratmill هذا الملخص استنادًا إلى المصدر الأصلي؛ وهو ليس نسخة منه.