Search

直接匹配不多,已補上最新動態。

Tag: #self-critique3 results

Adversarial Self-Critique for Safer AI Underwriting

Adversarial Self-Critique for Safer AI Underwriting

New agentic AI system for commercial insurance underwriting uses adversarial self-critique where a critic agent challenges primary decisions before human review. It reduces hallucinations from 11.3% to 3.8% and boosts accuracy from 92% to 96% on 500 expert cases. The human-in-the-loop design ensures oversight in regulated environments.

ArXiv AIResearchFeb 17#research#agentic-ai#self-critique
Ox Alpha:神秘模型公開亮相

Ox Alpha:神秘模型公開亮相

Ox Alpha 是一款透過 OpenRouter 與 OpenCode 提供的匿名模型,具備 100 萬 token 上下文、圖片與影片輸入、工具呼叫及免費使用等能力。早期測試顯示其推理與程式設計表現強勁,但開發者、參數量、訓練資料與正式基準排名仍未獲確認。