FRAUDSkill: Structured Frozen-Weight Skill Optimization for Audio Anti-Fraud Detection
The authors propose FRAUDSkill, a structured frozen-weight adaptation framework that optimizes an external layer of skill programs, route-specific policies, and decision rules without changing the underlying audio-language model. By combining structured output control with validation-guided multi-path inference, the framework ensures protocol-compliant predictions for service-scenario identification, fraud detection, and conditional fraud-type classification. Tested on the TeleAntiFraud benchmark, it significantly outperforms baseline models while drastically reducing invalid outputs, offering an adaptable solution for complex audio anti-fraud workflows.
FRAUDSkill achieves 73.50% Macro-F1 on the TeleAntiFraud benchmark.
Outperforms the shared frozen-model baseline by 31.96%.
Reduces invalid outputs to 1.94%.