Chinese-Jev: Bringing System One Model to Chinese-Language Tasks
The authors introduce Chinese-Jev, a System One model for Chinese-language decision tasks using a unified data processing and training pipeline. It adopts a lightweight encoder-only backbone with a pre-training phase on 10 million examples followed by domain-specific fine-tuning for medical, legal, and financial domains. Evaluated on the new CJ-Bench, Chinese-Jev exceeds the closed-source Jev model's general-domain accuracy by 1.24% while achieving a 20.3x speedup. In specialized domains, it yields a 4.0% accuracy improvement in medicine, a 17x speedup, an average latency of 15 ms per example, and on-device INT8 deployment with ~1.0 second latency.
After first-stage pre-training, Chinese-Jev exceeds the accuracy of the closed-source Jev model by 1.24% on general-domain tasks while achieving a 20.3x speedup.
Subsequent domain-specific fine-tuning yields a 4.0% accuracy improvement over Jev in medicine and achieves a 17x speedup with an average latency of only 15 ms per example.
An INT8-quantized model deployed on mobile devices achieves an inference latency of approximately 1.0 second per decision.