Abstract
System One models such as Jev offer an efficient alternative to generative language models for tasks that require decisions rather than open-ended responses. However, existing Jev models exhibit limited Chinese-language decision accuracy, restricting their utility in both general and specialized settings. In this paper, we introduce Chinese-Jev, a System One model that addresses this gap through a unified data processing and training pipeline. Our data processing protocol converts heterogeneous Chinese-language annotations into probability targets over candidate options, enabling a shared training formulation across domains and question formats. To enable efficient inference, Chinese-Jev adopts a lightweight encoder-only backbone for text encoding and learns to score candidate answers through decision-oriented training. To address the misalignment between the pre-training distribution and downstream Chinese-language scenarios, we first train the model on a general-purpose corpus of 10 million examples, then fine-tune it separately for the medical, legal, and financial domains. To evaluate decision accuracy and calibration in both general and domain-specific Chinese-language settings, we introduce Chinese-Jev Bench (CJ-Bench). After first-stage pre-training, Chinese-Jev exceeds the accuracy of the closed-source Jev model by 1.24% on general-domain tasks while achieving a 20.3x speedup. Subsequent domain-specific fine-tuning yields a 4.0% accuracy improvement over Jev in medicine and achieves 92% of Jev's average accuracy across specialized domains, with a 17x speedup and an average latency of only 15 ms per example. We further demonstrate on-device deployment of an INT8-quantized model on mobile devices, achieving an inference latency of approximately 1.0 second per decision. The project is available at https://gulucaptain.github.io/Chinese-Jev/.
Community
System One models such as Jev offer an efficient alternative to generative language models for tasks that require decisions rather than open-ended responses. In this paper, we introduce Chinese-Jev, a System One model that addresses the Chinese-QA domain gap for Jev through a unified data processing and training pipeline. We will open source all the training/inference code, model and checkpoints, and training datasets. The project page is at: https://gulucaptain.github.io/Chinese-Jev/.
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- C-HAT-Bench: Benchmarking Chinese AI-Text Detection Beyond Fully Generated Text (2026)
- Reference-Grounded Data Curation for Instruction-Following Thai-English Machine Translation (2026)
- Evaluating and Benchmarking the System One Model Jev (2026)
- Polish ModernBERT: The Long and Short of Polish Language Understanding (2026)
- BEAR-Bench: A Bilingual Enterprise and Academic Reasoning Benchmark for Multimodal Models (2026)
- EuroAlpaca: Task-Preserving Localisation of Instruction Data for European Languages (2026)
- IronLLM: Forging Compact Edge-Native Language Models for Real-Time Embodied Intelligence (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2609.36965 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper