Abstract
Typed decision models answer a declared question without generating text: a decision head returns a probability for each of the declared options in a single forward pass. A single pass is fast, intuitive System 1 thinking. We study what lies between one pass and generated reasoning: looping, in which the same layers are recursively applied several times before one typed readout. Each loop lets the model revise its hidden state before it commits to an answer, without generating a token; we call this System 1.5 thinking. We propose SanSi, which turns a pre-trained looped language model into a typed decision model. The option probabilities are read after every loop, and every loop is trained with a proper scoring rule, so that one model serves every budget from one loop to eight in a single run. On 10,027 test decisions from 59 sources, SanSi reaches 72.0% accuracy: 13.5 points above a non-looped model of the same shape trained with the same recipe, 5.3 points above a newer non-looped model of its size, and 1.8 points below one with three times the parameters. On two depth-controlled tasks, loops extend the solvable depth beyond the depths seen in training, where the larger single-pass model fails. Used as the judge for policy optimization with reinforcement learning, without gold answers, SanSi raises the generator's F1 by 7.7 points.
Community
TL;DR: Our work is a looped typed decision model that returns a probability for each declared option and can be read after any of its eight loops, so one model serves every compute budget without generating reasoning tokens. With the same data and recipe, the 1.4B SanSi comes within 1.8 points of Qwen3.5-4B (three times its parameters) and, unlike it, solves reasoning chains longer than any seen in training.
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- Benchmarking System One decision models against trained classifiers and language models for automated decision gates (2026)
- ufakzeka-karar: An Open Turkish Typed-Decision Model with Order-Invariant Option Scoring (2026)
- RecurTrace: Adaptive Latent Reasoning with Loop-Time Memory (2026)
- Bongard: Training Machine Intuition (2026)
- AnyJev Technical Report (2026)
- From Retrieval to Typed Decisions: Calibrated System One Models from Biomedical Sentence Encoders (2026)
- Labels Override Definitions in Jev-Style Typed Decision Models (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2610.07730 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 2
minnesotanlp/SanSi
Datasets citing this paper 0
No dataset linking this paper