HEAD TO HEAD
Bofeng Huang: docto-decision and Jared Palmer: kev, compared on what they decide, where they run, what they cost and what their publishers report.
| Property | bofeng-huang/docto-decision | jared-palmer/kev |
|---|---|---|
| Summary | French medical decision models by Bofeng Huang: pass a patient message or clinical note, a question and candidate answers, and get one probability per answer from a single forward pass. Choice, score and noul. A research model, not a medical device. | An open family of System One models. A LoRA adapter plus a pointer head on a frozen Qwen base returns a distribution per typed question in one forward pass, serves TypeSafe's /v1/systemone contract, and ships a fitted temperature with every checkpoint. |
| Decides | choice, score, noul | choice, score, noul, classify, route |
| Architecture | docto-decision | kev |
| Fine-tuned from | qwen/qwen3.5-4b | qwen/qwen3.5-4b-base |
| License | apache-2.0 | apache-2.0 |
| Availability | Open weights | Open weights |
| Hosted by | — | — |
| Input price | — | — |
| Decision accuracy | 88.9% | 83.8% |
| Calibration error | 0.079 | 0.042 |
| Valid action rate | — | — |
| Median latency | 47 ms | — |
| p95 latency | — | — |
| Evaluation suite | Docto Decision Bench fr v0.1 (12 tasks, mostly silver labels; the maker's own benchmark) | transfer-v4 (locked, out of domain) |
| Latest version | 0.1.0 | 2026.09 |
| Variants | LICENSE, adapters | — |
| Size of latest version | 8.8 GB | 152.2 MB |
| Files | 13 | 11 |
| Downloads | 0 | 1 |
| Stars | 0 | 0 |
| Tags | system-one, french, medical, qwen, lora, 4.66b | system-one, qwen, lora, pointer-head, 4b |
| Updated | Oct 7, 2026 | Oct 7, 2026 |
Figures are from each model’s manifest; accuracy and latency are what the publishers report, on their own suites and hardware. Add a third model.
docto-decision is from Bofeng Huang and kev from Jared Palmer. Both have open weights you can download and run. Both answer choice, score and noul questions. Only kev answers classify and route. kev is the smaller model, at 4.0B parameters to 4.7B.
They report on different suites — docto-decision 88.9% on Docto Decision Bench fr v0.1 (12 tasks, mostly silver labels; the maker's own benchmark), kev 83.8% on transfer-v4 (locked, out of domain) — so the numbers do not rank them. Test both on your own labelled examples.
docto-decision: Free (open weights). kev: Free (open weights). Open weights cost nothing per call beyond your own hardware.
Yes, both: systemone pull bofeng-huang/docto-decision and systemone pull jared-palmer/kev download the weights.