Instructions to use mertkayacs/Wahler-4B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use mertkayacs/Wahler-4B with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("text-classification", model="mertkayacs/Wahler-4B")# pip install -U transformers accelerate # Load model directly from transformers import AutoProcessor, AutoModelForMultimodalLM processor = AutoProcessor.from_pretrained("mertkayacs/Wahler-4B") model = AutoModelForMultimodalLM.from_pretrained("mertkayacs/Wahler-4B", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Wähler-4B: small German language model for local decisions
Wähler-4B is one of the best open German decision models at 4B parameters: a small German language model for text classification, routing and decisions with probabilities. Wähler-4B: 92.0% accuracy; Kev-4B: 81.1% on held-out German decisions. The test set comes from the same data pipeline as the training data.
Try a decision in the browser, or run the GGUF build locally below.
Run locally
The Q4_K_M build used 3.03 GB RAM at a 4k context. The JevAlt server uses llama.cpp to read decision probabilities and load the matching calibration.
pip install "jevalt[serve,gguf] @ git+https://github.com/mertkayacs/jevalt" && jevalt serve --model mertkayacs/Wahler-4B-GGUF --file Wahler-4B-Q4_K_M.gguf
Send a situation, question and options to the local server:
import requests
request = {
"state": "I was charged twice for my subscription. Please refund the duplicate.",
"questions": {
"team": {
"type": "choice",
"instructions": "Which team should handle this ticket?",
"criteria": {"billing": "payments, refunds", "technical": "bugs, outages"},
}
},
"reasoning": "off",
"abstain": False,
}
response = requests.post("http://127.0.0.1:8000/v1/systemone", json=request)
response.raise_for_status()
print(response.json())
Write the situation, question and options in German. The API also supports yes/no questions, ordered scores and an unknown option with abstain: true. The card widget shows a recorded answer from the full-precision model.
Use the full-precision weights with Transformers
This repository holds the full-precision weights. The JevAlt Transformers backend preserves the decision-token format and calibration:
pip install "jevalt[hf] @ git+https://github.com/mertkayacs/jevalt" && jevalt serve --backend hf --model mertkayacs/Wahler-4B
See the local run guide for memory and client settings.
Deutsche Zusammenfassung
Wähler-4B ist ein kleines deutsches Sprachmodell für Textklassifikation und Entscheidungen zwischen vorgegebenen Optionen. Bei den zurückgehaltenen deutschen Entscheidungsaufgaben erreicht Wähler-4B 92,0 % Genauigkeit, Kev-4B 81,1 %. Trainings- und Testdaten stammen aus derselben Datenpipeline. Die Q4_K_M-Version läuft lokal mit etwa 3 GB RAM.
Results and limits
On the same held-out German decisions, Wähler-4B's probability error (Brier score, lower is better) is 0.119, against 0.298 for Kev-4B. A separate live test on 4 October 2026 scored 122/130 for Wähler-4B. Comparison files and live requests, including misses record both tests.
The models were fine-tuned with LoRA from Intern-Decision-4B, using JevAlt training data. German is this model's focus.
- Refit calibration on your own data before using confidence thresholds. See JevOss.
- The model accepts text, with an 8k context. Long irrelevant text and date arithmetic remain weak spots. Leave reasoning off for German: it did not improve accuracy on the reported date, number and policy test.
- Planted instructions still change some answers. Keep authorization checks outside the model and review consequential decisions.
Downloads and project
GGUF files and measured agreement | Code and training | Wähler-4B project page | Training and evaluation details.
Apache-2.0. Cite the JevAlt repository for this release.
An Eschatia Labs project. Mert Kaya.
- Downloads last month
- 161