Thanks, and a tip: Jeb decides from label logprobs, not chat

#1
by jbrashear - opened

Hey bartowski, thank you for quantizing Jeb. Your repo has far more downloads than mine, so this is where a lot of people are meeting the model, and I really appreciate it.

One thing that might save them some confusion. Jebadiah isn't a chat model. It answers a typed question about some JSON (pick one option, yes or no, or a score) from the probability of each option label at one position, with thinking off. If someone loads it in a chat window with the <think> template, the replies will look strange, and that's expected.

The easy way to use your quants the way it was trained:

pip install jebadiah-decide
ollama pull hf.co/bartowski/frontier-infra_jebadiah-9b-v2-GGUF:Q8_0
jeb serve --model hf.co/bartowski/frontier-infra_jebadiah-9b-v2-GGUF:Q8_0

jeb renders the exact prompt the model was trained on, reads the label probabilities and applies the per-type temperatures from temperatures.json in frontier-infra/jebadiah-9b-v2, so the probabilities come out calibrated. It works with llama-server too (--backend llama-server).

If you're up for it, a line on your card pointing at that would help people a lot. Either way, thanks again.

Jason

Sign up or log in to comment