dolev31 commited on
Commit
45b629e
·
verified ·
1 Parent(s): b6e8077

Card: Results and Limitations sections, pointing to the adapter card's results

Browse files
Files changed (1) hide show
  1. README.md +16 -0
README.md CHANGED
@@ -50,6 +50,12 @@ This is training seed 1, the adapter at the root of the adapter repository. The
50
  and the weights are stored in bfloat16. On the adapter card's two-turn example, greedy decoding with
51
  this model returns the adapter's output character for character.
52
 
 
 
 
 
 
 
53
  ## How to use it
54
 
55
  The questioner reads the prompt template it was trained on, in [`prompts/`](prompts/), and replies
@@ -102,6 +108,16 @@ vllm serve dolev31/ProactiveInquirer-Qwen3-8B-Merged
102
  # "chat_template_kwargs": {"enable_thinking": false}, "temperature": 0}
103
  ```
104
 
 
 
 
 
 
 
 
 
 
 
105
  ## Citation
106
 
107
  ```bibtex
 
50
  and the weights are stored in bfloat16. On the adapter card's two-turn example, greedy decoding with
51
  this model returns the adapter's output character for character.
52
 
53
+ ## Results
54
+
55
+ The results are the trained questioner's, as the paper reports them: see the
56
+ [adapter card's Results](https://huggingface.co/dolev31/ProactiveInquirer-Qwen3-8B#results). This merged model reproduces the adapter's output on that card's example, as the paragraph
57
+ above says.
58
+
59
  ## How to use it
60
 
61
  The questioner reads the prompt template it was trained on, in [`prompts/`](prompts/), and replies
 
108
  # "chat_template_kwargs": {"enable_thinking": false}, "temperature": 0}
109
  ```
110
 
111
+ ## Limitations
112
+
113
+ - The questioner's own limitations, from the paper: it has learned what to ask more readily than when to
114
+ stop, the extra evidence it finds does not yet translate into better final answers, and its user-facing
115
+ results come from a simulated customer, not from real people.
116
+ - It is a component inside an agent, meant to be called with its prompt template. It is not a chat
117
+ assistant, and it was trained and evaluated in English.
118
+ - This is one training seed (seed 1), merged in float32 and stored in bfloat16. Its equality with the
119
+ adapter was checked on the card's example, not on a benchmark.
120
+
121
  ## Citation
122
 
123
  ```bibtex