Gnonymous commited on
Commit
bf760fd
·
verified ·
1 Parent(s): 2844448

Link the released ALFWorld model repos

Browse files
Files changed (1) hide show
  1. README.md +8 -7
README.md CHANGED
@@ -12,19 +12,20 @@ tags:
12
 
13
  # EVOKE: Eliciting World Knowledge in Agents for Transferable Decision-Making
14
 
15
- > **Coming soon.** Trained models will be released here. Follow [github.com/Gnonymous/EVOKE](https://github.com/Gnonymous/EVOKE) for updates.
16
 
17
  **[🌐 Project Page](https://gnonymous.github.io/EVOKE)** · **[💻 Code](https://github.com/Gnonymous/EVOKE)** · **[📑 Paper (arXiv:2609.38334)](https://arxiv.org/abs/2609.38334)**
18
 
19
  **EVOKE** is a post-training method that elicits the world knowledge already inside pretrained LLM agents, so that they decide by the consequences of their actions rather than by contextual habits. It holds the state fixed, swaps in alternative goals, and trains the policy to rank the same candidate actions under each goal, with no world-model module and no inference-time planning.
20
 
21
- ## Planned release
22
 
23
- | Backbone | Benchmarks |
24
- | --- | --- |
25
- | Qwen2.5-3B-Instruct | ALFWorld, WebShop, search-based QA |
26
- | Qwen2.5-7B-Instruct | ALFWorld, WebShop, search-based QA |
27
- | Qwen3-1.7B | ALFWorld, WebShop, search-based QA |
 
28
 
29
  ## Citation
30
 
 
12
 
13
  # EVOKE: Eliciting World Knowledge in Agents for Transferable Decision-Making
14
 
15
+ > **ALFWorld models are released.** See the repositories below; evaluation code is at [github.com/Gnonymous/EVOKE](https://github.com/Gnonymous/EVOKE).
16
 
17
  **[🌐 Project Page](https://gnonymous.github.io/EVOKE)** · **[💻 Code](https://github.com/Gnonymous/EVOKE)** · **[📑 Paper (arXiv:2609.38334)](https://arxiv.org/abs/2609.38334)**
18
 
19
  **EVOKE** is a post-training method that elicits the world knowledge already inside pretrained LLM agents, so that they decide by the consequences of their actions rather than by contextual habits. It holds the state fixed, swaps in alternative goals, and trains the policy to rank the same candidate actions under each goal, with no world-model module and no inference-time planning.
20
 
21
+ ## Models
22
 
23
+ | Model | Base model | Benchmark |
24
+ | --- | --- | --- |
25
+ | [Gnonymous/EVOKE-ALFWorld-7B](https://huggingface.co/Gnonymous/EVOKE-ALFWorld-7B) | Qwen2.5-7B-Instruct | ALFWorld |
26
+ | [Gnonymous/EVOKE-ALFWorld-3B](https://huggingface.co/Gnonymous/EVOKE-ALFWorld-3B) | Qwen2.5-3B-Instruct | ALFWorld |
27
+
28
+ Both repositories contain full models, each with its own license. Evaluation code: [github.com/Gnonymous/EVOKE](https://github.com/Gnonymous/EVOKE).
29
 
30
  ## Citation
31