Hanks1234 commited on
Commit
02355ff
·
verified ·
1 Parent(s): 65582b4

Update model card

Browse files
Files changed (1) hide show
  1. README.md +28 -1
README.md CHANGED
@@ -1,3 +1,30 @@
1
  ---
2
- {}
 
 
 
 
 
 
 
3
  ---
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
  ---
2
+ library_name: stable-baselines3
3
+ tags:
4
+ - reinforcement-learning
5
+ - battleship
6
+ - ppo
7
+ - maskable-ppo
8
+ - sb3-contrib
9
+ - custom-environment
10
  ---
11
+
12
+ # Battleship PPO Agent — Hanks1234/battleship-ppo-phase3
13
+
14
+ A MaskablePPO agent trained on a 10x20 Battleship board with custom
15
+ T-shaped and Z-shaped ships using [sb3-contrib](https://sb3-contrib.readthedocs.io/).
16
+
17
+ ## Environment
18
+
19
+ - **Board**: 10 columns x 20 rows
20
+ - **Ships**: 10 ships including T-shaped Battleships and Z-shaped Carriers
21
+ - **Observation**: 5-channel binary image (5, 20, 10)
22
+ - **Action**: Discrete(200) with action masking (no repeat shots)
23
+
24
+ ## Usage
25
+
26
+ ```python
27
+ from training.hub import load_model_from_hub
28
+
29
+ model = load_model_from_hub("Hanks1234/battleship-ppo-phase3")
30
+ ```