Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

acv1229
/
rl-clarify-orig-prompt-d1-1

Reinforcement Learning
Safetensors
ppo
lora
code-generation
clarification
Model card Files Files and versions
xet
Community
rl-clarify-orig-prompt-d1-1
2.58 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 11 commits
acv1229's picture
acv1229
Upload iter_0009
b933f43 verified 5 months ago
  • iter_0004
    Upload iter_0004 5 months ago
  • iter_0009
    Upload iter_0009 5 months ago
  • iter_0044
    Upload iter_0044 5 months ago
  • iter_0049
    Upload iter_0049 5 months ago
  • iter_0054
    Upload iter_0054 5 months ago
  • iter_0059
    Upload iter_0059 5 months ago
  • iter_0064
    Upload iter_0064 5 months ago
  • iter_0069
    Upload iter_0069 5 months ago
  • .gitattributes
    2.01 kB
    Upload iter_0009 5 months ago
  • README.md
    1.52 kB
    Add README 5 months ago