Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

acv1229
/
rl-clarify-orig-prompt-d1-1p5

Reinforcement Learning
Safetensors
ppo
lora
code-generation
clarification
Model card Files Files and versions
xet
Community
rl-clarify-orig-prompt-d1-1p5
5.15 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 18 commits
acv1229's picture
acv1229
Upload iter_0079
4633df7 verified 5 months ago
  • iter_0004
    Upload iter_0004 5 months ago
  • iter_0009
    Upload iter_0009 5 months ago
  • iter_0014
    Upload iter_0014 5 months ago
  • iter_0019
    Upload iter_0019 5 months ago
  • iter_0024
    Upload iter_0024 5 months ago
  • iter_0029
    Upload iter_0029 5 months ago
  • iter_0034
    Upload iter_0034 5 months ago
  • iter_0039
    Upload iter_0039 5 months ago
  • iter_0044
    Upload iter_0044 5 months ago
  • iter_0049
    Upload iter_0049 5 months ago
  • iter_0054
    Upload iter_0054 5 months ago
  • iter_0059
    Upload iter_0059 5 months ago
  • iter_0064
    Upload iter_0064 5 months ago
  • iter_0069
    Upload iter_0069 5 months ago
  • iter_0074
    Upload iter_0074 5 months ago
  • iter_0079
    Upload iter_0079 5 months ago
  • .gitattributes
    2.5 kB
    Upload iter_0079 5 months ago
  • README.md
    1.35 kB
    Add README 5 months ago