Instructions to use DARJYO/persadian_14B-GRPO with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use DARJYO/persadian_14B-GRPO with Transformers:
# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("DARJYO/persadian_14B-GRPO", device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 767 Bytes
d741ce3 0d7a541 d741ce3 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 | ---
license: apache-2.0
language:
- en
metrics:
- accuracy
base_model:
- unsloth/phi-4
library_name: transformers
tags:
- text-generation-inference
- reinforcement-learning
- trl
- vllm
- datasets
---
# Model
- **Developed by:** DARJYO
- **Base Type:** Fine-tuned language model
- **Finetuned model :** persadian_14B-GRPO
- **Base Architecture:** Transformer-based/Phi-4
This model is fine-tuned on datasets for tasks with [Unsloth](https://github.com/unslothai/unsloth) and Huggingface's TRL library.
It is based on the `unsloth/Phi-4` model and uses reinforcement learning for improved performance.
[<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>](https://github.com/unslothai/unsloth)
|