metadata
tags:
- fp32
- vision
- qwen
- qwen3 vl
- image to text
base_model:
- Qwen/Qwen3-VL-4B-Instruct
license: apache-2.0
Qwen3-VL-4B-Instruct-Uncensored
This model has been finetuned with image and text pairs at 1024px and shows high affinity with limited hallucination on NSFW task.
Full FP32 Training (AdamW NO 8bit Optimizers)
This model has limited video caption ability.