translategemma-27b-it-NVFP4A16 / generation_config.json
henry1477's picture
Add NVFP4A16 text-only quantization of google/translategemma-27b-it
7992414 verified
Raw
History Blame Contribute Delete
156 Bytes
{
"cache_implementation": "hybrid",
"do_sample": true,
"eos_token_id": [1, 106],
"top_k": 64,
"top_p": 0.95,
"transformers_version": "4.57.3"
}