guerilla7's picture
Add README with quantization metadata
869c26f verified
|
Raw
History Blame
798 Bytes
metadata
base_model: fdtn-ai/Foundation-Sec-8B-Instruct
tags:
  - quantized
  - nvfp4
  - tensorrt
  - foundation-sec-8b-instruct
  - cybersecurity

Foundation-Sec-8B-Instruct-NVFP4-quantized

This repository contains an NVFP4 quantized version of the fdtn-ai/Foundation-Sec-8B-Instruct model, optimized for NVIDIA Spark using TensorRT Model Optimizer.

Quantization Details

Loading

Refer to TensorRT-LLM or your deployment stack for loading NVFP4 artifacts.

License

(Inherit from base model if applicable, or specify your own)