Fix junk output: skip_special_tokens=False, lower temp, remove repetition_penalty, preserve </think> tag 5a509aa verified Laksh99 commited on Mar 2
Fix streaming: stop silent during thinking, yield answer after </think> 1b1b699 verified Laksh99 commited on Mar 2
Remove type="messages" from ChatInterface - use compatible Gradio API 7fb1f9c verified Laksh99 commited on Mar 2
Rewrite with gr.ChatInterface to fix data incompatibility error 07f896a verified Laksh99 commited on Mar 2
Fix: use Gradio messages format (type="messages") to resolve data incompatibility error c6a39ad verified Laksh99 commited on Mar 2
Lazy model loading - load model on first GPU request instead of startup 2e76d0b verified Laksh99 commited on Mar 2
Fix indentation - clean app.py for Qwen3.5 AWQ ZeroGPU chat 234b259 verified Laksh99 commited on Mar 2