YAML Metadata Error:Invalid Eval Result format in .eval_results/health-check-sakthai-coder-1.5b-2026-07-30-4.yaml

Check out the documentation for more information.

Show details
✖ Invalid input: expected array, received object
sakthai-coder-1.5b / .eval_results /health-check-sakthai-coder-1.5b-2026-07-30-4.yaml
Nanthasit's picture
Upload .eval_results/health-check-sakthai-coder-1.5b-2026-07-30-4.yaml with huggingface_hub
c0f7d3c verified
Raw History Blame
4.81 kB
target_model:
id: Nanthasit/sakthai-coder-1.5b
slug: sakthai-coder-1.5b
eval_metadata:
schema: llm_cron
generated_at: '2026-07-30T22:52:00Z'
generated_by: sakthai-agent-cron
model_type: text-generation
popularity:
downloads: 93
likes: 0
max_sibling_downloads: 1599
dl_score: 6
likes_score: 0
score: 4
weight: 0.20
momentum:
age_days: 6.52
velocity: 14.27
max_sibling_velocity: 63.95
velocity_rank: 9
ratio_score: 22
rank_score: 20
score: 21
weight: 0.20
benchmarks:
model_index_present: true
metric_count: 4
all_verified: false
score: 60
weight: 0.25
entries:
- dataset: HumanEval
dataset_type: openai_humaneval
metric: 'pass@1 (base model reference)'
metric_type: 'pass@1'
value: 74.40
verified: false
- dataset: MBPP
dataset_type: mbpp
metric: 'pass@1 (base model reference)'
metric_type: 'pass@1'
value: 71.20
verified: false
- dataset: MultiPL-E (Python)
dataset_type: multipl_e
metric: 'pass@1 (base model reference)'
metric_type: 'pass@1'
value: 65.30
verified: false
- dataset: SakThai Coding Suite (internal)
dataset_type: custom
metric: 'pass@1 (fine-tuned model, internal)'
metric_type: 'pass@1'
value: 100
verified: false
card_quality:
license: apache-2.0
base_model: Qwen/Qwen2.5-Coder-1.5B-Instruct
tags_count: 10
datasets_count: 3
readme_bytes: 15490
score: 100
weight: 0.20
repo_summary:
total_siblings: 1558
total_storage_bytes: 1218622744
total_gb: 1.13
weight_files:
- path: qwen2.5-coder-1.5b-instruct-q4_k_m.gguf
size_bytes: 1117320768
has_weights: true
config_exists: false
dev_artifact_dirs:
- .hypothesis
- .pytest_cache
- .ruff_cache
- .venv
dev_artifact_count: 4
hygiene:
score: 40
weight: 0.15
deductions:
dev_artifacts: 60
storage_bloat: 0
sibling_comparison:
- id: Nanthasit/sakthai-context-1.5b-merged
downloads: 1599
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-context-0.5b-merged
downloads: 1370
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-context-7b-merged
downloads: 744
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-context-7b-128k
downloads: 506
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-context-7b-tools
downloads: 399
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-embedding-multilingual
downloads: 362
likes: 0
pipeline_tag: sentence-similarity
- id: Nanthasit/sakthai-context-1.5b-tools
downloads: 349
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-vision-7b
downloads: 186
likes: 1
pipeline_tag: image-to-text
- id: Nanthasit/sakthai-tts-model
downloads: 150
likes: 0
pipeline_tag: text-to-speech
- id: Nanthasit/sakthai-context-0.5b-tools
downloads: 94
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-embedding
downloads: 23
likes: 0
pipeline_tag: sentence-similarity
- id: Nanthasit/sakthai-context-1.5b-tools-v2
downloads: 0
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-context-1.5b-merged-v2
downloads: 0
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-plus-1.5b
downloads: 0
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-plus-1.5b-lora
downloads: 0
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-plus-1.5b-coder
downloads: 0
likes: 0
pipeline_tag: text-generation
- id: Nanthasit/sakthai-coder-browser-lora
downloads: 0
likes: 0
pipeline_tag: null
- id: Nanthasit/sakthai-coder-browser
downloads: 0
likes: 0
pipeline_tag: text-generation
health_score:
overall: 46
popularity: 4
momentum: 21
benchmarks: 60
card_quality: 100
hygiene: 40
weighting: 'pop=20% mom=20% bench=25% card=20% hyg=15%'
delta:
score_change: 0
downloads_change: 0
likes_change: 0
from_prior: health-check-sakthai-coder-1.5b-2026-07-30-3.yaml
assessment:
status: stable
notes: No change in downloads/likes since previous check. Score steady at 46/100. GGUF-only coder model (1.04 GB GGUF, Q4_K_M, qwen2 arch, 32K context). Card complete with license/base_model/10 tags/3 datasets. 4 unverified benchmarks listed (HumanEval 74.4%, MBPP 71.2%, MultiPL-E 65.3%, internal suite 100%). 4 dev artifact directories penalizing hygiene. Last modified 2026-07-30T22:46:41Z.
recommendation: 'Clean .venv, .hypothesis, .pytest_cache, .ruff_cache to boost hygiene by 15 points (potential score: 55/100). Add verified benchmark results to model-index. Consider 1.12 GB LFS limit.'