Commit History

fix: normalize explicit TSQA choice labels
9dbdae6

fengxr93 commited on

Track remaining evaluation bundles
cb894ce

fengxr93 commited on

Align caption and TSQA evaluation pipelines
592a672

fengxr93 commited on

ChatTS OOS runner: drop English datasets (cap_en, qa_en), keep zh caption + QA; remove redundant run_chatts_qazh.py
8852ba0

fengxr93 commited on

GPT inference: align TS serialization to text-LM peers + streaming/resumable
cd8bbb9

fengxr93 Claude Opus 4.8 commited on

README: document the caption GLM-extraction step (metric key + statistical-property explanation)
3b8f190

fengxr93 Claude Opus 4.8 commited on

Caption-scoring pipeline fixes + with-definitions GLM extractor
60d6a4e

fengxr93 Claude Opus 4.8 commited on

Add repository README
c1b66a7

fengxr93 commited on

Add slimmed 4level_grpo_min2 + joint_grpo_min2 checkpoints; peer models -> placeholder
11423f4

fengxr93 commited on

TS-Align benchmark reproduction bundles + canonical eval data + dataset sources
c3efe57

fengxr93 Claude Opus 4.8 commited on