Verified Rewards Beat Exact Match
On three in-domain benchmarks, VSeek-VETL beats the base model it is post-trained from, the agentic baselines, and VSeek-EM, which sees only an exact-match reward. A 4B model trained this way overtakes GPT-5.2 on LongVideoBench and MLVU.