TW3Cast reaches third on GIFT‑Eval benchmark
TW3Cast is a time‑series forecasting system that achieved the third position out of 130 entries on the GIFT‑Eval benchmark as of 2026‑09‑14. The system does not use agents or language models; instead it relies on a frozen routing table built once on the training split. The table selects among four modes for each of 97 dataset‑frequency‑horizon configurations: a specialist (a LoRA or full…
Key points
- TW3Cast ranks third on GIFT‑Eval with a mean MASE rank of 19.4
- It uses a frozen routing table selecting among specialists, blends, and tournaments
- The system relies on lightly fine‑tuned Chronos‑2, TiRex, and Toto models
The best base model alone scores a mean MASE rank of 33.8, the tournament alone 38.0, and the full router 19.4. All routing tables, expert indices, and score files are released, allowing others to regenerate the leaderboard numbers with a single script.
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us. Published by Martin K., who runs Digest AI and handles corrections.
More in Research
All →- Researchers release COILD corpus for Indian language machine translation · 1 src
- Research finds script knowledge in LLMs emerges only in final layers · 1 src
- Researchers test how language models handle numerical formats in word problems · 2 src
- PTC-Bias improves speech LLM accuracy with phoneme-level bias correction · 1 src
- Researchers benchmark how LLMs handle political character attacks · 1 src
Comments
via GitHub Discussions