DigestAI news desk

Cut through the AI noise.

Research

Researchers introduce GoldiMask to improve diffusion language model fine-tuning

A new method called GoldiMask optimizes how discrete diffusion language models are fine-tuned by strategically selecting which tokens to reveal as context. Unlike random masking, GoldiMask uses a submodular objective to balance the benefit of visible tokens against their value as prediction targets. It then adjusts the weights of remaining targets based on their learning potential and context…

1 source primary source

Key points

  • GoldiMask selects tokens to reveal by maximizing a submodular objective balancing context benefit and target value
  • Method improves accuracy in reasoning and code generation across three backbones and datasets
  • Reduces decoding iterations on GSM8K and MATH-500 without sacrificing accuracy at higher thresholds

The approach, detailed in a paper on arXiv, shows improved accuracy across three model backbones and datasets, particularly in reasoning and code generation tasks. Ablation studies confirm both context selection and target weighting contribute to these gains. GoldiMask also reduces decoding iterations on GSM8K and MATH-500 while keeping accuracy stable at higher confidence thresholds.

Read the original at arXiv cs.AI · by Loay Mualem, Llu\'is Pastor-P\'erez, Vinh Tong, Andrei Manolache, Tanja Bien, Steffen Staab, Mathias Niepert primary sourceOpen source ↗
Topics · follow one to build your own front page
GoldiMask

The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us. Published by Martin K., who runs Digest AI and handles corrections.

Comments

via GitHub Discussions

More in Research

All →

Related stories