Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

TOPIC · ENTITY #5492

fine-tuning (~3600 examples)

Related event timeline, sources and context from the news index.

EVENT TIMELINE

1

RESEARCH · 1 SOURCE · arXiv cs.AI

Round-trip study finds lossy, asymmetric serialization of tree-structured expressions in language models

The arXiv paper (arXiv:2609.21509v1) proposes a round-trip protocol where one model generates a word problem from a procedurally created arithmetic expression and another model extracts the original expression; symbolic equivalence is used as an exact oracle. Evaluating 16 models pairwise, the study finds the natural-language channel is lossy and asymmetric (generation vs extraction can shift accuracy by up to 60.4 points), most failures stem from generation and tree structure (operator count, depth, right-branching) drives difficulty, and roughly 3,600 targeted fine-tuning examples substantially improve open-weight models above an untrained Gemini-3.1-Pro baseline.

7.0