reasoning-before-translation-enhancing-legal-machine-translation-with-structured-reasoning-f07be993·1 events·first seen Aliases: Reasoning Before Translation: Enhancing Legal Machine Translation with Structured Reasoning
Researchers evaluate multiple training paradigms for legal machine translation, comparing supervised fine-tuning and reinforcement learning with verifiable rewards (RLVR) on small models (Qwen3.5 4B/9B, Gemma 3 12B) against frontier reasoning models. Using the Swiss multilingual legal system as a testbed, they find RLVR outperforms SFT for legal NMT and brings small models close to frontier performance, though a gap remains. The study also observes diminishing returns from re-training as model size increases. Code and models are publicly released.