Loading…
StalePO: Anchored Token-Level Preference Optimization using Legacy Post-Edits in Machine Translation · Researchar