TY - RPRT TI - Learning in Continuous Games from Pairwise Preference Feedback AU - Anas Barakat PY - 2026 UR - https://arxiv.org/abs/2610.05428 ID - 2610.05428 ER -