📊 Статистика дайджестов

Всего дайджестов: 34022 Добавлено сегодня: 82

Последнее обновление: сегодня

📄 GTAlign: Game-Theoretic Alignment of LLM Assistants for Mutual Welfare

2025-10-14

Авторы:

Siqi Zhu, David Zhang, Pedro Cisneros-Velarde, Jiaxuan You

Саммари на русском не найдено
Доступные поля: ['id', 'arxiv_id', 'title', 'authors', 'abstract', 'summary_ru', 'categories', 'published_date', 'created_at']

Annotation:

Large Language Models (LLMs) have achieved remarkable progress in reasoning, yet sometimes produce responses that are suboptimal for users in tasks such as writing, information seeking, or providing practical guidance. Conventional alignment practices typically assume that maximizing model reward also maximizes user welfare, but this assumption frequently fails in practice: models may over-clarify or generate overly verbose reasoning when users prefer concise answers. Such behaviors resemble the...

ID: 2510.08872v1 cs.AI, cs.GT, cs.HC, cs.LG, cs.MA

arXiv PDF