Skip to Main Content (Press Enter)

Logo UNIMI
  • ×
  • Home
  • Persone
  • Attività
  • Ambiti
  • Strutture
  • Pubblicazioni
  • Terza Missione

Expertise & Skills
Logo UNIMI

|

Expertise & Skills

unimi.it
  • ×
  • Home
  • Persone
  • Attività
  • Ambiti
  • Strutture
  • Pubblicazioni
  • Terza Missione
  1. Pubblicazioni

Tuning LLM-Based Advisors for the Common Good: The Case for Direct Preference Optimization

Contributo in Atti di convegno
Data di Pubblicazione:
2026
Citazione:
Tuning LLM-Based Advisors for the Common Good: The Case for Direct Preference Optimization / L. Mauri, G. Sargsyan, E. Damiani - In: FLLM[s.l] : Institute of Electrical and Electronics Engineers (IEEE), 2026. - ISBN 979-8-3315-9410-7. - pp. 910-915 (( 3. International Conference on Foundation and Large Language Models : November, 25 - 28 Vienna 2025 [10.1109/fllm67465.2025.11391029].
Abstract:
Large Language Models (LLMs) are often deployed as advisors to consumers, e.g., recommending purchases, and to managers, e.g., suggesting new hires. In fact, LLMs provide advice based on cost or convenience, overlooking broader societal impacts (e.g., carbon footprint when recommending products to a potential customer, or fairness when recommending a new hire to a company). To align LLMs' advice with societal goals like environmental sustainability and gender parity, tuning strategies must integrate the notion of common good. We discuss why Direct Alignment tuning could be preferable to classic Reinforcement Learning from Human Feedback to achieve this integration. Then, we describe and compare two approaches to Direct Preference Optimization: (1) exposing the model tuning examples taken from recommendations and regulations, and (2) mythopoiesis, i.e., model tuning based on synthetic ``legends'', fictional success stories of regulatory compliance (also generated by LLMs). We present a pipeline to evaluate legends' effectiveness in reducing bias and fostering compliance. Our preliminary results suggest that legend-based tuning may enhance engagement and generalization, while direct exposure ensures factual accuracy but risks rigidity.
Tipologia IRIS:
03 - Contributo in volume
Keywords:
large language models; legend-based tuning; AI alignment; regulatory compliance; benchmarking LLMs;
Elenco autori:
L. Mauri, G. Sargsyan, E. Damiani
Autori di Ateneo:
DAMIANI ERNESTO ( autore )
Link alla scheda completa:
https://air.unimi.it/handle/2434/1233599
Link al Full Text:
https://air.unimi.it/retrieve/handle/2434/1233599/3300204/paper_FLLM2025.pdf
Titolo del libro:
FLLM
Progetto:
MUSA - Multilayered Urban Sustainability Actiona
  • Aree Di Ricerca

Aree Di Ricerca

Settori


Settore INFO-01/A - Informatica
  • Informazioni
  • Assistenza
  • Accessibilità
  • Privacy
  • Utilizzo dei cookie
  • Note legali

Realizzato con VIVO | Progettato da Cineca | 26.6.2.0