Thrilled to release the first LLM persuasion benchmark with user personas in our paper: Ψ-Bench: Evaluating Perso…
Thrilled to release the first LLM persuasion benchmark with user personas in our paper: Ψ-Bench: Evaluating Persona-Sensitive Influencing in Persuasive Dialogues!
Paper:
https://
arxiv.org/pdf/2606.02754
Code:
https://
github.com/Hanpx
How it ranks: Backlist reads my Twitter/X timeline, scores every tweet for substance with an LLM rubric (not engagement), and publishes the daily top picks with a one-line takeaway. Curated by Surya Dantuluri.