<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Decision Theory | Sandro Radovanović</title><link>https://sandror.netlify.app/tag/decision-theory/</link><atom:link href="https://sandror.netlify.app/tag/decision-theory/index.xml" rel="self" type="application/rss+xml"/><description>Decision Theory</description><generator>Wowchemy (https://wowchemy.com)</generator><language>en-us</language><copyright>© 2026 Sandro Radovanović</copyright><lastBuildDate>Fri, 09 Oct 2026 00:00:00 +0000</lastBuildDate><image><url>https://sandror.netlify.app/media/icon_hue52cc3259761a1f3d68e013d051cf4b8_149327_512x512_fill_lanczos_center_3.png</url><title>Decision Theory</title><link>https://sandror.netlify.app/tag/decision-theory/</link></image><item><title>Large Language Models and Decision-Making Theory</title><link>https://sandror.netlify.app/post/llm_decision_theory_tutorial_2026/</link><pubDate>Fri, 09 Oct 2026 00:00:00 +0000</pubDate><guid>https://sandror.netlify.app/post/llm_decision_theory_tutorial_2026/</guid><description>&lt;p>The next one in the &lt;a href="https://sandror.netlify.app/post/ai_tutorials_2026/">AI tutorial series&lt;/a>, and this time purely out of curiosity:&lt;/p>
&lt;p>📄 &lt;a href="https://drive.google.com/file/d/1g_fzS-S-JAqYDc9EyoHaJuAY81QoGWfj/view?usp=sharing" target="_blank" rel="noopener">&lt;strong>Large Language Models and Decision-Making Theory&lt;/strong>&lt;/a>&lt;/p>
&lt;p>I teach &lt;a href="https://sandror.netlify.app/teaching/">decision theory&lt;/a>, so the question behind this tutorial was irresistible: if you run the classic experiments on a large language model, what do you get? Does it satisfy the von Neumann–Morgenstern axioms? Does it violate independence the way people do in the Allais paradox? Is it loss averse? Does it update like a Bayesian?&lt;/p>
&lt;p>The tutorial covers the theory first — preference axioms, revealed preference and GARP, prospect theory, heuristics and biases, Bayesian updating, games — and then what has actually been found when these are tested on LLMs. The honest answer turns out to be &lt;em>&amp;ldquo;it depends&amp;rdquo;&lt;/em>: sometimes the model is more normatively consistent than any human, sometimes it reproduces human biases, and quite often it does something that is neither, which may be a genuinely non-human pattern or simply a measurement artefact.&lt;/p>
&lt;p>That last part is what I found most interesting. A good chunk of the tutorial is about how to test properly — option-order effects, reading logits versus sampled text, benchmark contamination, personas, reliability — because many published findings in this area are arguably measuring the harness rather than the model.&lt;/p>
&lt;p>As always: AI-generated, so treat it as a map rather than gospel. Enjoy.&lt;/p></description></item></channel></rss>