Kahneman-Tversky Optimization (KTO) is a preference learning method that only needs binary good/bad labels per output rather than paired comparisons, making data collection simpler than DPO. Inspired by prospect theory from behavioral economics, KTO models the asymmetry between gains and losses in human preference judgments.