Should we align models based on human preference or human behaviour? We explore Alignment Makes Language Models Normative, Not Descriptive (Shapira, et al., 2026) to find out.
Should we align models based on human preference or human behaviour? We explore Alignment Makes Language Models Normative, Not Descriptive (Shapira, et al., 2026) to find out.