TY - RPRT TI - DERELAB: Probing Defeasible Reasoning and Confirmation Bias in LLMs with a Generative Benchmark AU - Jayanta Sadhu AU - Sayem Shahad AU - Kenneth Marino PY - 2026 UR - https://arxiv.org/abs/2608.30413 ID - 2608.30413 ER -