Search arXivSearch

arXiv · 2609.00187

Ordinal Gates, Cardinal Bets: Matching LLM Confidence to the Financial Decision Operator

Abstract

LLM confidence scores are not independently deployable objects: their decision value depends on the downstream operator and exposure controller that consume them. Monotone recalibration cannot change a coverage-matched rank-based gate, whereas position sizing consumes score magnitude, so changing a confidence map can invalidate a scale fitted to the previous score distribution. We test this on FactSet news for Nasdaq-100 equities, fitting maps and scales on 2021 and evaluating nine open-weight LLMs out-of-sample on 2022--2023. Cross-applying raw and correctness maps with independently fitted scales shows that the two components are not portable alone: scale transfer reduces certainty-equivalent return (CER) in $8/9$ models and produces large risk-target errors. Matching each map with its fitted scale improves ensemble CER by $9.2$ percentage points per year under frozen-scale control ($p<0.001$), and the effect remains significant when the single largest-contributing model is excluded ($+5.5$pp/yr), so it is not driven by one case. Under an identical adaptive-volatility controller, however, the incremental effect falls to $+1.6$pp/yr, with a significant controller interaction. Annual walk-forward effects are smaller, although map--scale interaction remains positive in every fold. Confidence transformations should therefore be evaluated jointly with the downstream controllers that consume them.

Explore related subjects

Keep this discovery

BibTeXRIS

Rayansh Singh, Sara Rezaeimanesh. 2026-09-02. Ordinal Gates, Cardinal Bets: Matching LLM Confidence to the Financial Decision Operator. https://arxiv.org/abs/2609.00187

Cite the original work for its findings. Save a collection to share your selection of sources.

Discover connections

Connections use source metadata and explicit phrase matches, not verified experimental comparisons.

KEEP EXPLORING

Related papers

Hemispherical Ray-Casting Analysis for Milling Configuration Classification and Machinability Assessment

Early-stage manufacturability assessment reduces design iterations and costly downstream changes. This paper presents a GPU-accelerated geometric analysis method that evaluates CNC machinability and setup requirements directly from CAD files. A multi-stage pipeline uses hemispherical ray-casting to quantify surface accessibility, followed by directional variability analysis to extract dominant machining axes. Threshold-based classification determines 3-axis versus multi-axis machining. For 3-axis parts, orthogonal ray-casting and reachability optimization compute the minimum setup count and explicitly flag inaccessible surfaces. Validation cases show computation times sufficiently short to support interactive CAD workflows, enabling quantitative manufacturability feedback during the design stage without requiring manufacturing expertise.

cs.CE

Intrinsic Finite Element Methods for Fluids on Riemannian Manifolds Compared with Surface FEM

We present an intrinsic finite element formulation for the incompressible Navier--Stokes equations on Riemannian manifolds. We derive the corresponding weak formulation and prove that the backward Euler discretisation is energy stable. The proposed framework is validated on several representative manifolds, with particular attention paid to the long-time behaviour of the flow and its convergence to steady-state solutions represented by Killing vector fields. Comprehensive comparisons are performed with the surface finite element method and a corresponding eigenvalue formulation for Killing vector fields. The numerical results demonstrate that the intrinsic formulation provides an accurate, computationally efficient, and geometrically transparent alternative to embedded surface finite element formulations, while naturally extending to higher-dimensional Riemannian manifolds.

cs.CE