Search arXiv⌕ Search

arXiv · 2610.01008

Modeling Shipping Emissions: Machine Learning, Engineering, and Policy Counterfactuals

Abstract

Machine learning predicts outcomes well, but predictive accuracy does not ensure reliable counterfactual responses. We examine how to combine machine learning and theory for measurement and counterfactual analysis, using maritime CO2 emissions where physics provides a benchmark speed response. Matching hourly tracking data for dry bulk and container ships to annual fuel consumption reported under EU regulations, we compare engineering calculations, structural regressions, hybrid models, and machine learning. Out of sample, all estimated models predict within a few percent of reported totals, outperforming standard engineering calculations. Yet pure machine learning and unrestricted structural regression imply attenuated speed responses. Hybrids preserve the structural component's speed response by excluding speed-related inputs from the machine learning component. A cost-benefit analysis of speed reductions illustrates the policy stakes: an attenuated speed response can flip the sign of net benefits. Accurate aggregate predictions therefore cannot substitute for scrutiny of the restrictions determining counterfactual responses.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Hiroyuki Kasahara, Allen Peters, Oliver Xu. 2026-10-01. Modeling Shipping Emissions: Machine Learning, Engineering, and Policy Counterfactuals. https://arxiv.org/abs/2610.01008

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Hardening Soft Information: Evidence on Analyst Integration Costs

We examine how the cost of transforming qualitative information into precise numerical estimates--a form of integration cost--creates a structural friction in expectations formation. To isolate this integration cost from the costs of information awareness and acquisition, we exploit sell-side analyst reports, in which the same forecaster simultaneously produces textual narratives and numerical forecasts. Because the information underlying the text has already been acquired, any systematic gap between the two outputs can be attributed to integration costs. We document systematic quantification inefficiency: an analyst's textual tone negatively predicts her contemporaneous forecast errors and positively predicts her subsequent numerical revisions, revealing that analysts leave part of their qualitative insights unquantified until further evidence arrives. Consistent with this integration-friction explanation, this inefficiency intensifies when reports are linguistically vaguer, environmental uncertainty is higher, or analysts' processing capacity is more constrained, and it persists where strategic and behavioral explanations are weaker. Our findings provide direct, large-sample evidence that integration costs constitute a distinct economic friction, explaining why soft information carries value-relevant content beyond contemporaneous hard numbers.

econ.GN↗

AI as Coordination-Compressing Capital: Task Reallocation, Organizational Redesign, and the Regime Fork

Task-based models of AI hold organizational structure fixed. We model AI as agent capital that compresses managers' per-link coordination costs toward residual floors for consequential decisions. Positive floors bound spans of control; zero floors do not. For a fixed manager pool, proportional team sizes, equal mean team quality and fixed within-firm pay shares, and when every floor is positive, the floors determine long-run managerial wage dispersion, independent of execution-compression parameters. With a common positive floor, distinct skills and skill-biased execution savings, managerial inequality is purely transitional: zero at zero capital and in the long run, positive at every positive finite capital level. If every lower-floor manager approaches capacity faster than every higher-floor one, inequality exceeds its long-run level at sufficiently large finite capital and converges from above, a non-monotone path. Numerical illustrations, not estimates, include a simulation with skill-sorted workers, to which the wage and Gini results do not apply.

econ.GN↗

Daycare Matching with Siblings: Social Implementation and Welfare Evaluation

In centralized matching markets, agents may value joint assignment, as with siblings or couples. Standard preference estimation ignores such complementarities, complicating welfare analysis of priority rules for paired assignment. We develop an empirical framework incorporating these preferences and apply it to Japanese daycare assignment. Families face both additional commuting distance and a fixed disutility from split assignment. We estimate the latter at 4.61 commuting-kilometer equivalents. Our fixed-report counterfactual estimates that the reform increased mean welfare by 0.032 kilometer-equivalent units. Ignoring sibling complementarity understates welfare gains for households applying simultaneously for multiple children by about 27%.

econ.GN↗