Economics

2026-07-28 | | Total: 25

#1 Debiased Machine Learning: Identification, Estimation, and Shape Constraints [PDF] [Copy] [Kimi] [REL]

Authors: Qihui Chen, Ka Yan Cheng, Zheng Fang

We develop a general framework of identification and estimation for automatic debiased machine learning (DML) where the parameter of interest $θ_0$ is identified by a moment condition involving a nuisance $γ_0$ that may be high dimensional. DML leverages machine learning to estimate $γ_0$ while correcting for regularization and overfitting biases that may otherwise transmit to biased estimation of $θ_0$. We establish conditions under which the Riesz representer $α_0$, which is at the core of DML, is identified, and show that the identification occurs precisely when $α_0$ uniquely optimizes a quadratic functional. This characterization enables us to develop a general estimation procedure for $α_0$ that allows for generic $γ_0$ including those defined by models with endogeneity and encompasses both classical sieves and modern architectures such as deep neural networks. To improve estimation precision and mitigate the curse of dimensionality, we incorporate shape constraints on $γ_0$ by embedding them into a possibly nonlinear parameter space. We illustrate our estimation procedure through simulations and empirical applications.

Subjects: Econometrics , Statistics Theory , Methodology , Machine Learning

Publish: 2026-07-27 14:08:15 UTC


#2 How to Disrupt a Market [PDF] [Copy] [Kimi1] [REL]

Authors: Edoardo Gallo, Rebecca Heath, Jonathan Lusthaus, Federico Varese

Market design research in economics naturally focusses on how to improve market efficiency. Our objective here is exactly the opposite - how to design interventions that make a market less efficient. Our research is inspired by the growth of illicit markets online where reducing their efficiency may reduce societal harm. Using a web-based experiment, we find that a partial disruption to delivery is an effective method to decrease market efficiency. The decrease is borne by sellers who sell fewer goods and have lower earnings. A consequence of a disruption to delivery, however, is an increase in market concentration because it facilitates the emergence of a dominant seller. In contrast, we find that attacks on seller ratings are ineffective at reducing market efficiency. This study paves the way for evidence-based, causally driven investigations to aid policies to disrupt cybercrime and other illicit markets.

Subject: General Economics

Publish: 2026-07-27 13:04:10 UTC


#3 Randomness in large language models: What researchers need to know (and report) [PDF] [Copy] [Kimi] [REL]

Authors: Guillaume Coqueret, Joan Llull, Florian Oswald, Christophe Pérignon, Christoph Scheuch, Lars Vilhuber

Large language models (LLMs) are increasingly used to generate data for research. Typical use cases are classifications, annotations, information extraction, and generation of numerical scores. Unlike conventional measurements, LLM outputs can vary across repeated requests even when the prompt and apparent model settings remain unchanged. This variation arises from deliberate sampling, silent model updates, numerical rounding, or expert routing. Setting a dedicated temperature parameter to zero removes deliberate sampling when that option is available, but it does not eliminate the other sources of randomness. Exact reproduction is therefore generally not possible when using proprietary application programming interfaces. Local execution of open-weight models offers greater control, but reproducibility still depends on the complete hardware and software stack. We illustrate these issues through sentiment classifications of corporate filings and examine their consequences for downstream regression results. We then propose a reporting standard for articles and replication packages, as well as guidance for data editors and authors. Together, these findings and recommendations establish that LLM outputs should be treated as draws from a distribution rather than as fixed measurements.

Subject: General Economics

Publish: 2026-07-27 12:45:35 UTC


#4 A World of Ginis [PDF1] [Copy] [Kimi1] [REL]

Authors: Lidia Ceriani, Paolo Verme

The Gini index remains the most important measure of economic inequality worldwide, and accurate estimates of this index are essential for effective public policies. Yet, Gini estimates for the same country and year vary considerably across data sources, a problem that remains largely unresolved. The paper reviews the largest global and regional databases providing Gini estimates, surveys the related literature, and constructs a unified dataset of 122,351 Gini observations spanning 222 countries and territories and 158 years, from 1867 to 2024. The analysis of this new dataset shows that income-based Ginis exceed consumption-based ones by 4.7 points on average globally, and by as much as 10 points in some regions, with these gaps widening over time. The gross--net income distinction and the use of alternative equivalence scales together with several other measurement choices add further systematic differences. Based on these findings, the paper provides correction factors that can be used to harmonise Ginis built on different welfare concepts. We further show that overall divergence across databases has grown only modestly since 1960, and mainly through the proliferation of databases rather than through genuine divergence among long-standing sources. Thus, improving on the existing discrepancies across Ginis globally is possible, but ultimately depends on database administrators disclosing full details of Gini construction and on users selecting Ginis built on comparable measures.

Subjects: General Economics , Applications

Publish: 2026-07-27 09:00:40 UTC


#5 Robust estimation of the autocorrelation function via forward ratios [PDF] [Copy] [Kimi] [REL]

Authors: A. Montañés, E. Ruiz

It is obvious to say that an adequate estimation of the autocorrelation function is central in time series analysis. In this paper, we propose three new robust estimators based on ratios of observations, which offer strong resistance against outliers. While the first estimator, which is based on the median, is not efficient, the second is a Quasi Maximum Likelihood (QML) estimator with better efficiency properties. The third estimator is a plug-in estimator, which does not require numerical optimization and, consequently, is extremely simple from a computationally point of view, having similar efficiency to that of the ML estimator. We derive the asymptotic distribution of the first two estimators, when the true autocorrelations are zero. Furthermore, we also show that the asymptotic distribution of the plug-in estimator is rather close to that of the QML estimator, allowing for inference and, in particular, for the construction of point-wise significance bands for the autocorrelations. Using Monte Carlo simulations, we analyse the finite sample properties of the proposed estimators and compare them with those of the sample autocorrelations and alternative extant robust estimators based on ranks. Although the proposed estimators have larger dispersion than the sample autocorrelations in uncontaminated time series, they are highly robust in the presence of outliers. Also, they have better properties than popular alternative robust estimators based on ranks when estimating autocorrelations of order larger than one. The results are illustrated by estimating the correlogram of daily IBEX35 returns, quarterly US economic growth and monthly US inflation.

Subject: Econometrics

Publish: 2026-07-26 16:38:35 UTC


#6 Systemic Methodological Dysfunction in Statistical Research for Clinical Decisions [PDF] [Copy] [Kimi] [REL]

Author: Charles F. Manski

I critique a set of entrenched methodological conventions that collectively create systemic dysfunction in statistical research for clinical decisions. These include: (1) the prevalent use of hypothesis tests to compare treatments, (2) remoteness from patient care of the methods used to evaluate the accuracy of predictions of patient outcomes, (3) poor practice of meta-analysis to combine findings across studies, and (4) widespread research with incredible certitude. It appears that the dysfunction is held in place by three factors: (i) rudimentary instruction in statistical methodology received by medical students and residents, (ii) reliance of clinical researchers on consulting biostatisticians, wo act as statistical gatekeepers in evaluation of grant proposals and paper submissions, and (iii) institutional practices of research funding agencies, medical journals, and governmental bodies that regulate medical treatment. I conjecture that systemic changes are necessary to break the existing impasse, moving statistical research to a better equilibrium.

Subjects: Econometrics , Methodology

Publish: 2026-07-26 13:49:16 UTC


#7 Do Carbon Price Forecasts Improve Compliance Procurement? Evidence from European Union Allowances [PDF] [Copy] [Kimi] [REL]

Authors: Muzi Chen, Difang Huang, Shouyang Wang, Xinghan Xia

Firms covered by emissions trading systems need forecasts not only to value allowances, but also to decide when to buy them. This paper asks whether European Union Allowance (EUA) prices contain short-horizon predictability that survives a forecast-origin information design and improves simulated compliance procurement. Using daily data from 2019 to 2025, we produce direct forecasts for one to five trading days ahead. All predictors are observable at the forecast origin, and calibration and model-selection rules are fixed before the final holdout. The released forecast has the lowest point-estimate RMSE at every horizon among fourteen benchmarks, with the strongest loss-difference evidence at horizons three and four. Relative to a random walk, out-of-sample R^2 rises from 1.2% at one day to 15.5% at five days. We then use the forecast path in a constrained procurement problem with execution costs, market impact, capacity limits, and tail risk; sensitivity exercises add demand uncertainty. For a fixed 100,000-EUA order, optimized schedules lower average realized costs by 8.5 to 38.5 basis points relative to uniform execution across horizons h=2 to h=5. The gains come from reallocating purchases within a fixed window, not from reliable next-day directional timing.

Subject: General Economics

Publish: 2026-07-26 02:54:26 UTC


#8 Wrong and More Confident: A Field Experiment on Language Models Taking a Graduate Economics Exam [PDF] [Copy] [Kimi] [REL]

Author: Piyush Akimitsu

A red herring, an irrelevant passage added to a problem, makes a language model reason incorrectly and answer incorrectly far more often. Yet the model still writes out a full explanation, and the answer it gives remains consistent with the steps it shows. The red herring corrupts the reasoning, while leaving the explanation intact and coherent. I show this on the Graduate Economic Reasoning Benchmark (GERB), sixty graduate-level microeconomics problems, each a detailed setup with a verified answer and a step-by-step reference solution, presented with and without the red herring and answered by thirty-eight language models in a within-task $2\times2$ design. The red herring lowers the probability of a correct answer by 12.3 percentage points, about a quarter of what these models answer correctly, and thirty-seven of the thirty-eight are less accurate under it. Reasoning capability confers no protection. The effect does not differ detectably between models that reason by default and models with no reasoning mode, nor between open-weight and closed-weight models, though open-weight models reach comparable accuracy at a substantially lower cost per correct answer. The damage is largest on the problems the model rates as easy. Worse, if anything the red herring leads a model to rate a problem as easier, not harder, than its clean version, even as it answers that problem wrong more often. The clean problems are already hard, and on average the models answer fewer than three of every five correctly. A wrong answer still comes with a full explanation, is coherent with its own reasoning, and the model stays confident, so the error cannot be caught without checking it against the verified answer.

Subject: General Economics

Publish: 2026-07-26 02:46:04 UTC


#9 Low-Rank Payoffs and Limit Uniqueness in Global Games [PDF] [Copy] [Kimi] [REL]

Author: Dana Golden

When does the global game information structure select a unique equilibrium? Limit uniqueness in two-player supermodular games fails exactly when a risk-dominant better response cycle exists (Veiel, 2025). We show that rank-one factor structure on payoffs eliminates such cycles entirely, so every rank-one supermodular game admits a generalized ordinal potential and limit uniqueness follows for any number of actions. The boundary is sharp: an explicit three-action rank-two game carries a length-six cycle, no supermodular game carries a cycle of length four, and every game within a quantified sup-norm margin of a nondegenerate rank-one game is cycle-free. Rank-one structure can also be manufactured: when players compete across many independent markets with common latent payoffs, the stacked observation matrix is rank one plus sparse, and a Robust PCA estimator leaves residual noise that vanishes with the signal scale yet stays positive at any finite sample, even under partial observation.

Subject: Theoretical Economics

Publish: 2026-07-25 20:56:10 UTC


#10 Happy Birthday? Age Labels, Search Criteria, and Matching from Dating to Marriage [PDF] [Copy] [Kimi1] [REL]

Author: Suguru Otani

Age is a match trait and a prominent label on search platforms. Using confidential records from a large Japanese marriage platform, I study how a birthday age update affects consideration, applications, relationship progression, and engagement. Only displayed age updates at birthdays. Receivers enter some acceptable-age ranges as they exit others, leaving only modest overall eligibility changes. Application counts often increase. Applications shift from younger to older suitors. Proposal counts fall at nine of ten female receiver ages \(31\text{--}40\), with losses concentrated at entry into early dating. In stage-specific accounting for receivers in their thirties, the no-birthday proposal count is \(12.9\) percent above the benchmark for female receivers and \(6.7\) percent below it for male receivers; the younger-man share in the female-receiver channel is \(4.8\) percentage points higher. Age labels and filters are not neutral windows onto preferences: they govern who marries whom and how many marriages form.

Subject: General Economics

Publish: 2026-07-25 18:37:41 UTC


#11 Agentic AI Orchestration of Heterogeneous Economic Models for Rapid, Multi-scenario Analysis of Energy Crises [PDF] [Copy] [Kimi] [REL]

Authors: Dana Golden, Brett Indelicato, Lav R. Varshney, Carlos D. Messina, Suzanne Thornsbury

Rigorous economic models can take months to construct, yet energy crises demand decisions from policymakers within days or even hours. Any disruption in energy markets is not isolated but rapidly disseminates through interlinked global systems. Off-the-shelf models that already exist typically focus only on limited aspects of the system and are distributed across research groups, programming languages, software architectures not designed for model integration, and incompatible formats. Integrating these models manually can take longer than the crisis itself, forcing analysts to rely on whichever models are easiest to connect and leaving consequential scenarios unexplored. Policymakers must make rapid decisions with obstructed and limited information. We show that large language models can perform the critical integration directly. The system constructs internally consistent scenarios, translates assumptions into model-specific inputs, executes existing economic and physical models in dependency order, and synthesizes outputs tailored to policymakers. The language model generates no quantitative results: every reported value is reproduced directly from an underlying model run, remains traceable to its source and is subject to analyst approval at each stage. We develop a LLM framework that coordinates 16 models of oil, natural gas, shipping, water, helium, fertilizer and macroeconomic equilibrium. The framework is applied across five scenarios to assess the 2026 closure of the Strait of Hormuz and refreshed weekly for eight weeks as events on the ground continued to unfold. By linking models that already exist and reading them as a suite rather than in isolation, this architecture mobilizes distributed scientific models rapidly during energy and geopolitical disruptions while keeping any single model's assumptions from driving the conclusion.

Subject: General Economics

Publish: 2026-07-25 18:02:40 UTC


#12 Ranking-based competitive balance measures in Formula One [PDF] [Copy] [Kimi] [REL]

Authors: Dóra Gréta Petróczy, László Csató

Competitiveness in racing sports can be measured by comparing the start and finish rankings within races, as well as the start and finish rankings across races in a season. Since the importance of position changes is non-uniform and variance at the top of the ranking is more interesting than at the bottom of the ranking, we propose using weighted distances for this purpose. Therefore, two weighting schemes are applied and compared to the standard Kemeny distance to analyse the Formula One between 1950 and 2024. The evolution of competitive balance is unexpectedly robust to the weights, but competition is more intense if one focuses on the top positions. Competitive balance has been more unfavourable in the last two decades than ever before. Statistical tests uncover three structural breaks in competitive balance that are closely related to regulatory changes, highlighting the role of decision-makers in the evolution of competitiveness.

Subjects: General Economics , Physics and Society , Applications

Publish: 2026-07-25 17:27:45 UTC


#13 Do Preferences Matter in Balanced Task Allocation? [PDF] [Copy] [Kimi] [REL]

Author: Terence Highsmith

I model balanced task allocation where tasks stochastically arrive and must be matched to a fixed set of agents; the novel constraint is that agents must receive allocations that require the same level of average effort. Social work supervisors, call center managers, and courts all rotate allocation across workers to satisfy this constraint, but the Rotation mechanism is not Pareto efficient. I design the Dynamic Pseudomarket (DPM) mechanism, and it satisfies Pareto efficiency and asymptotic balance. I derive an explicit equation characterizing DPM's expected productivity gain over Rotation that can be estimated only from aggregate statistics in firm-level data. Simulation results indicate large average productivity gains. These results indicate that preference-based allocation can Pareto dominate the status quo.

Subject: Theoretical Economics

Publish: 2026-07-25 14:18:59 UTC


#14 Do decisions about outliers and influential effects matter? Evidence from 358 behavioral science meta-analyses [PDF] [Copy] [Kimi] [REL]

Authors: Tomas Havranek, Zuzana Irsova, Martina Luskova, T. D. Stanley

Meta-analysts routinely face estimates that look too large or extreme. Yet, how to handle them is left to the reviewer's judgment. The methods for detecting such estimates are well known. What is missing is an informed assessment of how much alternative handling choices might change a meta-analysis' conclusions. We fill this gap by analyzing the effects of four pre-registered handling treatments across 358 behavioral science meta-analyses with at least ten estimates. Each outlier handling treatment is estimated by two estimators (random effects and unrestricted weighted least squares), and compared to the 'do-nothing' baseline on three outcomes: the pooled effect, statistical significance, and whether the effect reaches the smallest effect size of interest (|d| >= 0.20). Our entire analysis and comparison pipelines were pre-registered. Alternative outlier handling treatments have little effect on the meta-analysis mean as the median absolute change in Cohen's d is at most 0.047 and often much less. Yet, at least one of these four treatments in combination with one of these estimators reverses the statistical significance of 11.5% of meta-analyses and the smallest-effect-of-interest assessment in 15.9%. Winsorizing has the least effect and DFBETAS the most. Categorical changes are found almost entirely among results already close to the decision boundary; strongly significant results essentially never change. These findings give applied meta-analysts, methods specialists, and reviewers a reference point for how much this under-reported choice matters and provide yet another reason for meta-analysts to publicly pre-specify their methods and handling treatments.

Subject: Econometrics

Publish: 2026-07-25 12:09:02 UTC


#15 Public Goods Game on Complex Networks: the interplay between conformity and topology [PDF] [Copy] [Kimi] [REL]

Authors: Ren Manfredi, Eugenio Vicario, Ennio Bilancini, Rossana Mastrandrea

Human cooperation is a phenomenon that has been extensively studied, and to date several explanations have been proposed, from network reciprocity to behavioral mechanisms that incorporate social and cognitive aspects. In this work, we studied the combined effect of conformity and network structure on the evolution of cooperation in the spatial Public Goods Game. By assigning agents different individual sensitivities to payoffs and neighborhood behavior, we explored the cooperative dynamics of this heterogeneous population on both regular and complex topologies. Our results show how the interaction between conformity and the distinctive features of each network can lead to very different outcomes, from the promotion of cooperation in regular topologies to null or negative effects in heterogeneous networks.

Subject: Theoretical Economics

Publish: 2026-07-25 10:19:44 UTC


#16 Attenuated Heterogeneity in Fixed-Effects Causal Forests, and a Cross-Fitted Correction [PDF] [Copy] [Kimi] [REL]

Author: Harry Aytug

Causal forests that estimate conditional average treatment effects by averaging honest leaf-level effects across trees are widely used in fixed-effects panel settings. We show that this averaging systematically attenuates the estimated heterogeneity: the raw prediction behaves like a + b*tau(x) with slope b < 1, so the spread of the CATEs is compressed toward the average effect, and the additive recentering used to report an unbiased average treatment effect does not fix it. Benchmarking against a similarity-weight generalized random forest on the same within-transformed signal, we find both estimators attenuate but the leaf-averaging construction attenuates materially more. We characterize how b moves with the design, worsening with lower signal-to-noise, smaller panels, and higher dimension; this diagnosis is our main contribution. As a remedy we adapt the best-linear-predictor calibration of Chernozhukov et al., estimating the de-attenuation slope out-of-bag so that it is self-contained within the observational panel and asymptotically inert under a homogeneous effect. In simulations the correction cuts CATE mean-squared error by 25-42% relative to the recentering default; on a standard county minimum-wage panel the attenuation is present but mild and the correction restores the imposed spread. We ship the method in the causalfe Python package.

Subjects: Econometrics , Machine Learning

Publish: 2026-07-24 20:18:11 UTC


#17 What should the encroaching supplier do?: A Stackelberg Game Approach [PDF] [Copy] [Kimi] [REL]

Authors: Gurkirat Wadhwa, Veeraruna Kavitha

Suppliers often encroach downstream by operating in-house production-units while continuing to supply independent production-units. We study the optimal configuration, including optimal pricing, for an encroaching supplier that balances these dual roles through a Stackelberg game. The integrated supplier determines the wholesale price charged to the outsourced production unit and the retail price of its own product, while the outsourced unit responds optimally. Customer demand-response incorporates both price-based substitutions (of the two production-units) and loyalty (towards individual units). With strong customer loyalty and luxury products, at the optimal choice for the coalition, both units co-exist profitably. In contrast, when the products become essential, the optimal strategy depends upon customer-fallback rates (fraction of the exiting production-unit's market that falls-back to other). Under low fallback, the coalition either sustains co-existence at maximum prices or disciplines the out-house to operate at break-even---with high fallback it is optimal to shut-down the in-house or eliminate the out-house---we derive two factors that identify the above. We further develop a numerical procedure to identify the optimal regime for any given set of parameters. Two surprising results are---higher market potential of the out-house can become a reason for it to operate at break-even---and the coalition may find it beneficial to operate its in-house at losses, particularly for products that are neither highly essential nor in the luxury category.

Subject: Econometrics

Publish: 2026-07-24 18:37:15 UTC


#18 Inference on counterfactual distributions using martingale posteriors [PDF] [Copy] [Kimi] [REL]

Authors: Gregor Steiner, Mark Steel

Causal inference is often focused on average effects, which can hide important aspects of the effect distributions. Here we consider the entire posterior effects distribution by estimating full counterfactual outcome distributions. We propose a methodology for inference on counterfactual distributions which builds upon the martingale posterior framework of Fong et al. (2023). This provides a highly flexible approach to estimating densities, distribution functions, and derived quantities such as quantiles, which coherently quantifies the epistemic uncertainty on any target estimand of interest. As the predictive recursions are based on an underlying nonparametric model (a Dirichlet process mixture model), our method naturally inherits robustness with respect to restrictive parametric assumptions. In addition, implementation of our method is typically very fast. This approach can be applied to marginal or conditional counterfactual distributions and is easily extended to an instrumental variables setup. Using the concept of almost conditionally identically distributed random variables, we prove convergence of the martingale posterior inference on the counterfactual outcome distributions for the causal models considered in the paper. We illustrate our approach on both simulated and real data. Using the latter, we investigate the effect of zinc lozenges on common cold duration, the impact of vitamin A supplementation on children's survival rates with one-sided non-compliance (analysed in Imbens and Rubin, 1997a) and the effect of job training (LaLonde, 1986).

Subjects: Methodology , Econometrics

Publish: 2026-07-27 08:23:09 UTC


#19 AI Strategy: How to Choose What AI Product to Implement [PDF] [Copy] [Kimi1] [REL]

Authors: Foster Provost, Panos Ipeirotis

Firms struggle to choose AI projects that pay off: two projects can look equally promising to smart, motivated stakeholders and yet deserve opposite decisions. At the residential real-estate brokerage Compass, one AI product (Likely-to-Sell recommendations) flagged sales outreach opportunities and went on to account for nine figures in annual gross commission revenue. Another championed AI product (a Time-on-Market pricing tool) was rightly shelved. A simple ROI estimate could not distinguish the two. We present expected ROI (eROI), a framework that decomposes each bet into three components and rates them separately: Value if Successful, Likelihood of Success, and Investment Required. Each maps to a question executives can answer before building: How valuable would it be if it worked? How likely is it to work? And what would it cost to implement? Separating the three breaks a common catch-22: teams cannot estimate ROI until they know whether a project will work, yet cannot know whether it will work without building it. Judging Value if Successful on its own dissolves the loop, letting a team argue that a product would be valuable if it worked while it weighs how likely that is. The framework also asks, before ranking anything, whether there are enough good ideas on the table. After ranking, it guides assembling a portfolio of bets rather than funding only the single top-ranked project. We illustrate eROI on Compass's candidate AI products. Precise ROI estimates are hard to make given the inherent uncertainty of AI projects. Coarse business-level ratings of the three components are enough to tell strong bets from weak ones.

Subjects: Computers and Society , Artificial Intelligence , Machine Learning , General Economics , Applications

Publish: 2026-07-26 16:05:39 UTC


#20 The One-Period Kyle (1985) Model Has a Unique Equilibrium: A Monotone Gaussian Bayes inverse-rigidity theorem [PDF] [Copy] [Kimi] [REL]

Author: Rabee Tourky

Let $V$ and $U$ be independent standard normal random variables. For any Borel map $φ\colon\mathbb{R}\to\mathbb{R}$, set $Y_φ=φ(V)+U$, and define $P_φ(y)=\mathbb{E}[V\mid Y_φ=y]$ and $F_φ(x)=\mathbb{E}[P_φ(x+U)]$. We prove that, if for every $v\in\mathbb{R}$, the quantity $φ(v)$ maximises $x(v-F_φ(x))$ over $x\in\mathbb{R}$, then $φ$ is the identity function. This is the normalised one-period Kyle (1985) model of insider trading. It follows that Kyle's closed-form affine strategy is the unique equilibrium of the model for arbitrary Gaussian location and scale, and that its canonical competitive pricing rule is necessarily linear. This settles a long-standing question in financial economics. Boulatov, Kyle and Livdan (2005, 2013) introduced complex-analytic techniques to the problem. McLennan, Monteiro and Tourky (2017) proved a linear growth bound for $P_φ$, real-entire regularity of $F_φ$, and uniqueness when an equilibrium strategy agrees locally with a uniquely continuable analytic function. The present proof requires no regularity assumption on $φ$ beyond Borel measurability. Its argument is real-variable and probabilistic: the maximisation forces sharp upper and lower bounds for Gaussian-randomised monotone functions to coincide, and the corresponding equality cases admit only the identity function.

Subjects: Probability , Econometrics , Theoretical Economics , Functional Analysis

Publish: 2026-07-26 10:27:29 UTC


#21 Bitcoin Price Direction Prediction via Regime-Aware Multi-Modal Fusion of Social Sentiment and Technical Features [PDF] [Copy] [Kimi] [REL]

Author: Muhammad Abdullah Haroon

Bitcoin price prediction on sub-daily timescales is a hard open problem in computational finance. Bitcoin exhibits fat-tailed returns, non-stationary dynamics, and a price discovery process influenced by social discourse on Reddit and Twitter. Conventional approaches fuse OHLCV technical features with sentiment via static concatenation, applying identical fusion weights regardless of market state. This is inconsistent with the behavioural finance literature, which shows that retail sentiment is most predictive during volatile periods and noisy during calm ones. This paper proposes Regime-Aware Multi-Modal Learning (RAML), which conditions fusion of sentiment and price features on a dynamically detected binary market regime. Rolling 24-hour volatility partitions observations into stable and volatile regimes; a learnable sigmoid gate adjusts the weight of the sentiment embedding relative to the price embedding, trusting sentiment more during volatility and price dynamics more during stable phases. The system is evaluated on 3,491 hourly observations (July 2024-September 2025), combining Bitcoin OHLCV data with Reddit /r/Bitcoin FinBERT sentiment. Four models are compared - price-only BiLSTM, sentiment-only classifier, static-concatenation BiLSTM, and RAML - across 3-hour and 6-hour horizons, with an ablation study isolating the sentiment branch, regime detection, and adaptive fusion. RAML achieves macro-F1 of 0.5474 (3h) and 0.5513 (6h), with the highest AUC at 3 hours (0.5084), indicating better calibration. Ablation confirms every component is necessary, and replacing adaptive weighting with concatenation causes recall collapse at 6 hours (F1: 0.14). These results establish regime-conditioned adaptive fusion as a necessary design principle for multi-modal financial forecasting.

Subjects: Machine Learning , Computational Engineering, Finance, and Science , Econometrics , Computation

Publish: 2026-07-25 21:25:12 UTC


#22 Fair Division with Strictly Increasing Valuations: A Tight Threshold for Two-Agent EF1 and PO [PDF] [Copy] [Kimi] [REL]

Author: Nicholas Teh

We study whether strictly positive marginal values restore the compatibility of envy-freeness up to one good (EF1) and Pareto optimality (PO) for indivisible goods. For two agents, we identify the exact threshold in the number of goods. Every instance with at most seven goods and strictly increasing valuations admits an allocation that is both EF1 and PO, without any submodularity assumption. In contrast, we construct an eight-good instance with normalized, integer-valued, strictly increasing, submodular valuations in which every EF1 allocation is strictly Pareto dominated. Thus, eight goods are necessary and sufficient for a two-agent counterexample. Finally, we strengthen the three-agent NP-hardness result of Chandramouleeswaran and Nimbhorkar (2026): deciding whether an EF1 and PO allocation exists remains NP-hard for normalized, integer-valued, monotone submodular valuations even when zero marginals are confined to eight fixed agent-good pairs, all involving a single agent.

Subjects: Computer Science and Game Theory , Artificial Intelligence , Theoretical Economics

Publish: 2026-07-25 21:17:38 UTC


#23 Online Fair Division with Budget Constraints [PDF] [Copy] [Kimi] [REL]

Authors: Saar Cohen, Nicholas Teh, Paul W. Goldberg, Michael J. Wooldridge

We study an online variant of discrete fair division under generalized assignment budget constraints. Goods arrive one at a time and must be assigned irrevocably to a feasible agent or to charity, which holds all unallocated goods, while fairness is evaluated only against budget-feasible subsets of every recipient's bundle. We first show that, without additional structure, no deterministic online algorithm can guarantee any fixed approximation to feasible envy-freeness, even in highly symmetric instances. We then identify bounded density spread as a structural condition that restores meaningful guarantees, obtaining approximation algorithms for arbitrary item sizes and showing that, under common valuations and sufficiently small goods, these guarantees can be strengthened to an optimal deterministic frontier. We further study resource augmentation, where the online algorithm is allowed slightly larger budgets than the fairness benchmark, and characterize the resulting improvement in the achievable guarantees. Finally, we develop a learning-augmented framework based on predicting joint value-size types, proving consistency under perfect predictions, robustness to prediction error, and showing that separate predictions of value and size marginals are insufficient to recover strong fairness guarantees.

Subjects: Computer Science and Game Theory , Artificial Intelligence , Machine Learning , Multiagent Systems , Theoretical Economics

Publish: 2026-07-25 17:53:10 UTC


#24 Towards Optimal Estimators for Randomized Control Trials [PDF] [Copy] [Kimi] [REL]

Authors: Harsh Parikh, Gabriel Levin-Konigsberg, Nilesh Tripuraneni, Dhruv Madeka, Michael I. Jordan, Dean Foster, Dominique Perrault-Joncas, Alexander Volfovsky

Randomized controlled trials (RCTs) are fundamental tools for causal inference across technology companies, pharmaceutical research, and federal agencies. While the standard difference-in-means estimator provides unbiased treatment effect estimates, it often lacks precision, particularly when treatment effects are heterogeneous or outcomes exhibit heavy-tailed distributions. Although numerous precision-enhancing methods exist---from covariate adjustment techniques to variance reduction strategies---recent research demonstrates that no single estimator performs optimally across all datasets. Rather than seeking the best estimator for individual RCTs, which risks compromising scientific validity through convenient selection, we propose a principled framework for identifying optimal estimators within families of RCTs based on specific analytical goals. Our approach uses sample splitting to estimate the distribution of evaluation metrics (e.g., mean squared error, regret) across RCT families, enabling systematic comparisons between estimators while maintaining asymptotic guarantees. We demonstrate this framework using a sample of Amazon's Supply Chain Optimization Technology trials and the Strengthening Democracy Challenge dataset (25 interventions). Results reveal that optimal estimators vary significantly by analytical objective: weighted least squares performs best for inference goals, while difference-in-means minimizes regret for decision-making contexts. This work provides actionable guidance for estimator selection while preserving methodological rigor across diverse research applications.

Subjects: Applications , Econometrics , Methodology

Publish: 2026-07-25 15:28:57 UTC


#25 Risk Aversion in the Small and in the Large: Beyond Arrow-Pratt A Wiener Chaos Hierarchy of Dynamic Risk Premia [PDF] [Copy] [Kimi] [REL]

Author: Christian Oliver Ewald

The Arrow-Pratt approximation is one of the cornerstones of expected utility theory, providing the classical local approximation of certainty equivalents and risk premia in terms of absolute risk aversion. Despite its widespread use, its mathematical scope and relationship to higher-order risk preferences remain only partially understood. This paper develops a new framework for the analysis of certainty equivalents and dynamic risk premia based on Malliavin calculus and Wiener chaos analysis. We first show that the classical Arrow-Pratt approximation is not asymptotically valid for arbitrary sequences of vanishing risks, thereby identifying precise limitations of the traditional theory. Motivated by this observation, we formulate certainty equivalents dynamically by considering the progressive revelation of uncertainty through a Brownian filtration. Combining Itô calculus, the Clark--Ocone representation and the Wiener chaos decomposition, we derive a complete hierarchy of higher-order dynamic risk premia and obtain explicit representations of the corresponding coefficients in terms of Malliavin derivatives. For mixed Wiener chaos expansions, higher-order preference measures, including prudence and temperance, emerge naturally through interactions between chaos components and are characterised using Bell polynomial representations. Explicit results for quadratic Gaussian functionals and the Vasicek interest-rate model illustrate the theory and identify a broad class of regular Wiener functionals for which the classical Arrow-Pratt approximation is recovered as the leading-order term. The results establish a unified framework linking expected utility theory, stochastic analysis and Wiener chaos expansions, opening a new perspective on higher-order certainty equivalents and the dynamic measurement of risk.

Subjects: Mathematical Finance , General Economics

Publish: 2026-07-25 11:34:27 UTC