Institution profile

Didi Chuxing

Industry researchasia · cn
Official website
Research library92linked papers
Opportunities0open roles
Selected work

Representative Papers

Treatment Allocation under Uncertain Costs

Mar 20, 2021

This paper addresses the problem of optimal treatment allocation under a budget constraint when treatment costs vary heterogeneously with covariates. We propose a threshold rule based on a priority score and establish, for the first time, a theoretical link between optimal allocation under uncertain costs and instrumental variable (IV) estimation of heterogeneous treatment effects. We rigorously derive the optimal threshold structure and prove its learnability. Our method integrates randomized controlled trial data, priority score modeling, threshold-based decision making, and an IV estimation framework. Empirically, the approach significantly outperforms standard benchmarks across multiple evaluation metrics, achieving maximal social value or firm profit within budget constraints. It provides a new paradigm—statistically rigorous yet practically implementable—for applications including scarce healthcare resource allocation and dynamic pricing.

14 citations1 influentialRead paper

Latent Chain-of-Thought as Planning: Decoupling Reasoning from Verbalization

Jan 29, 2026

This work addresses the limitations of traditional chain-of-thought methods, which incur high computational costs in discrete token spaces and often converge to a single reasoning path, as well as existing implicit reasoning approaches that rely on fixed step counts and lack dynamic termination mechanisms. The authors reformulate implicit reasoning as a planning process with adaptive termination by decoupling reasoning from language generation. Reasoning occurs in a continuous latent space via deterministic state trajectories, while a separate decoder produces textual outputs on demand. This approach is the first to enable adaptive reasoning length, substantially enhancing reasoning diversity and scalability. Experimental results show that, although greedy accuracy is slightly lower, the model explores a broader solution space, providing a transparent and flexible foundation for search during inference.

1 citationsRead paper
Recent publications

Latest Papers