Skip to content

Author

Dingyan Shang

4 papers indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Preprint Aug 2026

DACRI: Decision-Aware Causal Intervention Ranking for Critical Supply Chains

Detecting or attributing a supply-chain disruption is not the same as selecting the intervention that maximizes recoverable net value. We present CriticalSCM-Bench v1, a controlled synthetic benchmark with causal ground truth, paired factual/counterfactual rollouts, and an explicit net-value objective. Relative to a full-information train-selected static benchmark, LambdaMART improves median normalized net value by 5.7--16.2\%, with paired statistical support on the semiconductor and critical-material archetypes but not on digital infrastructure. On digital infrastructure, a domain-informed constant-buffer policy remains stronger, showing that greater model complexity is not uniformly justified. Across partial and delayed settings, LambdaMART retains 33--75\% of full-clamp value. Stress tests further show that intervention fidelity, timing, cost, and held-out disruptions can alter policy ordering. Critical materials show the weakest out-of-distribution retention. Separately, a guarded explanation study over 540 generations preserves every fixed intervention decision after deterministic validation and template fallback, although exact wording remains unstable. Within this controlled setting, the results identify regimes in which adaptive ranking adds value and those in which simpler structural policies remain preferable.

Shi Zhuo Huang, Jiani He, Dingyan Shang et al. · 0 citations
Jul 2026

Context-Masked Truncated Reasoning Audits for Answer-Key Dependence in LLM Tutors

Context masking is established as necessary for attributing early answer availability to an explanation rather than its hidden input when early-prefix evidence disappears after masking.

Bo-Nan Shen, Ding-Yan Shang, You Wang et al. · 0 citations
Jul 2026

Safety, or Just Capability? A Validity Audit of Agent-Safety Benchmarks

Naming the benchmark, metric, target behavior, and model panel is the minimum a safety claim needs, and both instruments score harmful compliance, so this is evidence of convergent validity rather than general safety.

You Wang, Xiao Han, Ding-Yan Shang et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.