A Mamba state-space model trained only for next-step prediction appears to recover Granger-causal structure through a simple readout $S = |W_{out} W_{in}|$, with early experiments suggesting the phenomenon generalized across architectures and benefited from interventional data at $p < 10^{-5}$. We package the protocol used to test that claim -- standardized synthetic generators (VAR/Lorenz/CauseMe-style), three intervention semantics ($do(X=c)$, soft-noise, random-forcing), edge-provenance cards on three real datasets, and size-matched control arms -- as a reusable falsification benchmark, and walk the claim through it in five stages. The method-level claim does not survive: (i) a plain linear bottleneck does as well or better; (ii) tuned Lasso beats the bottleneck on synthetic CauseMe-style benchmarks, and on Lorenz-96 (the only real benchmark with unambiguous ground truth) classical PCMCI and Granger lead a tight cluster in which the bottleneck trails; (iii) the headline intervention advantage is roughly 60% a sample-size confound, and the residual disappears under standard $do(X=c)$ interventions, surviving only under a non-standard random-forcing scheme; (iv) even that residual reproduces, with a larger effect, in classical bivariate Granger -- the effect is method-agnostic. What survives is a narrow characterization result; the benchmark is the lasting artifact, and each stage above is one of its control arms.


翻译:仅针对下一步预测训练的Mamba状态空间模型,似乎能通过简单读取操作$S=|W_{out}W_{in}|$恢复格兰杰因果结构。早期实验表明,该现象可跨架构泛化,并在$p<10^{-5}$的干预数据中受益。我们将验证该主张的标准化协议体系——包括标准合成数据生成器(VAR/Lorenz/CauseMe风格)、三种干预语义($do(X=c)$、软噪声、随机强迫)、三个真实数据集上的边缘溯源卡片以及规模匹配的对照组——打包为可复用的证伪基准,分五个阶段对该主张进行验证。方法层面的主张未能成立:(i)普通线性瓶颈模型表现相当或更优;(ii)调优的Lasso在合成CauseMe风格基准和Lorenz-96系统(唯一具有明确真实值的真实基准)上均优于瓶颈模型,而经典PCMCI和格兰杰检验在此处形成紧密聚类,瓶颈模型落后于该聚类;(iii)标题所示的干预优势中约60%源于样本量混杂,剩余部分在标准$do(X=c)$干预下消失,仅存在非标准随机强迫方案中;(iv)该剩余效应在经典双变量格兰杰检验中复现且效应更大——表明该效应与具体方法无关。最终保留的仅为狭隘的表征结论;本研究所建立的基准体系才是持久性成果,而上述每个阶段均构成其对照组之一。

0
下载
关闭预览

相关内容

【博士论文】因果发现与预测:方法与算法,101页pdf
专知会员服务
59+阅读 · 2023年9月24日
【匹兹堡大学博士论文】数据限制下的因果推理,147页pdf
【NeurIPS 2020】基于因果干预的小样本学习
专知会员服务
70+阅读 · 2020年10月6日
【NeurIPS2020】可处理的反事实推理的深度结构因果模型
专知会员服务
49+阅读 · 2020年9月28日
基于深度元学习的因果推断新方法
图与推荐
12+阅读 · 2020年7月21日
您可以相信模型的不确定性吗?
TensorFlow
14+阅读 · 2020年1月31日
你的算法可靠吗? 神经网络不确定性度量
专知
40+阅读 · 2019年4月27日
你真的懂时间序列预测吗?
腾讯大讲堂
104+阅读 · 2019年1月7日
相关性≠因果:概率图模型和do-calculus
论智
31+阅读 · 2018年10月29日
用模型不确定性理解模型
论智
11+阅读 · 2018年9月5日
国家自然科学基金
5+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
21+阅读 · 2013年12月31日
国家自然科学基金
26+阅读 · 2011年12月31日
VIP会员
最新内容
印度精确打击与指挥架构的断层
专知会员服务
4+阅读 · 7月20日
美空军AI完成F-16战斗机自主空战历史性试飞
专知会员服务
5+阅读 · 7月20日
深入Project Maven:为何人工智能在战场上依然失灵
锻造未来士兵:外骨骼、基因工程与赛博格
专知会员服务
7+阅读 · 7月19日
《无人机蜂群通信技术研究》50页
专知会员服务
10+阅读 · 7月19日
相关基金
国家自然科学基金
5+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
4+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
21+阅读 · 2013年12月31日
国家自然科学基金
26+阅读 · 2011年12月31日
Top
微信扫码咨询专知VIP会员