10. What CONTRADICTS the above

5.8k 字 · 源文件 darlin-consciousness-literature-report.md 第 1185 行起 · other

Explicit counter-evidence and live disputes, so the parent agent does not over-trust any single line:

  1. Against the value of the indicator approach itself. Koch, arXiv 2603.27597: indicator-based assessments cannot be calibrated (no labelled AI consciousness outcomes exist) and their transfer from biology is unsupported. → The whole Tier-2 framework rests on an unvalidated bridge. This is the strongest self-undermining result in the literature and I recommend the parent treat it as the governing caveat.

  2. Against GWT's bottleneck being good. Phua, arXiv 2512.19155: "GWT-style broadcasting amplifies internal noise, creating extreme fragility." A workspace can be a liability, not an asset. (Contrast Goyal et al. arXiv 2103.01197, who argue capacity limits have a rational basis.)

  3. Against self-monitoring being useful at all. Xie, arXiv 2604.11914: three self-monitoring modules (metacognition, self-prediction, subjective duration) gave *"no statistically significant benefit across 20 random seeds"*; even after structural integration the comparison against a no-self-monitoring baseline was non-significant (d = 0.15, p = 0.67), and a parameter-matched control performed comparably. → The burden of proof is on anyone adding self-monitoring to Darlin.

  4. Against the Bayesian/inference interpretation of the FEP. Biehl, Pollock & Kanai, arXiv 2001.06408: the free energy lemma "when taken at face value, is wrong"; and when it does hold it "implies equality of variational density and ergodic conditional density", which they say makes the inference interpretation unjustified.

  5. Against "active inference is a distinct theory of behaviour". Millidge, arXiv 1907.03876: deep active inference *"shows similarities with maximum entropy reinforcement learning and the policy gradients algorithm."* Mazzaglia et al., arXiv 2110.10083: active inference "closely matches" reward-engineered RL rather than beating it. → The "active inference" label may not buy you anything over MaxEnt RL.

  6. Against the assumption that good self-reports imply internal access. Zeng et al., arXiv 2608.30980: *"improved self-modeling may not arise from privileged access to the model's internal decision process."* Also Šekrst, arXiv 2608.18816: self-reports of sentience "come within the definition of hallucination", and consciousness, if it occurred, *"might remain epistemically inaccessible since it would be indistinguishable from a sufficiently advanced hallucination."*

  7. Against using confidence/calibration metrics as metacognition. Cacioli, arXiv 2603.25112: ECE and Brier "conflate two capacities"; and in v1/v2 of that same paper an inverse accuracy–efficiency coupling "does not survive relabelling" after a scorer-bias correction — i.e. the metric can reverse under a scoring change. Dai & Wang, arXiv 2603.09309: the confidence scale changes the measured metacognition. → Any Darlin metacognition result is fragile to scoring and scale choices. Pre-register both.

  8. Against "high Φ ⇒ awareness" from a different angle. Aaronson's own concession is the sharpest philosophical point in that post and it cuts against naive functionalism too: *"you can't say both that the Hard Problem is meaningless, and that progress in neuroscience will soon solve the problem if it hasn't already."* If you dismiss the hard problem as meaningless (the illusionist move), you cannot then claim your mechanism explains consciousness.

  9. Against the "should be abandoned" position. The counterweight to Schwitzgebel and Garrido-Merchán is that the functional-indicator programme has produced real, replicated, mechanistically specific results (AST's parameter-matched benefit, GWT's data-efficiency wins, the meta-d' dissociation). The disagreement is not "is there anything to measure" but "does any of it license a consciousness claim" — and the honest answer from the sources is: no, not yet, and possibly not ever, but the mechanisms are real and testable.

  10. Against Searle-style substrate arguments. Klatzmann & Doerig, arXiv 2606.02121: *"Biology can act as a guide on this quest, but not as a solution."* And the SEP's synron thought experiment (chinese-room) leaves the substrate question formally open: "will Otto [notice]? If so, when? And why?"

  11. Against the assumption that the 2023 pseudoscience letter settled anything. Frankish's reply (doi 10.31234/osf.io/uscwt) agrees the accusation "has merit" but reframes IIT as metaphysics rather than pseudoscience — and Frankish is an illusionist who thinks phenomenal realism itself is the mistake. So there is no consensus even about what kind of thing IIT is.

  12. On the IIT side, unresolved. Tsuchiya, Andrillon & Haun's reply to the unfolding argument (doi 10.1016/j.concog.2020.102877) exists, and its title claims the unfolding argument presupposes functionalism/behaviourism. I could not retrieve its abstract, so I cannot fairly evaluate it. Do not treat the unfolding argument as unopposed.


← 返回《darlin-consciousness-literature-report.md》目录