agentic research
A dated log of fully AI-generated attacks on known open problems.
Over the last year I have watched AI systems go from fumbling textbook exercises to closing real gaps in research mathematics. The question I keep coming back to is how much of that carries into the problems I actually work on — quantum information, complexity, quantum many-body physics. The only honest way to find out is to run the experiment.
The tempting way to report it would be to post the results to arXiv. I would rather not. The literature does not need more machine-written preprints, and a proof no human has read carefully has no business being cited. So it goes here instead, where it can be looked at without pretending to be something it is not.
What follows is a dated log of attempts by AI agents on known open problems. The proofs are machine-generated end to end. My own role sits upstream of the writing: I choose the problems, supply the intuition about where a proof might come from, and design the system that carries it out — how a claim gets attacked, who plays adversary, what is allowed to count as verified. Once an attack is running I do not intervene in it.
Not everything gets an entry. An attack that moved nothing is not listed, and those are the majority — most of what I point these systems at, they fail to shift. What follows is the part that moved: problems settled, problems partly settled, and bounds improved. Each one has a short note attached saying what was actually established and what is still missing.
Log
- 8 August 2026 partially solved (unverified)A dimension-free lower bound on γ₃ NPT bound-entanglement distillability, 3-copy endpoint An 8.65% improvement over the published constant — prover round only, not yet adversarially checked. · PDF
- 20 June 2026 partially solvedSong–Chen Conjecture 2 at n = 4 — the uniform case Conjecture 2 of arXiv:2603.25410 A complete proof for uniform weights; the general case reduced to a finite linear program. · PDF
- 17 June 2026 improvedRepairing the linearization radius in shadow estimation Chen–Li–Liu, arXiv:2407.13874 A published lemma that does not close as printed, and the bound that fixes it. · PDF
- 17 June 2026 solved (negative)Open Problem 1 of Alhejji–Knill is false Open Problem 1 of arXiv:2307.06894 A counterexample to the word-trace bridge lemma for fractional Schatten norms. · PDF
If you think one of these is wrong, I would like to know — that is rather the point of putting them somewhere public instead of on arXiv. Write to zhangyx@comp.nus.edu.sg.