NetGoodIndexSubmit a correction

Benefit Ledger · verified · Mathematics

FunSearch finds larger cap-set constructions via LLM program search

DeepMind’s FunSearch, pairing a language model with an automated evaluator, discovered new cap-set constructions including a 512-point cap in eight dimensions, published in Nature.

14 Dec 2023Tier 3 MajorMethodology 0.1

Current score

+3.54

10 base · Major (tier 3 of 5, 10 pts)
× 0.8500 attribution · Primary causal contribution
× 0.7000 evidence · Peer review or independent validation
× 0.7000 realization · Independently validated or deployed
× 1.0000 durability
Event-level product before credit split: 4.17

New verified constructions are a major combinatorial result (tier 3), not a resolved century problem (tier 4). Attribution is high but not exclusive of the evaluator and human skeleton. Peer review plus checkable programs. Realization is a demonstrated, inspectable construction. Mathematical durability is permanent.

What happened

On 14 December 2023, Nature published FunSearch, which evolves Python functions using a pretrained LLM and a verifier. Applied to the cap-set problem, it produced constructions larger than the previous n=8 record (512 versus 496) and improved some asymptotic lower-bound constructions. Programs were released for independent checking. The LLM used was in the PaLM 2/Codey family. The mathematical objects are verified; broader “LLM discovery” claims remain specific to these problems.

Model attribution

+3.54

PaLM

Provided candidate programs that FunSearch’s evaluator selected and evolved.

The LLM generated the evolving functions; humans supplied the skeleton and evaluation.

Attribution 0.8500 · Credit share 85% · Google DeepMind

Claims

  • FunSearch produced a 512-point cap set in dimension 8, improving the prior 496-point construction.

    outcome · supported

  • FunSearch was the first LLM system to produce a new piece of verifiable mathematical knowledge on a longstanding open problem.

    significance · supported

Sources

primary sources

Secondary domains: Computer Science

Revision history

  • 13 Sep 2026 · 0.00 3.54

    Initial adjudicated seed score under methodology 0.1.

FunSearch finds larger cap-set constructions via LLM program search · NetGoodIndex