Daniel Franzen & Jan Disselhoff - ARC Prize 2024 winners

Topics covered
Popular Clips
Questions from this episode
- Asked by 34 people
- Asked by 28 people
- Asked by 24 people
- Asked by 22 people
- Asked by 12 people
Episode Highlights
Solution Design
and , winners of the ARC Prize 2024, shared their innovative approach to solving ARC tasks using large language models (LLMs). They began with a 12 billion parameter LLM, tokenizing tasks into text and feeding them directly into the model without pre-processing. This approach, combined with test-time fine-tuning, significantly improved their model's performance. explained, "We do another training process on the examples of the validation set, which we get during inference without the final challenge, which we then try to predict with the model, and that gives a big improvement to the score." 1 2.
Refinement
The iterative refinement of their approach was crucial, involving constant testing and adaptation. noted that many ideas initially seemed promising but didn't pan out, leading them to focus on strategies that reliably improved performance. The depth-first search algorithm and scoring process were pivotal, enhancing their model's capability to generate and select solutions effectively. remarked, "The DFS was a very natural extension of the scoring, I think after we had a good working scoring and selection algorithm, the challenge became to generate good candidates." 3 4.
Augmentation Techniques
Data augmentation played a critical role in enhancing model accuracy and generalization. They employed symmetry augmentations and test-time training to maximize data utility, given the limited examples per challenge. explained, "We heavily use augmentation in basically all our inference steps. Also in training. We also use it for the pre training in rehab, actually. But it helps a lot to get enough data during the retraining or the test time training." 5 6.
Related Episodes


Francois Chollet - ARC reflections - NeurIPS 2024
Answers 383 questions

New 50% ARC result and current winners interviewed
Answers 383 questions

How Do AI Models Actually Think? - Laura Ruis
Answers 383 questions
It's Not About Scale, It's About Abstraction - Francois Chollet
Answers 383 questions

Ryan Greenblatt - Solving ARC with GPT4o
Answers 383 questions

#114 - Secrets of Deep Reinforcement Learning (Minqi Jiang)
Answers 383 questions

Robert Lange on NN Pruning and Collective Intelligence
Answers 383 questions

Pattern Recognition vs True Intelligence - Francois Chollet
Answers 383 questions

Subbarao Kambhampati - Do o1 models search?
Answers 383 questions

Decompiling Dreams: A New Approach to ARC? - Alessandro Palmarini
Answers 383 questions

Nicholas Carlini (Google DeepMind)
Answers 383 questions

Dr. Paul Lessard - Categorical/Structured Deep Learning
Answers 383 questions

Jürgen Schmidhuber - Neural and Non-Neural AI, Reasoning, Transformers, and LSTMs
Answers 383 questions

Sepp Hochreiter - LSTM: The Comeback Story?
Answers 383 questions

#106 - Prof. KARL FRISTON 3.0 - Collective Intelligence [Special Edition]
Answers 383 questions
