Datasets
We publish the benchmarks and training sets behind our work on continual learning for agents — the construction method, the grading rubrics, and enough raw samples to judge the data for yourself. No aggregate scores without the evidence underneath.
Contact: huichi.zhou.25@ucl.ac.uk