AOL4FOLTRA

Creators
Publication date 18-06-2025
Description
AOL4FOLTR is the first learning-to-rank (LTR) dataset designed specifically for evaluating federated online learning-to-rank (FOLTR) algorithms. Including user identifiers and timestamps, this dataset allows for the simulation of real user behavior with heterogeneous data and in asynchronous federated learning settings. The dataset consists of two files letor.txt.gz (55G uncompressed) metadata.csv letor.txt contains the query-document pairs for all query logs in standard LETOR format. Each query-document pair holds a binary label derived from user clicks, and is further represented by a 103-dimensional vector. We document the features in our code repository. The query logs are cross-referenced (by qid) in metadata.csv, where contextual information is provided. This includes the user, timestamp, raw query, the target document ID, and a list of 20 candidate documents. The document IDs and user IDs directly map to the AOL-IA dataset; the query IDs do not. For access to the raw document contents, please refer to this dataset.
Publisher Zenodo
Organisations
  • Faculty of Science (FNWI) - Informatics Institute (IVI)
Document type Dataset
DOI https://doi.org/10.5281/zenodo.15678397
Permalink to this page
Back