Reads the dataset, extracts context, question, answer, tokenizes them, and calculates answer span in terms of token indices. Note: due to tokenization issues, and the fact that the original answer spans are given in terms of characters, some examples are discarded because we cannot g
(dataset, tier, out_dir)
source not stored for this graph (policy: none)
no test coverage detected