TODO: - Understand types of questions asked (computation, visual understanding, etc.) -- try to cluster them into question templates - Extract shared computations that are common among multiple questions (i.e., index candidates) - Statistics about retrieval count (number of documents in Oracle, etc.) - Anything you might notice as you look through the dataset
TODO: