Repository navigation
Queries and Datasets (from old wiki) #3980
Closed
chenlica
started this conversation in
archived-wiki
Replies: 0 comments
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Uh oh!
There was an error while loading. Please reload this page.
From the page https://github.com/apache/texera/wiki/Queries-and-Datasets (may be dangling)
====
Datasets
1. A snippet of the Twitter dataset.
Each tweet is stored in Json format. To friendly visualize Json format, we suggest some online Json viewer, such as JsonViewer.
2. A snippet of the COCO dataset.
3. A snippet of the UCF101 dataset.
Queries
1. Ten queries on the Twitter dataset.
To study the behavior of CORE with different numbers of predicates, we randomly select five queries with two strong correlated predicates and five queries with three strong correlated predicates.
2. Ten queries on the COCO dataset.
To study the behavior of CORE with different orders of predicates, we randomly select four pairs of queries. Each pair of queries contains two queries with different orders, such as q2 and q3.
3. Ten queries on the UCF101 dataset.
For the UCF101dataset, we randomly select ten queries with strong correlations.
All reactions