
The Scalability dataset is part of the Scalability Workload of the Geographica 2 GeoSPARQL benchmark which aims at discovering the limits of spatial RDF stores as the number of triples in the dataset increase. It data is sourced from two real-world datasets that are synthesized in a specific manner to achieve the desired goals. The OSM data concern the following list of countries: Wales, Scotland, Greece, Northern Ireland, England and Germany. The feature classes selected are: buildings, landuse, natural, places, points of interest, railways, roads, traffic, transport, water and waterways. The CLC-2012 is the 2012 version of the CLC dataset and its data covers the 33 European Environment Agency member countries and six cooperating countries. The logic and steps taken are described in the original paper "Evaluating Geospatial RDF stores Using the Benchmark Geographica 2". The actual distribution format of the dataset is the form of a compressed reference dataset file (500M triples, 48M features of which 24M polygons) and an accompanying script which allows for the dynamic generation of 6 increacingly larger (10K, 100K, 1M, 10M, 100M, 500M) proper subsets of the reference dataset.
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 0 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Average | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Average |
