What this is
100 rows drawn from the paid product, with the same 24 columns and the same meanings. It exists so you can judge the classification quality and the shape of the long tail before spending anything.
How the 100 were chosen
A systematic spread across the full demand ranking, topped up so every value of every classification appears at least once. Facet representatives are picked pseudo-randomly within each facet rather than by highest volume - picking the top member of each rare category would have pulled the sample to 61% lowest-band against the product's true 84%, making the dataset look denser than it is. The selection is deterministic: this sample does not change between requests.
What it is honest about
- 80% of these rows are in the 10-30 monthly-searches band, against 84% in the full product. Google's reporting floor is 10, so the long tail is not further resolvable. The value of the full dataset is coverage and classification, not volume precision.
- Around 30% of terms are insufficient_data for trend - too few active months to make a claim. That is visible here too.
- clinical_focus is general for about two thirds of terms. Most search terms name no specific aspect of the disease.
Column meanings
Identical to the full product - see the paid listing for the column table, the volume-banding method and the trend method.
The full product
26,250 terms, same schema, one request, CSV export enabled.
