The CAFE collection at a glance
The BUSPH–HSPH CAFE Research Coordinating Center collection on Harvard Dataverse, summarised across its dataset-level metadata. Every figure on this page excludes unpublished drafts.
Where the datasets live
Datasets reach the CAFE collection three ways, and all three count as part of it: they are owned by a collection in the CAFE tree, linked in individually, or sit in a whole collection that is linked into CAFE.
Datasets per subcollection
Top 18 of — subcollections holding at least one published dataset.
Deposited locally or harvested
Harvested datasets are indexed from another repository. They carry no local publication date, which is why several panels on this site have a smaller denominator than the collection.
Route into the CAFE collection
A dataset can be owned by a CAFE collection, linked in on its own, or carried in as part of a whole collection linked into CAFE.
Subjects
Datasets per subject
The N/A placeholder is excluded from this chart and reported beside it.
Deposits per year
Shape of the collection
Size
Files
Keywords
Authors
Per-dataset summary statistics
Most-used keywords and most prolific authors
Both are parsed rather than taken as stored. Keywords are case-folded and whitespace-collapsed; descriptor labels such as Federal Agency are removed from author lists before counting. The Explorer shows both under any filter.