Aboutness Overview

Aboutness is what the things in OpenAlex are about. Most works are “about” something, and that aboutness aggregates up to characterize authors, institutions, sources, and other entities. OpenAlex offers several distinct aboutness signals, each usable as a filter; which one fits depends on your question.

Two properties help you choose. Granularity (the number of groups) sets how fine-grained an analysis can be. Familiarity is how recognizable the scheme is to others — how easily you can compare or share results.

Signal# groupsFamiliarityFit to custom areas
SDGs17HighLow
Domains4HighLow
Fields26HighLow
Subfields252HighMedium
Topics4,516LowMedium-high
Keywords~1.9 millionMediumHigh
Concepts (deprecated)~65,000HighVariable
Text search∞LowHigh
Semantic search∞LowHigh

A rough guide: the topics hierarchy (domains → fields → subfields → topics) is the supported general-purpose system — pick the level whose granularity matches your question. Keywords fit narrower, more specific slices, and find works whatever words their authors used. SDGs map research onto the UN Sustainable Development Goals and little else. Concepts are deprecated — kept for continuity with Microsoft Academic Graph, no longer maintained; see Concepts. Text search fits custom areas no scheme covers, at the cost of comparability.

Not a subject signal, but close by: study designs say how the research inside a work was done (randomized controlled trial, observational study, systematic review and four more), not what it is about. They combine well with any of the signals above, e.g. every randomized controlled trial in the topic Cancer Immunotherapy and Biomarkers: filter=topics.id:T10158,study_designs.id:randomized-controlled-trial.

Aboutness doesn’t have to go through a labeling scheme at all. OpenAlex embeds the title and abstract of every work as a vector using an open-source embedding model, so works that are about similar things sit near each other in vector space — no categories required. Semantic search queries these embeddings directly: describe what you’re looking for in plain language (a sentence, or even a whole abstract or grant description) and get back the works closest in meaning, even when the wording differs. It’s the best fit when your research area doesn’t line up with any predefined scheme and keyword matching is too brittle. Embeddings also help build the keyword vocabulary: they surface candidate synonyms, which a judge then confirms or rejects before they are merged.

Aboutness for your own text

For the topics hierarchy and keywords, you can supply your own custom text — the title and abstract of an unpublished article, say, or a grant proposal — and get back topics (with their subfield, field, and domain) and keywords in exactly the form OpenAlex assigns them to works. See the text aboutness endpoint.

  • Topics — the four-level hierarchy and how it’s assigned
  • Keywords — how keyword tagging works
  • SDGs — the Sustainable Development Goals tagger
  • Study designs: how the research inside a work was done
  • FWCI — the field-normalized citation metric built on subfields

Last updated

View as Markdown