The Noosphere corpus
Noosphere does not search the internet: it answers from a document base of its own. This page states exactly what that base contains, where each part came from and how it was built, so anyone can judge the worth of an answer before trusting it.
Live figures, read from the graph itself when the page loads. The breakdown below is as of 31 July 2026.
Where it comes from
No single source dominates, and that is deliberate. Biomedical literature brings experimental rigour; the project's own archive brings what rarely reaches an indexed journal — harm-reduction reports, conference proceedings, dissertations, specialist press; OpenAlex brings the fields a biomedical search never reaches, such as anthropology or law.
| Origin | Documents | What it contributes |
|---|---|---|
| PubMed / Europe PMC | 3.973 | Peer-reviewed biomedical literature |
| Own library | 2.017 | Reference books and monographs |
| Document archive | 1.268 | Reports, proceedings, theses and grey literature |
| OpenAlex | 753 | Social sciences, botany, law |
| Cannabis Magazine | 700 | Own back catalogue, 1997 onwards |
| Open and reference web | 80 | Public bodies and organisations |
Scope
How it is built
Automated harvesters query PubMed, Europe PMC and OpenAlex every week across forty-four thematic axes, from receptor pharmacology to botanical taxonomy and drug policy. What they find is filtered for relevance, dropped if already held, and anything new is split into passages, indexed by meaning and by word, and analysed to extract the concepts it mentions and the relations between them. That graph is what lets a question about a receptor also surface what has been written about the plant that activates it.
Answering is not a single search: the question is rewritten, retrieval runs along two routes — meaning and literal match — the results are fused, and a local reranking model decides which passages earn a place in the answer. It is slower than a search engine, and it is why the citations match what is claimed.
What you can and cannot read here
The index, the bibliographic metadata and the citations are freely consultable. Noosphere does not republish the full text of third-party works: when an answer rests on a document, it links to the original so it can be read at source. Full text is kept only when the work is open access or belongs to the project itself.
How to cite it
If you use Noosphere in your work, cite the corpus rather than the answer: answers are generated on the spot and are not reproducible word for word. The sources they cite are, and those are what belongs in an academic citation.
Psiconáutica (2026). Corpus documental multidisciplinar sobre sustancias psicoactivas (Noosphere). https://brain.psiconautica.org/corpus.html
Contributing documents
The corpus grows mainly by contribution. If you hold relevant literature — theses, agency reports, conference proceedings, harm-reduction material, back catalogue — write to noosphere@psiconautica.org and we will review it before taking it in. We are especially interested in what is indexed nowhere and gets lost: the grey, the local, and what was not published in English.
Ask it something you already know
It is the only honest way to measure a corpus: ask a question whose answer you already know, and check where it says it got it.