Open catalogs Lattuce connects to — scholarly search, OA links, and books
Inventory of open catalogs Lattuce connects for literature search, OA PDF resolve, public-domain books, and Learner corpora — with links to each provider.
Lattuce does not host the world’s scholarly corpus. It connects to open catalogs so you can discover papers, resolve legal open-access PDFs, and pull public-domain books into your library — then cite them while you write.
A shorter attribution list also lives on the public welcome catalogs strip. This page is the fuller inventory of what is wired today.

For the end-to-end review workflow, see Literature review with Lattuce.
What “connected” means
| Mode | What Lattuce does |
|---|---|
| Scholarly search | Literature Search / Library Discover / missions query the catalog and return ranked hits you can import or attach |
| Resolve & OA | Given a DOI (or landing URL), look up metadata and the best legal open-access PDF link |
| Books | Discover public-domain titles (Gutenberg, Open Library, Internet Archive) for library or Learner reading |
| Language learning | Tatoeba (dump or live API), Wikipedia, Wikisource, Wiktionary, Leipzig sentence samples for Learner |
Operators can add more catalogs to Literature Search when they are ready. These are not partnerships or endorsements — third-party marks and names belong to their owners.
Scholarly search
| Catalog | Role in Lattuce |
|---|---|
| arXiv | Preprint search and PDF download |
| OpenAlex | Works metadata, topics, related literature |
| Semantic Scholar | Papers, citation graph, related explore |
| OSF Preprints | SocArXiv, PsyArXiv, and other OSF hosts |
| Europe PMC | Life-sciences literature |
| CORE | Open-access full text |
| Crossref | Bibliographic search and DOI metadata |
| Zenodo | CERN-hosted papers, datasets, software |
| bioRxiv / medRxiv | Biology and health preprints |
| DOAJ | Directory of open-access journals |
| OpenAIRE | EU research graph |
| OAPEN | Open-access academic monographs |
| DOAB | OA book discovery directory |
| DPLA | US cultural heritage (requires API key) |
| Europeana | European cultural heritage (requires API key) |
| HathiTrust | Bibliographic ISBN/title lookup |
| Chronicling America | US historical newspaper pages |
Resolve & open access
Books
| Catalog | Role in Lattuce |
|---|---|
| Project Gutenberg | Public-domain books via Gutendex / ephemeral Learner read |
| Open Library | Book search with Internet Archive links |
| Internet Archive | Text search and download landing pages |
Language learning (Learner)
| Catalog | Role in Lattuce |
|---|---|
| Tatoeba | Example sentences (CSV dump or live API) |
| Wikipedia | Articles to read |
| Wikisource | Free-text excerpts |
| Wiktionary | Dictionary extracts |
| Leipzig Wortschatz | Corpus sentence samples |
Optional harvest tools
Operators can seed corpora from open OAI-PMH endpoints and OPDS feeds (for example Standard Ebooks new releases). That path is for hosting teams — not everyday Literature Search. Skip retired OPDS hosts such as Feedbooks (shut down 2024).
Related
- Literature review with Lattuce
- Research signals and feeds
- In-app help: Literature review with real references