Library of Congress Linked Data Services and Datasets
The Library of Congress provides a sophisticated infrastructure for accessing standardized bibliographic data. By transforming traditional cataloging records into linked data, the library enables researchers and developers to navigate complex relationships between subjects, names, and classifications across the web.
Key Facts
- Provides access to major authority files including LCSH and LCNAF.
- Supports multiple machine-readable formats including JSON, RDF/XML, and N-Triples.
- Utilizes MADS/RDF and SKOS (Simple Knowledge Organization System) for data presentation.
- Allows users to create custom datasets via library application profiles (APIs).
- Offers full vocabulary downloads for offline use.
Available Datasets
The Library of Congress hosts a variety of authoritative datasets that serve as the backbone for library classification worldwide. These include the Library of Congress Subject Headings (LCSH) and the Library of Congress Name Authority File (LCNAF).
Beyond these primary files, the service includes the Library of Congress Thesaurus for Graphic Materials, various preservation vocabularies, and a wide array of MARC (Machine-Readable Cataloging) codes.
The Challenge of LC Classification
While most authority files map smoothly to linked data formats, the Library of Congress Classification presented a unique technical challenge. Because LC Classification utilizes a different MARC format than the LC Authorities, mapping it to MADS/RDF (Metadata Application Schema) was more complex than the process used for LCSH or LCNAF.
[ไม่มีภาพประกอบ]Data Formats and Accessibility
To ensure maximum interoperability, the service presents data using MADS/RDF and SKOS. Additionally, the library employs its own proprietary ontology to more accurately describe the specific relationships and resources associated with its classification systems.
Users can access individual records through content negotiation, which allows the server to deliver the data in the format most suitable for the requester. Supported formats include:
- XHTML/RDFa
- RDF/XML
- N-Triples
- JSON
For those requiring bulk data, each vocabulary is available for download in its entirety. It is important to note that id.loc.gov does not currently provide a SPARQL endpoint for direct querying.
Summary of Technical Specifications
| Category | Details |
|---|---|
| Primary Datasets | LCSH, LCNAF, LC Classification, Graphic Materials Thesaurus |
| Standard Schemas | MADS/RDF, SKOS, Custom LC Ontology |
| Output Formats | JSON, RDF/XML, N-Triples, XHTML/RDFa |
| Customization | Available via library application profiles (APIs) |
Frequently Asked Questions
What are the primary datasets available through the Library of Congress?
The primary datasets include the Library of Congress Subject Headings (LCSH), the Library of Congress Name Authority File (LCNAF), the Library of Congress Classification, and the Thesaurus for Graphic Materials.
In what formats can I retrieve individual records?
Individual records are available via content negotiation in JSON, RDF/XML, N-Triples, and XHTML/RDFa.
Can I create my own custom datasets?
Yes, the Library of Congress allows users to create their own datasets using library application profiles (APIs).
Does id.loc.gov support SPARQL queries?
No, id.loc.gov does not currently provide a SPARQL endpoint.
Why was LC Classification harder to map than LCSH?
Mapping LC Classification to MADS/RDF was more difficult because it uses a different MARC format than the LC Authorities files.