CLIO Library Catalog Open Data

Columbia University Libraries publishes its full catalogue — bibliographic and holdings records from the integrated library system behind CLIO — as gzipped MARCXML bulk extracts under a CC0 1.0 Public Domain Dedication, refreshed monthly, alongside a deletes file. 108 files were present in the extract directory at probe time, served from an open Apache directory index with no credential and no rate limit. Covers books, serials, music, video and manuscripts; excludes Law Library and ReCAP partner records. There is no manifest, no checksum, no change feed and no harvesting protocol — a consumer diffs the directory listing.

Work with this as data

Every API here is available over the APIs.io API and to AI agents over MCP.

MCP server

One button, every client — Claude, Cursor, VS Code and the rest.

https://apis.io/mcp

Tools for apis

7 MCP tools reach this
  • find_apisBrowse and filter every API in the catalog.
  • get_api_artifactsOne API's artifacts, grouped by type.
  • get_openapiThe primary OpenAPI for this API.
  • find_similar_apisAPIs that look like this one.
  • apis_io_searchSTART HERE — APIs, providers and tags for one query, each with its total.
  • resolveTurn a domain, URL or GitHub org into the provider it belongs to.
  • find_cohortsEvery scored population of providers in the catalog.
All 92 tools →

Call it yourself

curl for this page
This API
curl "https://apis.io/api/v1/apis/clio-opendata"
All apis
curl "https://apis.io/api/v1/apis?limit=25"

Discovery needs no key. Ratings and market analysis are Pro.

Get an API key

Free tier, no form to fill in. Signing in shares your email address with us — we store it to create your key and to recognise you if you sign in with another provider. See our Privacy Policy and Terms.

A second provider on the same verified email joins the account you already have.

API entry from apis.yml

apis.yml Raw ↑
aid: columbia:clio-opendata
name: CLIO Library Catalog Open Data
description: Columbia University Libraries publishes its full catalogue — bibliographic and holdings records
  from the integrated library system behind CLIO — as gzipped MARCXML bulk extracts under a CC0 1.0 Public
  Domain Dedication, refreshed monthly, alongside a deletes file. 108 files were present in the extract
  directory at probe time, served from an open Apache directory index with no credential and no rate limit.
  Covers books, serials, music, video and manuscripts; excludes Law Library and ReCAP partner records.
  There is no manifest, no checksum, no change feed and no harvesting protocol — a consumer diffs the
  directory listing.
humanURL: https://library.columbia.edu/bts/clio-data.html
baseURL: https://lito.cul.columbia.edu/extracts/ColumbiaLibraryCatalog/full/
tags:
- Library
- Catalog
- MARCXML
- Open Data
- Bulk Download
- CC0
properties:
- type: Documentation
  url: https://library.columbia.edu/bts/clio-data.html
- type: GitHubRepository
  url: https://github.com/cul/clio-spectrum
- type: Lifecycle
  url: lifecycle/columbia-lifecycle.yml
x-operator: institution
x-operator-evidence: lito.cul.columbia.edu resolves to lito-apache-prod2.cul.columbia.edu / 128.59.222.64,
  in Columbia University Libraries' own address space. The extracts are generated from Columbia's own
  ILS and released under Columbia's own CC0 dedication.
x-access: public
x-status: 200, 108 extract files listed, probed 2026-08-19