Extract CTDL course data from a live catalog
Deterministic extraction with source verification, the approach we propose for scaling CTDL xTRA.
Paste a catalog page to begin
Point this at a college's course listing page and it will do what we propose CTDL xTRA should do on every source.
- Identify the platform. Most US college catalogs run on a handful of systems, so a per-platform parser handles the bulk of them.
- Extract the courses. Code, title, credits, description and prerequisites, mapped to CTDL Course fields.
- Verify every value against the source. Anything that cannot be found in the page text is sent to human review instead of being published.
- Emit a bulk-upload CSV ready for the Credential Registry.