Build a repeatable Coursera course-inventory extractor
Project brief
Need a small tool to collect inventory data from about 50 Coursera course URLs and put it into a usable table. This is meant to be run again every quarter, not just a one-off copy-and-paste job.
We can agree the public URLs and the exact data columns before work starts. Please make the run repeatable, keep each row tied to its source URL, and flag pages that could not be read rather than silently leaving gaps. The result should be straightforward for someone else to run at the next quarterly update. Work only with publicly accessible pages and within the site's permitted access rules.
Deliverables & acceptance
What you'll deliver
- A runnable extraction tool and brief instructions for rerunning it on an agreed list of about 50 public Coursera URLs
- One structured table of the agreed course-inventory fields, with source URLs and a list of any failed or missing records
What the result must meet
- The tool can be rerun with the agreed URL list without manually copying course data
- The table contains one traceable row per successfully processed URL and the agreed data columns are populated or explicitly marked unavailable
- Failures are reported clearly, and the workflow does not bypass login, paywalls, rate limits or access controls