Time:
9am–noon, 1:30pm–4pm
Live notes:
Etherpad
Topic:
Today, you will be working in teams on a project of your own. Your goal is to extract useful data from a website, either via API querying (if the option is available) or via web scraping (if your site has no API).
Try to pair up with one or more people having a similar interest.
Of course, we will be with you all day to help!
It would be fantastic if you could chose a website that contains data useful for your research or that is of interest to you. If you don’t have any idea however, here is a list of sites that have well-built APIs:
- Hathi Trust Hathifiles
- Hathi Trust bibliographic API
- New York Times
- Internet Archive
- Spotify
- Getty Museum
- Library of Congress
- World Bank
- NASA
- Smithsonian Institution Open Access API - requires an API key (needs registration)
