Day 1 — Fetching and parsing
- A refresher on HTTP: methods, headers, status codes
- Requests and httpx
- Encodings, or why your accents turned into gibberish
- Parsing HTML with BeautifulSoup, lxml and selectolax
- Finding your way with CSS selectors and XPath
- Forms, sessions and cookies
- The API hidden behind the page: reading the network tab
- Sites full of JavaScript: driving a browser with Playwright
- Storing the results in CSV and SQLite