Use whenever the user wants to collect, gather, scrape, crawl, harvest, or research data from the web, research papers, arXiv, GitHub, social media, forums, news, or any online source — even if they just say find me data on X or get everything about Y. Covers APIs, scraping, headless browsers, hidden JSON APIs, feeds, academic APIs, GitHub APIs, social APIs, and large-scale crawling.