An Introduction to Web Mining
with Applications in R
Springer
ISBN 978-3-031-96637-8
Standardpreis
Bibliografische Daten
Fachbuch
Buch. Softcover
2025
14 s/w-Abbildungen, 14 Farbabbildungen.
In englischer Sprache
Umfang: xxi, 251 S.
Format (B x L): 15,5 x 23,5 cm
Verlag: Springer
ISBN: 978-3-031-96637-8
Weiterführende bibliografische Daten
Das Werk ist Teil der Reihe: Use R!
Produktbeschreibung
Through the book, readers will learn how to
- scrape static and dynamic/JavaScript-heavy websites
- use web APIs for structured data extraction from web sources
- build fault-tolerant crawlers and cloud-based scraping pipelines
- navigate CAPTCHAs, rate limits, and authentication hurdles
- integrate AI-driven tools to speed up every stage of the workflow
- apply ethical, legal, and scientific guidelines to their web mining activities
Part I explains why web data matters and leads the reader through a first “hello-scrape” in R while introducing HTML, HTTP, and CSS. Part II explores how the modern web works and shows, step by step, how to move from scraping static pages to collecting data from APIs and JavaScript-driven sites. Part III focuses on scaling up: building reliable crawlers, dealing with log-ins and CAPTCHAs, using cloud resources, and adding AI helpers. Part IV looks at ethical, legal, and research standards, offering checklists and case studies, enabling the reader to make responsible choices. Together, these parts give a clear path from small experiments to large-scale projects.
This valuable guide is written for a wide readership — from graduate students taking their first steps in data science to seasoned researchers and analysts in economics, social science, business, and public policy. It will be a lasting reference for anyone with an interest in extracting insight from the web — whether working in academia, industry, or the public sector.
Autorinnen und Autoren
Produktsicherheit
Hersteller
Springer Nature Customer Service Center GmbH
ProductSafety@springernature.com