Getting Started with Web Scraping in C#
Learn a responsible C# scraping workflow: fetch pages with HttpClient, parse HTML with AngleSharp, and use Playwright only when browser execution is needed.
Blog
How to take clean screenshots, turn HTML into images and PDFs, run headless browsers, and give AI agents eyes. Written by the team that builds ScreenshotNeo.
Learn a responsible C# scraping workflow: fetch pages with HttpClient, parse HTML with AngleSharp, and use Playwright only when browser execution is needed.
Capture fetch and XHR traffic in Puppeteer with page events, wait for a specific response, and troubleshoot failures without unnecessary interception.
Learn why X link cards lose headlines, how to fix missing previews, and what to do when headline display is controlled by X.
Build a website chat interface with a server-side LLM endpoint, streamed replies, practical security controls, and clear privacy choices.
A compliance-first guide to collecting public beIN Sports metadata with robots.txt, Python, cURL and Node.js—plus safer page capture options.
Build a maintainable design-editor library with foundations, components, variants, documentation, publishing workflows, governance, and practical examples.
Learn how to identify a website’s DNS provider, registrar, CDN and likely origin host with dig, WHOIS, IP lookups and practical checks.
Use fit when every pixel matters and fill when the frame matters. Learn CSS contain, cover and fill, responsive patterns, cropping control and API workflows.
Build a production-ready MCP client: choose transports, negotiate protocol eras, discover tools, route model calls, secure data, and clean up reliably.
Convert scraped pages into a valid RSS 2.0 feed with stable IDs, safe XML, validation, scheduling, and reliable publishing.
Learn how to benchmark remote browsers fairly with lifecycle latency, failure rates, concurrency tests, and reproducible code.
Pass spider-owned values between Scrapy callbacks with cb_kwargs. Learn when to use meta or spider.state, how to handle errbacks and JOBDIR, and how to troubleshoot common mistakes.
Learn how to bulk export consistent screenshots with Playwright, deterministic filenames, full-page capture, batching, retries, and a hosted API option.
Set Axios headers per request, per client, or with an interceptor. Learn precedence, CORS, FormData, token safety, and common fixes.
Use Task.WhenAll for a small batch of HTTP calls and Parallel.ForEachAsync for a collection with bounded concurrency. Learn how to reuse HttpClient and handle failures safely.
Build a Make scenario that requests a public page, checks its HTML, extracts fields, and routes the results—with guidance for JavaScript-rendered sites.
Use an authorized login flow, save Playwright state securely, reuse it for protected pages, and handle cookies, storage, expiry, and errors.
Choose Playwright for cross-browser testing, more language options, or a built-in test runner. Choose Puppeteer when its Node.js workflow and browser coverage fit your project.