MaskWright
All notes

Scraping

Public pages, official feeds, and a slow hand. A coherent fingerprint is not permission.

What a scraping browser is 1

2026-08-13 · Morgan Ellis · Scraping

What a scraping browser is

What a scraping browser is. Best-of lives under antidetect-for-scraping. This definition keeps a headed collector from being sold as an operator profile.

Web scraping with Python, without a stealth stack 1

2026-08-10 · Morgan Ellis · Scraping

Web scraping with Python, without a stealth stack

Overview. Official libraries and APIs. No stealth stack. If you collect with Python, start from the documented door, not a random Chrome costume.

How to judge a web scraping tool 1

2026-08-07 · Morgan Ellis · Scraping

How to judge a web scraping tool

2023-2025 lists get refreshed as criteria, not a 15-tool clone. Judge permission, pace, and where the session lives, then pick a tool you can defend.

We will not hide a scraper 1

2026-08-05 · Morgan Ellis · Scraping

We will not hide a scraper

We will not hide a scraper. Undetectable is a sales word. If the site forbade the collection, a coherent fingerprint is not permission. Authorized work only.

User-agent honesty in research 1

2026-08-01 · Morgan Ellis · Scraping

User-agent honesty in research

Honest UA for research clients. Not a random Chrome string. If you collect, say what you are. A costume UA is not research hygiene. Local Windows notes only.

Storing research files 1

2026-07-29 · Jordan Hale · Scraping

Storing research files

Where research files live. Adjacent hygiene. Keep dumps off the work rooms that hold logins, and treat the folder as data you are responsible for.

Scraping ethics we follow 1

2026-07-27 · Morgan Ellis · Scraping

Scraping ethics we follow

Ethics they underwrite because they sell bypass. Our rules: public pages, official feeds, and a stop when the site forbids the method. Local Windows notes only.

Scraping behind a login is not research 1

2026-07-23 · Morgan Ellis · Scraping

Scraping behind a login is not research

Behind-login collection is not research. If you needed a password, you are in a different legal and ethical job than a public-page note. Authorized work only.

robots.txt and terms come first 1

2026-07-20 · Morgan Ellis · Scraping

robots.txt and terms come first

robots.txt and terms come first. No crawl page on their side treats this as the rule. This one does, before any talk of headed collection. Authorized work only.

Rate limits are not a puzzle 1

2026-07-17 · Morgan Ellis · Scraping

Rate limits are not a puzzle

Rate limits are a contract, not a puzzle. This page refuses the sport of squeezing one more request out of a door that already told you to wait.

Public pages, slowly 1

2026-07-15 · Morgan Ellis · Scraping

Public pages, slowly

Public pages, slowly. They optimize for volume. We optimize for staying inside what the site already offers the public, at a human pace. Authorized work only.

Playwright for pages you own 1

2026-07-08 · Morgan Ellis · Scraping

Playwright for pages you own

Playwright on properties you own. Product does not ship Playwright. This how-to keeps official automation on your own sites, outside MaskWright.

Personal data and collection 1

2026-07-06 · Jordan Hale · Scraping

Personal data and collection

Personal data collection is refused. They scrape emails from Shopify stores. We will not. This page is the privacy line for the scraping desk.

Proxies for web scraping 1

2026-07-06 · Priya Nair · Scraping

Proxies for web scraping

When a proxy is enough, and when a headed profile is the wrong tool. This commercial page keeps exits and collectors from being sold as the same product.

Official APIs versus headed collection 1

2026-07-02 · Morgan Ellis · Scraping

Official APIs versus headed collection

If there is an official API, that is the door. Headed collection is a different job with a different permission story. Compare them before you open a tab.

We will not scrape LinkedIn inboxes 1

2026-06-30 · Morgan Ellis · Scraping

We will not scrape LinkedIn inboxes

We will not scrape LinkedIn inboxes. Official API and public pages only. This commercial query gets a refusal, not another tool list. Local Windows notes only.

Instagram public research versus a scrape 1

2026-06-26 · Morgan Ellis · Scraping

Instagram public research versus a scrape

2023 code guide is the weak competitor page. Public research versus scrape. No code stealth. Look with a cold room, or use official channels.

Facebook Ads Library research locally 1

2026-06-24 · Morgan Ellis · Scraping

Facebook Ads Library research locally

Public Ads Library. Manual and official tools first. No login scrape. This commercial page is how a local room looks at ads the Library already shows.

Captchas are a stop sign 1

2026-06-21 · Morgan Ellis · Scraping

Captchas are a stop sign

Distinct from the captcha-solver commercial page. Here a captcha stops collection. It is not a feature request and not a solver bake-off. Authorized work only.

Anti-bot pages and official channels 1

2026-06-18 · Morgan Ellis · Scraping

Anti-bot pages and official channels

Their page is a bypass guide. Ours is official channels only. A Cloudflare wall is a stop, not a puzzle this blog will help you around. Authorized work only.

A local automation bench next to a closed script folder

2026-05-18 · Morgan Ellis · Scraping

What web scraping is

What collection is, and what it is not. Official channels first. This 2023-definition refresh is the scraping pillar, not a stealth-stack advertisement.