Headless Chrome isn't why your web agent gets blocked — fingerprints are

Share
Headless Chrome isn't why your web agent gets blocked — fingerprints are

Your agent's model is fine. Its browser disguise is what's blowing your scraping runs — and a maintainer of an open-source automation library has the test numbers to show it.

The most useful agent engineering post of the day didn't come from a lab or a vendor — it came from a Reddit thread by thalissonvs, maintainer of Pydoll, an MIT-licensed Python library that drives Chromium directly over the Chrome DevTools Protocol with no WebDriver layer. His claim: web agents don't get blocked because the model is weak or because headless mode is detectable. They get blocked because their browser fingerprint is incoherent. He tested it with live bot-score checks: plain headless Chrome scored 100 out of 100 — maximum bot. The same run with a fingerprint profile matched to the host machine dropped to 15, indistinguishable from a normal user's browser.

The failure mode that should make every agent builder wince: a mismatched profile scores worse than no spoofing at all. A Windows profile running on his Mac scored 57 — because one signal that contradicts all the others is itself the tell. Detection engines like CreepJS and commercial bot managers don't check whether your User-Agent looks human; they check whether a dozen signals agree, and disagreement is what flags you.

There's a hard limit worth stating plainly: none of this touches the network layer. Your TLS handshake and egress IP stay exactly what they are, so this is identity coherence for the browser layer only — which is why the advice is to match the profile to the machine you actually run on instead of faking everything. Pydoll's own documentation is blunt about it: an inconsistent fingerprint is more detectable than an unmodified browser.

Why it matters: browser-use agents are shipping everywhere right now, and most teams debug blocks by swapping models, adding retries, or rotating proxies — when the actual defect sits in the browser identity layer nobody instrumented. The fix is architectural, not expensive: coherent profile, matched to hardware, applied before first navigation.

Our take: treat this as infrastructure knowledge, not scraping tricks. Every team running production browsing agents will need fingerprint-coherence work within the year — and the counter-pressure from detection vendors is about to get interesting.

What to watch: whether major bot-detection vendors start scoring cross-layer coherence explicitly (some already do), and whether agent frameworks bake fingerprint matching in as a default rather than leaving it to each developer.

If your web agent keeps getting blocked despite a frontier model behind it — is the browser layer the first place you'd look now? Tell us in the comments.

Sources: Pydoll on GitHub

Read more

South Korea bets $3.49B on its own frontier AI model

South Korea bets $3.49B on its own frontier AI model

Sovereign-model money is getting serious, and the hardware money is following it. Today's inbox: Korea's nine-figure upgrade to its homegrown model push, a physics-simulation startup priced like a chip designer, and Google turning a geospatial model loose on public health. South Korea is putting 4.7 trillion won — about $3.49 billion — of state equity behind a homegrown frontier AI model. The Ministry of Science and ICT confirmed the figure as part of its proposed 2027 budget, split into two t

Mistral's Le Chonk puts Europe's sovereignty bet on a download date

Mistral's Le Chonk puts Europe's sovereignty bet on a download date

Mistral's biggest model ever is real, benchmarked and for sale today — but the thing that makes it matter to Europe's sovereignty argument, the weights, is still three weeks out. The preview settles who built it; the release will settle whether it counts. What Mistral actually shipped Mistral opened a public preview of Mistral Large 4 — unofficially ML4, officially le Chonk — a 1 trillion-parameter mixture-of-experts model with 49 billion active parameters and native multimodal input. The p

Mistral unveils Le Chonk: a 1T-parameter open-weights model

Mistral unveils Le Chonk: a 1T-parameter open-weights model

The biggest open-weight release outside China lands in public preview today, and the country that spent the week promising its own frontier model just put a price on the ambition. Mistral has opened a public preview of Mistral Large 4 — codenamed "le Chonk" — a 1 trillion-parameter mixture-of-experts model with 49 billion active parameters, natively multimodal, which the company calls its largest and most capable model to date. The preview API is live today on Mistral Studio at $1.36 per milli

Google signs 3.6 GW power deal, a quarter of it new nuclear

Google signs 3.6 GW power deal, a quarter of it new nuclear

Grid capacity, not chips, is becoming the binding constraint on the AI buildout — and on the same day, the labs told an Australian inquiry they can live with mandatory incident reporting. Google has contracted 3.6 GW of power from Constellation Energy across the PJM grid — the largest electricity deal in the region's history, and the biggest single power commitment any AI company has made. The agreement covers 3,590 megawatts over 13 states, with 890 megawatts of new nuclear capacity coming fr