Affect3DStore: use python cloudscraper with proxies to circumvent Cloudflare - #2863
Merged
Maista6969 merged 1 commit intoSep 3, 2026
Merged
Conversation
Collaborator
|
I'll merge this, but keep in mind that Stash v0.32 is adding support for TLS impersonation that makes it possible to keep using XPath for this 🙂 |
Contributor
Author
It will be nice to use a pure xPath scraper again if it can work with Cloudflare-protected sites. I tried to use the xPath selectors in the python code as much as possible, so it should be pretty easy to bring it back over to an xPath scraper (when Stash can work with Cloudflare protected sites) |
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Scraper type(s)
Examples to test
sceneByURL
sceneByName / sceneByFragment / sceneByQueryFragment
TBC
Short description
The site now has Cloudflare protection. It is possible to use cloudscraper to bypass Cloudflare protection. This has worked well for other scrapers when used together with free-proxy, so is used here too.
There is a text search on the site, so sceneByName and sceneByFragment/sceneByQueryFragment are probably possible too (so I'll look into that)