scrapeunblocker 0.5.0 → 0.6.0
This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
- checksums.yaml +4 -4
- data/CHANGELOG.md +6 -0
- data/README.md +1 -1
- data/lib/scrapeunblocker/client.rb +6 -4
- data/lib/scrapeunblocker/version.rb +1 -1
- metadata +2 -2
checksums.yaml
CHANGED
|
@@ -1,7 +1,7 @@
|
|
|
1
1
|
---
|
|
2
2
|
SHA256:
|
|
3
|
-
metadata.gz:
|
|
4
|
-
data.tar.gz:
|
|
3
|
+
metadata.gz: 3076debfbde8d8f4aac22c77b69c7d9117d2050743f7c5794e856b4c8f0618f4
|
|
4
|
+
data.tar.gz: 19dd437fccd00ae77927c1863695bfd0599f20d6a424e61ead7aff0abffbfd04
|
|
5
5
|
SHA512:
|
|
6
|
-
metadata.gz:
|
|
7
|
-
data.tar.gz:
|
|
6
|
+
metadata.gz: e0e42653b97107a96cd99ef480633c455b5833d38b57b1e18942232ff46684c48a7d616a3953c9ba887bf5eef848fa17cdafcba912b3014618d9dc45775570b2
|
|
7
|
+
data.tar.gz: c7809f130d23a492dd1faec14b243a77b8d1618f796998fc2aad9c05fa6d3c681f175fc54444e59eed818c5fb0906596daa697260f3b4c29b4bde25b689f5257
|
data/CHANGELOG.md
CHANGED
|
@@ -1,5 +1,11 @@
|
|
|
1
1
|
# Changelog
|
|
2
2
|
|
|
3
|
+
## 0.6.0 (2026-09-24)
|
|
4
|
+
|
|
5
|
+
- `google_images` takes `pages:` (1-5): fetch up to five Google Images result pages of ~100 results each in one call. Each page fetched is billed as one request; the response's `pagesFetched` says how many. The Google market now follows `proxy_country` automatically, so `gl` is only an optional override, and `max_results` is an optional cap up to 500.
|
|
6
|
+
|
|
7
|
+
No breaking changes.
|
|
8
|
+
|
|
3
9
|
## 0.5.0 (2026-09-17)
|
|
4
10
|
|
|
5
11
|
- Added `google_images(q, ...)` for the new Google Images plugin (`POST /images/google-search`). Given a keyword `q` it returns Google Images results as a Hash - each with the full-size `imageUrl` and its `sourceDomain`, plus the source page URL, title, source name, thumbnail URL, pixel dimensions and file size. Optional `gl` (ISO-2 lowercase market), `max_results` (1-100) and `proxy_country` (ISO-2) refine the search.
|
data/README.md
CHANGED
|
@@ -131,7 +131,7 @@ local["results"].each { |biz| puts "#{biz['name']} #{biz['rating']} #{biz['addre
|
|
|
131
131
|
## Google Images
|
|
132
132
|
|
|
133
133
|
```ruby
|
|
134
|
-
images = su.google_images("golden retriever puppy", proxy_country: "US",
|
|
134
|
+
images = su.google_images("golden retriever puppy", proxy_country: "US", pages: 3)
|
|
135
135
|
images["results"].each { |img| puts "#{img['imageUrl']} #{img['sourceDomain']} #{img['title']}" }
|
|
136
136
|
```
|
|
137
137
|
|
|
@@ -111,11 +111,13 @@ module ScrapeUnblocker
|
|
|
111
111
|
# Returns image results, each with the full-size +imageUrl+ and its
|
|
112
112
|
# +sourceDomain+, plus the source page URL, title, source name, thumbnail
|
|
113
113
|
# URL, pixel dimensions and file size. +q+ is the search keyword. Set
|
|
114
|
-
# +proxy_country+
|
|
115
|
-
# +
|
|
116
|
-
|
|
114
|
+
# +proxy_country+ to target a market (the Google market follows it; +gl+
|
|
115
|
+
# is an optional override), +pages+ (1-5, ~100 results each; each page
|
|
116
|
+
# fetched is billed as one request, see +pagesFetched+) and +max_results+
|
|
117
|
+
# (1-500) to cap the count.
|
|
118
|
+
def google_images(q, proxy_country: nil, pages: nil, gl: nil, max_results: nil)
|
|
117
119
|
post_json("/images/google-search",
|
|
118
|
-
q: q, gl: gl, max_results: max_results, proxy_country: proxy_country)
|
|
120
|
+
q: q, pages: pages, gl: gl, max_results: max_results, proxy_country: proxy_country)
|
|
119
121
|
end
|
|
120
122
|
|
|
121
123
|
# Fetch an advertiser's Meta (Facebook) Ad Library ads and return them as a Hash.
|
metadata
CHANGED
|
@@ -1,14 +1,14 @@
|
|
|
1
1
|
--- !ruby/object:Gem::Specification
|
|
2
2
|
name: scrapeunblocker
|
|
3
3
|
version: !ruby/object:Gem::Version
|
|
4
|
-
version: 0.
|
|
4
|
+
version: 0.6.0
|
|
5
5
|
platform: ruby
|
|
6
6
|
authors:
|
|
7
7
|
- ScrapeUnblocker
|
|
8
8
|
autorequire:
|
|
9
9
|
bindir: bin
|
|
10
10
|
cert_chain: []
|
|
11
|
-
date: 2026-09-
|
|
11
|
+
date: 2026-09-24 00:00:00.000000000 Z
|
|
12
12
|
dependencies: []
|
|
13
13
|
description: JS-rendered pages that bypass Cloudflare, DataDome, PerimeterX and Akamai,
|
|
14
14
|
plus Google SERP and Skyscanner flights/hotels/car-hire scraping as JSON.
|