hasdata-google-scholar-mcp 1.0.0__tar.gz

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
@@ -0,0 +1,21 @@
1
+ ---
2
+ name: Something in the README is wrong
3
+ about: A tool table, a response sample or a documented behaviour does not match reality
4
+ labels: documentation
5
+ ---
6
+
7
+ **Where in the README**
8
+
9
+ Section or heading.
10
+
11
+ **What it says**
12
+
13
+ Quote the line.
14
+
15
+ **The call you made**
16
+
17
+ Tool name and arguments, or the equivalent REST URL with your key removed.
18
+
19
+ **What came back**
20
+
21
+ Trimmed response, with anything private removed.
@@ -0,0 +1,31 @@
1
+ # The tool contract is checked on a schedule as well as on push, because the upstream tool list
2
+ # can change without a single commit in this repository.
3
+ name: tool contract
4
+
5
+ on:
6
+ push:
7
+ branches: [main]
8
+ pull_request:
9
+ schedule:
10
+ - cron: '0 6 * * 1'
11
+ workflow_dispatch:
12
+
13
+ permissions:
14
+ contents: read
15
+
16
+ jobs:
17
+ contract:
18
+ runs-on: ubuntu-latest
19
+ timeout-minutes: 5
20
+ steps:
21
+ - uses: actions/checkout@v4
22
+ - uses: actions/setup-node@v4
23
+ with:
24
+ node-version: '22'
25
+ # Forks cannot read repository secrets. The suite skips its live checks when the key is
26
+ # absent, so a pull request from a fork stays green instead of failing for a reason the
27
+ # contributor cannot fix.
28
+ - name: Assert the tool list still matches the README
29
+ env:
30
+ HASDATA_API_KEY: ${{ secrets.HASDATA_API_KEY }}
31
+ run: npm test
@@ -0,0 +1,75 @@
1
+ # Publishes the npm and PyPI wrapper packages on a version tag, using OIDC
2
+ # trusted publishing. No NPM_TOKEN or PYPI_TOKEN is stored anywhere: GitHub
3
+ # mints a short-lived OIDC token per run, and npmjs.org / pypi.org accept it
4
+ # because this repo + workflow are configured as trusted publishers.
5
+ #
6
+ # One-time setup, done once per package on the registries (not in this repo):
7
+ # npmjs.org -> package settings -> Trusted Publisher -> GitHub Actions,
8
+ # repo HasData/google-scholar-mcp, workflow publish.yml
9
+ # pypi.org -> the hasdata org -> Publishing -> add a trusted publisher
10
+ # (pending publisher works before the first release),
11
+ # repo HasData/google-scholar-mcp, workflow publish.yml
12
+ #
13
+ # The MCP registry entry (com.hasdata/google-scholar) is NOT published here. It uses
14
+ # domain auth, which would need the namespace-wide Ed25519 key as a secret in
15
+ # every repo. That key stays off CI; the registry entry is published by hand
16
+ # when server.json changes, after the package versions below are live.
17
+ #
18
+ # Release: bump nothing by hand. Tag the commit `vX.Y.Z` and push the tag; the
19
+ # tag is the single source of the version and is written into both manifests.
20
+
21
+ name: publish
22
+
23
+ on:
24
+ push:
25
+ tags: ['v*.*.*']
26
+
27
+ permissions:
28
+ contents: read
29
+ id-token: write
30
+
31
+ jobs:
32
+ npm:
33
+ runs-on: ubuntu-latest
34
+ steps:
35
+ - uses: actions/checkout@v4
36
+ # No registry-url here on purpose. With it, setup-node writes an .npmrc
37
+ # carrying _authToken=${NODE_AUTH_TOKEN}, which resolves to a placeholder
38
+ # when no token is passed. npm then authenticates with that garbage instead
39
+ # of falling back to OIDC, and a scoped package answers 404.
40
+ - uses: actions/setup-node@v4
41
+ with:
42
+ node-version: '24'
43
+ # OIDC trusted publishing landed in npm 11.5.1. Node 24 already ships a
44
+ # newer npm than that, but pinning the upgrade here keeps the job working
45
+ # if the runner image drifts back.
46
+ - name: Upgrade npm for OIDC trusted publishing
47
+ run: |
48
+ npm install -g npm@latest
49
+ npm -v
50
+ - name: Set version from the tag
51
+ run: npm version "${GITHUB_REF_NAME#v}" --no-git-tag-version --allow-same-version
52
+ - name: Publish to npm (OIDC, no token)
53
+ run: npm publish --access public
54
+
55
+ pypi:
56
+ runs-on: ubuntu-latest
57
+ steps:
58
+ - uses: actions/checkout@v4
59
+ - uses: actions/setup-python@v5
60
+ with:
61
+ python-version: '3.12'
62
+ - name: Set version from the tag
63
+ run: |
64
+ python - "${GITHUB_REF_NAME#v}" <<'PY'
65
+ import re, sys
66
+ v = sys.argv[1]
67
+ p = "pyproject.toml"
68
+ t = open(p, encoding="utf-8").read()
69
+ t = re.sub(r'(?m)^version = ".*"$', f'version = "{v}"', t, count=1)
70
+ open(p, "w", encoding="utf-8", newline="\n").write(t)
71
+ PY
72
+ - name: Build the wheel and sdist
73
+ run: pipx run build
74
+ - name: Publish to PyPI (OIDC, no token)
75
+ uses: pypa/gh-action-pypi-publish@release/v1
@@ -0,0 +1,11 @@
1
+ node_modules/
2
+ package-lock.json
3
+ dist/
4
+ build/
5
+ *.egg-info/
6
+ __pycache__/
7
+ *.pyc
8
+ .env
9
+ .env.*
10
+ .DS_Store
11
+ *.log
@@ -0,0 +1,21 @@
1
+ MIT License
2
+
3
+ Copyright (c) 2026 HasData
4
+
5
+ Permission is hereby granted, free of charge, to any person obtaining a copy
6
+ of this software and associated documentation files (the "Software"), to deal
7
+ in the Software without restriction, including without limitation the rights
8
+ to use, copy, modify, merge, publish, distribute, sublicense, and/or sell
9
+ copies of the Software, and to permit persons to whom the Software is
10
+ furnished to do so, subject to the following conditions:
11
+
12
+ The above copyright notice and this permission notice shall be included in all
13
+ copies or substantial portions of the Software.
14
+
15
+ THE SOFTWARE IS PROVIDED "AS IS", WITHOUT WARRANTY OF ANY KIND, EXPRESS OR
16
+ IMPLIED, INCLUDING BUT NOT LIMITED TO THE WARRANTIES OF MERCHANTABILITY,
17
+ FITNESS FOR A PARTICULAR PURPOSE AND NONINFRINGEMENT. IN NO EVENT SHALL THE
18
+ AUTHORS OR COPYRIGHT HOLDERS BE LIABLE FOR ANY CLAIM, DAMAGES OR OTHER
19
+ LIABILITY, WHETHER IN AN ACTION OF CONTRACT, TORT OR OTHERWISE, ARISING FROM,
20
+ OUT OF OR IN CONNECTION WITH THE SOFTWARE OR THE USE OR OTHER DEALINGS IN THE
21
+ SOFTWARE.
@@ -0,0 +1,391 @@
1
+ Metadata-Version: 2.5
2
+ Name: hasdata-google-scholar-mcp
3
+ Version: 1.0.0
4
+ Summary: MCP server for Google Scholar through HasData's hosted API. 1,000 free credits every month.
5
+ Project-URL: Homepage, https://docs.hasdata.com/apis/google-scholar/scholar
6
+ Project-URL: Repository, https://github.com/HasData/google-scholar-mcp
7
+ License: MIT
8
+ License-File: LICENSE
9
+ Keywords: academic,bibliometrics,citations,google-scholar,hasdata,mcp,model-context-protocol,research
10
+ Requires-Python: >=3.10
11
+ Requires-Dist: mcp-proxy>=0.12.0
12
+ Requires-Dist: mcp<2,>=1.17
13
+ Description-Content-Type: text/markdown
14
+
15
+ # Google Scholar MCP Server
16
+
17
+ <!-- mcp-name: com.hasdata/google-scholar -->
18
+
19
+ A hosted Model Context Protocol (MCP) server that gives Claude, Cursor, Windsurf and any other MCP client two read-only Google Scholar tools. Search the literature with Scholar's own operators, year ranges and cited-by lookups, then pull a paper's citation in five styles with its BibTeX and EndNote export links, both as structured JSON, with nothing to host.
20
+
21
+ It reads public Google Scholar pages that a signed-out visitor can see.
22
+
23
+ **1,000 free credits every month, no card required**, which is 100 Scholar calls at the 10-credit rate.
24
+
25
+ ```
26
+ https://mcp.hasdata.com/api/mcp?apis=google_scholar
27
+ ```
28
+
29
+ [![Glama score](https://glama.ai/mcp/servers/HasData/google-scholar-mcp/badges/score.svg)](https://glama.ai/mcp/servers/HasData/google-scholar-mcp)
30
+ [![tool contract](https://github.com/HasData/google-scholar-mcp/actions/workflows/contract.yml/badge.svg)](https://github.com/HasData/google-scholar-mcp/actions/workflows/contract.yml)
31
+ [![MCP](https://img.shields.io/badge/MCP-remote%20%7C%20streamable%20HTTP-6366f1?style=flat-square)](https://mcp.hasdata.com/api/mcp?apis=google_scholar)
32
+ [![Tools](https://img.shields.io/badge/tools-2-10b981?style=flat-square)](#tools)
33
+ [![npm](https://img.shields.io/npm/v/@hasdata/google-scholar-mcp?style=flat-square&logo=npm&label=npm&color=cb3837)](https://www.npmjs.com/package/@hasdata/google-scholar-mcp)
34
+ [![PyPI](https://img.shields.io/pypi/v/hasdata-google-scholar-mcp?style=flat-square&logo=pypi&logoColor=white&label=PyPI&color=3775a9)](https://pypi.org/project/hasdata-google-scholar-mcp/)
35
+ [![License](https://img.shields.io/badge/license-MIT-blue?style=flat-square)](LICENSE)
36
+
37
+ ## Contents
38
+
39
+ - [What you need](#what-you-need)
40
+ - [Quick start](#quick-start)
41
+ - [Example prompts](#example-prompts)
42
+ - [Tools](#tools)
43
+ - [Errors and failure paths](#errors-and-failure-paths)
44
+ - [Pricing, free tier and limits](#pricing-free-tier-and-limits)
45
+ - [Tool selection](#tool-selection)
46
+ - [How it compares](#how-it-compares)
47
+ - [FAQ](#faq)
48
+ - [HasData links](#hasdata-links)
49
+ - [Development](#development)
50
+ - [Contributing](#contributing)
51
+ - [License](#license)
52
+
53
+ ## What you need
54
+
55
+ An MCP client and a HasData API key from the [dashboard](https://app.hasdata.com/sign-up?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp), free to create with no card, and the free tier covers about 100 calls a month at the 10-credit rate. This is a remote server, so the simplest path is a URL and an `x-api-key` header, with no container to run. A client that only speaks stdio reaches it through a thin launcher, published as `@hasdata/google-scholar-mcp` on npm and `hasdata-google-scholar-mcp` on PyPI, shown below.
56
+
57
+ ## Quick start
58
+
59
+ The server URL is the same for every client. We run it hands-on in Claude Code and Claude Desktop. The other blocks follow each client's own documented format for a remote server.
60
+
61
+ | Field | Value |
62
+ | :--- | :--- |
63
+ | URL | `https://mcp.hasdata.com/api/mcp?apis=google_scholar` |
64
+ | Transport | HTTP, streamable |
65
+ | Auth header | `x-api-key: HASDATA_API_KEY` |
66
+
67
+ Clients with OAuth support can add the same URL as a connector and sign in without putting a key in a config file.
68
+
69
+ <details>
70
+ <summary><b>Claude Code</b></summary>
71
+
72
+ ```bash
73
+ claude mcp add --transport http google-scholar "https://mcp.hasdata.com/api/mcp?apis=google_scholar" \
74
+ --header "x-api-key: HASDATA_API_KEY"
75
+ ```
76
+
77
+ </details>
78
+
79
+ <details>
80
+ <summary><b>Claude Desktop</b></summary>
81
+
82
+ Settings, then Connectors, then Add custom connector, then paste `https://mcp.hasdata.com/api/mcp?apis=google_scholar` and sign in.
83
+
84
+ For the config-file route, Claude Desktop loads only local (stdio) servers, so it reaches a remote server through a stdio launcher. The `@hasdata/google-scholar-mcp` package is that launcher, and it reads the key from the environment. Add this to `claude_desktop_config.json`:
85
+
86
+ ```json
87
+ {
88
+ "mcpServers": {
89
+ "google-scholar": {
90
+ "command": "npx",
91
+ "args": ["-y", "@hasdata/google-scholar-mcp"],
92
+ "env": { "HASDATA_API_KEY": "YOUR_KEY" }
93
+ }
94
+ }
95
+ }
96
+ ```
97
+
98
+ For Python instead of Node, swap the launcher for the PyPI package, which `uvx` runs without a manual install:
99
+
100
+ ```json
101
+ {
102
+ "mcpServers": {
103
+ "google-scholar": {
104
+ "command": "uvx",
105
+ "args": ["hasdata-google-scholar-mcp"],
106
+ "env": { "HASDATA_API_KEY": "YOUR_KEY" }
107
+ }
108
+ }
109
+ }
110
+ ```
111
+
112
+ </details>
113
+
114
+ <details>
115
+ <summary><b>Cursor</b></summary>
116
+
117
+ `~/.cursor/mcp.json` for every project, or `.cursor/mcp.json` for one:
118
+
119
+ ```json
120
+ {
121
+ "mcpServers": {
122
+ "google-scholar": {
123
+ "url": "https://mcp.hasdata.com/api/mcp?apis=google_scholar",
124
+ "headers": { "x-api-key": "HASDATA_API_KEY" }
125
+ }
126
+ }
127
+ }
128
+ ```
129
+
130
+ </details>
131
+
132
+ <details>
133
+ <summary><b>Windsurf</b></summary>
134
+
135
+ `~/.codeium/windsurf/mcp_config.json`. Windsurf calls the field `serverUrl`, not `url`:
136
+
137
+ ```json
138
+ {
139
+ "mcpServers": {
140
+ "google-scholar": {
141
+ "serverUrl": "https://mcp.hasdata.com/api/mcp?apis=google_scholar",
142
+ "headers": { "x-api-key": "HASDATA_API_KEY" }
143
+ }
144
+ }
145
+ }
146
+ ```
147
+
148
+ </details>
149
+
150
+ <details>
151
+ <summary><b>VS Code</b></summary>
152
+
153
+ `.vscode/mcp.json` in the workspace:
154
+
155
+ ```json
156
+ {
157
+ "servers": {
158
+ "google-scholar": {
159
+ "type": "http",
160
+ "url": "https://mcp.hasdata.com/api/mcp?apis=google_scholar",
161
+ "headers": { "x-api-key": "HASDATA_API_KEY" }
162
+ }
163
+ }
164
+ }
165
+ ```
166
+
167
+ </details>
168
+
169
+ ## Example prompts
170
+
171
+ Each of these lands on one tool, or on two in sequence when the second needs an id the first returns.
172
+
173
+ - Find papers on transformer architectures published since 2023 and sort them by citation count.
174
+ - Who has cited this paper, and how has that grown year on year?
175
+ - Give me the BibTeX for this paper.
176
+ - Find review articles only on this topic, excluding citations without full records.
177
+ - Show me every indexed version of this paper and which of them have a PDF.
178
+ - Find recent work by this author on this topic.
179
+
180
+ A prompt naming a paper takes two calls, one search to reach its `resultId` and one citation lookup. A prompt about who cites a paper also takes two, because the second call reuses `citedBy.citesId` from the first.
181
+
182
+ ## Tools
183
+
184
+ Two tools, 10 credits per successful call.
185
+
186
+ ### Get Scholar search results
187
+
188
+ [`hasdata_google_scholar_scholar_getScholarSearchResults`](https://docs.hasdata.com/apis/google-scholar/scholar?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp)
189
+
190
+ A page of Scholar results.
191
+
192
+ | Parameter | Type | Required | Notes |
193
+ | :--- | :--- | :--- | :--- |
194
+ | `q` | string | yes | The query. Scholar operators such as `author:` and `source:` work here |
195
+ | `asYlo` / `asYhi` | number | | Published from and up to these years |
196
+ | `start` | number | | Result offset, where `0` is the first result |
197
+ | `num` | number | | Results per page |
198
+ | `scisbd` | number | | `1` sorts abstracts by date, `2` sorts everything by date. Omit for relevance |
199
+ | `cites` | string | | Find articles citing this one, using a `citedBy.citesId` |
200
+ | `cluster` | string | | Find every indexed version of an article, using a `versions.clusterId` |
201
+ | `asSdt` | string | | Search type, `0,5` for articles, `4` for case law, `0` or `7` for patents |
202
+ | `asRr` | number | | `1` returns review articles only |
203
+ | `asVis` | number | | `1` excludes citations, `0` includes them |
204
+ | `hl` | string | | Interface language, one of 159 |
205
+ | `lr` | array | | Restrict to these content languages |
206
+ | `safe` | string | | `active` or `off` |
207
+ | `filter` | number | | `1` keeps Google's similar and omitted result filters, `0` drops them |
208
+
209
+ Returns `searchInformation` with `totalResults`, `queryDisplayed` and the time Scholar reported, an `organicResults` array, and `pagination`.
210
+
211
+ Each result carries `position`, `resultId`, `title`, `link`, `snippet`, a `publicationInfo` object, a `resources` array, `citedBy`, `versions`, `relatedPagesLink` and `citeHasdataLink`.
212
+
213
+ `publicationInfo.authors` is the part worth knowing about. Alongside the raw `summary` line it lists the authors Scholar has profiles for, each with a `name`, a profile `link` and an `authorId`, which is how you follow one author rather than parsing a byline.
214
+
215
+ `citedBy` and `versions` are the two ids that make this tool compose with itself. `citedBy.citesId` goes back into `cites` to walk a citation graph, and `versions.clusterId` goes into `cluster` to see every indexed copy of the same paper.
216
+
217
+ ```json
218
+ {
219
+ "position": 1,
220
+ "resultId": "A7L9JolPKkoJ",
221
+ "title": "A historical survey of advances in transformer architectures",
222
+ "link": "https://www.mdpi.com/2076-3417/14/10/4316",
223
+ "snippet": "… of the Vision Transformer (ViT) opening a new realm of architectures which build … transformer architecture, it becomes pertinent to examine in detail the architecture of the transformer …",
224
+ "publicationInfo": {
225
+ "summary": "AR Sajun, I Zualkernan, D Sankalpa - Applied Sciences, 2024 - mdpi.com",
226
+ "authors": [
227
+ { "name": "AR Sajun", "authorId": "k6zWX4EAAAAJ", "link": "https://scholar.google.com/citations?user=k6zWX4EAAAAJ&hl=en" }
228
+ ]
229
+ },
230
+ "resources": [{ "fileFormat": "Html", "title": "mdpi.com", "link": "https://www.mdpi.com/2076-3417/14/10/4316" }],
231
+ "citedBy": { "total": 97, "citesId": "5344171358311789059" },
232
+ "versions": { "total": 7, "clusterId": "5344171358311789059" }
233
+ }
234
+ ```
235
+
236
+ ### Get Scholar citation formats
237
+
238
+ [`hasdata_google_scholar_cite_getScholarCitationFormats`](https://docs.hasdata.com/apis/google-scholar/cite?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp)
239
+
240
+ The citation block for one paper.
241
+
242
+ | Parameter | Type | Required | Notes |
243
+ | :--- | :--- | :--- | :--- |
244
+ | `q` | string | yes | A `resultId` from a search result, not a search query |
245
+ | `hl` | string | | Interface language |
246
+
247
+ Returns `citations`, the formatted string in MLA, APA, Chicago, Harvard and Vancouver, and `links`, the export URLs for BibTeX, EndNote, RefMan and RefWorks.
248
+
249
+ ```json
250
+ {
251
+ "citations": [
252
+ {
253
+ "title": "APA",
254
+ "snippet": "Sajun, A. R., Zualkernan, I., & Sankalpa, D. (2024). A historical survey of advances in transformer architectures. Applied Sciences, 14(10), 4316."
255
+ }
256
+ ],
257
+ "links": [
258
+ { "name": "BibTeX", "link": "https://scholar.googleusercontent.com/scholar.bib?q=info:A7L9JolPKkoJ:scholar.google.com/&output=citation..." }
259
+ ]
260
+ }
261
+ ```
262
+
263
+ ## Errors and failure paths
264
+
265
+ Plan for these rather than assuming a happy path.
266
+
267
+ **On the citation tool, `q` is a paper id rather than a query.** It takes the `resultId` from a search result, such as `A7L9JolPKkoJ`. Passing a title or a DOI there returns nothing useful, and the shared parameter name is the reason people get this wrong.
268
+
269
+ **`citesId` and `clusterId` can hold the same value, and they are not interchangeable.** They were identical on the paper above. One goes into `cites` to find papers citing this one, the other into `cluster` to find copies of this one. Sending the right number to the wrong parameter returns a plausible page of the wrong thing.
270
+
271
+ **`type` arrives on some results and not others.** It was present on one result in five, describing the format of the primary resource. Read `resources[].fileFormat` when you need to know whether a PDF exists.
272
+
273
+ **A BibTeX link is a Scholar URL with a signature in it, not the BibTeX itself.** The `links` array gives you addresses to fetch, and those carry expiring `scisig` tokens, so fetch them promptly rather than storing them for later.
274
+
275
+ **`totalResults` is Scholar's estimate and it is very rough.** The query above reported 2,080,000. Treat it as an order of magnitude, never as a count.
276
+
277
+ **Scholar counts citations, not quality, and it indexes preprints, theses and citing-only records.** `asVis: 1` drops citation-only entries when you need records with full metadata.
278
+
279
+ **Paging deep gets thin.** Scholar limits how far a result set goes and starts repeating or blocking well before the estimate suggests, so narrow with `asYlo`, `asYhi` or `asSdt` rather than walking `start` upward.
280
+
281
+ Results that carry data also carry a `requestMetadata.id` worth quoting in support.
282
+
283
+ ## Pricing, free tier and limits
284
+
285
+ Each Scholar tool costs **10 credits per successful call**. Response size does not change the price, so raising `num` is the cheap way to widen a search.
286
+
287
+ The free tier is **1,000 credits every month with no card**, which is 100 Scholar calls at the base rate. It renews with the billing cycle, so a low-volume agent runs on the free tier indefinitely.
288
+
289
+ Paid plans start at **$49 a month** for 200,000 credits, which is 20,000 calls. The unit price falls with volume, from **$2.45 per 1,000 calls** on the entry plan to **$1.00** on Business, **$0.84** on Growth and **$0.74** on the largest [high-volume plans](https://hasdata.com/prices?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp).
290
+
291
+ Your plan also sets concurrency. The free tier allows 1 request at a time, Startup 15, Business 30, Growth 50, and the high-volume plans run from 200 to 1,500. Retry on the 429 with a backoff in anything unattended, because an agent that walks a citation graph will reach the ceiling before you do.
292
+
293
+ A request that comes back non-200 is not billed. A successful call that finds nothing is still a call.
294
+
295
+ ## Tool selection
296
+
297
+ Start from what the prompt gives you. A topic, an author or a year range goes to the search tool. A `resultId` you already hold goes straight to the citation tool.
298
+
299
+ Then think about which id the next step needs. A literature sweep is one search call with a large `num`. A citation graph is a search call followed by one `cites` call per paper you follow. A deduplication pass across preprints and published versions is a `cluster` call per paper. Each of those reuses an id the first response already gave you, so a second search call is usually wasted.
300
+
301
+ Raise `num` before you page. Cost is per call rather than per result, so one wide page beats three narrow ones.
302
+
303
+ ## How it compares
304
+
305
+ Google Scholar has no public API, so the comparison worth making is against the open bibliographic APIs.
306
+
307
+ | | OpenAlex or Semantic Scholar | This server |
308
+ | :--- | :--- | :--- |
309
+ | Eligibility | Open, no key for basic use | An API key |
310
+ | Coverage | Large, curated, DOI-centred | What Scholar indexes, including theses and preprints |
311
+ | Citation counts | Their own, computed from their graph | Scholar's, as displayed |
312
+ | Citation strings | Build them yourself from metadata | MLA, APA, Chicago, Harvard, Vancouver, as Scholar formats them |
313
+ | Full-text links | DOI and open-access locations | The resource links Scholar shows, including PDFs |
314
+ | Structured metadata | Rich and typed | As the page presents it |
315
+
316
+ The row that decides it is whose citation count you need. For bibliometrics on typed, stable metadata, OpenAlex and Semantic Scholar are better instruments and they are free. Reach for this one when the question is specifically about what Google Scholar shows, which is what most researchers actually look at, or when you want the formatted citation rather than the fields to build one.
317
+
318
+ ## FAQ
319
+
320
+ ### Is there an official Google Scholar MCP server?
321
+
322
+ Google does not publish one, and Scholar has no public API either. This one is maintained by HasData and reads public Scholar pages.
323
+
324
+ ### What is a Google Scholar MCP server?
325
+
326
+ An MCP server exposes tools an AI client can call. This one turns Scholar search results and citation blocks into JSON an agent can reason over, without a browser or a scraping library in your stack.
327
+
328
+ ### Do I need a Google account?
329
+
330
+ No. The only credential is your HasData key.
331
+
332
+ ### How do I get the BibTeX for a paper?
333
+
334
+ Two calls. Search to get the paper's `resultId`, then pass that id as `q` to the citation tool, and take the BibTeX URL from `links`.
335
+
336
+ ### How do I find everything that cites a paper?
337
+
338
+ Take `citedBy.citesId` from the search result and send it back as the `cites` parameter. The response is the citing papers, paged like any other search.
339
+
340
+ ### What is the difference between `cites` and `cluster`?
341
+
342
+ `cites` finds papers that cite the one you named. `cluster` finds other indexed versions of that same paper, such as a preprint next to the published article. The two ids often look identical, so pick by what you want rather than by the number.
343
+
344
+ ### Can I search by author?
345
+
346
+ Yes, with Scholar's own operator, as `author:"J Dean"` in `q`. Search results also carry `authorId` for authors with a Scholar profile, which is the stabler handle.
347
+
348
+ ### Can I use this together with other HasData APIs?
349
+
350
+ Yes. One key covers everything, and one endpoint serves them all through the `apis` parameter. Point a client at `?apis=google_scholar,google_serp` to get both tool sets in one connection, or at [`mcp.hasdata.com/api/mcp`](https://docs.hasdata.com/mcp-server?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp) for the full catalogue.
351
+
352
+ ### Is HasData affiliated with Google?
353
+
354
+ No. HasData is an independent service and is not affiliated with, endorsed by, or sponsored by Google. Google Scholar is a trademark of its respective owner. The tools work with publicly available data only, and you are responsible for using the results in line with Google's terms and the law that applies to you.
355
+
356
+ ### Compliance and personal data
357
+
358
+ Author names, affiliations and Scholar profile ids are personal data, even though they are published as part of the scholarly record. Building a profile of one researcher's output is a different act from counting citations on a topic, and it is the one that needs a second thought about purpose and retention. The papers themselves stay under their own licences, so a link is not permission to redistribute a PDF.
359
+
360
+ ## HasData links
361
+
362
+ - [Google Scholar API documentation](https://docs.hasdata.com/apis/google-scholar/scholar?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp), the REST endpoints behind these tools
363
+ - [Citation formats endpoint](https://docs.hasdata.com/apis/google-scholar/cite?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp)
364
+ - [MCP server documentation](https://docs.hasdata.com/mcp-server?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp)
365
+ - [Pricing](https://hasdata.com/prices?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp)
366
+ - [Dashboard](https://app.hasdata.com/sign-up?utm_source=github&utm_medium=syndication&utm_campaign=google-scholar-mcp)
367
+
368
+ Other HasData MCP servers: [Google Search](https://github.com/HasData/google-search-mcp), [Google Images](https://github.com/HasData/google-images-mcp), [Google Maps](https://github.com/HasData/google-maps-mcp), [Google Trends](https://github.com/HasData/google-trends-mcp), [Bing](https://github.com/HasData/bing-mcp), [DuckDuckGo](https://github.com/HasData/duckduckgo-mcp), [YouTube](https://github.com/HasData/youtube-mcp), [TikTok](https://github.com/HasData/tiktok-mcp), [Instagram](https://github.com/HasData/instagram-mcp), [Amazon](https://github.com/HasData/amazon-mcp), [Walmart](https://github.com/HasData/walmart-mcp), [Shopify](https://github.com/HasData/shopify-mcp), [Yelp](https://github.com/HasData/yelp-mcp), [Yellow Pages](https://github.com/HasData/yellowpages-mcp), [Zillow](https://github.com/HasData/zillow-mcp), [Redfin](https://github.com/HasData/redfin-mcp), [Airbnb](https://github.com/HasData/airbnb-mcp), [Booking.com](https://github.com/HasData/booking-mcp), [Indeed](https://github.com/HasData/indeed-mcp), [Glassdoor](https://github.com/HasData/glassdoor-mcp).
369
+
370
+ ## Development
371
+
372
+ The launcher is a thin stdio bridge to the remote server, so there is nothing to build.
373
+
374
+ ```bash
375
+ npm install
376
+ HASDATA_API_KEY=your_key_here npm test
377
+ ```
378
+
379
+ The tests in `test/` assert the tool contract, the part that can break without a commit here. They check that `?apis=google_scholar` returns the two expected tools, that no name changed, that both still require `q` and carry descriptions, that the search parameters this README documents are still in the schema, and that the key in use is actually accepted.
380
+
381
+ Two tests go further. One asserts that a live search still returns `resultId`, `citedBy.citesId` and `versions.clusterId`, because those three ids are what let the tools compose and nothing else in the response would reveal their loss. The other feeds a `resultId` straight into the citation tool, which is the two-call workflow this README documents, and checks the five styles come back. Together they cost 20 credits a run, which is the price of a canary that can fail for the right reason.
382
+
383
+ The contract suite also runs weekly on a schedule, because the upstream tool list can change without anyone touching this repository.
384
+
385
+ ## Contributing
386
+
387
+ A tool table, a response sample or a documented behaviour that does not match reality is worth an issue. There is a template for exactly that. Pull requests are welcome for the same, and for anything in the launcher.
388
+
389
+ ## License
390
+
391
+ MIT, see [LICENSE](LICENSE).