Documentation Python quickstart Blog Free tools hello@quanticdata.ioLog in

Glassdoor Proxies: 624 Words, Skip the Browser

What the Glassdoor home page returns with and without a browser, measured through residential proxies on 26 and 28 September 2026: 624 words over plain HTTP from the United States and 589 from the United Kingdom, against 54 words and a redirect to the login page for 705,071 bytes when rendered
What the Glassdoor home page returns with and without a browser, measured through residential proxies on 26 and 28 September 2026: 624 words over plain HTTP from the United States and 589 from the United Kingdom, against 54 words and a redirect to the login page for 705,071 bytes when rendered

Glassdoor proxies get bought for a headless browser that the measurement says you should never launch. On 26 September 2026 the Glassdoor home page answered a plain HTTP client through a US residential exit with 200 OK and 624 words; rendered in a browser, the same URL returned 54 words and a canonical of /member/profile/login. On 28 September the UK edition gave a plain client 589 words with the h1 and the Organization schema, while the rendered fetch landed on the login page for 705,071 bytes and 61 seconds. The setting is residential, plain HTTP, country pinned to the edition, and no browser at all; for listings with salaries as numbers, the jobs collector of the same operator returns ten rows in 7.3 seconds.

The keyword is a company called Proxy

Google has no autocomplete for "glassdoor proxies", and the first page for the term is a homonym: seven of ten results are Glassdoor's own pages about companies named Proxy and Proxy Live Solutions, 3.9 stars from 27 reviews, plus a search page for 5,020 jobs with proxy in the title. The related searches, "glassdoor proxies legit" and "glassdoor proxies reviews", are people wondering whether to work there. The two results that are not Glassdoor are vendor guides, and the top one teaches a Python and Playwright scraper, on the reasoning that a simple HTML parser will miss dynamically loaded data.

The real demand is one word over. "Glassdoor scraper" autocompletes to selenium, github, api, extension, reddit, python, review scraper, job scraper, interview scraper and salary scraper; "scrape glassdoor" to reviews, interview questions, data, jobs and the question "can you scrape glassdoor". Those searchers want reviews, salaries and listings as rows for compensation benchmarking, employer-brand monitoring and labour-market research. That is what this post measures, and the first finding is that the tool the top guide recommends is the one that returns the least.

Plain HTTP gets the page, the browser gets the login form

We audited glassdoor.com from a United States residential exit on 26 September with our SEO audit, which fetches once as a pure HTTP client and once rendered, repeated it from a United Kingdom exit on 28 September against the UK edition, and weighed the rendered fetch separately.

FetchStatusWordsTitle or h1Canonical
Home, plain HTTP, US, 26 September200624Glassdoor homeglassdoor.com
Home, rendered, US, 26 September20054Log In | Glassdoor/member/profile/login
Home, plain HTTP, UK, 28 September200589You deserve a job that loves you backglassdoor.co.uk/index.htm
Home, rendered, UK, 28 Septembern/a600600 words, no canonicalabsent
Home, rendered and weighed, US, 28 Septembern/a44Log In | Glassdoor/member/profile/login

Read the two US rows together. Without JavaScript the server hands over the home page: navigation, search, the headline, the promotional copy, the footer, 624 words of it. With JavaScript the page's own scripts decide that a browser without a member session belongs on the login form, and the canonical link on what arrives is https://www.glassdoor.com/member/profile/login. The page was delivered, and then the page sent the browser away. That is what a harmful render looks like, and it is the reason the verdict on this target is not "useless" but "harmful": rendering does not just cost more, it replaces the data with a form.

The UK rows say the same thing in a different shape. Over plain HTTP the UK home page arrived with a canonical, an h1, a meta description and two JSON-LD blocks, WebSite and Organization, for 589 words. Rendered, it returned 600 words with no canonical, no h1 from the page and no JSON-LD: the same notice repeated in seven languages. A word count alone would call that a success; the missing canonical is what tells you it is not. On Glassdoor, classify on the canonical link and the h1, and treat a canonical ending in /member/profile/login as "rendered the wrong page", not as "missing item".

The weighed render from the US puts a price on the mistake: 705,071 bytes and 60.8 seconds of browser time to arrive at a login form. Glassdoor puts the reason in the query string of that URL, and the reason is the browser.

The exit country picks the edition

Glassdoor runs national editions on their own domains: glassdoor.com, glassdoor.co.uk, glassdoor.de and others. The exit country does not silently move you between them the way Indeed does; you choose the domain, and the edition sets the currency, the salary format and the review pool. What the exit does change is whether a residential address looks like it belongs to that edition's market, and the two editions we measured agreed on the shape of the answer:

EditionExitWords, plain HTTPWords, renderedRendered document
glassdoor.comUnited States62454Login form
glassdoor.co.ukUnited Kingdom589600600 words, no canonical

Pin the exit to the edition: country=us for glassdoor.com, country=gb for glassdoor.co.uk. A US address reading the UK edition is a mismatch the site can score, and there is no reason to give it one. And whatever the edition, do not render. The 35-word difference between the two plain fetches is copy; the 570-word difference between plain and rendered on the US edition is the data.

robots.txt says how deep the site wants you to go

glassdoor.com/robots.txt is 6,250 bytes and it is an honest map of what the site is willing to serve. Every user agent is disallowed from /api/, /graph, /search/, /member/ and /profile/, and from pagination: /Reviews/*_P*.htm, /Jobs/*_P*.htm and /Interview/*_P*.htm are all disallowed, with page 2 of a reviews search and pages 2 to 5 of a salaries listing allowed back in by name. The login page, /member/profile/login, is explicitly allowed, which is why the browser's destination is crawlable and the data behind it is not.

Then a block headed "Partially Block High-Quality AI/LLM Bots" names GPTBot, Google-Extended, Amazonbot, anthropic-ai, ClaudeBot, Perplexity, Cohere, Applebot-Extended and Google-CloudVertexBot and disallows them from everything except /blog/, /Award/, /About/ and /employers/; a second block fully blocks CCBot, Bytespider, Diffbot, FacebookBot and others. Reviews, salaries and interviews are off limits to every AI crawler by name, and the pagination rules cap what anyone else can page through. The file opens with a recruiting joke for humans reading it, and it closes the door on machines reading past page two of reviews and page five of salaries.

Whose terms these are: Indeed, Inc.

Glassdoor's Terms of Use, revised 1 July 2026, open with the operator: Indeed, Inc. provides the services Glassdoor.com and Fishbowlapp.com. The House Rules require use solely for lawful purposes consistent with the Terms and incorporate the Community Guidelines. Glassdoor's own security notice says the rest in one line: Glassdoor has been built on the contributions of real employees and job seekers, and it uses advanced security systems to prevent misuse or unauthorized access. The same operator's Site Rules for Indeed say do not access the site through any means other than its public interfaces, and do not access any data, especially personal data, by automated means without permission.

So be accurate about what a proxied fetch is. The home page is served to any visitor, and reading what a server hands the public is a technical fact. Reviews and salaries are contributed by identifiable employees, and Glassdoor's model is that you contribute to read; the login wall the browser lands on is that model enforced. Permission to collect and reuse those contributions is contractual, and Glassdoor has not published a data API for it. Anyone telling you a residential IP settles that question is selling you an IP. Where the honest answer is that the terms forbid it, the answer is that the terms forbid it, and the next section is about what you can build instead.

What to spend on instead of the browser

There is no Glassdoor collector in our catalogue, and given the terms and the login wall we are not building one. The jobs are a different matter, because the same operator publishes them on Indeed, and we ran the Indeed jobs collector the same afternoon: ten "data engineer" listings in New York in 7.3 seconds, each with the salary parsed into salary_min, salary_max and a period, the company's rating and review count, and a sponsored flag. Three of the ten were sponsored, and the ratings came from the same review pool that Glassdoor gates. For a second source with the same input shape, the Google Jobs collector returns listings aggregated across boards. And for the page as it is, the web scraping API over plain HTTP returns the 624 words as Markdown at $0.0002 a page and never renders unless you ask it to. We measured the operator's own board in Indeed through residential proxies, and it points the same way: plain HTTP wins, the browser loses words.

Cost: what the browser bill buys on Glassdoor

Prices from our pricing page: residential Basic $0.80/GB, the web scraping API from $0.0002 per page and $0.001 rendered, the jobs collector $0.001 per delivered listing. A gigabyte is counted as 10^9 bytes.

ApproachBytes per pagePages per GBCost per 1,000What you get
Rendered, residential Basic705,0711,418$0.5654 words and a login form, 61 seconds each
Web scraping API, renderedn/an/a$1.00The same login form as Markdown
Web scraping API, plain fetchn/an/a$0.20624 words, h1, canonical, JSON-LD
Indeed jobs collectorn/an/a$1.001,000 listings with parsed salaries, same operator

Every rendered row in that table is money spent to reach a form. The plain fetch is the only row that returns the page, and it is also the cheapest. Byte figures are decoded bodies and exclude TLS overhead. If you are sizing a job before you buy bandwidth, how much proxy data you need does the arithmetic from the other side; on Glassdoor the answer is "less than the browser guide told you".

Where we stop

Everything above is about the public home page and public listings, for employer-brand monitoring, compensation research at the level Glassdoor publishes to visitors, and checking how an edition reads from its own market. It does not cover accounts. We will not help with creating accounts to pass the contribute-to-read wall, submitting reviews or salaries you did not experience, or automating a logged-in session. Glassdoor's Community Guidelines govern contributions, its operator's rules forbid fake accounts and automated account creation, and reviews are the words of identifiable employees, which makes them personal data wherever your users live. A proxy does not change who wrote a review or whether you were allowed to keep it.

The setting that works on Glassdoor

  • Network: residential proxies, Basic line at $0.80/GB, one exit per edition. The plain fetch returned 624 words from a US exit and 589 from a UK one; nothing in the plain path rewarded a mobile address at $2.30/GB or an ISP address at $2.50 per IP per month.
  • Fetch mode: engine: tls, plain HTTP, never rendered. 624 words with h1, canonical and Organization JSON-LD; rendered, 54 words and a canonical of /member/profile/login for 705,071 bytes.
  • Country: pin the exit to the edition, country=us for glassdoor.com and country=gb for glassdoor.co.uk. The plain word count moves by 35 between editions; the rendered one loses 570.
  • When the proxy is not enough: there is no Glassdoor collector. For listings with salaries as numbers use the operator's board through the indeed_jobs collector, 10 rows in 7.3 seconds at $0.001 per listing; for the page as Markdown, the web scraping API at $0.0002 per page, plain fetch.
  • Free tier: every account gets $2 of free API usage per month, which is 10,000 plain-fetch pages or 2,000 parsed listings before you pay anything.

Sources & further reading

FAQ

Quick answers on glassdoor proxies.

Something else? Ask us →

Do I need a headless browser to scrape Glassdoor?

No, and it is the one thing that loses you the page. On 26 September 2026 the home page returned 624 words to a plain HTTP client and 54 words to a rendered browser, whose canonical became /member/profile/login. On 28 September the UK edition gave a plain client 589 words and a rendered browser 600 words with no canonical and no page. Fetch it plain.

Why does my rendered Glassdoor page turn into the login form?

Because Glassdoor's own scripts send a browser without a member session there, and the query string of the URL it lands on names the browser as the reason. The weighed render cost 705,071 bytes and 61 seconds to reach a form with 44 words. Classify on the canonical: /member/profile/login means you rendered the wrong page, not that the item is missing.

Does the exit country change what Glassdoor returns?

The edition is chosen by domain, not by IP, and each edition sets currency and review pool. Over plain HTTP glassdoor.com returned 624 words from a US exit and glassdoor.co.uk 589 from a UK exit, both with an h1 and JSON-LD; rendered, both lost the page. Pin the exit to the edition you fetch.

Is scraping Glassdoor allowed?

Glassdoor's robots.txt disallows most pagination for everyone, reviews past page 2 and salaries past page 5, and blocks AI crawlers from reviews, salaries and interviews by name. Its Terms of Use, published by its operator Indeed, Inc. and revised 1 July 2026, require lawful use under the House Rules, and Indeed's Site Rules forbid accessing data by automated means without permission. Reviews are contributions from identifiable employees.

How do I get job listings with salaries if there is no Glassdoor collector?

Use the same operator's board. The Indeed jobs collector returned 10 "data engineer" listings in New York in 7.3 seconds with salary_min, salary_max and period parsed, plus the company rating and review count, at $0.001 per listing. The Google Jobs collector takes the same query and location for a second source.

What does rendering Glassdoor cost compared with a plain fetch?

At $0.80/GB residential Basic, a rendered page is 705,071 bytes, 1,418 pages per gigabyte, $0.56 per thousand pages, each ending on a login form after 61 seconds. The web scraping API fetches the plain page, 624 words, for $0.0002, five times less than its rendered rate and with the page still there.

Fetch it plain, and keep the words

Every number here came from one platform: fetch any URL over plain HTTP or through a browser, compare the two views, and see which one still has the page. Every account gets $2 of free API usage each month, and failed requests are never billed.

Related reading