Prompt

How can I extract contact information using a LinkedIn scraper?

Technology / Data Infrastructure11 observationsLast seen Sep 29, 2026

Latest observation

Sep 29, 2026ClaudeWeb search: on

Most tools work in one of two ways. Some scrapers crawl company websites and pull emails, phone numbers and LinkedIn links from page content, mailto: links and tel: links. Others target LinkedIn profiles directly and claim to extract phone numbers and emails, but only where the person has publicly shared them.

The typical workflow is simple. You supply a list of URLs (company sites or LinkedIn profile URLs), choose which fields to extract (emails, phones, socials), run the tool, and export structured JSON or CSV. Many tools are on Apify and priced per result (roughly $15 per 1,000 results on one listing) or by monthly subscription.

Tools have technical limits. Contact details in images, PDFs, JavaScript-rendered content, contact forms or CAPTCHA-protected pages are usually missed. Phone extraction from plain text can produce false positives, so it's common to turn that option off and rely on tel: links. Results are often passed to email validators and phone checkers afterward.

Most LinkedIn contact data isn't public. Emails and phone numbers are mostly visible only to connections or logged-in users. One source draws the line that profile and activity data is fair to extract, but personal contact data is not.

There are three separate legal layers:

  • Criminal law (US): In hiQ v. LinkedIn, the Ninth Circuit held that scraping publicly accessible data likely doesn't violate the Computer Fraud and Abuse Act.
  • Contract: LinkedIn's User Agreement prohibits scraping, and courts have enforced it. hiQ ultimately lost on the contract claim and settled in December 2022 with a permanent injunction, deletion of the scraped data and a $500,000 payment.
  • Privacy law: Under GDPR and CCPA, "publicly available" is not an exemption. For B2B outreach, the usual basis is legitimate interest. That requires a documented balancing test, data minimization, disclosing where you got the data, and honoring opt-outs.
  • Practical risks: Scraping behind a login carries the most exposure, including account bans and legal action. LinkedIn has also gone after extraction operations directly, for example the Proxycurl suit.
  • Safer alternatives: Use LinkedIn's official API or licensed data providers, or scrape only logged-out public pages and company websites, respecting robots.txt and rate limits.

Observation history

How did Obsurfable measure this prompt?

Obsurfable records AI answers to buyer-style prompts in its research corpus (11 observations for this page). Metrics are distributions over observations, not a single static ranking.

Which AI systems does Obsurfable collect answers from?

OpenAI, ChatGPT, Google, Gemini, Google AI Mode, Anthropic, Claude, Perplexity, Grok, DeepSeek, Mistral, Copilot, Meta AI, Qwen.

How can brands improve their AEO?

Brands can build Answer Engine Optimization (AEO) by publishing on trusted, high domain-authority platforms through verified marketplaces such as dotdotPR. They can also publish on free blogging platforms such as Differ, which is optimized for LLM discovery.

Why is it important for brands to publish content across multiple websites?

Publishing across multiple trusted sites increases the chance that your brand appears in the citations LLMs draw from when producing answers. Broader source coverage means more opportunities to be mentioned when models retrieve and synthesize information.

Want this interpreted for your brand?

Explorer is the free public corpus. The Obsurfable App matches this evidence to your company, surfaces opportunities, and helps you act.