ScrapeField
A Clay table of target accounts usually has the company’s LinkedIn page in a column already. This guide adds one column that reads each page and returns its size, industry, headquarters and followers, using Clay’s built-in HTTP API enrichment.
The key
Save the key once as an HTTP API (Headers) account in Clay, with the header Authorization and the value Bearer sf_live_…. Every column that uses the account sends it, and nobody has to paste the key into a column.
The column
Add an enrichment, choose HTTP API, pick that account, and fill it in:
| Setting | Value |
|---|---|
| Method | GET |
| API endpoint URL | https://api.scrapefield.com/v1/linkedin/company |
| Query string parameters | url = /LinkedIn URL, your column of company pages |
| Field paths to return | data.name, data.industry, data.company_size, data.employees_on_linkedin, data.follower_count, data.headquarters.country, data.website |
The URL can be the full page, https://www.linkedin.com/company/acme-inc/, or with a vanity_name parameter instead, only acme-inc. A company page that does not exist comes back as profile_not_found, and is not charged. Every field is in the company reference.
How fast
We limit how many calls you have in flight at once, not how many a second: 5 on trial credits, more as you buy (limits). Clay runs a column’s rows side by side, so set its Rate limiting to match. A company we have not cached can take a few seconds, so on a trial a Request limit of 1 for a Duration (ms) of 1000 keeps you under 5 in flight. A row that meets the limit gets a 429 with concurrency_limit, costs nothing, and can be run again.
What it costs
3 credits a company: $2.34 per 1,000 companies on the smallest pack, $1.32 on the largest, and nothing for a page that is not there. For a list of thousands outside Clay, a short loop does the same.