apify-booking-host-leads

作者: apify

從Booking.com尋找並豐富B2B潛在客戶——包括酒店、公寓和度假租賃——並提取每位房東或物業管理者的真實聯絡資訊(電子郵件、……)

npx skills add https://github.com/apify/awesome-skills --skill apify-booking-host-leads

Booking.com host & operator lead finder

Turn a Booking.com destination search into a clean table of accommodation hosts / property managers with their real contact details.

The key insight

When you scrape Booking.com, the host's contact email is sometimes already in the data under traderInfo (a trader-transparency disclosure). But this is region-dependent: strong in the EU (Zurich apartments: 80%) and Australia (Melbourne apartments: 82% in a 50-property test), and increasingly present in Asia (Beijing apartments: 65% as of June 2026, typically a QQ, 163, or Gmail address plus a Chinese company name). It is historically sparser in the US and UK. A naive chained Google Maps email scraper found emails for only ~10% and often matched the wrong business.

So use a waterfall: try the cheapest source first, fall through to more expensive ones only for leads still missing a contact. Never start with a Google Maps email extractor.

TierSourceGetsFires when
1Booking traderInfoemail, phone, companyalways (free, already scraped)
2Google Maps Email Extractor (by name)email, phone, socials, websitelead still has no email
3Google Search (by name)a candidate websitelead still has no website
4Website contact scraper (on the website)email / phone / socialslead has a website but still no email

Prerequisites

Workflow

  1. Collect inputs. Ask the user for: destination (city/area), property type (Apartments / Guest houses / Holiday homes / Villas surface the most independent operators; none = all), max properties, and output format (CSV / JSON / Google Sheet).
  2. Tier 1 - Scrape Booking.com. Run voyager/booking-scraper. Build a lead row per property from traderInfo and hostInfo (see Mapping below). Mark rows where traderInfo.email is present as contactSource: booking-trader.
  3. Tier 2 - Google Maps (only rows still missing an email). Run lukaskrivka/google-maps-with-contact-details with searchStringsArray = ["<name>, <destination>", ...]. Fill email/phone/socials and capture any website it returns. Mark filled rows google-maps.
  4. Tier 3 - Google Search (only rows still missing a website). Run apify/google-search-scraper with one query per lead ("<name> <city> official website"). Take the first non-OTA/non-social organicResults[].url as the candidate website.
  5. Tier 4 - scrape the website (only rows with a website but no email). Default: vdrmota/contact-info-scraper (deterministic, ~14s/site). Optional: apify/ai-web-scraper with a prompt for email/phone/contact-form URL - but in live testing it was slow and failed on multiple sites, so prefer the contact scraper. Mark filled rows website-contact (or website-ai). Cap the count to control cost.
  6. Deliver. Output one row per property. Report total rows, % with an email, and the contactSource breakdown. Rows still empty keep BookingURL (contact via Booking messaging).

Mapping (Booking item → lead row)

Output columnSource field
name, type, address, city, country, rating, ratingLabel, stars, reviews, BookingURLtop-level Booking fields (address.full, address.city, address.country, rating, ratingLabel, reviews, url)
email, phonetraderInfo.email / traderInfo.phone → fallback Google Maps emails[0] / phones[0]
companyName, registrationNumber, traderAddresstraderInfo.companyName / .registrationNumber / .address
hostName, managedPropertieshostInfo.name / hostInfo.managedProperties (operator-size signal)
website, socialsGoogle Maps fallback only; drop OTA domains (booking.com, agoda, expedia, etc.)
contactSourcebooking-trader | google-maps | none

Put email and phone immediately after address in the output.

Always return BookingURL (link to the property), rating, ratingLabel (e.g. "Superb", "Exceptional"), and reviews (review count) for every lead. These give each row its own quality signal, so leads can be ranked and prioritized without re-opening Booking.

Actor routing

Waterfall tierActor IDMaintainerBest for
1 - Booking + trader contactsvoyager/booking-scrapercommunityPrimary. traderInfo holds host email/phone/company
2 - email/social/website by namelukaskrivka/google-maps-with-contact-detailscommunityEmails Booking doesn't disclose; also returns a website
3 - discover a websiteapify/google-search-scraperapifyFinding the host's official site when Maps has none
4 - extract from the website (default)vdrmota/contact-info-scrapercommunityDeterministic email/phone/social extraction; fast and reliable
4 - extract from the website (alt)apify/ai-web-scraperapifyLLM extraction via prompt; flexible but slow/unreliable in testing

Prefer apify-maintained Actors where available.

Calling Actors — Apify CLI

Three flags on every call (--json, --user-agent, 2>/dev/null):

# 1. Scrape Booking.com
apify actors call "voyager/booking-scraper" \
  -i '{"search":"Melbourne","maxItems":50,"propertyType":"Apartments","currency":"USD","language":"en-us","sortBy":"distance_from_search","rooms":1,"adults":2}' \
  --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads \
  2>/dev/null

# Tier 2 - enrich names that have no traderInfo email
apify actors call "lukaskrivka/google-maps-with-contact-details" \
  -i '{"searchStringsArray":["<name>, Melbourne"],"locationQuery":"Melbourne","maxCrawledPlacesPerSearch":1,"language":"en"}' \
  --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads \
  2>/dev/null

# Tier 3 - discover a website for leads that still have none
apify actors call "apify/google-search-scraper" \
  -i '{"queries":"<name> Melbourne official website","maxPagesPerQuery":1,"resultsPerPage":5,"countryCode":"us","languageCode":"en"}' \
  --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads \
  2>/dev/null

# Tier 4 - AI-extract a contact email / form URL from the website
# (confirm the prompt/instructions field name against the Actor's input schema)
apify actors call "apify/ai-web-scraper" \
  -i '{"startUrls":[{"url":"<website>"}],"prompt":"Extract the contact email, phone, and contact-form URL. Return JSON keys: email, phone, contactFormUrl."}' \
  --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads \
  2>/dev/null

# Inspect input schema / fetch results
apify actors info "voyager/booking-scraper" --input --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads 2>/dev/null
apify datasets get-items DATASET_ID --format json \
  --user-agent apify-awesome-skills/apify-booking-host-leads 2>/dev/null

The Apify MCP server (https://mcp.apify.com) and any MCP client (e.g. mcpc) are equivalent alternatives.

Troubleshooting

  • No emails returned → you are probably reading Google Maps output instead of traderInfo. Re-check the Booking item's traderInfo field first.
  • Emails for the wrong business → that is the Google Maps fallback matching a generic place. Trust contactSource: booking-trader rows; treat google-maps rows as lower confidence.
  • Many contactSource: none rows → those hosts disclosed no trader info and have no Google Business match. Use the BookingURL to contact via Booking, or accept the blank.
  • Run cost climbs → skip the Google Maps fallback entirely; traderInfo alone covers ~80%.
  • See references/gotchas.md for cost notes.

來自 apify 的更多技能

bug-triage
apify
分類處理 apify/apify-mcp-server 上的未解決錯誤問題。分析、草擬回覆、取得批准、發布。
official
apify-influencer-brand-collabs
apify
探索Instagram品牌與創作者的合作關係,透過串聯Apify Actors。當使用者詢問某品牌與誰合作、某創作者曾與哪些品牌進行付費合作時使用…
official
dig
apify
用於在 Apify MCP 伺服器上探索、規劃與規格化工作的靈活技能。請勿編輯原始檔案——此技能僅供理解與規劃使用。
official
apify-financial-news
apify
探索並提取追蹤投資組合公司在33個經認證的一級來源(彭博、路透、金融時報、華爾街日報、IntelliNews、ČTK、PAP、BTA…)中的財經新聞。
official
apify-actor-development
apify
建立、除錯及部署無伺服器雲端程式,用於網頁爬取、自動化及資料處理。支援 JavaScript、TypeScript 及 Python 範本,內建 Crawlee、Playwright 與 Cheerio 函式庫,適用於 HTTP 及瀏覽器爬取。包含透過 apify run 進行本地測試(具備隔離儲存)、輸入/輸出結構驗證,以及透過 apify push 部署至 Apify 平台。需進行 Apify CLI 驗證,並在 .actor/actor.json 中強制加入 generatedBy 元資料以供 AI 使用...
official
apify-actorization
apify
將現有專案轉換為無伺服器 Apify Actors,並整合語言專屬 SDK。支援 JavaScript/TypeScript(使用 Actor.init() / Actor.exit())、Python(非同步上下文管理器),以及透過 CLI 包裝器的任何語言。提供結構化工作流程:使用 apify init 建立專案骨架、套用 SDK 包裝、設定輸入/輸出架構、以 apify run 進行本地測試,再透過 apify push 部署。包含輸入與輸出架構驗證、Docker 容器化,以及可選的按事件付費...
official
apify-generate-output-schema
apify
為 Apify Actor 分析其原始碼,生成輸出結構(dataset_schema.json、output_schema.json、key_value_store_schema.json)。用於…
official
apify-ultimate-scraper
apify
自動化網頁爬蟲,為55多個平台選擇最佳Actor,包括Instagram、TikTok、YouTube、Facebook、Google地圖等。涵蓋8大主要平台的55多個預配置Actor,並提供針對特定使用案例的選擇指引(潛在客戶開發、網紅發現、品牌監控、競爭對手分析、趨勢研究)。支援三種輸出格式:快速聊天顯示、CSV匯出或JSON匯出,並可自訂結果數量限制。包含多Actor工作流程模式,適用於複雜...
official