apify-booking-host-leads

por apify

Encuentra y enriquece leads B2B de Booking.com: hoteles, apartamentos y alquileres vacacionales, y extrae los datos de contacto reales de cada anfitrión o administrador de propiedades (correo electrónico,…

npx skills add https://github.com/apify/awesome-skills --skill apify-booking-host-leads

Booking.com host & operator lead finder

Turn a Booking.com destination search into a clean table of accommodation hosts / property managers with their real contact details.

The key insight

When you scrape Booking.com, the host's contact email is sometimes already in the data under traderInfo (a trader-transparency disclosure). But this is region-dependent: strong in the EU (Zurich apartments: 80%) and Australia (Melbourne apartments: 82% in a 50-property test), and increasingly present in Asia (Beijing apartments: 65% as of June 2026, typically a QQ, 163, or Gmail address plus a Chinese company name). It is historically sparser in the US and UK. A naive chained Google Maps email scraper found emails for only ~10% and often matched the wrong business.

So use a waterfall: try the cheapest source first, fall through to more expensive ones only for leads still missing a contact. Never start with a Google Maps email extractor.

TierSourceGetsFires when
1Booking traderInfoemail, phone, companyalways (free, already scraped)
2Google Maps Email Extractor (by name)email, phone, socials, websitelead still has no email
3Google Search (by name)a candidate websitelead still has no website
4Website contact scraper (on the website)email / phone / socialslead has a website but still no email

Prerequisites

Workflow

  1. Collect inputs. Ask the user for: destination (city/area), property type (Apartments / Guest houses / Holiday homes / Villas surface the most independent operators; none = all), max properties, and output format (CSV / JSON / Google Sheet).
  2. Tier 1 - Scrape Booking.com. Run voyager/booking-scraper. Build a lead row per property from traderInfo and hostInfo (see Mapping below). Mark rows where traderInfo.email is present as contactSource: booking-trader.
  3. Tier 2 - Google Maps (only rows still missing an email). Run lukaskrivka/google-maps-with-contact-details with searchStringsArray = ["<name>, <destination>", ...]. Fill email/phone/socials and capture any website it returns. Mark filled rows google-maps.
  4. Tier 3 - Google Search (only rows still missing a website). Run apify/google-search-scraper with one query per lead ("<name> <city> official website"). Take the first non-OTA/non-social organicResults[].url as the candidate website.
  5. Tier 4 - scrape the website (only rows with a website but no email). Default: vdrmota/contact-info-scraper (deterministic, ~14s/site). Optional: apify/ai-web-scraper with a prompt for email/phone/contact-form URL - but in live testing it was slow and failed on multiple sites, so prefer the contact scraper. Mark filled rows website-contact (or website-ai). Cap the count to control cost.
  6. Deliver. Output one row per property. Report total rows, % with an email, and the contactSource breakdown. Rows still empty keep BookingURL (contact via Booking messaging).

Mapping (Booking item → lead row)

Output columnSource field
name, type, address, city, country, rating, ratingLabel, stars, reviews, BookingURLtop-level Booking fields (address.full, address.city, address.country, rating, ratingLabel, reviews, url)
email, phonetraderInfo.email / traderInfo.phone → fallback Google Maps emails[0] / phones[0]
companyName, registrationNumber, traderAddresstraderInfo.companyName / .registrationNumber / .address
hostName, managedPropertieshostInfo.name / hostInfo.managedProperties (operator-size signal)
website, socialsGoogle Maps fallback only; drop OTA domains (booking.com, agoda, expedia, etc.)
contactSourcebooking-trader | google-maps | none

Put email and phone immediately after address in the output.

Always return BookingURL (link to the property), rating, ratingLabel (e.g. "Superb", "Exceptional"), and reviews (review count) for every lead. These give each row its own quality signal, so leads can be ranked and prioritized without re-opening Booking.

Actor routing

Waterfall tierActor IDMaintainerBest for
1 - Booking + trader contactsvoyager/booking-scrapercommunityPrimary. traderInfo holds host email/phone/company
2 - email/social/website by namelukaskrivka/google-maps-with-contact-detailscommunityEmails Booking doesn't disclose; also returns a website
3 - discover a websiteapify/google-search-scraperapifyFinding the host's official site when Maps has none
4 - extract from the website (default)vdrmota/contact-info-scrapercommunityDeterministic email/phone/social extraction; fast and reliable
4 - extract from the website (alt)apify/ai-web-scraperapifyLLM extraction via prompt; flexible but slow/unreliable in testing

Prefer apify-maintained Actors where available.

Calling Actors — Apify CLI

Three flags on every call (--json, --user-agent, 2>/dev/null):

# 1. Scrape Booking.com
apify actors call "voyager/booking-scraper" \
  -i '{"search":"Melbourne","maxItems":50,"propertyType":"Apartments","currency":"USD","language":"en-us","sortBy":"distance_from_search","rooms":1,"adults":2}' \
  --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads \
  2>/dev/null

# Tier 2 - enrich names that have no traderInfo email
apify actors call "lukaskrivka/google-maps-with-contact-details" \
  -i '{"searchStringsArray":["<name>, Melbourne"],"locationQuery":"Melbourne","maxCrawledPlacesPerSearch":1,"language":"en"}' \
  --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads \
  2>/dev/null

# Tier 3 - discover a website for leads that still have none
apify actors call "apify/google-search-scraper" \
  -i '{"queries":"<name> Melbourne official website","maxPagesPerQuery":1,"resultsPerPage":5,"countryCode":"us","languageCode":"en"}' \
  --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads \
  2>/dev/null

# Tier 4 - AI-extract a contact email / form URL from the website
# (confirm the prompt/instructions field name against the Actor's input schema)
apify actors call "apify/ai-web-scraper" \
  -i '{"startUrls":[{"url":"<website>"}],"prompt":"Extract the contact email, phone, and contact-form URL. Return JSON keys: email, phone, contactFormUrl."}' \
  --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads \
  2>/dev/null

# Inspect input schema / fetch results
apify actors info "voyager/booking-scraper" --input --json \
  --user-agent apify-awesome-skills/apify-booking-host-leads 2>/dev/null
apify datasets get-items DATASET_ID --format json \
  --user-agent apify-awesome-skills/apify-booking-host-leads 2>/dev/null

The Apify MCP server (https://mcp.apify.com) and any MCP client (e.g. mcpc) are equivalent alternatives.

Troubleshooting

  • No emails returned → you are probably reading Google Maps output instead of traderInfo. Re-check the Booking item's traderInfo field first.
  • Emails for the wrong business → that is the Google Maps fallback matching a generic place. Trust contactSource: booking-trader rows; treat google-maps rows as lower confidence.
  • Many contactSource: none rows → those hosts disclosed no trader info and have no Google Business match. Use the BookingURL to contact via Booking, or accept the blank.
  • Run cost climbs → skip the Google Maps fallback entirely; traderInfo alone covers ~80%.
  • See references/gotchas.md for cost notes.

Más skills de apify

bug-triage
apify
Triage de issues de bugs abiertos en apify/apify-mcp-server. Analizar, redactar respuestas, obtener aprobación, publicar.
official
apify-influencer-brand-collabs
apify
Descubre asociaciones entre marcas y creadores de Instagram encadenando Apify Actors. Úsalo cuando el usuario pregunte quién colabora con una marca, qué marcas ha hecho pagadas un creador…
official
dig
apify
Habilidad flexible para explorar, planificar y especificar trabajo en el servidor Apify MCP. NO edites archivos fuente — esta habilidad es solo para comprensión y planificación.
official
apify-financial-news
apify
Descubre y extrae noticias financieras para empresas de cartera rastreadas en 33 fuentes verificadas de primer nivel (Bloomberg, Reuters, FT, WSJ, IntelliNews, ČTK, PAP, BTA,…
official
apify-actor-development
apify
Crear, depurar e implementar programas en la nube sin servidor para web scraping, automatización y procesamiento de datos. Compatible con plantillas de JavaScript, TypeScript y Python con las bibliotecas integradas Crawlee, Playwright y Cheerio para rastreo HTTP y basado en navegador. Incluye pruebas locales mediante apify run con almacenamiento aislado, validación de esquemas para entradas/salidas e implementación en la plataforma Apify mediante apify push. Requiere autenticación de Apify CLI y metadatos obligatorios generatedBy en .actor/actor.json para IA...
official
apify-actorization
apify
Convierte proyectos existentes en Apify Actors sin servidor con integración de SDK específica para cada lenguaje. Soporta JavaScript/TypeScript (con Actor.init() / Actor.exit()), Python (gestor de contexto asíncrono) y cualquier lenguaje mediante envoltorio CLI. Proporciona un flujo de trabajo estructurado: apify init para crear la estructura, aplicar envoltorio SDK, configurar esquemas de entrada/salida, probar localmente con apify run y luego desplegar con apify push. Incluye validación de esquemas de entrada y salida, contenedorización Docker y opcional pago por evento...
official
apify-generate-output-schema
apify
Generar esquemas de salida (dataset_schema.json, output_schema.json, key_value_store_schema.json) para un Apify Actor analizando su código fuente. Úselo cuando…
official
apify-ultimate-scraper
apify
Raspador web automatizado que selecciona los Actores óptimos para más de 55 plataformas, incluyendo Instagram, TikTok, YouTube, Facebook, Google Maps y más. Cubre más de 55 Actores preconfigurados en 8 plataformas principales con orientación de selección específica para casos de uso (generación de leads, descubrimiento de influencers, monitoreo de marca, análisis de competencia, investigación de tendencias). Admite tres formatos de salida: visualización rápida en chat, exportación CSV o exportación JSON con límites de resultados personalizables. Incluye patrones de flujo de trabajo con múltiples Actores para casos complejos...
official