Extract competitor and customer intelligence from any company's landing page HTML. Discovers tech stack, analytics tools, ad pixels, customer logos, SEO metadata, CTAs, hidden elements, and more. No API keys required.
npx gooseworks install --all # then, in Claude Code, Cursor, or Codex: /gooseworks use the landing-page-intel skill
Extract GTM-relevant intelligence from any company's landing page by scraping its HTML source.
Only dependency is pip install requests. No API key needed.
# Basic scan of a single URL
python3 skills/landing-page-intel/scripts/scrape_landing_page.py \
--url "https://example.com"
# Scan multiple pages of the same site
python3 skills/landing-page-intel/scripts/scrape_landing_page.py \
--url "https://example.com" --pages "/,/pricing,/about"
# Output as summary table instead of JSON
python3 skills/landing-page-intel/scripts/scrape_landing_page.py \
--url "https://example.com" --output summary
# Save full report to file
python3 skills/landing-page-intel/scripts/scrape_landing_page.py \
--url "https://example.com" --output json > report.json| Category | Details |
|---|---|
| Tech Stack | Analytics (GA4, Mixpanel, Amplitude, PostHog, Heap), marketing automation (HubSpot, Marketo, Pardot), chat widgets (Intercom, Drift, Crisp, Zendesk), A/B testing (Optimizely, VWO, LaunchDarkly), session recording (Hotjar, FullStory, LogRocket), CDPs (Segment, Clearbit, 6sense) |
| Ad Pixels | Meta Pixel, Google Ads, LinkedIn Insight Tag, TikTok pixel, Twitter pixel |
| Customer Logos | Image URLs from "trusted by" / logo carousel sections, grouped by directory |
| SEO Metadata | Title, meta description, Open Graph tags, Twitter Cards, canonical URL, structured data (JSON-LD), hreflang tags |
| CTAs & Sales Motion | All CTA button text and links — reveals PLG vs sales-led motion |
| Social Proof | Testimonials, customer counts, case study links, badge images |
| Integrations | Links to integration/partner pages, embedded third-party widgets |
| Hidden Elements | Content in display:none, hidden, or HTML comments that may reveal upcoming features |
| Infrastructure | CMS platform (Webflow, WordPress, Next.js, etc.), detected from HTML signatures |
| Flag | Default | Description |
|---|---|---|
--url | required | Target website URL |
--pages | / | Comma-separated paths to scan (e.g., /,/pricing,/about) |
--output | json | Output format: json or summary |
--timeout | 15 | Request timeout in seconds |
Free. No API keys required. Uses only HTTP requests to fetch public HTML.
Discover rising category conversations, formats, sounds, questions, and creator patterns across social platforms, then separate durable demand signals from short-lived noise.
Turn TikTok, Instagram, YouTube, Facebook, X, LinkedIn, Reddit, or Rumble transcripts into timestamped hooks, claims, objections, proof, calls to action, sponsorship signals, and reusable content atoms. Use directly for transcript analysis or as support for creator, competitor, trend, demand, and repurposing work.
Produce a decision-ready brief of current brand, product, category, and competitor conversations across social platforms, including sentiment drivers, questions, risks, and growth opportunities.