Tutorials
Step-by-step tutorials on Python scrapers, Reddit APIs, Google Trends, RAG pipelines, and job data. Real production code, tested patterns.
98 articles
Pull the Top 100 Best Sellers for Every Amazon Subcategory in One Run
Walk an Amazon department with Amazon Bestsellers Scraper and get the top 100 of every subcategory in one run, with a hard cap on lists and cost.
Benchmark LinkedIn Post Engagement Across Company Pages
Compare reactions and comments per post across LinkedIn company pages, scaled by follower count, with the LinkedIn Company Posts Scraper. No login.
Build a Labeled Image Dataset From Google Images Search Terms
Turn one search term per class into a labeled image manifest with Google Images Scraper: full size URLs, width, height and the source page to check.
Build a List of Funded Startups That Are Hiring on Wellfound
Turn Wellfound job listings into one row per startup with website, stage, total raised and latest round, using the Wellfound Jobs Scraper.
Compare an App's App Store Reviews Across Countries
Read one app's App Store reviews in many countries in one run with App Store Reviews Scraper, then compare ratings and top complaints country by country.
Compare Salary Bands in Lakhs by City and Experience on Foundit
Build a salary table for one role across Indian cities and experience bands from Foundit.in listings that publish pay, and drop the placeholder figures.
Compare Pay for a Role by Experience Level on Shine
Run one Shine.com search per experience band, keep the listings that show pay, and turn them into salary ranges in lakhs for each band.
Compare Startup Salary and Equity Ranges by Role on Wellfound
Pull Wellfound jobs for several roles and cities, then compute median salary and equity ranges per role with the Wellfound Jobs Scraper and Python.
Post a Daily Feed of New Tech Jobs to Slack or Telegram
Send each morning's new Hirist.tech jobs for your stack and cities to a Slack or Telegram channel, with repeats filtered out by job id, using Python.
Export Every Image and Caption From Instagram Carousel Posts
Export every slide image link, slide type and caption from Instagram carousel posts to a CSV, by profile or by post link, with Instagram Post Scraper.
Export a Creator's Instagram Reels With Play Counts to a Spreadsheet
Export one creator's reels with plays, likes, comments, length and audio to a CSV, then compute median plays to vet them, with Instagram Reel Scraper.
Export a YouTube Video's Comments and Replies to CSV
Pull a YouTube video's comments and their replies into one CSV with YouTube Comments Scraper, each thread rebuilt from parent_comment_id.
Find the Decision Makers at Your Target Accounts by Job Title
Turn a list of company domains into a CRM-ready CSV of people by job title, with LinkedIn URLs, labeled emails and a company check on every row.
Find Entry-Level AI and ML Jobs in Bangalore Posted This Week
Pull fresh AI and ML jobs for 0 to 2 years of experience from Hirist.tech, merge several searches and rank them by seniority and age with Python.
Find Paid Partnership Posts on Any Instagram Profile
List the paid partnership and branded posts on any public Instagram profile, and the brands behind them, using Instagram Post Scraper and Python.
Find the Skills Employers Ask for Most in a Role on Foundit
Count the skill tags on Foundit.in (Monster India) jobs for one role and city, merge spelling variants, and rank skills by how many employers want them.
Find the Highest-Reach TikTok Ads in Any EU Country
Rank the TikTok ads that reached the most people in an EU country over any date range, with per-country reach, using the TikTok Ads Library Scraper.
Map Which Companies Are Hiring for a Tech Stack Across Indian Cities
Turn Hirist.tech job listings into a table of employers by city for one tech stack, with job counts, ratings and top skills, using Python.
Monitor New App Store Reviews for Your App Every Day
Get each day's new App Store reviews of your app in several countries with App Store Reviews Scraper, plus an alert for every 1 and 2 star review.
Monitor Competitor LinkedIn Company Posts on a Schedule
Get every new post from competitor LinkedIn company pages each day or week, with only-new mode and an Apify schedule. No LinkedIn login or cookies.
Monitor New LinkedIn Ads for a Keyword Every Week
Get a weekly list of new LinkedIn ads that mention your keywords, from any advertiser, with the LinkedIn Ad Library Scraper and only-new mode.
Monitor the Newest Comments on Your YouTube Videos Every Day
Read the newest comments on your YouTube videos every day with YouTube Comments Scraper, keep only unseen ones, and flag questions you have not answered.
See a Competitor's TikTok Ad Targeting and Reach in the EU
Pull every TikTok ad a competitor ran in the EU, with interests, age groups, audience size and reach by country, using the TikTok Ads Library Scraper.
See Every LinkedIn Ad a Competitor Is Running, With Impressions
List a competitor's LinkedIn ads with landing pages, run dates, impressions by country and targeting, using the LinkedIn Ad Library Scraper. No login.
Track Daily Rank Changes in an Amazon Best Sellers Category
Run Amazon Bestsellers Scraper every morning, diff each ASIN's rank against yesterday, and flag new entrants, drop outs and big movers in Python.
Track Competitor Instagram Reels Weekly With View Counts
Log every new reel your competitors post each week, with plays, likes and audio, using Instagram Reel Scraper on a weekly Apify schedule.
Track New Foundit Job Postings Every Day With Monitor Mode
Get only the Foundit.in (Monster India) jobs you have not seen before, every morning, and pay only for the new rows. Input, Python code and real costs.
Track New People at Your Target Accounts Every Week
Schedule B2B Leads Finder in monitor mode so each weekly run returns only people you have not been sent before, and never bills the same person twice.
Track New Shine.com Jobs for a Role Every Day
Get only the Shine.com jobs posted since your last check, for one role and city, using freshness sort, monitor mode and a daily schedule.
Track Which Sites Rank in Google Images for Your Keywords
Use Google Images Scraper to see which domains hold the top image positions for your keywords in each country, then rerun weekly to catch movement.
GitHub Repo Intelligence: Vet Any Dependency Before You Add It
Pull stars, forks, contributors, commit recency, latest release and full README for any public GitHub repo in one call. No token required, MCP-ready for agents doing dependency checks.
Bilibili Scraper: Video, Creator and Search Data in Python
Pull Bilibili videos by keyword, URL or creator ID: titles, view and like counts, duration, author stats. Plain HTTP, no cookies, no login required.
SEEK Jobs Scraper: Australian Job Listings in Python
How to pull SEEK.com.au job listings by keyword, location, work type and salary band in Python: title, company, salary range, apply link, no login or API key.
Weibo Scraper: Posts and Profiles by User ID
Pull Weibo posts and profile data by numeric user ID: text, repost and like counts, images, follower stats. Why a China residential IP is mandatory, not optional.
How to Scrape CutShort Jobs for India Tech Hiring Data (No API)
CutShort is where Indian startups post engineering and product roles, and it has no public jobs API. Learn how to extract titles, companies, salary ranges, skills, and experience bands as structured JSON for recruiting feeds and talent market research.
Airbnb Scraper: Listing Prices, Ratings, and Availability Without the API
Airbnb has no public listings API. Here's how to pull nightly price, rating, superhost status and coordinates from Airbnb search results using Python.
Maps Leads: Verified Email Extraction from Google Maps Business Listings
How to pull B2B leads from Google Maps with MX-verified emails, past the official API's 120-result cap, and pay only for contactable businesses.
Realtor.com Scraper: Property and Agent Data Without the MLS Paywall
Scrape Realtor.com listings in Python: price, beds, baths, county, and the listing agent + brokerage office on every record. No login, no API key.
Redfin Scraper: For-Sale and Sold Property Data in Python (No API Key)
Scrape Redfin listings by city, ZIP, or URL in Python: price, beds, baths, sqft, agent, MLS status, and coordinates. No login, no browser, no unblocker.
Trustpilot Business Search: Discover Companies by Category with TrustScore Data
Search Trustpilot by category or keyword in Python to build company lists with TrustScore, review count and verification status, no login required.
Zillow Property Details API: Zestimate, Tax History, and Agent Data by URL
Pull full Zillow property records by URL or ZPID in Python: price history, tax history, schools, HOA, and agent data the official API never exposed.
Zillow Rental Listings API: Rent Estimates and Availability at Scale
How to pull Zillow for-rent listings by location in Python: monthly rent, availability date, beds, baths, and Rent Zestimate, with no login or API key.
How to Scrape Zillow Listings in Python (2026 Guide, No Blocking)
Scrape Zillow for-sale listings by city or ZIP in Python: price, beds, baths, Zestimate, and days on market. No API key, no browser, no blocking.
Amazon Product Research in Python: ASIN Data, BSR and Price Without the SP-API
How to scrape Amazon product listings by ASIN or keyword: price, Best Sellers Rank, star rating, review count, seller and feature bullets, across 8 marketplaces.
Amazon Reviews Scraper: Collect Customer Ratings and Sentiment Without the SP-API
How to pull Amazon customer reviews at scale by ASIN: star rating, verified badge, full review text, helpful votes and images, across 8 marketplaces.
Bulk Email Verification in Python: MX, SMTP and Catch-All Detection Without an API
How to verify thousands of email addresses for free using DNS MX lookups and optional SMTP handshakes. No API key, no monthly subscription.
Google Maps Reviews Scraper: Extract Ratings, Text and Owner Replies Without an API
How to pull Google Maps place reviews at scale: star ratings, full text, reviewer details and owner responses, for any hotel, restaurant, shop or attraction.
Scrape Business Emails, Phones and Social Profiles from Any Website
How to extract contact information from a list of company domains at scale: emails, phone numbers, LinkedIn, X, Instagram and more, with no API key.
TripAdvisor Reviews Scraper: Hotels, Restaurants and Attractions Without an API
How to pull TripAdvisor reviews at scale: star rating, full text, trip type, owner response and sub-ratings for any listing, with no TripAdvisor API key.
YouTube Channel Scraper: Subscribers, Video Stats and Channel Data Without a YouTube API Key
How to pull YouTube channel metadata and video lists: subscriber count, view counts, publish dates, video descriptions, without using the YouTube Data API or paying for quota.
Company KYB API: Resolve LEI and SEC CIK, Then Check EU VAT
Resolve a list of company names to GLEIF LEIs and SEC EDGAR CIKs in one run, flag EU entities, then validate their VAT numbers with VIES. Python included.
EU VAT Validator: Bulk VIES Verification Without an API Key
Validate EU VAT numbers in bulk via the official VIES service. Returns validity, registered company name, and address. No API key required. Empty results are never charged.
GLEIF LEI Lookup API: Resolve Legal Entity Identifiers Without an API Key
Look up Legal Entity Identifiers (LEIs) and registered company data from the GLEIF global registry. Search by company name or LEI code. No API key required.
The Mine Works MCP Server: Add 29 Data Tools to Claude Desktop in 2 Minutes
Connect LinkedIn, Reddit, SEC filings, B2B leads, PubMed, arXiv, Google Trends, and more to Claude Desktop or Cursor via a single MCP endpoint. No subscriptions: billing flows through your Apify account.
arXiv Scraper: Search AI, Physics, and Biology Preprints with PDF Links via API
Search arXiv preprints by keyword, category, or author. Returns title, abstract, authors, categories, and PDF links. Track AI research before peer review. No API key required.
B2B Leads Finder: Business Emails and LinkedIn Profiles Without Apollo or ZoomInfo
Find business emails, LinkedIn profiles, and job titles for decision-makers at target companies. No Apollo, ZoomInfo, or Lusha API key required. $0.003 per lead.
CMS Hospital Quality Data: Star Ratings, HCAHPS Scores, and Complication Rates via API
Pull CMS hospital quality metrics for 4,500+ US hospitals: overall star ratings, patient satisfaction scores, complication rates, and readmission rates. No API key required.
FDA 510(k) Clearances API: Medical Device Intelligence by Company, Device, or Product Code
Search FDA 510(k) premarket clearances by company, device name, or product code. Returns applicant, device, decision, and dates for medtech competitive intelligence. No API key required.
LinkedIn Company Scraper: Size, Industry, Website, and Followers Without Login
Scrape LinkedIn company pages for employee count, industry, HQ, founding year, website, follower count, and specialties. No LinkedIn login or API key required.
LinkedIn Jobs Scraper: Job Listings by Keyword and Location Without Login
Scrape LinkedIn job listings by keyword and location: job title, company, seniority, applicant count, and job description. No LinkedIn login required.
LinkedIn Post Search: Find Posts by Keyword Without Login
How to search LinkedIn posts without an account: find a post from one person, trace the author of a quote, and pull matching posts in Python.
LinkedIn Profile Scraper: Experience, Education, and Skills Without Login
Extract full LinkedIn profile data: work history, education, skills, connections, from a list of profile URLs. No LinkedIn cookies or login required.
Medicare Part D Drug Spending Data: Unit Cost, Claims, and Beneficiary Counts via API
Pull Medicare Part D drug spending from CMS: total cost, claims, beneficiaries, and unit price by drug name or manufacturer. Track pharmaceutical pricing trends. No API key required.
NIH RePORTER API: Search Grant Funding, Award Amounts, and Principal Investigators
Search NIH grant awards by topic, agency, institution, or state. Returns project title, abstract, award amount, and PI data for research funding intelligence. No API key required.
AliExpress Product Data API: Prices, Ratings, and Orders in Python
AliExpress affiliate API has restricted coverage. Learn how to scrape AliExpress product listings for prices, ratings, order counts, and seller data as structured JSON, no affiliate approval needed.
How to Scrape AmbitionBox Company Reviews and Ratings
AmbitionBox is India largest employer review platform with 300,000 companies. Learn how to pull ratings, review counts, salary data, and dimension scores as structured JSON without any official API.
ClinicalTrials.gov API v2: How to Search 600,000 Studies and Track Trial Status
How to use the ClinicalTrials.gov v2 REST API: what changed from v1, which filters actually exist, and Python for bulk pulls and status monitoring.
CourtListener API: How to Search US Court Records and Case Law Programmatically
CourtListener exposes 10M+ court opinions and dockets via a free REST API. Here is how to query it, what the rate limits actually are, and when a scraper is faster.
Crossref API: 150 Million DOIs, Citation Counts, and Bibliographic Data for Free
Crossref's free REST API returns journal and article metadata, reference lists and citation counts for 150M+ DOIs, with no API key needed.
How to Scrape Crunchbase Company Profiles in Python (Funding, Investors, No API Key)
Crunchbase has no free public API. Learn how to extract company names, total funding raised, investor counts, latest rounds, and headquarters as structured JSON using an Apify scraper with residential proxy bypass.
FDA Recall Data API: How to Monitor Drug, Device, and Food Recalls Programmatically
openFDA serves drug, device and food recalls plus adverse events and 510(k) clearances over a free REST API. How the endpoints and fields work.
Federal Register API: How to Track US Rules, Proposed Rules, and Executive Orders
The Federal Register publishes every US executive action, proposed rule, and final rule via a REST API. Here is how to query it and what the data contains.
How to Scrape Google News in Python (No API Key Required)
Google killed its News API in 2013. Learn how to pull headlines, sources, and publication dates from Google News in Python using the RSS feed, the GNews approach, and a pay-per-result scraper.
Google Trends API for Python in 2025: pytrends vs Scraper
Google Trends has no official API. Learn why pytrends breaks, how the SERP API approach works, and the fastest way to pull trend data into Python without getting rate-limited.
India Government Data API: How to Pull Any data.gov.in Dataset Without the Documentation Confusion
data.gov.in has 10,000+ datasets including mandi prices, foreign trade, and census data. The OGD API works but has quirks that are not documented anywhere.
How to Scrape IndiaMART B2B Suppliers in Python (Phone, Price, Leads)
IndiaMART is India's largest B2B marketplace with no public API. Learn how to extract supplier names, phone numbers, cities, product categories, and price indications as structured JSON for sales prospecting and market research.
Instagram Profile Data Without the Meta API: Followers, Bio, and Posts at Scale
Meta restricts the Instagram Graph API to your own accounts. For researching public third-party profiles at scale, here is what data is available and how to collect it.
How to Scrape JustDial Business Listings in Python (Phone, Address, Ratings)
JustDial is India's largest local business directory with no public API. Learn how to extract business names, phone numbers, addresses, geo-coordinates, and review data by city and category as structured JSON.
How to Scrape LinkedIn Employees Without Login or Sales Navigator
LinkedIn has no public API for employee data. Learn how to pull B2B leads, employee lists, and org chart data from LinkedIn company pages without a LinkedIn account or Sales Navigator subscription.
How to Scrape the Meta Ad Library in Python (Facebook and Instagram Ads, No Login)
The Meta Ad Library has no official scraping API. Learn how to extract ad copy, creative URLs, advertiser details, platforms, and run dates from Facebook and Instagram ads by keyword or advertiser, as structured JSON.
NPI Registry API: How to Look Up Any US Healthcare Provider Programmatically
CMS publishes the National Provider Identifier registry as a free API. Here is how to search by provider name, specialty, location, and NPI number, and what the data contains.
OpenAlex API: 250 Million Research Papers, Free, No Rate-Limit Workarounds Needed
OpenAlex replaced the defunct Microsoft Academic Graph with 250M+ scholarly works. The API is free, well-documented, and returns structured data including citations and author affiliations.
How to Query Norway BRREG Business Register in Python (Companies, Officers, AML)
Norway's BRREG business register is public and free, but navigating the API to get companies plus officer roles requires multiple calls per entity. Learn how to extract company status, industry, address, and full officer rosters as structured JSON.
How to Scrape Pinterest Profiles in Python (Followers, Pins, Boards Without Login)
Pinterest has no public API for scraping profiles. Learn how to extract follower counts, monthly views, board lists, recent pins with save counts, and profile metadata as structured JSON without login or API key.
SEC EDGAR Full-Text Search API: Search Every Filing Since 2001
The free EFTS endpoint searches every SEC filing since 2001. The parts that break scripts: a required User-Agent, 100 hits per page, and a hard 10,000-result cap.
Socrata API: How to Pull CDC, HHS, NYC, and 200+ Government Data Portals
Socrata powers data portals for the CDC, HHS, Chicago, New York City, Texas, and 200+ other government entities. One API, same query syntax, all of them.
Threads Has No Public API in 2026: Here Is How to Get Post and Profile Data Anyway
Every field you can collect from Threads without an API: post text, engagement counts, media, profiles. No login required, from $1 per 1,000 posts, plus a comparison of the scrapers that do it.
How to Scrape Trustpilot Reviews by Company Domain (Python Guide)
Trustpilot has no public API for review data. Learn how to pull business reviews, star ratings, trust scores, and business replies from any Trustpilot company page using Python.
USASpending.gov API: How to Pull Federal Contracts, Grants, and Awards Programmatically
USASpending.gov tracks every federal dollar spent. The API is public and free but the endpoint structure is non-obvious. Here is how to actually use it in Python.
World Bank API in Python 2025: GDP, Inflation, and 1,400 Indicators Without the SOAP Hell
The World Bank has a REST API but it returns XML by default, uses quirky pagination, and has undocumented quirks. Here is how to actually use it in Python.
World Bank Trade Data API: How to Pull Global Import and Export Statistics
The World Bank WITS database covers bilateral trade flows between 200+ countries. Here is how to access it programmatically and what the data actually contains.
How to Scrape Yellow Pages US Business Listings in Python (Phone, Address, Website)
YellowPages.com has no public API. Learn how to extract US local business names, phone numbers, addresses, websites, and categories by city and business type as structured JSON for sales prospecting and local research.
The Agentic Data Stack 2025: How to Pick the Right Scrapers for Your AI Workflow
A practical guide to building grounded AI agents on live scraped data, and which data sources matter for which kind of agent.
Building a RAG Pipeline on SEC EDGAR Filings: A Step-by-Step Guide
How to scrape SEC EDGAR filings, chunk them for vector search, and build a provenance-aware Q&A system that cites specific filing sections using Claude.
How to Aggregate Job Postings from 500+ Companies Using Public ATS APIs
Greenhouse, Lever and Ashby expose public job board APIs with no auth. Build one aggregator that pulls from all three, dedupes, and flags new roles.
How to Build a RAG Pipeline Using Web-Scraped Content
A complete guide to turning any website into LLM context, from crawling and chunking to embedding, retrieval, and keeping the index fresh.
Naukri API 2025: How to Programmatically Access India's Largest Job Board
Naukri has no public API. This guide covers the session-warming approach that gets past Akamai bot detection and returns structured job data.
How to Scrape Reddit Without an API Key in 2026
The old reddit.com .json endpoints now return 403 and commercial API access is enterprise-only. Every method that still works in 2026, with code you can use today.