Now liveThe Skillselion MCP - thousands of ranked skills, loaded into your agent mid-task. No install.Get it →
brightdata avatar

Brightdata Cli

  • 1 installs
  • 3 repo stars
  • Updated March 30, 2026
  • brightdata/opencode-brightdata

This is a copy of brightdata-cli by brightdata - installs and ranking accrue to the original listing.

brightdata-cli is a skill for using the Bright Data CLI (brightdata / bdata) to scrape URLs, search Google/Bing/Yandex, and extract structured data from 40+ platforms from the terminal.

About

brightdata-cli is a guide for using the Bright Data CLI (brightdata / bdata) to collect web data from the terminal. It scrapes any URL as markdown, HTML, JSON, or screenshot, searches Google/Bing/Yandex, and extracts structured data from 40+ platforms like Amazon, LinkedIn, Instagram, and YouTube via pipelines. A developer uses it to do web data collection from the terminal with automatic anti-bot bypass, CAPTCHA handling, and proxy zones after a single OAuth login.

  • Wraps the Bright Data CLI (brightdata / bdata) for scraping, SERP search, and structured extraction from 40+ platforms
  • Handles auth, proxy zones, anti-bot bypass, CAPTCHA solving, and JS rendering automatically after one login
  • Covers scrape, search, pipelines, status, budget, and zones commands with format and geo options

Brightdata Cli by the numbers

  • 1 all-time installs (skills.sh)
  • Data as of Jul 28, 2026 (Skillselion catalog sync)
At a glance

brightdata-cli capabilities & compatibility

Free CLI install; requires a Bright Data account and login, with usage billed by Bright Data (check via bdata budget).

Capabilities
web scraping · web search · structured data extraction · proxy zones
Works with
chrome · linkedin
Use cases
web scraping · web search · research · data analysis
Runs
Runs locally
Pricing
Bring your own API key
From the docs

What brightdata-cli says it does

The Bright Data CLI (`brightdata` or `bdata`) gives you full access to Bright Data's web data platform from the terminal.
SKILL.md
It handles authentication, proxy zones, anti-bot bypass, CAPTCHA solving, and JavaScript rendering automatically — the user just needs to log in once.
SKILL.md
Requires Node.js >= 20. After install, both `brightdata` and `bdata` (shorthand) are available.
SKILL.md
npx skills add https://github.com/brightdata/opencode-brightdata --skill brightdata-cli

Add your badge

Show developers this skill is listed on Skillselion. Paste this into your README.

Listed on Skillselion
Installs1
repo stars3
Last updatedMarch 30, 2026
Repositorybrightdata/opencode-brightdata

What it does

Scrape URLs, search engines, and extract structured data from 40+ platforms via the Bright Data CLI.

Who is it for?

Developers who want terminal-based web scraping and structured data extraction with automatic anti-bot handling.

Skip if: Users needing an in-code SDK (use the Python or JS SDK skills) rather than a CLI.

When should I use this skill?

The user wants to scrape a URL, search a web engine, extract data from a major platform, or check their Bright Data balance or zones.

What you get

Clean scraped pages, structured SERP results, or platform datasets pulled from the terminal with automatic bot handling.

  • Scraped pages as markdown, HTML, JSON, or screenshots
  • Structured platform data via pipelines
  • SERP search results

By the numbers

  • 40+ platform pipeline types
  • 3 search engines (Google/Bing/Yandex)
  • requires Node.js >= 20

Files

SKILL.mdMarkdownGitHub ↗

Bright Data CLI

The Bright Data CLI (brightdata or bdata) gives you full access to Bright Data's web data platform from the terminal. It handles authentication, proxy zones, anti-bot bypass, CAPTCHA solving, and JavaScript rendering automatically — the user just needs to log in once.

Installation

If the CLI is not installed yet, guide the user:

macOS / Linux:

curl -fsSL https://cli.brightdata.com/install.sh | bash

Windows or manual install (any platform):

npm install -g @brightdata/cli

Without installing (one-off usage):

npx --yes --package @brightdata/cli brightdata <command>

Requires Node.js >= 20. After install, both brightdata and bdata (shorthand) are available.

First-Time Setup

Before anything else, check if the user is authenticated. If they haven't logged in yet, guide them through the one-time setup:

# One-time login — opens the browser for OAuth, then everything is automatic
bdata login

This single command: 1. Opens the browser for secure OAuth authentication 2. Saves the API key locally (never needs to be entered again) 3. Auto-creates required proxy zones (cli_unlocker, cli_browser) 4. Sets default configuration

After login, every subsequent command works without any manual intervention.

For headless/SSH environments where no browser is available:

bdata login --device

For direct API key authentication (non-interactive):

bdata login --api-key <key>

To verify setup is complete, run:

bdata config

Command Reference

Read references/commands.md for the full command reference with all flags, options, and examples for every command.

Read references/pipelines.md for the complete list of 40+ pipeline types (Amazon, LinkedIn, Instagram, TikTok, YouTube, Reddit, and more) with their specific parameters.

Quick Command Overview

bdata is the shorthand for brightdata. Both work identically.

CommandPurpose
bdata scrape <url>Scrape any URL as markdown, HTML, JSON, or screenshot
bdata search "<query>"Search Google/Bing/Yandex with structured results
bdata pipelines <type> [params]Extract structured data from 40+ platforms
bdata pipelines listList all 40+ available pipeline types
bdata status <job-id>Check async job status
bdata zonesList proxy zones
bdata budgetView account balance and costs
bdata skill addInstall AI agent skills
bdata skill listList available skills
bdata configView/set configuration
bdata loginAuthenticate with Bright Data
bdata versionShow CLI version and system info

How to Use Each Command

Scraping

Scrape any URL with automatic bot bypass, CAPTCHA handling, and JS rendering:

# Default: returns clean markdown
bdata scrape https://example.com

# Get raw HTML
bdata scrape https://example.com -f html

# Get structured JSON
bdata scrape https://example.com -f json

# Take a screenshot
bdata scrape https://example.com -f screenshot -o page.png

# Geo-targeted scrape from the US
bdata scrape https://amazon.com --country us

# Save to file
bdata scrape https://example.com -o page.md

# Async mode for heavy pages
bdata scrape https://example.com --async

Searching

Search engines with structured JSON output (Google returns parsed organic results, ads, People Also Ask, and related searches):

# Google search with formatted table
bdata search "web scraping best practices"

# Get raw JSON for piping
bdata search "typescript tutorials" --json

# Search Bing
bdata search "bright data pricing" --engine bing

# Localized search
bdata search "restaurants berlin" --country de --language de

# News search
bdata search "AI regulation" --type news

# Extract just URLs
bdata search "open source tools" --json | jq -r '.organic[].link'

Pipelines (Structured Data Extraction)

Extract structured data from 40+ platforms. These trigger async jobs that poll until results are ready:

# LinkedIn profile
bdata pipelines linkedin_person_profile "https://linkedin.com/in/username"

# Amazon product
bdata pipelines amazon_product "https://amazon.com/dp/B09V3KXJPB"

# Instagram profile
bdata pipelines instagram_profiles "https://instagram.com/username"

# Amazon search
bdata pipelines amazon_product_search "laptop" "https://amazon.com"

# YouTube comments (top 50)
bdata pipelines youtube_comments "https://youtube.com/watch?v=..." 50

# Google Maps reviews (last 7 days)
bdata pipelines google_maps_reviews "https://maps.google.com/..." 7

# Output as CSV
bdata pipelines amazon_product "https://amazon.com/dp/..." --format csv -o product.csv

# List all available pipeline types
bdata pipelines list

Checking Status

For async jobs (from --async scrapes or pipelines):

# Quick status check
bdata status <job-id>

# Wait until complete
bdata status <job-id> --wait

# With custom timeout
bdata status <job-id> --wait --timeout 300

Budget & Zones

# Quick account balance
bdata budget

# Detailed balance with pending charges
bdata budget balance

# All zones cost/bandwidth
bdata budget zones

# Specific zone costs
bdata budget zone my_zone

# Date range filter
bdata budget zones --from 2024-01-01T00:00:00 --to 2024-02-01T00:00:00

# List all zones
bdata zones

# Zone details
bdata zones info cli_unlocker

Configuration

# View all config
bdata config

# Set defaults
bdata config set default_zone_unlocker my_zone
bdata config set default_format json

Installing AI Agent Skills

# Interactive picker — choose skills and target agents
bdata skill add

# Install a specific skill
bdata skill add scrape

# List available skills
bdata skill list

Output Modes

Every command supports multiple output formats:

FlagEffect
(none)Human-readable formatted output with colors
--jsonCompact JSON to stdout
--prettyIndented JSON to stdout
-o <path>Write to file (format auto-detected from extension)

When piped (stdout is not a TTY), colors and spinners are automatically disabled.

Chaining Commands

The CLI is pipe-friendly:

# Search → extract first URL → scrape it
bdata search "top open source projects" --json \
  | jq -r '.organic[0].link' \
  | xargs bdata scrape

# Scrape and view with markdown reader
bdata scrape https://docs.github.com | glow -

# Amazon product data to CSV
bdata pipelines amazon_product "https://amazon.com/dp/xxx" --format csv > product.csv

Environment Variables

These override stored configuration:

VariablePurpose
BRIGHTDATA_API_KEYAPI key (skips login entirely)
BRIGHTDATA_UNLOCKER_ZONEDefault Web Unlocker zone
BRIGHTDATA_SERP_ZONEDefault SERP zone
BRIGHTDATA_POLLING_TIMEOUTPolling timeout in seconds

Troubleshooting

ErrorFix
CLI not foundInstall with npm i -g @brightdata/cli or `curl -fsSL https://cli.brightdata.com/install.sh \
"No Web Unlocker zone specified"bdata config set default_zone_unlocker <zone> or re-run bdata login
"Invalid or expired API key"bdata login
"Access denied"Check zone permissions in the Bright Data control panel
"Rate limit exceeded"Wait and retry, or use --async for large jobs
Async job timeoutIncrease with --timeout 1200 or BRIGHTDATA_POLLING_TIMEOUT=1200

Key Design Principles

  • One-time auth: After bdata login, everything is automatic. No tokens to manage, no keys to pass.
  • Zones auto-created: Login creates cli_unlocker and cli_browser zones automatically.
  • Smart defaults: Markdown output, auto-detected formats from file extensions, colors only in TTY.
  • Pipe-friendly: JSON output + jq for automation. Colors/spinners disabled in pipes.
  • Async support: Heavy jobs can run in background with --async + status --wait.
  • npm package: @brightdata/cli — install globally or use via npx.

Related skills

FAQ

Does it need a browser to log in?

Normally yes (bdata login opens the browser for OAuth), but for headless/SSH environments use bdata login --device or --api-key.

How many platforms can pipelines extract from?

The skill states 40+ platforms including Amazon, LinkedIn, Instagram, TikTok, YouTube, and Reddit.

This week in AI coding

Five minutes, every Monday - the tools, releases and tactics for developers.

unsubscribe anytime.