Option
HeimHeim Skill API-Entwicklung brightdata-cli

brightdata-cli

brightdata/skills brightdata/skills

Anleitung zur Verwendung der Bright Data-Befehlszeilenschnittstelle (CLI) (`brightdata` / `bdata`) zum Scrapen von Websites, für die Suche im Internet, zum Extrahieren strukturierter Daten aus über 40 Plattformen, zum Verwalten von Proxy-Zonen und zum Überprüfen des Kontoguthabens. Verwenden Sie diese Funktion immer dann, wenn der Nutzer eine URL auslesen, bei Google/Bing/Yandex suchen, Daten von Amazon/LinkedIn/Instagram/TikTok/YouTube/Reddit oder einer anderen Plattform extrahieren, sein Bright Data-Guthaben oder seine Zonen überprüfen oder sonstige Aufgaben im Zusammenhang mit der Erfassung von Webdaten über das Terminal ausführen möchte. Lösen Sie die Funktion außerdem aus, wenn der Nutzer Folgendes erwähnt:

...Alle erweitern
76
Zeit aktualisiert 29. Juni 2026

Über brightdata-cli

Die Bright Data CLI ist ein Befehlszeilentool, das direkten Terminalzugriff auf die Webdatenplattform von Bright Data ermöglicht. Es vereinfacht das Web-Scraping erheblich, indem es die Authentifizierung, die Proxy-Verwaltung, die Umgehung von Anti-Bot-Maßnahmen, das Lösen von CAPTCHAs und das Rendern von JavaScript automatisch übernimmt. Benutzer melden sich einfach einmalig über OAuth an, und die CLI wickelt alle nachfolgenden Vorgänge nahtlos ab – ohne manuelle Konfiguration oder Verwaltung von API-Schlüsseln.

Diese Funktion ermöglicht das Scraping beliebiger URLs in verschiedenen Formaten (Markdown, HTML, JSON, Screenshot), die Durchführung von Websuchen bei Google, Bing und Yandex mit strukturierten Ergebnissen sowie die Extraktion von Daten aus über 40 Plattformen, darunter Amazon, LinkedIn, Instagram, TikTok, YouTube und Reddit. Die CLI unterstützt sowohl synchrone als auch asynchrone Vorgänge sowie geotargetetes Scraping aus bestimmten Ländern und bietet Funktionen zur Kontoverwaltung wie Budgetüberprüfung und Konfiguration von Proxy-Zonen. Alle Vorgänge werden über einfache Befehle wie „bdata scrape“ oder „bdata search“ ausgeführt.

Das Tool richtet sich an Entwickler, Datenanalysten und Forscher, die eine zuverlässige Web-Datenerfassung ohne zusätzlichen Infrastrukturaufwand benötigen. Ob es um die Gewinnung von Wettbewerbsinformationen, die Überwachung von Inhalten, die Durchführung von SEO-Analysen oder das Sammeln von Trainingsdaten geht – die Bright Data CLI übernimmt die technischen Komplexitäten und bietet gleichzeitig eine übersichtliche Befehlsschnittstelle. Sie läuft unter macOS, Linux und Windows, erfordert lediglich Node.js 20+ und kann für einmalige Aufgaben sogar ohne Installation über npx ausgeführt werden.

FAQ

Wie authentifiziere ich mich bei der Bright Data CLI?

Führen Sie einmalig „bdata login“ aus, um sich über OAuth in Ihrem Browser zu authentifizieren. Dadurch werden Ihre Anmeldedaten lokal gespeichert und die erforderlichen Proxy-Zonen automatisch konfiguriert. Für Headless-Umgebungen verwenden Sie „bdata login --device“ oder „bdata login --api-key “ für die nicht-interaktive Authentifizierung.

Von welchen Plattformen kann ich Daten extrahieren?

Die CLI unterstützt die Extraktion strukturierter Daten von über 40 Plattformen über Pipelines, darunter Amazon, LinkedIn, Instagram, TikTok, YouTube, Reddit und viele andere. Verwenden Sie „bdata pipelines list“, um alle verfügbaren Plattformtypen anzuzeigen.

Welche Ausgabeformate werden für das Scraping unterstützt?

Das Scraping unterstützt mehrere Formate: Markdown (Standard), HTML, JSON für strukturierte Daten und Screenshot-Bilder. Geben Sie das Format mit dem Flag -f an, z. B. „bdata scrape -f html“ oder „bdata scrape -f screenshot -o image.png“.

Muss ich Proxys verwalten oder CAPTCHAs selbst bearbeiten?

Nein. Die CLI übernimmt automatisch die Proxy-Verwaltung, die Umgehung von Anti-Bot-Maßnahmen, das Lösen von CAPTCHAs und das Rendern von JavaScript. Nach dem ersten „bdata login“ werden alle technischen Komplexitäten aus dem Prozess ausgeblendet.

Kann ich von bestimmten geografischen Standorten aus scrapen?

Ja. Verwenden Sie das Flag „--country“, um ein geografisch gezieltes Scraping festzulegen, zum Beispiel „bdata scrape https://amazon.com --country us“, um von einer US-amerikanischen IP-Adresse aus zu scrapen.

Auf GitHub ansehen

Bright Data CLI

The Bright Data CLI (brightdata or bdata) gives you full access to Bright Data's web data platform from the terminal. It handles authentication, proxy zones, anti-bot bypass, CAPTCHA solving, and JavaScript rendering automatically — the user just needs to log in once.

Installation

If the CLI is not installed yet, guide the user:

macOS / Linux:

curl -fsSL https://cli.brightdata.com/install.sh | bash

Windows or manual install (any platform):

npm install -g @brightdata/cli

Without installing (one-off usage):

npx --yes --package @brightdata/cli brightdata <command>

Requires Node.js >= 20. After install, both brightdata and bdata (shorthand) are available.

First-Time Setup

Before anything else, check if the user is authenticated. If they haven't logged in yet, guide them through the one-time setup:

# One-time login — opens the browser for OAuth, then everything is automaticbdata login

This single command:

  1. Opens the browser for secure OAuth authentication
  2. Saves the API key locally (never needs to be entered again)
  3. Auto-creates required proxy zones (cli_unlocker, cli_browser)
  4. Sets default configuration

After login, every subsequent command works without any manual intervention.

For headless/SSH environments where no browser is available:

bdata login --device

For direct API key authentication (non-interactive):

bdata login --api-key <key>

To verify setup is complete, run:

bdata config

Command Reference

Read references/commands.md for the full command reference with all flags, options, and examples for every command.

Read references/pipelines.md for the complete list of 40+ pipeline types (Amazon, LinkedIn, Instagram, TikTok, YouTube, Reddit, and more) with their specific parameters.

Quick Command Overview

bdata is the shorthand for brightdata. Both work identically.

CommandPurpose
bdata scrape <url>Scrape any URL as markdown, HTML, JSON, or screenshot
bdata search "<query>"Search Google/Bing/Yandex with structured results
bdata pipelines <type> [params]Extract structured data from 40+ platforms
bdata pipelines listList all 40+ available pipeline types
bdata status <job-id>Check async job status
bdata zonesList proxy zones
bdata budgetView account balance and costs
bdata skill addInstall AI agent skills
bdata skill listList available skills
bdata configView/set configuration
bdata loginAuthenticate with Bright Data
bdata versionShow CLI version and system info

How to Use Each Command

Scraping

Scrape any URL with automatic bot bypass, CAPTCHA handling, and JS rendering:

# Default: returns clean markdownbdata scrape https://example.com# Get raw HTMLbdata scrape https://example.com -f html# Get structured JSONbdata scrape https://example.com -f json# Take a screenshotbdata scrape https://example.com -f screenshot -o page.png# Geo-targeted scrape from the USbdata scrape https://amazon.com --country us# Save to filebdata scrape https://example.com -o page.md# Async mode for heavy pagesbdata scrape https://example.com --async

Searching

Search engines with structured JSON output (Google returns parsed organic results, ads, People Also Ask, and related searches):

# Google search with formatted tablebdata search "web scraping best practices"# Get raw JSON for pipingbdata search "typescript tutorials" --json# Search Bingbdata search "bright data pricing" --engine bing# Localized searchbdata search "restaurants berlin" --country de --language de# News searchbdata search "AI regulation" --type news# Extract just URLsbdata search "open source tools" --json | jq -r '.organic[].link'

Pipelines (Structured Data Extraction)

Extract structured data from 40+ platforms. These trigger async jobs that poll until results are ready:

# LinkedIn profilebdata pipelines linkedin_person_profile "https://linkedin.com/in/username"# Amazon productbdata pipelines amazon_product "https://amazon.com/dp/B09V3KXJPB"# Instagram profilebdata pipelines instagram_profiles "https://instagram.com/username"# Amazon searchbdata pipelines amazon_product_search "laptop" "https://amazon.com"# YouTube comments (top 50)bdata pipelines youtube_comments "https://youtube.com/watch?v=..." 50# Google Maps reviews (last 7 days)bdata pipelines google_maps_reviews "https://maps.google.com/..." 7# Output as CSVbdata pipelines amazon_product "https://amazon.com/dp/..." --format csv -o product.csv# List all available pipeline typesbdata pipelines list

Checking Status

For async jobs (from --async scrapes or pipelines):

# Quick status checkbdata status <job-id># Wait until completebdata status <job-id> --wait# With custom timeoutbdata status <job-id> --wait --timeout 300

Budget & Zones

# Quick account balancebdata budget# Detailed balance with pending chargesbdata budget balance# All zones cost/bandwidthbdata budget zones# Specific zone costsbdata budget zone my_zone# Date range filterbdata budget zones --from 2024-01-01T00:00:00 --to 2024-02-01T00:00:00# List all zonesbdata zones# Zone detailsbdata zones info cli_unlocker

Configuration

# View all configbdata config# Set defaultsbdata config set default_zone_unlocker my_zonebdata config set default_format json

Installing AI Agent Skills

# Interactive picker — choose skills and target agentsbdata skill add# Install a specific skillbdata skill add scrape# List available skillsbdata skill list

Output Modes

Every command supports multiple output formats:

FlagEffect
(none)Human-readable formatted output with colors
--jsonCompact JSON to stdout
--prettyIndented JSON to stdout
-o <path>Write to file (format auto-detected from extension)

When piped (stdout is not a TTY), colors and spinners are automatically disabled.

Chaining Commands

The CLI is pipe-friendly:

# Search → extract first URL → scrape itbdata search "top open source projects" --json \  | jq -r '.organic[0].link' \  | xargs bdata scrape# Scrape and view with markdown readerbdata scrape https://docs.github.com | glow -# Amazon product data to CSVbdata pipelines amazon_product "https://amazon.com/dp/xxx" --format csv > product.csv

Environment Variables

These override stored configuration:

VariablePurpose
BRIGHTDATA_API_KEYAPI key (skips login entirely)
BRIGHTDATA_UNLOCKER_ZONEDefault Web Unlocker zone
BRIGHTDATA_SERP_ZONEDefault SERP zone
BRIGHTDATA_POLLING_TIMEOUTPolling timeout in seconds

Troubleshooting

ErrorFix
CLI not foundInstall with npm i -g @brightdata/cli or curl -fsSL https://cli.brightdata.com/install.sh | bash
"No Web Unlocker zone specified"bdata config set default_zone_unlocker <zone> or re-run bdata login
"Invalid or expired API key"bdata login
"Access denied"Check zone permissions in the Bright Data control panel
"Rate limit exceeded"Wait and retry, or use --async for large jobs
Async job timeoutIncrease with --timeout 1200 or BRIGHTDATA_POLLING_TIMEOUT=1200

Key Design Principles

  • One-time auth: After bdata login, everything is automatic. No tokens to manage, no keys to pass.
  • Zones auto-created: Login creates cli_unlocker and cli_browser zones automatically.
  • Smart defaults: Markdown output, auto-detected formats from file extensions, colors only in TTY.
  • Pipe-friendly: JSON output + jq for automation. Colors/spinners disabled in pipes.
  • Async support: Heavy jobs can run in background with --async + status --wait.
  • npm package: @brightdata/cli — install globally or use via npx.

brightdata-cli installieren

Laden Sie die Skill-Dateien herunter und entpacken Sie sie in Ihr Verzeichnis „.claude/skills/“.

ZIP herunterladen

Klonen Sie das Repository und kopieren Sie die Skill-Dateien in Ihr Projekt.

git clone https://github.com/brightdata/skills/blob/main/skills/brightdata-cli/SKILL.md # Copy SKILL.md to your .claude/skills/ directory

Kopieren Kopieren
Schnelle Einrichtung: Kopiere den Skill-Ordner nach .claude/skills/. Claude erkennt den Skill automatisch und nutzt ihn.
Repository brightdata/skills

Ähnliche Skills

humanize
Zeit aktualisiert 7. Juli 2026
agentwallet
Zeit aktualisiert 7. Juli 2026
korean-stock-search
Zeit aktualisiert 8. Juli 2026
trello
Zeit aktualisiert 1. Juli 2026
OR