direkt zum inhalt

7. Oktober 2026kiunternehmen

Agent Reach und Scrapling: Claude Code liest das InternetAgent Reach and Scrapling: Claude Code reads the internet

Agent Reach holt YouTube, Feeds, GitHub und Websuche in Claude Code, Scrapling knackt einzelne geschützte Seiten. Ich zeige, wie ich beide auf Mac und Server eingerichtet habe, welche Kanäle ich bewusst sperre und mit welchen acht Prompts ich arbeite.Agent Reach brings YouTube, feeds, GitHub and web search into Claude Code, Scrapling handles single protected pages. I show how I set up both on my Mac and server, which channels I block on purpose and the eight prompts I work with.

Claude Code baut dir eine Funktion in zehn Minuten. Ein YouTube-Transkript holt es ohne Hilfe nicht.

Genau da setzen zwei Open-Source-Projekte an. Agent Reach gibt einem KI-Agenten Lesezugriff auf YouTube, RSS-Feeds, GitHub und eine semantische Websuche. Scrapling holt einzelne Webseiten, die normale Abrufe abblocken. Den Anstoß gab mir die Anleitung von Alan Langhammer (@alan.buildz). Ich habe beide Werkzeuge am 7. Oktober 2026 auf meinem Mac und auf meinem Server eingerichtet, in meinen Arbeitsablauf eingebaut und dabei zwei Stellen gefunden, an denen ich die Voreinstellungen nicht übernehme.

Fünf Glaslinsen auf Messingständern stehen in einer Reihe vor einem aufgeklappten Laptop, eine Hand liegt auf der Tastatur
KI-generiertes Bild in 4K, erstellt für diesen Beitrag. Jede Linse steht für einen Kanal, den Claude Code lesen kann.

Was Agent Reach macht

Agent Reach ist kein eigener Scraper. Das Projekt sucht für jede Plattform das freie Werkzeug heraus, das gerade funktioniert, installiert es und prüft regelmäßig, ob der Weg noch offen ist. Für YouTube ist das yt-dlp, für Feeds feedparser, für die Websuche Exa über mcporter, für GitHub die offizielle gh-CLI. Fällt ein Weg aus, nimmt Agent Reach den nächsten aus einer festen Kette.

fällt ein weg aus, übernimmt der nächste

  1. Bilibili
    yt-dlp (gesperrt)bili-cli (aktiv)OpenCLI
  2. Twitter / X
    twitter-cli (aktiv)OpenCLIbird
  3. Xiaohongshu
    OpenCLI (aktiv)xiaohongshu-mcpxhs-cli

Beispiel aus dem Projekt: Als Bilibili im Juni 2026 yt-dlp aussperrte, wechselte Agent Reach auf bili-cli. Welcher Weg gerade aktiv ist, zeigt agent-reach doctor.

Den Zustand aller Kanäle zeigt ein einziger Befehl: agent-reach doctor. Er listet jeden Kanal, den gerade aktiven Weg und fehlende Anmeldungen. Bei jedem Problem ist das der erste Befehl.

Stand 7. Oktober 2026

Zwei Open-Source-Projekte, fast 180.000 Sterne, keine API-Rechnung.

92,8k

GitHub-Sterne für Agent Reach (MIT-Lizenz, Release v1.5.0 vom 11. Juni 2026).

github.com/Panniantong/Agent-Reach
86,0k

GitHub-Sterne für Scrapling (BSD-3-Lizenz, Release v0.4.15 vom 23. August 2026).

github.com/D4Vinci/Scrapling
13 Kanäle

kennt Agent Reach insgesamt. Die Hälfte davon funktioniert nur mit deinem eigenen Login.

agent-reach doctor
0 €

API-Gebühren. Kostenlos heißt nicht garantiert: Eine Plattform kann einen Weg jederzeit schließen.

Projektbeschreibung
Kanäle, die nach meiner Installation in agent-reach doctor bereitstehen, ganz ohne Login
Mac (Desktop)9 / 13
Server herrlichAI7 / 13

Sterne zeigen Aufmerksamkeit, keine tägliche Nutzung. Der Mac zählt mehr, weil eine ältere OpenCLI-Installation X, Reddit und Xiaohongshu als verfügbar meldet. Genau diese Kanäle sperre ich, siehe unten.

Wo Scrapling ins Spiel kommt

Agent Reach denkt in Plattformen. Für eine einzelne Seite, die Bots abweist, reicht das nicht. Scrapling von Karim Shoair (D4Vinci) schließt diese Lücke. Der StealthyFetcher kommt laut Projekt an Cloudflare Turnstile vorbei. Mit auto_save=True merkt sich Scrapling, wie ein Element aussah, und findet es mit adaptive=True nach einem Umbau der Seite wieder. Für Claude Code bringt Scrapling einen MCP-Server mit. Er schneidet Seiten per CSS-Selektor zu und entfernt laut Dokumentation versteckten Text, der den Agenten umlenken soll.

Firecrawl bleibt bei mir für Suche mit strukturierten Datenquellen zuständig. Die drei Werkzeuge überschneiden sich weniger, als die Produktseiten vermuten lassen.

die entscheidungsmatrix

Drei Werkzeuge, drei Aufgaben. Die Überschneidung ist kleiner, als sie wirkt.

Aufgabe, empfohlenes Werkzeug und Grund
aufgabeAgent ReachScraplingFirecrawl
YouTube-Transkript ziehenyt-dlp über Agent Reach, Untertitel als Datei●
RSS-Feeds täglich auswertenfeedparser, ohne Login, ohne Kosten●
Breit recherchierenExa-Suche per mcporter, liefert Quellen mit Auszug●
Öffentliches Repo verstehengh-CLI, Issues und Releases inklusive●
Geschützte Seite lesenStealthyFetcher kommt an Cloudflare Turnstile vorbei●
Preise einer Seite regelmäßig prüfenadaptive Selektoren überstehen Layout-Wechsel●
Suche plus strukturierte DatenFirecrawl-Connector mit Alexandria-Datenquellen●
Eingeloggt auf Reddit, X, Instagramtechnisch möglich, bei mir gesperrt×

Faustregel: Plattform-Inhalte an Agent Reach, eine einzelne störrische Seite an Scrapling, Suche mit strukturierten Daten an Firecrawl.

Installation mit einer Zeile oder mit vier Befehlen

Der bequeme Weg steht in der README von Agent Reach. Du fügst Claude Code eine einzige Zeile mit der Adresse der Installationsanleitung ein, der Agent liest sie und arbeitet sie ab. Das funktioniert. Es installiert aber den aktuellen Stand von main, also Code, den vorher niemand als Release markiert hat.

Ich installiere deshalb eine feste Version in eine eigene Python-Umgebung. Vorher habe ich das Paket von v1.5.0 heruntergeladen und den Installer nach sudo, Pipes in die Shell und globalen Installationen durchsucht. Gefunden habe ich nur Hinweistexte, keine automatischen Systemeingriffe.

installation in vier blöcken

in Claude Code einfügenDen Agenten installieren lassen
Install Agent Reach: https://raw.githubusercontent.com/Panniantong/agent-reach/main/docs/install.md
TerminalFeste Version statt main.zip
python3 -m venv ~/.agent-reach-venv
~/.agent-reach-venv/bin/pip install https://github.com/Panniantong/agent-reach/archive/refs/tags/v1.5.0.zip
mkdir -p ~/.local/bin && ln -sf ~/.agent-reach-venv/bin/agent-reach ~/.local/bin/agent-reach
TerminalProbelauf, installieren, prüfen
agent-reach install --env=auto --dry-run   # zeigt nur, was passieren würde
agent-reach install --env=auto             # installiert die fehlenden Basis-Werkzeuge
agent-reach doctor                         # Status jedes Kanals
TerminalScrapling mit MCP-Server
python3 -m venv ~/.scrapling-venv
~/.scrapling-venv/bin/pip install "scrapling[ai]==0.4.15"
~/.scrapling-venv/bin/scrapling install
claude mcp add ScraplingServer ~/.scrapling-venv/bin/scrapling-mcp

Agent Reach v1.5.0 kennt den Schalter --system noch nicht, die README auf main beschreibt ihn bereits. Prüf vor dem Kopieren von Schaltern agent-reach --help in deiner Version.

Auf meinem Mac wollte der Probelauf nur mcporter installieren und die Exa-Suche einrichten. Danach meldete doctor noch einen Hinweis für YouTube. yt-dlp braucht eine JavaScript-Laufzeit, eine Zeile in seiner Konfigurationsdatei löst das. Auf dem Server lief ein YouTube-Abruf danach in drei Sekunden durch, die Untertiteldatei hatte 14,5 KB.

Zwei Voreinstellungen, die ich nicht übernehme

Erstens schreibt agent-reach doctor ungefragt einen Skill in die Skill-Ordner von Claude Code, OpenClaw und weiteren Agenten. Dessen Beschreibung beansprucht jede Webadresse für sich. In meinem Setup lesen Firecrawl, Scrapling und ein Browser-Werkzeug normale Seiten. Ich habe die Beschreibung deshalb auf YouTube, Feeds, Exa-Suche und öffentliches GitHub eingegrenzt. Ein Update überschreibt diese Eingrenzung, also aktualisiere ich nur von Hand.

Zweitens die Login-Kanäle. Auf meinem Mac lag aus einer früheren Installation OpenCLI. Damit meldet Agent Reach X, Reddit und Xiaohongshu als verfügbar, und zwar über meine eingeloggte Chrome-Sitzung. Das sind meine Hauptkonten. Das Projekt rät selbst zu Testaccounts, weil Plattformen automatisierte Zugriffe erkennen und Konten sperren. Ich sperre diese Kanäle deshalb in der Skill-Beschreibung und in meinen globalen Regeln.

was sofort geht

Die Hälfte läuft ohne Einrichtung. Die andere Hälfte will dein Konto.

ohne Login 8

  • Webseiten (Jina Reader)
  • YouTube: Untertitel, Suche
  • RSS und Atom
  • Exa-Websuche
  • GitHub öffentlich
  • V2EX
  • einzelne Tweets
  • Bilibili Basis

mit Login 7

  • X: Suche, Timeline, Threadsgesperrt
  • Redditgesperrt
  • Instagramgesperrt
  • Facebookgesperrt
  • LinkedIngesperrt
  • Xiaohongshugesperrt
  • GitHub privat

Mein Setup: OpenCLI würde die eingeloggte Chrome-Sitzung mit meinen Hauptkonten nutzen. Deshalb bleiben diese Kanäle aus.

TikTok steht nicht in der offiziellen Kanaltabelle. Einzelne TikTok-Seiten lassen sich wie jede Webseite lesen, eine echte Suche gibt es nicht.

Raster aus weißen Karten auf Leinen, ein Teil ist mit goldenen Fäden an eine Keramikscheibe in der Mitte angebunden, zwei Karten liegen umgedreht am Rand
KI-generiertes Bild in 4K. Angebunden wird nur, was ohne fremdes Konto läuft.

So hängt es in meinem Arbeitsablauf

Ich arbeite mit einer festen Schleife. Das Hauptmodell plant und prüft, ein kleineres Modell setzt einzelne Schritte um. Agent Reach und Scrapling sitzen vor dem Plan. Hängt eine Aufgabe an fremden Quellen, holt das Hauptmodell sie zuerst und legt sie als Datei ab. Erst dann entsteht der Plan, und der ausführende Agent liest nur noch Dateien.

quellen-schritt vor dem plan

1quelle holen

Transkript, Feed oder Seite über Agent Reach oder Scrapling

2als datei ablegen

im Scratch-Ordner, nie im Repo

3test -s

leere Datei heißt nicht geholt, nicht nichts gefunden

4plan

das Hauptmodell plant auf der Datei

5executor

liest nur Dateien, recherchiert nicht

Der Prüfbefehl ist bewusst schlicht: test -s datei. Eine leere Datei bedeutet, dass die Quelle nicht geholt wurde. Ohne diese Prüfung plant ein Modell auf einer leeren Datei und meldet hinterher, es habe nichts gefunden.

Auf beiden Rechnern ist die Version fest auf v1.5.0, Scrapling läuft in v0.4.15 als MCP-Server. Den Zustand prüfe ich mit agent-reach doctor, einmal lokal und einmal per SSH auf dem Server.

drei fragen, eine antwort

  1. 01
    Liegt der Inhalt auf einer Plattform wie YouTube, in einem Feed oder auf GitHub?

    ja: Agent Reachnein: weiter zur nächsten Frage

  2. 02
    Geht es um eine einzelne Seite, die normale Abrufe blockiert?

    ja: Scrapling (StealthyFetcher)nein: weiter zur nächsten Frage

  3. 03
    Brauchst du Suchergebnisse oder strukturierte Datensätze?

    ja: Firecrawlnein: ein einfacher Abruf reicht

Acht Prompts, die ich wirklich nutze

Ab hier redest du normal mit Claude Code. Jeder Prompt nennt die Plattform, das gewünschte Ergebnis und eine Grenze, an der Claude aufhören soll. Ohne Plattform greift das Modell gern zur allgemeinen Websuche und liefert eine Zusammenfassung über den Thread statt den Thread.

acht prompts zum kopieren

Nenne die Plattform, das Ergebnis und die Grenze.

prompt 01YouTube-Transkript
Hol das Transkript von [LINK] mit Agent Reach und speichere es als Datei im Scratch-Ordner. Fasse danach zusammen: die These in zwei Sätzen, die fünf konkretesten Aussagen mit Zeitstempel und jede genannte Zahl mit ihrer Quelle. Was nicht im Transkript steht, lässt du weg.
prompt 02RSS-Morgenradar
Lies diese Feeds: [URLS]. Zeig mir nur Einträge der letzten 24 Stunden, die eines dieser Stichworte treffen: [STICHWORTE]. Pro Treffer: Titel, ein Satz Relevanz für mein Geschäft, Link. Keine Treffer heißt: schreib „keine Treffer“ und nenne die Zahl der gelesenen Einträge.
prompt 03Exa-Recherche
Recherchiere [THEMA] mit der Exa-Suche über Agent Reach, mindestens acht Quellen. Trenne Herstellerangaben von unabhängigen Quellen. Gib mir eine Tabelle: Aussage, Quelle, Datum, wie belastbar. Am Ende drei offene Fragen, die die Quellen nicht beantworten.
prompt 04Repo-Analyse
Analysiere [REPO-URL] mit gh: Was macht das Projekt tatsächlich, welche Abhängigkeiten zieht es, wie oft erscheinen Releases, welche Issues sind am häufigsten offen? Prüf den Installer auf sudo, curl-Pipes und globale npm-Installationen. Schluss: einsetzen, mit Auflagen einsetzen oder lassen.
prompt 05Geschützte Seite
Die Seite [URL] blockiert normale Abrufe. Hol sie mit dem Scrapling-MCP über stealthy_fetch, nur den Bereich [CSS-SELEKTOR], als Markdown. Wenn eine Anmeldung nötig wäre, brich ab und sag es mir, statt es zu versuchen.
prompt 06Preis-Monitoring
Schreib ein kurzes Python-Skript mit Scrapling: StealthyFetcher holt [URL], liest Produktname und Preis per CSS mit auto_save=True und schreibt eine Zeile mit Datum in preise.csv. Beim nächsten Lauf mit adaptive=True, damit ein Layout-Wechsel das Skript nicht bricht. Kein Login, höchstens ein Abruf pro Stunde.
prompt 07Kombi-Briefing
Briefing zu [THEMA] auf einer Seite. Nutze Exa-Suche, YouTube-Transkripte und öffentliche GitHub-Repos über Agent Reach, für einzelne geschützte Seiten Scrapling. Gliederung: Konsens, Streitpunkte, die drei meistgenannten Werkzeuge, Widersprüche zwischen den Quellen. Jede Aussage mit Quelle und Plattform.
prompt 08Loop-Quellen-Schritt
Bevor du planst: Hol alle Quellen für diese Aufgabe mit Agent Reach oder Scrapling und leg sie als Dateien im Scratch-Ordner ab. Prüf jede Datei mit test -s. Leere Datei heißt nicht geholt, nicht „nichts gefunden“. Erst danach schreibst du den Plan, und der executor arbeitet nur mit diesen Dateien.

Ersetze die eckigen Klammern. Jeder Prompt sagt, wo Claude suchen soll und wann es aufhören soll. Das ist der Unterschied zwischen einem echten Thread und der Zusammenfassung einer Zusammenfassung.

Was du dir damit ins Haus holst

Agent Reach bündelt fremde Kommandozeilenwerkzeuge und MCP-Server. Du vertraust damit einer ganzen Kette von Projekten. Darum lohnt sich der Blick in den Installer und der Probelauf mit --dry-run vor jeder Änderung am System.

Cookie-Zugriffe auf Plattformen bleiben eine Grauzone der Nutzungsbedingungen. Ein Testaccount ist die Mindestmaßnahme, ein Hauptkonto gehört nie hinein. Die Sternezahl misst Aufmerksamkeit. Wie viele Menschen ein Werkzeug täglich einsetzen, verrät sie nicht. Und TikTok fehlt in der offiziellen Kanaltabelle, auch wenn es in manchen Videos anders klingt.

„Keine API-Gebühren“ heißt kostenlos, nicht dauerhaft. Bilibili hat im Juni 2026 einen Weg geschlossen, X und Reddit ändern ihre Sperren regelmäßig. Für Recherche ist das in Ordnung. Ein Kundenprozess, der jeden Morgen laufen muss, braucht eine offizielle Schnittstelle.

Wie ich Werkzeuge wie diese in einen festen Ablauf mit Prüfung und zweitem Modell einbaue, beschreibt der Beitrag Claude Code und Codex orchestrieren.

Häufige Fragen

Was ist Agent Reach?

Ein Open-Source-Projekt unter MIT-Lizenz. Es wählt pro Plattform ein funktionierendes freies Werkzeug aus, installiert es und prüft mit agent-reach doctor, ob der Kanal noch läuft. Fällt ein Weg aus, schaltet es auf den nächsten um.

Wofür brauche ich Scrapling, wenn ich Agent Reach habe?

Agent Reach arbeitet plattformweise. Scrapling holt einzelne Webseiten, auch hinter Cloudflare Turnstile, und findet Elemente nach einem Layout-Wechsel wieder. Über seinen MCP-Server ruft Claude Code es direkt auf.

Kostet Agent Reach etwas?

Es fallen keine API-Gebühren an. Die Wege funktionieren aber nur, solange die Plattformen sie offen lassen. Für Geschäftsprozesse, die nicht ausfallen dürfen, ist das keine Grundlage.

Sollte ich Login-Kanäle wie Reddit oder Instagram freischalten?

Nur mit einem eigenen Testaccount. Das Projekt rät selbst davon ab, Hauptkonten zu verwenden, weil Plattformen automatisierte Zugriffe erkennen und Konten sperren können.

Funktioniert Agent Reach auch auf einem Server?

Ja. Auf meinem Server liefen nach der Installation sieben von dreizehn Kanälen ohne Login, darunter YouTube-Untertitel, RSS, Exa-Suche und GitHub. Login-Kanäle gehören nicht auf einen Server.

Quellen und Stand

Stand: 7. Oktober 2026. Versionen, Kanäle und Installationsschritte können sich ändern.

  1. Agent Reach auf GitHub, README, Installationsanleitung und Release v1.5.0, gelesen am 7. Oktober 2026.
  2. Scrapling auf GitHub, README und Release v0.4.15, gelesen am 7. Oktober 2026.
  3. Scrapling-Dokumentation: MCP-Server, gelesen am 7. Oktober 2026.
  4. Alan Langhammer (@alan.buildz): Anleitung „Claude bekommt Augen fürs ganze Internet“, September 2026. Anstoß für diesen Beitrag, Text und Prompts hier sind eigene Fassungen.

Bildhinweis: Beide Bilder sind KI-generiert, erstellt für diesen Beitrag. Zahlen zu Sternen und Kanälen stammen aus GitHub und aus agent-reach doctor auf meinem Mac und meinem Server am 7. Oktober 2026.

Claude Code builds you a feature in ten minutes. It cannot fetch a YouTube transcript on its own.

Two open-source projects close that gap. Agent Reach gives an AI agent read access to YouTube, RSS feeds, GitHub and a semantic web search. Scrapling fetches single web pages that block normal requests. The prompt came from Alan Langhammer's guide (@alan.buildz). I set up both tools on my Mac and on my server on 7 October 2026, wired them into my workflow and found two defaults I do not keep.

Five glass lenses on brass stands lined up in front of an open laptop, one hand resting on the keyboard
AI-generated 4K image created for this article. Each lens stands for a channel Claude Code can read.

What Agent Reach does

Agent Reach is not a scraper of its own. For each platform it picks the free tool that currently works, installs it and checks regularly whether the route is still open. For YouTube that is yt-dlp, for feeds feedparser, for web search Exa via mcporter, for GitHub the official gh CLI. If a route fails, Agent Reach takes the next one from a fixed chain.

if one route closes, the next one takes over

  1. Bilibili
    yt-dlp (blocked)bili-cli (active)OpenCLI
  2. Twitter / X
    twitter-cli (active)OpenCLIbird
  3. Xiaohongshu
    OpenCLI (active)xiaohongshu-mcpxhs-cli

Example from the project: when Bilibili locked out yt-dlp in June 2026, Agent Reach switched to bili-cli. Which route is active right now is shown by agent-reach doctor.

One command shows the state of every channel: agent-reach doctor. It lists each channel, the active route and any missing logins. Whatever goes wrong, run it first.

as of 7 October 2026

Two open-source projects, almost 180,000 stars, no API bill.

92.8k

GitHub stars for Agent Reach (MIT licence, release v1.5.0 of 11 June 2026).

github.com/Panniantong/Agent-Reach
86.0k

GitHub stars for Scrapling (BSD-3 licence, release v0.4.15 of 23 August 2026).

github.com/D4Vinci/Scrapling
13 channels

Agent Reach knows in total. Half of them only work with your own login.

agent-reach doctor
0 €

API fees. Free does not mean guaranteed: a platform can close a route at any time.

project description
Channels that report ready in agent-reach doctor after my installation, without any login
Mac (desktop)9 / 13
herrlichAI server7 / 13

Stars show attention, not daily use. The Mac counts more because an OpenCLI installation from earlier already reports X, Reddit and Xiaohongshu as available. I block exactly these channels, see below.

Where Scrapling comes in

Agent Reach thinks in platforms. That is not enough for a single page that turns bots away. Scrapling by Karim Shoair (D4Vinci) fills this gap. According to the project, its StealthyFetcher gets past Cloudflare Turnstile. With auto_save=True Scrapling remembers what an element looked like and finds it again with adaptive=True after the page has been rebuilt. For Claude Code it ships an MCP server. The server trims pages with a CSS selector and, according to the documentation, strips hidden text meant to hijack the agent.

In my setup Firecrawl stays in charge of search with structured data providers. The three tools overlap less than their product pages suggest.

the decision matrix

Three tools, three jobs. Overlap is smaller than it looks.

Task, recommended tool and reason
taskAgent ReachScraplingFirecrawl
Pull a YouTube transcriptyt-dlp via Agent Reach, subtitles as a file●
Scan RSS feeds every dayfeedparser, no login, no cost●
Research broadlyExa search via mcporter, returns sources with excerpts●
Understand a public repogh CLI, issues and releases included●
Read a protected pageStealthyFetcher gets past Cloudflare Turnstile●
Check a page's prices regularlyadaptive selectors survive layout changes●
Search plus structured dataFirecrawl connector with Alexandria data providers●
Logged in on Reddit, X, Instagramtechnically possible, blocked in my setup×

Rule of thumb: platform content goes to Agent Reach, a single stubborn page to Scrapling, search with structured data to Firecrawl.

Installation with one line or with four commands

The convenient route is in the Agent Reach README. You paste a single line with the address of the install guide into Claude Code, and the agent reads and follows it. That works. It also installs the current state of main, which is code nobody has tagged as a release.

So I install a pinned version into its own Python environment. Before that I downloaded the v1.5.0 package and searched the installer for sudo, pipes into the shell and global installs. I only found hint texts, no automatic changes to the system.

installation in four blocks

into Claude CodeLet the agent install it
Install Agent Reach: https://raw.githubusercontent.com/Panniantong/agent-reach/main/docs/install.md
TerminalPinned version instead of main.zip
python3 -m venv ~/.agent-reach-venv
~/.agent-reach-venv/bin/pip install https://github.com/Panniantong/agent-reach/archive/refs/tags/v1.5.0.zip
mkdir -p ~/.local/bin && ln -sf ~/.agent-reach-venv/bin/agent-reach ~/.local/bin/agent-reach
TerminalDry run, install, check
agent-reach install --env=auto --dry-run   # only shows what would happen
agent-reach install --env=auto             # installs the missing base tools
agent-reach doctor                         # status of every channel
TerminalScrapling with MCP server
python3 -m venv ~/.scrapling-venv
~/.scrapling-venv/bin/pip install "scrapling[ai]==0.4.15"
~/.scrapling-venv/bin/scrapling install
claude mcp add ScraplingServer ~/.scrapling-venv/bin/scrapling-mcp

Agent Reach v1.5.0 has no --system flag yet; the README on main already documents it. Check agent-reach --help on your version before you copy flags.

On my Mac the dry run only wanted to install mcporter and set up Exa search. Afterwards doctor still flagged YouTube. yt-dlp needs a JavaScript runtime, and one line in its config file fixes that. On the server a YouTube fetch then finished in three seconds, and the subtitle file came to 14.5 KB.

Two defaults I do not keep

First, agent-reach doctor writes a skill into the skill folders of Claude Code, OpenClaw and other agents without asking. Its description claims every web address. In my setup Firecrawl, Scrapling and a browser tool read normal pages. I narrowed the description to YouTube, feeds, Exa search and public GitHub. An update overwrites that change, so I only update by hand.

Second, the login channels. My Mac still had OpenCLI from an earlier installation. With it, Agent Reach reports X, Reddit and Xiaohongshu as available, through my logged-in Chrome session. Those are my main accounts. The project itself recommends test accounts because platforms detect automated access and suspend accounts. So I block these channels in the skill description and in my global rules.

what works right away

Half runs without setup. The other half wants your account.

without login 8

  • web pages (Jina Reader)
  • YouTube: subtitles, search
  • RSS and Atom
  • Exa web search
  • public GitHub
  • V2EX
  • single tweets
  • Bilibili basic

with login 7

  • X: search, timeline, threadsblocked
  • Redditblocked
  • Instagramblocked
  • Facebookblocked
  • LinkedInblocked
  • Xiaohongshublocked
  • private GitHub

My setup: OpenCLI would use the logged-in Chrome session with my main accounts. That is why these channels stay off.

TikTok is not in the official channel table. Single TikTok pages can be read like any web page, a real search does not exist.

Grid of white cards on linen, some tied to a ceramic disc in the centre with gold threads, two cards turned face down at the edge
AI-generated 4K image. Only what runs without someone else's account gets connected.

How it fits into my workflow

I work with a fixed loop. The main model plans and reviews, a smaller model carries out single steps. Agent Reach and Scrapling sit in front of the plan. If a task depends on outside sources, the main model fetches them first and stores them as files. Only then is the plan written, and the executing agent reads files and nothing else.

source step before the plan

1fetch the source

transcript, feed or page via Agent Reach or Scrapling

2save it as a file

in the scratch folder, never in the repo

3test -s

an empty file means not fetched, not nothing found

4plan

the main model plans on the file

5executor

only reads files, does no research

The check command is deliberately plain: test -s file. An empty file means the source was not fetched. Without that check a model plans on an empty file and later reports that it found nothing.

Both machines are pinned to v1.5.0, and Scrapling runs as an MCP server in v0.4.15. I check the state with agent-reach doctor, once locally and once on the server via SSH.

three questions, one answer

  1. 01
    Is the content on a platform like YouTube, a feed or GitHub?

    yes: Agent Reachno: next question

  2. 02
    Is it a single page that blocks normal requests?

    yes: Scrapling (StealthyFetcher)no: next question

  3. 03
    Do you need search results or structured data records?

    yes: Firecrawlno: a plain fetch is enough

Eight prompts I actually use

From here on you talk to Claude Code normally. Each prompt names the platform, the result you want and a limit where Claude should stop. Without a platform the model tends to reach for general web search and hands you a summary of the thread instead of the thread.

eight prompts to copy

Name the platform, name the output, name the limit.

prompt 01YouTube transcript
Fetch the transcript of [LINK] with Agent Reach and save it as a file in the scratch folder. Then summarise: the thesis in two sentences, the five most concrete claims with timestamps and every number mentioned with its source. Leave out anything that is not in the transcript.
prompt 02RSS morning radar
Read these feeds: [URLS]. Show me only entries from the last 24 hours that match one of these keywords: [KEYWORDS]. Per hit: title, one sentence on why it matters for my business, link. No hits means: write “no hits” and state how many entries you read.
prompt 03Exa research
Research [TOPIC] with Exa search via Agent Reach, at least eight sources. Separate vendor claims from independent sources. Give me a table: claim, source, date, how solid. End with three open questions the sources do not answer.
prompt 04Repo analysis
Analyse [REPO URL] with gh: what does the project actually do, which dependencies does it pull in, how often do releases appear, which issues are most often open? Check the installer for sudo, curl pipes and global npm installs. Conclusion: adopt, adopt with conditions or skip.
prompt 05Protected page
The page [URL] blocks normal requests. Fetch it with the Scrapling MCP via stealthy_fetch, only the area [CSS SELECTOR], as Markdown. If a login would be needed, stop and tell me instead of trying.
prompt 06Price monitoring
Write a short Python script with Scrapling: StealthyFetcher fetches [URL], reads product name and price via CSS with auto_save=True and appends a dated line to prices.csv. On the next run use adaptive=True so a layout change does not break the script. No login, one request per hour at most.
prompt 07Combined briefing
One-page briefing on [TOPIC]. Use Exa search, YouTube transcripts and public GitHub repos via Agent Reach, Scrapling for single protected pages. Structure: consensus, points of dispute, the three most mentioned tools, contradictions between sources. Every claim with source and platform.
prompt 08Loop source step
Before you plan: fetch every source for this task with Agent Reach or Scrapling and store them as files in the scratch folder. Check each file with test -s. An empty file means not fetched, not “nothing found”. Only then write the plan, and the executor works with these files only.

Replace the square brackets. Each prompt names where Claude should look and when it should stop. That is the difference between a real thread and a summary of a summary.

What you are bringing in

Agent Reach bundles third-party command-line tools and MCP servers. You trust a whole chain of projects. That is why the look at the installer and the --dry-run before any change to the system are worth the time.

Cookie access to platforms stays a grey area under their terms of service. A test account is the minimum, a main account never belongs there. The star count measures attention. It does not tell you how many people use a tool every day. And TikTok is missing from the official channel table, even if some videos make it sound otherwise.

“No API fees” means free, not permanent. Bilibili closed a route in June 2026, and X and Reddit change their blocks regularly. For research that is fine. A customer process that has to run every morning needs an official API.

How I build tools like these into a fixed flow with checks and a second model is covered in orchestrating Claude Code and Codex.

Frequently asked questions

What is Agent Reach?

An open-source project under the MIT licence. For each platform it picks a working free tool, installs it and uses agent-reach doctor to check whether the channel still runs. If a route fails, it switches to the next one.

Why do I need Scrapling if I have Agent Reach?

Agent Reach works per platform. Scrapling fetches single web pages, including ones behind Cloudflare Turnstile, and finds elements again after a layout change. Claude Code calls it directly through its MCP server.

Does Agent Reach cost anything?

There are no API fees. The routes only work as long as the platforms leave them open, though. That is no basis for business processes that must not fail.

Should I unlock login channels such as Reddit or Instagram?

Only with a dedicated test account. The project itself advises against main accounts because platforms detect automated access and can suspend accounts.

Does Agent Reach work on a server?

Yes. On my server seven of thirteen channels ran without login after installation, among them YouTube subtitles, RSS, Exa search and GitHub. Login channels do not belong on a server.

Sources and status

As of 7 October 2026. Versions, channels and installation steps can change.

  1. Agent Reach on GitHub, README, install guide and release v1.5.0, read on 7 October 2026.
  2. Scrapling on GitHub, README and release v0.4.15, read on 7 October 2026.
  3. Scrapling documentation: MCP server, read on 7 October 2026.
  4. Alan Langhammer (@alan.buildz): guide on giving Claude eyes for the whole internet, September 2026. The starting point for this article; text and prompts here are my own.

Image note: both images are AI-generated for this article. Star and channel figures come from GitHub and from agent-reach doctor on my Mac and my server on 7 October 2026.

geschrieben vonwritten by · business transformation, ki und führung im mittelstand. methode: core+.business transformation, ai and leadership for mid-sized companies. method: core+.

weiterlesenkeep reading

← zurück zum herrlichblog← back to the herrlichblog