返回目录

GITHUB TOPIC

web-scraping

6个项目包含此标签

6 个项目

GitHub Topic 精确匹配

Agent 与会话

deepspider

ma-pony

AI 原生智能爬虫与 JavaScript 逆向工程平台,基于 DSH、Patchright/CDP 与独立语义运行时,从浏览器证据恢复参数生成逻辑并交付可验证 Solver。 | AI-native web scraping and JavaScript reverse-engineering platform powered by DSH, Patchright/CDP, and an independent semantic runtime—from browser evidence to recovered parameter-generation logic and verifiable Solvers.

已验证
ai-agentanti-detectautomationcaptcha
文件与数据

dsh-read-url

2672243194

DeepSeek Harness URL reader: fetch any page and return clean main-content text/Markdown. Auto charset (GBK/GB2312/UTF-8/Big5), token-efficient (6000-char cap, cache, offset), zero deps, no API key. 网页一键读全文 → 干净正文 / 结构化 Markdown

已验证
aiai-agentai-agentsai-coding
文件与数据

dsh-firecrawl

firecrawl

Firecrawl-backed web_search and web_fetch providers for the DeepSeek Harness web capability seam (ctx.web)

已验证
deepseek-harnessdsh-pluginfirecrawlweb-scraping
开发工具

dsh-crw

fastcrw

fastCRW-backed web_search and web_fetch providers for DeepSeek Harness (ctx.web)

已验证
agent-toolscordisdeepseek-harnessdsh-plugin
开发工具

dsh-webfetch

TYEclipse

Web page reader for DeepSeek Harness (dsh): fetch any URL and extract clean markdown or plain text, inventory links, read RSS/Atom feeds, inspect HTTP headers without the body, and extract HTML tables as structured rows — zero runtime dependencies

已验证
deepseek-harnessdsh-pluginfeed-readerhtml-table
开发工具

dsh-fetch-third-party

tallahandsome-ux

Safe third-party web fetch for DSH — delegates page crawling to user-configured services with no direct URL access (no SSRF), API keys in the managed credentials store, per-session budget, and an auto-managed local web scraping service stack.

已验证
crawl4aideepseek-harnessdsh-pluginweb-scraping