Hacktoberfest 2026:メンテナが10月に向けて印を付けた、オープンで初心者向けの issue。 Hacktoberfest の issue を見る

Bring Docker API server to feature parity with Crawl4AI v0.7.x

オープン
#1,452 コメント 3 件 リアクション 4 件 担当者 0 名 GitHub で見る

メンテナーはふだん 1 日以内に返信

まだ誰も着手していません。

評価

難易度
5/5
見積もり時間
1週間以上
初心者へのやさしさ
25/100
issue の種類
機能追加
明瞭さ
説明が足りない
活発さ
停滞
技術スタック
docker, python

調査の方向性

Start by reviewing the existing Docker API routes and schemas.py, then compare them with the listed Crawl4AI library feature groups. Trace the requested endpoints, request schemas, documentation, examples, and minimal end-to-end tests. Done means the missing feature groups are exposed with documented parameters and validated happy and error paths.

索引モデルが issue の本文から書いたものです。

説明

⚙️ Under Review ✨ Enhancement

Description

The Docker API server lags behind the Python library. This issue tracks adding endpoints/parameters to expose the following library features:

1. Adaptive crawling
  • AdaptiveCrawler, AdaptiveConfig, CrawlState, CrawlStrategy, StatisticalStrategy
  • Missing: endpoints to run/tune adaptive crawls
2. C4A Script language
  • c4a_compile, c4a_validate, c4a_compile_file, CompilationResult, ValidationResult, ErrorDetail
  • Missing: submit/validate/execute script endpoints
3. URL seeding
  • AsyncUrlSeeder, SeedingConfig
  • Missing: sitemap/common-crawl/discovery endpoints
4. Chunking
  • ChunkingStrategy, RegexChunking
  • Missing: chunking configuration
5. Browser adapters
  • BrowserAdapter, PlaywrightAdapter, UndetectedAdapter
  • Missing: adapter/stealth selection
6. Proxy rotation
  • ProxyRotationStrategy, RoundRobinProxyStrategy
  • Missing: rotation strategy selection (beyond raw proxy)
7. Dispatchers
  • SemaphoreDispatcher, BaseDispatcher
  • Missing: dispatcher selection (only MemoryAdaptive used internally)
8. Link preview
  • LinkPreview, LinkPreviewConfig
  • Missing: link preview/scoring endpoint
9. Profiling/monitoring
  • BrowserProfiler, CrawlerMonitor
  • Missing: profiling/monitoring endpoints
10. HTTP-only crawling
  • HTTPCrawlerConfig
  • Missing: HTTP crawler methods/params (non-browser). API uses browser-based crawling with LXMLWebScrapingStrategy
11. Virtual scroll
  • VirtualScrollConfig
  • Missing: infinite-scroll capture configuration
12. Undetected/stealth browser
  • UndetectedAdapter; browser_config/browser_type='undetected'; stealth options
  • Missing: explicit stealth mode controls
Acceptance criteria
1. New/extended endpoints and/or request schemas added
  • New endpoints: Add missing API routes (e.g., /adaptive/crawl, /deep-crawl, /c4a-script/compile, /hub/crawlers)
  • Extended schemas: Enhance existing endpoints to accept new parameters (e.g., add virtual_scroll_config to /crawl, add table_extraction_strategy options)
  • Request schemas: Update schemas.py to include new request models for the missing features
2. Docs and examples updated
  • API documentation: Update the docs to show new endpoints and parameters
  • Parameter documentation: Add descriptions, examples, and validation rules for new fields
  • Examples: Add working code examples showing how to use each new feature.
3. Minimal e2e tests per feature group
  • Test coverage: Create integration tests that verify each new feature works end-to-end
  • Happy path: Test successful usage of each feature
  • Validation: Test error handling (invalid parameters, edge cases, etc.)
  • Feature groups: Organize tests by category (adaptive crawling, deep crawling, C4A scripts, etc.)
主要言語
Python
スター
84.5k
フォーク
8.7k
平均マージ
3日 9時間
マージ済み PR(30日)
17

環境構築

はじめの一歩

  1. issue を最後まで読み、次にプロジェクトのコントリビューションガイドを読みます。
  2. 着手することを issue にコメントします — 二人が同じ作業をするのを防げます。
  3. リポジトリをフォークし、ブランチを切って変更します。
  4. issue 番号を参照したプルリクエストを送ります。

unclecode/crawl4ai のほかの issue

unclecode/crawl4ai の issue をすべて見る

似ている issue

Python の issue をもっと見る

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。