PDFを選択してください
Read text coordinates, image positions and shared resources per page.
Native text is read directly, while every eligible page image is recognized with PP-OCRv6. Overlap with hidden OCR layers is deduplicated.
これらの事実は現在の製品を説明するものであり、未出荷のロードマップ作業を説明するものではありません。
DataDance 製品およびエンジニアリング チームによるレビュー · 公開 · 更新されました
エクスポートする前に重要な設定を確認してください。処理場所は常に開示されます。
Read text coordinates, image positions and shared resources per page.
Auto uses Professional on desktop and lower-memory Fast on mobile; desktop can also select Fast, Professional or Ultimate.
Merge native text, images and non-duplicate OCR into a self-contained long page.
Yes. Native words come from PDF text operators; scans and images inside mixed pages use ImgIng PP-OCRv6. Both streams are merged by page coordinates.
Auto chooses Professional on desktop and Fast on mobile; desktop users can choose Fast, Professional or Ultimate. On first use the selected model is downloaded and verified from multiple model CDN sources, automatically falls back when one fails, and is then reused from the browser cache.
Normalized OCR output is compared with the page’s native text. Existing phrases are not inserted twice, while unmatched text and the source image remain available for review.
HTML keeps page markers, paragraphs, images and recognized image text together. TXT and Markdown provide lighter content-only alternatives.
Updated 2026-08-30 · Live capability detection inside the tool is authoritative
これらの目に見える答えは、現在の製品の動作と構造化データと一致します。
Yes. Every eligible page image is OCR-processed and merged with native text.
Overlapping normalized text is removed automatically.
No. PDF parsing, OCR, layout reconstruction and export stay local.
各タスク ページには、キーワードを交換した重複ではなく、実際の設定、制限、形式に関するアドバイスが文書化されています。