RAG/LLMパイプライン向けにPDFコンテンツをLlamaIndex JSON形式で抽出する。
クリックしてファイルを選択します またはドラッグ&ドロップ
1つ以上のPDFファイル
ファイルがデバイスの外に出ることはありません。
Output Format:
Each PDF will be extracted as a JSON file containing an array of LlamaIndex Document objects with:
text
text - Extracted text content per page
metadata
metadata - Page number, headings, and document info
extra_info
extra_info - Additional context for RAG systems
処理中...
ファイルをクリックまたはドラッグ&ドロップして始めます
プロセスボタンをクリックしてスタートします
処理したファイルをすぐに保存します
Yes! TuHe PDF is 100% free with no hidden fees, no signup required, and unlimited file processing.
Absolutely! All processing happens in your browser. Your files never leave your device, ensuring complete privacy.
No! Process files of any size, as many times as you want, completely free.