gpu-document-processing

Use when processing large PDFs, document collections, or bulk text extraction tasks that benefit from GPU-accelerated processing. Triggers when the user…

npx skills add https://github.com/langchain-ai/deepagents --skill gpu-document-processing

GPU Document Processing Skill

Process large documents and document collections using GPU-accelerated tools. This skill uses the sandbox-as-tool pattern: the agent runs on CPU for reasoning, and sends document processing work to a GPU-equipped environment.

When to Use This Skill

Use this skill when:

  • Processing large PDF files (50+ pages)
  • Analyzing collections of documents (10+ files)
  • Extracting structured data from unstructured documents
  • Performing bulk text extraction and chunking
  • Generating embeddings for large document sets
  • The user uploads or references large documents for analysis

Architecture: Sandbox as Tool

This skill follows the sandbox-as-tool pattern for GPU execution:

  1. Agent reasons on CPU - planning, synthesis, report writing
  2. Processing sent to GPU sandbox - document parsing, embedding, extraction
  3. Results returned to agent - structured output for further analysis

This separation ensures:

  • API keys stay outside the sandbox (security)
  • Agent state persists independently of processing jobs
  • Processing can be parallelized across documents
  • Cost-efficient: GPU used only during processing, not during reasoning

Capabilities

PDF Text Extraction

Extract text content from PDF documents with layout preservation:

  • Headers, paragraphs, lists, and tables detected separately
  • Page numbers and section boundaries preserved
  • Multi-column layout handling

Tabular Data Extraction

Extract tables from documents into structured formats:

  • PDF tables to CSV/DataFrames using GPU-accelerated parsing
  • Automatic column type detection
  • Handles merged cells and multi-row headers

Document Chunking

Split large documents into meaningful chunks for analysis:

  • Semantic chunking (by topic/section boundaries)
  • Fixed-size chunking with overlap for embedding
  • Configurable chunk sizes (default: 512 tokens)

Embedding Generation

Generate vector embeddings for document chunks:

  • Uses NVIDIA NeMo Retriever NIM for GPU-accelerated embedding
  • Supports batch processing for large document sets
  • Compatible with standard vector stores (Milvus, ChromaDB)

Workflow

  1. Receive document reference from the orchestrator
  2. Determine processing type (extraction, analysis, embedding)
  3. Send to GPU sandbox for processing
  4. Collect structured results (text, tables, embeddings)
  5. Write findings to /shared/ for the orchestrator to synthesize

Processing Large Document Collections

For multiple documents:

  1. Process documents in parallel batches (3-5 concurrent)
  2. Extract key metadata first (title, date, author, page count)
  3. Generate per-document summaries
  4. Cross-reference findings across documents
  5. Write consolidated findings with per-document citations

Output Format

When reporting document processing results:

  • Include document metadata (filename, pages, size)
  • Structure extracted content by section/chapter
  • Format tables as markdown tables
  • Include page references for all extracted content
  • Note any extraction quality issues (scanned images, corrupted pages)

Integration with NVIDIA NIM

For production deployments, GPU document processing can leverage:

  • NVIDIA NeMo Retriever: GPU-accelerated embedding and retrieval
  • NVIDIA RAPIDS cuDF: Tabular data processing from extracted tables
  • NVIDIA Triton: Scalable inference for document classification models

See NVIDIA's NIM documentation for self-hosted deployment options.

Lebih banyak skill dari langchain-ai

langgraph-docs
langchain-ai
Mengakses dokumentasi LangGraph untuk membangun agen stateful dan alur kerja multi-agen. Mengambil dokumentasi resmi LangGraph Python yang mencakup mesin state, desain agen berbasis grafik, dan pola human-in-the-loop. Memprioritaskan dokumentasi yang relevan berdasarkan jenis kueri: panduan implementasi untuk pertanyaan cara, halaman konsep untuk teori, tutorial untuk contoh ujung ke ujung, dan referensi API untuk detail teknis. Secara otomatis memilih 2–4 URL dokumentasi yang paling relevan dan mengambil kontennya untuk menjawab...
official
langgraph-human-in-the-loop
langchain-ai
Jeda eksekusi graf untuk peninjauan, persetujuan, atau validasi manusia, lalu lanjutkan dengan masukan mereka. Membutuhkan tiga komponen: checkpointer (InMemorySaver atau PostgresSaver), ID thread dalam konfigurasi, dan payload interupsi yang dapat diserialisasi JSON. interrupt(value) menjeda dan menampilkan data; Command(resume=value) melanjutkan dan mengembalikan nilai tersebut ke node yang dijeda. Semua kode sebelum interrupt() akan dieksekusi ulang saat melanjutkan, sehingga efek samping harus idempoten (gunakan upsert, bukan insert). Mendukung alur kerja persetujuan,...
official
web-research
langchain-ai
Gunakan keterampilan ini untuk permintaan yang terkait dengan riset web; ini menyediakan pendekatan terstruktur untuk melakukan riset web yang komprehensif.
official
langchain-oss-primer
langchain-ai
SELALU MULAI DI SINI untuk proyek pembuatan agen LangChain, Deep Agents, atau LangGraph apa pun. Titik awal yang diperlukan sebelum memilih keterampilan lain atau menulis apa pun…
official
skill-creator
langchain-ai
Panduan untuk membuat skill yang efektif guna memperluas kemampuan agen dengan pengetahuan khusus, alur kerja, atau integrasi alat. Gunakan skill ini ketika pengguna…
official
social-media
langchain-ai
Menyusun draf posting media sosial khusus platform dengan konten berbasis riset dan gambar pendamping yang dihasilkan. Mendukung posting LinkedIn (1.300 karakter dengan nada profesional) dan utas Twitter/X (280 karakter per tweet dengan format 1/🧵). Memerlukan delegasi riset ke subagen sebelum menulis, kemudian membaca temuan untuk memastikan akurasi dan relevansi. Menghasilkan gambar sosial yang menarik secara otomatis menggunakan alat generate_social_image dengan komposisi tebal dan kontras tinggi yang dioptimalkan untuk ukuran kecil...
official
deep-agents-memory
langchain-ai
Backend memori dan file yang dapat dipasang untuk Deep Agents dengan opsi perutean sementara, persisten, dan hibrida. Empat jenis backend: StateBackend (berlaku dalam thread, sementara), StoreBackend (persisten lintas sesi), FilesystemBackend (akses disk nyata untuk pengembangan lokal), dan CompositeBackend (merutekan jalur berbeda ke backend berbeda). FilesystemMiddleware menyediakan enam alat operasi file: ls, read_file, write_file, edit_file, glob, grep. CompositeBackend menggunakan pencocokan prefiks terpanjang untuk merutekan...
official
deep-agents-orchestration
langchain-ai
We need to translate the given English text into Indonesian. The text describes an agent skill for orchestrating subagents, planning tasks, requiring human approval, delegating work, etc. We must preserve product names, protocol names, URLs, numbers, technical terms. The name "deep-agents-orchestration" is not in the text, so we don't include it. We translate only the text inside <text>. No extra commentary, labels, etc. Let's translate step by step: "Orchestrate subagents, plan multi-step tasks, and require human approval for sensitive operations." -> "Orkestrasi subagen, rencanakan tugas multi-langkah, dan minta persetujuan manusia untuk operasi sensitif." "Delegate work to specialized subagents via the task tool; custom subagents support isolated tool sets and system prompts, while the default "general-purpose" subagent inherits main agent configuration" -> "Delegasikan pekerjaan ke subagen khusus melalui alat tugas; subagen kustom mendukung set alat dan prompt sistem yang terisolasi, s
official