Results for “document-ingestion”

57 skills
More results
tianhao909
llamaindex
Data framework for building LLM applications with RAG. Specializes in document ingestion (300+ connectors), indexing, and querying. Features vector indices, query engines, agents, and multi-modal support. Use for document Q&A, chatbots, knowledge retrieval, or building RAG pipelines. Best for data-centric LLM applications.
1 · bundle
pawbytes
paw-pa-library
Ingests case studies, past proposals, and boilerplate from library/inbox into structured indexes. Use when the user requests to 'index proposal library', 're-index case studies', 'ingest inbox docs', or 'validate library'.
85 · bundle
theheavenlyd3mon
ingest
Team-Wiki Ingest
28
omer-metin
document-ai
Comprehensive patterns for AI-powered document understanding including PDF parsing, OCR, invoice/receipt extraction, table extraction, multimodal RAG with vision models, and structured data output. Use when "document parsing, PDF extraction, OCR, invoice processing, receipt extraction, document understanding, LlamaParse, Unstructured, vision document, table extraction, structured output from PDF, " mentioned.
128 · bundle
drnabeelkhan
wiki-ingest
Converts raw, unstructured sources into structured wiki pages with YAML frontmatter, Counter-Arguments sections, and bidirectional wikilinks, storing them in MemPalace.
2
georgeqle
quiz-me
Adversarial one-question-at-a-time questioning to verify deep reading comprehension of alignment pages, specs, or any document
1 · bundle
nvidia
vss-search-archive
Search archived video using natural language, ingest video files or RTSP streams, and manage ingested sources.
2.2k · bundle
nvidia
nemo-retriever
Index folders of PDFs and other documents into LanceDB for vector search, then query them with semantic search, page filters, verbatim quotes, and cross-document aggregation.
2.2k · bundle
github
acquire-codebase-knowledge
Maps, documents, and onboards into an existing codebase by generating seven structured documents covering stack, structure, architecture, conventions, integrations, testing, and concerns.
36.2k · bundle
drnabeelkhan
notebooklm-integration
Wraps the notebooklm-py CLI to ingest sources and generate synthesized artifacts like podcasts, slide decks, and quizzes from Google NotebookLM.
2
scoheart
resource-gatherer
Acquires resources from URLs, PDFs, and local files, categorizes them, and organizes content and media into a structured workspace with a lowercase assets folder.
2
bitwikiorg
document
Imported skill document from anthropic
3
mukul975
detecting-indirect-prompt-injection
Detect and defend against prompt injection hidden in documents, web pages, and images consumed by an agent.
24.6k · bundle
bitwikiorg
readme
Imported skill readme from openai
3
a5c-ai
feature-intake
Feature Intake
1.7k · bundle
joshuashepherd
book-ingest
Upserts a validated MDX book corpus into Supabase via Drizzle, hydrating books, chapters, sections, and chunks tables while preserving stable bookmark anchors and only re-embedding changed content.
1
salacoste
bmad-distillator
Lossless LLM-optimized compression of source documents. Use when the user requests to 'distill documents' or 'create a distillate'.
1 · bundle
affaan-m
videodb
Ingest, index, search, edit, and generate video and audio content from files, URLs, live streams, or desktop capture.
226k · bundle
peteedoo
brain-to-docs
Extract project vision and decisions from a user through iterative Q&A, then convert them into clear repo documentation (README, ADRs, lessons). Use when user asks to document ideas in their head.
0
jarbitechture
docx
Use this skill whenever the user wants to create, read, edit, or manipulate Word documents (.docx files). Triggers include: any mention of "Word doc", "word document", ".docx", or requests to produce professional documents with formatting like tables of contents, headings, page numbers, or letterheads. Also use when extracting or reorganizing content from .docx files, inserting or replacing images in documents, performing find-and-replace in Word files, working with tracked changes or comments, or converting content into a polished Word document. If the user asks for a "report", "memo", "letter", "template", or similar deliverable as a Word or .docx file, use this skill. Do NOT use for PDFs, spreadsheets, Google Docs, or general coding tasks unrelated to document generation.
0 · bundle
mhassan0000
docx
Creates, edits, and analyzes Word documents (.docx/.dotx) using docx-js, XML manipulation, and conversion tools, including tracked changes and comments.
1 · bundle
seb1n
context-injection
Place trusted contextual information into prompts or agent state using explicit boundaries, provenance, and templates. Use when relevant context has already been selected and must be inserted safely; use context-retrieval to find it or context-optimization to choose and order it.
159
sinhoneyy
docx
Comprehensive document creation, editing, and analysis with support for tracked changes, comments, formatting preservation, and text extraction. When Claude needs to work with professional documents (.docx files) for: (1) Creating new documents, (2) Modifying or editing content, (3) Working with tracked changes, (4) Adding comments, or any other document tasks
11 · bundle
winbda
intake-form
Create patient intake forms with history and screening. TRIGGERS - Use when user needs help with intake-form related tasks.
3
memento-teams
docx
Create, read, edit, and manipulate Word documents (.docx) with formatting, tables, images, headers, footers, tracked changes, and comments.
1.5k · bundle
sakamoto-family-smile
videodb
Ingests video and audio from files, URLs, RTSP feeds, or desktop capture; indexes and searches moments with timestamps; transcodes, edits timelines, generates media assets, and creates real-time alerts for live streams.
0 · bundle
seaworld008
tome
Converting repository changes into detailed learning documents. Use when turning diffs into teaching materials, recording design decisions, or creating onboarding materials for new members.
65 · bundle
michaelschecht
docx
Use this skill whenever the user wants to create, read, edit, or manipulate Word documents (.docx files). Triggers include: any mention of 'Word doc', 'word document', '.docx', or requests to produce professional documents with formatting like tables of contents, headings, page numbers, or letterheads. Also use when extracting or reorganizing content from .docx files, inserting or replacing images in documents, performing find-and-replace in Word files, working with tracked changes or comments, or converting content into a polished Word document. If the user asks for a 'report', 'memo', 'letter', 'template', or similar deliverable as a Word or .docx file, use this skill. Do NOT use for PDFs, spreadsheets, Google Docs, or general coding tasks unrelated to document generation.
0 · bundle
lambenthan
init
基于用户素材与可选外部发现搭建 ΩmegaWiki,并用并行 `/ingest` 完成最终论文集的消化
77 · bundle
lambenthan
ingest
把一篇论文 ingest 进 wiki —— 建立 papers + concepts + people + claims 页面,并完成所有双向交叉引用与 graph edge。当用户说 "ingest"、"加入这篇论文"、丢 `.pdf` / `.tex` / arXiv URL 或要求把论文折叠进知识库时触发。
77 · bundle
joshuashepherd
pathway-organizer
Organizes discovered pathway, portal, and theme content into a docs repository by copying new files, resolving duplicates, updating the README, and generating per-portal indexes.
1
lambenthan
daily-arxiv
每日从 arXiv 拉取新论文,过滤相关性,auto-ingest 高优先级论文,检测 SOTA 更新
77
zero-yx
music-ingestion-publisher
Ingests music into a LanceDB via ncmdump-rs and sf-cli, searching and downloading from Netease Cloud Music or Bilibili, decrypting local NCM files, and extracting metadata, lyrics, and cover art.
0 · bundle