Document Search Index

Document Search Index

Files and index live only in this browser tab. Image-only PDF scans need OCR. Opening a PDF page depends on the browser viewer's #page support.

Build document index

Up to 8 files · 10 MiB each · 24 MiB total · 50 PDF pages per file · UTF-8 text only

Indexed: 0 files · 0 PDF pages · 0 locations · 0 source bytes

No indexed documents yet.

Search index

Case-insensitive Korean and English search. A phrase must be in the same PDF page or text line; all-word mode allows any order. Up to 80 characters and six words.

Comments & questions

Document Search Index

Build an in-memory index of selectable text from multiple local documents by PDF page or text line. Search an exact phrase or all terms, read a context snippet and open the corresponding source PDF page. Unlike a single-text replacement or a predefined book-term index, this explores an ad hoc file collection.

Key features

  • In-memory inverted index for up to eight documents with PDF pages and TXT/MD lines
  • Korean and English exact phrase or all-term search with context snippets
  • Open a PDF result at its original local page or inspect a text line number
  • Removing a document deletes its posting lists and resets stale results
  • Explicit file, page, character and result bounds with a JSON search manifest

How to use

  1. Choose multiple local PDF, TXT or MD files, or load the example.
  2. Review indexed documents and page or line counts.
  3. Enter terms, choose phrase or all-word matching, and search.
  4. Inspect snippets and positions; open an original PDF page when needed.
  5. Remove an unneeded document, export JSON or clear the whole index.

Use cases

  • Find a contract phrase across several meeting PDFs
  • Search Korean notes and a PDF report in one session
  • Remove one document and compare refreshed results
  • Export document, location and excerpt evidence as JSON

Frequently asked questions

Does it search image-only scanned PDFs?

No. It reads selectable text already embedded in the PDF. Run OCR to make a scanned PDF searchable first, then check OCR accuracy against the source.

Where are my documents and index stored?

Only in this tab's memory. Closing or refreshing the tab removes the index. This tool does not upload your documents.

Can removed documents still appear in results?

No. Removal clears that document's page or line segments and inverted-index postings, and invalidates the previous result list.

Can extracted PDF text order differ from the visible page?

Yes. Multi-column layouts, tables and vertical writing can produce a different extraction order. Use the page and snippet as clues and check the original page.

What formats and limits apply?

Up to eight PDF, UTF-8 TXT or MD files. Each may be 10 MiB, all files together 24 MiB; PDFs are limited to 50 pages each and 120 total. Extracted text is bounded too. Encrypted and image-only PDFs cannot be searched.

Privacy

Documents and postings stay in this tab's browser memory; they are not uploaded or saved in browser storage. PDF code and worker load from this site's static assets. Exported JSON contains document names and excerpts; review it before sharing.

References

Related Tools

Searchable Scan PDF MakerPDF Merge/SplitCorpus ConcordancerDOCX Style AuditorCitation ManagerSigned PDF InspectorPDF/A Archival PreflightPrint Preflight AuditorPublication Accessibility AuditorResearch Evidence MatrixPDF Form DesignerPDF Annotation StudioPDF Object InspectorPDF Outline EditorEPUB Authoring WorkbenchDocument Version ComparatorBooklet Imposition DesignerSensitive Text RedactorHanja Reading WorkbenchGrammar Production LabDialogue Script EditorBook Index BuilderBilingual QA CheckerText Diagram EditorANSI Art StudioEmail Thread ExplorerPDF Redaction StudioRich Text SanitizerFixed Width Record DesignerPresentation Rehearsal StudioText Tokenization LabThree Way Text MergeKorean & English Braille ConverterText to SpeechTyping speed testKorean Text Pattern ReviewKorean Spelling QuizFont Preview and ComparisonHTML to MarkdownMojibake RepairSort LinesFind & ReplaceReverse TextRoman Numeral ConverterHangul Jamo ConverterKorean Initial ConsonantsAdd or Remove Line NumbersNumber to Korean WordsMorse Code ConverterROT13 & Caesar ConverterURL Slug GeneratorHTML Tag RemoverEmoji CollectionCharacter CounterUnit ConverterFile Size ConverterColor Code ConverterText DiffBusiness Day CalculatoriCalendar Rule LabvCard Address Book EditorJPG to PDFKorean Name RomanizerOrganize PDF PagesFancy Font GeneratorContrast CheckerColor Palette GeneratorPDF to JPGLorem Ipsum GeneratorText Template WorkbenchKorean-English Typo ConverterText CleanerOnline NotepadWord Cloud GeneratorPDF UnlockEnglish Address ConverterPDF Compressor
Explore all Text/Convert tools →Image/Media →Life/Fun →Dev Tools →