PageIndex · Issues· 105 open
Open on GitHubLocally synced open issues (discussions stay on GitHub)
- #513
Implementing PageIndex for Tabular Code Datasets in a GitHub RAG Chatbot
Updated Sep 17, 2026 - #478
Please replace the deprecated PyPDF2 with PyPDF
Updated Sep 7, 2026 - #486
[Bug]: Bare except in page_index_md.py masks underlying import and syntax errors
Updated Sep 5, 2026 - #474
[Resource] GeoMind: Pre-indexed Knowledge Base Compatible with PageIndex Approach
Updated Sep 3, 2026 - #467
Oversized first page produces an empty chunk that reaches the model
Updated Sep 2, 2026 - #426
PyPDF2.errors.DependencyError: PyCryptodome is required for AES algorithm — crash on permission-only-encrypted PDFs (Flash preview)
Updated Aug 26, 2026 - #316
Feature Request: Incremental Index Updates for Large Documents
Updated Aug 3, 2026 - #355
Markdown path: below --summary-token-threshold, node text is copied verbatim into summary (undocumented; diverges from PDF path)
Updated Jul 16, 2026 - #330
Bug: `get_leaf_nodes` raises KeyError on leaf nodes due to missing `nodes` key
Updated Jul 13, 2026 - #175
Hello everyone, I have some questions regarding Hybrid Tree Search and would appreciate your insights.
Updated Jul 9, 2026 - #341
No-TOC fallback uses entire source sentence as node title when no short heading exists
Updated Jul 1, 2026 - #340
Sibling nodes on the same page get identical text → duplicate/near-duplicate summaries
Updated Jul 1, 2026 - #137
can web just start a pageindex backend service,and use pageindex-py-sdk to manipulate the pageindex service。 And especially docker compose is better
Updated Jun 15, 2026 - #295
Header elements overlap at widths above 1024px
Updated May 25, 2026 - #240
Security Policy missing
Updated Apr 22, 2026