PageIndex gives you accurate, trustworthy answers across long, professional documents your work depends on. Bring in your entire document set, ask your hardest question, and click any citation to jump to the exact highlighted source line, so you can verify it in seconds.
@henry_habib Thanks Henry! Exact-line citations are one of my favorite parts too — we really want every answer to be easy to verify, not just sound convincing. Appreciate the support!
the one click jump back to the exact source line is the killer peice 🙌 nice work
Report
Looks like a life changer. How well can it handle unusual documents like textbooks is going to be nice to see. 👋😄
Report
Congrats on launching PageIndex! Verifying answers with clickable citations solves a real pain point. How do you handle documents with scanned or image based pages?
@jackthompson68 Thanks Jack! Scanned and image-based PDFs are supported too. PageIndex runs OCR automatically, then indexes the extracted content and document structure as usual. You still get exact line references, so you can click a citation and jump straight to the precise source lines behind the answer.
Congrats! Avoiding embeddings is an ambitious architectural choice. I'm especially curious whether this makes the results more interpretable, since each search decision could potentially be traced through the document tree.
@william_wang24 Exactly — that’s one of the things we care about most. Because retrieval happens over the document structure rather than an embedding space, we can make the reasoning path much easier to inspect, and then ground the final answer back to the exact source lines. Thanks William!
Love that I don't have to keep scrolling through a 500-page PDF anymore!
PageIndex
@lantian Haha yes, that’s one of the best parts 😄
Voquill
I like the exact-line citations. Congrats!
PageIndex
@henry_habib Thanks Henry! Exact-line citations are one of my favorite parts too — we really want every answer to be easy to verify, not just sound convincing. Appreciate the support!
FunBlocks MindMax
Really good product from popular opensource lib. Congrats on this launch!
PageIndex
@peng_wood Thank you! Happy to keep contributing to open source :)
Macaly
the one click jump back to the exact source line is the killer peice 🙌 nice work
Looks like a life changer. How well can it handle unusual documents like textbooks is going to be nice to see. 👋😄
Congrats on launching PageIndex! Verifying answers with clickable citations solves a real pain point. How do you handle documents with scanned or image based pages?
PageIndex
@jackthompson68 Thanks Jack! Scanned and image-based PDFs are supported too. PageIndex runs OCR automatically, then indexes the extracted content and document structure as usual. You still get exact line references, so you can click a citation and jump straight to the precise source lines behind the answer.
Acti
Congrats! Avoiding embeddings is an ambitious architectural choice. I'm especially curious whether this makes the results more interpretable, since each search decision could potentially be traced through the document tree.
PageIndex
@william_wang24 Exactly — that’s one of the things we care about most. Because retrieval happens over the document structure rather than an embedding space, we can make the reasoning path much easier to inspect, and then ground the final answer back to the exact source lines. Thanks William!