mirror of
https://github.com/VectifyAI/PageIndex.git
synced 2026-10-02 07:44:37 +08:00
BREAKING: resolve_citations(), released in 0.2.17, is renamed
get_citations(); same arguments, same list. resolve_citations() now returns
{'answer', 'citations'}: the answer with each citation tag replaced by
[[i]](#pageindex-citation-0i), and the get_citations() entries led by
'anchor' and 'index'. Numbers follow the tags as written, so a repeated
citation reuses its number and a block cited under the wrong page still
gets its link while its entry carries the block's real page. One shared
_citation_key() parses a matched tag for both the parser and the rewrite.
highlight_region(image, bbox, scale=1000) draws the cited region on a page
image (PIL.Image or bytes in, PIL.Image out). Pillow becomes a dependency.
get_page_image(doc_id, page) and get_document_image(doc_id, img_id) return
short-lived URLs. Cloud-only. They need the compute routes that return
presigned URLs; against an older server they raise PageIndexAPIError.
get_document_path, get_folder_path and get_folder_id translate between ids
and readable paths. Paths are built from list_folders() parent links,
because the server's `path` field exists only on /docs listing rows. A path
shared by two folders raises instead of picking one.
16 lines
309 B
Plaintext
16 lines
309 B
Plaintext
litellm==1.97.0
|
|
openai>=1.70.0
|
|
requests>=2.28.0
|
|
# MCP retries use Retry(allowed_methods=...), added in urllib3 1.26.
|
|
urllib3>=1.26
|
|
openai-agents>=0.18.1
|
|
mcp>=1.19.0,<3
|
|
# pymupdf # optional
|
|
PyPDF2==3.0.1
|
|
pypdfium2==5.13.0
|
|
python-dotenv==1.2.2
|
|
pyyaml==6.0.2
|
|
regex>=2024.0.0
|
|
sortedcontainers==2.4.0
|
|
Pillow>=9.0
|