- Change DocumentRecord.docling_document_json (str) to docling_document (bytes) - Use large_binary Arrow type (64-bit offsets) to avoid 2GB column limit - Decompress in Document.get_docling_document() using gzip - Update DocumentRepository field mappings |
||
|---|---|---|
| .. | ||
| __init__.py | ||
| chunk.py | ||
| document.py | ||