Metadata-Version: 2.4
Name: docextract-core
Version: 0.1.0
Summary: Shared document-extraction substrate: hashing, JSON codec, archives, LLM protocol
Author: Ivan Ortega
License-Expression: Apache-2.0
Project-URL: Homepage, https://github.com/iortega10/word-extract
Project-URL: Source, https://github.com/iortega10/word-extract
Project-URL: Issues, https://github.com/iortega10/word-extract/issues
Classifier: Programming Language :: Python :: 3
Classifier: Programming Language :: Python :: 3.11
Classifier: Programming Language :: Python :: 3.12
Classifier: Programming Language :: Python :: 3.13
Classifier: Programming Language :: Python :: 3.14
Requires-Python: >=3.11
Description-Content-Type: text/markdown
License-File: LICENSE
License-File: NOTICE
Provides-Extra: dev
Requires-Dist: pytest>=8; extra == "dev"
Dynamic: license-file

# docextract-core

Shared substrate for the document-extraction packages ([`word-extract`](https://pypi.org/project/word-extract/)
and its siblings): content hashing, a strict canonical-JSON codec, archive helpers, a
content-addressed `Collection`, and a small LLM-client protocol. It has no runtime dependencies.

Most users want `word-extract`, which depends on this package.

Licensed under the Apache License, Version 2.0.
