Product
Introducing License Enforcement in Socket
Ensure open-source compliance with Socket’s License Enforcement Beta. Set up your License Policy and secure your software!
Docling Core is a library that defines the data types in Docling, leveraging pydantic models.
To use Docling Core, simply install docling-core
from your package manager, e.g. pip:
pip install docling-core
To develop for Docling Core, you need Python 3.9 / 3.10 / 3.11 / 3.12 / 3.13 and Poetry. You can then install from your local clone's root dir:
poetry install
To run the pytest suite, execute:
poetry run pytest test
You can validate your JSON objects using the pydantic class definition.
from docling_core.types import Document
data_dict = {...} # here the object you want to validate, as a dictionary
Document.model_validate(data_dict)
data_str = {...} # here the object as a JSON string
Document.model_validate_json(data_str)
You can generate the JSON schema of a model with the script generate_jsonschema
.
# for the `Document` type
generate_jsonschema Document
# for the use `Record` type
generate_jsonschema Record
Docling supports 3 main data types:
The data schemas are defined using pydantic models, which provide built-in processes to support the creation of data that adhere to those models.
Please read Contributing to Docling Core for details.
If you use Docling Core in your projects, please consider citing the following:
@techreport{Docling,
author = "Deep Search Team",
month = 8,
title = "Docling Technical Report",
url = "https://arxiv.org/abs/2408.09869",
eprint = "2408.09869",
doi = "10.48550/arXiv.2408.09869",
version = "1.0.0",
year = 2024
}
The Docling Core codebase is under MIT license. For individual model usage, please refer to the model licenses found in the original packages.
FAQs
A python library to define and validate data types in Docling.
We found that docling-core demonstrated a healthy version release cadence and project activity because the last version was released less than a year ago. It has 1 open source maintainer collaborating on the project.
Did you know?
Socket for GitHub automatically highlights issues in each pull request and monitors the health of all your open source dependencies. Discover the contents of your packages and block harmful activity before you install or update your dependencies.
Product
Ensure open-source compliance with Socket’s License Enforcement Beta. Set up your License Policy and secure your software!
Product
We're launching a new set of license analysis and compliance features for analyzing, managing, and complying with licenses across a range of supported languages and ecosystems.
Product
We're excited to introduce Socket Optimize, a powerful CLI command to secure open source dependencies with tested, optimized package overrides.