- 25 Jul, 2024 1 commit
-
-
myhloli authored
fix(pdf_extract_kit): specify utf-8 encoding when reading model configEnsure the model configuration file is read with utf-8 encoding to support non-ASCII characters and prevent potential encoding errors.
-
- 24 Jul, 2024 5 commits
-
-
myhloli authored
Specify utf-8 encoding when opening the configuration file to ensure compatibility with files containing non-ASCII characters, avoiding potentialencoding errors.
-
赵小蒙 authored
-
myhloli authored
-
myhloli authored
fix(magic-pdf): add default values and improve warning logs for config optionsEnsure that 'temp-output-dir', 'models-dir', and 'device-mode' have sensible default values in case they are not specified in the config file.
-
myhloli authored
-
- 23 Jul, 2024 6 commits
- 22 Jul, 2024 3 commits
- 19 Jul, 2024 3 commits
- 18 Jul, 2024 1 commit
-
-
myhloli authored
-
- 17 Jul, 2024 4 commits
- 15 Jul, 2024 1 commit
-
-
myhloli authored
-
- 14 Jul, 2024 2 commits
-
-
myhloli authored
Improve the model loading mechanism in magic_pdf by implementing a Singleton pattern to reduce redundant model instantiation. Additionally, enhance the command-line interface to support input from list files, allowing batch processing of multiple PDF documents.
-
myhloli authored
Introduce a Singleton pattern to manage custom models in the magic_pdf module. This change improves the efficiency by ensuring that a single instance of the custom model is created and reused, thereby reducing the overhead of multiple instantiate calls for the same model configuration.
-
- 13 Jul, 2024 2 commits
- 12 Jul, 2024 7 commits
-
-
myhloli authored
-
myhloli authored
-
zhaoxiaomeng authored
-
myhloli authored
-
myhloli authored
-
myhloli authored
Add new configuration options for custom model directories and device modeselection. This allows users to specify the directory where models are stored and choose between CPU and GPU modes for model inference. The configurations are read from a JSON file and can be easily extended to support additional options in the future.
-
myhloli authored
-
- 11 Jul, 2024 5 commits
-
-
myhloli authored
Introduce a new feature that allows users to choose between a "lite" and a "full" model mode for PDF document analysis. The "lite" mode uses a faster, less accurate model, while the "full" mode employs a higher-precision model at the cost of speed. This selection can be made through the CLI or API, providing flexibility for different use cases.
-
myhloli authored
update:Add md make mode config in do_parse.You can control whether the produced md is for NLP or MM by changing the value of f_make_md_mode
-
myhloli authored
-
myhloli authored
-
myhloli authored
-