- 17 Jul, 2024 4 commits
- 15 Jul, 2024 1 commit
-
-
myhloli authored
-
- 14 Jul, 2024 2 commits
-
-
myhloli authored
Improve the model loading mechanism in magic_pdf by implementing a Singleton pattern to reduce redundant model instantiation. Additionally, enhance the command-line interface to support input from list files, allowing batch processing of multiple PDF documents.
-
myhloli authored
Introduce a Singleton pattern to manage custom models in the magic_pdf module. This change improves the efficiency by ensuring that a single instance of the custom model is created and reused, thereby reducing the overhead of multiple instantiate calls for the same model configuration.
-
- 13 Jul, 2024 2 commits
- 12 Jul, 2024 7 commits
-
-
myhloli authored
-
myhloli authored
-
zhaoxiaomeng authored
-
myhloli authored
-
myhloli authored
-
myhloli authored
Add new configuration options for custom model directories and device modeselection. This allows users to specify the directory where models are stored and choose between CPU and GPU modes for model inference. The configurations are read from a JSON file and can be easily extended to support additional options in the future.
-
myhloli authored
-
- 11 Jul, 2024 5 commits
-
-
myhloli authored
Introduce a new feature that allows users to choose between a "lite" and a "full" model mode for PDF document analysis. The "lite" mode uses a faster, less accurate model, while the "full" mode employs a higher-precision model at the cost of speed. This selection can be made through the CLI or API, providing flexibility for different use cases.
-
myhloli authored
update:Add md make mode config in do_parse.You can control whether the produced md is for NLP or MM by changing the value of f_make_md_mode
-
myhloli authored
-
myhloli authored
-
myhloli authored
-
- 10 Jul, 2024 3 commits
-
-
zhaoxiaomeng authored
-
zhaoxiaomeng authored
-
myhloli authored
-
- 09 Jul, 2024 2 commits
- 08 Jul, 2024 2 commits
- 07 Jul, 2024 1 commit
-
-
myhloli authored
-
- 05 Jul, 2024 1 commit
-
-
赵小蒙 authored
fix: The presence of ".pdf" multiple times in the pdf_path results in model_path not matching the expected.
-
- 28 Jun, 2024 4 commits
- 26 Jun, 2024 4 commits
- 25 Jun, 2024 2 commits