Hello Numind team,
First of all, I would like to highlight that the work done with NuExtract really make sense from my point of view.
Being able to have a consistent schema for each document extraction is really important when processing millions of documents, particularly for industry use-cases.
With the recent release of Deepseek-OCR https://huggingface.co/deepseek-ai/DeepSeek-OCR and their new encoder compression, can we expected a similar fine-tune from you?
As well https://huggingface.co/PaddlePaddle/PaddleOCR-VL is an interesting one to consider.
Cheers
Hello Numind team,
First of all, I would like to highlight that the work done with NuExtract really make sense from my point of view.
Being able to have a consistent schema for each document extraction is really important when processing millions of documents, particularly for industry use-cases.
With the recent release of Deepseek-OCR https://huggingface.co/deepseek-ai/DeepSeek-OCR and their new encoder compression, can we expected a similar fine-tune from you?
As well https://huggingface.co/PaddlePaddle/PaddleOCR-VL is an interesting one to consider.
Cheers