Add Real5-OmniDocBench evaluation results for NaviDC-OCR

#1

Real5-OmniDocBench evaluation results for NaviDC-OCR

This PR adds the evaluation file for NaviDC-OCR under the Hugging Face
evaluation-results workflow.

This is based on the new evaluation results feature: https://huggingface.co/docs/hub/eval-results.

Dataset

File

  • .eval_results/real5_omnidocbench.yaml

Tasks

  • overall
  • scanning
  • warping
  • screen_photography
  • illumination
  • skew
StarDoc-AI org

Thank you for your interest in our work. However, there is a significant discrepancy between your evaluation results and our own evaluation results. If you are willing to share your evaluation results with us, we would be more than happy to analyze the issue. The table below shows our evaluation results.
image

StarDoc-AI org

Thank you for your interest. This is the Markdown file from our local evaluation. I’m curious to understand where the discrepancy comes from. If you are willing to provide some detailed evaluation results, that would be very helpful for analyzing the discrepancy.
https://drive.google.com/file/d/1XihsGFTfmhqTzogMJkV07K7XINNsbAQO/view

Thank you very much for your reply. Based on the results you provided, we conducted a repeat test and verification, and the results are largely consistent with the metrics reported in our PR. Therefore, the issue may lie in your evaluation environment.

Real5-OmniDocBench is strictly aligned with OmniDocBench V1.5. Therefore, both the evaluation environment and evaluation code should strictly follow the OmniDocBench V1.5 version for evaluation.

caipeng328 changed pull request status to closed

Sign up or log in to comment