I am setting up a document ingestion pipeline that handles both text extraction and web-based document…
I am setting up a document ingestion pipeline that handles both text extraction and web-based document…: a task in SkillNet-Gym (Harbor dataset). Because the document contains very small text, standard low-resolution processing is insufficient. Please execute the following pipeline:
Part of zjunlp/SkillNet-Gym.