Wu, Daiqing
- Track the Answer: Extending TextVQA from Image to Video with Spatio-Temporal Clues
2024/12/17 by Yan Zhang, Zhang, Yan, Gangyan Zeng +9 · 4 citations
Computer Science · #Natural Language Processing Techniques #Topic Modeling #Multimodal Machine Learning Applications
- Char-SAM: Turning Segment Anything Model into Scene Text Segmentation Annotator with Character-level Visual Prompts
2024/12/27 by Enze Xie, Xie, Enze, Lyu, Jiaho +6 · 1 citation
Computer Science · #Topic Modeling #Advanced Text Analysis Techniques #Text and Document Classification Technologies
- Gather and Trace: Rethinking Video TextVQA from an Instance-oriented Perspective
2025/08/06 by Zhang, Yan, Zeng, Gangyan, Wu, Daiqing +5 · 2 citations
#Artificial Intelligence (cs.AI) #Computer Vision and Pattern Recognition (cs.CV) #FOS: Computer and information sciences