Lv2
118 积分 2024-12-24 加入
OmniParser V2: Structured-Points-of-Thought for Unified Visual Text Parsing and Its Generality to Multimodal Large Language Models
4个月前
已完结
A 119.64 GOPs/W FPGA-Based ResNet50 Mixed-Precision Accelerator Using the Dynamic DSP Packing
5个月前
已完结
HTR-VT: Handwritten text recognition with vision transformer
6个月前
已完结
Scene table structure recognition with segmentation collaboration and alignment
6个月前
已完结
Multiple object detection and tracking from drone videos based on GM-YOLO and multi-tracker
7个月前
已完结