UI-Mate: Advancing Open-Weight Foundation GUI Agents with In-Context Demonstrations Paper • 2608.15930 • Published 4 days ago • 40
From Documents to Spans: Scalable Supervision for Evidence-Based ICD Coding with LLMs Paper • 2603.15270 • Published May 7 • 1
From Perception to Cognition: A Survey of Vision-Language Interactive Reasoning in Multimodal Large Language Models Paper • 2509.25373 • Published Sep 29, 2025
U-Bench: A Comprehensive Understanding of U-Net through 100-Variant Benchmarking Paper • 2510.07041 • Published Oct 8, 2025 • 5
Hi-End-MAE: Hierarchical encoder-driven masked autoencoders are stronger vision learners for medical image segmentation Paper • 2502.08347 • Published Feb 12, 2025