Submitted by bingyi 15 TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment Google 3
1 CoGate-LSTM: Prototype-Guided Feature-Space Gating for Mitigating Gradient Dilution in Imbalanced Toxic Comment Classification Google
Submitted by Zhaochong An 64 VGGRPO: Towards World-Consistent Video Generation with 4D Latent Reward Google 3
Submitted by taesiri 9 VQQA: An Agentic Approach for Video Evaluation and Quality Improvement Google 1
Submitted by Zorik 76 Thinking to Recall: How Reasoning Unlocks Parametric Knowledge in LLMs Google 4
Submitted by Liuyanqing 5 CAST: Modeling Visual State Transitions for Consistent Video Retrieval Google 2
Submitted by taesiri 17 Discovering Multiagent Learning Algorithms with Large Language Models Google 2