VERA: Identifying and Leveraging Visual Evidence Retrieval Heads in Long-Context Understanding Paper • 2602.10146 • Published Feb 9
Towards a Densing Law for User Representation Learning at Billion-Scale Capacity Paper • 2608.23392 • Published 12 days ago • 28
LATTEArena: An Evaluation Framework for LLM-powered Tabular Feature Engineering (Extended Version) Paper • 2606.09004 • Published Jun 16 • 2
Efficient Inference for Large Vision-Language Models: Bottlenecks, Techniques, and Prospects Paper • 2604.05546 • Published Apr 14 • 1
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare Paper • 2605.11814 • Published May 12 • 2
Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models Paper • 2605.09681 • Published May 10 • 10
Token Economics for LLM Agents: A Dual-View Study from Computing and Economics Paper • 2605.09104 • Published May 9 • 6
CoIDO: Efficient Data Selection for Visual Instruction Tuning via Coupled Importance-Diversity Optimization Paper • 2510.17847 • Published Oct 11, 2025
AnyTalker: Scaling Multi-Person Talking Video Generation with Interactivity Refinement Paper • 2511.23475 • Published Nov 28, 2025 • 44
Time-VLM: Exploring Multimodal Vision-Language Models for Augmented Time Series Forecasting Paper • 2502.04395 • Published Feb 6, 2025
AlayaDB: The Data Foundation for Efficient and Effective Long-context LLM Inference Paper • 2504.10326 • Published Apr 14, 2025 • 25
Train Small, Infer Large: Memory-Efficient LoRA Training for Large Language Models Paper • 2502.13533 • Published Feb 19, 2025 • 13
Draft & Verify: Lossless Large Language Model Acceleration via Self-Speculative Decoding Paper • 2309.08168 • Published Sep 15, 2023
AnyTalker: Scaling Multi-Person Talking Video Generation with Interactivity Refinement Paper • 2511.23475 • Published Nov 28, 2025 • 44
Towards a Densing Law for User Representation Learning at Billion-Scale Capacity Paper • 2608.23392 • Published 12 days ago • 28
LATTEArena: An Evaluation Framework for LLM-powered Tabular Feature Engineering (Extended Version) Paper • 2606.09004 • Published Jun 16 • 2
MedMemoryBench: Benchmarking Agent Memory in Personalized Healthcare Paper • 2605.11814 • Published May 12 • 2
Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Models Paper • 2605.09681 • Published May 10 • 10
Token Economics for LLM Agents: A Dual-View Study from Computing and Economics Paper • 2605.09104 • Published May 9 • 6