SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models Paper • 2608.10538 • Published 10 days ago • 15
Power law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference Paper • 2608.10288 • Published 11 days ago • 7
FocusMem: Factorizing Content, Readout, and Trust in Latent GUI Memory Paper • 2608.04530 • Published 16 days ago • 14
Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing Paper • 2608.02711 • Published 17 days ago • 90
QQWorld: Quantile-Quantile Matching for World Model Regularization Paper • 2607.28415 • Published 22 days ago • 30