Weixian Waylon Li, Jiaxin Zhang, Xianan Jim Yang, Tiejun Ma, Yiwen Guo
RoMem models time in agent memory with continuous phase rotation and a learned volatility score so persistent facts stay stable while outdated facts are rotated out of phase instead of being deleted. It…
HuggingFacedaily curated papers2026-04-13Upvotes: 1
Zeyue Tian, Binxin Yang, Zhaoyang Liu, Jiexuan Zhang, Ruibin Yuan
Audio-Omni proposes an end-to-end framework that unifies audio understanding, generation, and editing across speech, music, and general sound, backed by a million-pair AudioEdit dataset. It matters because…
HuggingFacedaily curated papers2026-04-12Upvotes: 1
Yuzhen Mao, Qitong Wang, Martin Ester, Ke Li
IceCache clusters semantically related tokens and combines that layout with PagedAttention to keep long-context LLM inference accurate while moving more cache data off GPU. It matters because KV-cache memory…
HuggingFacedaily curated papers2026-04-12Upvotes: 0
Samuel Sameer Tanguturi
ATANT defines continuity as a measurable systems property and introduces a 250-story benchmark for testing whether AI memory stacks can preserve, update, and retrieve narrative truth over time without…
HuggingFacedaily curated papers2026-04-08Upvotes: 0
Hangoo Kang, Tarun Suresh, Jon Saad-Falcon, Azalia Mirhoseini
TRACE turns failed agent trajectories into capability-targeted synthetic RL environments, then trains specialized LoRA adapters and routes the agent to the right one at inference time. It matters because agent…
HuggingFacedaily curated papers2026-04-07Upvotes: 11
Charles Arnal, Vivien Cabannes, Taco Cohen, Julia Kempe, Remi Munos
This study argues that strict on-policy sampling is often wasteful for LLM post-training and shows that well-designed replay buffers can reuse experience without sacrificing final performance. It matters…
HuggingFacedaily curated papers2026-04-09Upvotes: 9
Yoonsang Lee, Howard Yen, Xi Ye, Danqi Chen
AggAgent treats parallel agent trajectories as an environment of their own and gives an aggregation agent tools to inspect, compare, and synthesize them rather than simply voting on final answers. It matters…
HuggingFacedaily curated papers2026-04-13Upvotes: 10
Krzysztof Ociepa, Łukasz Flis, Remigiusz Kinas, Krzysztof Wróbel, Adrian Gwoździej
This report shows how a Polish-specific tokenizer and adapted pretraining curriculum improve efficiency and usable context length for Bielik v3 compared with more universal vocabularies. It matters because…
HuggingFacedaily curated papers2026-04-12Upvotes: 4
Ivan Sedykh, Nikita Sorokin, Valentin Malykh
This paper shows masked diffusion language models can swap in smaller models for less sensitive denoising steps, cutting sampling FLOPs while preserving much of the quality. It matters because diffusion LMs…
HuggingFacedaily curated papers2026-04-11Upvotes: 6