Long-Horizon Audio-Visual Generation for Persistent Stories and Interactive Worlds Paper • 2608.23383 • Published 8 days ago • 19
view post Post 110 Omni-Rewriter Replay: feed it a local clip, get a formatted H3 prompt, then generate if you want.Gallery: Wayne-King/omni-rewriter-replayCode: https://github.com/WayneJin0918/Omni-Rewriter See translation 👍 1 1 + Reply
Symbol-LLM: Towards Foundational Symbol-centric Interface For Large Language Models Paper • 2311.09278 • Published Nov 15, 2023 • 9
On the Efficacy of Eviction Policy for Key-Value Constrained Generative Language Model Inference Paper • 2402.06262 • Published Feb 9, 2024
A Controlled Study on Long Context Extension and Generalization in LLMs Paper • 2409.12181 • Published Sep 18, 2024 • 45
Context Compression for Auto-regressive Transformers with Sentinel Tokens Paper • 2310.08152 • Published Oct 12, 2023 • 1
Audio Turing Test: Benchmarking the Human-likeness of Large Language Model-based Text-to-Speech Systems in Chinese Paper • 2505.11200 • Published May 16, 2025
EvalTalker: Learning to Evaluate Real-Portrait-Driven Multi-Subject Talking Humans Paper • 2512.01340 • Published Dec 1, 2025
EMO: Earth Mover Distance Optimization for Auto-Regressive Language Modeling Paper • 2310.04691 • Published Oct 7, 2023 • 3
WBench: A Comprehensive Multi-turn Benchmark for Interactive Video World Model Evaluation Paper • 2605.25874 • Published May 25 • 107
ClawMark: A Living-World Benchmark for Multi-Turn, Multi-Day, Multimodal Coworker Agents Paper • 2604.23781 • Published Apr 26 • 34
Next-Embedding Prediction Makes Strong Vision Learners Paper • 2512.16922 • Published Dec 18, 2025 • 91
Next-Embedding Prediction Makes Strong Vision Learners Paper • 2512.16922 • Published Dec 18, 2025 • 91
Does Understanding Inform Generation in Unified Multimodal Models? From Analysis to Path Forward Paper • 2511.20561 • Published Nov 25, 2025 • 33
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy Paper • 2510.13778 • Published Oct 15, 2025 • 17
SemiReward: A General Reward Model for Semi-supervised Learning Paper • 2310.03013 • Published Oct 4, 2023 • 2