Studying Image Tokenizers as Visual Languages in Unified Multimodal Models Paper • 2609.09143 • Published 5 days ago • 18
Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions Paper • 2608.29109 • Published 15 days ago • 17
Procedural Graphs: Self-Evolving Execution Structures for LLM Agents Paper • 2609.09153 • Published 5 days ago • 37
FlashRender: Few-Step Generative Rendering via Camera-Controlled Video MeanFlow Paper • 2609.03563 • Published 10 days ago • 19
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published 17 days ago • 155
SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation Paper • 2608.17426 • Published 26 days ago • 159
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published 30 days ago • 282
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 263
Consistency-Driven Co-Evolution for Self-Supervised Cross-Representation Learning Paper • 2608.04926 • Published Aug 5 • 8
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks Paper • 2608.02023 • Published Aug 3 • 158