EviRank: Structured Relevance Evidence for Multimodal Image Re-ranking Paper • 2608.20886 • Published 6 days ago • 11
CForce: Boosting Parallel Decoding for dLLMs via Consistency Forcing Paper • 2608.13925 • Published 13 days ago • 1
ForgeWM: Progressive Causal Training for Few-Step Action-Conditioned Video World Models Paper • 2608.14022 • Published 13 days ago • 24
Progressive Agent Skill Generation via Reinforcement Learning Paper • 2608.01678 • Published 24 days ago • 60
LEDGERMIND: Provenance-Constrained Multimodal Agentic Reasoning with a Structured Evidence Ledger Paper • 2607.28374 • Published 28 days ago • 15
CineMobile: On-Device Image-to-Video Diffusion for Cinematic Camera Motion Generation Paper • 2607.03803 • Published Jul 4 • 19
WTF GENIUS PAPERS Collection Papers that made me appreciate my major and my life a little more. obs=Observation, innov=Innovation. Most papers are abt improving tiny models. • 307 items • Updated 2 days ago • 73
PhysRAG: Enhancing Physics-Awareness in Video Generation via Retrieval-Augmented Generation Paper • 2606.26916 • Published Jun 25 • 4
ProductWebGen: Benchmarking Multimodal Product Webpage Generation Paper • 2606.01022 • Published May 31 • 5
LatentUM: Unleashing the Potential of Interleaved Cross-Modal Reasoning via a Latent-Space Unified Model Paper • 2604.02097 • Published Apr 2 • 32
CoBA-RL: Capability-Oriented Budget Allocation for Reinforcement Learning in LLMs Paper • 2602.03048 • Published Feb 3 • 32
Think-Then-Generate: Reasoning-Aware Text-to-Image Diffusion with LLM Encoders Paper • 2601.10332 • Published Jan 15 • 32
LoPA: Scaling dLLM Inference via Lookahead Parallel Decoding Paper • 2512.16229 • Published Dec 18, 2025 • 17
DEER: Draft with Diffusion, Verify with Autoregressive Models Paper • 2512.15176 • Published Dec 17, 2025 • 45