100 Coder/Programming - MOE, Reasoning, Reg, Imatrix, Fused. Collection Models (0.8B to 87B) in regular, "reasoning", "Brainstorm", MOE (1x to 8x / 128 experts), and expanded to create better and stronger code, faster. • 60 items • Updated about 3 hours ago • 50
DavidAU/Qwen3.6-27B-Fable-Fusion-711-Uncensored-Heretic-NM-DAU-MTP Image-Text-to-Text • 28B • Updated 1 day ago • 18.6k • • 64
DavidAU/Qwen3.6-40B-Grand-Intelligence-Six-711-717-BETA1-Tune Image-Text-to-Text • 40B • Updated about 2 hours ago • 5 • 49
DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF Image-Text-to-Text • 9B • Updated 6 days ago • 323k • 282
CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition Paper • 2607.25294 • Published 9 days ago • 49
BM25 Wins at Scale: A Scaling Study of Retrieval-Augmented Generation Paradigms Paper • 2607.26497 • Published 7 days ago • 49
Beacon: Knowing When and How to Perform Agentic Visual Reasoning Paper • 2607.28595 • Published 7 days ago • 52
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory Paper • 2607.27919 • Published 7 days ago • 58
DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space Paper • 2607.25675 • Published 9 days ago • 65
VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System Paper • 2607.27380 • Published 8 days ago • 70
HumanCLAW: Can Vision-Language Models Act Through a Body? Paper • 2607.27180 • Published 8 days ago • 76
DistillAlign: Coordinating Mode Covering and Mode Seeking in Autoregressive Video Distillation Paper • 2607.26811 • Published 8 days ago • 92
TurboVLA: Real-Time Vision-Language-Action Model at 32 Hz on an RTX 4090 with <1 GB VRAM Paper • 2607.27205 • Published 8 days ago • 139
PhiZero: A World Model Built Around Physical Language Paper • 2607.28624 • Published 7 days ago • 166