Cool-Research-Mad Hogwild! Inference: Parallel LLM Generation via Concurrent Attention Paper • 2504.06261 • Published Apr 8, 2025 • 110 QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs Paper • 2510.11696 • Published Oct 13, 2025 • 183
Hogwild! Inference: Parallel LLM Generation via Concurrent Attention Paper • 2504.06261 • Published Apr 8, 2025 • 110
QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs Paper • 2510.11696 • Published Oct 13, 2025 • 183
Cool-Research-Mad Hogwild! Inference: Parallel LLM Generation via Concurrent Attention Paper • 2504.06261 • Published Apr 8, 2025 • 110 QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs Paper • 2510.11696 • Published Oct 13, 2025 • 183
Hogwild! Inference: Parallel LLM Generation via Concurrent Attention Paper • 2504.06261 • Published Apr 8, 2025 • 110
QeRL: Beyond Efficiency -- Quantization-enhanced Reinforcement Learning for LLMs Paper • 2510.11696 • Published Oct 13, 2025 • 183
cinnybun02/gpt-oss-9.0b-specialized-math-pruned-moe-only-12-experts-Q5_K_M-GGUF Text Generation • 9B • Updated Aug 14, 2025 • 25 • 1
cinnybun02/gpt-oss-9.0b-specialized-math-pruned-moe-only-12-experts-Q8_0-GGUF Text Generation • 9B • Updated Aug 14, 2025 • 18
cinnybun02/gpt-oss-12.0b-specialized-science-pruned-moe-only-17-experts-Q8_0-GGUF Text Generation • 12B • Updated Aug 14, 2025 • 19 • 1
cinnybun02/gpt-oss-6.6b-specialized-all-pruned-moe-only-8-experts-Q8_0-GGUF Text Generation • 7B • Updated Aug 14, 2025 • 22