FLUX 3 Action Collection FLUX 3 Action base and shared encoders, SO-101 policy, and DROID policy with optimization variants. • 3 items • Updated 3 days ago • 32
Not All Prompts Are Equal: Exploration-Guided Prompt Scaffolding for Multimodal Reinforcement Post-Training Paper • 2609.15051 • Published 13 days ago • 14
PlaidQ Collection Continuous diffusion LM for code: the 0.7B teacher plus its few-step and one-step distilled students. • 3 items • Updated 20 days ago • 1
ForgeWM: Progressive Causal Training for Few-Step Action-Conditioned Video World Models Paper • 2608.14022 • Published Aug 14 • 24
Moebius: 0.2B Lightweight Image Inpainting Framework with 10B-Level Performance Paper • 2606.19195 • Published Jun 17 • 80
Foundation Text-Generation Models Below 360M Parameters Collection Great candidates for fine-tuning targeting Wllama and Transformers.js for mobile devices, ordered by number of parameters. • 101 items • Updated 12 days ago • 47
Mistral Small 4 Collection A state-of-the-art model, open-weight, with a granular Mixture-of-Experts architecture that fuses instruct, reasoning and agentic skills. • 3 items • Updated Mar 16 • 80
FastRTC Custom UIs Collection A collection of FastRTC demos that showcase how to built a Custom UI for your server • 4 items • Updated Apr 7, 2025 • 2
MOE/Mixture of Experts Models (see also "source" cll) Collection Mixture of Expert Models by me. This leverages the power of multiple models at the same time during generation for next level performance. • 62 items • Updated Jul 9 • 20
MobileLLM Collection Optimizing Sub-billion Parameter Language Models for On-Device Use Cases (ICML 2024) https://arxiv.org/abs/2402.14905 • 49 items • Updated Mar 2 • 139
LLM in a flash: Efficient Large Language Model Inference with Limited Memory Paper • 2312.11514 • Published Dec 12, 2023 • 265