SSAE Training and evaluation dataset, model checkpoints in 'Step-Level Sparse Autoencoder for Reasoning Process Interpretation' Collection by TorresYang Mar 4 - Miaow-Lab/SSAE-Dataset Viewer • Updated Mar 4 • 1.28M • 51 Miaow-Lab/SSAE-Checkpoints Feature Extraction • Updated Mar 4 Step-Level Sparse Autoencoder for Reasoning Process Interpretation Paper • 2603.03031 • Published Mar 3
Step-Level Sparse Autoencoder for Reasoning Process Interpretation Paper • 2603.03031 • Published Mar 3
MuscleMimic Collection of datasets, demo datasets, checkpoints. Collection by amathislab 7 days ago 1 amathislab/musclemimic-retargeted Updated May 7 • 1.68k • 8 amathislab/musclemimic-bimanual-retargeted Updated Jul 24 • 56 amathislab/demo_dataset Updated May 7 • 66 • 7 amathislab/mm-fullbody-base Updated May 11 • 84
Struct-SQL Distilled Query-Plan CoT to an SLM Collection by craterlabs Mar 27 - craterlabs/struct-sql-data Viewer • Updated Jan 28 • 1.3k • 27 craterlabs/Struct-SQL Text Generation • 4B • Updated Jan 28 • 25 Knowledge Distillation with Structured Chain-of-Thought for Text-to-SQL Paper • 2512.17053 • Published Dec 18, 2025 heegyu/bird-sql-mini-dev Viewer • Updated Jul 26, 2024 • 500 • 343 • 1
Knowledge Distillation with Structured Chain-of-Thought for Text-to-SQL Paper • 2512.17053 • Published Dec 18, 2025
Teprocessor.batch_decode(outputs Collection by Lennie29 Dec 30, 2025 - LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation Paper • 2512.23576 • Published Dec 29, 2025 • 66
LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation Paper • 2512.23576 • Published Dec 29, 2025 • 66
TwinFlow A collection of TwinFlow-accelerated diffusion models Collection by inclusionAI 8 days ago 6 TwinFlow: Realizing One-step Generation on Large Models with Self-adversarial Flows Paper • 2512.05150 • Published Dec 3, 2025 • 76 inclusionAI/TwinFlow Text-to-Image • Updated Dec 29, 2025 • 15 • 124 inclusionAI/TwinFlow-Z-Image-Turbo Text-to-Image • Updated Dec 29, 2025 • 14 • 213
TwinFlow: Realizing One-step Generation on Large Models with Self-adversarial Flows Paper • 2512.05150 • Published Dec 3, 2025 • 76
BERT Models Text & DNA Models trained for the paper "Entropy, Disagreement, and the Limits of Foundation Models in Genomics". Collection by mrochk Mar 3 - mrochk/bert-90M-text-bpe-1 Fill-Mask • 89.2M • Updated Feb 19 • 6 mrochk/bert-90M-text-bpe-2 Fill-Mask • 89.2M • Updated Feb 19 • 6 mrochk/bert-90M-text-bpe-3 Fill-Mask • 89.2M • Updated Feb 19 • 6 mrochk/bert-90M-text-bpe-4 Fill-Mask • 89.2M • Updated Feb 19 • 6
POPE Collection by CMU-AIRe Mar 2 2 CMU-AIRe/POPE-source-128x32k Viewer • Updated Jan 31 • 2.52k • 21 CMU-AIRe/POPE-HARD-w-oracle-solution Viewer • Updated Jan 31 • 601 • 85 CMU-AIRe/POPE-more-64x32k Viewer • Updated Oct 27, 2025 • 91k • 162 CMU-AIRe/POPE-HARD-w-guide Viewer • Updated Jan 31 • 512 • 25
LingBot-VLA Vision-Language-Action Foundation Model Collection by robbyant Mar 9 15 robbyant/lingbot-vla-4b-depth 4B • Updated Apr 30 • 96 • 19 robbyant/lingbot-vla-4b 4B • Updated Apr 30 • 506 • 32 robbyant/gm100 Preview • Updated Feb 28 • 2.37k • 19 robbyant/lingbot-vla-4b-posttrain-robotwin 4B • Updated Apr 30 • 461 • 3
LASIK Eye Surgery Vs Contact Lenses: Which is Right for You? If you’ve ever fumbled with contact lenses early in the morning or felt tired of depending on glasses, you’ve probably wondered if LASIK eye surgery Collection by Dranisha Dec 27, 2025 -
Olg Collection by kyrgan Apr 10 - Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper • 2511.14993 • Published Nov 19, 2025 • 236 black-forest-labs/FLUX.2-klein-9B Image-to-Image • 9B • Updated Feb 24 • 220k • • 1.39k
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper • 2511.14993 • Published Nov 19, 2025 • 236
SSAE Training and evaluation dataset, model checkpoints in 'Step-Level Sparse Autoencoder for Reasoning Process Interpretation' Collection by TorresYang Mar 4 - Miaow-Lab/SSAE-Dataset Viewer • Updated Mar 4 • 1.28M • 51 Miaow-Lab/SSAE-Checkpoints Feature Extraction • Updated Mar 4 Step-Level Sparse Autoencoder for Reasoning Process Interpretation Paper • 2603.03031 • Published Mar 3
Step-Level Sparse Autoencoder for Reasoning Process Interpretation Paper • 2603.03031 • Published Mar 3
BERT Models Text & DNA Models trained for the paper "Entropy, Disagreement, and the Limits of Foundation Models in Genomics". Collection by mrochk Mar 3 - mrochk/bert-90M-text-bpe-1 Fill-Mask • 89.2M • Updated Feb 19 • 6 mrochk/bert-90M-text-bpe-2 Fill-Mask • 89.2M • Updated Feb 19 • 6 mrochk/bert-90M-text-bpe-3 Fill-Mask • 89.2M • Updated Feb 19 • 6 mrochk/bert-90M-text-bpe-4 Fill-Mask • 89.2M • Updated Feb 19 • 6
MuscleMimic Collection of datasets, demo datasets, checkpoints. Collection by amathislab 7 days ago 1 amathislab/musclemimic-retargeted Updated May 7 • 1.68k • 8 amathislab/musclemimic-bimanual-retargeted Updated Jul 24 • 56 amathislab/demo_dataset Updated May 7 • 66 • 7 amathislab/mm-fullbody-base Updated May 11 • 84
POPE Collection by CMU-AIRe Mar 2 2 CMU-AIRe/POPE-source-128x32k Viewer • Updated Jan 31 • 2.52k • 21 CMU-AIRe/POPE-HARD-w-oracle-solution Viewer • Updated Jan 31 • 601 • 85 CMU-AIRe/POPE-more-64x32k Viewer • Updated Oct 27, 2025 • 91k • 162 CMU-AIRe/POPE-HARD-w-guide Viewer • Updated Jan 31 • 512 • 25
Struct-SQL Distilled Query-Plan CoT to an SLM Collection by craterlabs Mar 27 - craterlabs/struct-sql-data Viewer • Updated Jan 28 • 1.3k • 27 craterlabs/Struct-SQL Text Generation • 4B • Updated Jan 28 • 25 Knowledge Distillation with Structured Chain-of-Thought for Text-to-SQL Paper • 2512.17053 • Published Dec 18, 2025 heegyu/bird-sql-mini-dev Viewer • Updated Jul 26, 2024 • 500 • 343 • 1
Knowledge Distillation with Structured Chain-of-Thought for Text-to-SQL Paper • 2512.17053 • Published Dec 18, 2025
LingBot-VLA Vision-Language-Action Foundation Model Collection by robbyant Mar 9 15 robbyant/lingbot-vla-4b-depth 4B • Updated Apr 30 • 96 • 19 robbyant/lingbot-vla-4b 4B • Updated Apr 30 • 506 • 32 robbyant/gm100 Preview • Updated Feb 28 • 2.37k • 19 robbyant/lingbot-vla-4b-posttrain-robotwin 4B • Updated Apr 30 • 461 • 3
Teprocessor.batch_decode(outputs Collection by Lennie29 Dec 30, 2025 - LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation Paper • 2512.23576 • Published Dec 29, 2025 • 66
LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation Paper • 2512.23576 • Published Dec 29, 2025 • 66
LASIK Eye Surgery Vs Contact Lenses: Which is Right for You? If you’ve ever fumbled with contact lenses early in the morning or felt tired of depending on glasses, you’ve probably wondered if LASIK eye surgery Collection by Dranisha Dec 27, 2025 -
TwinFlow A collection of TwinFlow-accelerated diffusion models Collection by inclusionAI 8 days ago 6 TwinFlow: Realizing One-step Generation on Large Models with Self-adversarial Flows Paper • 2512.05150 • Published Dec 3, 2025 • 76 inclusionAI/TwinFlow Text-to-Image • Updated Dec 29, 2025 • 15 • 124 inclusionAI/TwinFlow-Z-Image-Turbo Text-to-Image • Updated Dec 29, 2025 • 14 • 213
TwinFlow: Realizing One-step Generation on Large Models with Self-adversarial Flows Paper • 2512.05150 • Published Dec 3, 2025 • 76
Olg Collection by kyrgan Apr 10 - Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper • 2511.14993 • Published Nov 19, 2025 • 236 black-forest-labs/FLUX.2-klein-9B Image-to-Image • 9B • Updated Feb 24 • 220k • • 1.39k
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper • 2511.14993 • Published Nov 19, 2025 • 236