documents Collection by weibeu Jan 4, 2024 - DocLLM: A layout-aware generative language model for multimodal document understanding Paper • 2401.00908 • Published Dec 31, 2023 • 192
DocLLM: A layout-aware generative language model for multimodal document understanding Paper • 2401.00908 • Published Dec 31, 2023 • 192
LLM Collection by BM999 Jan 4, 2024 - mistralai/Mixtral-8x7B-v0.1 47B • Updated Jul 24, 2025 • 48.3k • 1.82k
Encoders Collection by CCMat Sep 11, 2024 - EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters Paper • 2402.04252 • Published Feb 6, 2024 • 31 Scaling (Down) CLIP: A Comprehensive Analysis of Data, Architecture, and Training Strategies Paper • 2404.08197 • Published Apr 12, 2024 • 30
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters Paper • 2402.04252 • Published Feb 6, 2024 • 31
Scaling (Down) CLIP: A Comprehensive Analysis of Data, Architecture, and Training Strategies Paper • 2404.08197 • Published Apr 12, 2024 • 30
df Collection by imsarfaraz Jan 4, 2024 - Running 122 DeepFakeAI 👀 122 Generate deepfake videos from images and audio
model-llm Collection by cang Jan 4, 2024 - Skywork/Skywork-13B-base Text Generation • Updated Nov 24, 2023 • 287 • 70
music_generation Collection by mohit-ak Jan 4, 2024 - facebook/musicgen-stereo-melody-large Text-to-Audio • 3B • Updated Apr 24, 2024 • 252 • 74
llm_models Collection by feilaichao Jan 4, 2024 - huantian2415/vicuna-13b-chinese-4bit-ggml Updated Apr 28, 2023 • 11
Self-Learning AI Collection by admarcosai Jan 4, 2024 - Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models Paper • 2401.01335 • Published Jan 2, 2024 • 69
Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models Paper • 2401.01335 • Published Jan 2, 2024 • 69
Prompt Engineering Collection by therealchrisbrown Jan 4, 2024 - Contrastive Chain-of-Thought Prompting Paper • 2311.09277 • Published Nov 15, 2023 • 35 From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting Paper • 2309.04269 • Published Sep 8, 2023 • 34 Prompt Engineering a Prompt Engineer Paper • 2311.05661 • Published Nov 9, 2023 • 23
From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting Paper • 2309.04269 • Published Sep 8, 2023 • 34
7B Collection by 3D-DLW Jan 4, 2024 - berkeley-nest/Starling-LM-7B-alpha Text Generation • 7B • Updated Mar 20, 2024 • 2.25k • • 560 mistralai/Mistral-7B-Instruct-v0.2 Text Generation • 7B • Updated Jul 24, 2025 • 1.31M • • 3.19k zai-org/chatglm3-6b 6B • Updated Dec 5, 2024 • 92.1k • 1.17k
berkeley-nest/Starling-LM-7B-alpha Text Generation • 7B • Updated Mar 20, 2024 • 2.25k • • 560
mistralai/Mistral-7B-Instruct-v0.2 Text Generation • 7B • Updated Jul 24, 2025 • 1.31M • • 3.19k
documents Collection by weibeu Jan 4, 2024 - DocLLM: A layout-aware generative language model for multimodal document understanding Paper • 2401.00908 • Published Dec 31, 2023 • 192
DocLLM: A layout-aware generative language model for multimodal document understanding Paper • 2401.00908 • Published Dec 31, 2023 • 192
music_generation Collection by mohit-ak Jan 4, 2024 - facebook/musicgen-stereo-melody-large Text-to-Audio • 3B • Updated Apr 24, 2024 • 252 • 74
LLM Collection by BM999 Jan 4, 2024 - mistralai/Mixtral-8x7B-v0.1 47B • Updated Jul 24, 2025 • 48.3k • 1.82k
llm_models Collection by feilaichao Jan 4, 2024 - huantian2415/vicuna-13b-chinese-4bit-ggml Updated Apr 28, 2023 • 11
Encoders Collection by CCMat Sep 11, 2024 - EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters Paper • 2402.04252 • Published Feb 6, 2024 • 31 Scaling (Down) CLIP: A Comprehensive Analysis of Data, Architecture, and Training Strategies Paper • 2404.08197 • Published Apr 12, 2024 • 30
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters Paper • 2402.04252 • Published Feb 6, 2024 • 31
Scaling (Down) CLIP: A Comprehensive Analysis of Data, Architecture, and Training Strategies Paper • 2404.08197 • Published Apr 12, 2024 • 30
Self-Learning AI Collection by admarcosai Jan 4, 2024 - Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models Paper • 2401.01335 • Published Jan 2, 2024 • 69
Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models Paper • 2401.01335 • Published Jan 2, 2024 • 69
df Collection by imsarfaraz Jan 4, 2024 - Running 122 DeepFakeAI 👀 122 Generate deepfake videos from images and audio
Prompt Engineering Collection by therealchrisbrown Jan 4, 2024 - Contrastive Chain-of-Thought Prompting Paper • 2311.09277 • Published Nov 15, 2023 • 35 From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting Paper • 2309.04269 • Published Sep 8, 2023 • 34 Prompt Engineering a Prompt Engineer Paper • 2311.05661 • Published Nov 9, 2023 • 23
From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting Paper • 2309.04269 • Published Sep 8, 2023 • 34
model-llm Collection by cang Jan 4, 2024 - Skywork/Skywork-13B-base Text Generation • Updated Nov 24, 2023 • 287 • 70
7B Collection by 3D-DLW Jan 4, 2024 - berkeley-nest/Starling-LM-7B-alpha Text Generation • 7B • Updated Mar 20, 2024 • 2.25k • • 560 mistralai/Mistral-7B-Instruct-v0.2 Text Generation • 7B • Updated Jul 24, 2025 • 1.31M • • 3.19k zai-org/chatglm3-6b 6B • Updated Dec 5, 2024 • 92.1k • 1.17k
berkeley-nest/Starling-LM-7B-alpha Text Generation • 7B • Updated Mar 20, 2024 • 2.25k • • 560
mistralai/Mistral-7B-Instruct-v0.2 Text Generation • 7B • Updated Jul 24, 2025 • 1.31M • • 3.19k