-
meta-llama/Llama-3.2-90B-Vision-Instruct
Image-Text-to-Text • 89B • Updated • 151k • 362 -
meta-llama/Llama-3.2-11B-Vision-Instruct
Image-Text-to-Text • 11B • Updated • 60.4k • 1.65k -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.84M • • 2.63k -
meta-llama/Llama-3.2-1B-Instruct
Text Generation • 1B • Updated • 6.92M • • 1.69k
🔄 In a Training Loop
Justin
jxtngx
AI & ML interests
None yet
Organizations
Papers
-
Attention Is All You Need
Paper • 1706.03762 • Published • 141 -
LLaMA: Open and Efficient Foundation Language Models
Paper • 2302.13971 • Published • 25 -
Efficient Tool Use with Chain-of-Abstraction Reasoning
Paper • 2401.17464 • Published • 21 -
MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts
Paper • 2407.21770 • Published • 22
Models
-
meta-llama/Llama-3.2-90B-Vision-Instruct
Image-Text-to-Text • 89B • Updated • 151k • 362 -
meta-llama/Llama-3.2-11B-Vision-Instruct
Image-Text-to-Text • 11B • Updated • 60.4k • 1.65k -
meta-llama/Llama-3.2-3B-Instruct
Text Generation • 3B • Updated • 1.84M • • 2.63k -
meta-llama/Llama-3.2-1B-Instruct
Text Generation • 1B • Updated • 6.92M • • 1.69k
Papers
-
Attention Is All You Need
Paper • 1706.03762 • Published • 141 -
LLaMA: Open and Efficient Foundation Language Models
Paper • 2302.13971 • Published • 25 -
Efficient Tool Use with Chain-of-Abstraction Reasoning
Paper • 2401.17464 • Published • 21 -
MoMa: Efficient Early-Fusion Pre-training with Mixture of Modality-Aware Experts
Paper • 2407.21770 • Published • 22
models 0
None public yet
datasets 0
None public yet