Collections
Discover the best community collections!
Collections trending this week
-
sepidmnorozy/Vietnamese_sentiment
Viewer • Updated • 3.4k • 92 • 7 -
Distillation Contrastive Decoding: Improving LLMs Reasoning with Contrastive Decoding and Distillation
Paper • 2402.14874 • Published • 4 -
Viet-Mistral/Vistral-7B-Chat
Text Generation • 7B • Updated • 1.01k • 151 -
linhtran92/viet_bud500
Viewer • Updated • 649k • 1.03k • 72
-
Large Language Model Alignment: A Survey
Paper • 2309.15025 • Published • 2 -
Aligning Large Language Models with Human: A Survey
Paper • 2307.12966 • Published • 1 -
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Paper • 2305.18290 • Published • 71 -
SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF
Paper • 2310.05344 • Published • 1
-
sepidmnorozy/Vietnamese_sentiment
Viewer • Updated • 3.4k • 92 • 7 -
Distillation Contrastive Decoding: Improving LLMs Reasoning with Contrastive Decoding and Distillation
Paper • 2402.14874 • Published • 4 -
Viet-Mistral/Vistral-7B-Chat
Text Generation • 7B • Updated • 1.01k • 151 -
linhtran92/viet_bud500
Viewer • Updated • 649k • 1.03k • 72
-
Large Language Model Alignment: A Survey
Paper • 2309.15025 • Published • 2 -
Aligning Large Language Models with Human: A Survey
Paper • 2307.12966 • Published • 1 -
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Paper • 2305.18290 • Published • 71 -
SteerLM: Attribute Conditioned SFT as an (User-Steerable) Alternative to RLHF
Paper • 2310.05344 • Published • 1