Blind Men and the Elephant: Probing the Epistemic Myopia of LLMs under Long-Tail Divergent Knowledge
🤝 Open to Collab
Junrulu
AI & ML interests
None yet
Recent Activity
upvoted a paper about 20 hours ago
How Far Are We from Removing the Visual Encoder? Scaling Laws for Encoder-Free Multimodal Pretraining upvoted a collection 23 days ago
RoleMRC updated a collection 29 days ago
Youtu-LLMOrganizations
SSA
Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space
RoleMRC
A Fine-Grained Composite Benchmark for Role-Playing and Instruction-Following
MemoChat
Tuning LLMs to Use Memos for Consistent Long-Range Open-Domain Conversation
ContextPilot
Teaching Agents for Proactive Context Management via Fine-grained RL
-
tencent/ContextPilot-E4B
Text Generation • 8B • Updated • 970 • 8 -
tencent/ContextPilot-8B
Text Generation • 8B • Updated • 809 • 11 -
tencent/ContextPilot-14B
Text Generation • 15B • Updated • 1.06k • 23 -
ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL
Paper • 2608.28476 • Published • 27
Youtu-LLM
Unlocking the Native Agentic Potential for Lightweight Large Language Models
-
tencent/Youtu-LLM-2B
Text Generation • 2B • Updated • 10.9k • 231 -
tencent/Youtu-LLM-2B-Base
Text Generation • 2B • Updated • 2.07k • 43 -
tencent/Youtu-LLM-2B-GGUF
Text Generation • 2B • Updated • 498 • 30 -
Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language Models
Paper • 2512.24618 • Published • 156
SamPO
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
-
jiazhengli/Pythia-2.8B-HH-RLHF-Iterative-SamPO
Text Generation • 3B • Updated • 20 -
jiazhengli/Pythia-2.8B-TLDR-Iterative-SamPO
Text Generation • 3B • Updated • 24 -
Junrulu/Llama-3-8B-Instruct-Iterative-SamPO
Text Generation • 8B • Updated • 14 • 1 -
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
Paper • 2406.10957 • Published • 2
ElephantBench
Blind Men and the Elephant: Probing the Epistemic Myopia of LLMs under Long-Tail Divergent Knowledge
ContextPilot
Teaching Agents for Proactive Context Management via Fine-grained RL
-
tencent/ContextPilot-E4B
Text Generation • 8B • Updated • 970 • 8 -
tencent/ContextPilot-8B
Text Generation • 8B • Updated • 809 • 11 -
tencent/ContextPilot-14B
Text Generation • 15B • Updated • 1.06k • 23 -
ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL
Paper • 2608.28476 • Published • 27
SSA
Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space
Youtu-LLM
Unlocking the Native Agentic Potential for Lightweight Large Language Models
-
tencent/Youtu-LLM-2B
Text Generation • 2B • Updated • 10.9k • 231 -
tencent/Youtu-LLM-2B-Base
Text Generation • 2B • Updated • 2.07k • 43 -
tencent/Youtu-LLM-2B-GGUF
Text Generation • 2B • Updated • 498 • 30 -
Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language Models
Paper • 2512.24618 • Published • 156
RoleMRC
A Fine-Grained Composite Benchmark for Role-Playing and Instruction-Following
SamPO
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
-
jiazhengli/Pythia-2.8B-HH-RLHF-Iterative-SamPO
Text Generation • 3B • Updated • 20 -
jiazhengli/Pythia-2.8B-TLDR-Iterative-SamPO
Text Generation • 3B • Updated • 24 -
Junrulu/Llama-3-8B-Instruct-Iterative-SamPO
Text Generation • 8B • Updated • 14 • 1 -
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
Paper • 2406.10957 • Published • 2
MemoChat
Tuning LLMs to Use Memos for Consistent Long-Range Open-Domain Conversation