Blind Men and the Elephant: Probing the Epistemic Myopia of LLMs under Long-Tail Divergent Knowledge
🤝 Open to Collab
Junrulu
AI & ML interests
None yet
Recent Activity
upvoted a collection 11 days ago
RoleMRC updated a collection 17 days ago
Youtu-LLM updated a collection 17 days ago
Youtu-LLMOrganizations
SSA
Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space
RoleMRC
A Fine-Grained Composite Benchmark for Role-Playing and Instruction-Following
MemoChat
Tuning LLMs to Use Memos for Consistent Long-Range Open-Domain Conversation
ContextPilot
Teaching Agents for Proactive Context Management via Fine-grained RL
-
tencent/ContextPilot-E4B
Text Generation • 8B • Updated • 1.1k • 7 -
tencent/ContextPilot-8B
Text Generation • 8B • Updated • 904 • 10 -
tencent/ContextPilot-14B
Text Generation • 15B • Updated • 1.13k • 22 -
ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL
Paper • 2608.28476 • Published • 26
Youtu-LLM
Unlocking the Native Agentic Potential for Lightweight Large Language Models
-
tencent/Youtu-LLM-2B
Text Generation • 2B • Updated • 11k • 231 -
tencent/Youtu-LLM-2B-Base
Text Generation • 2B • Updated • 2.64k • 43 -
tencent/Youtu-LLM-2B-GGUF
Text Generation • 2B • Updated • 366 • 30 -
Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language Models
Paper • 2512.24618 • Published • 156
SamPO
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
-
jiazhengli/Pythia-2.8B-HH-RLHF-Iterative-SamPO
Text Generation • 3B • Updated • 17 -
jiazhengli/Pythia-2.8B-TLDR-Iterative-SamPO
Text Generation • 3B • Updated • 30 -
Junrulu/Llama-3-8B-Instruct-Iterative-SamPO
Text Generation • 8B • Updated • 17 • 1 -
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
Paper • 2406.10957 • Published • 2
ElephantBench
Blind Men and the Elephant: Probing the Epistemic Myopia of LLMs under Long-Tail Divergent Knowledge
ContextPilot
Teaching Agents for Proactive Context Management via Fine-grained RL
-
tencent/ContextPilot-E4B
Text Generation • 8B • Updated • 1.1k • 7 -
tencent/ContextPilot-8B
Text Generation • 8B • Updated • 904 • 10 -
tencent/ContextPilot-14B
Text Generation • 15B • Updated • 1.13k • 22 -
ContextPilot: Teaching Agents for Proactive Context Management via Fine-grained RL
Paper • 2608.28476 • Published • 26
SSA
Sparse Sparse Attention by Aligning Full and Sparse Attention Outputs in Feature Space
Youtu-LLM
Unlocking the Native Agentic Potential for Lightweight Large Language Models
-
tencent/Youtu-LLM-2B
Text Generation • 2B • Updated • 11k • 231 -
tencent/Youtu-LLM-2B-Base
Text Generation • 2B • Updated • 2.64k • 43 -
tencent/Youtu-LLM-2B-GGUF
Text Generation • 2B • Updated • 366 • 30 -
Youtu-LLM: Unlocking the Native Agentic Potential for Lightweight Large Language Models
Paper • 2512.24618 • Published • 156
RoleMRC
A Fine-Grained Composite Benchmark for Role-Playing and Instruction-Following
SamPO
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
-
jiazhengli/Pythia-2.8B-HH-RLHF-Iterative-SamPO
Text Generation • 3B • Updated • 17 -
jiazhengli/Pythia-2.8B-TLDR-Iterative-SamPO
Text Generation • 3B • Updated • 30 -
Junrulu/Llama-3-8B-Instruct-Iterative-SamPO
Text Generation • 8B • Updated • 17 • 1 -
Eliminating Biased Length Reliance of Direct Preference Optimization via Down-Sampled KL Divergence
Paper • 2406.10957 • Published • 2
MemoChat
Tuning LLMs to Use Memos for Consistent Long-Range Open-Domain Conversation