AI & ML interests

LLM

Recent Activity

sinwangĀ  submitted a paper about 12 hours ago
Multi-hop Reasoning via Early Knowledge Alignment
AuraithmĀ  updated a model about 13 hours ago
OpenMOSS-Team/DiRL-8B-Instruct
lkdhyĀ  updated a dataset 5 days ago
OpenMOSS-Team/VideoThinkBench
View all activity

OpenMOSS-Team 's collections 10

MHA2MLA
The MHA2MLA model published in the paper "Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-Based LLMs"
MHA2MLA-refactor
The MHA2MLA model published in the paper "Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-Based LLMs"
MHA2MLA-refactor
The MHA2MLA model published in the paper "Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-Based LLMs"
MHA2MLA
The MHA2MLA model published in the paper "Towards Economical Inference: Enabling DeepSeek's Multi-Head Latent Attention in Any Transformer-Based LLMs"