Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
mengfanxu's picture

mengfanxu

fxmeng
5 11 128
Le222's profile picture zuoke's profile picture Jinnan's profile picture
·
https://fxmeng.github.io
  • fxmeng

AI & ML interests

None yet

Recent Activity

updated a dataset 14 days ago
fxmeng/UltraData-SFT-2605-no-think-8k-32k
updated a dataset 14 days ago
fxmeng/UltraData-SFT-2605-no-think-32k-200k
published a dataset 14 days ago
fxmeng/UltraData-SFT-2605-no-think-32k-200k
View all activity

Organizations

None yet

commented a paper 11 months ago

TPLA: Tensor Parallel Latent Attention for Efficient Disaggregated Prefill \& Decode Inference

Paper • 2508.15881 • Published Aug 21, 2025 • 10 •
2
commented 3 papers over 1 year ago

TransMLA: Multi-head Latent Attention Is All You Need

Paper • 2502.07864 • Published Feb 11, 2025 • 69 •
9

TransMLA: Multi-head Latent Attention Is All You Need

Paper • 2502.07864 • Published Feb 11, 2025 • 69 •
9

TransMLA: Multi-head Latent Attention Is All You Need

Paper • 2502.07864 • Published Feb 11, 2025 • 69 •
9
New activity in MMMU/MMMU over 2 years ago

Question about "Text as Input"

#4 opened over 2 years ago by
fxmeng
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs