Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
momo sheng's picture

momo sheng

ImmortalMo
1 6 3
·
  • ImmortalSdm

AI & ML interests

None yet

Recent Activity

new activity 18 days ago
opendatalab/Sci-Base:It seems that part-00001 of textbook is missing
liked a model about 1 year ago
omni-research/Tarsier2-7b-0115
upvoted a paper about 1 year ago
Show-o2: Improved Native Unified Multimodal Models
View all activity

Organizations

None yet

upvoted a paper about 1 year ago

Show-o2: Improved Native Unified Multimodal Models

Paper • 2506.15564 • Published Jun 18, 2025 • 31
upvoted 4 papers over 1 year ago

Packing Input Frame Context in Next-Frame Prediction Models for Video Generation

Paper • 2504.12626 • Published Apr 17, 2025 • 52

InternLM-XComposer2.5-OmniLive: A Comprehensive Multimodal System for Long-term Streaming Video and Audio Interactions

Paper • 2412.09596 • Published Dec 12, 2024 • 97

Interleaved Scene Graph for Interleaved Text-and-Image Generation Assessment

Paper • 2411.17188 • Published Nov 26, 2024 • 20

ShowUI: One Vision-Language-Action Model for GUI Visual Agent

Paper • 2411.17465 • Published Nov 26, 2024 • 90
upvoted a paper almost 2 years ago

NaturalBench: Evaluating Vision-Language Models on Natural Adversarial Samples

Paper • 2410.14669 • Published Oct 18, 2024 • 39
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs