Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Pengjun Fang's picture

Pengjun Fang

Frank2416
1
P(doom) <0.1%

AI & ML interests

Computer Vision, Multimodal Generation

Recent Activity

authored a paper about 19 hours ago
MMTrail: A Multimodal Trailer Video Dataset with Language and Music Descriptions
authored a paper about 19 hours ago
AC-Foley: Reference-Audio-Guided Video-to-Audio Synthesis with Acoustic Transfer
authored a paper about 19 hours ago
HelixWorld: A Real-time Interactive Audio-Visual World Model
View all activity

Organizations

None yet

Papers 4

arxiv:2610.08760
arxiv:2609.38123
arxiv:2603.15597
arxiv:2407.20962

models 2

Frank2416/svc_hjy

Updated Feb 16

Frank2416/svc_models

Updated Feb 9

datasets 1

Frank2416/Sphere360-YT-Ambigen

Viewer • Updated Aug 18 • 61.3k • 3.11k
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs