Zixuan Jiang
Andrew0425
AI & ML interests
Vision Language Model, Multimodel Language Model, Remote sensing
Recent Activity
upvoted a paper 3 days ago
OmniEcho: Spatial Audio Understanding for Embodied Agents authored a paper 18 days ago
AuK Technical Report: An Open-Source Foundational Model for Speech Generation and Editing authored a paper 18 days ago
AgenticASR: Refining Speech Recognition in Real-World Scenarios via an Agentic Approach