TaoMate Digital Human
Real-time audio-video digital human generation from text
None defined yet.
Real-time audio-video digital human generation from text
Cross-scale pathology VLM with multi-magnification reasoning
Restore damaged Hanja in Joseon Dynasty historical documents
Fix self-collisions in SMPL-H poses via neural fields
Separate audio into vocals and instruments with BS-Roformer
Add fine detail to images with a Krea 2 edit LoRA
985K-param prompt router - domain, complexity, flags, tier
Persian OCR with Bina 0.1 Koochik VLM
Reference-driven multi-speaker audio scene generation
Zero-shot PII detection with GLiNER streaming-span model
Live multi-step Fara browser-agent loop on real websites
Unified AR-LM-based speech enhancement & separation
Multilingual document image to Markdown OCR
Animate a first frame with Cosmos3-Super I2V (64B, 4-step)
Web computer-use agent β screenshot + task to next action
Embodied AI VLM for visual grounding and planning
Text-to-image with Microsoft Mage-Flow-Base (4.1B MMDiT)
Persian OCR with Bina 0.1 vision-language model
Temporal reasoning over multi-temporal satellite imagery
3D spatial reasoning VLM with video-diffusion 3D priors
Hierarchical driving world model β predict future frames
Watch & think simultaneously β streaming video reasoning
Dense 1.3B embodied video generation β T2V, I2V, T2I
NVIDIA Cosmos3-Edge β reason over images & video