Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Rider Jones's picture

Rider Jones

mazuj2
5 2 5
21world's profile picture
·

AI & ML interests

None yet

Recent Activity

reacted to FredyRivera-dev's post with ❤️ 1 day ago
We wrote a full technical guide on how to train a bilingual (ES/EN) LLM from scratch: TinyQwen. Covers: - Hybrid architecture based on Qwen3.5 - Pre-training with 15B tokens - Cost benchmark between H200 and B200 - Post-training with SFT + LoRA - Full code and data, open source With ~$11 of compute on an H200 we ran an initial training run, enough to validate the full architecture and pipeline. Blog post: https://aquiles-ai.vercel.app/blog/tinyqwen-from-scratch Technical feedback welcome, especially from anyone looking to replicate the pipeline with more compute.
liked a model 4 days ago
mindlab-research/Macaron-V1-Tall
new activity 23 days ago
unsloth/Qwen3.6-27B-MTP-GGUF:FAST!!!! 39tps!
View all activity

Organizations

None yet

liked a model 4 days ago

mindlab-research/Macaron-V1-Tall

Text Generation • 36B • Updated 6 days ago • 604 • 52
liked 4 models 5 months ago

turboderp/Qwen3.5-35B-A3B-exl3

Updated Mar 2 • 30 • 22

turboderp/Qwen3-VL-32B-Instruct-exl3

Updated Nov 9, 2025 • 6 • 7

turboderp/Qwen3-Next-80B-A3B-Instruct-exl3

Updated Nov 1, 2025 • 3 • 27

turboderp/Qwen3-VL-30B-A3B-Instruct-exl3

Updated Nov 9, 2025 • 18 • 5
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs