AI & ML interests

We research Diffusions, LLMs and other ML.

arudradeyย 
posted an update 7 days ago
eienmojikiย 
posted an update 3 months ago
adamm-hfย 
posted an update 11 months ago
adamm-hfย 
posted an update 11 months ago
view post
Post
909
The new King ๐Ÿ‘‘has arrived!

Moonshot AI now the top model on Hugging Face ๐Ÿ”ฅ
moonshotai/Kimi-K2-Thinking
adamm-hfย 
posted an update 11 months ago
view post
Post
2942
๐Ÿ’ธ๐Ÿค‘You donโ€™t need 100 GPUs to train something amazing!

Our Smol Training Playbook teaches you a better path to world-class LLMs, for free!

Check out the #1 trending space on ๐Ÿค— :
HuggingFaceTB/smol-training-playbook
adamm-hfย 
posted an update 12 months ago
view post
Post
2404
Cool stuff these past weeks on huggingface! ๐Ÿค— ๐Ÿš€ !
โ€ข ๐Ÿ“ˆTrackio, local-first W&B alternative
https://github.com/gradio-app/trackio/issues
โ€ข ๐ŸŒEmbeddingGemma, 300M-param, multilingual embeddings, on-device
https://huggingface.co/blog/embeddinggemma
โ€ข ๐Ÿ’ปOpen LLMs in VS Code (Inference Providers)
https://x.com/reach_vb/status/1966185427582497171
โ€ข ๐Ÿค–Smol2Operator GUI agents
https://huggingface.co/blog/smol2operator
โ€ข ๐Ÿ–ผ๏ธGradio visible watermarking
https://huggingface.co/blog/watermarking-with-gradio
ehristoforuย 
posted an update about 1 year ago
view post
Post
2790
๐Ÿš€Hello from the Project Fluently team!

โœจ We are happy to share with you our new universal LLM models based on Qwen3 1.7B and 4B โ€” powerful, multilingual and ready to solve a wide range of problems!

๐Ÿ› ๏ธ We have conducted additional training and carefully merged them to achieve even better results and maximize the potential of the models.

๐Ÿ†“ And most importantly โ€” the models are completely open and free under the Apache-2.0 license!

๐Ÿ”— Links to repositories:
- FluentlyQwen3-4B: fluently/FluentlyQwen3-4B
- FluentlyQwen3-1.7B: fluently/FluentlyQwen3-1.7B

๐Ÿ˜ We will be very glad to hear your feedback and impressions! Your opinion is very important to us!
zamalย 
posted an update about 1 year ago
view post
Post
4380
Hey all
Finally it's happening. DeepGit lite is back now, running on cpu only devices. Just smartly search across Github and spin up conversational agents in the background and have grounded conversation with repositories
Try it out now!!!! zamal/DeepGit
  • 1 reply
ยท
zamalย 
posted an update over 1 year ago
view post
Post
1673
Say hallo to GermaNER ๐Ÿ’ชโ€“ a lightweight, high-accuracy NER model for German texts, powered by XLM-RoBERTa + LoRA adapters!
โšก Fast, efficient, and open-source โ€“ perfect for tagging names, places & orgs in real-world German data.
Try it now on Hugging Face ๐Ÿ‘‰ fau/GermaNER
zamalย 
posted an update over 1 year ago
view post
Post
4454
๐Ÿš€ Videoxity is live on Hugging Face! ๐ŸŽž๏ธ
A powerful, modular toolkit for intelligent video manipulation and scene editing.

With Videoxity, you can:

๐Ÿ–ผ๏ธ Auto-caption keyframes with BLIP

๐Ÿง  Filter scenes using natural language (e.g. โ€œremove dog scenesโ€)

โœ‚๏ธ Seamlessly trim videos with FFmpeg

๐Ÿ“Š Generate frame-based summaries

Powered by Groq LLM + LangChain, OpenCV, BLIP, and SentenceTransformers, Videoxity bridges vision and language to give developers full control over video content.
๐Ÿ”ง Built for developers. Feedback welcome!


๐Ÿ‘‰ Try it out here fau/videoxity
zamalย 
posted an update over 1 year ago
view post
Post
2005
๐Ÿš€ DeepGit Lite is live! ๐Ÿ”โœจ

Hey folks!
Just launched DeepGit Lite โ€” a lighter version of DeepGit with fewer components under the hood.
It wonโ€™t perform quite like the full powerhouse, but itโ€™s great for a quick peek and first-hand feel! โš™๏ธ๐Ÿ‘€

Give it a spin and tell us what you think!
๐Ÿ‘‰ Try it here https://huggingface.co/spaces/zamal/DeepGit-lite
#opensource #DeepGit #gradio #githubresearch
  • 3 replies
ยท
zamalย 
posted an update over 1 year ago
view post
Post
2630
DeepGit: Your GitHub Gold Digger! ๐Ÿ’ฐ๐Ÿš€
Hey Hugging Face gang! Meet DeepGitโ€”my open-source sidekick that rips through GitHub to snag repos that fit you. Done with dead-end searches? Me too. Built it with LangGraph and some dope tricks:
Embeddings grab the good stuff (HF magic, baby!)

Re-ranking nails the best picks

Snoops docs, code, and buzz in one slick flow

Drops a clean list of hidden gems ๐Ÿ’Ž

Unearth that sneaky ML lib or Python gemโ€”run python app.py or langgraph dev and boom! Peek it at https://github.com/zamalali/DeepGit. Fork it, tweak it, love itโ€”Dockerโ€™s in, HF vibes are strong. Drop a ๐ŸŒŸ or a crazy ideaโ€”Iโ€™m pumped to jam with you all! ๐Ÿช‚
zamalย 
posted an update over 1 year ago
view post
Post
2059
๐Ÿš€ ftBoost is LIVE โ€“ Stop Struggling with Fine-Tuning Data!

Alright folks, if youโ€™re tired of manually crafting fine-tuning datasets, ftBoost is here to do the heavy lifting. One-click, LangChain-Groq-powered data augmentation that scales your training data in OpenAI, Gemini, Mistral, and LLaMA formatsโ€”automatically.

๐Ÿ”ฅ Whatโ€™s inside?
โœ… Smart Augmentations โ€“ Paraphrasing, back translation, synonym swapping & synthetic noise.
โœ… No more JSONL headaches โ€“ Auto-formats everything for OpenAI, Gemini, Mistral & LLaMA.
โœ… Custom tuning โ€“ Adjust similarity, diversity, and fluency in real-time.
โœ… Upload, generate, download โ€“ Thatโ€™s it.

โšก If youโ€™re fine-tuning LLMs, this will save you hours.

๐Ÿš€ Try it now: ๐Ÿ‘‰ zamal/Finetune-Boost

๐ŸŒŸ Give us a star on GitHub!

Let me know what you think & how it boosts your workflow! ๐Ÿ”ฅ
ehristoforuย 
posted an update over 1 year ago
view post
Post
4570
Introducing our first standalone model โ€“ FluentlyLM Prinum

Introducing the first standalone model from Project Fluently LM! We worked on it for several months, used different approaches and eventually found the optimal one.

General characteristics:
- Model type: Causal language models (QwenForCausalLM, LM Transformer)
- Number of parameters: 32.5B
- Number of parameters (not embedded): 31.0B
- Number of layers: 64
- Context: 131,072 tokens
- Language(s) (NLP): English, French, Spanish, Russian, Chinese, Japanese, Persian (officially supported)
- License: MIT

Creation strategy:
The basis of the strategy is shown in Pic. 2.
We used Axolotl & Unsloth for SFT-finetuning with PEFT LoRA (rank=64, alpha=64) and Mergekit for SLERP and TIES mergers.

Evolution:
๐Ÿ† 12th place in the Open LLM Leaderboard ( open-llm-leaderboard/open_llm_leaderboard) (21.02.2025)

Detailed results and comparisons are presented in Pic. 3.

Links:
- Model: https://huggingface.co/fluently-lm/FluentlyLM-Prinum
- GGUF version: mradermacher/FluentlyLM-Prinum-GGUF
- Demo on ZeroGPU: ehristoforu/FluentlyLM-Prinum-demo
  • 7 replies
ยท
zamalย 
posted an update over 1 year ago
view post
Post
654
๐Ÿš€ Try Out RAG Demo! ๐Ÿš€

A Hugging Face Space where you can compare DeepSeek-R1 vs Llama-3 using Stuff RAG (Retrieval-Augmented Generation)!

๐Ÿ” Upload a PDF, ask questions, and see how both models perform in real-time!

Try out now:
zamal/Deepseek-R1-vs-LLama3
  • 1 reply
ยท
ameerazam08ย 
posted an update over 1 year ago
zamalย 
posted an update over 1 year ago
view post
Post
1533
zamal/Multimodal-Chat-PDF

๐Ÿš€ Introducing Chat PDF Multimodal ๐Ÿ’ฌ

Interact with your PDF documents like never before! ๐Ÿคฏ
Extract text & images, then ask context-aware questions based on both. Powered by RAG techniques & multimodal LLMs. Perfect for studying, research & more! ๐Ÿ“๐Ÿ‘€
Try it out now!!!! โœ๏ธ

#LlavaNext #MultimodalAI #Transformers
ehristoforuย 
posted an update almost 2 years ago
view post
Post
4693
โœ’๏ธ Ultraset - all-in-one dataset for SFT training in Alpaca format.
fluently-sets/ultraset

โ“ Ultraset is a comprehensive dataset for training Large Language Models (LLMs) using the SFT (instruction-based Fine-Tuning) method. This dataset consists of over 785 thousand entries in eight languages, including English, Russian, French, Italian, Spanish, German, Chinese, and Korean.

๐Ÿคฏ Ultraset solves the problem faced by users when selecting an appropriate dataset for LLM training. It combines various types of data required to enhance the model's skills in areas such as text writing and editing, mathematics, coding, biology, medicine, finance, and multilingualism.

๐Ÿค— For effective use of the dataset, it is recommended to utilize only the "instruction," "input," and "output" columns and train the model for 1-3 epochs. The dataset does not include DPO or Instruct data, making it suitable for training various types of LLM models.

โ‡๏ธ Ultraset is an excellent tool to improve your language model's skills in diverse knowledge areas.
adamm-hfย 
posted an update almost 2 years ago
zamalย 
posted an update almost 2 years ago
view post
Post
1870
๐Ÿš€ Announcement for the Lovely community! ๐Ÿš€

Just launched the zamal/DeepSeek-VL-1.3B-Chat on Hugging Face, and it's ready for YOU to explore! ๐Ÿ’ฌ๐Ÿ–ผ๏ธ

This full-fledged model is perfect for advanced image and text interactions, with zero GPU required. The Deepseek VL-1.3B Chat typically needs around 8 GB of VRAM and storage of almost 4 GB, but now you can experience it hassle-free right on our space!

Want something lighter? Weโ€™ve also uploaded a 4 bit quantized version (just around 1GB!), available on my profile. Perfect for those with limited hardware. ๐ŸŒ๐Ÿ”

Come try it now and see what this model can do! ๐Ÿš€โœจ