AI & ML interests

🤖🤗multi media inputs and outputs to create augmented culture and better outcomes for humans everywhere.❤️🚀

Reubencf 
posted an update 2 days ago
view post
Post
1352
🕷️ I made a Spider-Man: Miles Morales web-swinging game that runs right in your browser!

Swing, wall-crawl and web-zip through a Spider-Verse style city at sunset: ink outlines, halftone shading, comic caption boxes, "THWIP!"s, wall crashes and all.

▶️ Play it here: Reubencf/spiderman-miles-morales

Built with three.js, and made with Claude Opus 5.5 🤖: the city generator, swing physics, the Mixamo animation pipeline and the comic-book shader.

🎮 Hold left click to swing · E to web-zip · Space to jump / grab walls · Shift to sprint

Fan project, not affiliated with Marvel or Sony.
  • 5 replies
·
Reubencf 
posted an update 3 days ago
view post
Post
61
🖼️ Nano Banana Editor is now Portrait Editor, now with Qwen-Image-2.1

Some updates to the node-based photo editor 👇

✨ What's new
- New name: Nano Banana Editor → Portrait Editor
- Qwen-Image-2.1 is the HuggingFace model. It runs on its own ZeroGPU Gradio Space and handles both image editing and text-to-image. FLUX.1-Kontext and Qwen-Image-Edit have been removed.
- HuggingFace is the default mode. Sign in with HF and start editing. No API key is needed, and it uses your own ZeroGPU quota.
- Bug fix: after signing in with HuggingFace, the app used to switch back to Gemini. You now stay on HuggingFace.

🧩 Gemini and GPT modes are still available for multi-image MERGE nodes.

👉 Try it: Reubencf/Nano_Banana_Editor
🔌 Qwen-Image-2.1 API Space: Reubencf/qwen-image-2.1
Reubencf 
posted an update 12 days ago
view post
Post
55
dropping something intresting
sheet to score convert audio clips to instrumental
Reubencf/Score-Studio
  • 1 reply
·
Nymbo 
posted an update about 2 months ago
view post
Post
2310
Anthropic gave me six months of Claude Max 20x through the Claude for Open Source program, granted based on my Hugging Face work. Thank you
Anthropic
for supporting open source.

So far I've been pointing it at Markdown Minimap, an Obsidian plugin that adds a scrollable IDE-style minimap to your notes. This week I've been clearing a backlog of user-reported issues on it, with Claude often handling them end to end.

https://github.com/Nymbo/Markdown-Minimap — issues and PRs welcome.
Nymbo 
posted an update 2 months ago
view post
Post
6023
Introducing Inflect-v2, two exceptionally small, open-weight English TTS models at just 3.9M and 9.3M parameters. Both generate speech multiple times faster than real-time on CPU. Despite their size, Inflect-v2 delivers quality that is competitive with much larger lightweight TTS systems, including KittenTTS, Piper, and Supertonic-3.

CPU, CUDA, PyTorch, and ONNX are supported. Apache 2.0.

See it for yourselves:
owensong/Inflect-Micro-v2
owensong/Inflect-Nano-v2

Try the Demos:
Nymbo/Inflect-TTS (unlimited CPU usage)
owensong/Inflect-v2 (ultra-fast ZeroGPU usage)
  • 6 replies
·
Reubencf 
posted an update 2 months ago
view post
Post
229
Open Video Craft v1.0.2 is live 🎬

I built an open-source, local-first screen recorder and timeline editor for macOS and Windows. Record your screen, camera and audio; trim clips, add subtitles and export without an account.

GitHub + downloads: https://github.com/Reubencfernandes/Open-Video-Craft/releases/tag/v1.0.2

I’d love feedback from the Hugging Face community—especially ideas for useful captioning and AI-assisted editing workflows.
  • 1 reply
·
Reubencf 
posted an update 3 months ago
Reubencf 
posted an update 3 months ago
Reubencf 
posted an update 3 months ago
Reubencf 
posted an update 3 months ago
view post
Post
3780
Shadows of Tomorrow is finally live on Hugging Face Spaces with Gradio.

It’s a browser-playable RPG built with Godot, set in a post-nuclear future where players explore Magnus Province, collect medicinal plants, craft medicine, and help cure NPCs.

Play it here: Reubencf/Shadows_of_Tomorrow
  • 11 replies
·
Reubencf 
posted an update 3 months ago
view post
Post
137
Tiny Maple is now live on iOS and Android.

The app introduces the @Coherelabs Tiny Aya series of multilingual AI models to mobile devices. This release is significant as it enhances access to multilingual AI from anywhere, particularly for users who prefer offline capabilities.

I invite you to try it and share your feedback.

App Store: https://apps.apple.com/bn/app/tiny-maple/id6774123088

Google Play: https://play.google.com/store/apps/details?id=com.reubencf.tinyaya
KingNish 
posted an update 3 months ago
view post
Post
4941
We trained an open-source Mythos like cybersecurity LLM for the Build Small Hackathon meet OpenMythos

Trained in two stages: SFT on ~1.84K filtered ArXiv cs.CR papers + real CVE data, then RLVR using paired with past vulnerabilities GitHub repos with a verifier model checking outputs against ground truth.

Trained on: H100s from Modal

The RLVR stage made the biggest difference responses got more precise and less prone to confusing similar vulnerability classes.

Everything is open:
🤖 Demo → build-small-hackathon/OpenMythos
🧠 Model → build-small-hackathon/OpenMythos
📦 CVE Dataset → build-small-hackathon/CVE_Vulnerailities_Detailed
📄 ArXiv Dataset → himanshu17HF/ArvixImport-Filtered-Final

Try it out and let us know where it breaks 🙏
  • 2 replies
·
Abhaykoul 
posted an update 3 months ago
view post
Post
453
Shipped v0.1.2 of vtx — a minimalist coding agent for the terminal.

Most agentic CLIs ship 10k+ token system prompts. Vtx is ~2,200. Less prompt overhead means more room for your code in the model's context window.

Vtx is a from-scratch Python implementation of the design philosophy behind pi-mono — same principles, pure Python, no transpiled runtime.

What ships out of the box:

→ Textual TUI + headless CLI (vtx -p "fix the failing test")
→ 49 LLM provider gateways, all declared in a single provider.yaml
→ 5 core tools (read / edit / write / bash / find) plus web search and fetch
→ Session tree with compaction, handoff, and resume
→ AGENTS.md / CLAUDE.md auto-discovery
→ Skills system — drop SKILL.md files in .agents/skills/ and they become slash commands
→ Two OAuth flows (GitHub Copilot device flow, OpenAI Codex PKCE)
→ Two-mode permissions: prompt (default) or auto, with a safe-command allowlist

This release adds a proper extension system. Register new LLM-callable tools, intercept tool calls, hook lifecycle events, and add slash commands from a single register(api) function in a Python file under ~/.vtx/agent/extensions/. Extensions can override built-in tools by name and chain handler logic across subscribers.

Apache 2.0. uv tool install vtx-coding-agent and you're running.

GitHub: https://github.com/OEvortex/vtx-coding-agent
PyPI: https://pypi.org/project/vtx-coding-agent

Built in the open. Feedback, extensions, and PRs welcome.
Reubencf 
posted an update 4 months ago
view post
Post
2073
Millions speak Konkani. The internet barely knows it.

Today's major LLMs struggle with regional languages. They can't read, write or even recognize Konkani. So I built one that can.

Here is a working demo of the Konkani LLM I've been training. 👇

https://youtu.be/8K04ylbXh6k
Reubencf 
posted an update 4 months ago
view post
Post
4707
I have improved my Portfolio please do check it out
Reubencf/Portfolio
  • 6 replies
·
Tonic 
posted an update 5 months ago
view post
Post
3380
🙋🏻‍♂️ Hey there folks ,

Turns out : if we predict 🌏 earth we can save a lot of time looking for interesting things and less time looking at things that we expect to see.

Sentinel-2 imagery 🛰️basically takes a long time to download towards earth. so our "near real time" systems are quite far from that in practical terms.

meanwhile , if we "predict" what we will see , based on what we do see , we can send down much less data in a timely way , and prioritize 📡earth-bound response .

I'm talking about illegal fishing , logging , mining or building in nature reserves , the more of that we predict early the more we're able to stop it on time.

At least that's the concept !

check out the blog : https://huggingface.co/blog/Tonic/save-patagonia-by-predicting-earth


- Collection: https://huggingface.co/collections/NuTonic/earth-observation-with-temporal-and-general-understanding
- Code: https://github.com/Josephrp/Nutonic
- Dataset: NuTonic/sat-vl-sft-training-ready-v1
- Model: NuTonic/lspace
- Training: NuTonic/lspace-trackio
- Evals: NuTonic/Patagonia_Eval
  • 2 replies
·
Tonic 
posted an update 5 months ago
view post
Post
4493
🙋🏻‍♂️ Hey there folks,

since everyone liked my previous announcement post ( https://huggingface.co/posts/Tonic/338509028435394 ) so much , i'm back with more high quality proceedural datasets in the Geospacial domain for SFT training !

Check this one out :
NuTonic/sat-bbox-metadata-sft-v1

the goal is to be able to train vision models on multiple images for remote sensing analysis with one shot .

hope you like it ! 🚀
  • 2 replies
·
Tonic 
posted an update 5 months ago
view post
Post
3751
🙋🏻‍♂️ Hey there folks ,

I'm sharing huggingface's largest dataset of annotated statelite images today.

check it out here : NuTonic/sat-image-boundingbox-sft-full

I hope you like it , the idea is to be able to use this with small vision models 🚀
Parveshiiii 
posted an update 6 months ago
view post
Post
658
🚀 Sonic: A lightweight Python audio processing library with tempo matching, BPM detection, time-stretching, resampling & track blending — now with GPU (CUDA) acceleration for 10x speed!

Perfect for quick remixes, batch edits or syncing tracks.

👉 https://github.com/Parveshiiii/Sonic

#Python #AudioProcessing #OpenSource #PyTorch