Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Pedro Henrique Luz de Araujo
peluz
Follow
webxos's profile picture
alexandremoraisdarosa's profile picture
arthrod's profile picture
4 followers
·
5 following
https://peluz.github.io/
AI & ML interests
None yet
Recent Activity
updated
a model
about 1 hour ago
peluz/ppo-LunarLander-v2
updated
a model
about 3 hours ago
peluz/tongue-tied-baker-qwen3.5-0.8b
updated
a Space
about 7 hours ago
peluz/tongue-tied-baker
View all activity
Organizations
peluz
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
updated
a model
about 1 hour ago
peluz/ppo-LunarLander-v2
Reinforcement Learning
•
Updated
about 1 hour ago
•
1
updated
a model
about 3 hours ago
peluz/tongue-tied-baker-qwen3.5-0.8b
Updated
about 2 hours ago
updated
a Space
about 7 hours ago
Running
on
Zero
Agents
Tongue Tied Baker
🥐
SafeRL toy example: a baker that cannot mention cakes
published
a Space
about 7 hours ago
Running
on
Zero
Agents
Tongue Tied Baker
🥐
SafeRL toy example: a baker that cannot mention cakes
published
a model
about 20 hours ago
peluz/tongue-tied-baker-qwen3.5-0.8b
Updated
about 2 hours ago
updated
a model
4 days ago
peluz/poca-SoccerTwos
Reinforcement Learning
•
Updated
4 days ago
•
42
published
a model
4 days ago
peluz/poca-SoccerTwos
Reinforcement Learning
•
Updated
4 days ago
•
42
updated
a model
about 1 month ago
peluz/a2c-PandaReachDense-v3
Reinforcement Learning
•
Updated
Aug 20
published
a model
about 1 month ago
peluz/a2c-PandaReachDense-v3
Reinforcement Learning
•
Updated
Aug 20
updated
a model
about 1 month ago
peluz/ppo-Pyramids
Reinforcement Learning
•
Updated
Aug 20
•
18
published
a model
about 1 month ago
peluz/ppo-Pyramids
Reinforcement Learning
•
Updated
Aug 20
•
18
updated
a model
about 1 month ago
peluz/ppo-SnowballTarget
Reinforcement Learning
•
Updated
Aug 20
•
30
published
a model
about 1 month ago
peluz/ppo-SnowballTarget
Reinforcement Learning
•
Updated
Aug 20
•
30
updated
a model
2 months ago
peluz/Reinforce-Pixelcopter-PLE-v0
Reinforcement Learning
•
Updated
Jul 20
published
a model
2 months ago
peluz/Reinforce-Pixelcopter-PLE-v0
Reinforcement Learning
•
Updated
Jul 20
updated
a model
2 months ago
peluz/Reinforce-CartPole-v1
Reinforcement Learning
•
Updated
Jul 20
published
a model
2 months ago
peluz/Reinforce-CartPole-v1
Reinforcement Learning
•
Updated
Jul 20
updated
a model
2 months ago
peluz/dqn-SpaceInvadersNoFrameskip-v4
Reinforcement Learning
•
Updated
Jul 12
•
1
published
a model
2 months ago
peluz/dqn-SpaceInvadersNoFrameskip-v4
Reinforcement Learning
•
Updated
Jul 12
•
1
updated
a Space
3 months ago
Sleeping
Agents
Qwen3 0.6b Cat Lingo Grpo
👀
🐾 Qwen3-Cat — GRPO Cat-Lingo Demo
Load more