Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Kashish gupta
kashishgupta
1
Follow
0 followers
·
1 following
Kashish415
AI & ML interests
None yet
Recent Activity
updated
a model
about 10 hours ago
kashishgupta/qwen2.5-1.5b-anti-sycophancy-lora
published
a model
about 10 hours ago
kashishgupta/qwen2.5-1.5b-anti-sycophancy-lora
updated
a dataset
4 days ago
kashishgupta/anti-sycophancy-dpo-cleaned
View all activity
Organizations
kashishgupta
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
updated
a model
about 10 hours ago
kashishgupta/qwen2.5-1.5b-anti-sycophancy-lora
Text Generation
•
Updated
about 6 hours ago
•
7
published
a model
about 10 hours ago
kashishgupta/qwen2.5-1.5b-anti-sycophancy-lora
Text Generation
•
Updated
about 6 hours ago
•
7
updated
a dataset
4 days ago
kashishgupta/anti-sycophancy-dpo-cleaned
Viewer
•
Updated
4 days ago
•
2.83k
•
56
published
a dataset
4 days ago
kashishgupta/anti-sycophancy-dpo-cleaned
Viewer
•
Updated
4 days ago
•
2.83k
•
56
updated
a Space
3 months ago
Sleeping
CSFAQ Project
💻
Answer FAQ questions instantly with AI
published
a Space
3 months ago
Sleeping
CSFAQ Project
💻
Answer FAQ questions instantly with AI
updated
a model
3 months ago
kashishgupta/ML-Agents-Pyramids
Reinforcement Learning
•
Updated
Jun 21
•
15
upvoted
an
article
4 months ago
view article
Article
Deep Q-Learning with Space Invaders
ThomasSimonini
•
Jun 7, 2022
•
2
published
a model
5 months ago
kashishgupta/ML-Agents-Pyramids
Reinforcement Learning
•
Updated
Jun 21
•
15
updated
a model
5 months ago
kashishgupta/ML-Agents-SnowballTarget
Reinforcement Learning
•
Updated
May 10
•
26
published
a model
5 months ago
kashishgupta/ML-Agents-SnowballTarget
Reinforcement Learning
•
Updated
May 10
•
26
updated
a model
5 months ago
kashishgupta/Reinforce-PixelCopter-policy-gradient
Reinforcement Learning
•
Updated
May 9
published
a model
5 months ago
kashishgupta/Reinforce-PixelCopter-policy-gradient
Reinforcement Learning
•
Updated
May 9
updated
a model
5 months ago
kashishgupta/Reinforce-CartPolev1-policy-gradient
Reinforcement Learning
•
Updated
May 9
published
a model
5 months ago
kashishgupta/Reinforce-CartPolev1-policy-gradient
Reinforcement Learning
•
Updated
May 9
updated
a model
5 months ago
kashishgupta/a2c-PandaReachDense-Robotic-arm
Reinforcement Learning
•
Updated
Apr 30
•
2
•
1
published
a model
5 months ago
kashishgupta/a2c-PandaReachDense-Robotic-arm
Reinforcement Learning
•
Updated
Apr 30
•
2
•
1
updated
a model
5 months ago
kashishgupta/MARL-Agents-AI-vs-AI-2x2-SoccerTwos
Reinforcement Learning
•
Updated
Apr 28
•
7
published
a model
5 months ago
kashishgupta/MARL-Agents-AI-vs-AI-2x2-SoccerTwos
Reinforcement Learning
•
Updated
Apr 28
•
7
updated
a model
5 months ago
kashishgupta/PPO-CleanRL-LunarLander-v3
Reinforcement Learning
•
Updated
Apr 28
Load more