Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
In a Training Loop
Tarun Jain
lucifertrj
11
7
16
Follow
Charliee37's profile picture
Zept's profile picture
vinhphu3000's profile picture
40 followers
·
12 following
https://youtube.com/@aiwithtarun
TRJ_0751
lucifertrj
AI & ML interests
Deep Learning, FPGA, and ML
Recent Activity
liked
a model
about 17 hours ago
google/embeddinggemma-2
replied
to
their
post
16 days ago
You can now automate EDD (eval-driven development) with Coding Harness Agents > build a baseline LLM-based application > score every change with judge evals > keep what improves, reject what regresses I made a tutorial on what EDD is, how it works, and how to use eval scores across experiments to improve an LLM app. It builds on Jeffrey's (Confident AI) article on EDD and Eugene Yan's write-up on product evals. > Setup: a baseline RAG app using Qdrant and Gemini that every experiment starts from > Step 1: a binary-labelled dataset with critiques > Step 2: aligning the LLM-as-a-judge evaluator with Opik evals > Step 3: a harness loop that runs each experiment and scores it against the baseline. Tracing and experiment comparison then show what improved, what regressed, and what to tweak next. Source code is open source. Full guide (source code linked in the description): https://www.youtube.com/watch?v=e6akw_fKWPk
liked
a model
17 days ago
kaividlabs/LFM2.5-2.6B-litertlm
View all activity
Organizations
lucifertrj
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a model
about 17 hours ago
google/embeddinggemma-2
Feature Extraction
•
0.7B
•
Updated
3 days ago
•
29.2k
•
1.28k
liked
a model
17 days ago
kaividlabs/LFM2.5-2.6B-litertlm
Updated
17 days ago
•
1
liked
a model
2 months ago
kaividlabs/Qwen3.5-4B-litertlm
Updated
Jul 24
•
1
liked
a dataset
over 1 year ago
datalab-to/marker_benchmark
Viewer
•
Updated
Jan 29, 2025
•
2.14k
•
111
•
5
liked
a model
over 1 year ago
aiplanet/LuxLlama
Text Generation
•
8B
•
Updated
May 15, 2025
•
26
•
5
liked
a dataset
about 2 years ago
RedHenLabs/news-euro-2016
Viewer
•
Updated
Aug 20, 2024
•
19.9k
•
39
•
1
liked
a model
about 2 years ago
RedHenLabs/news-reporter-euro-3b
Text Generation
•
4B
•
Updated
Oct 24, 2024
•
22
•
1
liked
a dataset
about 2 years ago
RedHenLabs/qa-news-2016
Viewer
•
Updated
Aug 1, 2024
•
43.7k
•
52
•
3
liked
a model
about 2 years ago
RedHenLabs/news-reporter-3b
Text Generation
•
4B
•
Updated
Oct 24, 2024
•
33
•
4
liked
a dataset
about 2 years ago
aiplanet/buddhi-dataset
Viewer
•
Updated
Jul 31, 2024
•
99.2k
•
231
•
4
liked
a model
about 2 years ago
aiplanet/buddhi-indic
Text Generation
•
9B
•
Updated
Sep 5, 2024
•
33
•
7
liked
a model
over 2 years ago
aiplanet/buddhi-128k-chat-7b
Text Generation
•
7B
•
Updated
Aug 15, 2024
•
126
•
18
liked
a Space
over 2 years ago
Build error
Agents
2
Aiplanet Panda Coder 13B
🏆
2
liked
2 models
about 3 years ago
aiplanet/panda-coder-13B
Text Generation
•
13B
•
Updated
Apr 10, 2024
•
42
•
•
14
aiplanet/effi-13b
Text Generation
•
Updated
Aug 20, 2023
•
57
•
10
liked
a model
over 3 years ago
keras-io/neural-decision-forest
Updated
Sep 10, 2022
•
15
•
2