Tonyo
El1iasss
AI & ML interests
None yet
Recent Activity
updated a collection 15 days ago
SLM updated a collection 16 days ago
SLM liked a model 16 days ago
harshatheg/Qwen-2.5-1B-RLCDOrganizations
fix: arxiv:2508.08221 license metadata
2
#465 opened 3 months ago
by
El1iasss
source: arxiv:2508.08221 — Tricks or Traps? A Deep Dive into RL for LLM Reasoning (Lite PPO)
3
#464 opened 3 months ago
by
bfuzzy1
source: arxiv:2501.17161 — SFT Memorizes, RL Generalizes
5
#463 opened 3 months ago
by
bfuzzy1
source: arxiv:2402.05749 - Generalized Preference Optimization
2
#459 opened 3 months ago
by
El1iasss
source: arxiv:2509.08827 — A Survey of RL for Large Reasoning Models
4
#452 opened 3 months ago
by
bfuzzy1
source: arxiv:2412.10400 — Reinforcement Learning Enhanced LLMs: A Survey
4
#451 opened 3 months ago
by
bfuzzy1
source: arxiv:2311.05821 — Let's Reinforce Step by Step
9
#429 opened 3 months ago
by
bfuzzy1
source: arxiv:2204.14146 - Training Language Models with Language Feedback
3
#444 opened 3 months ago
by
El1iasss
source: arxiv:2110.03111 - Cut the CARP
2
#443 opened 3 months ago
by
El1iasss
source: arxiv:2109.10862 - Recursively Summarizing Books with Human Feedback
2
#442 opened 3 months ago
by
El1iasss
source: arxiv:2007.12626 - SummEval
3
#441 opened 3 months ago
by
El1iasss
topic: adversarial-robustness-and-jailbreaks — add GPT-4 fine-tuning attack + RLHF data-poisoning surface; developing → comprehensive
5
#428 opened 3 months ago
by
bfuzzy1
source: arxiv:2308.06385 — ZYN: Zero-Shot Reward Models with Yes-No Questions for RLAIF
4
#417 opened 3 months ago
by
bfuzzy1
source: arxiv:2309.16155 — Trickle-down Impact of Reward (In-)consistency on RLHF
4
#415 opened 3 months ago
by
bfuzzy1
source: arxiv:2405.01525 — FLAME: Factuality-Aware Alignment for LLMs
4
#439 opened 3 months ago
by
bfuzzy1
source: arxiv:2502.19328 — Agentic Reward Modeling (RewardAgent)
2
#438 opened 3 months ago
by
bfuzzy1
source: arxiv:2311.14743 — reward models under distribution shift
4
#437 opened 3 months ago
by
bfuzzy1
source: arxiv:2311.09641 — RLHFPoison (RankPoison reward poisoning)
4
#436 opened 3 months ago
by
bfuzzy1