SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation Paper • 2608.04419 • Published 5 days ago • 11
DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimization Paper • 2605.31455 • Published May 29 • 6
MASPOB: Bandit-Based Prompt Optimization for Multi-Agent Systems with Graph Neural Networks Paper • 2603.02630 • Published Mar 3
Words & Weights: Streamlining Multi-Turn Interactions via Co-Adaptation Paper • 2603.01375 • Published Mar 2 • 1
T-POP: Test-Time Personalization with Online Preference Feedback Paper • 2509.24696 • Published Sep 29, 2025 • 1
Adaptive Batch-Wise Sample Scheduling for Direct Preference Optimization Paper • 2506.17252 • Published Jun 8, 2025 • 2
Federated Zeroth-Order Optimization using Trajectory-Informed Surrogate Gradients Paper • 2308.04077 • Published Aug 8, 2023
Prompt Optimization with EASE? Efficient Ordering-aware Automated Selection of Exemplars Paper • 2405.16122 • Published May 25, 2024
Use Your INSTINCT: INSTruction optimization for LLMs usIng Neural bandits Coupled with Transformers Paper • 2310.02905 • Published Oct 2, 2023
Robustifying and Boosting Training-Free Neural Architecture Search Paper • 2403.07591 • Published Mar 12, 2024
Training-Free Neural Active Learning with Initialization-Robustness Guarantees Paper • 2306.04454 • Published Jun 7, 2023
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation Paper • 2608.04419 • Published 5 days ago • 11