RayOrch: Programming and Executing Lineage-Controlled Multi-Grain Dataflows for Foundation-Model Data Preparation Paper ⢠2609.18703 ⢠Published 22 days ago ⢠56
FuseReg: Regularizing Layer Fusion Mitigates the Reconstruction-Generation Gap in Representation Autoencoders Paper ⢠2609.31620 ⢠Published 13 days ago ⢠160
Disaggregated Quantization: Specializing LLM Prefill and Decode Paper ⢠2609.26333 ⢠Published 16 days ago ⢠93
Morphometric Imitation: From Morphology and Contact Aware Hand Retargeting to Sim-to-Real Visuomotor Policy Paper ⢠2609.28660 ⢠Published 15 days ago ⢠16
Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs Paper ⢠2609.29845 ⢠Published 14 days ago ⢠105
Coding Agents for Generalized Task and Motion Planning Problems Paper ⢠2609.30233 ⢠Published 14 days ago ⢠28
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego ⢠Sep 3 ⢠147
yassiracharki/Yelp_Reviews_for_Binary_Senti_Analysis Viewer ⢠Updated Jul 26, 2024 ⢠598k ⢠59 ⢠1
yassiracharki/Amazon_Reviews_for_Sentiment_Analysis_fine_grained_5_classes Viewer ⢠Updated Jul 26, 2024 ⢠3.65M ⢠117 ⢠4
yassiracharki/Amazon_Reviews_Binary_for_Sentiment_Analysis Viewer ⢠Updated Jul 26, 2024 ⢠4M ⢠139
yassiracharki/Yelp_Reviews_for_Sentiment_Analysis_fine_grained_5_classes Viewer ⢠Updated Jul 26, 2024 ⢠700k ⢠73
yassiracharki/Yahoo_Answers_10_categories_for_NLP Viewer ⢠Updated Jul 26, 2024 ⢠1.46M ⢠71 ⢠3