-
Reward-Augmented Decoding: Efficient Controlled Text Generation With a Unidirectional Reward Model
Paper • 2310.09520 • Published • 11 -
When can transformers reason with abstract symbols?
Paper • 2310.09753 • Published • 3 -
Improving Large Language Model Fine-tuning for Solving Math Problems
Paper • 2310.10047 • Published • 6 -
LLaVA-Interactive: An All-in-One Demo for Image Chat, Segmentation, Generation and Editing
Paper • 2311.00571 • Published • 42
Harry Xie
Hackiey
AI & ML interests
None yet
Recent Activity
liked a dataset 7 days ago
XiaomiMiMo/MiMo-V2.6-RL-oss upvoted a collection 11 days ago
MiMo-V2.6 liked a dataset 4 months ago
wdndev/webnovel-chineseOrganizations
None yet