-
Reward-Augmented Decoding: Efficient Controlled Text Generation With a Unidirectional Reward Model
Paper • 2310.09520 • Published • 11 -
When can transformers reason with abstract symbols?
Paper • 2310.09753 • Published • 3 -
Improving Large Language Model Fine-tuning for Solving Math Problems
Paper • 2310.10047 • Published • 6 -
LLaVA-Interactive: An All-in-One Demo for Image Chat, Segmentation, Generation and Editing
Paper • 2311.00571 • Published • 42
Harry Xie
Hackiey
AI & ML interests
None yet
Recent Activity
upvoted a collection about 16 hours ago
MiMo-V2.6 liked a dataset 4 months ago
wdndev/webnovel-chinese liked a dataset over 1 year ago
open-thoughts/OpenThoughts-114kOrganizations
None yet