Stoney Kang
sikang99
AI & ML interests
Remote Control based on Vision
Recent Activity
upvoted a paper about 23 hours ago
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement upvoted a paper about 23 hours ago
QQWorld: Quantile-Quantile Matching for World Model Regularization