arxiv:2512.22322
Shaofei Cai
phython96
AI & ML interests
Embodied Decision Making, Computer Vision, Game AI, LLM Agents
Recent Activity
updated a model about 14 hours ago
CraftJarvis/ROCKET-3-1.5x authored a paper 9 months ago
Learn the Ropes, Then Trust the Wins: Self-imitation with Progressive
Exploration for Agentic Reinforcement Learning authored a paper 9 months ago
DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models