Zhongxiang Dai
dzxagent
AI & ML interests
Agents, LLMs
Recent Activity
upvoted a paper about 1 hour ago
COBRA-Skills: Contextual Bandit-Guided Evolution for Agent Skill Optimization submitted a paper about 1 month ago
SPOT: Sparse Probing and Outcome Calibration for On-Policy Distillation authored a paper about 1 month ago
SPOT: Sparse Probing and Outcome Calibration for On-Policy DistillationOrganizations
None yet