Large Language Models · Agents · Reasoning
Kou Shi
My work focuses on LLM agents, tool use, lifelong skill learning, and scientific agents.
Education
Education
University of Science and Technology of China USTC
M.S. Student
Hangzhou Dianzi University HDU
Bachelor's Degree
Research
Publications & Technical Reports
-
2026
Intern-S2-Preview: Scientific Agentic Foundation Model
A technical report on scientific agentic foundation models for multimodal understanding, reasoning, generation, and long-horizon tasks.
arXiv:2608.13505 -
2026
AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios
A benchmark for asynchronous function calling, dependency tracking, and coordination under multi-task workloads and tool latency.
arXiv:2605.27995 -
2026
SaaSBench: Exploring the Boundaries of Coding Agents in Long-Horizon Enterprise SaaS Engineering
A benchmark for coding agents in long-horizon enterprise SaaS engineering across heterogeneous stacks and system-level integration.
arXiv:2605.17526 -
2026
SkillFlow: Benchmarking Lifelong Skill Discovery and Evolution for Autonomous Agents
A benchmark for how autonomous agents discover, repair, transfer, and maintain skills over sequential experience.
arXiv:2604.17308 -
2026
Internalizing Meta-Experience into Memory for Guided Reinforcement Learning in Large Language Models
Internalizing error-derived meta-experience into parametric memory to improve reusable reasoning knowledge.
arXiv:2602.10224
Internships
Internships
vivo Hangzhou R&D Center
Large Model Algorithm Intern
AI Lab
Foundation Model Intern
Collaborative Research
Collaborative Research
Research involving my broader group is listed separately from papers on which I am a named author.