Efficient Switchable Safety Control in LLMs via Magic-Token-Guided Co-Training Paper • 2508.14904 • Published Aug 12, 2025 • 2
NL2Repo-Bench: Towards Long-Horizon Repository Generation Evaluation of Coding Agents Paper • 2512.12730 • Published Dec 14, 2025 • 52
TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment Paper • 2601.18292 • Published Jan 26 • 12
Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows Paper • 2605.27922 • Published May 27 • 1
ARC: Active and Reflection-driven Context Management for Long-Horizon Information Seeking Agents Paper • 2601.12030 • Published Jan 17
RealClawBench: Live OpenClaw Benchmarks from Real Developer-Agent Sessions Paper • 2606.03889 • Published Jun 5 • 1
S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement? Paper • 2608.31100 • Published 11 days ago • 39
S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement? Paper • 2608.31100 • Published 11 days ago • 39
Training Agents to Evolve with Their Harness: TaoLive Digital Avatar Agent Technical Report Paper • 2608.15763 • Published 20 days ago • 54
TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment Paper • 2601.18292 • Published Jan 26 • 12