SAEVerbalizer Collection SAEVerbalizer: Generating Explanations for Sparse Autoencoder Features via Representation Verbalization • 3 items • Updated about 21 hours ago
HarnessDev: Can LLMs Create and Evolve Their Own Agent Harness? Paper • 2609.01437 • Published 20 days ago • 119
S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement? Paper • 2608.31100 • Published 21 days ago • 41
CogEvol: Towards Efficient and Reliable Learning Environment Generation Paper • 2608.30968 • Published 21 days ago • 32
EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery Paper • 2606.13662 • Published Jun 11 • 33
EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery Paper • 2606.13662 • Published Jun 11 • 33
EurekAgent: Agent Environment Engineering is All You Need For Autonomous Scientific Discovery Paper • 2606.13662 • Published Jun 11 • 33
Diagnosing Harmful Continuation in Answer-Correct Long-CoT Training Traces Paper • 2605.29288 • Published May 28 • 8
LongTraceRL Collection LongTraceRL: Learning Long-Context Reasoning from Search Agent Trajectories with Rubric Rewards • 5 items • Updated Jun 1 • 1