Argonne National Laboratory, U.S. Department of Energy science and engineering research lab
| Name | Categories | Topics | References | Credibility | CreateDate | UpdateDate | Title | RelatedNotes | NavigationOrder | Eleventy | TemplateEngineOverride | Created |
|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 20260626_progress_advantage_llm_agents | paper_analysis | reinforce_learning, reward_modeling, test_time_scaling, agent_architecture, llm, reasoning | university_of_wisconsin_madison, argonne_national_laboratory | 100 | Neglected Free Lunch from Post-training: Progress Advantage for LLM Agents | rl_long_horizon_reasoning_llm_expressiveness, vector_policy_optimization_vpo, trace_unified_rollout_budget_allocation_agentic_rl | [object Object] | [object Object] | md |