ScaleLogic: synthetic logical reasoning framework for studying RL scaling with long-horizon reasoning problems

表格 1 results
NameCategoriesTopicsReferencesCredibilityCreateDateUpdateDateTitleRelatedNotesNavigationOrderEleventyTemplateEngineOverrideCreated
20260511_rl_long_horizon_reasoning_llm_expressivenesspaper_analysisreinforce_learning, reasoning, llm, test_time_scaling, symbolic_reasoningpurdue_university, georgia_tech, scalelogic100Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Keychain_of_thought_prompting_elicits_reasoning, reasoningbank_scaling_agent_self_evolving_reasoning_memory[object Object][object Object]md
Powered by Forestry.md