Research on agents that can self-improve and evolve their capabilities through experience and reasoning

表格 26 results
NameCategoriesTopicsReferencesCredibilityCreateDateUpdateDateTitleRelatedNotesNavigationOrderEleventyTemplateEngineOverrideCreated
20260419_memskill_learning_evolving_memory_skillspaper_analysisself_evolving_agents, memory_mechanism, agent_architecture, llm, long_term_memorynanyang_technological_university, tsinghua_university, university_of_illinois_urbana_champaign100MemSkill: Learning and Evolving Memory Skills for Self-Evolving AgentsGEMS: Agent-Native Multimodal Generation with Memory and Skills[object Object][object Object]md
20260504_reasoningbank_scaling_agent_self_evolving_reasoning_memorypaper_analysisself_evolving_agents, reasoning_memory, test_time_scaling, agent_architecture, memory_mechanismgoogle_cloud_ai_research, webarena, mind2web, swebench_verified100ReasoningBank: Scaling Agent Self-Evolving with Reasoning MemorySelf-Evolving Agents: A Survey[object Object][object Object]md
20260504_self_evolving_agents_survey_asipaper_analysisself_evolving_agents, artificial_super_intelligence, survey, agent_architecture, memory_mechanismprinceton_university, tsinghua_university, sjtu, google_cloud_ai_research100A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super IntelligenceReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory[object Object][object Object]md
20260506_openclaw_rl_train_any_agent_simply_by_talkingpaper_analysisreinforce_learning, agent_architecture, llm, self_evolving_agents, reasoninggen_verse100OpenClaw-RL: Train Any Agent Simply by Talking[object Object][object Object]md
20260509_skillos_learning_skill_curation_self_evolving_agentspaper_analysisagent_architecture, memory_mechanism, reinforce_learning, self_evolving_agents, skill_curationgoogle_cloud_ai_research, grpo, skillos, alfworld100SkillOS: Learning Skill Curation for Self-Evolving Agentsself_evolving_agents_survey_asi, memory_os_of_ai_agent, agentevolver_self_evolving_agent, reasoningbank_scaling_agent_self_evolving_reasoning_memory[object Object][object Object]md
20260515_harnessing_agentic_evolutionpaper_analysisagent_architecture, self_evolving_agents, reasoning, memory_mechanism, reinforce_learning, workflow_optimizationdeepwisdom, hkust_gz, sjtu, tsinghua_university, nanyang_technological_university100harnessing_agentic_evolutionself_evolving_agents_survey_asi, skillos_learning_skill_curation_self_evolving_agents, aflow_automating_agentic_workflow_generation[object Object][object Object]md
20260520_amr_sd_token_level_credit_assignmentpaper_analysisreinforce_learning, reasoning, llm, reward_modeling, self_evolving_agentsgrpo, dapo, meituan, sciknoweval100AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignmentrl_long_horizon_reasoning_llm_expressiveness, skillos_learning_skill_curation_self_evolving_agents[object Object][object Object]md
20260523_moss_self_evolution_source_level_rewritingpaper_analysisself_evolving_agents, agent_architecture, llm, multi_agent_systemsclaw, hkust_gz100MOSS: Self-Evolution through Source-Level Rewriting in Autonomous Agent Systemsmethodology_selecting_composing_runtime_architecture_patterns_production_llm_agents, skillos_learning_skill_curation_self_evolving_agents, agent_memory_and_reasoning_frontier_survey_2026[object Object][object Object]md
20260526_skillopt_executive_strategy_self_evolving_agent_skillspaper_analysisself_evolving_agents, skill_curation, agent_architecture, reasoning, workflow_optimizationmicrosoft_research_asia, sjtu100SkillOpt: Executive Strategy for Self-Evolving Agent Skillsskillos, moss, harnessing_agentic_evolution, self_evolving_agents_survey[object Object][object Object]md
20260527_scaling_harness_agentic_aipaper_analysisagent_architecture, memory_mechanism, reasoning, self_evolving_agents, multi_agent_systemsuc_berkeley, cheetahclaws100From Model Scaling to System Scaling: Scaling the Harness in Agentic AIMemLineage, MemGPT, SkillOS, CalMem[object Object][object Object]md
20260528_CORE_contrastive_reflection_reasoningpaper_analysisreasoning, memory_mechanism, self_evolving_agents, contrastive_reflection, reinforce_learning, cognitive_sciencestanford_iris_lab, grpo, memgpt100CORE: Contrastive Reflection Enables Rapid Improvements in ReasoningCan RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key, MemLineage: Lineage-Guided Enforcement for LLM Agent Memory, AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment[object Object][object Object]md
20260528_muse_autoskill_self_evolving_skill_memorypaper_analysisagent_architecture, memory_mechanism, self_evolving_agents, skill_curation, continual_learningbytedance_seed, skillsbench, muse_autoskill100MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluationskillos_learning_skill_curation_self_evolving_agents, skillopt_executive_strategy_self_evolving_agent_skills, agent_memory_and_reasoning_frontier_survey_2026[object Object][object Object]md
20260529_agent_lifespan_engineering_for_deployed_systemspaper_analysismemory_mechanism, long_term_memory, agent_architecture, evaluation, self_evolving_agentsut_austin100Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed SystemsYour Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems[object Object][object Object]md
20260603_agentcl_continual_learning_language_agentspaper_analysiscontinual_learning, memory_mechanism, agent_architecture, evaluation, self_evolving_agentsagentcl, ohio_state_university100AGENTCL: Toward Rigorous Evaluation of Continual Learning in Language Agentslearning_fast_slow_llms_adapt_continually, meme_multi_entity_evolving_memory_evaluation, muse_autoskill_self_evolving_skill_memory[object Object][object Object]md
20260604_language_models_need_sleep_self_modify_consolidate_memoriespaper_analysismemory_mechanism, self_evolving_agents, continual_learning, reinforce_learning, reasoning, llmgoogle_brain, knowledge_seeding, sleep_paradigm100Language Models Need Sleep: Learning to Self-Modify and Consolidate Memoriesmemory_os_of_ai_agent, hipporag_neurobiologically_inspired_long_term_memory, muse_autoskill_self_evolving_skill_memory[object Object][object Object]md
20260605_aip_graph_representation_for_learning_and_governing_agent_skillspaper_analysisagent_architecture, skill_curation, self_evolving_agents, reasoning, workflow_optimizationskillsbench, agent_instruction_protocol95AIP: A Graph Representation for Learning and Governing Agent Skillsskillos_learning_skill_curation_self_evolving_agents, muse_autoskill_self_evolving_skill_memory, aflow_automating_agentic_workflow_generation, apwa_parallelizable_agentic_workflows[object Object][object Object]md
20260609_socratic_swe_self_evolving_coding_agents_trace_derived_skillspaper_analysisself_evolving_agents, memory_mechanism, skill_curation, reinforce_learning, reasoningsjtu, swebench_verified, terminalbench_2, grpo, skillos100Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skillsskillos_learning_skill_curation_self_evolving_agents, skillopt_executive_strategy_self_evolving_agent_skills, aflow_automating_agentic_workflow_generation, scaling_harness_agentic_ai, meme_multi_entity_evolving_memory_evaluation[object Object][object Object]md
20260617_evolvenav_proactive_preflection_self_evolving_memorypaper_analysisembodied_ai, semantic_navigation, self_evolving_agents, llm, memory_mechanismhkust_gz, nanyang_technological_university, habitat_simulator100EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navigation[object Object][object Object]md
20260624_self_harnesspaper_analysisagent_architecture, self_evolving_agents, llm, reasoning, test_time_scalingshanghai_ai_lab, terminalbench_2100Self-Harness: Harnesses That Improve Themselves[object Object][object Object]md
20260627_joint_learning_experiential_rules_policies_llm_agentspaper_analysismemory_mechanism, reinforce_learning, self_evolving_agents, reasoning, agent_architecturealfworld, grpo, sun_yat_sen_university100Joint Learning of Experiential Rules and Policies for Large Language Model Agentsunlocking_working_memory_latent_reasoning, memory_os_of_ai_agent, memskill_learning_evolving_memory_skills[object Object][object Object]md
20260701_severa_verified_synthesis_self_evolving_agentspaper_analysisagent_architecture, self_evolving_agents, reinforce_learning, reasoning, agent_securityuniversity_of_illinois_urbana_champaign, grpo, tau_squared_bench100SEVerA: Verified Synthesis of Self-Evolving Agentsskillos_learning_skill_curation_self_evolving_agents, muse_autoskill_self_evolving_skill_memory, trace_unified_rollout_budget_agentic_rl, amr_sd_token_level_credit_assignment, agent_memory_characterization_system_implications[object Object][object Object]md
20260703_self_evolving_world_models_llm_agent_planningpaper_analysisworld_model, self_evolving_agents, memory_mechanism, embodied_ai, reasoningnus, alfworld, deepmind100Self-Evolving World Models for LLM Agent Planningfrom_tokens_to_states_llms_world_models, world_models_in_pieces_structural_certification, memory_r1_enhancing_llm_agents_manage_utilize_memories_rl, stop_comparing_llm_agents_without_disclosing_harness[object Object][object Object]md
20260706_demopsd_disagreement_modulated_policy_self_distillationpaper_analysisreinforce_learning, reasoning, llm, self_evolving_agentsgrpo, sciknoweval, kl_distillation, gpqa100DemoPSD: Disagreement-Modulated Policy Self-Distillationrl_long_horizon_reasoning, vector_policy_optimization[object Object][object Object]md
20260707_evopolicygym_evaluating_autonomous_policy_evolutionpaper_analysisembodied_ai, reinforce_learning, self_evolving_agents, evaluation, agent_architecturecuhk_shenzhen, sjtu, tsinghua_university, ustc100EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environmentsdemopsd_disagreement_modulated_policy_self_distillation, llm_agents_social_structure_latent_objective_emergence, recontext_recursive_evidence_replay, evolvenav_proactive_preflection_self_evolving_memory[object Object][object Object]md
20260716_immersion_in_github_universe_scaling_coding_agents_masterypaper_analysismulti_agent_systems, software_engineering, coding_agents, self_evolving_agents, data_scalingrenmin_university, bytedance_seed, swe_bench100Immersion in the GitHub Universe: Scaling Coding Agents to Mastery20260710_loger_long_context_geometric_reconstruction, 20260708_solving_million_step_llm_zero_errors[object Object][object Object]md
20260725_evolving_commitments_self_adaptive_stspaper_analysismulti_agent_systems, agent_architecture, software_engineering, self_evolving_agents, reasoningfudan_university, the_open_university, university_of_trento100Evolving Commitments for Self-Adaptive Socio-Technical Systems[object Object][object Object]md
Powered by Forestry.md