| 20260419_memskill_learning_evolving_memory_skills | paper_analysis | self_evolving_agents, memory_mechanism, agent_architecture, llm, long_term_memory | nanyang_technological_university, tsinghua_university, university_of_illinois_urbana_champaign | 100 | | | MemSkill: Learning and Evolving Memory Skills for Self-Evolving Agents | GEMS: Agent-Native Multimodal Generation with Memory and Skills | [object Object] | [object Object] | md | |
| 20260504_reasoningbank_scaling_agent_self_evolving_reasoning_memory | paper_analysis | self_evolving_agents, reasoning_memory, test_time_scaling, agent_architecture, memory_mechanism | google_cloud_ai_research, webarena, mind2web, swebench_verified | 100 | | | ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory | Self-Evolving Agents: A Survey | [object Object] | [object Object] | md | |
| 20260504_self_evolving_agents_survey_asi | paper_analysis | self_evolving_agents, artificial_super_intelligence, survey, agent_architecture, memory_mechanism | princeton_university, tsinghua_university, sjtu, google_cloud_ai_research | 100 | | | A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence | ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory | [object Object] | [object Object] | md | |
| 20260506_openclaw_rl_train_any_agent_simply_by_talking | paper_analysis | reinforce_learning, agent_architecture, llm, self_evolving_agents, reasoning | gen_verse | 100 | | | OpenClaw-RL: Train Any Agent Simply by Talking | | [object Object] | [object Object] | md | |
| 20260509_skillos_learning_skill_curation_self_evolving_agents | paper_analysis | agent_architecture, memory_mechanism, reinforce_learning, self_evolving_agents, skill_curation | google_cloud_ai_research, grpo, skillos, alfworld | 100 | | | SkillOS: Learning Skill Curation for Self-Evolving Agents | self_evolving_agents_survey_asi, memory_os_of_ai_agent, agentevolver_self_evolving_agent, reasoningbank_scaling_agent_self_evolving_reasoning_memory | [object Object] | [object Object] | md | |
| 20260515_harnessing_agentic_evolution | paper_analysis | agent_architecture, self_evolving_agents, reasoning, memory_mechanism, reinforce_learning, workflow_optimization | deepwisdom, hkust_gz, sjtu, tsinghua_university, nanyang_technological_university | 100 | | | harnessing_agentic_evolution | self_evolving_agents_survey_asi, skillos_learning_skill_curation_self_evolving_agents, aflow_automating_agentic_workflow_generation | [object Object] | [object Object] | md | |
| 20260520_amr_sd_token_level_credit_assignment | paper_analysis | reinforce_learning, reasoning, llm, reward_modeling, self_evolving_agents | grpo, dapo, meituan, sciknoweval | 100 | | | AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment | rl_long_horizon_reasoning_llm_expressiveness, skillos_learning_skill_curation_self_evolving_agents | [object Object] | [object Object] | md | |
| 20260523_moss_self_evolution_source_level_rewriting | paper_analysis | self_evolving_agents, agent_architecture, llm, multi_agent_systems | claw, hkust_gz | 100 | | | MOSS: Self-Evolution through Source-Level Rewriting in Autonomous Agent Systems | methodology_selecting_composing_runtime_architecture_patterns_production_llm_agents, skillos_learning_skill_curation_self_evolving_agents, agent_memory_and_reasoning_frontier_survey_2026 | [object Object] | [object Object] | md | |
| 20260526_skillopt_executive_strategy_self_evolving_agent_skills | paper_analysis | self_evolving_agents, skill_curation, agent_architecture, reasoning, workflow_optimization | microsoft_research_asia, sjtu | 100 | | | SkillOpt: Executive Strategy for Self-Evolving Agent Skills | skillos, moss, harnessing_agentic_evolution, self_evolving_agents_survey | [object Object] | [object Object] | md | |
| 20260527_scaling_harness_agentic_ai | paper_analysis | agent_architecture, memory_mechanism, reasoning, self_evolving_agents, multi_agent_systems | uc_berkeley, cheetahclaws | 100 | | | From Model Scaling to System Scaling: Scaling the Harness in Agentic AI | MemLineage, MemGPT, SkillOS, CalMem | [object Object] | [object Object] | md | |
| 20260528_CORE_contrastive_reflection_reasoning | paper_analysis | reasoning, memory_mechanism, self_evolving_agents, contrastive_reflection, reinforce_learning, cognitive_science | stanford_iris_lab, grpo, memgpt | 100 | | | CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning | Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key, MemLineage: Lineage-Guided Enforcement for LLM Agent Memory, AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment | [object Object] | [object Object] | md | |
| 20260528_muse_autoskill_self_evolving_skill_memory | paper_analysis | agent_architecture, memory_mechanism, self_evolving_agents, skill_curation, continual_learning | bytedance_seed, skillsbench, muse_autoskill | 100 | | | MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation | skillos_learning_skill_curation_self_evolving_agents, skillopt_executive_strategy_self_evolving_agent_skills, agent_memory_and_reasoning_frontier_survey_2026 | [object Object] | [object Object] | md | |
| 20260529_agent_lifespan_engineering_for_deployed_systems | paper_analysis | memory_mechanism, long_term_memory, agent_architecture, evaluation, self_evolving_agents | ut_austin | 100 | | | Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems | Your Agents Are Aging Too: Agent Lifespan Engineering for Deployed Systems | [object Object] | [object Object] | md | |
| 20260603_agentcl_continual_learning_language_agents | paper_analysis | continual_learning, memory_mechanism, agent_architecture, evaluation, self_evolving_agents | agentcl, ohio_state_university | 100 | | | AGENTCL: Toward Rigorous Evaluation of Continual Learning in Language Agents | learning_fast_slow_llms_adapt_continually, meme_multi_entity_evolving_memory_evaluation, muse_autoskill_self_evolving_skill_memory | [object Object] | [object Object] | md | |
| 20260604_language_models_need_sleep_self_modify_consolidate_memories | paper_analysis | memory_mechanism, self_evolving_agents, continual_learning, reinforce_learning, reasoning, llm | google_brain, knowledge_seeding, sleep_paradigm | 100 | | | Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories | memory_os_of_ai_agent, hipporag_neurobiologically_inspired_long_term_memory, muse_autoskill_self_evolving_skill_memory | [object Object] | [object Object] | md | |
| 20260605_aip_graph_representation_for_learning_and_governing_agent_skills | paper_analysis | agent_architecture, skill_curation, self_evolving_agents, reasoning, workflow_optimization | skillsbench, agent_instruction_protocol | 95 | | | AIP: A Graph Representation for Learning and Governing Agent Skills | skillos_learning_skill_curation_self_evolving_agents, muse_autoskill_self_evolving_skill_memory, aflow_automating_agentic_workflow_generation, apwa_parallelizable_agentic_workflows | [object Object] | [object Object] | md | |
| 20260609_socratic_swe_self_evolving_coding_agents_trace_derived_skills | paper_analysis | self_evolving_agents, memory_mechanism, skill_curation, reinforce_learning, reasoning | sjtu, swebench_verified, terminalbench_2, grpo, skillos | 100 | | | Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills | skillos_learning_skill_curation_self_evolving_agents, skillopt_executive_strategy_self_evolving_agent_skills, aflow_automating_agentic_workflow_generation, scaling_harness_agentic_ai, meme_multi_entity_evolving_memory_evaluation | [object Object] | [object Object] | md | |
| 20260617_evolvenav_proactive_preflection_self_evolving_memory | paper_analysis | embodied_ai, semantic_navigation, self_evolving_agents, llm, memory_mechanism | hkust_gz, nanyang_technological_university, habitat_simulator | 100 | | | EvolveNav: Proactive Preflection and Self-Evolving Memory for Zero-Shot Object Goal Navigation | | [object Object] | [object Object] | md | |
| 20260624_self_harness | paper_analysis | agent_architecture, self_evolving_agents, llm, reasoning, test_time_scaling | shanghai_ai_lab, terminalbench_2 | 100 | | | Self-Harness: Harnesses That Improve Themselves | | [object Object] | [object Object] | md | |
| 20260627_joint_learning_experiential_rules_policies_llm_agents | paper_analysis | memory_mechanism, reinforce_learning, self_evolving_agents, reasoning, agent_architecture | alfworld, grpo, sun_yat_sen_university | 100 | | | Joint Learning of Experiential Rules and Policies for Large Language Model Agents | unlocking_working_memory_latent_reasoning, memory_os_of_ai_agent, memskill_learning_evolving_memory_skills | [object Object] | [object Object] | md | |
| 20260701_severa_verified_synthesis_self_evolving_agents | paper_analysis | agent_architecture, self_evolving_agents, reinforce_learning, reasoning, agent_security | university_of_illinois_urbana_champaign, grpo, tau_squared_bench | 100 | | | SEVerA: Verified Synthesis of Self-Evolving Agents | skillos_learning_skill_curation_self_evolving_agents, muse_autoskill_self_evolving_skill_memory, trace_unified_rollout_budget_agentic_rl, amr_sd_token_level_credit_assignment, agent_memory_characterization_system_implications | [object Object] | [object Object] | md | |
| 20260703_self_evolving_world_models_llm_agent_planning | paper_analysis | world_model, self_evolving_agents, memory_mechanism, embodied_ai, reasoning | nus, alfworld, deepmind | 100 | | | Self-Evolving World Models for LLM Agent Planning | from_tokens_to_states_llms_world_models, world_models_in_pieces_structural_certification, memory_r1_enhancing_llm_agents_manage_utilize_memories_rl, stop_comparing_llm_agents_without_disclosing_harness | [object Object] | [object Object] | md | |
| 20260706_demopsd_disagreement_modulated_policy_self_distillation | paper_analysis | reinforce_learning, reasoning, llm, self_evolving_agents | grpo, sciknoweval, kl_distillation, gpqa | 100 | | | DemoPSD: Disagreement-Modulated Policy Self-Distillation | rl_long_horizon_reasoning, vector_policy_optimization | [object Object] | [object Object] | md | |
| 20260707_evopolicygym_evaluating_autonomous_policy_evolution | paper_analysis | embodied_ai, reinforce_learning, self_evolving_agents, evaluation, agent_architecture | cuhk_shenzhen, sjtu, tsinghua_university, ustc | 100 | | | EvoPolicyGym: Evaluating Autonomous Policy Evolution in Interactive Environments | demopsd_disagreement_modulated_policy_self_distillation, llm_agents_social_structure_latent_objective_emergence, recontext_recursive_evidence_replay, evolvenav_proactive_preflection_self_evolving_memory | [object Object] | [object Object] | md | |
| 20260716_immersion_in_github_universe_scaling_coding_agents_mastery | paper_analysis | multi_agent_systems, software_engineering, coding_agents, self_evolving_agents, data_scaling | renmin_university, bytedance_seed, swe_bench | 100 | | | Immersion in the GitHub Universe: Scaling Coding Agents to Mastery | 20260710_loger_long_context_geometric_reconstruction, 20260708_solving_million_step_llm_zero_errors | [object Object] | [object Object] | md | |
| 20260725_evolving_commitments_self_adaptive_sts | paper_analysis | multi_agent_systems, agent_architecture, software_engineering, self_evolving_agents, reasoning | fudan_university, the_open_university, university_of_trento | 100 | | | Evolving Commitments for Self-Adaptive Socio-Technical Systems | | [object Object] | [object Object] | md | |