| 20260509_skillos_learning_skill_curation_self_evolving_agents | paper_analysis | agent_architecture, memory_mechanism, reinforce_learning, self_evolving_agents, skill_curation | google_cloud_ai_research, grpo, skillos, alfworld | 100 | | | SkillOS: Learning Skill Curation for Self-Evolving Agents | self_evolving_agents_survey_asi, memory_os_of_ai_agent, agentevolver_self_evolving_agent, reasoningbank_scaling_agent_self_evolving_reasoning_memory | [object Object] | [object Object] | md | |
| 20260627_joint_learning_experiential_rules_policies_llm_agents | paper_analysis | memory_mechanism, reinforce_learning, self_evolving_agents, reasoning, agent_architecture | alfworld, grpo, sun_yat_sen_university | 100 | | | Joint Learning of Experiential Rules and Policies for Large Language Model Agents | unlocking_working_memory_latent_reasoning, memory_os_of_ai_agent, memskill_learning_evolving_memory_skills | [object Object] | [object Object] | md | |
| 20260703_self_evolving_world_models_llm_agent_planning | paper_analysis | world_model, self_evolving_agents, memory_mechanism, embodied_ai, reasoning | nus, alfworld, deepmind | 100 | | | Self-Evolving World Models for LLM Agent Planning | from_tokens_to_states_llms_world_models, world_models_in_pieces_structural_certification, memory_r1_enhancing_llm_agents_manage_utilize_memories_rl, stop_comparing_llm_agents_without_disclosing_harness | [object Object] | [object Object] | md | |
| 20260717_experience_memory_graph_one_shot_error_correction_for_agents | paper_analysis | memory_mechanism, agent_architecture, reasoning, long_term_memory, knowledge_graph | alfworld, scienceworld, uestc | 100 | | | Experience Memory Graph: One-Shot Error Correction for Agents | memlineage_lineage_guided_llm_agent_memory, shared_selective_persistent_memory_agentic_llm, agent_memory_and_reasoning_frontier_survey_2026 | [object Object] | [object Object] | md | |