代睿 Wiki
Search
搜索
暗色模式
亮色模式
阅读模式
知识导航
首页
/
知识库
来源摘要
2026年8月28日
1分钟阅读
论文、文章与其他原始资料的可追溯摘要。
此文件夹下有20条笔记。
2026年8月28日
14年老编辑真诚分享:十大通用选题模板
wiki/source
content-creation
self-media
2026年8月28日
Harness Engineering for Self-Improvement
wiki/source
llm-agent
harness-engineering
2026年8月28日
Agent Plugins
wiki/source
llm-agent
plugin
interoperability
2026年8月28日
Shepherd: Enabling Programmable Meta-Agents via Reversible Agentic Execution Traces
wiki/source
llm-agent
meta-agent
agent-runtime
2026年8月28日
WebArena: A Realistic Web Environment for Building Autonomous Agents
wiki/source
paper
browser-use
benchmark
agent-evaluation
2026年8月28日
OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments
wiki/source
computer-use
benchmark
multimodal-agent
2026年8月28日
Meta-Harness: End-to-End Optimization of Model Harnesses
wiki/source
paper
harness-engineering
context-engineering
2026年8月28日
Agentic Harness Engineering: Observability-Driven Automatic Evolution of Coding-Agent Harnesses
wiki/source
paper
harness-engineering
self-improvement
2026年8月28日
Shepherd: Enabling Programmable Meta-Agents via Reversible Agentic Execution Traces
wiki/source
llm-agent
meta-agent
agent-runtime
agentic-rl
2026年8月28日
Self-Harness: Harnesses That Improve Themselves
wiki/source
paper
harness-engineering
self-improvement
2026年8月28日
Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures
wiki/source
paper
llm-agent
failure-analysis
2026年8月28日
HarnessOpt-Bench: Evaluating LLMs at Harness Optimization
wiki/source
paper
harness-engineering
benchmark
2026年8月28日
FACET: Preserving Source Intent and Executable State in Terminal Task Synthesis
wiki/source
paper
terminal-agent
executable-verification
harness-engineering
2026年8月28日
SkillForge: Evolving Verifiable Skills for Reinforcement Learning Agents
wiki/source
paper
agentic-rl
skill-learning
2026年8月28日
StarHarness: Evolving Harnesses with Stratified Search for Enterprise Environments
wiki/source
paper
harness-engineering
enterprise-agent
2026年8月28日
Recursive Experiential-Working Memory Evolution for Long-Horizon Agent Harnesses
wiki/source
paper
harness-engineering
recursive-self-improvement
2026年8月28日
PILOT in the Loop: Live Self-Improvement for Long-Horizon Agents
wiki/source
paper
harness-engineering
self-improvement
2026年8月28日
Astar: Learning to Propose Evolution Directions for Self-Evolving Industrial AI Systems
wiki/source
paper
agentic-rl
self-improvement
2026年8月28日
Verify Smarter, Evolve Further: Efficient Harness Evolution through Behavior-Aware Verification
wiki/source
paper
harness-engineering
self-improvement
2026年8月28日
Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Memory Requirements
wiki/source
paper
agentic-rl
long-horizon-agent