This skill should be used for advanced LLM evaluation: LLM-as-judge systems, direct scoring, pairwise comparison, rubric calibration, evaluator bias mitigation,
所属插件包:context-engineering
正在加载项目…
A comprehensive collection of Agent Skills for context engineering, multi-agent architectures, and production agent systems. Use when building, optimizing, or debugging agent systems that require effective context management.
共 22 项匹配内容,示例会单独标注;未声明归属的内容不代表插件自动包含。
This skill should be used for advanced LLM evaluation: LLM-as-judge systems, direct scoring, pairwise comparison, rubric calibration, evaluator bias mitigation,
所属插件包:context-engineering
This skill should be used when modeling agent mental states with BDI concepts: beliefs, desires, intentions, RDF-to-belief transformations, rational agency trac
所属插件包:context-engineering
This skill should be used when long-running agent sessions need context compression, structured summarization, compaction, token-per-task optimization, or durab
所属插件包:context-engineering
This skill should be used for diagnosing and mitigating context degradation: lost-in-middle failures, context poisoning, context clash, context confusion, atten
所属插件包:context-engineering
A comprehensive collection of Agent Skills for context engineering, harness engineering, multi-agent architectures, and production agent systems. Use when build
其他仓库内容 · 未声明插件包归属
This skill should be used to explain or reason about the foundational concepts of context engineering: what context is, the anatomy of a context window, how att
所属插件包:context-engineering
This skill should be used for improving context efficiency: context budgeting, observation masking, prefix or KV-cache strategy, partitioning, token-cost reduct
所属插件包:context-engineering
This skill should be used when building agent evaluation systems: deterministic checks, regression suites, multi-dimensional rubrics, quality gates, production
所属插件包:context-engineering
This skill should be used when agent work needs file-backed context: durable scratchpads, tool-output offloading, just-in-time discovery, cross-agent handoff fi
所属插件包:context-engineering
This skill should be used when designing autonomous agent harnesses: research loops, evaluation scaffolds, locked and editable surfaces, durable logs, novelty g
所属插件包:context-engineering
This skill should be used when designing hosted or background agent infrastructure: sandboxed execution, remote coding environments, warm pools, session persist
所属插件包:context-engineering
This skill should be used when the user asks to "share memory between agents", "KV cache compaction for multi-agent", "orchestrator worker context", "latent bri
所属插件包:context-engineering
This skill should be used when writing, enhancing, or evaluating the launch prompt for a long-running autonomous agent or a parallel multi-agent orchestration a
所属插件包:context-engineering
This skill should be used for persistent semantic memory in agent systems: cross-session knowledge retention, entity tracking, temporal validity, graph or vecto
所属插件包:context-engineering
This skill should be used when designing multi-agent systems that need context isolation, supervisor or swarm coordination, explicit handoffs, parallel executio
所属插件包:context-engineering
This skill should be used for project-level decisions about LLM-powered systems: whether an LLM is the right primitive for the task at hand, the shape of a mult
所属插件包:context-engineering
This skill should be used when the harness, scaffold, workflow, or optimizer itself is the optimization target: recursive self-improvement (RSI) loops, meta-har
所属插件包:context-engineering
This skill should be used for the tool-interface layer of an agent system specifically: writing tool descriptions agents can route on, designing tool schemas an
所属插件包:context-engineering
This skill should be used for book-to-SFT pipelines: ePub extraction, literary segmentation, author-voice dataset construction, style-transfer training, LoRA wo
其他仓库内容 · 未声明插件包归属
Ensure thorough validation, error recovery, and transparent reasoning in research tasks with multiple tool calls
其他仓库内容 · 未声明插件包归属
This skill should be used for personal operating-system workflows: content creation, voice consistency, relationship lookup, meeting preparation, weekly review,
其他仓库内容 · 未声明插件包归属
Debug and optimize AI agents by analyzing reasoning traces, context degradation, tool confusion, instruction drift, repeated task failures, and performance regr
其他仓库内容 · 未声明插件包归属