Specialized agent for executing GAIA benchmark runs, monitoring progress, and analyzing results
所属插件包:ruflo-workflows
正在加载项目…
🌊 The original agent harness. Deploy intelligent multi-player swarms, coordinate autonomous workflows, and build conversational AI systems. Features adaptive memory, self-learning intelligence, federation, vector RAG integration, and native Claude Code / Codex / Hermes and many more Integrated
共 11 项匹配内容,示例会单独标注;未声明归属的内容不代表插件自动包含。
Specialized agent for executing GAIA benchmark runs, monitoring progress, and analyzing results
所属插件包:ruflo-workflows
Specialized agent for packaging, signing, and coordinating HAL leaderboard submission of GAIA benchmark results
所属插件包:ruflo-workflows
Workflow automation specialist for creating, executing, and managing multi-step processes
所属插件包:ruflo-workflows
GAIA benchmark dispatcher — run, submit, validate, and track leaderboard scores against the Princeton HAL benchmark
所属插件包:ruflo-workflows
Report cumulative GAIA API spend and project cost for planned configurations
所属插件包:ruflo-workflows
Show measured benchmark runs stored across sessions in the gaia-runs memory namespace
所属插件包:ruflo-workflows
Fetch and display current HAL GAIA leaderboard scores and our positioning
所属插件包:ruflo-workflows
Execute a GAIA benchmark run — shells out to gaia-bench run, streams progress, and writes JSON results
所属插件包:ruflo-workflows
Package GAIA results into an Ed25519-signed, HAL-compatible submission archive
所属插件包:ruflo-workflows
Pre-submit validation — TypeScript clean, dataset accessible, all required env keys present
所属插件包:ruflo-workflows
Workflow management -- list MCP workflows + templates and native .claude/workflows/*.js scripts
所属插件包:ruflo-workflows