ai-agent-book 精选快照(<2MB 代码与文档,来自 github.com/bojieli/ai-agent-book)
Build latest book artifacts / build (push) Canceled after 0s
dependency resolution / resolve (3.11) (push) Canceled after 0s
dependency resolution / resolve (3.13) (push) Canceled after 0s
deploy-pages / build (push) Canceled after 0s
deploy-pages / deploy (push) Canceled after 0s
i18n consistency check / check (push) Canceled after 0s
provider adoption tests / test (chapter2/context-compression) (push) Canceled after 0s
provider adoption tests / test (chapter2/prompt-injection) (push) Canceled after 0s
provider adoption tests / test (chapter2/system-hint) (push) Canceled after 0s
provider adoption tests / test (chapter3/log-sanitization) (push) Canceled after 0s
web-search-agent tests / test (push) Canceled after 0s
web-search-agent tests / agentbook (push) Canceled after 0s
Build latest book artifacts / build (push) Canceled after 0s
dependency resolution / resolve (3.11) (push) Canceled after 0s
dependency resolution / resolve (3.13) (push) Canceled after 0s
deploy-pages / build (push) Canceled after 0s
deploy-pages / deploy (push) Canceled after 0s
i18n consistency check / check (push) Canceled after 0s
provider adoption tests / test (chapter2/context-compression) (push) Canceled after 0s
provider adoption tests / test (chapter2/prompt-injection) (push) Canceled after 0s
provider adoption tests / test (chapter2/system-hint) (push) Canceled after 0s
provider adoption tests / test (chapter3/log-sanitization) (push) Canceled after 0s
web-search-agent tests / test (push) Canceled after 0s
web-search-agent tests / agentbook (push) Canceled after 0s
This commit is contained in:
@@ -0,0 +1,33 @@
|
||||
# Contributing Agent Tasks
|
||||
|
||||
Contribute your own agent tasks and we test if the agent solves them for CI testing!
|
||||
|
||||
## How to Add a Task
|
||||
|
||||
1. Create a new `.yaml` file in this directory (`tests/agent_tasks/`).
|
||||
2. Use the following format:
|
||||
|
||||
```yaml
|
||||
name: My Task Name
|
||||
task: Describe the task for the agent to perform
|
||||
judge_context:
|
||||
- List criteria for success, one per line
|
||||
max_steps: 10
|
||||
```
|
||||
|
||||
## Guidelines
|
||||
- Be specific in your task and criteria.
|
||||
- The `judge_context` should list what counts as a successful result.
|
||||
- The agent's output will be judged by an LLM using these criteria.
|
||||
|
||||
## Running the Tests
|
||||
|
||||
To run all agent tasks:
|
||||
|
||||
```bash
|
||||
pytest tests/ci/test_agent_real_tasks.py
|
||||
```
|
||||
|
||||
---
|
||||
|
||||
Happy contributing!
|
||||
@@ -0,0 +1,7 @@
|
||||
name: Amazon Laptop Search
|
||||
task: Go to amazon.com, search for 'laptop', and return the first result
|
||||
judge_context:
|
||||
- The agent must navigate to amazon.com
|
||||
- The agent must search for 'laptop'
|
||||
- The agent must return name of the first laptop
|
||||
max_steps: 10
|
||||
@@ -0,0 +1,5 @@
|
||||
name: Find pip install command for browser-use
|
||||
task: Find the pip installation command for the browser-use repo
|
||||
judge_context:
|
||||
- The output must include the command ('pip install browser-use')
|
||||
max_steps: 10
|
||||
Reference in New Issue
Block a user