# CaptainAgent
By: Jiale Liu, Linxin Song, Jieyu Zhang, Shaokun Zhang

In this notebook, we introduce CaptainAgent, an agent enhanced with the capability to call AutoBuild to break down and solve complex tasks. AutoBuild can initiate a group chat between a group of experts and converse to solve the task. The experts in nested chat can be retrieved from agent library. The agents can be equipped with tools for advanced coding.

````{=mdx}
:::info Requirements
Install `pyautogen`:
```bash
pip install pyautogen[autobuild]
```

For more information, please refer to the [installation guide](/docs/installation/).
:::
````

## Setup API endpoint
In order to setup API, you should create a OAI_CONFIG_LIST file. The config list should look like the following:
```python
config_list = [
    {
        'model': 'gpt-4o-mini',
        'api_key': '<your OpenAI API key here>',
    },
    {
        'model': 'gpt-3.5-turbo',
        'api_key': '<your Azure OpenAI API key here>',
        'base_url': '<your Azure OpenAI API base here>',
        'api_type': 'azure',
        'api_version': '2024-08-01-preview',
    },
    {
        'model': 'gpt-3.5-turbo-16k',
        'api_key': '<your Azure OpenAI API key here>',
        'base_url': '<your Azure OpenAI API base here>',
        'api_type': 'azure',
        'api_version': '2024-08-01-preview',
    },
]
```

````{=mdx}
:::tip
Learn more about configuring LLMs for agents [here](/docs/topics/llm_configuration).
:::
````

In [1]:
import autogen

config_path = "OAI_CONFIG_LIST"
llm_config = {"temperature": 0}
config_list = autogen.config_list_from_json(
    config_path, filter_dict={"model": ["gpt-4o"]}
)  # You can modify the filter_dict to select your model

## Using CaptainAgent without libraries
We begin with demonstrating how to use CaptainAgent without retrieving from libaries. In this case, CaptainAgent will automatically generate a set of experts according to its identified subtask and initate the group chat.

In [None]:
from autogen import UserProxyAgent
from autogen.agentchat.contrib.captainagent import CaptainAgent

llm_config = {"temperature": 0, "config_list": config_list}

## build agents
captain_agent = CaptainAgent(
    name="captain_agent",
    llm_config=llm_config,
    code_execution_config={"use_docker": False, "work_dir": "groupchat"},
    agent_config_save_path=None,  # If you'd like to save the created agents in nested chat for further use, specify the save directory here
)
captain_user_proxy = UserProxyAgent(name="captain_user_proxy", human_input_mode="NEVER")

In [3]:
result = captain_user_proxy.initiate_chat(
    captain_agent,
    message="Find a recent paper about large language models on arxiv and find its potential applications in software.",
    max_turns=1,
)

[33mcaptain_user_proxy[0m (to captain_agent):

Find a recent paper about large language models on arxiv and find its potential applications in software.

--------------------------------------------------------------------------------
[34m
********************************************************************************[0m
[34mStarting a new chat....[0m
[34m
********************************************************************************[0m
[33mExpert_summoner[0m (to CaptainAgent):

Find a recent paper about large language models on arxiv and find its potential applications in software.

--------------------------------------------------------------------------------
[31m
>>>>>>>> USING AUTO REPLY...[0m
[33mCaptainAgent[0m (to Expert_summoner):

To find a recent paper about large language models on arXiv and explore its potential applications in software, I will follow these steps:

1. **Search for a Recent Paper**: I will search for a recent paper on arXiv related to larg

## Building Agents from library & Retrieve tools from tool library
One significant feature of CaptainAgent is that the agents and tools can be retrieved from a dedicated library. When CaptainAgent starts building experts for group chat, it will retrieve and select from agent library, then assign tools retreived from tool library to the experts.

For agent library, refer to `captainagent_expert_library.json` for samples. You can refer to [docs](https://ag2ai.github.io/ag2/docs/topics/captainagent/agent_library) on how to customize your own expert library.

For tool library, we provide a set of tools [here](https://github.com/ag2ai/ag2/tree/main/autogen/agentchat/contrib/captainagent/tools/README.md), the tools are categorized into three types: data analysis, information_retrieval, math.

### Using Agent Library Only
Below is an example that retrieves experts from library and build nested chat accordingly.


In [4]:
from autogen import UserProxyAgent
from autogen.agentchat.contrib.captainagent import CaptainAgent

nested_config = {
    # this is used to configure the autobuild building process
    "autobuild_build_config": {
        # this is used to configure the user proxy within nested chat
        "library_path_or_json": "captainagent_expert_library.json",
    }
}

Now let's instantiate the agents and ask it to solve a problem.

In [11]:
# build agents
captain_agent = CaptainAgent(
    name="captain_agent",
    llm_config=llm_config,
    nested_config=nested_config,
    code_execution_config={"use_docker": False, "work_dir": "groupchat"},
)
captain_user_proxy = UserProxyAgent(name="captain_user_proxy", human_input_mode="NEVER")

query = "find papers on LLM applications from arxiv in the last week, create a markdown table of different domains. After collecting the data, point out future research directions in light of the collected data."

result = captain_user_proxy.initiate_chat(captain_agent, message=query)

[33mcaptain_user_proxy[0m (to captain_agent):

find papers on LLM applications from arxiv in the last week, create a markdown table of different domains. After collecting the data, point out future research directions in light of the collected data.

--------------------------------------------------------------------------------
[34m
********************************************************************************[0m
[34mStarting a new chat....[0m
[34m
********************************************************************************[0m
[33mExpert_summoner[0m (to CaptainAgent):

find papers on LLM applications from arxiv in the last week, create a markdown table of different domains. After collecting the data, point out future research directions in light of the collected data.

--------------------------------------------------------------------------------
[31m
>>>>>>>> USING AUTO REPLY...[0m
[33mCaptainAgent[0m (to Expert_summoner):

To address this task, I will follow t

## Using Both Agent Library and Tool Library
Now let's retrieve from both agent library and tool library while building experts for nested chat.

To run the following demo, it is required to install the dependencies from the tool library and obtain BING search api and Rapid API key for tools in library to fully function. Please follow the instructions [here](https://ag2ai.github.io/ag2/docs/topics/captainagent/tool_library).

In [12]:
from autogen import UserProxyAgent
from autogen.agentchat.contrib.captainagent import CaptainAgent

nested_config = {
    # this is used to configure the autobuild building process
    "autobuild_build_config": {
        "library_path_or_json": "captainagent_expert_library.json",
    },
    # this is used to configure the tool library
    "autobuild_tool_config": {
        "tool_root": "default",  # Setting the value to default will use the tool library we provide
    },
}

In [None]:
# The function requires BING api key and Rapid API key to work. You can follow the instructions from docs to get one.
import os

os.environ["BING_API_key"] = ""  # set your bing api key here, if you donot need search engine, you can skip this step
os.environ["RAPID_API_KEY"] = (
    ""  # set your rapid api key here, in order for this example to work, you need to subscribe to the youtube transcription api(https://rapidapi.com/solid-api-solid-api-default/api/youtube-transcript3)
)

## build agents
captain_agent = CaptainAgent(
    name="captain_agent",
    llm_config=llm_config,
    nested_config=nested_config,
    code_execution_config={"use_docker": False, "work_dir": "groupchat"},
    agent_config_save_path=None,  # If you'd like to save the created agents in nested chat for further use, specify the save directory here
)
captain_user_proxy = UserProxyAgent(name="captain_user_proxy", human_input_mode="NEVER")

query = """# Task
Your task is to solve a question given by a user.

# Question
Examine the video at https://www.youtube.com/watch?v=1htKBjuUWec.

What does Teal'c say in response to the question "Isn't that hot?"
""".strip()
result = captain_user_proxy.initiate_chat(captain_agent, message=query)

[33mcaptain_user_proxy[0m (to captain_agent):

# Task
Your task is to solve a question given by a user.

# Question
Examine the video at https://www.youtube.com/watch?v=1htKBjuUWec.

What does Teal'c say in response to the question "Isn't that hot?"

--------------------------------------------------------------------------------
[34m
********************************************************************************[0m
[34mStarting a new chat....[0m
[34m
********************************************************************************[0m
[33mExpert_summoner[0m (to CaptainAgent):

# Task
Your task is to solve a question given by a user.

# Question
Examine the video at https://www.youtube.com/watch?v=1htKBjuUWec.

What does Teal'c say in response to the question "Isn't that hot?"

--------------------------------------------------------------------------------
[31m
>>>>>>>> USING AUTO REPLY...[0m
[33mCaptainAgent[0m (to Expert_summoner):

To solve this task, I will follow the



[33mComputer_terminal[0m (to chat_manager):

exitcode: 0 (execution succeeded)
Code output: {'success': True, 'transcript': "Wow this coffee's great I was just thinking that yeah is that cinnamon chicory tea oak [Music] isn't that hot extremely"}
Wow this coffee's great I was just thinking that yeah is that cinnamon chicory tea oak [Music] isn't that hot extremely


--------------------------------------------------------------------------------
[32m
Next speaker: APIProcessing_Expert
[0m
[33mAPIProcessing_Expert[0m (to chat_manager):

Based on the retrieved captions, Teal'c's response to the question "Isn't that hot?" is: 

"extremely"

This confirms the accuracy of the transcription regarding the specific dialogue. If there are any additional questions or tasks, feel free to let me know!

--------------------------------------------------------------------------------
[32m
Next speaker: APIProcessing_Expert
[0m
[33mAPIProcessing_Expert[0m (to chat_manager):

TERMINATE

----



Creating agent VideoTranscription_Expert...
Adding user console proxy...
[33mAPIProcessing_Expert[0m (to chat_manager):

You are given: (1) a task and advises from your manager with a specific plan and (2) a general task.
Collect information from the general task, follow the suggestions from manager to solve the task.

# General Task
# Task
Your task is to solve a question given by a user.

# Question
Examine the video at https://www.youtube.com/watch?v=1htKBjuUWec.

What does Teal'c say in response to the question "Isn't that hot?"

# Task and suggestions from manager
## Task description\nExamine the video at https://www.youtube.com/watch?v=1htKBjuUWec and find out what Teal'c says in response to the question "Isn't that hot?"\n\n## Plan for solving the task\n- Access the video and locate the moment when the question "Isn't that hot?" is asked.\n- Transcribe Teal'c's response accurately.\n\n## Output format\n- Provide the exact quote of Teal'c's response.\n\n## Constraints and condi



[33mComputer_terminal[0m (to chat_manager):

exitcode: 0 (execution succeeded)
Code output: {'success': True, 'transcript': "Wow this coffee's great I was just thinking that yeah is that cinnamon chicory tea oak [Music] isn't that hot extremely"}
Wow this coffee's great I was just thinking that yeah is that cinnamon chicory tea oak [Music] isn't that hot extremely


--------------------------------------------------------------------------------




[32m
Next speaker: APIProcessing_Expert
[0m
[33mAPIProcessing_Expert[0m (to chat_manager):

The captions from the video confirm that in response to the question "Isn't that hot?" Teal'c indeed replies, "extremely." This directly matches the task requirements, ensuring that the quote is accurate and faithful to the video content.

If there are no further questions or tasks related to this, I believe we can conclude the conversation. Do you agree, or is there anything else that needs attention?

--------------------------------------------------------------------------------




[32m
Next speaker: APIProcessing_Expert
[0m
[33mAPIProcessing_Expert[0m (to chat_manager):

TERMINATE

--------------------------------------------------------------------------------
[33mExpert_summoner[0m (to CaptainAgent):

[33mExpert_summoner[0m (to CaptainAgent):

[32m***** Response from calling tool (call_ycX3M2UaldCUrHgxaIF6kYqq) *****[0m
# Response from seek_agent_help: 
## Task
The task is to examine a YouTube video (https://www.youtube.com/watch?v=1htKBjuUWec) and find out what the character Teal'c says in response to the question "Isn't that hot?"

## Results
Teal'c's response to the question "Isn't that hot?" is "extremely," as confirmed by retrieving and reviewing the video captions using an API function.

## Reason for the results
The result was derived by accessing the video's captions through an API, which accurately reflected the dialogue in the video, confirming that the response "extremely" was indeed what Teal'c said in response to the question.

## Errors

The ground truth answer to the question is 'Extremely', which the CaptainAgent answers correctly. Notably, with the assistance of tools, the agent can answer video-related questions. This shows the huge potential of what user customized tools can bring.

# Further Reading
For a full list of the configurables and the functionalities in CaptainAgent, please refer [here](https://ag2ai.github.io/ag2/docs/topics/captainagent/configurations).

For how to customize your own agent library, check [here](https://ag2ai.github.io/ag2/docs/topics/captainagent/agent_library).

For how to implement your own tool library that suits your need, check the documents [here](https://ag2ai.github.io/ag2/docs/topics/captainagent/tool_library).
