<a href="https://colab.research.google.com/github/CMBAgents/cmbagent/blob/main/docs/notebooks/cmbagent_colab.ipynb" target="_parent"><img src="https://colab.research.google.com/assets/colab-badge.svg" alt="Open In Colab"/></a>

# Installation

To ensure package dependencies are consistant first do:

In [17]:
pip install --upgrade openai httpx

Collecting openai
  Downloading openai-1.57.4-py3-none-any.whl.metadata (24 kB)
Downloading openai-1.57.4-py3-none-any.whl (390 kB)
[2K   [90m━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━━[0m [32m390.3/390.3 kB[0m [31m7.7 MB/s[0m eta [36m0:00:00[0m
[?25hInstalling collected packages: openai
  Attempting uninstall: openai
    Found existing installation: openai 1.54.5
    Uninstalling openai-1.54.5:
      Successfully uninstalled openai-1.54.5
Successfully installed openai-1.57.4


After running this code, restart your runtime (i.e., restart session).

Then the latest openai version should show up from the cell below.

In [1]:
import openai
openai.__version__

'1.57.4'

## OpenAI API key

The key can be set in the secrets panels (see key icon) in left panel of colab workspace.



In [2]:
import json
from openai import OpenAI
import os
from google.colab import userdata

It can the be acccessed via:

In [3]:
userdata.get('OAIKEY')

## Install cmbagent from PyPi

In [3]:
pip install cmbagent



Check that cmbagent imports correctly:

In [4]:
import cmbagent

Dask dataframe query planning is disabled because dask-expr is not installed.

You can install it with `pip install dask[dataframe]` or `conda install dask`.
This will raise in a future version.



CMBAgent version: 0.0.1post1
CMBAGENT_DATA not set.
Defaulting to: /root


To run machine learning-accelerated calculation, let us initialize the class_sz package. To do so, we simply import it.

In [5]:
import classy_sz

PATH_TO_CLASS_SZ_DATA not set. Setting it now...
Defaulting to: /root/class_sz_data_directory




## Get the RAG data for cosmology calculations

In [7]:
# Set the environment variable
os.environ["CMBAGENT_DATA"] = "/content/cmbagent_data"  # Path where data will be (or is) stored

In [10]:
# Clone the repository
!git clone https://github.com/CMBAgents/cmbagent_data.git $CMBAGENT_DATA

# Verify the repository is cloned
!ls /content/cmbagent_data

Cloning into '/content/cmbagent_data'...
remote: Enumerating objects: 88, done.[K
remote: Counting objects: 100% (88/88), done.[K
remote: Compressing objects: 100% (78/78), done.[K
remote: Total 88 (delta 8), reused 82 (delta 7), pack-reused 0 (from 0)[K
Receiving objects: 100% (88/88), 37.24 MiB | 24.83 MiB/s, done.
Resolving deltas: 100% (8/8), done.
data  LICENSE  README.md


# Cosmology calculations

To run a cosmology calculation, you first decide which agent you need and then instantiate cmbagent.

In [9]:
from cmbagent import CMBAgent
cmbagent = CMBAgent(llm_api_key = userdata.get('OAIKEY'),
                    agent_list = ['classy_sz','act'],
                    # make_vector_stores=['act'],
                    skip_memory = False,
                    verbose=False)
cmbagent.clear_cache()

Checking agent: classy_sz_agent
Agent classy_sz_agent exists in available assistants


Checking agent: act_agent
Agent act_agent exists in available assistants


Checking agent: memory_agent
Agent memory_agent exists in available assistants


	- admin
	- planner
	- engineer
	- executor
	- classy_sz_agent
	- act_agent
	- memory_agent
assistant not set-up properly, relaying error message
Vector store not found for memory_agent
re-instantiating with make_vector_stores=['memory'],
setting make_vector_stores=['memory'],
Updating vector store for memory_agent
Files to upload:
	 - summary_2024-11-21_23:59:56.json
completed
FileCounts(cancelled=0, completed=1, failed=0, in_progress=0, total=1)
memory_agent: uploaded assistant data to vector store with id:  vs_rMeCEjk0rGu5mAfbV5jjT0XL


vector stores updated
'memory_agent': 'vs_rMeCEjk0rGu5mAfbV5jjT0XL',


  self.push_vector_stores([agent.name.removesuffix('_agent')], chunking_strategy, verbose = verbose)


Then pick a task and solve it.

In [10]:
task = "Compute a gaussian centered on best-fit H0 value from ACT."
cmbagent.solve(task)

admin (to chat_manager):

Compute a gaussian centered on best-fit H0 value from ACT.

--------------------------------------------------------------------------------

Next speaker: planner



/usr/local/lib/python3.10/dist-packages/autogen/oai/client.py:834: PydanticDeprecatedSince20: `schema_json_of` is deprecated. Use `pydantic.TypeAdapter.json_schema` instead. Deprecated in Pydantic V2.0 to be removed in V3.0. See Pydantic V2 Migration Guide at https://errors.pydantic.dev/2.10/migration/
  {**params, **{"response_format": schema_json_of(params["response_format"])}}
/usr/local/lib/python3.10/dist-packages/pydantic/deprecated/tools.py:101: PydanticDeprecatedSince20: `schema_of` is deprecated. Use `pydantic.TypeAdapter.json_schema` instead. Deprecated in Pydantic V2.0 to be removed in V3.0. See Pydantic V2 Migration Guide at https://errors.pydantic.dev/2.10/migration/
  schema_of(type_, title=title, by_alias=by_alias, ref_template=ref_template, schema_generator=schema_generator),


Model,agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o-2024-08-06,planner,0.00462,988,215,1203


planner (to chat_manager):

To solve the task of computing a Gaussian centered on the best-fit H0 value from ACT, we will break it down into the following steps:

- Main task: Compute a Gaussian centered on the best-fit H0 value from ACT.

- Step 1:
  * sub-task: Retrieve the best-fit H0 value from ACT data.
  * agent: act_agent

- Step 2:
  * sub-task: Check if there are any past tasks related to computing a Gaussian centered on a best-fit value.
  * agent: memory_agent

- Step 3:
  * sub-task: Write Python code to compute a Gaussian centered on the retrieved best-fit H0 value.
  * agent: engineer

- Step 4:
  * sub-task: Ensure the code aligns with guidelines and is free of conflicts.
  * agent: engineer

- Step 5:
  * sub-task: Execute the Python code to compute the Gaussian.
  * agent: executor

Please let me know if we should proceed with this plan or if any modifications are needed.

--------------------------------------------------------------------------------

Next speaker: a

/usr/local/lib/python3.10/dist-packages/autogen/oai/client.py:834: PydanticDeprecatedSince20: `schema_json_of` is deprecated. Use `pydantic.TypeAdapter.json_schema` instead. Deprecated in Pydantic V2.0 to be removed in V3.0. See Pydantic V2 Migration Guide at https://errors.pydantic.dev/2.10/migration/
  {**params, **{"response_format": schema_json_of(params["response_format"])}}
/usr/local/lib/python3.10/dist-packages/pydantic/deprecated/tools.py:101: PydanticDeprecatedSince20: `schema_of` is deprecated. Use `pydantic.TypeAdapter.json_schema` instead. Deprecated in Pydantic V2.0 to be removed in V3.0. See Pydantic V2 Migration Guide at https://errors.pydantic.dev/2.10/migration/
  schema_of(type_, title=title, by_alias=by_alias, ref_template=ref_template, schema_generator=schema_generator),


Model,agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o-2024-08-06,speaker_selection_agent,0.0022,870,2,872



Next speaker: act_agent



Model,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o,0.10523,20537,170,20707


act_agent (to chat_manager):

- **Part 1**: Retrieved the best-fit H0 value from ACT data.
- **Part 2**: [2304.05203v1.pdf]
- **Part 3**:

**H_0**: 68.3 km s⁻¹ Mpc⁻¹ (Equation 6 from 2304.05203v1.pdf) 

- **Part 4**: We will now check if there are any past tasks related to computing a Gaussian centered on a best-fit value with the memory_agent.


--------------------------------------------------------------------------------

Next speaker: admin

Replying as admin. Provide feedback to chat_manager. Press enter to skip and use auto-reply, or type 'exit' to end the conversation: skip memory agent
admin (to chat_manager):

skip memory agent

--------------------------------------------------------------------------------


Model,agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o-2024-08-06,speaker_selection_agent,0.00251,997,2,999



Next speaker: engineer



Model,agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o-2024-08-06,engineer,0.00641,2223,85,2308


engineer (to chat_manager):

- **Part 1**: Skipped the memory agent step as per the admin's instruction.
- **Part 2**: No conflicts to resolve since the memory agent step was skipped.
- **Part 3**: Proceeding to write Python code to compute a Gaussian centered on the best-fit H0 value.
- **Part 4**: Next step is to write the Python code with the engineer agent.

--------------------------------------------------------------------------------

Next speaker: admin

Replying as admin. Provide feedback to chat_manager. Press enter to skip and use auto-reply, or type 'exit' to end the conversation: proceed
admin (to chat_manager):

proceed

--------------------------------------------------------------------------------


Model,agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o-2024-08-06,speaker_selection_agent,0.00276,1097,2,1099



Next speaker: engineer



Model,agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o-2024-08-06,engineer,0.01538,4546,402,4948


engineer (to chat_manager):

- **Part 1**: Writing Python code to compute a Gaussian centered on the best-fit H0 value from ACT.
- **Part 2**: No conflicts were found in the parameters.
- **Part 3**:
```python
import numpy as np
import matplotlib.pyplot as plt

# Best-fit H0 value from ACT
H0_best_fit = 68.3  # km s⁻¹ Mpc⁻¹

# Standard deviation for the Gaussian
sigma = 1.0  # You can adjust this value as needed

# Generate H0 values around the best-fit value
H0_values = np.linspace(H0_best_fit - 5*sigma, H0_best_fit + 5*sigma, 1000)

# Compute the Gaussian distribution
gaussian = (1/(sigma * np.sqrt(2 * np.pi))) * np.exp(-0.5 * ((H0_values - H0_best_fit) / sigma)**2)

# Plot the Gaussian distribution
plt.figure(figsize=(8, 5))
plt.plot(H0_values, gaussian, label=f'Gaussian centered at H0 = {H0_best_fit}')
plt.xlabel('H0 (km s⁻¹ Mpc⁻¹)')
plt.ylabel('Probability Density')
plt.title('Gaussian Distribution of H0')
plt.legend()
plt.grid(True)
plt.savefig('gaussian_H0_distribution.png')
plt

Model,agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o-2024-08-06,speaker_selection_agent,0.00358,1429,1,1430



Next speaker: executor


>>>>>>>> EXECUTING CODE BLOCK (inferred language is python)...
executor (to chat_manager):

exitcode: 0 (execution succeeded)
Code output: 

--------------------------------------------------------------------------------

Next speaker: admin

Replying as admin. Provide feedback to chat_manager. Press enter to skip and use auto-reply, or type 'exit' to end the conversation: exit




Unnamed: 0,Agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
0,planner,0.00462,988.0,215.0,1203.0
1,engineer,0.021792,6769.0,487.0,7256.0
2,act,0.105235,20537.0,170.0,20707.0
3,admin (speaker_selection_agent),0.002195,870.0,2.0,872.0
4,admin (speaker_selection_agent),0.002512,997.0,2.0,999.0
5,admin (speaker_selection_agent),0.002763,1097.0,2.0,1099.0
6,admin (speaker_selection_agent),0.003583,1429.0,1.0,1430.0
Total,,0.1427,32687.0,879.0,33566.0


Do you want to save this task summary to the "memory agent" vector stores? This will aid you and others in solving similar tasks in the future. Please only save the task if it has been completed successfully. Type "yes" or "no". yes
Asking planner to generate summary
The summary will be json formatted.



Prepared group chat with 10 messages, the last speaker is executor
planner (to chat_manager):


            Based on the conversation history, write a synthetic executive summary of the session.
            Specifically highlight the steps that required revision by admin and what was done to reach a successful answer that 
            was accepted the admin (i.e., who asked to proceed). 

            Your answer is formatted in json. 
            

--------------------------------------------------------------------------------

Next speaker: planner



/usr/local/lib/python3.10/dist-packages/autogen/oai/client.py:834: PydanticDeprecatedSince20: `schema_json_of` is deprecated. Use `pydantic.TypeAdapter.json_schema` instead. Deprecated in Pydantic V2.0 to be removed in V3.0. See Pydantic V2 Migration Guide at https://errors.pydantic.dev/2.10/migration/
  {**params, **{"response_format": schema_json_of(params["response_format"])}}
/usr/local/lib/python3.10/dist-packages/pydantic/deprecated/tools.py:101: PydanticDeprecatedSince20: `schema_of` is deprecated. Use `pydantic.TypeAdapter.json_schema` instead. Deprecated in Pydantic V2.0 to be removed in V3.0. See Pydantic V2 Migration Guide at https://errors.pydantic.dev/2.10/migration/
  schema_of(type_, title=title, by_alias=by_alias, ref_template=ref_template, schema_generator=schema_generator),


Model,agent,Cost,Prompt Tokens,Completion Tokens,Total Tokens
gpt-4o-2024-08-06,planner,0.01199,2995,450,3445


planner (to chat_manager):

{"main_task":"Compute a Gaussian centered on the best-fit H0 value from ACT.","results":"Successfully computed and plotted a Gaussian distribution centered on the best-fit H0 value from ACT data.","summary":[{"sub_task":"Retrieve the best-fit H0 value from ACT data.","result":"Retrieved the best-fit H0 value as 68.3 km s⁻¹ Mpc⁻¹ from ACT data.","feedback":"Proceed to the next step.","agent":"act_agent"},{"sub_task":"Check if there are any past tasks related to computing a Gaussian centered on a best-fit value.","result":"Skipped as per admin's instruction.","feedback":"Proceed to the next step.","agent":"memory_agent"},{"sub_task":"Write Python code to compute a Gaussian centered on the retrieved best-fit H0 value.","result":"Python code was written to compute and plot the Gaussian distribution.","feedback":"Proceed to execute the code.","agent":"engineer"},{"sub_task":"Execute the Python code to compute the Gaussian.","result":"Code executed successfully, a

  self.push_vector_stores(['memory'], None, verbose = True)
  self.push_vector_stores(['memory'], None, verbose = True)
