In [1]:
from autosearch.functions.text_analysis import chunk_pdf

from autosearch.database.paper_database import PaperDatabase
from autosearch.analysis.document_analyzer import DocumentAnalyzer
from autosearch.research_project import ResearchProject
from autosearch.write_blog import WriteBlog

import autogen
from typing import List, Dict, Any

In [2]:
import os
from dotenv import load_dotenv
from azure.core.exceptions import HttpResponseError

# Load environment variables
load_dotenv()

# Retrieve Azure credentials from environment variables
config={
    'doc_api_key': os.getenv("DOCUMENT_INTELLIGENCE_KEY"),
    'doc_endpoint': os.getenv("DOCUMENT_INTELLIGENCE_ENDPOINT")
}

In [3]:
title = "Exploring the Intricacies of Polymer Representation: Unraveling Complexity"

In [4]:
blog_project = WriteBlog(
    project_id = "polymer_representation",
    version= "0.2",
    config=config,
    config_file="../OAI_CONFIG_LIST-sweden-505",
    initiate_db= False,
    funcClsList = ["FactualCheck", "GetPDF", "GetPDFs", "UrlCheck", "AcademicRetriever", "AcademicSearch"],
    communiteList = ["outline_agents", "write_section_agents", "instructor_agents"],
    local_papers_dir="./papers"
)
project_config = blog_project.ProjectConfig
print(project_config.logging_session_id)


Logging session ID: 6888af83-9538-487b-b93e-485e9f9b808e
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
Equipping function 'academic_retriever' to agent 'blog_editor-in-chief'
Equipping function 'academic_search' to agent 'blog_editor-in-chief'
Equipping function 'get_pdf' to agent 'blog_editor-in-chief'
Equipping function 'get_pdfs' to agent 'blog_editor-in-chief'
Equipping function 'factual_check' to agent 'content_strategist'
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
Equipping function 'factual_check' to agent 'content_review_specialist'
Equipping function 'academic_retriever' to agent 'content_review_specialist'
Equipping function 'academic_search' to agent 'content_review_specialist'
Equipping function 'get_pdf' to agent 'content_review_specialist'
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
Equipping function 'academic_search' to agent 'topic_expert'
Equipping function 'academic_retriever' to



Equipping function 'academic_retriever' to agent 'research_resource_expert'
Equipping function 'get_pdf' to agent 'research_resource_expert'
Processing local PDFs...
Processing 2205.13757v1.pdf...
Skipping 2205.13757v1.pdf as it is already in the database.
Processing 2102.08134v2.pdf...
Skipping 2102.08134v2.pdf as it is already in the database.
Processing 2312.04013v3.pdf...
Skipping 2312.04013v3.pdf as it is already in the database.
Processing 1-s2.0-S0927796X2030053X-am.pdf...
Created 54 docs with a total of 31619 tokens. Largest doc has 2947 tokens.
Error searching google_scholar: Cannot Fetch from Google Scholar.
Error copying local PDF for Version of Record: https://www.sciencedirect.com/science/article/pii/S0927796X2030053X Manuscript_c681a8a49ed2ede90aaca95b72c753e1: './polymer_representation/0.2/output/1-s2.0-S0927796X2030053X-am.pdf' and './polymer_representation/0.2/output/1-s2.0-S0927796X2030053X-am.pdf' are the same file
Error processing ./papers/1-s2.0-S0927796X2030053X-a

In [5]:
blog_project.run(
    title=title,
    target_audience="expert in experimental polymer science and machine learning experts",
)

[33mtopic_expert[0m (to chat_manager):


        We need to create comprehensive instructions for the Editor-in-Chief to write an article titled:
        "Exploring the Intricacies of Polymer Representation: Unraveling Complexity"

        The target audience for this article is: expert in experimental polymer science and machine learning experts

        Each agent should contribute their expertise to craft a well-rounded instruction,
        keeping in mind the specific needs and background of the target audience.
        Topic Expert, please start by analyzing the title and providing key insights,
        considering how they should be presented to this audience.
        

--------------------------------------------------------------------------------
[32m
Next speaker: structure_specialist
[0m
[33mstructure_specialist[0m (to chat_manager):

Analyzing the title "Exploring the Intricacies of Polymer Representation: Unraveling Complexity," it's clear that the article aims to de

Add of existing embedding ID: 9
Add of existing embedding ID: 16
Add of existing embedding ID: 22
Add of existing embedding ID: 25
Add of existing embedding ID: 25
Add of existing embedding ID: 27
Add of existing embedding ID: 27
Add of existing embedding ID: 28
Add of existing embedding ID: 33
Add of existing embedding ID: 38
Add of existing embedding ID: 46
Add of existing embedding ID: 53
Add of existing embedding ID: 60
Add of existing embedding ID: 75
Add of existing embedding ID: 102
Add of existing embedding ID: 107
Add of existing embedding ID: 134
Add of existing embedding ID: 138
Add of existing embedding ID: 143
Add of existing embedding ID: 144
Add of existing embedding ID: 150
Add of existing embedding ID: 160
Add of existing embedding ID: 168
Add of existing embedding ID: 192
Add of existing embedding ID: 198
Add of existing embedding ID: 221
Add of existing embedding ID: 224
Add of existing embedding ID: 247
Add of existing embedding ID: 263
huggingface/tokenizers: The c

[33mresearch_resource_expert[0m (to chat_manager):

As the Topic Expert crafting instructions for an article on "Exploring the Intricacies of Polymer Representation: Unraveling Complexity," aimed at experts in experimental polymer science and machine learning, it's essential to include resources and research methodologies tailored to the needs of this sophisticated readership. Here's how you can structure your guidance for the Editor-in-Chief:

### Key Authoritative Sources:

1. **Prominent Researchers and Papers:**
   - **Ryo Yoshida** – His work on machine learning-assisted exploration of high-conductivity polymers is highly relevant.
   - **Ruimin Ma and Junichiro Shiomi** – Known for their contribution to thermally conductive polymer exploration via high-throughput molecular dynamics simulations.
   - **Evan R. Antoniuk, Peggy Li, Bhavya Kailkhura** – Authors of "Representing Polymers as Periodic Graphs with Learned Descriptors for Accurate Polymer Property Predictions."
   - **C

The return type of the function 'write_section' is not annotated. Although annotating it is optional, the function should return either a string, a subclass of 'pydantic.BaseModel'.


[33meditor_user[0m (to chat_manager):

For the Editor-in-Chief, the article titled "Exploring the Intricacies of Polymer Representation: Unraveling Complexity" requires a comprehensive synthesis to match the expert-level understanding of an audience deeply versed in experimental polymer science and machine learning. To construct this article, the following points must be integrated:

1. **Introduction to Polymer Complexity** - Open with a discussion on the multidimensional nature of polymers that creates a challenge for representation. Assume the reader's familiarity with polymer chemistry and provide a nuanced perspective on why representation is more complex for these substances. (Word Count: 300-400)

2. **Historical Overview of Polymer Representation** - Provide a concise history of polymer representation methods to contextualize modern techniques, recognizing the audience's probable knowledge of these developments. (Word Count: 200-300)

3. **Molecular Descriptors in Polymer Rep

huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Avoid using `tokenizers` before the fork if possible
	- Explicitly set the environment variable TOKENIZERS_PARALLELISM=(true | false)
huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Avoid using `tokenizers` before the fork if possible
	- Explicitly set the environment variable TOKENIZERS_PARALLELISM=(true | false)


Error searching google_scholar: Cannot Fetch from Google Scholar.
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m


huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Avoid using `tokenizers` before the fork if possible
	- Explicitly set the environment variable TOKENIZERS_PARALLELISM=(true | false)
huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Avoid using `tokenizers` before the fork if possible
	- Explicitly set the environment variable TOKENIZERS_PARALLELISM=(true | false)


Error searching google_scholar: Cannot Fetch from Google Scholar.
Error searching google_scholar: Cannot Fetch from Google Scholar.
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m


Could not close webdriver cleanly: Remote end closed connection without response


Error searching google_scholar: Cannot Fetch from Google Scholar.
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m
[92m    Location = ./polymer_representation/0.2/db/uid_text_dict.pkl[0m


huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Avoid using `tokenizers` before the fork if possible
	- Explicitly set the environment variable TOKENIZERS_PARALLELISM=(true | false)
huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Avoid using `tokenizers` before the fork if possible
	- Explicitly set the environment variable TOKENIZERS_PARALLELISM=(true | false)
huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Avoid using `tokenizers` before the fork if possible
	- Explicitly set the environment variable TOKENIZERS_PARALLELISM=(true | false)
huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Av

[31m
>>>>>>>> USING AUTO REPLY...[0m


huggingface/tokenizers: The current process just got forked, after parallelism has already been used. Disabling parallelism to avoid deadlocks...
	- Avoid using `tokenizers` before the fork if possible
	- Explicitly set the environment variable TOKENIZERS_PARALLELISM=(true | false)


[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
Error searching google_scholar: Cannot Fetch from Google Scholar.
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m


Add of existing embedding ID: 277
Insert of existing embedding ID: 277


[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m


Add of existing embedding ID: 278
Insert of existing embedding ID: 278


[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m


Add of existing embedding ID: 284
Insert of existing embedding ID: 284


[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
[31m
>>>>>>>> USING AUTO REPLY...[0m
