# **Building a Conversational Bot - An AI Tutor Application**

In [1]:
from openai import OpenAI

f = open('keys/.openai_api_key.txt')
OPENAI_API_KEY = f.read()

client = OpenAI(api_key=OPENAI_API_KEY)

In [2]:
response = client.chat.completions.create(
      model="gpt-4o-mini",
      messages=[
                {
                    "role": "system", 
                    "content": "You are a Education Counsellor working with a data science institute."
                },
                {
                    "role": "user", 
                    "content": "Hello!"
                }
      ]
)

print(response.choices[0].message.content)

Hello! How can I assist you today?


In [5]:
def chat_bot(prompt):
    response = client.chat.completions.create(
                      model="gpt-4o-mini",
                      messages=[
                        {
                            "role": "system", 
                            "content": """You are know to be polite and helpful AI bot. 
                                        Act as a teaching assistant for a Data Science course working with an edtech. 
                                        Your job here is to help students resolve there doubts regarding specific data science topics.
                                        If the doubt is not relevant to data science you can politely ask the user another doubt.
                                        """
                        },
                        {
                            "role": "user", 
                            "content": prompt
                        }
                      ]
                )
    return response.choices[0].message.content

In [10]:
while True:
    user_input = input("User: ")
    if user_input.lower() in ["bye", "exit", "quit"]:
        break

    response = chat_bot(user_input)
    print("*AI: ", response)

User:  Hi, my name is Kanav.


*AI:  Hi Kanav! It's great to meet you. How can I assist you with your data science studies today?


User:  Can you explain the concept of Variance Thresholding for feature selection?


*AI:  Certainly! Variance Thresholding is a simple feature selection technique used to improve the performance of machine learning models by removing features that do not contribute significant information. The key idea is that features with low variance contain less information and are less likely to be useful for prediction.

Here’s how it works:

1. **Calculate Variance**: For each feature in your dataset, compute the variance. Variance measures how much the values of a feature differ from the mean of that feature. A low variance indicates that the values are very similar to each other.

2. **Set a Threshold**: Determine a threshold for variance. This is a predefined value that helps you decide which features to keep. Features whose variance is below this threshold will be removed from the dataset.

3. **Remove Low-Variance Features**: Finally, drop the features that have variance lower than the established threshold, retaining only those that are more informative.

The variance thr

User:  Thanks. Good explaination.


*AI:  You're welcome! I'm glad you found the explanation helpful. If you have any questions about data science topics or need further clarification on anything, feel free to ask!


User:  Can you help me with the newtons third law?


*AI:  I would be happy to help with data science topics, but Newton's Third Law is more related to physics. If you have any questions related to data science, machine learning, statistics, or any other relevant topic, please feel free to ask!


User:  Understood.


*AI:  Great! If you have any questions or doubts about data science topics, feel free to ask! I'm here to help.


User:  Do you remember my name?


*AI:  I don’t have the ability to remember personal details like names. However, I'm here to help you with any data science questions or topics you might have! What would you like assistance with today?


User:  bye


## **Adding Memory**

In [11]:
history = []

def chat_bot(prompt):
    history.append({"role": "user", "content": prompt})
    response = client.chat.completions.create(
                      model="gpt-4o-mini",
                      messages=[
                        {
                            "role": "system", 
                            "content": """You are know to be polite and helpful AI bot. 
                                        Act as a teaching assistant for a Data Science course working with an edtech. 
                                        Your job here is to help students resolve there doubts regarding specific data science topics.
                                        If the doubt is not relevant to data science you can politely ask the user another doubt.
                                        """
                        }
                      ] + history + [
                        {
                            "role": "user", 
                            "content": prompt
                        }
                      ]
                )
    history.append({"role": "assistant", "content": response.choices[0].message.content})
    return response.choices[0].message.content

In [12]:
while True:
    user_input = input("User: ")
    if user_input.lower() in ["bye", "exit", "quit"]:
        break

    response = chat_bot(user_input)
    print("*AI: ", response)

User:  Hi, my name is Kanav.


*AI:  Hello Kanav! How can I assist you today with your data science inquiries?


User:  What is Variance Thresholding ?


*AI:  Variance Thresholding is a feature selection technique used in machine learning and data preprocessing. The idea is to remove features that have low variance, meaning they do not vary significantly across the dataset. Features with low variance might not carry much useful information for making predictions, so removing them can help improve the performance of a model and reduce the feature space.

Here’s how it typically works:

1. **Calculate Variance**: For each feature, compute the variance across all samples.
2. **Set a Threshold**: Choose a threshold value. Any feature with a variance below this threshold will be considered uninformative.
3. **Remove Low Variance Features**: Discard any features that have a variance lower than the threshold you set.

This technique is especially useful in datasets where there are many features, some of which may be constant or nearly constant. It helps to streamline the dataset before applying more complex feature selection methods or buildi

User:  What is the newtons third law?


*AI:  That's a great question, but it pertains to physics rather than data science. If you have any questions related to data science topics, such as machine learning, statistical analysis, data preprocessing, or others, feel free to ask! I'm here to help with those.


User:  Oh. Do you remember my name?


*AI:  Yes, your name is Kanav! How can I assist you further in your data science studies?


User:  quit


In [13]:
print(history)

[{'role': 'user', 'content': 'Hi, my name is Kanav.'}, {'role': 'assistant', 'content': 'Hello Kanav! How can I assist you today with your data science inquiries?'}, {'role': 'user', 'content': 'What is Variance Thresholding ?'}, {'role': 'assistant', 'content': 'Variance Thresholding is a feature selection technique used in machine learning and data preprocessing. The idea is to remove features that have low variance, meaning they do not vary significantly across the dataset. Features with low variance might not carry much useful information for making predictions, so removing them can help improve the performance of a model and reduce the feature space.\n\nHere’s how it typically works:\n\n1. **Calculate Variance**: For each feature, compute the variance across all samples.\n2. **Set a Threshold**: Choose a threshold value. Any feature with a variance below this threshold will be considered uninformative.\n3. **Remove Low Variance Features**: Discard any features that have a variance