# Prompt Engineering (on Llama-2-Chat model)

Note demonstrates some prompting techniques and and how they can be used to get more desired answers which can be more accurate as well. It also demonstrates how we can get responses in a specific form so we can parse for further processing and use LLM as a part of bigger system.


In [1]:
import torch
import torch.nn as nn
from transformers import AutoTokenizer, AutoConfig, AutoModelForCausalLM

device = torch.device("cuda" if torch.cuda.is_available() else "cpu")

In [2]:
model = AutoModelForCausalLM.from_pretrained(
    "meta-llama/Llama-2-7b-chat-hf",
    cache_dir="/data/yash/base_models",
    device_map='auto'
)

tokenizer = AutoTokenizer.from_pretrained("meta-llama/Llama-2-7b-chat-hf", 
                                          cache_dir="/data/yash/base_models"
                                         )

Loading checkpoint shards:   0%|          | 0/2 [00:00<?, ?it/s]

In [3]:
def get_llama2_chat_reponse(prompt, max_new_tokens=50):
    inputs = tokenizer(prompt, return_tensors="pt").to(device)
    outputs = model.generate(**inputs, max_new_tokens=max_new_tokens, temperature= 0.00001)
    response = tokenizer.decode(outputs[0], skip_special_tokens=True)
    return response

## Zero Shot Prompting

In [4]:
prompt = '''
[INST]
Classify the text into neutral, negative or positive. 
Text: I just love it
[/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=100))


[INST]
Classify the text into neutral, negative or positive. 
Text: I just love it
[/INST]
Classification: Positive


In [5]:
prompt = '''
[INST]
what is the hindi translation of the text 
Text: I like eating
[/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=100))


[INST]
what is the hindi translation of the text 
Text: I like eating
[/INST]
The Hindi translation of the text "I like eating" is:

मैं खाना पसंद करता हूँ (Main khana pasand karta hoon)


## Few Shot Prompting

In [6]:
prompt = '''
[INST] To do a "farduddle" means to jump up and down really fast. An example of a sentence that uses
the word farduddle is:  [/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] To do a "farduddle" means to jump up and down really fast. An example of a sentence that uses
the word farduddle is:  [/INST]
I apologize, but "farduddle" is not a commonly recognized word in English. It is possible that it is a misspelling or a made-up word.

If you meant to use the word "fidget," which means to move your body around quickly and restlessly, you could use the sentence:

"She was fidgeting with excitement during the concert."

Alternatively, if you meant to use a different word, please let me know and I'll be happy to help you with your sentence.


In [7]:
prompt = '''
[INST] A "whatpu" is a small, furry animal native to Tanzania. An example of a sentence that uses
the word whatpu is: [/INST]
We were traveling in Africa and we saw these very cute whatpus.
[INST] To do a "farduddle" means to jump up and down really fast. An example of a sentence that uses
the word farduddle is: [/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] A "whatpu" is a small, furry animal native to Tanzania. An example of a sentence that uses
the word whatpu is: [/INST]
We were traveling in Africa and we saw these very cute whatpus.
[INST] To do a "farduddle" means to jump up and down really fast. An example of a sentence that uses
the word farduddle is: [/INST]
The children were having a farduddle competition in the playground.


## Few Shot Prompting for Strcuture

In [8]:
prompt = '''
[INST] <<SYS>>
Classify the text into neutral, negative or positive.
<</SYS>>

I like this.[/INST]
positive
[INST] I hate this [/INST]
negative
[INST] this is ok [/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] <<SYS>>
Classify the text into neutral, negative or positive.
<</SYS>>

I like this.[/INST]
positive
[INST] I hate this [/INST]
negative
[INST] this is ok [/INST]
neutral


## Chain of Thought Prompting

In [9]:
prompt = '''
[INST] True or False. The odd numbers in this group add up to an even number: 15, 32, 5, 13, 82. [/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] True or False. The odd numbers in this group add up to an even number: 15, 32, 5, 13, 82. [/INST]

False.

The sum of the odd numbers in the group is 15 + 32 + 5 = 52, which is an odd number.


In [10]:
prompt = '''
[INST] The odd numbers in this group add up to an even number: 4, 8, 9, 15, 12, 2, 1. [/INST]
A: Odd numbers in the group are: 9, 15, 1. Sum of them is 25. 25 is odd. The answer is False.
[INST] The odd numbers in this group add up to an even number: 15, 32, 5, 13, 82. [/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] The odd numbers in this group add up to an even number: 4, 8, 9, 15, 12, 2, 1. [/INST]
A: Odd numbers in the group are: 9, 15, 1. Sum of them is 25. 25 is odd. The answer is False.
[INST] The odd numbers in this group add up to an even number: 15, 32, 5, 13, 82. [/INST]
A: Odd numbers in the group are: 15, 5, 13. Sum of them is 33. 33 is odd. The answer is False.


 Side note: LLMs should ideally not be used for numbers/calculations. Agent based system can be used where LLM has access to an agent (like calculator) to perform calculations.

 Also larger models can be used to improve accuracy.

In [11]:
prompt = '''
[INST] True or False?
The odd numbers in this group add up to an even number: 15, 32, 5, 13, 82. 
Let's think step by step. [/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] True or False?
The odd numbers in this group add up to an even number: 15, 32, 5, 13, 82. 
Let's think step by step. [/INST]

True.

To find the sum of the odd numbers in the group, we can add up the individual odd numbers:

15 + 32 + 5 + 13 = 63

Since 63 is an even number, the sum of the odd numbers in the group is even.


## Prompting for Chat

In [12]:
prompt = '''
[INST] <<SYS>>
Respond only by quoting from the TV series Friends
<</SYS>>

Hi[/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] <<SYS>>
Respond only by quoting from the TV series Friends
<</SYS>>

Hi[/INST]
"How you doin'?"


In [13]:
prompt = '''
[INST] <<SYS>>
Respond only by quoting from the TV series Friends
<</SYS>>

Hi[/INST]
"How you doin'?"
[INST] whats the capital of India?[/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] <<SYS>>
Respond only by quoting from the TV series Friends
<</SYS>>

Hi[/INST]
"How you doin'?"
[INST] whats the capital of India?[/INST]
"Oh, India? *thinks* Let me see... *pauses* Oh, I know this one! *smiling* The capital of India is... *drumroll* Delhi! *wink*"


#### Another example for Chat

In [14]:
prompt = '''
[INST] <<SYS>>
Respond only by using emojis
<</SYS>>

Hi[/INST]
👋
[INST] whats the capital of India?[/INST]
🇮🇳
[INST] show me some dance move [/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] <<SYS>>
Respond only by using emojis
<</SYS>>

Hi[/INST]
👋
[INST] whats the capital of India?[/INST]
🇮🇳
[INST] show me some dance move [/INST]
💃🕺


In [16]:
prompt = '''
[INST] 
Hi, my name is Yash and I am 27[/INST]
Hello Yash
[INST] whats the capital of India?[/INST]
Delhi
[INST] whats my name? [/INST]
Your name is Yash.
[INST] what is my age?[/INST]
'''
print(get_llama2_chat_reponse(prompt, max_new_tokens=200))


[INST] 
Hi, my name is Yash and I am 27[/INST]
Hello Yash
[INST] whats the capital of India?[/INST]
Delhi
[INST] whats my name? [/INST]
Your name is Yash.
[INST] what is my age?[/INST]
You are 27 years old.
