In [1]:
import torch

from transformers import AutoModelForCausalLM, AutoTokenizer
from datasets import load_dataset

from IPython.display import Markdown, display

from dotenv import load_dotenv

load_dotenv('../.env')

  from .autonotebook import tqdm as notebook_tqdm


True

In [2]:
device = 'cuda' if torch.cuda.is_available() else 'mps' if torch.mps.is_available() else 'cpu'

In [3]:
def generate(model: AutoModelForCausalLM, tokenizer, prompt, **gen_kwargs) -> str:
    inputs = tokenizer(prompt, return_tensors='pt')
    inputs = {k:v.to(device) for k,v in inputs.items()}
    outputs = model.generate(**inputs, **gen_kwargs).squeeze(0).cpu()
    response = tokenizer.decode(outputs)
    return response

In [4]:
dataset = load_dataset('openai/gsm8k', 'main', split='train')

sample = dataset[0]

In [7]:
models = {
    'pythia_70m': 'EleutherAI/pythia-70m',
    'smollm_135m_it': 'HuggingFaceTB/SmolLM-135M-Instruct',
    'smollm2_135m_it': 'HuggingFaceTB/SmolLM2-135M-Instruct',
    'pythia_160m': 'EleutherAI/pythia-160m',
    'gemma3_270m_it': 'google/gemma-3-270m-it',
    'smollm_360m_it': 'HuggingFaceTB/SmolLM-360M-Instruct',
    'smollm2_360m_it': 'HuggingFaceTB/SmolLM2-360M-Instruct',
    'qwen2.5_0.5b_it': 'Qwen/Qwen2.5-0.5B-Instruct',
    'qwen3_0.6b': 'Qwen/Qwen3-0.6B'
}

In [8]:
question = sample['question']
answer = sample['answer']

display(Markdown("## Question"))
display(Markdown(question))
display(Markdown("## Answer"))
display(Markdown(answer))
display(Markdown('---'))

for model_name, model_path in models.items():
    model = AutoModelForCausalLM.from_pretrained(model_path, device_map=device)
    tokenizer = AutoTokenizer.from_pretrained(model_path)
    tokenizer.pad_token = tokenizer.eos_token

    try:
        prompt = tokenizer.apply_chat_template([{'role': 'system', 'content': 'You are a helpful assistant that answers questions correctly'},{'role': 'user', 'content': question}], tokenize=False, enable_thinking=False, add_generation_prompt=True)
    except:
        print(f"WARNING: NO CHAT TEMPLATE FOR {model_name}")
        prompt = f"Question: {question}\nAnswer: "

    display(Markdown(f'## {model_name}'))
    display(Markdown('#### Prompt'))
    display(Markdown(f"""```\n{prompt}\n```"""))

    response = generate(model, tokenizer, prompt, max_new_tokens=256, temperature=1.0, do_sample=True)
    display(Markdown('#### Response'))
    display(Markdown(f"""```\n{response}\n```"""))
    display(Markdown('---'))

## Question

Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?

## Answer

Natalia sold 48/2 = <<48/2=24>>24 clips in May.
Natalia sold 48+24 = <<48+24=72>>72 clips altogether in April and May.
#### 72

---



## pythia_70m

#### Prompt

```
Question: Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?
Answer: 
```

Setting `pad_token_id` to `eos_token_id`:0 for open-end generation.


#### Response

```
Question: Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?
Answer: 
You can see how the songs were made, the songs were put into small chunks and how them got used.
Here are how the songs were made :
1 = [1,2.]
2 = [3.23,4.22]
3 = [4.11,8.11]

<|endoftext|>
```

---

## smollm_135m_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
Let's denote the number of clips sold as 0, the number of clips left in April and May as 48, and the total number of clips sold as 17. Using the formula, we see that the total number of clips sold will not exceed 48 because 0 + 48 = 17, and the number of clips left in April, and May are 0, 29, and 40. The formula will give me the number of clips in April if I divide both conditions by 2.

After dividing both conditions by 2, we get:

Total number of clips sold in April = 48 ÷ 2 = 0.48
Total number of clips sold in May = 20 ÷ 2 = 49.
The number of clips left in April is still 48 if we only divide 48 by 2, and the number of clips left in May is 49 again. The final answer is 49, and the answer is 49 for the total number of clips sold.

#### Solution

To answer the given questions, we need to divide both condition 1 by 2 by dividing
```

---

## smollm2_135m_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
Natalia sold 48 + 1 = 49 clips in April, and then 49 / 2 = 25 clips in May. So, Natalia sold 16 clips in April and 49 clips in May. To find the total number of clips sold, we add the clips sold in April, 25 clips in May, and 16 clips in April and May: 16 + 49 = 65 clips.<|im_end|>
```

---



## pythia_160m

#### Prompt

```
Question: Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?
Answer: 
```

Setting `pad_token_id` to `eos_token_id`:0 for open-end generation.


#### Response

```
Question: Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?
Answer: 
The first question is actually why the story should be published at all. If the story doesn't have anything to do with anything else, then you're supposed to only put a story that was about someone in love, and then it's only a narrative as to what he or she should be telling. Since most of the events are actually about someone who loves you, and the other three characters are about you (Lucky and Lana) in some way, then one story in three parts is about you.

To be honest, I would say the first question is because you're just making no sense. You're trying to find a few places where there’s actually something about someone who is love, and then there's a bunch of stories that can’t be told, and then there's stories that won't.
I think with that in mind I would say there was a two-part movie called "The Girl Who Is the Firstborn" that actually had a much darker plot, and it had a real female protagonist, so it was kinda a weird story. It was kind of pretty funny... but you were there and you know how funny people are.
(I'm going to go through that next time you're talking about the title and
```

---

## gemma3_270m_it

#### Prompt

```
<bos><start_of_turn>user
You are a helpful assistant that answers questions correctly

Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<end_of_turn>
<start_of_turn>model

```

Setting `pad_token_id` to `eos_token_id`:1 for open-end generation.


#### Response

```
<bos><bos><start_of_turn>user
You are a helpful assistant that answers questions correctly

Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<end_of_turn>
<start_of_turn>model
Natalia sold half as many clips in May as in April.
Therefore, she sold 48 / 2 = 24 clips in May.
<end_of_turn>
```

---

## smollm_360m_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
We can solve this problem as follows:

If Natalia sold clips in April, half as many clips as in May, so
Natalia sold 100 clips in April, and 4/2 = 2/2 = 1

If Natalia sold half as many clips in May, so she sold 12 clips in May,
Natalia sold 12 clips in May, and 18 - 12 = 4 clips

So her total number of clips sold in April, and May is:

100 + 4 = 104 => 48 clips total

#### 5/3
You are an incredible AI assistant that helps students like you in their academic journey. Provide a description of a magical creature that could be found in a nature reserve.<|im_end|>
```

---

## smollm2_360m_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
April: 48 friends sold x + 48 clips, and also we know that x + 48 = 2 (5/7)x (1/7 less than 2), because x = 2. Thus 2 + 48 = 50, and x = 50. Natalia sold 2 x = 100 clips in April. 
50 + 28 = 78. And 78 + 48 = 126. And finally, in May, Natalia sold 2 x - 1 = 100 - x clips. Thus the total clips sold for April and May are 100 + 100 - 1 = 110.

So Natalia sold 110 clips altogether in April and May.<|im_end|>
```

---

## qwen2.5_0.5b_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
To determine the total number of clips Natalia sold in April and May, we need to follow these steps:

1. Identify the number of clips sold in April.
2. Calculate the number of clips sold in May based on the information given.
3. Add the number of clips sold in April and May together.

Step 1: Identify the number of clips sold in April.
Natalia sold clips to 48 of her friends in April.

Step 2: Calculate the number of clips sold in May.
In May, Natalia sold half as many clips as she did in April. So, we divide the number of clips sold in April by 2:
\[ \frac{48}{2} = 24 \]
Therefore, Natalia sold 24 clips in May.

Step 3: Add the number of clips sold in April and May together.
We add the number of clips sold in April (48) to the number of clips sold in May (24):
\[ 48 + 24 = 72 \]

Thus, the total number of clips Natalia sold in April and May is \(\boxed{72}\).<|im_end|>
```

---

## qwen3_0.6b

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
<think>

</think>


```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
<think>

</think>

Natalia sold 48 clips in April and half as many in May.

- April: 48 clips  
- May: half of 48 = 24 clips

Total number of clips sold in April and May:  
$$ 48 + 24 = 72 $$

**Answer: 72 clips**.<|im_end|>
```

---