In [1]:
import torch

from transformers import AutoModelForCausalLM, AutoTokenizer
from datasets import load_dataset

from IPython.display import Markdown, display

from dotenv import load_dotenv

load_dotenv('../.env')

  from .autonotebook import tqdm as notebook_tqdm


True

In [2]:
device = 'cuda' if torch.cuda.is_available() else 'mps' if torch.mps.is_available() else 'cpu'

In [3]:
def generate(model: AutoModelForCausalLM, tokenizer, prompt, **gen_kwargs) -> str:
    inputs = tokenizer(prompt, return_tensors='pt')
    inputs = {k:v.to(device) for k,v in inputs.items()}
    outputs = model.generate(**inputs, **gen_kwargs).squeeze(0).cpu()
    response = tokenizer.decode(outputs)
    return response

In [4]:
dataset = load_dataset('openai/gsm8k', 'main', split='train')

sample = dataset[0]

In [5]:
models = {
    'pythia_70m': 'EleutherAI/pythia-70m',
    'smollm_135m_it': 'HuggingFaceTB/SmolLM-135M-Instruct',
    'smollm2_135m_it': 'HuggingFaceTB/SmolLM2-135M-Instruct',
    'pythia_160m': 'EleutherAI/pythia-160m',
    'gemma3_270m_it': 'google/gemma-3-270m-it',
    'smollm_360m_it': 'HuggingFaceTB/SmolLM-360M-Instruct',
    'smollm2_360m_it': 'HuggingFaceTB/SmolLM2-360M-Instruct',
    'qwen2.5_0.5b_it': 'Qwen/Qwen2.5-0.5B-Instruct',
    'qwen3_0.6b': 'Qwen/Qwen3-0.6B'
}

In [6]:
question = sample['question']
answer = sample['answer']

display(Markdown("## Question"))
display(Markdown(question))
display(Markdown("## Answer"))
display(Markdown(answer))
display(Markdown('---'))

for model_name, model_path in models.items():
    model = AutoModelForCausalLM.from_pretrained(model_path, device_map=device)
    tokenizer = AutoTokenizer.from_pretrained(model_path)
    tokenizer.pad_token = tokenizer.eos_token

    try:
        prompt = tokenizer.apply_chat_template([{'role': 'system', 'content': 'You are a helpful assistant that answers questions correctly'},{'role': 'user', 'content': question}], tokenize=False, enable_thinking=False, add_generation_prompt=True)
    except:
        print(f"WARNING: NO CHAT TEMPLATE FOR {model_name}")
        prompt = f"Question: {question}\nAnswer: "

    display(Markdown(f'## {model_name}'))
    display(Markdown('#### Prompt'))
    display(Markdown(f"""```\n{prompt}\n```"""))

    response = generate(model, tokenizer, prompt, max_new_tokens=256, temperature=1.0, do_sample=True)
    display(Markdown('#### Response'))
    display(Markdown(f"""```\n{response}\n```"""))
    display(Markdown('---'))

## Question

Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?

## Answer

Natalia sold 48/2 = <<48/2=24>>24 clips in May.
Natalia sold 48+24 = <<48+24=72>>72 clips altogether in April and May.
#### 72

---



## pythia_70m

#### Prompt

```
Question: Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?
Answer: 
```

Setting `pad_token_id` to `eos_token_id`:0 for open-end generation.


#### Response

```
Question: Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?
Answer: 
You should consult us and we'll tell you a bit about one about what things we're interested in about our partners. And a little bit about what we'll do by contacting our partners if applicable.
More

A:

From the comments above, the following example should provide an explanation with regard to the use of audio tape using a tape measure (10 Hz, 1.8 Hz, 1.9 Hz, and 2.9 Hz), "D" with "D" showing how the tape measure is applied.
I'll state that I didn't use a tape measure to read the tape, but I will also assume my tape measure is not appropriate when trying to write my tape.
If I were to put three clips on audio tape, they would be as follows. First of all, the length of each one in the tape will decrease with each clip, so that they will get the last one, which can be modified with the width equal to the size of the tape it will put on the tape. Then, the angle of each tape, which the tape measure is, varies, and then varies according to the lengths of the first and second tape. Thus, each clip would change one of the lengths of the tape during the first clip
```

---

## smollm_135m_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
We can use the given data to calculate the total number of clips sold in April and May. Let $n$ be this number of clips sold.

* During April: The total number of clips sold is $2n + 1$
* During May: The total number of clips sold is $2n - 1$
* Total number of clips sold in April and May: $2n + 1$ = $2n - 1$ and $2n - 1 - 1 = 2n - 1 = 2n - 1 = 3*n - 1 = 3n - 1$ Let $x$ be the number of clips sold in April.
* Total number of clips sold in May: Total number of clips sold in April and May is $2n*(3n-1) = 2n(3n-1) - 1 = 2n*(3n-1) - 1 = 6*n - 2 = 6*n - 3$
* Total number of clips sold in April and May: $y$ = 100(2x + 2x^2 - 3x)

```

---

## smollm2_135m_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
Natalia sold 48 clips.
So, she sold 48/2 = 24 clips in the first month, and then 24/4 = 6 clips in the same month.
In April, Natalia sold 24/4 = 6 clips.
In May, Natalia sold 6/4 = 1 clip.
#### 1
The answer is: 1<|im_end|>
```

---



## pythia_160m

#### Prompt

```
Question: Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?
Answer: 
```

Setting `pad_token_id` to `eos_token_id`:0 for open-end generation.


#### Response

```
Question: Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?
Answer: ied
"This is exactly what was discussed on April 16, 2009 during the hearing on behalf of the defendant, Natalia Williams, at an April 21, 1998 hearing in Cincinnati. The defense counsel explained to the jury at the time of trial that the first half of September in June 1979, when the defendant and her friend were working at the B&P Company for thirty miles around Cincinnati, Ohio, the girl could bring a few clips to her car and then drive to her mother's house in Cincinnati. The defendant testified that at one point her friends told her that she could bring clips because people were there all week. Defendant said, "That was about as cool as they can get. I mean, we all had to work a little bit" in the car. She then continued to use a car in their kitchen, where she kept a box and an envelope. During the first six months after the trial, she and her friends told her that he kept videos and she found them very interesting. She gave them to a friend, and she got a few clips in the freezer to keep her friends with. The rest of the time, just as it was done at the prison, she was in the living room and playing with her friends. She watched the videos and she
```

---

## gemma3_270m_it

#### Prompt

```
<bos><start_of_turn>user
You are a helpful assistant that answers questions correctly

Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<end_of_turn>
<start_of_turn>model

```

Setting `pad_token_id` to `eos_token_id`:1 for open-end generation.


#### Response

```
<bos><bos><start_of_turn>user
You are a helpful assistant that answers questions correctly

Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<end_of_turn>
<start_of_turn>model
Natalia sold half as many clips in May as in April.
Therefore, she sold 48 / 2 = 24 clips in May.
<end_of_turn>
```

---

## smollm_360m_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
Let's solve this problem:

48 clips sold in April
(150 + 25 * 48) = ?
52 clips sold in May
(150 + 25 * 48) = ?
52 clips sold in May
128 clips sold in July
(150 + 25 * 48) = ?
128 clips sold in July
So, Natalia sold 128 clips in May, 128 clips in July, and 128 clips in August.
Here's the correct answer:
128 clips sold in May
138 clips sold in July
170 clips sold in August
#### 245-128

245 is the answer.<|im_end|>
```

---

## smollm2_360m_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
April clips sold: 48
April clips sold: 48/2 = 24
May clips sold: 24 + 4 = 28

So, Natalia sold a total of 28 + 24 = 52 clips in April and May.
The answer is 52.<|im_end|>
```

---

## qwen2.5_0.5b_it

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant

```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
In April, Natalia sold 48/2 = 24 clips.
Thus, Natalia sold 48 + 24 = 72 clips in April and May.

Therefore, the answer is 72.<|im_end|>
```

---

## qwen3_0.6b

#### Prompt

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
<think>

</think>


```

#### Response

```
<|im_start|>system
You are a helpful assistant that answers questions correctly<|im_end|>
<|im_start|>user
Natalia sold clips to 48 of her friends in April, and then she sold half as many clips in May. How many clips did Natalia sell altogether in April and May?<|im_end|>
<|im_start|>assistant
<think>

</think>

Natalia sold **48 clips in April** and **half as many in May** as 48. Half of 48 is 24, so she sold **24 clips in May**.

Now, we add the number of clips sold in April and May:  
**48 + 24 = 72**.

So, Natalia sold **72 clips in total in April and May**.<|im_end|>
```

---