In [1]:
from transformers import AutoTokenizer, AutoModelForCausalLM
import torch

In [8]:
import transformers
print(transformers.__version__)
print(torch.__version__)

4.46.3
2.5.1+cu124


In [2]:
def load_model(model_path):
    device = "cuda"
    tokenizer = AutoTokenizer.from_pretrained(model_path, local_files_only=True)
    model = AutoModelForCausalLM.from_pretrained(
    model_path,
    device_map="auto",
    torch_dtype=torch.bfloat16)
    return model, tokenizer

In [3]:
def prompt_model(model, tokenizer, prompt_user, prompt_system=None, verbose=True, max_new_token=128, temperature=0.5, top_p=0.9, do_sample=True):
    messages = []
    if prompt_system:
        messages.append({"role": "user", "content": prompt_system})
    messages.append({"role": "user", "content": prompt_user})
    input_text=tokenizer.apply_chat_template(messages, tokenize=False)
    inputs = tokenizer.encode(input_text, return_tensors="pt").to(model.device)
    outputs = model.generate(inputs, max_new_tokens=max_new_token, temperature=temperature, top_p=top_p, do_sample=do_sample)
    response = tokenizer.decode(outputs[0])
    if verbose:
        print(response)
    return response

### Llama-3.2-3B-Instruct

In [4]:
#model, tokenizer = load_model("/model-weights/Llama-3.2-3B-Instruct")
model, tokenizer = load_model("/h/ws_yuehuan_he/green-ai-explore/models/Llama-3.2-3b-Instruct")

Loading checkpoint shards:   0%|          | 0/2 [00:00<?, ?it/s]

In [5]:
prompt_model(model, tokenizer, "Write a poem about Machine Learning and Large Language Models.", max_new_token=1024)

The attention mask and the pad token id were not set. As a consequence, you may observe unexpected behavior. Please pass your input's `attention_mask` to obtain reliable results.
Setting `pad_token_id` to `eos_token_id`:None for open-end generation.
The attention mask is not set and cannot be inferred from input because pad token is same as eos token. As a consequence, you may observe unexpected behavior. Please pass your input's `attention_mask` to obtain reliable results.


<|begin_of_text|><|begin_of_text|><|start_header_id|>system<|end_header_id|>

Cutting Knowledge Date: December 2023
Today Date: 05 Feb 2025

<|eot_id|><|start_header_id|>user<|end_header_id|>

Write a poem about Machine Learning and Large Language Models.<|eot_id|><|start_header_id|>assistant<|end_header_id|>

In silicon halls, a new mind stirs,
A machine that learns, and adapts, and swears,
To understand the language, complex and grand,
A symphony of data, at its command.

It's trained on vast amounts of text and more,
A never-ending feast, to fuel its core,
The algorithms dance, a waltz of light,
As it absorbs and processes, day and night.

With neural networks, a web of might,
It weaves a tapestry, of insight bright,
Pattern recognition, a skill so fine,
It learns to predict, with uncanny design.

Large language models, a marvel to behold,
Their capabilities, a story to be told,
From chatbots to translation, they roam,
A world of possibilities, in their digital home.

They learn fro

"<|begin_of_text|><|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nCutting Knowledge Date: December 2023\nToday Date: 05 Feb 2025\n\n<|eot_id|><|start_header_id|>user<|end_header_id|>\n\nWrite a poem about Machine Learning and Large Language Models.<|eot_id|><|start_header_id|>assistant<|end_header_id|>\n\nIn silicon halls, a new mind stirs,\nA machine that learns, and adapts, and swears,\nTo understand the language, complex and grand,\nA symphony of data, at its command.\n\nIt's trained on vast amounts of text and more,\nA never-ending feast, to fuel its core,\nThe algorithms dance, a waltz of light,\nAs it absorbs and processes, day and night.\n\nWith neural networks, a web of might,\nIt weaves a tapestry, of insight bright,\nPattern recognition, a skill so fine,\nIt learns to predict, with uncanny design.\n\nLarge language models, a marvel to behold,\nTheir capabilities, a story to be told,\nFrom chatbots to translation, they roam,\nA world of possibilities, in their d

### Google-Gemma-2-2b-it

In [4]:
model, tokenizer = load_model("/h/ws_yuehuan_he/green-ai-explore/models/google-gemma-2-2b-it")

Loading checkpoint shards:   0%|          | 0/2 [00:00<?, ?it/s]

In [5]:
prompt_model(model, tokenizer, "Write a poem about Machine Learning and Large Language Models.", max_new_token=1024)

<bos><bos><start_of_turn>user
Write a poem about Machine Learning and Large Language Models.<end_of_turn>


In the realm of data, a revolution's begun,
Machine learning, a power that's second to none.
Algorithms dance, patterns they discern,
From bits and bytes, a new world to learn.

Large language models, a marvel to behold,
With vast knowledge, stories untold.
From whispers of code, they weave words with grace,
A symphony of language, in a digital space.

They learn from the past, a constant evolution,
With every interaction, their understanding's elevation.
From simple queries, to complex tasks,
They learn and adapt, breaking down all masks.

But with great power, comes a weight to bear,
Bias and misinformation, a silent snare.
Ensuring fairness, a crucial decree,
For the future of knowledge, we must set it free.

So let us embrace, this powerful tool,
With wisdom and care, to make it cool.
For Machine Learning, and Large Language Models,
Hold the key to progress, in a world that e

'<bos><bos><start_of_turn>user\nWrite a poem about Machine Learning and Large Language Models.<end_of_turn>\n\n\nIn the realm of data, a revolution\'s begun,\nMachine learning, a power that\'s second to none.\nAlgorithms dance, patterns they discern,\nFrom bits and bytes, a new world to learn.\n\nLarge language models, a marvel to behold,\nWith vast knowledge, stories untold.\nFrom whispers of code, they weave words with grace,\nA symphony of language, in a digital space.\n\nThey learn from the past, a constant evolution,\nWith every interaction, their understanding\'s elevation.\nFrom simple queries, to complex tasks,\nThey learn and adapt, breaking down all masks.\n\nBut with great power, comes a weight to bear,\nBias and misinformation, a silent snare.\nEnsuring fairness, a crucial decree,\nFor the future of knowledge, we must set it free.\n\nSo let us embrace, this powerful tool,\nWith wisdom and care, to make it cool.\nFor Machine Learning, and Large Language Models,\nHold the key

### SmolLM2-1.7B-Instruct

In [6]:
model, tokenizer = load_model("/h/ws_yuehuan_he/green-ai-explore/models/SmolLM2-1.7B-Instruct")

In [7]:
prompt_model(model, tokenizer, "Write a poem about Machine Learning and Large Language Models.", max_new_token=1024)

The attention mask is not set and cannot be inferred from input because pad token is same as eos token. As a consequence, you may observe unexpected behavior. Please pass your input's `attention_mask` to obtain reliable results.


<|im_start|>system
You are a helpful AI assistant named SmolLM, trained by Hugging Face<|im_end|>
<|im_start|>user
Write a poem about Machine Learning and Large Language Models.<|im_end|>
<|im_start|>assistant
In the realm of data, vast and wide,
Lies a tale of Machine Learning's pride.
A journey of code, of algorithms true,
In the heart of the digital, it does grow.

A spark of intelligence, a spark of might,
In the vast expanse of the digital night.
A symphony of patterns, a dance so fine,
In the realm of Large Language Models, it shines.

The language of machines, in code it's born,
A symphony of words, a song to adorn.
With every byte, with every line,
In the heart of the digital, it does shine.

Large Language Models, they learn and grow,
In the realm of data, they do know.
A symphony of words, a dance so fine,
In the realm of Machine Learning, they shine.

They learn from the vast, from the deep,
From the data that's so hard to keep.
They learn from the noise, from the glitch,
In

"<|im_start|>system\nYou are a helpful AI assistant named SmolLM, trained by Hugging Face<|im_end|>\n<|im_start|>user\nWrite a poem about Machine Learning and Large Language Models.<|im_end|>\n<|im_start|>assistant\nIn the realm of data, vast and wide,\nLies a tale of Machine Learning's pride.\nA journey of code, of algorithms true,\nIn the heart of the digital, it does grow.\n\nA spark of intelligence, a spark of might,\nIn the vast expanse of the digital night.\nA symphony of patterns, a dance so fine,\nIn the realm of Large Language Models, it shines.\n\nThe language of machines, in code it's born,\nA symphony of words, a song to adorn.\nWith every byte, with every line,\nIn the heart of the digital, it does shine.\n\nLarge Language Models, they learn and grow,\nIn the realm of data, they do know.\nA symphony of words, a dance so fine,\nIn the realm of Machine Learning, they shine.\n\nThey learn from the vast, from the deep,\nFrom the data that's so hard to keep.\nThey learn from th

### Qwen2.5-1.5B-Instruct

In [7]:
model, tokenizer = load_model("/model-weights/Qwen2.5-1.5B-Instruct")

In [8]:
prompt_model(model, tokenizer, "Write a poem about Machine Learning and Large Language Models.", max_new_token=1024)

<|im_start|>system
You are Qwen, created by Alibaba Cloud. You are a helpful assistant.<|im_end|>
<|im_start|>user
Write a poem about Machine Learning and Large Language Models.<|im_end|>
<|im_start|>subject
Machine learning, a tool so vast,
Transforms data into knowledge it's capable.
From text to images, from speech to code,
It learns with each passing day.

Large language models, they too have their role,
Understanding complex queries, without any doubt.
With every update, they grow in might,
Learning new languages, words anew.

But like all tools, they can be used for good or harm,
So let us use them wisely, always our goals.
To solve problems, to help others, to guide the way,
In this world of technology, where machines do sway.

For as long as we live, these models will evolve,
Helping humanity, towards progress and growth.
Together, let's make sure that AI is kind,
And serves only what is right and just.<|im_end|>


"<|im_start|>system\nYou are Qwen, created by Alibaba Cloud. You are a helpful assistant.<|im_end|>\n<|im_start|>user\nWrite a poem about Machine Learning and Large Language Models.<|im_end|>\n<|im_start|>subject\nMachine learning, a tool so vast,\nTransforms data into knowledge it's capable.\nFrom text to images, from speech to code,\nIt learns with each passing day.\n\nLarge language models, they too have their role,\nUnderstanding complex queries, without any doubt.\nWith every update, they grow in might,\nLearning new languages, words anew.\n\nBut like all tools, they can be used for good or harm,\nSo let us use them wisely, always our goals.\nTo solve problems, to help others, to guide the way,\nIn this world of technology, where machines do sway.\n\nFor as long as we live, these models will evolve,\nHelping humanity, towards progress and growth.\nTogether, let's make sure that AI is kind,\nAnd serves only what is right and just.<|im_end|>"

### DeepSeek-R1-Distill-Qwen-1.5B

In [4]:
model, tokenizer = load_model("/model-weights/DeepSeek-R1-Distill-Qwen-1.5B")

In [6]:
prompt_model(model, tokenizer, "Write a poem about Machine Learning and Large Language Models.<|Assistant|>", max_new_token=1024)

The attention mask and the pad token id were not set. As a consequence, you may observe unexpected behavior. Please pass your input's `attention_mask` to obtain reliable results.
Setting `pad_token_id` to `eos_token_id`:151643 for open-end generation.


<｜begin▁of▁sentence｜><｜begin▁of▁sentence｜><｜User｜>Write a poem about Machine Learning and Large Language Models.<|Assistant|> I need to write a poem about Machine Learning and Large Language Models. I'm a bit new to this, so I need to figure out how to approach it. Maybe I can start by thinking about what ML and LLMs are, and then find metaphors or imagery that relate to them. I remember that LLMs are a big thing now, so maybe I can use some metaphors about technology, innovation, or the process of learning. I should also consider the emotions involved, like the excitement or the challenge of working with these models. Maybe I can structure the poem with a rhyme scheme, perhaps something like ABAB or AABB, and include imagery related to data, patterns, predictions, and the unseen universe or something like that. I'm not entirely sure, but I think I can get started by brainstorming some ideas and then organizing them into a poem.
</think>

**The Machine Learning Machine**

In the vast e

"<｜begin▁of▁sentence｜><｜begin▁of▁sentence｜><｜User｜>Write a poem about Machine Learning and Large Language Models.<|Assistant|> I need to write a poem about Machine Learning and Large Language Models. I'm a bit new to this, so I need to figure out how to approach it. Maybe I can start by thinking about what ML and LLMs are, and then find metaphors or imagery that relate to them. I remember that LLMs are a big thing now, so maybe I can use some metaphors about technology, innovation, or the process of learning. I should also consider the emotions involved, like the excitement or the challenge of working with these models. Maybe I can structure the poem with a rhyme scheme, perhaps something like ABAB or AABB, and include imagery related to data, patterns, predictions, and the unseen universe or something like that. I'm not entirely sure, but I think I can get started by brainstorming some ideas and then organizing them into a poem.\n</think>\n\n**The Machine Learning Machine**\n\nIn the 