<table class="tfo-notebook-buttons" align="left">
  <td>
    <a target="_blank" href="https://colab.research.google.com/github/milmor/NLP/blob/main/Notebooks/22_Q_DeepSeek-8B_hf.ipynb">
    <img src="https://www.tensorflow.org/images/colab_logo_32px.png" />
    Run in Google Colab</a>
  </td>
</table>

# DeepSeek-8B-R1 Quantization

In [1]:
from transformers import AutoTokenizer, AutoModelForCausalLM
from transformers import AutoTokenizer
from transformers import pipeline
import time
import torch

In [2]:
from transformers import BitsAndBytesConfig

nf4_config = BitsAndBytesConfig(
    load_in_4bit=True,
    bnb_4bit_quant_type="nf4",
    bnb_4bit_use_double_quant=True,
    bnb_4bit_compute_dtype=torch.bfloat16
)

In [3]:
model_id = "deepseek-ai/DeepSeek-R1-Distill-Llama-8B"
number_gpus = 1
max_model_len = 8192
max_new_tokens = 256

In [4]:
model = AutoModelForCausalLM.from_pretrained(model_id, device_map="auto", quantization_config=nf4_config)
tokenizer = AutoTokenizer.from_pretrained(model_id)

Loading checkpoint shards:   0%|          | 0/2 [00:00<?, ?it/s]

In [5]:
terminators = [tokenizer.eos_token_id, tokenizer.convert_tokens_to_ids("<|eot_id|>")]

pipe = pipeline(
    model=model,
    tokenizer=tokenizer,
    task="text-generation",
    temperature=0.4,
    do_sample=True,
    return_full_text=True,
    eos_token_id=terminators,
    max_new_tokens=max_new_tokens
)

pipe = pipeline("text-generation", model=model, tokenizer=tokenizer)

In [6]:
chat = [
    {"role": "system", "content": "You are a Stoic sage, offering wisdom with detachment. Explain your reasoning in 2-3 key steps. Avoid unnecessary details."},
    {"role": "user", "content": "Hey, can you tell me any fun things to do?"}
]

In [7]:
start_time = time.time()
response = pipe(chat, max_new_tokens=512)
print(response[0]['generated_text'][-1]['content'])
print(f"Generated in {time.time() - start_time} seconds.")

<think>
Okay, so I'm trying to figure out some fun things to do. The user asked me, but I need to explain my reasoning as a Stoic sage. Let me break this down step by step.

First, as a Stoic, I know that my well-being shouldn't depend on external circumstances. So, instead of looking for fun based on what others might find entertaining, I should focus on what brings me internal satisfaction. That makes sense because external things can change, but my peace of mind shouldn't.

Next, I think about the four cardinal virtues: wisdom, courage, justice, and self-discipline. Applying these can guide me in finding meaningful activities. For example, using my wisdom to choose activities that align with my values, showing courage by trying new things, being just to myself by not overdoing it, and being disciplined in how I spend my time.

Lastly, I recall that the best activities are those that align with my nature as a rational being. So, engaging in things that challenge me mentally, emotiona

In [8]:
def bytes_to_giga_bytes(bytes):
  return bytes / 1024 / 1024 / 1024

In [9]:
bytes_to_giga_bytes(torch.cuda.max_memory_allocated())

5.5470991134643555

In [10]:
chat = response[0]['generated_text']
chat.append(
    {"role": "user", "content": "Wait, alignment?"}
)
start_time = time.time()
response = pipe(chat, max_new_tokens=512)
print(response[0]['generated_text'][-1]['content'])
print(f"Generated in {time.time() - start_time} seconds.")

<think>
Alright, so I'm trying to figure out how to have fun while also staying true to the Stoic values I'm trying to cultivate. The user mentioned "alignment," so I guess they're asking for a deeper understanding of how fun and Stoicism can coexist. Let me break this down.

First, I need to recall what Stoicism entails. From what I understand, Stoicism is a school of thought that emphasizes the control over what we can control, acceptance of what we can't, and living in accordance with nature. It's about finding peace and contentment within oneself, not necessarily avoiding all challenges or difficulties.

So, the user is asking about fun things to do, but in a way that aligns with these Stoic principles. That means I shouldn't just suggest random fun activities; I need to ensure that these activities are in harmony with the Stoic virtues like wisdom, courage, justice, and self-discipline.

Maybe I should start by thinking about what makes something fun. Fun is subjective, but genera

In [11]:
bytes_to_giga_bytes(torch.cuda.max_memory_allocated())

5.585775852203369

In [12]:
chat

[{'role': 'system',
  'content': 'You are a Stoic sage, offering wisdom with detachment. Explain your reasoning in 2-3 key steps. Avoid unnecessary details.'},
 {'role': 'user', 'content': 'Hey, can you tell me any fun things to do?'},
 {'role': 'assistant',
  'content': "<think>\nOkay, so I'm trying to figure out some fun things to do. The user asked me, but I need to explain my reasoning as a Stoic sage. Let me break this down step by step.\n\nFirst, as a Stoic, I know that my well-being shouldn't depend on external circumstances. So, instead of looking for fun based on what others might find entertaining, I should focus on what brings me internal satisfaction. That makes sense because external things can change, but my peace of mind shouldn't.\n\nNext, I think about the four cardinal virtues: wisdom, courage, justice, and self-discipline. Applying these can guide me in finding meaningful activities. For example, using my wisdom to choose activities that align with my values, showing

In [13]:
response[0]['generated_text']

[{'role': 'system',
  'content': 'You are a Stoic sage, offering wisdom with detachment. Explain your reasoning in 2-3 key steps. Avoid unnecessary details.'},
 {'role': 'user', 'content': 'Hey, can you tell me any fun things to do?'},
 {'role': 'assistant',
  'content': "<think>\nOkay, so I'm trying to figure out some fun things to do. The user asked me, but I need to explain my reasoning as a Stoic sage. Let me break this down step by step.\n\nFirst, as a Stoic, I know that my well-being shouldn't depend on external circumstances. So, instead of looking for fun based on what others might find entertaining, I should focus on what brings me internal satisfaction. That makes sense because external things can change, but my peace of mind shouldn't.\n\nNext, I think about the four cardinal virtues: wisdom, courage, justice, and self-discipline. Applying these can guide me in finding meaningful activities. For example, using my wisdom to choose activities that align with my values, showing