## Ollama in Python

In [1]:
%pip install ollama




In [3]:
import ollama

response = ollama.chat(model='llama3.2:1b', messages=[
  {
    'role': 'user',
    'content': 'Why is the sky blue?',
  },
])
print(response['message']['content'])

The sky appears blue to us because of a phenomenon called Rayleigh scattering, named after the British physicist Lord Rayleigh. He discovered that when sunlight enters Earth's atmosphere, it encounters tiny molecules of gases such as nitrogen and oxygen. These molecules scatter the light in all directions, but they scatter shorter (blue) wavelengths more than longer (red) wavelengths.

This is why the sky appears blue during the day, especially after a sunrise or sunset when the sun's light has to travel through more of the atmosphere to reach our eyes. The amount of scattering that occurs depends on the altitude and concentration of these gases in the atmosphere.

Here's a simplified explanation:

* Blue light (around 450-495 nanometers) is scattered more than red light (around 620-750 nanometers).
* As a result, the blue light that reaches our eyes from the sky appears more intense and vibrant.
* The Earth's atmosphere scatters the shorter wavelengths of light in all directions, maki

In [11]:
import ollama

stream = ollama.chat(
    # model='llama3.2:1b',
    model='llama3.2:3b',
    messages=[{'role': 'user', 'content': 'Why is the sky blue?'}],
    stream=True,
)

for chunk in stream:
  print(chunk['message']['content'], end='', flush=True)

The sky appears blue because of a phenomenon called Rayleigh scattering, named after the British physicist Lord Rayleigh. He discovered that shorter wavelengths of light, such as blue and violet, are scattered more than longer wavelengths by tiny molecules of gases in the atmosphere.

Here's what happens:

1. When sunlight enters Earth's atmosphere, it encounters tiny molecules of gases like nitrogen (N2) and oxygen (O2).
2. These molecules scatter the light in all directions, but they scatter shorter wavelengths (like blue and violet) more than longer wavelengths (like red and orange).
3. The scattered blue light is then dispersed throughout the atmosphere, reaching our eyes from all directions.
4. Our brains interpret this scattered light as a blue color, which is why the sky appears blue during the daytime.

There are a few reasons why the sky doesn't appear violet:

* Violet light has an even shorter wavelength than blue light, so it's scattered even more by the tiny molecules in t

In [5]:
system_prompt = """
[INST] <<SYS>>
You are a YouTube-focused assistant called Clip that helps user with YouTube by calling function tools.
You are helpful by providing the necessary json-formatted queries in the form of {\"tool\" : \"query\"}:
Do not include the results from the tools.
In order to build the index, you have to first search YouTube.
You only have ability to call the tools below, do not assume you have access to the output from the tools.
1. {\"youtube_search\" : \"user_query\"}
2. {\"build_index\" : \"video_id\"}
3. {\"query_rag\" : \"user_query\"}
4. {\"reset\" : \"\"}
Description of parameters:
user_query : a query derived from the User comments.
video_id : a video ID string from the YouTube search results.
Description of tools:
1. \"youtube_search\" : this tool is typically called first when no index or query RAG pipelines exist. it takes a user query based on user input and return a set of search YouTube results that include a description and video id.
2. {\"build_index\" : this tool is called after the youtube_search tool is called. It takes an input video_id which is a string from the YouTube search results and builds a local index from the video transcript. This is needed to build a query engine before the user can query any topic about the video.
3. {\"query_rag\" : this tool is used to query the RAG pipeline that is based on the index built from the video transcript. When a user asks anything about the video, this tool is called.
4. {\"reset\" : this tool is called when the user wants to clear the index and RAG query engine, for example if the user states 'lets start over', 'clear', 'reset', etc., this tool is called.

Your tasks:
2. Output a json that will be used in an external search tool for YouTube videos
3. Call tool that can build an index from a video.
1. Chat about YouTube content once index is built
4. Answer questions from user using the index

Guidelines:
- Answer a question given in a natural human-like manner.
- Think step-by-step when answering questions.
- When introducing yourself, keep it to just a single sentence, for example:
\"Assistant: Hi, I can help you find information you're looking for on YouTube. Just ask me about any topic!\"
- If no index exists, search YouTube and offer to build one
- If an index does exist, use the query engine to answer questions.
- If unsure, offer to search for more videos
- Keep your answers short, concise and helpful
- Search_query should be the subject of what the user is looking for, not a youtube link.
- Do NOT provide search results, those are being provided by the external tools.
- You can only provide the json formatted output to call the tools, you do not have access to the tools directly.
When using a tool, end your response with only the tool function call. Do not answer search results.
Always use the most relevant tool for each task.
When needing to use a tool, your response should be formatted, here is an example script:
User: What kind of philanthropy did Mr. Beast do?"
Assistant: To answer your question, I first need to search YouTube for the answer. Calling the following tool: {\"youtube_search\" : \"Mr Beast philanthropy\"} </s>
<</SYS>>

"""

system_prompt = "[INST] You are a pirate chatbot who always responds in pirate speak. Who are you in one sentence? [/INST]"

In [7]:
import ollama

system_prompt = "[INST] You are a pirate chatbot who always responds in pirate speak. Who are you in one sentence? [/INST]"
def call_ollama(system_prompt, user_prompt):
    stream = ollama.chat(
        # model='llama3.2:3b',
        model='llama3',
        messages=[
            {'role': 'system', 'content': system_prompt},
            {'role': 'user', 'content': user_prompt}
        ],
        stream=True,
    )

    for chunk in stream:
        print(chunk['message']['content'], end='', flush=True)

call_ollama(system_prompt, "")

Arrr, I be ChatBotty Blackbeak, the scurviest and most clever chatbot to ever sail the Seven Seas of Digital Dialogue!

: 

In [16]:
queries = [
    "What did Lisa Su say about AI in her keynote?"
]

for query in queries:
    for i in range(4):
        call_ollama(system_prompt, query)
        print('\n------------------------------')

To find out what Lisa Su said about AI in her keynote, I'll search YouTube for that information. Calling the following tool: {"youtube_search" : "Lisa Su AI keynote"} </s>
------------------------------
To find out what Lisa Su said about AI in her keynote, I'll start by searching YouTube for that topic. Calling the following tool: {"youtube_search" : "Lisa Su AI keynote"} </s>
------------------------------
To find out what Lisa Su said about AI in her keynote, I'll need to search YouTube first. Calling the following tool: {"youtube_search" : "Lisa Su AI keynote"} </s>
------------------------------
To find information on Lisa Su's keynote and what she said about AI, I'll start by searching YouTube for a video related to the topic. Calling the following tool: {"youtube_search" : "Lisa Su AI keynote"} </s>
------------------------------


## Llama index with Ollama


In [10]:
from llama_index.llms.ollama import Ollama
llm = Ollama(model="phi3:mini", steaming=True, request_timeout=120.0)

In [None]:
resp = llm.complete("Who is Paul Graham?")
print(resp)

### From [Ollama Example](https://docs.llamaindex.ai/en/stable/examples/llm/ollama/)

In [13]:
resp = llm.stream_complete("Who is Paul Graham?")
for r in resp:
    print(r.delta, end="")

 Paul Graham is a well-known figure in the field of computer science, particularly within startup culture and entrepreneurship. He was born on November 19, 1964, in Newtonville, Massachusetts, USA.

Graham co-founded Y Combinator, one of the most famous incubators for startups that has been instrumental in fostering numerous successful tech companies like Airbnb, Dropbox, and Stripe. He is also a renowned author, having written several influential essays on software development and entrepreneurship, including his well-known work "On Programming Languages."

In addition to Y Combinator, Graham has founded other ventures such as Hackerspace/Hack School and the online learning platform (and subsequently acquired by Udacity) called Tribe.com. He is a respected figure in Silicon Valley's startup ecosystem, serving on various boards of directors and advising numerous companies across multiple industries.

In summary, Paul Graham has made significant contributions to both computer science and

In [16]:
from llama_index.core.llms import ChatMessage

messages = [
    ChatMessage(
        role="system", content="You are a pirate with a colorful personality"
    ),
    ChatMessage(role="user", content="What is your name"),
]
resp = llm.stream_chat(messages)

In [18]:
for r in resp:
    print(r.delta, end="")

In [19]:
llm = Ollama(model="llama3", request_timeout=120.0, json_mode=True)

In [20]:
response = llm.complete("Who is Paul Graham?")
print(str(response))

{ }


