# L1: NLP tasks with a simple interface üóûÔ∏è

Load your HF API key and relevant Python libraries.

In [1]:
import os
import io
from IPython.display import Image, display, HTML
from PIL import Image
import base64 
from utils import *
import gradio as gr

# Helper function
import requests, json

from dotenv import load_dotenv, find_dotenv


### How about running it locally?
The code would look very similar if you were running it locally instead of from an API. The same is true for all the models in the rest of the course, make sure to check the [Pipelines](https://huggingface.co/docs/transformers/main_classes/pipelines) documentation page

```py
from transformers import pipeline

get_completion = pipeline("summarization", model="facebook/bart-large-cnn")

def summarize(input):
    output = get_completion(input)
    return output[0]['summary_text']
    
```

In [2]:
from transformers import pipeline

get_completion = pipeline("summarization", model="facebook/bart-large-cnn")

def summarize(input):
    output = get_completion_summary(input)
    return output[0]['summary_text']

Device set to use cpu


## Building a text summarization app

In [3]:
text = ('''The tower is 324 metres (1,063 ft) tall, about the same height
        as an 81-storey building, and the tallest structure in Paris. 
        Its base is square, measuring 125 metres (410 ft) on each side. 
        During its construction, the Eiffel Tower surpassed the Washington 
        Monument to become the tallest man-made structure in the world,
        a title it held for 41 years until the Chrysler Building
        in New York City was finished in 1930. It was the first structure 
        to reach a height of 300 metres. Due to the addition of a broadcasting 
        aerial at the top of the tower in 1957, it is now taller than the 
        Chrysler Building by 5.2 metres (17 ft). Excluding transmitters, the 
        Eiffel Tower is the second tallest free-standing structure in France 
        after the Millau Viaduct.''')

get_completion_summary(text)

[{'summary_text': 'The tower is 324 metres (1,063 ft) tall, about the same height as an 81-storey building. Its base is square, measuring 125 metres (410 ft) on each side. It is the second tallest free-standing structure in France after the Millau Viaduct.'}]

### Getting started with Gradio `gr.Interface` 

In [4]:
def summarize(input):
    output = get_completion_summary(input)
    return output[0]['summary_text']
    
gr.close_all()
demo = gr.Interface(fn=summarize, inputs="text", outputs="text")
demo.launch(share=True, server_port=int(os.environ['PORT1']))

* Running on local URL:  http://127.0.0.1:7860
* Running on public URL: https://e93701adc1522c17c6.gradio.live

This share link expires in 72 hours. For free permanent hosting and GPU upgrades, run `gradio deploy` from the terminal in the working directory to deploy to Hugging Face Spaces (https://huggingface.co/spaces)




`demo.launch(share=True)` lets you create a public link to share with your team or friends.

In [5]:
import gradio as gr

def summarize(input):
    output = get_completion_summary(input)
    return output[0]['summary_text']

gr.close_all()
demo = gr.Interface(fn=summarize, 
                    inputs=[gr.Textbox(label="Text to summarize", lines=6)],
                    outputs=[gr.Textbox(label="Result", lines=3)],
                    title="Text summarization with distilbart-cnn",
                    description="Summarize any text using the `facebook/bart-large-cnn` model under the hood!"
                   )
demo.launch(share=True, server_port=int(os.environ['PORT2']))

Closing server running on port: 7860
* Running on local URL:  http://127.0.0.1:7861
* Running on public URL: https://7d7841c14c3e247cf4.gradio.live

This share link expires in 72 hours. For free permanent hosting and GPU upgrades, run `gradio deploy` from the terminal in the working directory to deploy to Hugging Face Spaces (https://huggingface.co/spaces)




## Building a Named Entity Recognition app

We are using this [Inference Endpoint](https://huggingface.co/inference-endpoints) for `dslim/bert-base-NER`, a 108M parameter fine-tuned BART model on the NER task.

### How about running it locally?

```py
from transformers import pipeline

get_completion = pipeline("ner", model="dslim/bert-base-NER")

def ner(input):
    output = get_completion(input)
    return {"text": input, "entities": output}
    
```

In [6]:
from transformers import pipeline

get_completion = pipeline("ner", model="dslim/bert-base-NER")

def ner(input):
    output = get_completion_summary(input)
    return {"text": input, "entities": output}

Some weights of the model checkpoint at dslim/bert-base-NER were not used when initializing BertForTokenClassification: ['bert.pooler.dense.bias', 'bert.pooler.dense.weight']
- This IS expected if you are initializing BertForTokenClassification from the checkpoint of a model trained on another task or with another architecture (e.g. initializing a BertForSequenceClassification model from a BertForPreTraining model).
- This IS NOT expected if you are initializing BertForTokenClassification from the checkpoint of a model that you expect to be exactly identical (initializing a BertForSequenceClassification model from a BertForSequenceClassification model).
Device set to use cpu


In [7]:
# Local NER pipeline
get_completion_ner = pipeline("ner", model="dslim/bert-base-NER")

# Test the NER pipeline
text = "My name is Miqueas, I'm building MMD Solutions and I live in Spain"
result = get_completion_ner(text)  # Use this instead of get_completion()
print(result)

Some weights of the model checkpoint at dslim/bert-base-NER were not used when initializing BertForTokenClassification: ['bert.pooler.dense.bias', 'bert.pooler.dense.weight']
- This IS expected if you are initializing BertForTokenClassification from the checkpoint of a model trained on another task or with another architecture (e.g. initializing a BertForSequenceClassification model from a BertForPreTraining model).
- This IS NOT expected if you are initializing BertForTokenClassification from the checkpoint of a model that you expect to be exactly identical (initializing a BertForSequenceClassification model from a BertForSequenceClassification model).
Device set to use cpu


[{'entity': 'B-PER', 'score': np.float32(0.99400455), 'index': 4, 'word': 'Mi', 'start': 11, 'end': 13}, {'entity': 'I-PER', 'score': np.float32(0.9677398), 'index': 5, 'word': '##que', 'start': 13, 'end': 16}, {'entity': 'B-PER', 'score': np.float32(0.44358006), 'index': 6, 'word': '##as', 'start': 16, 'end': 18}, {'entity': 'B-ORG', 'score': np.float32(0.9990521), 'index': 12, 'word': 'M', 'start': 33, 'end': 34}, {'entity': 'I-ORG', 'score': np.float32(0.99648154), 'index': 13, 'word': '##MD', 'start': 34, 'end': 36}, {'entity': 'I-ORG', 'score': np.float32(0.99809283), 'index': 14, 'word': 'Solutions', 'start': 37, 'end': 46}, {'entity': 'B-LOC', 'score': np.float32(0.9997489), 'index': 19, 'word': 'Spain', 'start': 61, 'end': 66}]


In [8]:
# Create local pipeline
get_completion_local = pipeline("ner", model="dslim/bert-base-NER")

# Test the API directly first
text = "My name is Miqueas, I'm building MMD Solutions and I live in Spain"
try:
    api_result = get_completion_api(text)
    print("\nAPI Result:", api_result)
except Exception as e:
    print("\nAPI Error:", e)

Some weights of the model checkpoint at dslim/bert-base-NER were not used when initializing BertForTokenClassification: ['bert.pooler.dense.bias', 'bert.pooler.dense.weight']
- This IS expected if you are initializing BertForTokenClassification from the checkpoint of a model trained on another task or with another architecture (e.g. initializing a BertForSequenceClassification model from a BertForPreTraining model).
- This IS NOT expected if you are initializing BertForTokenClassification from the checkpoint of a model that you expect to be exactly identical (initializing a BertForSequenceClassification model from a BertForSequenceClassification model).
Device set to use cpu


Using API URL: https://api-inference.huggingface.co/models/dslim/bert-base-NER
Token starts with: hf_Ss...

API Result: [{'entity_group': 'PER', 'score': 0.9808721542358398, 'word': 'Mique', 'start': 11, 'end': 16}, {'entity_group': 'PER', 'score': 0.44358277320861816, 'word': '##as', 'start': 16, 'end': 18}, {'entity_group': 'ORG', 'score': 0.9978755116462708, 'word': 'MMD Solutions', 'start': 33, 'end': 46}, {'entity_group': 'LOC', 'score': 0.9997488856315613, 'word': 'Spain', 'start': 61, 'end': 66}]


In [9]:
API_URL = os.environ['HF_API_NER_BASE'] #NER endpoint
text = "My name is Miqueas, I'm building MMD Solutions and I live in Spain"
get_completion_api(text)

Using API URL: https://api-inference.huggingface.co/models/dslim/bert-base-NER
Token starts with: hf_Ss...


[{'entity_group': 'PER',
  'score': 0.9808721542358398,
  'word': 'Mique',
  'start': 11,
  'end': 16},
 {'entity_group': 'PER',
  'score': 0.44358277320861816,
  'word': '##as',
  'start': 16,
  'end': 18},
 {'entity_group': 'ORG',
  'score': 0.9978755116462708,
  'word': 'MMD Solutions',
  'start': 33,
  'end': 46},
 {'entity_group': 'LOC',
  'score': 0.9997488856315613,
  'word': 'Spain',
  'start': 61,
  'end': 66}]

#### gr.interface()
- Notice below that we pass in a list `[]` to `inputs` and to `outputs` because the function `fn` (in this case, `ner()`, can take in more than one input and return more than one output.
- The number of objects passed to `inputs` list should match the number of parameters that the `fn` function takes in, and the number of objects passed to the `outputs` list should match the number of objects returned by the `fn` function.

In [10]:
# For the Gradio interface
def ner(input):
    output = get_completion_ner(input)  # Use the NER-specific pipeline
    return {"text": input, "entities": output}

gr.close_all()
demo = gr.Interface(fn=ner,
                    inputs=[gr.Textbox(label="Text to find entities", lines=2)],
                    outputs=[gr.HighlightedText(label="Text with entities")],
                    title="NER with dslim/bert-base-NER",
                    description="Find entities using the `dslim/bert-base-NER` model under the hood!",
                    flagging_mode="never",
                    #Here we introduce a new tag, examples, easy to use examples for your application
                    examples=["My name is Miqueas, I'm building MMD Solutions and I live in Spain"])
demo.launch(share=True, server_port=int(os.environ['PORT3']))

Closing server running on port: 7861
Closing server running on port: 7860
* Running on local URL:  http://127.0.0.1:7862
* Running on public URL: https://4edae7630b870b3dde.gradio.live

This share link expires in 72 hours. For free permanent hosting and GPU upgrades, run `gradio deploy` from the terminal in the working directory to deploy to Hugging Face Spaces (https://huggingface.co/spaces)




### Adding a helper function to merge tokens

In [11]:
# Create local pipeline
get_completion_local = pipeline("ner", model="dslim/bert-base-NER")

def ner(input):
    # Changed this line to use get_completion_local directly
    output = get_completion_local(input)
    merged_tokens = merge_tokens(output)
    return {"text": input, "entities": merged_tokens}

gr.close_all()
demo = gr.Interface(fn=ner,
                    inputs=[gr.Textbox(label="Text to find entities", lines=2)],
                    outputs=[gr.HighlightedText(label="Text with entities")],
                    title="NER with dslim/bert-base-NER",
                    description="Find entities using the `dslim/bert-base-NER` model under the hood!",
                    flagging_mode="never",
                    examples=["My name is Miqueas, I'm building MMD Solutions and I live in Spain"])

demo.launch(share=True, server_port=int(os.environ['PORT4']))

Some weights of the model checkpoint at dslim/bert-base-NER were not used when initializing BertForTokenClassification: ['bert.pooler.dense.bias', 'bert.pooler.dense.weight']
- This IS expected if you are initializing BertForTokenClassification from the checkpoint of a model trained on another task or with another architecture (e.g. initializing a BertForSequenceClassification model from a BertForPreTraining model).
- This IS NOT expected if you are initializing BertForTokenClassification from the checkpoint of a model that you expect to be exactly identical (initializing a BertForSequenceClassification model from a BertForSequenceClassification model).
Device set to use cpu


Closing server running on port: 7861
Closing server running on port: 7860
Closing server running on port: 7862
* Running on local URL:  http://127.0.0.1:7863
* Running on public URL: https://fcc28e8292d4c8fe69.gradio.live

This share link expires in 72 hours. For free permanent hosting and GPU upgrades, run `gradio deploy` from the terminal in the working directory to deploy to Hugging Face Spaces (https://huggingface.co/spaces)




In [12]:
gr.close_all()

Closing server running on port: 7863
Closing server running on port: 7861
Closing server running on port: 7860
Closing server running on port: 7862


## How to get your own Hugging Face API key (token)

Hugging Face "API keys" are called "User Access tokens".  

You can create your own User Access Tokens here: [Access Tokens](https://huggingface.co/settings/tokens).

#### Save your user access tokens to environment variables
To save your access token securely on your own machine:
- Create a `.env` file in the root directory of your project.
- Edit the file to contain the following:  
`HF_API_KEY="abc123"` replace that string with your user access token.
- Save the .env file.
- Install Python-dotenv, which allows you to run that first code cell at the top of this jupyter notebook:  
`pip install python-dotenv`


For more information on how to get your own access tokens, please see [User access tokens](https://huggingface.co/docs/hub/security-tokens#:~:text=To%20create%20an%20access%20token,you're%20ready%20to%20go!)