<a href="https://colab.research.google.com/github/colesmcintosh/llama-4-scout-arc-agi-sample/blob/main/LLama_4_Scout_ARC_AGI_Sample.ipynb" target="_parent"><img src="https://colab.research.google.com/assets/colab-badge.svg" alt="Open In Colab"/></a>

# LLama 4 Scout ARC AGI Sample

This notebook demonstrates a sample implementation for solving ARC AGI puzzles using Llama 4 Scout model from Meta, integrated with LangChain, powered by OpenRouter.

## Setup and Installation

First, we install the necessary LangChain packages to work with the LLama 4 Scout model and process ARC puzzle data.

In [22]:
!pip install langchain langchain-openai langchain-core langchain-community

Collecting langchain-community
  Downloading langchain_community-0.3.21-py3-none-any.whl.metadata (2.4 kB)
Collecting langchain-core
  Downloading langchain_core-0.3.51-py3-none-any.whl.metadata (5.9 kB)
Collecting langchain
  Downloading langchain-0.3.23-py3-none-any.whl.metadata (7.8 kB)
Collecting dataclasses-json<0.7,>=0.5.7 (from langchain-community)
  Downloading dataclasses_json-0.6.7-py3-none-any.whl.metadata (25 kB)
Collecting pydantic-settings<3.0.0,>=2.4.0 (from langchain-community)
  Downloading pydantic_settings-2.8.1-py3-none-any.whl.metadata (3.5 kB)
Collecting httpx-sse<1.0.0,>=0.4.0 (from langchain-community)
  Downloading httpx_sse-0.4.0-py3-none-any.whl.metadata (9.0 kB)
Collecting langchain-text-splitters<1.0.0,>=0.3.8 (from langchain)
  Downloading langchain_text_splitters-0.3.8-py3-none-any.whl.metadata (1.9 kB)
Collecting marshmallow<4.0.0,>=3.18.0 (from dataclasses-json<0.7,>=0.5.7->langchain-community)
  Downloading marshmallow-3.26.1-py3-none-any.whl.metadata 

## Importing Libraries

We import the required modules from LangChain and other dependencies to set up the model and handle inputs.

In [8]:
from langchain_openai import ChatOpenAI
from langchain_core.prompts import ChatPromptTemplate
from google.colab import userdata

## Model Initialization

Here, we initialize the LLama 4 Scout model using the ChatOpenAI class, configured with an API key and base URL from Colab's userdata.

In [None]:
llm = ChatOpenAI(
  openai_api_key=userdata.get("OPENROUTER_API_KEY"),
  openai_api_base=userdata.get("OPENROUTER_BASE_URL"),
  model_name="meta-llama/llama-4-scout",
)

## ARC Puzzle Image

We define the URL of an example ARC puzzle image to be solved. The image is displayed below for reference.

In [9]:

image_url = "https://arcprize.org/media/images/arc-example-task.jpg"

In [10]:
from IPython.display import Image

Image(url=image_url)


## Loading ARC Guidelines

We fetch the ARC AGI guidelines from the official ARC Prize website to provide context for solving the puzzle.

In [23]:
import bs4
from langchain_community.document_loaders import WebBaseLoader

page_url = "https://arcprize.org/guide"

loader = WebBaseLoader(web_paths=[page_url])
docs = []
for doc in loader.load():
    docs.append(doc)



## Defining the Prompt

We create a `ChatPromptTemplate` to structure the input for the model, including the ARC guidelines and the puzzle image.

In [32]:
image_message = ChatPromptTemplate([
    {
        "role": "system",
        "content": "Here are the ARC AGI guidelines: {guidelines}"
    },
    {
      "role": "user",
      "content": [
        {
          "type": "text",
          "text": "Solve the following ARC puzzle image based on the guidelines."
        },
        {
          "type": "image_url",
          "image_url": {
            "url": "{image_url}"
          }
        }
      ]
    }
  ]
)

## Creating the Chain and Running the Model

We chain the prompt with the LLama 4 Scout model and invoke it with the image URL and guidelines to get the solution.

In [33]:
chain = image_message | llm

In [34]:
result = chain.invoke({"image_url": image_url, "guidelines": docs})

In [35]:
result.content

'## Step 1: Analyze the given ARC-AGI task\nThe task consists of six 8x8 grids, with the first four grids being the input and the last two grids being the output. The goal is to determine the output for the test input.\n\n## Step 2: Identify the pattern in the input and output grids\nUpon examining the grids, it appears that the output grids are generated by "filling in" or "coloring" certain areas of the input grids.\n\n## Step 3: Determine the specific transformation applied to each input grid\nLooking closely, the transformation seems to involve expanding or growing the green shapes in the input grids.\n\n## Step 4: Analyze the first input-output pair\nThe first input grid has a green shape that resembles a "6" or a "9". The corresponding output grid has this shape filled in with yellow.\n\n## 5: Analyze the second input-output pair\nThe second input grid has a green shape that resembles a "U" or an inverted "L". The corresponding output grid has this shape filled in with yellow.\n\