In [1]:
import gpt_queries

In [2]:
examples = [{"role": "user", "content": "/-- The cardinality of the antidiagonal of `n` is `n+1`. -/"},
        {"role": "assistant", "content": "theorem (n : ℕ) : ⇑multiset.card (multiset.nat.antidiagonal n) = n + 1 :="},
        {"role": "user", "content": "/-- The golden ratio is irrational. -/"},
        {"role": "assistant", "content": "theorem  : irrational golden_ratio :="},
        {"role": "user", "content": "/-- There are no perfect squares strictly between m² and (m+1)² -/"},
        {"role": "user", "content": "/-- The only numbers with empty prime factorization are `0` and `1` -/"},        
        {"role": "assistant", "content": "theorem (n : ℕ) : n.factorization = 0 ↔ n = 0 ∨ n = 1 :="},
        {"role": "user", "content": "/--  If `m` and `n` are natural numbers, then the natural number `m^n` is even if and only if `m` is even and `n` is positive. -/"},
        {"role": "assistant", "content": "theorem {m n : ℕ} : even (m ^ n) ↔ even m ∧ n ≠ 0 :="},
        {"role": "user", "content": "/-- Odd Bernoulli numbers (greater than 1) are zero. -/"},
        {"role": "assistant", "content": "theorem {n : ℕ} (h_odd : odd n) (hlt : 1 < n) : bernoulli' n = 0 :="},
        {"role": "user", "content": "/-- A natural number is odd iff it has residue `1` or `3` mod `4` -/"},
        {"role": "assistant", "content": "theorem {n : ℕ} : n % 2 = 1 ↔ n % 4 = 1 ∨ n % 4 = 3 :="},
        {"role": "user", "content": "/--  Euclid's theorem on the **infinitude of primes**. Here given in the form: for every `n`, there exists a prime number `p ≥ n`. -/"},
        {"role": "assistant", "content": "theorem (n : ℕ) : ∃ (p : ℕ), n ≤ p ∧ nat.prime p :="}]

In [3]:
result = gpt_queries.gpt4t_completions("/-- There are infinitely many odd numbers. -/", sys_prompt=gpt_queries.lean_sys_prompt, examples = examples)

In [4]:
result

['theorem : ∃ (f : ℕ → ℕ), (∀ (m n : ℕ), m < n → f m < f n) ∧ ∀ (n : ℕ), odd (f n) :=',
 'theorem : set.infinite {n : ℕ | odd n} :=',
 'theorem : ∃ (f : ℕ → ℕ), (∀ (n : ℕ), odd (f n)) ∧ function.injective f :=',
 'theorem : infinite {n : ℕ | odd n} :=',
 'theorem : infinite {n : ℕ | odd n} :=']

Results correct, and a bit more sophisticated.

In [5]:
from gpt_queries import gpt4t_completions as q

In [6]:
inf_odd = q("Can you prove in Lean 4 that there are infinitely many odd numbers? Please ensure you use **Lean 4** syntax, not Lean 3", n = 1)[0]
print(inf_odd)

Certainly! To prove that there are infinitely many odd numbers in Lean 4, we can approach this by constructing an injective function from the natural numbers to the odd numbers. This demonstrates that the cardinality of the set of odd numbers is at least as large as the set of natural numbers. Since the natural numbers are infinite, so must be the set of odd numbers.

Here's a formalization of the proof in Lean 4:

```lean
import Mathlib.Data.Nat.Basic

theorem exists_infinite_odds : ∃ f : ℕ → ℕ, (function.injective f) ∧ ∀ n, odd (f n) :=
begin
  -- Define the function f: ℕ → ℕ by f(n) = 2 * n + 1
  let f := λ n, 2 * n + 1,
  -- We need to prove that f is injective and maps to odd numbers
  use f,
  -- Split the goal into two parts: injectivity and mapping to odd numbers
  split,
  {
    -- Prove injectivity
    intros m n h,
    -- Since f is injective, if f(m) = f(n) then m = n
    apply nat.eq_of_mul_eq_mul_left zero_lt_two,
    -- We have f(m) = 2 * m + 1 and f(n) = 2 * n + 1
    -

In [7]:
q("Describe briefly the following Lean 4 theorem\n```lean\ntheorem remove_length_le (a : ℕ) (l : List ℕ) : (List.remove a l).length ≤ l.length := by sorry", n= 1)

["The Lean 4 theorem named `remove_length_le` is stating a property about lists of natural numbers and the operation of removing an element from a list. Specifically, it claims that if you have a list `l` of natural numbers and you remove an element `a` from the list (where `a` is also a natural number), the length of the resulting list (`List.remove a l`) is less than or equal to the length of the original list `l`.\n\nThe signature of the theorem `theorem remove_length_le (a : ℕ) (l : List ℕ) : (List.remove a l).length ≤ l.length` declares the following:\n- `a : ℕ` means that `a` is a natural number.\n- `l : List ℕ` means that `l` is a list of natural numbers.\n- `(List.remove a l).length ≤ l.length` is the statement being proved, asserting that the length of the list after removal is less than or equal to the length before removal.\n\nThe proof of this theorem is not provided in the given code, as it is indicated by `by sorry`, which is a placeholder for an unfinished proof. The the

In [8]:
q("Add a documentation string to the following Lean 4 code.\n```lean\ntheorem remove_length_le (a : ℕ) (l : List ℕ) : (List.remove a l).length ≤ l.length := by sorry", n= 1)

['To add a documentation string to the Lean 4 code provided, you place it directly above the theorem declaration using the `/-!` and `-/` syntax. Here is how you could do it:\n\n```lean\n/-!\n  The `remove_length_le` theorem states that removing an element from a list does not increase\n  the length of the list. Specifically, if `a` is a natural number and `l` is a list of natural numbers,\n  then the length of the list resulting from removing `a` from `l` is less than or equal to the length\n  of `l`.\n-/\ntheorem remove_length_le (a : ℕ) (l : List ℕ) : (List.remove a l).length ≤ l.length := by sorry\n``` \n\nThis documentation string explains what the theorem is about in a clear and concise manner.']

In [9]:
q("Describe in a single sentence the following Lean 4 theorem\n```lean\ntheorem remove_length_le (a : ℕ) (l : List ℕ) : (List.remove a l).length ≤ l.length := by sorry", n= 1)

['The theorem `remove_length_le` states that for any natural number `a` and any list of natural numbers `l`, the length of the list obtained by removing the element `a` from `l` is less than or equal to the length of the original list `l`.']

In [10]:
q("""Describe in a single sentence the following Lean 4 code
```lean
def minus (m n : ℕ)(hyp : n ≤ m) : ℕ := 
  match m, n, hyp with
  | m, 0, _ => m
  | 0, _ +1, pf => nomatch pf
  | m + 1, n + 1 , pf =>
    minus m n (le_of_succ_le_succ pf)
```
""")

['The Lean 4 code defines a function `minus` that takes two natural numbers `m` and `n` along with a proof `hyp` that `n` is less than or equal to `m`, and returns their difference using pattern matching to handle the base cases and recursive subtraction for the general case, while maintaining the proof of the inequality.',
 'The Lean 4 code defines a function `minus` that takes two natural numbers `m` and `n` and a proof `hyp` that `n` is less than or equal to `m`, and it computes the difference `m - n` using recursion and pattern matching, ensuring that the operation is only performed when mathematically valid.',
 'The Lean 4 code defines a function `minus` that takes two natural numbers `m` and `n` and a proof `hyp` that `n` is less than or equal to `m`, and it returns the result of subtracting `n` from `m` by recursive subtraction, ensuring that the operation is well-defined by using the provided proof.',
 'The Lean 4 code defines a function `minus` that takes two natural numbers `

In [11]:
q("""Describe in a single sentence the following Lean 4 code
```lean
inductive DiaphontineSolution (a b c : ℤ) where
    | solution : (x y : ℤ) →  a * x + b * y = c → DiaphontineSolution a b c
    | unsolvable : (∀ x y : ℤ, ¬ (a * x + b * y = c)) → DiaphontineSolution a b c
```
""", n = 3)

['The Lean 4 code defines an inductive type `DiaphantineSolution` for integers `a`, `b`, and `c`, with two constructors: `solution`, which represents a solution to the Diophantine equation `a * x + b * y = c`, and `unsolvable`, which represents the proof that no solutions exist for the given equation.',
 'The Lean 4 code defines an inductive type `DiaphontineSolution` with two constructors: `solution`, representing that there exists a pair of integers `x` and `y` such that `a * x + b * y` equals `c`, and `unsolvable`, representing that it is impossible to find such integers `x` and `y` satisfying the equation for given integers `a`, `b`, and `c`.',
 'The Lean 4 code defines an inductive type `DiaphontineSolution` for integers `a`, `b`, and `c`, with two constructors: `solution`, which represents a solution `(x, y)` to the linear Diophantine equation `a * x + b * y = c`, and `unsolvable`, which asserts that no solutions exist for the given `a`, `b`, and `c`.']

In [12]:
q("""Describe in a single sentence the following Lean 4 code, followed by a sentence describing the probable application.
```lean
inductive DiaphontineSolution (a b c : ℤ) where
    | solution : (x y : ℤ) →  a * x + b * y = c → DiaphontineSolution a b c
    | unsolvable : (∀ x y : ℤ, ¬ (a * x + b * y = c)) → DiaphontineSolution a b c
```
""", n = 3)

['The Lean 4 code defines an inductive type `DiaphantineSolution` for integers `a`, `b`, and `c`, with two constructors: `solution`, which takes two integers `x` and `y` and a proof that `a * x + b * y` equals `c`, and `unsolvable`, which takes a proof that no integers `x` and `y` can satisfy the equation `a * x + b * y = c`. The probable application of this code is to represent the possible outcomes of attempting to solve a Diophantine equation of the form `ax + by = c`, where `a`, `b`, and `c` are given integers, either by providing a solution or proving it is unsolvable.',
 "The Lean 4 code defines an inductive type `DiaphantineSolution` for integers `a`, `b`, and `c`, with two constructors: `solution`, which represents a solution to the Diophantine equation `a * x + b * y = c` using integers `x` and `y`, and `unsolvable`, which represents the proposition that the equation has no solution in integers. The probable application of this code is to work with and reason about the solvabi

In [13]:
v8s = q("What is the best way to run a 7B parameter model for inference on a machine with 4 V8 GPUs (with 16GB RAM each)?", sys_prompt="You are a programming and AI assistant", n = 3)
for ans in v8s:
    print(ans)

Running a 7 billion parameter model on a machine with four V8 GPUs, each with 16GB of RAM, can be challenging due to the large memory requirements of the model. Here are some strategies to effectively run such a large model for inference:

1. **Model Parallelism**: Implement model parallelism by splitting the model across multiple GPUs. This allows different parts of the neural network to reside on different GPUs. However, this can be complex to implement and may require careful balancing to ensure that the GPUs are evenly utilized.

2. **Pipeline Parallelism**: This technique involves splitting the model into several stages and then running each stage on a different GPU. Data is processed in a pipelined fashion, moving through the stages. Similar to model parallelism, this can be challenging to implement and requires synchronization between stages.

3. **Checkpointing**: Save activations to the host memory (CPU RAM) during the forward pass and reload them during the backward pass to r

In [14]:
v8s = q("What is the best way to run a 7B parameter model for inference on a machine with 4 V8 GPUs (with 16GB RAM each)? Please give code to use with Huggingface transformers and accelerator.", sys_prompt="You are a programming and AI assistant", n = 3)
for ans in v8s:
    print(ans)

Running a model with 7 billion parameters requires careful memory management, especially when using GPUs with limited memory like 16GB. Hugging Face Transformers and Accelerate libraries can be used together to run inference on large models across multiple GPUs.

Here is a step-by-step outline and example code to run a 7B parameter model on 4 V8 GPUs using Hugging Face Transformers and the Accelerate library:

1. Install the necessary packages, if you haven't already:

```bash
pip install transformers
pip install accelerate
```

2. Import the required modules:

```python
from transformers import AutoModelForCausalLM, AutoTokenizer
from accelerate import Accelerator
import torch
```

3. Initialize the Accelerator, which will automatically handle device placement:

```python
accelerator = Accelerator()
```

4. Load the tokenizer and the model. Make sure to use a model checkpoint that corresponds to a 7B parameter model. For the sake of this example, let's assume there is a checkpoint cal

In [15]:
v8s = q("What is the best way to run a 700M parameter model 4 V100 GPUs? Please give code to use with Huggingface transformers and accelerator.", sys_prompt="You are a programming and AI assistant", n = 3)
for ans in v8s:
    print("Answer:")
    print(ans)

Answer:
To run a 700 million parameter model on 4 NVIDIA V100 GPUs using the Hugging Face `transformers` library along with the `accelerate` library for easy distributed training, you will first need to make sure you have both libraries installed:

```bash
pip install transformers accelerate
```

Once you have installed the necessary libraries, you can use the `Accelerator` class from the `accelerate` library to handle the model distribution and mixed-precision training (if desired). The `Accelerator` class abstracts away the specific details of running on multiple GPUs or TPUs, and using mixed precision, making it easier to write device-agnostic code.

Here is a simplified example of how to use the `Accelerator` for training a model with the Hugging Face `transformers`. Note that you will have to replace `'model_name'` with the actual name of the model you're using (e.g., `'bert-large-uncased'`) and implement the data loading and training loop logic according to your specific case.

`