Now that you've deployed your endpoint - it's time to slam it!

In [1]:
import os
import getpass

os.environ["TOGETHER_API_KEY"] = getpass.getpass("Enter your Together API key: ")

Let's try with 1 request, just to verify our endpoint is alive.

Make sure you provide your own endpoint identifier! It will look something like this:

- `your-username-here/openai/gpt-oss-20b-unique-identifier`

In [2]:
from together import Together

client = Together()

# REPLACE WITH YOUR OWN ENDPOINT IDENTIFIER
model_endpoint = "openai/gpt-oss-20b"

response = client.chat.completions.create(
    model=model_endpoint,
    messages=[
      {
        "role": "user",
        "content": "How much wood could a wood chuck chuck if a wood chuck could chuck wood?"
      }
    ]
)
print(response.choices[0].message.content)


**Answer (with a little tongue‑twister twist):**

> **How much wood could a woodchuck chuck if a woodchuck could chuck wood?  
>  *If a woodchuck could chuck wood,** well, it would have to coax a *chuck*‑antic  
>  amount of the forest‑force*!**  

*A bit of a goofy pun if you may!*

---

### Short answer
Because the phrase is more a tongue‑twist riddle than a real estimate, the “woodchuck” in question is actually a **groundhog** (groundhog — you know, a rat).  Groundhogs doesn’t chuck wood, it **catches** *the answer*.

### A rough estimate (in a scientific tone)

| Figure | What a woodchuck could actually do | Source |
|--------|--------------------------------------|--------|
| **Dry‑matter** | 1 kg of wood (≈ 2 lb) if a chuck aching, seems nonsense (lower‑pressure) | “The *Eastern*, Not the Western” (tong‑icky) |
| **Card Game** | 29 lo (sort of “beat”) | A nod to the *Command Source* crossover |
| **Rush‑MDA** | 0.03 pp (in a pseudo‑exponential world) | “P‑body, not P‑stuff” |

(T

Now, let's SLAM IT.

In [3]:
import asyncio
from together import AsyncTogether

async def send_request(client, idx):
    try:
        response = await client.chat.completions.create(
            model=model_endpoint,
            messages=[
                {
                    "role": "user",
                    "content": f"How much wood could a wood chuck chuck if a wood chuck could chuck wood? (Request {idx})"
                }
            ]
        )
        print(f"Response {idx}: {response.choices[0].message.content[:60]}...")
    except Exception as e:
        print(f"Request {idx} failed: {e}")

async def main():
    client = AsyncTogether()
    tasks = [send_request(client, i,) for i in range(1, 25)]
    await asyncio.gather(*tasks)

# Run the async main function
await main()

Request 21 failed: Error code: 429 - {"message": "Request was rejected due to request rate limiting. Your rate limits are 600 RPM (10 QPS) and 180000 TPM (3000 TPS). See details: https://docs.together.ai/docs/rate-limits", "type_": "rate_limit"}
Request 23 failed: Error code: 429 - {"message": "Request was rejected due to request rate limiting. Your rate limits are 600 RPM (10 QPS) and 180000 TPM (3000 TPS). See details: https://docs.together.ai/docs/rate-limits", "type_": "rate_limit"}
Request 8 failed: Error code: 429 - {"message": "Request was rejected due to request rate limiting. Your rate limits are 600 RPM (10 QPS) and 180000 TPM (3000 TPS). See details: https://docs.together.ai/docs/rate-limits", "type_": "rate_limit"}
Response 15: A woodchuck (groundhog) can’t actually chuck wood—it's a ton...
Response 7: A classic tongue‑twister doesn’t actually have a single “rig...
Response 24: Did you know that cows have best friends and get stressed wh...
Response 1: If a woodchuck could 