> ## Documentation Index
> Fetch the complete documentation index at: https://docs.getimpala.ai/llms.txt
> Use this file to discover all available pages before exploring further.

# Quickstart

> Get your endpoint and key, then send your first request.

## 1. Get your endpoint and key

Sign up at [platform.getimpala.ai](https://platform.getimpala.ai) or talk to your Impala contact to get a serverless endpoint with the details below.

| Value      | What it is                                                                                                                      |
| ---------- | ------------------------------------------------------------------------------------------------------------------------------- |
| `BASE_URL` | Your inference endpoint. Endpoints differ between accounts — use yours exactly as provided, without adding or stripping a path. |
| `MODEL`    | The model your endpoint is provisioned for                                                                                      |
| `API_KEY`  | Your account's API key. Send it as a bearer token on every request.                                                             |

You can create more keys under **API Keys** on the platform.

## 2. Install the SDK

Impala is OpenAI-compatible, so use the official `openai` package.

```bash theme={null}
pip install openai
```

## 3. Send your first request

Fill in your three values and run it.

<CodeGroup>
  ```python Python theme={null}
  from openai import OpenAI

  client = OpenAI(
      base_url="<BASE_URL>",
      api_key="<API_KEY>",
      timeout=600.0,
  )

  resp = client.chat.completions.create(
      model="<MODEL>",
      messages=[
          {"role": "system", "content": "You are a concise assistant."},
          {"role": "user", "content": "What is 17 times 3? Answer with just the number."},
      ],
  )
  print(resp.choices[0].message.content)
  ```

  ```bash cURL theme={null}
  curl "<BASE_URL>/chat/completions" \
    -H "Authorization: Bearer <API_KEY>" \
    -H "Content-Type: application/json" \
    --max-time 600 \
    -d '{
      "model": "<MODEL>",
      "messages": [{"role": "user", "content": "What is 17 times 3?"}]
    }'
  ```
</CodeGroup>

## 4. Check it worked

A response with no error means your endpoint, key and model are working end to end.

## Next

<CardGroup cols={2}>
  <Card title="Chat completions" icon="bolt" href="/send-a-request">
    The request shape, streaming, timeouts and parameters.
  </Card>

  <Card title="Run a batch" icon="layer-group" href="/run-your-first-batch">
    File in, results out, at the lowest cost per token.
  </Card>
</CardGroup>

Bring Your Own Cloud installs differently — see [Install Impala in your VPC](/getting-started-infrastructure-setup).
