Depends on the model you’re using. Claude 3.5 Sonnet has the biggest input context window. Albeit it tends to give more concise responses, especially with larger input context (from experience). Summarised for you below:
| Model | Input Context Window | Maximum Output Tokens |
|---|---|---|
| o1-mini | 128K tokens | 65.5K tokens |
| o1 | 128K tokens | 65.5K tokens |
| Claude 3.5 Sonnet | 200K tokens | 8,192 tokens |
| GPT-4o | 128K tokens | 16.4K tokens |
[ 1,000 tokens ≈ 750 words , though this depends on the specific text]