The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...
Details
| Field | Value |
|---|---|
| id | o1-pro-2025-03-19 |
| context_length | 200000 |
| max_output_tokens | 100000 |
| quantization | unknown |