Humanlike AI
All models
OpenAI logo

GPT-6

text model · Announced

OpenAI

Frontier reasoning with an effort dial, from instant to deliberate.

GPT-6Idle · sample soon

Announced model. All pricing, benchmarks and evals are estimates, replace before this page is indexed.

Fit and price

Recruitment fit

93

Per hour

$0.698

Per minute

$0.012

30-min interview

$0.349

About

GPT-6 is the current frontier of the GPT line. The headline change over GPT-5 is not raw benchmark score but controllability: a single effort parameter moves the same model from sub-second replies to multi-minute deliberate reasoning, so one integration covers both a live chat widget and an overnight batch job. For hiring teams that matters because screening and structured interview scoring have opposite latency budgets, and you no longer need two models to serve them.

  • Reasoning effort control
  • Tool use
  • Structured outputs
  • Vision
  • Streaming
  • Prompt caching
  • 90+ languages

Strengths and watch-outs

Strengths

  • Full transcript plus job spec in context
  • One integration, effort dial
  • ATS-ready structured outputs

Watch out for

  • 24× the cost of Mini at max
  • Still needs audited rubrics
  • Unpublished pricing

Best for Rubric-based interview scoring · Shortlist vs job spec · Draft candidate feedback

5 variants

  • OpenAI logo

    Max effortMax

    Longest deliberation. Reserve it for work where a wrong answer is expensive, final scoring, offer-risk analysis, contract review.

    42.0s

    latency

    24×

    rel. cost

    98

    quality

  • OpenAI logo

    HighHigh

    The default for anything a human will act on. Strong multi-step reasoning without the max-effort latency tax.

    9.0s

    latency

    8×

    rel. cost

    94

    quality

  • OpenAI logo

    MediumBalanced

    The workhorse. Good judgement at a price that survives volume, this is the rung most production traffic should sit on.

    3.2s

    latency

    3×

    rel. cost

    88

    quality

  • OpenAI logo

    LowFast

    Near-instant. Skips deliberation entirely, right for classification, routing and anything a user is waiting on.

    900ms

    latency

    1.4×

    rel. cost

    79

    quality

  • OpenAI logo

    MiniMini

    Smaller sibling. The cheapest way to run GPT-6-shaped prompts when the task is mechanical rather than judgemental.

    600ms

    latency

    1×

    rel. cost

    71

    quality

The numbers

HumanLike evals

CV field extraction (F1)
0.94Estimate

Projection. Not yet run against GPT-6.

Time to first token (medium effort)
640 msEstimate

Compare with