models.rip · plot reserved

GPT-4 Turbo

2024–present · Plot reserved · 2 years, 4 months so far

A hundred and twenty-eight thousand tokens of context at a third of GPT-4's input price and half its output, which is why so much production code was written against this exact model string and never moved. Its shutdown has a date on it. The endpoint is still answering.