models.rip · plot reserved
GPT-4 Turbo
2024–present · Plot reserved · 2 years, 4 months so far
A hundred and twenty-eight thousand tokens of context at a third of GPT-4's input price and half its output, which is why so much production code was written against this exact model string and never moved. Its shutdown has a date on it. The endpoint is still answering.