Faster inference would solve nearly all of the issues I have with AI models. I really liked working with Opus 4.8 and I would greatly prefer a 100x faster version of it, with 100x the tokens, at the same price, to any of the openai or anthropic models that have come out since.
Faster inference would solve nearly all of the issues I have with AI models. I really liked working with Opus 4.8 and I would greatly prefer a 100x faster version of it, with 100x the tokens, at the same price, to any of the openai or anthropic models that have come out since.
4.8 at 10k tokens a second would be more groundbreaking than a Opus 6 IMO.