Use Ultrafast mode with GPT-6 Astra over WebSockets, with SDK examples and an HTTP alternative.
2 comments
ademup1 day ago
Faster inference would solve nearly all of the issues I have with AI models. I really liked working with Opus 4.8 and I would greatly prefer a 100x faster version of it, with 100x the tokens, at the same price, to any of the openai or anthropic models that have come out since.
prpl1 day ago
4.8 at 10k tokens a second would be more groundbreaking than a Opus 6 IMO.
Read the full thread on Hacker News →
Related stories
- OpenAI says planned GPT-6.1 is too insecure to releasearstechnica.comArs Technica · 0 points · 1 day ago
- GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligenceartificialanalysis.aiHacker News · 47 points · about 18 hours ago
- Hacker News · 2 points · 7 days ago
- Hacker News · 4 points · 6 days ago
- Addendum to GPT-6 Astra System Card: GPT-6.1 Soldeploymentsafety.openai.comHacker News · 3 points · 1 day ago
- GPT-6 Astra Breaks an Old Enigma Messageschneier.comHacker News · 2 points · 7 days ago