
OpenAI's GPT-6 Astra Is Shockingly Good at Almost Everything
Early testers praise OpenAI's newly launched GPT-6 Astra model following its launch weekend, highlighting its impressive capabilities across 3D cities, playable…

Early testers praise OpenAI's newly launched GPT-6 Astra model following its launch weekend, highlighting its impressive capabilities across 3D cities, playable…
1 editorial report · 3 verified social mentions. The most authoritative report leads while later evidence completes the story.
People are now using Claude Code and Codex, two of the leading coding agents, to do almost everything, including tasks that have more to do with language than coding, such as negotiation. For example, OpenAI recently highlighted a use case where Codex negotiated with customer service to get a refund on behalf of a user . But can you really trust an agent to represent your best interests? And if so, which agent should you trust? That actually raises an interesting question. We already have countless benchmarks for LLMs, but surprisingly, we still don't have one that systematically measures how well they negotiate using the very thing they're built around: language. There's only one way to find out. The same way we evaluate human negotiators. Put them in a negotiation competition with carefu
Open mentionGPT-6 Astra Changes Everything
Open mentionPrivacy have been concern of many of us to have their own hardware to run llms, and here's another reason why: two mathematicians spent a year cracking one of the hardest problems in math and fed every draft of their works into Codex. A few days before they could publish, OpenAI suddenly showed up with the same solutions. When asked if their model (Sol and Astra) was trained on the pair's private chats, OpenAI did not answer the question. Full statement from them https://cims.nyu.edu/~tristanb/statement.pdf Feels like big labs believe everything you did with the help of their models is theirs. submitted by /u/bakawolf123 [link] [comments]
Open mention