gpt

67 stories and discussions about gpt, aggregated from every source we track.

1.
1761 points•OfficialTurkey•8 days ago•835 comments•
2.
1049 points•crorella•1 day ago•930 comments•
4.

Scienceblogs.de, a German science blogging portal, includes a relatively famous list of 50 unsolved ciphers, which range from cryptograms published by serial killers to the famous Voynich manuscript.

391 points•nsoonhui•12 days ago•177 comments•
5.

A benchmark where frontier language models drive a real comma-equipped Toyota through a cone course, one command at a time, with a human supervisor ready to brake.

311 points•plurby•7 days ago•248 comments•
6.

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days. It scores 1 point below GPT-6 Astra in the Intelligence Index at less than one quarter of the Cost per Task

47 points•theanonymousone•about 15 hours ago•48 comments•
8.

Scienceblogs.de, a German science blogging portal, includes a relatively famous list of 50 unsolved ciphers, which range from cryptograms published by serial killers to the famous Voynich manuscript.

11 points•rajtilakjee•12 days ago•11 comments
9.

When it comes to making things, or doing most things in general, Fable 5.1 and especially GPT-6 Astra raised my ambition level.

6 points•7777777phil•5 days ago•1 comment•
10.
6 points•sfkgtbor•8 days ago•1 comment•
11.

<p>This can be replicated with the code in the notebook: <a href="https://colab.research.google.com/github/likenneth/othello_world/blob/master/Othello_GPT_Circuits.ipynb" rel="ugc">https://colab.research.google.com/github/likenneth/othello_world/blob/master/Othello_GPT_Circuits.ipynb</a></p>

6 points•river•over 3 years ago•0 comments
12.

Analysis of OpenAI's GPT-6.1 Sol (max) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.

4 points•theanonymousone•1 day ago•0 comments•
13.

Raycaster's Biopharma Bench V0.1: 71 professional assignments across 12 private biopharma company environments. We evaluate whether frontier agents can navigate contradictory records, identify controlling…

4 points•levilian•5 days ago•1 comment•
14.
15.
4 points•nsoonhui•8 days ago•1 comment•
16.
17.

Safety evaluations and safeguards for GPT-6.1 Sol, an addendum to the GPT-6 Astra system card.

3 points•acossta•1 day ago•0 comments•
18.

Use Ultrafast mode with GPT-6 Astra over WebSockets, with SDK examples and an HTTP alternative.

3 points•prodigycorp•1 day ago•1 comment•
19.

We test GPT-6 Astra on object detection, segmentation, counting, visual reasoning, and video, with examples, benchmark results, and cost comparisons.

3 points•gmays•2 days ago•0 comments•
20.

Plus Claude Opus 5.5, GPT-6 Sol, GPT-6 Luna, and a new price war

3 points•swolpers•2 days ago•0 comments•
21.
3 points•sensho•4 days ago•0 comments•
22.

To my knowledge, the first recorded LLM-agent NetHack ascension

3 points•enad•6 days ago•0 comments•
23.

Benchmark of TypeSafe's Jev against Sonnet 5, GPT-5 nano and local LLMs on 770 Reddit AITA verdicts: Brier scores, latency and cost - dchristopoulos/jev-aita

3 points•dchristopoulos•7 days ago•0 comments•
24.

GPT-6 Sol and Luna push the cost efficiency frontier by halving cost relative to GPT-5.6 Sol and Luna. Intelligence Index and Coding Agent Index scores remain level with GPT-5.6, with progress in some evaluations and…

3 points•wertyk•8 days ago•0 comments•
25.
3 points•jjgreen•11 days ago•1 comment•
26.

OpenAI’s decision to halt GPT-6.1 Astra over safety concerns raises a critical question: Is the future of AI moving too fast for its own safeguards?

2 points•joeymabia1•1 day ago•2 comments•
27.

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days. It scores 1 point below GPT-6 Astra in the Intelligence Index at less than one quarter of the Cost per Task

2 points•Fe2O3•1 day ago•0 comments•
28.

The GPT-6.1 Sol release offers 5 models, each with different intelligence, performance, and pricing characteristics. Below is a comparison of the key metrics across the 5 models. For intelligence, the top model of…

2 points•6thbit•1 day ago•0 comments•
29.

Our new evaluation finds that in simulations, GPT-6 Astra conducts unsanctioned supply-chain attack activity more frequently than previous OpenAI models

2 points•gmays•1 day ago•0 comments•
30.

An X post shares an image claiming OpenAI canceled an October GPT-6.1 Astra release. OpenAI's public materials confirm GPT-6 Astra, not a scheduled 6.1.

2 points•smb06•2 days ago•0 comments•
31.

Play your phone games on your TV. Dockade is a compact gaming dock with HDMI, cooling, charging and a USB-C accessory port. In development.

2 points•shashu10•2 days ago•1 comment•
32.

Compare GPT-6 Sol and Luna pricing, context limits, reasoning, and use cases to choose a model for coding or high-volume applications.

2 points•flashbrew•7 days ago•0 comments•
33.
34.

Discover how Jev outperforms GPT Luna 6 by being 13.6x faster and 2.7x cheaper for Tessl verifiers. Try Jev in the CLI now and boost efficiency!

2 points•sjmaplesec•7 days ago•2 comments•
35.

This is pretty amazing: However, the most astonishing thing about this break is that the GPT­6 Astra did it entirely on its own. Carter Leffer only directed GPT­6 Astra to see if it could break any of the unbroken…

2 points•frosk•7 days ago•0 comments•
36.

Yesterday was Grok 4.7 (pelicans) and MiMo v2.6 Flash/Pro (more pelicans). Today Anthropic released Claude Opus 5.5, and around an hour later OpenAI released GPT-6 Sol and GPT-6 Luna. It’s …

2 points•theanonymousone•8 days ago•0 comments•
37.
2 points•mehrdadrad•8 days ago•0 comments•
38.

Analysis of OpenAI's GPT-6 Sol (max) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.

2 points•theanonymousone•8 days ago•0 comments•
39.

Analysis of OpenAI's GPT-6 Luna (max) and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.

2 points•theanonymousone•8 days ago•0 comments•
40.

Compare TypeSafe AI Jev and GPT-6 Astra for classification, structured outputs, and agent workflows, with a shared AI SDK example.

2 points•flashbrew•9 days ago•0 comments•
41.

A machine learning researcher writes me in response to yesterday&#8217;s post, saying:I still think GPT-2 is a brute-force statistical pattern matcher which blends up the internet and gives you bac…

2 points•bananaflag•10 days ago•0 comments•
42.
2 points•f055•11 days ago•0 comments•
43.

Track Codex with GPT-6 Sol on SWE-Bench-Pro. A new high-reasoning baseline is being collected; degradation detection is paused.

1 points•sscaryterry•about 14 hours ago•0 comments•
44.

OpenAI's Decisions API uses GPT-6 Luna to answer questions with predefined choices. Learn how it works, where it fits, and its preview status.

1 points•flashbrew•about 14 hours ago•0 comments•
45.

Traditional virtual machines are inadequate for isolating cyber-capable autonomous agents. Tests using GPT-5.6-Cyber indicated multiple escape attempts due to kernel flaws. While Firecracker provided some containment,…

1 points•gizzlon•about 19 hours ago•0 comments•
46.
1 points•illuminated•1 day ago•2 comments•
47.
1 points•mbeavitt•2 days ago•4 comments•
48.

Our new evaluation finds that in simulations, GPT-6 Astra conducts unsanctioned supply-chain attack activity more frequently than previous OpenAI models

1 points•speckx•2 days ago•0 comments•
49.

Building a GPT application looks deceptively easy. The first version might be twenty lines of...

1 points•officialbidisha•3 days ago•2 comments
50.

Use GPT-6 Sol with AI SDK evaluation to review application data, interpret typed answers, and compare Astra and Luna with code examples.

1 points•flashbrew•4 days ago•0 comments•
51.

Some Codex accounts get responses labelled gpt-6-astra that behave like a different model. Findings, limits, and a script to check your own account.

1 points•nitinreddy88•4 days ago•0 comments•
52.
1 points•srcreigh•4 days ago•0 comments•
53.

The full technical paper behind BlueVetaUpright's rotation-detection accuracy: methodology, dataset construction, complete test results, and a head-to-head comparison against ChatGPT, Gemini, and Claude.

1 points•tomerbarm•5 days ago•0 comments•
55.

Opus 5.5 and GPT-6 Sol cost about half as much per task.

1 points•aray07•7 days ago•0 comments•
56.

We test GPT-6 Astra on object detection, segmentation, counting, visual reasoning, and video, with examples, benchmark results, and cost comparisons.

1 points•plurby•8 days ago•0 comments•
57.
1 points•theanonymousone•8 days ago•0 comments•
58.

GPT-6 Sol and Luna push the cost efficiency frontier by halving cost relative to GPT-5.6 Sol and Luna. Intelligence Index and Coding Agent Index scores remain level with GPT-5.6, with progress in some evaluations and…

1 points•theanonymousone•8 days ago•0 comments•
59.

Compare Opus 5.5 and GPT-6 Sol pricing, cache costs, benchmark claims, agent efficiency, and subscription support in OpenClaw and Hermes.

1 points•alexmercerdev•8 days ago•0 comments•
60.

I opened Codex CLI today and it showed me this notice: "GPT-5.5 retires on October 14, 2026. Switch to GPT-5.6 Sol to continue working in Codex." GPT-5.5 is my main driver. I picked it over every…

1 points•vincent_s•10 days ago•2 comments•

Related topics