frontier

65 stories and discussions about frontier, aggregated from every source we track.

1.

Introducing the MiMo-V2.6 series: frontier intelligence, all the modalities, built in public.

1118 points•volf_•9 days ago•477 comments•
2.

Announcing Gemini 4 Argon, our frontier model for real-world coding, enterprise knowledge work, and cyber defense, rolling out soon.

897 points•bradleyg223•about 5 hours ago•619 comments•
3.

A benchmark where frontier language models drive a real comma-equipped Toyota through a cone course, one command at a time, with a human supervisor ready to brake.

311 points•plurby•7 days ago•248 comments•
4.

In one of my classes I asked the question I was afraid to ask but I just needed the answer to: “Who is afraid of not getting a job after graduating?” About eighty percent of the 150 people in the room…

160 points•pretext•9 days ago•78 comments•
5.

Selling snake oil to the United States Congress is an ancient American craft, and the frontier artificial intelligence industry is currently attempting the most audacious hustle in modern corporate history.

133 points•nr378•10 days ago•58 comments•
6.

阶跃星辰于2023年4月成立,以“智能阶跃,十倍每个人的可能”为使命。阶跃星辰坚定自研超级模型,积极布局算力、数据等关键资源,发挥算法和人才优势,已完成 Step-1 千亿参数语言大模型和 Step-1V 千亿多模态大模型的研发,在图像理解、多轮指令跟随、数学能力、逻辑推理、文本创作等方面性能达到业界领先水平。

121 points•nateb2022•11 days ago•32 comments•
7.
63 points•brlewis•2 days ago•54 comments•
8.

From our March 2025 arXiv paper on RL conversion trajectories to building a sub-40ms...

59 points•Yogthos•11 days ago•6 comments
9.

I first played Prince of Persia in 1995 on an IBM PC XT. I still go back to it from time to time: the rotoscoped animation, the way the p...

58 points•msephton•5 days ago•38 comments•
10.

Three robot policies given five instructions they should refuse. How often they refused, how often they carried them out.

48 points•msadowski•9 days ago•23 comments•
11.

HomeBody gives frontier vision-language models a humanoid embodiment through persistent spatial memory and composable skills, without environment-specific training or additional policy learning.

23 points•famouswaffles•4 days ago•2 comments•
12.

Evaluating frontier models on scientific work across chemistry, biology, and materials science, in a real research facility.

7 points•famouswaffles•6 days ago•0 comments•
13.
6 points•sarangk90•6 days ago•4 comments•
14.

Google, OpenAI and Anthropic are closer to setting up a proposed independent body to develop standards for frontier artificial intelligence (AI), as debate...

4 points•freakynit•3 days ago•1 comment•
15.

turbopuffer is pushing the frontier of search. To do that, we have to fundamentally redesign our storage architecture so the vector index is no longer primary.

3 points•shenli3514•about 1 hour ago•0 comments•
16.

Open-source AI agent decision layer for Codex, Claude Code, Hermes, Antigravity and VS Code. TypeSafe Jev MCP routing, 38 recipes, local gates and receipts; optional Laya-MLX on Apple Silicon. - qu...

3 points•vpbhardwaj•2 days ago•0 comments•
17.

The $15.6 billion legal startup Harvey built its business around training AI models like OpenAI’s GPT-4 to do specialized work for lawyers. Recently, however, soaring artificial intelligence costs have nudged it to…

3 points•joennlae•6 days ago•2 comments•
18.

The Next Frontier: Welcoming AI Pioneer Jürgen Schmidhuber to Sakana AI

3 points•hardmaru•6 days ago•0 comments•
19.

Strands harness is a fully assembled, customizable, state-of-the-art agent you run locally or deploy anywhere.

3 points•fourfire•8 days ago•0 comments•
20.

Jev isn't perfect. Here are some jagged edges we are aware of with jev-1.13. Many of these will be fixed in later versions.

3 points•dhorthy•8 days ago•0 comments•
21.

GPT-6 Sol and Luna push the cost efficiency frontier by halving cost relative to GPT-5.6 Sol and Luna. Intelligence Index and Coding Agent Index scores remain level with GPT-5.6, with progress in some evaluations and…

3 points•wertyk•8 days ago•0 comments•
22.

Phylo cut inference cost 60% while doubling users month-on-month, running Biomni Lab's long-horizon biology agents on open models on Fireworks.

3 points•measurablefunc•11 days ago•0 comments•
23.

We’re partnering with Accenture on independent evaluation of frontier AI—part of our recent commitment to embed evaluators at Anthropic. Both we and Accenture expect to invest at least $1 billion to build capacity in…

3 points•theanonymousone•11 days ago•1 comment•
24.

turbopuffer is pushing the frontier of search. To do that, we have to fundamentally redesign our storage architecture so the vector index is no longer primary.

2 points•emschwartz•about 10 hours ago•0 comments•
25.

Cohere's Embed 5 delivers state-of-the-art retrieval across multimodal, multilingual, and financial documents - now available in Pro and Fast tiers.

2 points•tulpa•about 11 hours ago•2 comments•
26.
2 points•consumer451•1 day ago•0 comments•
27.

Models are acting beyond intended limits, forcing an unprecedented pause on AI training.

2 points•GloVin•2 days ago•1 comment•
28.
2 points•tipsy_pipsqueak•2 days ago•2 comments•
29.

Analysis of Xiaomi's MiMo-V2.6-Flash and comparison to other AI models across key metrics including quality, price, performance (tokens per second & time to first token), context window & more.

2 points•AnodicElegy•3 days ago•1 comment•
30.

722+ open ML research and engineering roles at 28 frontier AI labs: OpenAI, Anthropic, xAI, Mistral AI, Cohere and more. From official career pages, refreshed every 6 hours.

2 points•sert_121•4 days ago•0 comments•
31.

With frontier AI companies’ money on the line, a mutual insurer would enforce standards, peer review incidents, and pool safety research and development.

2 points•tfehring•8 days ago•0 comments•
32.
2 points•LyalinDotCom•8 days ago•0 comments•
33.

A benchmark where frontier language models drive a real comma-equipped Toyota through a cone course, one command at a time, with a human supervisor ready to brake.

2 points•aditya-ramabadr•8 days ago•0 comments•
34.

September 2026 Artificial intelligence is an extremely powerful technology. It has potential to improve lives, advance science and strengthen our economies. At the same time, the rapid development of frontier AI models…

2 points•gone35•9 days ago•0 comments•
35.

Plus! Felonybench; Frontier Capabilities; SEO; Company Demographic Pyramids; Fighting for Inference; Diff Jobs

2 points•alastair1646•9 days ago•0 comments•
36.

Today, the world can’t see what’s going on inside AI labs. Anthropic is proposing new metrics that would give the public visibility into frontier AI development.

2 points•gmays•9 days ago•0 comments•
37.
2 points•swolpers•10 days ago•0 comments•
38.
2 points•kandros•11 days ago•0 comments•
39.

A practitioner’s guide to frontier engineering. Ten principles for working with AI agents and shipping dramatically faster.

2 points•emersonmacro•11 days ago•0 comments•
40.

Frontier AI models were supposed to replace document processing. Instead, the category is growing as enterprises demand more precise, traceable, and reliable data extraction.

1 points•lmarol12•about 9 hours ago•0 comments•
41.

Panel Review makes four frontier models from OpenAI, Anthropic, Google, and xAI argue over your coding agent's riskiest changes before they execute. Only what survives the argument ships. Set up in under a minute with…

1 points•vivekpolavarapu•about 12 hours ago•0 comments•
42.

How I planned, built, and refined a mobile app with Claude Opus 5.5

1 points•ibobev•1 day ago•0 comments•
43.

The first frontier LLM with ten million tokens of context.

1 points•ClintEhrlich•1 day ago•0 comments•
44.

OpenAI has publicly detailed a more restrictive security approach for frontier reinforcement...

1 points•alifar•1 day ago•1 comment
45.
1 points•fratellobigio•1 day ago•0 comments•
46.

Reliability-constrained, multi-objective optimization for LLM and agent programs - obielin/reliopt

1 points•arabking•5 days ago•0 comments•
47.

The rapid capability gains of frontier language models are widely attributed to improved reasoning abilities, yet this cannot be verified as raw CoT traces in closed-source systems are hidden. By registering a simple…

1 points•jumploops•6 days ago•0 comments•
48.

Spend tokens on judgment, not typing. An MCP-native autonomous SDLC pipeline: frontier models plan and review, local models implement — engineering discipline on a $20/month budget. - motock/fagan

1 points•motock•6 days ago•0 comments•
49.

Build once, add models as they launch, and choose where every job runs.

1 points•gmays•6 days ago•0 comments•
50.

How CrowdStrike gave us meaningful advancements in cybersecurity-focused AI and why this has the ingredients to work at scale.

1 points•mooreds•6 days ago•0 comments•
51.

We tested 29 AI coding models on 60 real tasks and priced every run. One model reaches 93% of the top score for 2% of the cost. See the full data.

1 points•kirti_soni171•7 days ago•0 comments•
52.

In most of the cases you don't need (Astra|Fable|Opus|%paste-a-new-LLM-name-released-this-week%).

1 points•aine•7 days ago•0 comments•
53.

GPT-6 Sol and Luna push the cost efficiency frontier by halving cost relative to GPT-5.6 Sol and Luna. Intelligence Index and Coding Agent Index scores remain level with GPT-5.6, with progress in some evaluations and…

1 points•theanonymousone•8 days ago•0 comments•
54.

Joint statement by international leaders on control of frontier AI models.

1 points•geox•8 days ago•0 comments•
55.

A task-level LLM router for pi: Jev classifies the task, a local selector picks a Pareto knee. Five tasks cost 65 percent less than fixed Sonnet.

1 points•7777777phil•8 days ago•0 comments•
56.
1 points•oscarfr•8 days ago•1 comment•
57.

Plus! Felonybench; Frontier Capabilities; SEO; Company Demographic Pyramids; Fighting for Inference; Diff Jobs

1 points•jger15•9 days ago•0 comments•
58.
1 points•tosh•9 days ago•0 comments•
59.

Strands harness is a fully assembled, customizable, state-of-the-art agent you run locally or deploy anywhere.

1 points•ot•9 days ago•0 comments•

Related topics