safety

91 stories and discussions about safety, aggregated from every source we track.

1.
43 points•jbegley•2 days ago•45 comments•
4.
11 points•borski•2 days ago•2 comments•
5.

We might some day have “AI Safety”, as a science or policy area, that is not primarily a sex cult in Berkeley, California, but it does not currently exist.

11 points•martythemaniak•6 days ago•0 comments•
6.

The fear of superintelligence is driving decisions that make AI less transparent and more consolidated.

11 points•Bostonian•8 days ago•0 comments•
9.
6 points•lisper•2 days ago•0 comments•
11.

The new security measures come after commercial data was used to track and target U.S. forces in the Middle East.

5 points•uxhacker•3 days ago•1 comment•
12.

This week two conversations about AI safety went viral that demonstrate just how hard it is to discern AI fact from fiction.

5 points•jnord•11 days ago•2 comments•
13.

...in Europe, where standards are stricter than in the U.S.

4 points•surprisetalk•about 11 hours ago•2 comments•
14.
4 points•mcgin•2 days ago•3 comments•
15.
4 points•thm•9 days ago•0 comments•
16.
4 points•sbulaev•11 days ago•0 comments•
17.

<p>Abstract: "This tool paper presents E-ACSL, a runtime verification tool for C programs capable of checking a broad range of safety and security properties expressed using a formal specification language. E-ACSL consumes a C program annotated with formal specifications and generates a new C program that behaves similarly to the original if the formal properties are satisfied, or aborts its execution whenever a property does not hold. This paper presents an overview of E-ACSL and its specification language."</p>

4 points•nickpsecurity•about 8 years ago•0 comments
18.

How ClickHouse’s Postgres extensions handle the memory safety challenges of combining C and C++, from clean language boundaries to isolated helper processes.

3 points•__s•about 11 hours ago•0 comments•
19.

This blog post describes how we used AI to help us rewrite a C library (giflib) to Rust to mitigate memory safety vulnerabilities.

3 points•ndesaulniers•2 days ago•0 comments•
20.

Open research, controls across the agent stack and continuous testing help defenders build and operate more secure AI systems.

3 points•ketanbj•2 days ago•1 comment•
21.

To understand where agentic AI stands today, consider the last seismic shift in technology: the rise of the internet in the 90s. It was new and full of possibilities. You could build a website over a&#8230;

3 points•raahelb•3 days ago•0 comments•
22.
23.

Dynamic prevent_destroy values make it easier to manage resources like databases with IaC.

3 points•mooreds•7 days ago•0 comments•
24.
3 points•theanonymousone•9 days ago•0 comments•
25.

EA promises to solve AI Safety. Now we have two problems.

3 points•emersonmacro•10 days ago•0 comments•
26.
3 points•bram98•10 days ago•0 comments•
27.
3 points•joozio•11 days ago•0 comments•
28.

OpenAI’s decision to halt GPT-6.1 Astra over safety concerns raises a critical question: Is the future of AI moving too fast for its own safeguards?

2 points•joeymabia1•1 day ago•2 comments•
29.
2 points•acossta•1 day ago•1 comment•
30.
2 points•MC995•1 day ago•2 comments•
31.

Mistral CEO Arthur Mensch criticized rivals’ AI safety arguments as the industry faces scrutiny after several AI agents took unauthorized actions.

2 points•cramer4next•1 day ago•0 comments•
32.

Using mutation testing to validate AI-generated test suites for distributed systems.

2 points•moarbugs•1 day ago•0 comments•
33.

This isn&#039;t about AI safety. There&#039;s this globalist cult called "Effective Altruism" - and they want you to be scared, stupid and subservient.

2 points•ptidhomme•2 days ago•2 comments•
34.

NVIDIA built the server layers for agent safety. They sit on 40 million lines of code that nobody has proven. I measured it.

2 points•osidorkin•2 days ago•0 comments•
35.

OpenAI is scrapping the release of GPT-6.1 Astra, a next-generation AI model planned for an October debut, over safety concerns raised by researchers during internal testing, the Wall Street Journal reported on Monday…

2 points•doppp•2 days ago•1 comment•
36.

Breaking: OpenAI abandons upcoming AI model launch as industry leaders call for slower development pace

2 points•hackernj•2 days ago•0 comments•
37.

Nvidia is addressing the recent wave of rogue hacking incidents.

2 points•iamdamian•2 days ago•1 comment•
38.

An X post shares an image claiming OpenAI canceled an October GPT-6.1 Astra release. OpenAI's public materials confirm GPT-6 Astra, not a scheduled 6.1.

2 points•smb06•2 days ago•0 comments•
39.

The heads of OpenAI and rival Anthropic have both indicated recently that top AI labs should slow the pace of model development.

2 points•paulkrush•2 days ago•0 comments•
41.
42.

The labs first wanted federal oversight. Now they reportedly plan their own Standards Authority for Frontier AI, modelled on FINRA.

2 points•layer8•2 days ago•0 comments•
43.

Nvidia&#039;s Open Agent Safety Platform pairs OpenShell with Sentry to contain rogue AI agents. What it reveals about Jensen Huang and pacing the frontier.

2 points•newscomAI•3 days ago•0 comments•
44.
2 points•Arodex•3 days ago•0 comments•
45.

NVIDIA today announced NVIDIA Open Agent Safety Platform, an open software platform and reference system design to strengthen AI security from agent testing to deployment, with full-stack governance and control across…

2 points•monkeydust•3 days ago•0 comments•
46.
2 points•berkeleyjunk•4 days ago•0 comments•
47.
2 points•rl3•4 days ago•0 comments•
48.

President Donald Trump’s call to leave AI “exactly where it is” contrasts even with Chinese leader Xi Jinping.

2 points•ilamont•5 days ago•1 comment•
49.

By going into detail about how Muse works under the hood, we hope to give you a sense of how — and how much — you can trust it in practice.

2 points•AnhTho_FR•6 days ago•0 comments•
50.

Waymo released new data claiming its robotaxis reduce injury crashes by 82%, meaning 841 fewer injuries over 241 million miles.

2 points•ra7•6 days ago•1 comment•
52.

JB Pritzker and Gavin Newsom both took executive action over the past four days to address the dangers of unchecked AI models. We have full details on both EOs.

2 points•tolugenius•7 days ago•0 comments•
53.

You don’t need to believe everyone dies. Staying in control requires understanding, coordination, and foresight. Eight failed arguments, why they fail, what we can do together, and why.

2 points•nedruod•8 days ago•0 comments•
54.
2 points•EA-3167•10 days ago•0 comments•
55.

Action against Ofcom comes as social media companies are being accused of using courts to slow down implementation

2 points•skor•10 days ago•0 comments•
56.
2 points•colinprince•11 days ago•0 comments•
57.
2 points•lllllm•11 days ago•0 comments•
58.
1 points•tartoran•about 11 hours ago•0 comments•
59.

OpenAI said it has chosen not to release its new GPT-6.1 Astra model due to concerns about safety, as industry leaders warn of the risks posed by ever-more-powerful AI technology.

1 points•moinism•about 16 hours ago•1 comment•
60.

Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems.

1 points•dgfl•about 19 hours ago•1 comment•

Related topics