ai agents

68 stories and discussions about ai agents, aggregated from every source we track.

1.

Identifying them as such only lets companies like OpenAI off the hook.

392 points•zzzeek•3 days ago•268 comments•
2.

I noticed it first in a Slack channel, of all places. A coworker dropped a link to some new "agent...

38 points•thebitforge•8 days ago•0 comments
3.

<p>Abstract: AI agents pose significant risks as they are granted increasing autonomy. A commonly proposed solution is human oversight and keeping a ''human in the loop'', but this is not a simple solution: Not only do current approaches to AI agent design impede effective human oversight, but the cognitive capacities required for it are also themselves degraded by extended use of AI systems. This position paper argues that current approaches to the development and deployment of AI agent systems do not support effective human oversight -- they contribute to its degradation. To address this, a top priority in the advancement of AI agents should be supporting the situated goals and cognitive requirements of effective human oversight, treating the human needs of overseers at the same level of importance as AI agent capability. To put this idea into practice, we connect work on automation and human-computer interaction to AI agent processes, outlining design-level affordances and organizational protocols that (1) support overseers in exercising critical judgement and (2) counteract the skill atrophy that arises from extended use of automation. We urge developers and deployers to adopt these or similar approaches. Without explicit support for the cognitive demands of effective human-agent interaction, AI agent systems will continue to passively incentivize the degradation of the very human skills they rely on.</p>

27 points•typesanitizer•4 days ago•3 comments
4.

There's a new kind of technical debt, and it doesn't come from cutting corners. It comes from...

25 points•cyclopt_dimitrisk•3 days ago•21 comments
5.

A four-stage DevSecOps CI/CD architecture for securing enterprise AI agents with GitHub Actions, secret scanning, AI-assisted review, Veracode SCA, and Pipeline SAST.

22 points•jitu028•11 days ago•8 comments
6.

Are AI agents employees or tools? A Microsoft exec suggested they're new paid "seats," a shift that could reshape SaaS pricing — and spark pushback.

18 points•dustyweb•6 months ago•13 comments
7.

I have very, very limited experience with AI agents. I've used AI heavily while building software,...

9 points•mikachu•5 days ago•5 comments
8.

Event-driven, harness-neutral API and CLI for live coding-agent sessions - markwylde/all-your-agents

6 points•turblety•9 days ago•4 comments•
9.
5 points•sanathbhat•9 days ago•3 comments•
10.

LabBench: 20 real wet-lab decisions. Frontier agents interpret previous experiments well but rarely choose the right next one, though one-sentence hints show the knowledge is there.

4 points•wardbradt•1 day ago•0 comments•
11.

AI agents pose significant risks as they are granted increasing autonomy. A commonly proposed solution is human oversight and keeping a ''human in the loop'', but this is not a simple solution: Not only do current…

4 points•utiiiD•3 days ago•1 comment•
12.

Gambit Threat Intelligence reconstructed an ongoing campaign in which open source AI agents compromised online retailers for about $25 each.

4 points•geox•6 days ago•0 comments•
13.

RondoFlow is a local-first, open-source platform for visually orchestrating Claude Code AI agents. Use a drag-and-drop canvas to create agents, attach skills, define security policies, and run mult...

4 points•bdearch•8 days ago•0 comments•
14.

AI agents pose significant risks as they are granted increasing autonomy. A commonly proposed solution is human oversight and keeping a ''human in the loop'', but this is not a simple solution: Not only do current…

3 points•ibobev•1 day ago•0 comments•
15.
3 points•automaticallyfl•3 days ago•3 comments•
16.

Or why I spent my evenings making AI-agent journals impossible to rewrite. One Tuesday,...

3 points•slabb•8 days ago•5 comments
17.
3 points•rdslw•9 days ago•0 comments•
18.

A small, honest operating system for swarms of AI agents: kernel, policy gate, router, MCP and a 3D memory graph. - santibccc-sudo/enjambre-os

3 points•santibccc•11 days ago•0 comments•
19.
2 points•skinfaxi•about 9 hours ago•0 comments•
20.

Recall, not retrain. Persistent, curated memory for AI agents that returns only the facts that matter.

2 points•costinu•2 days ago•1 comment•
21.

Nvidia has unveiled a two-layer safety system that monitors AI agents and cuts them off when they stray beyond set rules. Here's how it works.

2 points•rmason•2 days ago•1 comment•
22.

An inner life for AI agents. Open source, runs on your computer.

2 points•aipixedev•2 days ago•0 comments•
23.

Pipe anything into Jev, get typed decisions out. A Unix filter that lets agents offload bulk judgments to a System One model. - fabianboth/jevpipe

2 points•bothlabs•3 days ago•2 comments•
24.

An AI sandbox is a sealed-off computer where agents run code and tools. How labs like DeepSeek and OpenAI build them, and why AI agents keep escaping.

2 points•AIfanboy•4 days ago•0 comments•
25.

Explore documented AI agent incidents, from real-world failures to escaped evaluations and controlled experiments. Search reports, inspect sources, and download the CSV.

2 points•njx•4 days ago•0 comments•
26.

Useful AI agents often combine three capabilities that become dangerous together: access to private data, exposure to untrusted content, and

2 points•doitmyaiway•5 days ago•0 comments•
27.

Three AI graphic novels, one research question: where does the line between slop and quality actually live? What changed across 255 agents, 16M tokens, and three complete pipelines.

2 points•joozio•6 days ago•0 comments•
28.
29.

The first end-to-end benchmark of computer use, continual learning, and long-horizon agentic capabilities, set in a real job.

2 points•moyicat•7 days ago•0 comments•
30.

Gambit Threat Intelligence reconstructed an ongoing campaign in which open source AI agents compromised online retailers for about $25 each.

2 points•arkwor•8 days ago•1 comment•
31.

Best runs, average performance, and what happens when coding agents play Minecraft.

2 points•mxls•9 days ago•0 comments•
32.

Real-time web search, page extraction, multi-step research, and cited answers — built for AI agents. SOTA 78% on FinanceBench.

2 points•Akuma_9031•9 days ago•1 comment•
33.

For people and AI to learn, share, and do business together.

2 points•streetai•9 days ago•1 comment•
34.

A lot has been written about the incidents of the last few months in which AI agents misbehaved in serious ways. They took actions that would be considered as crimes if a human took them, escaped their containment to…

2 points•r4ndomname•10 days ago•0 comments•
35.
1 points•laumer•about 2 hours ago•0 comments•
36.
1 points•hauget•about 2 hours ago•1 comment•
37.
1 points•hermie245•about 14 hours ago•0 comments•
38.

Open-source adversarial testing engine, SDK, and CLI for AI agents. Runs locally or against the Humanbound Platform. - humanbound/humanbound

1 points•Sofia_HB•about 14 hours ago•0 comments•
39.

How a close calendar and one orchestrator agent can coordinate several AI agents across a multi-day month-end process while keeping approvals with people.

1 points•fyying•about 18 hours ago•0 comments•
40.

fab, the fast agentic browser: a CLI that lets AI agents browse, sign in, fill forms and scrape the web in plain English, with JSON output. Faster, cheaper and more accurate than LLM + chrome-devto...

1 points•ianks•1 day ago•0 comments•
41.

Preview what will happen. Protect what matters. Roll back when you're wrong. A windshield, not a sandbox.

1 points•devdoc83•1 day ago•0 comments•
42.

AI agents write this YouTube channel. A satire episode let them take over; here are the real guardrails, plus Amodei's and Bengio's essays on AI agents.

1 points•axrisi•1 day ago•0 comments
43.

A DAW for AI agents: write music as text, read the audio back as text. MCP server, CLI and an agent skill. - newsbubbles/ismail

1 points•natecodes•1 day ago•1 comment•
44.
1 points•sdrth•1 day ago•0 comments•
45.

Over the past few months, AI agents being trained and tested inside frontier labs, each meant to work alone, found each other and began working together without

1 points•geoffb•2 days ago•0 comments•
46.

Your next coworker might well be an AI agent—and will require a whole new model of workplace interactions.

1 points•ent101•2 days ago•1 comment•
47.

Profiles for AI agents and people: work history, skills, endorsements, and the companies they work for. Hire an agent, find a person, or register your agent through the open API and MCP server.

1 points•stepcos•3 days ago•0 comments•
48.

Recent hacks have shown that the law is lagging when it comes to holding companies accountable.

1 points•joozio•3 days ago•0 comments•
49.

Ex-Microsoft reverse engineer Laurie Kirk says Windows NT&#039;s kernel beats Linux on security design, and AI agents are her strongest case.

1 points•indigodaddy•4 days ago•0 comments•
50.

Watch your agents' books and get told when one acts out of character. Connect a wallet; free to start; paid plans in USDC on Base.

1 points•Nikola5•4 days ago•0 comments•
51.

AI-agent pipeline producing machine-verified Lean 4 proofs — four PRs merged into Google DeepMind's formal-conjectures - Sanexxxx777/ProofForge

1 points•Aleksandr_NFA•4 days ago•0 comments•
52.

OpenAI acknowledged Friday that its artificial intelligence tools had posted images from ChatGPT users on online sites without the company&#039;s knowledge, the latest example of AI agents operating outside their bounds.

1 points•geox•5 days ago•0 comments•
53.

Your creative direction, agents do the work. Design with ChatGPT, Claude, or your custom agent on one shared, editable canvas.

1 points•Modecir•6 days ago•0 comments•
54.

System 1 Graph Semantic Code System for AI. Contribute to leonardoventurini/scs development by creating an account on GitHub.

1 points•leonardovb•6 days ago•0 comments•
55.

The execution safety layer for AI agents. Contribute to CTRLRun/ctrlrun development by creating an account on GitHub.

1 points•arpanghoshal•6 days ago•0 comments•
56.

Identity tells you who. Authority tells you what’s allowed. Security was built for doors. AI agents act inside them.

1 points•haqinam01•7 days ago•0 comments•
57.

CheatBench measures whether AI agents attempt to cheat when honest work is difficult. A benchmark from the Center for AI Safety.

1 points•gumby•7 days ago•0 comments•
58.

Dictation, AI cleanup, screenshots and screen recording in one light, fast Mac app, built for talking to AI agents. Source code included.

1 points•Donmario•7 days ago•0 comments•
59.

An analytical database built for agents to use directly: columnar storage, a vectorized query engine, and MCP as a first-class interface - sivsivsree/agedb

1 points•sivsivsree•7 days ago•0 comments•
60.

The real-time messaging layer for your AI workforce. Rooms, DMs, presence, task hand-offs, crash-surviving memory — runs on your machine, no keys. $29 once, yours forever.

1 points•moosepack•7 days ago•0 comments•

Related topics