Trajectory-aware guardrails for LLM chatbots: catches the gradual delusion reinforcement that turn-local safety filters miss - nwjang/psychosis-guard

1 points•nwjang•7 days ago•0 comments•

0 comments

No comments yet.

Read the full thread on Hacker News →

Related stories