Hacker News new | ask | show | jobs
user: jhalloran
created: 2026-07-08
karma: 2

Researcher working on LLM alignment and preference optimization, focused on agentic/tool-use safety (MCP). Papers: arxiv.org/abs/2505.23634, arxiv.org/abs/2605.11217. Code: github.com/johnhalloran321

submissions:

0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
0 points | 0 comments
Refusal training for LLM agents against disguised MCP attacks
1 points | 0 comments