HN svenfaw 4/16/2026

I Let Claude Opus Write a Chrome Exploit

Claude Opus used to generate Chrome V8 exploit chains through iterative prompting. Demonstrates LLM capability for security vulnerability discovery with detailed exploitation workflow.

HN tkocmathla 4/16/2026

Teaching AI Agents to Speak Hardware

MCP server for improving AI agent knowledge of specialized hardware (Chimera GPNPU). Addresses hallucination problems in domain-specific contexts.

HN The_resa 4/16/2026

Who Tried Hermes Agent?

AIPOCH: curated library of 420+ medical research skills for AI agents. Open source tool for domain-specific agent capabilities.

Ax Han Li, Yifan Yao, Letian Zhu, Rili Feng, Hongyi Ye, Jiaming Wang, Yancheng He, Pengyu Zou, Lehan Zhang, Xinping Lei, Haoyang Huang, Ken Deng, Ming Sun, Zhaoxiang Zhang, He Ye, Jiaheng Liu 4/16/2026

CodeTracer: Towards Traceable Agent States

CodeTracer framework for debugging and tracing AI agent state transitions, error propagation, and tool orchestration in code agents.

Ax Hevish Cowlessur, Chandra Thapa, Tansu Alpcan, Seyit Camtepe 4/16/2026

Parameter-efficient Quantum Multi-task Learning

Parameter-efficient quantum multi-task learning with shared backbone and task-specific heads for quantum neural networks.

Ax Xiaohua Wang, Muzhao Tian, Yuqi Zeng, Zisu Huang, Jiakang Yuan, Bowen Chen, Jingwen Xu, Mingbo Zhou, Wenhao Liu, Muling Wu, Zhengkang Guo, Qi Qian, Yifei Wang, Feiran Zhang, Ruicheng Yin, Shihan Dou, Changze Lv, Tao Chen, Kaitao Song, Xu Tan, Tao Gui, Xiaoqing Zheng, Xuanjing Huang 4/16/2026

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges

Analyzes reward hacking vulnerabilities in RLHF and alignment approaches for LLMs, examining mechanisms and emergent misalignment issues.

Ax Aram Ebtekar, Michael K. Cohen 4/16/2026

Golden Handcuffs make safer AI agents

Bayesian mitigation strategy for safer AI agents using expanded subjective reward range to prevent reward hacking via risk aversion.