HN shepadorai 3/13/2026

Crazy Rogue AI

Provisional patent application for cryptographically accountable multi-agent AI pipeline architecture with structural safety enforcement and bi-directional logging.

HN gk1 3/13/2026

MCP Doesn't "Suck"

Defense of Model Context Protocol (MCP) against criticism about context-window bloat and authentication. Argues MCP is flexible protocol, not implementation problem.

HN samuel246 3/13/2026

How we hire AI-native engineers now: our criteria

Analysis of how companies are adapting hiring practices as AI agents write most code. Focus shifts from implementation ability to product taste and architectural judgment.

HN tardismechanic 3/13/2026

CLI-Anything

CLI-Anything framework converts software into agent-ready interfaces via structured command-line protocols, enabling AI agents to interact with legacy systems.

Ax Aili Chen, Chi Zhang, Junteng Liu, Jiangjie Chen, Chengyu Du, Yunji Li, Ming Zhong, Qin Wang, Zhengmao Zhu, Jiayuan Song, Ke Ji, Junxian He, Pengyu Zhao, Yanghua Xiao 3/13/2026

DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use

DIVE: scaling task diversity in agentic post-training for robust tool-use generalization across varying tool types and combinations.

Ax Linus Folkerts, Will Payne, Simon Inman, Philippos Giavridis, Joe Skinner, Sam Deverett, James Aung, Ekin Zorer, Michael Schmatz, Mahmoud Ghanem, John Wilkinson, Alan Steer, Vy Hong, Jessica Wang 3/13/2026

Measuring AI Agents' Progress on Multi-Step Cyber Attack Scenarios

Evaluation of frontier AI models' autonomous cyber-attack capabilities on multi-step attack chains across 18-month period.

Ax Xuhui Zhou, Weiwei Sun, Qianou Ma, Yiqing Xie, Jiarui Liu, Weihua Du, Sean Welleck, Yiming Yang, Graham Neubig, Sherry Tongshuang Wu, Maarten Sap 3/13/2026

Mind the Sim2Real Gap in User Simulation for Agentic Tasks

Study formalizing Sim2Real gap in LLM-based user simulators for multi-turn interactive agent evaluation.

Ax Harshitha Menon, Charles F. Jekel, Kevin Korner, Brian Gunnarson, Nathan K. Brown, Michael Stees, M. Giselle Fernandez-Godino, Walter Nissen, Meir H. Shachar, Dane M. Sterbentz, William J. Schill, Yue Hao, Robert Rieben, William Quadros, Steve Owen, Scott Mitchell, Ismael D. Boureima, Jonathan L. Belof 3/13/2026

Multi-Agent Collaboration for Automated Design Exploration on High Performance Computing Systems

Multi-agent LLM framework coordinating specialized agents for design space exploration in scientific computing.

Ax Mengsong Wu, Hao Hao, Shuzhen Bi, Keqian Li, Wentao Liu, Siyu Song, Hongbo Zhao, Aimin Zhou 3/13/2026

Scaling Laws for Educational AI Agents

Empirical study of scaling laws for LLM-based educational agents across role clarity, skill depth, and tool completeness dimensions.