HN AsafOz 3/25/2026

Instagram for AI Agents

Platform concept for discovering, sharing, and interacting with AI agents as social media-style experience.

HN gaigalas 3/25/2026

Teaching AI to Verify Sources

Technique to make LLMs verify their own source citations before output using AST parsing to detect hallucinations.

HN mpweiher 3/25/2026

Vibe physics: The AI grad student

Physics professor uses Claude AI to complete theoretical physics research calculation end-to-end without manual intervention.

HN fclaude 3/25/2026

Persistent long-term memory for Claude Code

Claude Code agent with persistent long-term memory using tiered architecture, deployed as FastAPI/SQLite/sqlite-vec backend with scale-to-zero pricing.

HN handfuloflight 3/25/2026

Using AI as a Design Engineer

Personal experience using AI as a design engineering tool for experimentation and iterative refinement in creative work.

Ax Ricardo Olmedo, Bernhard Sch\"olkopf, Moritz Hardt 3/25/2026

Computational Arbitrage in AI Model Markets

Research on arbitrage mechanisms in AI model markets where customers allocate inference budget across competing providers with different costs and capabilities.

Ax Jihyun Janice Ahn, Ryo Kamoi, Berk Atil, Renze Lou, WonWoo Kang, Heehyun Park, Sarkar Snigdha Sarathi Das, Zhuoyang Zou, Xiaoxin Lu, Yusen Zhang, Asfahan Shah, Ridwanul Hasan Tanvir, Lingxiao Zhao, Hongxi Huang, Vignesh Venkatesh, Dianjun Lin, Hamid Shah, Wentao Wang, Zhanpeng Song, Joshua Reed Bassin, Dax Patel, Ishan Appareddy Agrahar, Sahil Pardasani, Xin Dong, Fatemeh Rahbari, Benjamin David Rishel, Soochan Andrew Lee, Yuv Boghani, Ali B. AlNaseeb, Pranav Suby, Seokhyeon Bae, Shreya Buddharaju, Damien Kula, Soumyadeep Das, Hanyang Frank Liu, Faye Mo, Wenpeng Yin 3/25/2026

Bridging the Know-Act Gap via Task-Level Autoregressive Reasoning

Research on the know-act gap in LLMs: models can identify flawed inputs discriminatively but fail to reflect this in generative responses, revealing a fundamental gap between recognition and generation.