HN edf13 3/16/2026

Custom AI Smart Speaker

Product for building custom AI voice assistants deployable on hardware via Voice SDK. Marketing-focused with limited technical details.

HN tonysurfly 3/16/2026

WebMCP

Browser standard enabling websites to expose structured JavaScript tools to in-browser AI agents via navigator.modelContext.

HN rzk 3/16/2026

Wolfram LLM Benchmarking Project

Wolfram's LLM benchmarking project for evaluating language model performance. Research-focused evaluation framework.

HN Mopolo 3/16/2026

Docker Sandboxes and Docker Agent

Docker Sandboxes enables AI agents to autonomously handle multi-disciplinary development tasks. Frames agents replacing context-switching across product/design/engineering roles.

HN hpcaitech 3/16/2026

What's the Best LLM for Coding in 2026

Comparative analysis of LLMs for code generation and debugging, examining performance across reasoning, code generation, and general understanding tasks.

HN theorchid 3/16/2026

List of Rules for Cursor

Configuration system for Cursor AI editor defining custom rules and behaviors for code generation via .cursorrules files.

HN jkpe 3/16/2026

Code on the Fastest Largest AI Chip Ever Built

Announcement of GPT-5.3-Codex-Spark model for real-time coding in Cursor IDE, 1000+ tokens/sec, 128k context window, text-only. Details sparse, appears promotional.

HN vismit2000 3/16/2026

What Is an Agent Harness?

Agent harness concept: software infrastructure wrapping LLMs/agents for orchestrating tools, memory, workflows. Technical introduction with architectural focus.

HN teleforce 3/16/2026

Agentic Trust Framework (ATF)

Agentic Trust Framework: open security specification for Zero Trust governance of autonomous AI agents. Standards and governance for agent deployment.

LB sebastianraschka.com via refi64 3/16/2026

LLM Architecture Gallery

LLM Architecture Gallery: curated collection of architecture diagrams and specifications for major LLMs. Technical reference resource.

Ax Yulin Li, Tengyao Tu, Li Ding, Junjie Wang, Huiling Zhen, Yixin Chen, Yong Li, Zhuotao Tian 3/16/2026

Efficient Reasoning with Balanced Thinking

Research on Large Reasoning Models addressing overthinking/underthinking inefficiencies through balanced computational resource allocation for improved reasoning accuracy.

Ax Orit Shahnovsky, Rotem Dror 3/16/2026

AI Planning Framework for LLM-Based Web Agents

arXiv paper formalizing web task planning for LLM agents through sequential decision-making, mapping agent architectures to traditional planning paradigms.

Ax Zihan Wang, Zhongkui Ma, Xinguo Feng, Zhiyang Mei, Ethan Ma, Derui Wang, Minhui Xue, Guangdong Bai 3/16/2026

AI Model Modulation with Logits Redistribution

AIM model modulation paradigm enabling single LLM to exhibit diverse behaviors through utility and focus modulation modes.

Ax Smriti Jha, Vidhi Jain, Jianyu Xu, Grace Liu, Sowmya Ramesh, Jitender Nagpal, Gretchen Chapman, Benjamin Bellows, Siddhartha Goyal, Aarti Singh, Bryan Wilder 3/16/2026

Developing and evaluating a chatbot to support maternal health care

Development and evaluation of a phone-based chatbot for maternal health information in low-resource multilingual settings.

Ax I. de Zarz\`a, J. de Curt\`o, Jordi Cabot, Pietro Manzoni, Carlos T. Calafate 3/16/2026

Semantic Invariance in Agentic AI

Benchmark and evaluation framework for testing semantic invariance of LLM agents under equivalent input variations in reasoning tasks.

Ax Thomas Kleine Buening, Jonas H\"ubotter, Barna P\'asztor, Idan Shenfeld, Giorgia Ramponi, Andreas Krause 3/16/2026

Aligning Language Models from User Interactions

Method for learning from multi-turn user interactions to improve LLM alignment without explicit labels via implicit feedback.