Ax Amos Goldman (NVIDIA Corporation), Nimrod Boker (NVIDIA Corporation), Maayan Sheraizin (NVIDIA Corporation), Nimrod Admoni (NVIDIA Corporation), Artem Polyakov (NVIDIA Corporation), Subhadeep Bhattacharya (NVIDIA Corporation), Fan Yu (NVIDIA Corporation), Kai Sun (NVIDIA Corporation), Georgios Theodorakis (NVIDIA Corporation), Hsin-Chun Yin (NVIDIA Corporation), Peter-Jan Gootzen (NVIDIA Corporation), Aamir Shafi (NVIDIA Corporation), Assaf Ravid (NVIDIA Corporation), Salvatore Di Girolamo (NVIDIA Corporation), Manjunath Gorentla Venkata (NVIDIA Corporation), Gil Bloch (NVIDIA Corporation) 3/17/2026

NCCL EP: Towards a Unified Expert Parallel Communication API for NCCL

NCCL EP: unified expert parallel communication API for Mixture-of-Experts LLM training using GPU-initiated RDMA operations.

Ax Yao Wu, Kangping Yin, Liang Dong, Zhenxin Ma, Shuting Xu, Xuehai Wang, Yuxuan Jiang, Tingting Yu, Yunqing Hong, Jiayi Liu, Rianzhe Huang, Shuxin Zhao, Haiping Hu, Wen Shang, Jian Xu, Guanjun Jiang 3/17/2026

QuarkMedBench: A Real-World Scenario Driven Benchmark for Evaluating Large Language Models

QuarkMedBench realistic medical benchmark for evaluating LLMs on unstructured, ambiguous real-world healthcare queries beyond standardized exam questions.

Ax Alejandro Paredes La Torre, Barbara Flores, Diego Rodriguez 3/17/2026

Knowledge Distillation for Large Language Models

Knowledge distillation framework compressing Qwen 3B to 0.5B using chain-of-thought reinforcement learning across English, Spanish, and code datasets.

Ax Gwanwoo Song, Kwanyoung Park, Youngwoon Lee 3/17/2026

Chunk-Guided Q-Learning

Chunk-Guided Q-Learning algorithm for offline reinforcement learning addressing bootstrapping error and policy class restrictions.

Ax Gowtham, Sai Rupesh, Sanjay Kumar, Saravanan, Venkata Chaithanya 3/17/2026

FLUX: Data Worth Training On

FLUX: Data preprocessing pipeline for LLM training that balances scale and quality without sacrificing token efficiency or data integrity.

Ax Hossein Adeli, Seoyoung Ahn, Andrew Luo, Mengmi Zhang, Nikolaus Kriegeskorte, Gregory Zelinsky 3/17/2026

Human-like Object Grouping in Self-supervised Vision Transformers

Behavioral benchmark comparing self-supervised vision transformer object grouping with human perception using same/different judgments on naturalistic scenes.

Ax Federico Mirra, Matteo Boffa, Idilio Drago, Danilo Giordano, Marco Mellia 3/17/2026

Towards Agentic Honeynet Configuration

Dynamic honeypot deployment system using AI agents to adaptively select assets based on evolving attacker tactics and threat intelligence.