From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors
FALCON vision-language-action model incorporating 3D spatial foundation priors to improve reasoning and generalization in embodied AI tasks.
FALCON vision-language-action model incorporating 3D spatial foundation priors to improve reasoning and generalization in embodied AI tasks.
Interpretable operator-learning ML model for reconstructing electric field distributions from EFISH signal profiles in plasma physics.
Fairness-aware LoRA fine-tuning of vision-language models for medical imaging with differentiable MaxAccGap loss for demographic parity optimization.
ELERAG system enhancing retrieval-augmented generation with entity linking to improve factual accuracy in specialized domains like education.
ADHint method integrating difficulty-aware hints into reinforcement learning post-training to improve sample efficiency and reasoning generalization.
Theoretical analysis of distributed optimization with multiple local updates between communication rounds, proving acceleration guarantees.
Bayesian generative modeling framework for flexible conditional inference on arbitrary partitions of observed variables without fixed conditioning structure constraints.
Equivariant neural networks for robust object recognition under symmetric transformations and unusual viewing conditions.
FinTexTS dataset pairs financial time-series with semantic text data for multi-modal financial forecasting and analysis.
Analysis of performative chain-of-thought in reasoning models, showing models generate tokens without revealing internal beliefs via activation probing.
PolyBlocks: MLIR-based modular compiler infrastructure for AI programming frameworks and chips using affine analysis and analytical cost models.
VLN-Cache improves vision-language model inference efficiency for Vision-and-Language Navigation via semantic-aware token caching.
Megatron Core system optimizations for scaling Mixture-of-Experts model training across memory, communication, and computation constraints.
Covenant-72B: 72B parameter LLM trained via globally distributed, trustless peer-to-peer training over the internet without whitelisting.
Clinical feasibility study of AMIE, an LLM-based conversational AI for patient diagnostic history in real-world primary care workflows.
PostTrainBench benchmarks LLM agents' ability to automate post-training of language models, extending AI agents to AI research automation.
Discussion of prompt fatigue challenges when using LLMs for coding and writing, asking community for management strategies.
ClawSoc: Open-source framework for observing and testing AI agents in multi-agent scenarios with game-theoretic interactions.
Analysis of how AI pioneers Bengio, Hinton, and LeCun's divergent views on AI's future informed building TRACE platform.
Gemini CLI agent that orchestrates Google Workspace APIs to generate polished documents, sheets, and slides from natural language input.
MCP-compatible credit optimizer reducing Manus AI token usage 30-75% through prompt analysis and six optimization strategies.
OpenClaw plugin adding multi-mode orchestration (ask/delegate/autonomous) for Claude-based code generation with plan approval and session persistence.
Discussion on preventing runaway behavior in MCP-based agents through loop detection, tool call limits, and iteration constraints in production deployments.
IOA Core: open-source governance kernel for AI workflows with policy checks, audit trails, and quorum-style review patterns, provider-neutral execution controls.
Autonomous AI agent that optimizes control systems by independently writing code, training models, and iterating on research problems with minimal human guidance via Discord.
Polaris API provides structured, real-time intelligence from 160+ news sources for AI agents to query and reason over global events without web scraping.
Open marketplace indexing 45,000+ AI agent skills with semantic search. Works with Claude Code, Cursor, Windsurf and other agents.
Python library with drop-in adapters for translating embeddings between different model vector spaces. Enables interoperability without hacks.
MCP server providing 18 structured tools for AI agents to interact with Robinhood trading platform. Compatible with Claude Code and OpenClaw.
macOS utility that fixes prompt typos before sending to Claude, Codex, or Gemini. Reduces prompt noise in terminal AI sessions.
Claude Code skill that organizes problems into cross-functional teams and executes work in parallel using dependency-based waves and subagents.
Discussion thread on code review practices for AI-generated code. Explores tension between natural language prompting and artifact review.
Python library for creating plugin infrastructure. Enables code to automatically hook into contexts without direct dependencies.
Multi-agent software engineering framework using contracts-first architecture. Agents implement code in parallel with mechanical test validation.
MCP server for Hacker News that enables AI agents to discover relevant stories, identify credible voices using EigenTrust propagation, and understand ranking signals.
Curly-brace syntax prompting language for AI agents. JavaScript-like syntax for structured prompts with local LLM support.
Data agent middleware that builds semantic understanding of databases automatically. Sits between agents and databases to provide business logic context.
Research report analyzing AI economics: inference subsidies, energy constraints, semiconductor dependencies, labor disruption. 248k-word study.
Mumpix local-first AI infrastructure stack with database, memory, and state management for edge deployment. Open source developer tools.
Tutorial building deep research agents using DSPy Signatures and Modules. Covers agent design patterns and composable programs.
Tutorial on building MCP Server connecting LLMs to local databases using Model Context Protocol. Practical 10-minute setup guide.
Security report about root access vulnerability in Meta's AI infrastructure via prompt injection. Title only.
Title-only post about getting started with AI agent coding. No content provided.
Novel technique for fine-tuning local models via contrastive human feedback, reducing token usage 5.7× without technical background.
Announcement of Covenant-72B, a 72B parameter LLM trained via trustless peer-to-peer distributed pre-training.
Research on using LLMs for vulnerability discovery via AI-powered fuzzing, differential analysis, and automated harness generation.
Guide on agentic engineering patterns and maintaining code quality when using AI coding tools.
Discussion on inefficiencies in AI agent reasoning before code generation, lacks technical depth or evidence.
macOS push-to-talk transcription tool supporting Groq, OpenAI, and Deepgram models with sub-second latency.
Identity graph API with 330M+ verified B2B records to reduce hallucinations in AI agents, stress-testing available.