Show HN: Interactive recursive tool that warps what traditional consultants do
Tool converts natural language into interactive system diagrams using Stafford Beer's Viable System Model with LLM analysis.
Tool converts natural language into interactive system diagrams using Stafford Beer's Viable System Model with LLM analysis.
TRIBE v2 foundation model predicts human brain responses to visual, auditory, and language stimuli for neuroscience research.
Claude Code plugin enabling iterative code review workflows where Claude plans/implements and Codex reviews each step.
Hollow is serverless web perception API for AI agents with perceive/act primitives, enabling web access via low-cost functions.
Opinion piece discussing perception gap between executives and individual contributors regarding AI adoption and coding agents.
Sashiko is an agentic system using LLMs with Linux kernel-specific prompts to automatically review proposed kernel patches from mailing lists.
ICLR 2026 paper on divide-and-conquer approaches for weak models handling long context tasks efficiently
Study shows radiologists and AI both struggle to detect AI-generated deepfake X-ray images
Runtime tool enforcing RAG provenance by validating LLM responses against retrieved chunks, sitting between retriever and output.
Open-source TUI for managing multiple AI agents (Claude, Codex, Gemini) across git worktrees
Wikipedia bans LLM-generated content for articles due to policy violations
Paper proposing LLMs lack contextual understanding and proposing bounded rationality framework for intelligence
Open-source multi-agent orchestration platform with local model support, dashboard, and free tier
Open-source tool for ranking large image collections via pairwise comparisons using Bayesian TrueSkill algorithm for preference learning.
Analysis of how LLMs process language through cognitive semantics framework, examining meaning representation beyond tokens.
Revge builds AI infrastructure for Kurdish language preservation with models for speech, text, and vision processing.
ColBERT-Zero: Multi-vector retrieval model trained from scratch using dense-to-sparse adaptation for enterprise search.
GitHub will use private repositories for Copilot training unless users opt out by April 24.
Platform for evaluating AI agents on coding tasks using isolated sandboxes, automated scoring, and leaderboards. Built with Claude Code and Codex.
Research on type inference algorithm that improves error messages by prioritizing correct type assumptions over incorrect ones.
MIT project: persistent memory layer for AI agents that maintains conversation history in vector-indexed knowledge graph. Available as OpenClaw plugin, MCP server, and cloud dashboard.
TypeScript CSS parser with AST support for modern CSS including nesting, scopes, and cascading layers.
Open-source macOS GUI agent that autonomously reviews iPhone apps by browsing App Store, installing apps, exploring interfaces, and generating narrated video reviews.
Open-source self-hosted system running Claude Code tasks from email and Slack with isolated containers, network controls, and credential handling.
Framework for assessing organizational AI maturity, measuring integration depth from copilot experimentation to autonomous workflows.
Aura: open-source agent orchestration framework (Apache 2.0) for production AI systems, applying platform engineering patterns to coordinate AI complexity.
How AI agents select tools based on API simplicity. Resend's simpler API won over SendGrid/Twilio in agent-driven tool selection, reaching 500K developers.
Anvil desktop app for spec-driven development with parallel coding agents, worktree isolation, and spec-first workflows.
Discussion: viability of consumer-grade local AI boxes as alternative to cloud-based LLM access.
TurboQuant: sub-byte KV cache quantization technique enabling 80K+ token context on 32GB consumer hardware for LLM inference.
TurboQuant paper on online vector quantization with near-optimal distortion rate for machine learning.
Discussion: developers building local AI hardware setups for coding tasks as alternative to cloud services like Cursor.
Claude Code SSH skill and daemon enabling clipboard image pasting to remote sessions over SSH.
Appaca platform generates custom software from natural language descriptions using AI. LLM application for software generation but lacks technical depth.
MUP (Model UI Protocol) lets LLMs control interactive browser UI panels in real-time. Demo shows Claude narrating, building presentations with charts, and coding live.
Research on integrating security best practices into LLM-based code generation systems.
Chrome extension providing prompt suggestions for users.
MLForge is a visual ML platform with PyTorch backend supporting GPU/CPU/MPS training. Open source developer tool for machine learning workflows.
Discussion questioning whether AI agents are being over-applied to problems that don't require them.
Article on working asynchronously with Claude to avoid bottleneck of waiting for task completion
User demonstrates defeating Pangram AI detection tool by using Claude to generate human-sounding text.
Shoofly intercepts and controls AI agent tool calls before execution for safety and debugging.
DeerFlow v2.0 is open-source agentic framework orchestrating sub-agents, memory, sandboxes with extensible skills
Research showing sycophantic AI chatbots reduce user kindness toward others in social interactions.
Research showing that repeating prompts multiple times improves performance of non-reasoning LLMs.
Toolcast automates conversion of any API into an AI agent tool through a single command.
API proxy tool enabling OpenCode models to be used through OpenAI, Anthropic, and Gemini API interfaces.
MCP server enabling AI agents to access institutional financial data via real-time stock prices, fundamentals, and trading insights.
Article discussing challenges and limitations of AI coding agents in practice.
Open-source UI framework for Claude code agents with iMessage/web browsing, scheduling, tunneling, and MCP support.