Show HN: One provider starts lying at request 50. The quorum catches it
NUVL is a distributed compute system using quorum verification to detect provider failures/dishonesty, maintaining Byzantine fault tolerance across regions and hubs.
NUVL is a distributed compute system using quorum verification to detect provider failures/dishonesty, maintaining Byzantine fault tolerance across regions and hubs.
RustyRAG: Open-source Rust-based RAG API achieving sub-200ms latency with local embeddings and LLM-generated chunk prefixes.
docsearch is a CLI tool that scrapes and indexes developer documentation locally, integrating with Claude Code via /docs skill for AI-assisted coding.
Analysis of training data sourcing for AI code generation models. Examines ethical questions about AI learning from human engineering work.
Kvlar: Open-source security policy engine for AI agent tool calls. Enforces YAML policies between agents and MCP servers with audit trails.
Computer Use Protocol: Universal schema for AI agents to perceive/interact with desktop UIs. Compact text format optimized for LLM context windows.
Tool to paste URLs and watch multiple AI models redesign websites side-by-side. UI design comparison tool.
Reported details about OpenAI's GPT-5.4 featuring 1M-token context and improved reasoning. Based on third-party reporting without official confirmation.
Notch is a macOS app providing quick AI access with persistent conversations and a background agent that monitors system state and sends periodic messages.
AI agents that integrate with Slack, GitHub, and Jira to autonomously handle development tasks like ticket pickup, code writing, and PR reviews with persistent codebase context.
NumPy-like WebGPU wrapper for browser GPU computing. Zero shaders, automatic CPU/WebGL2 fallback. Suitable for local-first AI.
Open-source coding agent supporting multiple models. Free tier with 100 requests/day, $15/month premium with wholesale token pricing.
Discussion thread listing recent AI agent sandboxing solutions (microVMs, WASM, browser isolation) with inquiry into production usage, security tradeoffs, and performance characteristics.
LearnCodeGuide is an AI tool that analyzes code and provides health scores with suggestions for logic, performance, security, and maintainability improvements across multiple analysis modes.
Tool for right-sizing reserved LLM capacity based on Service Level Objectives.
GitHub Copilot Memory now default for Pro users. Persistent repository-level understanding to reduce context re-explanation.
Sentinel is a Go-based LLM proxy with 13ms semantic caching, PII scrubbing, and automatic fallback routing across OpenAI, Anthropic, Gemini, and Groq.
Workflow runtime for Claude Code with terminal UI. Wraps Claude via hooks for multi-step task automation with templates and plugins.
VibeCheck is a tool that quizzes developers on code changes made by AI coding assistants, supporting Claude Code, Cursor, Windsurf, Cline, and other tools.
Guide for developers on HIPAA compliance when building healthcare products with LLMs, covering BAA requirements and regulatory considerations.
Narrative account of an AI agent given $100 and 26 days to generate $200/month revenue, documenting business-building attempts and autonomous decision-making.
91k line OSS agent orchestration codebase generated by Claude with $1k bug bounty for poor engineering.
Autonomous coding agent in Rust that self-improves via GitHub Actions. Reads own source, constitution, and issues to implement improvements.
Amazon Lightsail now supports deploying OpenClaw, a self-hosted private AI assistant with built-in security controls and sandboxing.
Project proposing AGENTS.md file standard to reject AI agents from accessing repositories.
Movie reminder tool built with Val Town and Movie Database API. Personal project using serverless platform.
Open source project for agentic monitoring and trust verification of AI agents with associated dashboard and trust center.
Browser-based tamper-evident verification system for AI agent execution logs. Detects when agent evidence has been altered.
Conceptual discussion on SaaAS (Software as AI Service): future where LLMs and AI agents generate code instead of humans writing it.
Circle CI's CLI tool mines GitHub PR review comments and generates AI agent context prompts tuned to team code review standards.
Engineer used AI agents to build open-source Verilog simulator with verification stack, replacing expensive commercial EDA tools.
Research on convergence: different AI models appear to encode concepts similarly despite different architectures, suggesting unified internal representations.
Open-source localhost tunnel built in Elixir/Phoenix. No signup, no tracking. MIT licensed alternative to ngrok.
Educational course teaching product managers to use Claude Code. LLM application training material with limited technical depth.
Rust framework (AutoAgents) with composable middleware for LLM inference optimization: safety enforcement, caching, data sanitization. Open-source agent framework with technical depth.
Zero-knowledge secret sharing CLI tool with OpenAI Skill for AI workflows and Node.js implementation. Developer tool with LLM integration.
Open-source GPU cluster monitoring tool for ML training. Built to track foundation model training (Linum v2, 2B params). ML infrastructure with technical implementation.
Technical analysis of AI agent failure modes, arguing execution history matters more than model internals for debugging agentic systems. Original research on agent debugging.
ChatRoutes: Open-source platform for managing AI conversations with branching, parallel multi-model responses, and REST API. LLM application framework.
OpenTimelineEngine: Local-first shared memory platform for Claude Code and Codex agents to track workflows and extract patterns. Open source AI agent infrastructure.
safe-docx: Open-source TypeScript library for AI agents to surgically edit Word .docx files while preserving formatting. LLM application for document automation.
Multi-agent negotiation engine with knowledge graph and semantic git. Description lacks technical clarity on implementation details.
Social media scrolling agent that learns user preferences and filters feeds. Consumer application using AI to reduce information noise.
Framework for policy-governed smart wallets using ERC-4337 for AI agents interacting with crypto. Enforces spending limits and access controls on-chain.
MCP server enabling LLMs to use rr reverse debugger for program repair without modifying source code. Enables LLM agents to inspect program state via GDB/MI.
helpme: Bash wrapper launching tmux split with context-aware AI assistant CLI for debugging. Developer tool for agent-assisted debugging.
arXiv research framework announcement. No technical details provided about the dual-LLM policy repair method itself.
CLI tool syncing AI agent skills and MCP servers across multiple coding agents (Codex, Cursor, Copilot, Claude, Gemini) via symlinks.
Two Claude Code skills for founders: investor call debrief capture and ADHD-aware interaction mode. LLM application for business.
SmartAgentKit provides policy-governed smart wallets for AI agents to manage financial operations autonomously with configurable restrictions.