Show HN: Give your AI agent a preview link for files and diff
Diff4: tool converting AI agent file changes into shareable encrypted web preview links, compatible with OpenClaw and Hermes agents.
Diff4: tool converting AI agent file changes into shareable encrypted web preview links, compatible with OpenClaw and Hermes agents.
Browser extension using AI to filter Twitter/X feed by user-defined topics like crypto and rage politics with local model caching.
Personal account of how AI coding assistants changed development workflow and pace, shifting from manual linear processes to iterative exploration.
Multi-model workflow system for evaluating and selecting open source project ideas.
Mugib platform for deploying AI agents across multiple communication channels including chat, voice, web, and real-time data sources.
SuperHQ tool for running coding agents in isolated sandbox environments using microVMs, enabling parallel multi-agent execution without system risk.
Technical deep dive on training LLMs for event forecasting using Tinker, combining multiple approaches to achieve superforecaster-level accuracy on geopolitical predictions.
Tool for Claude to show source locations in PDFs when extracting values, enabling verification.
CUDA's dominance eroding as AI models generate low-level code (kernels, bindings) efficiently, reducing lock-in costs and enabling easier hardware switching.
Reinforcement learning approach using FP4 for exploration and BF16 for training.
Open-source resume evaluation tool using LLM system prompts to match CVs against job descriptions.
MythosAI announces early access to red team operating system for AI security testing.
Dataset project crawling millions of pricing pages and benchmarking ~50 LLMs on their ability to reconstruct pricing page structure.
Notification system for AI agents enabling event-driven alerts via cron jobs without token waste.
NeZha: Open-source agentic development environment for running multiple AI coding agents in parallel across projects using Claude, Codex, Git.
Brief title-only post discussing LLM preference for tool interfaces over pixel-based inputs.
Discussion of Zero Trust Architecture security for Dev In A Box, an AI-powered debugging/security vulnerability detection tool (~70% accuracy).
Formal: Tool for mathematically verifying AI-generated code correctness using Lean 4 theorem proving, works with any LLM backend.
LRTS is an open-source, local-first tool for regression testing of LLM prompts to ensure consistent behavior across model updates.
tasteID: Open-source tool capturing personal design fingerprint across 23 dimensions, portable to any AI tool, privacy-focused.
Local-first desktop/web app for browsing and analyzing AI agent coding sessions. Supports Claude Code, Codex, OpenCode and 11 other agents. Open source with GitHub releases.
Native macOS menu bar app for time-blocking with Claude AI integration, Google Calendar/Outlook sync, and drag-to-create scheduling interface.
Guide using Claude Code and Obsidian to build knowledge management systems with LLM assistance.
Analysis of memory systems for AI agents operating across multiple hosts, distinguishing durable profile state from chat transcripts.
Android AI agent that autonomously controls and operates mobile apps without root/adb access.
Reverse-engineered implementation of Cursor's tab completion client using Connect RPC over HTTP/2.
Local-first database system where LLM agents autonomously organize and structure data using minimal tools.
LLM-powered tool enabling users to create personal websites via chat interface, generates HTML.
CLI tool for Jupyter notebook manipulation optimized for AI agents with AI-friendly markdown format.
Open source framework for verifiable AI decision records with offline verification and append-only chains.
Tool enabling Claude AI to edit videos by analyzing footage and generating timelines for Final Cut, Premiere, and Resolve.
Open source app that monitors AI coding agent sessions and saves conversation history as markdown to repositories.
Demo of LLM-based wiki and RAG system for security knowledge base with MITRE frameworks.
Cost analysis comparing strong-model-first vs weak-model-first strategies for multi-step LLM agent workflows.
Guide to sandboxing AI agents using microVMs and Docker for safe execution environments.
AWS scaffolding tool for rapidly building agents, MCP servers, APIs, and websites.
Personal project: AI shell assistant with container isolation, semantic memory, and self-improvement loop for local developer use.
Loci is a Go-based knowledge store and grounding layer that adds persistent memory to stateless LLMs, enabling lifelong cognitive partnerships.
Analysis of GenAI.mil deployment challenges with classified networks due to air-gapped infrastructure.
Open source web app using vision AI and barcode scanning to decode food ingredients and nutrition.
Ask HN discussion about handling increased code review throughput caused by AI-accelerated development.
Ask HN discussion about emotional impact of agentic AI automation on developer work and learning.
CLI tool that scans codebases, indexes entities into SQLite, exports as structured format for LLMs to reduce tool calls during agent interactions.
Discussion thread on security concerns of sharing API keys and private credentials with AI agents.
Framework enabling LLMs to write TypeScript programs instead of sequential tool calls, improving agent orchestration and execution capabilities.
Analysis and visualization of how AI agent system deployments unintentionally evolve organizational structures through emergent routing and specialization.
Intel Arc Pro B70 GPU with 32GB VRAM for $949 targets local AI inference workloads, undercutting NVIDIA alternatives but facing software limitations.
Go framework for building AI agents with multi-provider LLM support, type-safe tools, agent handoffs, guardrails, MCP integration, and graph orchestration.
Proposal for standardized protocol enabling agents to execute multi-step website tasks with site owner consent, complementing MCP and A2A standards.
Catalog analyzing AI memory and RAG systems through biological memory parallels, mapping vector databases, knowledge graphs, and episodic memory architectures.