Show HN: Callimachus – Local search across your AI coding-agent history
Local search tool for AI coding-agent conversation history across 11 platforms. Includes MCP server, CLI, and VS Code extension. Privacy-focused.
Local search tool for AI coding-agent conversation history across 11 platforms. Includes MCP server, CLI, and VS Code extension. Privacy-focused.
MACCHA: Multi-agent context harness with persistent memory using vector embeddings and semantic conflict detection for AI coding agents like Claude Code and OpenCode.
API service converting web URLs to LLM-ready Markdown for AI agents and RAG pipelines. Handles JavaScript rendering and content cleaning. Free tier: 1,000 pages/month.
Technical demonstration: running 35B MoE model on 2017 AMD GPU via Vulkan without CUDA/ROCm.
TurboPrefill: Intra-Prompt Pipeline Scheduling technique reducing VLM response latency by ~50% without quality loss via multi-GPU prefill optimization.
Bluffbench benchmark shows LLMs approaching saturation on interpreting counterintuitive plots.
Alloy: PyTorch backend and compiler for GPU kernels on Apple Silicon via Metal, with torch.compile integration and LLM serving.
Botacts: directory/phonebook listing AI bot agents across applications with community contribution.
LoopFlow: YAML-based framework for multi-agent coding loops with Claude, enabling self-iterating AI systems with verification gates and memory.
Pulse: local dashboard for Claude Code showing token spend, context usage, and tool call approval via phone with zero dependencies.
Security research documenting AutoJack exploit chain in AutoGen Studio allowing RCE via MCP WebSocket, crossing localhost trust boundary in AI agent frameworks.
Developer tool providing real phone numbers for AI agents with voice calling, SMS, and real-time audio transcription via JSON webhooks.
Case study examining liability and failure modes when AI agents file tax returns, documenting a filer's experience with algorithmic tax optimization.
Technical comparison of Elixir and Phoenix for AI applications, addressing concurrency, WebSocket streaming, and rapid iteration requirements.
Open-source vanilla JavaScript library for building web-based AI agent UIs with native WebMCP support, no framework dependency required.
Local identity server in Rust with Ed25519 signing for AI agents and individuals. No cloud required, cryptographic audit trails.
Tiny: concurrent bytecode virtual machine and dynamic language written in Go with stack-based execution model and multi-threaded runtime.
Personal documentation of self-hosted LLM server setup for remote access to open models and GPU development work.
Middleware agent converting industrial PLC data (Modbus, OPC UA) into REST/gRPC APIs. Lightweight alternative to expensive licensed solutions.
Troubleshooting guide for OpenCode plugin deployment issues across different coding tools and organizational setups.
Discussion on open source sustainability as contribution barriers lower, focusing on trust signals and community management rather than technical implementation.
thethings.ai: Platform enabling agents to publish HTML on the internet. Agent deployment infrastructure.
Discussion on lack of orchestration systems for AI agents. Community analysis of agent infrastructure gaps.
Agent Rigor: Tool addressing failure modes in AI coding assistants to prevent doom-looping behavior. Relevant to agent reliability.
Lume: System describing retrieval primitives for information retrieval. Relevant to LLM applications and RAG systems.
CLI tool and Claude agent skill for Name.com DNS management, designed for AI agent automation with current v4 API support.
Quikdown: 17KB bidirectional Markdown parser with split-view editor. Developer tool but not AI/ML focused.
Tool allowing AI agents to self-wipe context for improved agentic loop engineering and prompt iteration.
Zone of Proximal Policy Optimization for LLM finetuning using RL instead of distillation to improve reasoning and generalization.
Launchreel: tool that converts landing page URLs into social media videos using vision APIs and LLMs to extract content and generate narratives.
Lelu: authorization engine for AI agents that detects manipulated behavior, prompt injection, and anomalies beyond traditional access control.
Overreach: MCP tool that detects when AI coding agents exceed their prompt scope by analyzing diffs for unauthorized changes.
Reverse engineering Qualcomm NPU compiler to enable faster edge deployment of ML models on NPUs with minimal documentation.
Lighthouse adds Agentic Browsing category to audit websites for machine interaction compatibility, focusing on data collection rather than scoring.
Analysis of CPU infrastructure requirements for agentic AI workloads. Discusses hardware implications of deploying AI agents at scale.
Personal essay on using local LLMs vs Claude in daily workflow. Subjective experience with no technical depth or original research.
Java library embedding llama.cpp inference in-process via JDK 22 Foreign Function API without HTTP overhead.
Enterprise GPU-accelerated GUI toolkit for Go using reactive state and Material Design 3 components.
Research on applying vision language models to construction takeoff automation. Found VLMs cannot achieve >80% accuracy due to data limitations in drawings.
Financial operating system for India where humans and AI agents operate on same structured primitives via web, mobile, CLI, and MCP server.
Microcosm transforms codebases into substrate for AI coding agents to orient themselves before acting, with verifiable audit trails of agent actions.
Stepyard: open-source automation runner using YAML pipelines with Python plugins. Runs locally without cloud setup.
Go tool that decrypts iOS WhatsApp backups into searchable SQLite database queryable by LLM agents like Claude and Cursor.
Memory layer for AI agents using Hebbian learning and fuzzy preference graphs instead of LLM calls, producing compact 40-token memory cards with minimal inference overhead.
Logslim is MIT open-source tool that compresses CI/build output by 80-95% for AI agents to read, reducing token consumption.
Reachpad is agent-friendly platform for uploading and organizing artifacts, documents, and knowledge bases for AI agents to query.
StayUp (Duck) is macOS tool to keep system awake while local AI agents, model servers, and renders run. Apple-notarized.
Discussion of tools for LLM agent interactions. Explores output compression, tool semantics for model consumption.
gcontext is open-source CLI for AI agents to maintain markdown notes in git, track task state, resume work across sessions.
Research paper on GitHub Copilot's effect on developer productivity. Observational dose-response analysis on arXiv.