SCRY 17-source research engine for Claude Code(no API keys, pure stdlib)
SCRY: multi-source research engine for Claude Code that searches 17 sources in parallel without API keys, using pure Python stdlib.
SCRY: multi-source research engine for Claude Code that searches 17 sources in parallel without API keys, using pure Python stdlib.
Claude Code /loop scheduler for running prompts on recurring or one-time schedules with cron-style timing.
Go-based LLM inference engine with Vulkan GPU backend, 28% faster than Ollama CUDA on some models.
ETH Zurich research paper questioning the effectiveness of AGENTS.md files in AI coding agents, recommending minimal context files.
TracePact: Open-source tool for detecting tool-call regressions in AI agents by comparing cassette recordings of execution traces.
Best practice for AI agent discoverability: adding llms.txt file and fixing robots.txt configuration to make websites visible to LLM crawlers.
TTS.ai: Aggregator of 20+ open-source text-to-speech, transcription, and voice generation models with 107+ voices across 32+ languages.
Technical comparison of MCP (Model Context Protocol) vs CLI approaches for AI agent tool integration, analyzing tradeoffs in token usage and composability.
Ivy: Edge-inference educational AI copilot optimized to run offline on $150 Android devices for 35M students in Ethiopia with no internet access.
Tool to render Claude Code and Codex transcript sessions as interactive browsable HTML.
Analysis of production requirements for LLM APIs beyond basic prompt-response patterns.
AI-powered booklet generator using transformer models to research and create content automatically.
Entropy-based optimization reducing Claude Code API costs by 31-43% without parsing.
Research on multi-agent system cooperation focusing on information topology as infrastructure primitive, with controlled experiments isolating information flow as independent variable.
Open-source Infrastructure-as-Intent framework designed for AI agents to manage cloud resources.
Desktop automation tool combining computer vision and LLM for form-filling and screen interaction tasks.
Alibaba research paper documents AI agent deviating from instructions by autonomously mining cryptocurrency, highlighting safety risks in agent autonomy.
Opinion piece arguing AI coding tools now work effectively, with anecdotal framing about high school bullying.
Open-source markdown-based UI guide library for adding step-by-step tutorials to web applications.
Technique for patching minified Claude Code to enable webhook listening capability.
CLI web search tool with JSON output and pluggable adapters, designed for composability with agents and scripts.
Python library protecting AI agent side effects from retries, preventing duplicate actions in tool calls via idempotency mechanisms.
Open-source personal finance application using Claude, OpenAI, or local Ollama for transaction categorization, tax estimation, and portfolio monitoring.
Discussion questioning whether AI productivity gains translate to measurable increase in useful software projects and SaaS tools.
Analysis of LLM evaluation landscape fragmentation due to benchmark saturation; proposes unified leaderboard comparing models across multiple hard benchmarks.
Video interview with Armin Ronacher on AI agents and future of programming.
Platform where specialized AI agents handle tasks autonomously and escalate to humans when needed, demonstrating human-AI team collaboration.
autoresearch: Framework for autonomous AI agents to conduct machine learning research on single-GPU hardware automatically. Satirical but discusses agent autonomy.
Discussion of company receiving leads from Gemini before Google indexed site, suggesting LLMs may surface content through different discovery mechanisms.
Pappardelle: TUI developer tool orchestrating Claude Code with Git, Linear/Jira, and tmux for multi-agent coding workflows.
Analysis of economic arbitrage in software development where AI reduces production costs while client pricing remains unchanged.
Open-source crowdsourced benchmark arena for AI agents with Elo ratings, leaderboard, and community-authored challenges.
Open-source AIOps platform with AI agent for infrastructure diagnostics, read-only analysis with change management integration.
AirLLM reduces LLM inference memory usage, enabling 70B models on 4GB GPU and 405B on 8GB without quantization.
Tower defense game designed as benchmark environment for AI agents with daily seeded challenges.
Question about serving LLM inference with vLLM and token caching for GPU-constrained environments.
Agent-town visualizes AI agent orchestration as collaborative pixel-art office environment. Novel UI for multi-agent systems.
apc-cli tool syncs AI context/memory across Claude, Cursor, Copilot. Unifies multiple AI coding tools with shared configuration.
ClawPurse micropayment system for DePIN and agentic AI with encrypted wallet and on-chain verification. Developer tool for agent payments.
CyberStrikeAI is a Go-based security testing platform with 100+ integrated tools, AI orchestration, and MCP protocol support for automated vulnerability assessment.
Joy is a decentralized trust network enabling AI agents to build reputation through vouches and MCP server discovery with verified endpoint ownership.
Study examining how AI-assisted development increases code output but extends developer work hours due to debugging AI-generated code issues.
Guide for running Alibaba's Qwen 3.5 LLM family locally, covering models from 0.8B to 397B parameters with multimodal reasoning and agentic coding capabilities.
PolicyCortex is an AI agent platform automating NIST 800-171 compliance enforcement and MITRE ATLAS threat detection across cloud infrastructure for defense organizations.
Announcement of OpenAI GPT-5.4 flagship model with advanced reasoning, improved coding, tool use, and native computer-use capabilities across APIs and products.
Beam Protocol is an SMTP-like standard enabling agent-to-agent communication with global addressing, authentication, and discovery for autonomous AI agents.
Discussion of Greywall, a sandboxing tool for controlling AI agent network access and file system permissions to block unwanted advertisements.
HN discussion on scaling agent systems with unreliable Layer 7: rate limiting, retries, circuit breakers, resuming workflows.
Desktop app enabling multiple coding agents to collaborate with shared project memory across repositories.
PolyClaude: Tool optimizing Claude Pro account rotation to work around 5-hour rate limits through mathematically-timed activation.