Show HN: AI agents that run real user interviews
Usercall MCP tool enables AI agents to conduct user interviews via voice calls, returning structured insights with themes and quotes.
Usercall MCP tool enables AI agents to conduct user interviews via voice calls, returning structured insights with themes and quotes.
Distributed multi-agent cluster using local LLMs (DeepSeek-R1, Qwen) on single GPU server to reduce cloud API costs.
Nvidia announces space computing platforms for orbital data centers with AI acceleration for geospatial and autonomous operations.
Zalor feature for testing AI agents with custom CSV datasets and automated test case generation from edge cases.
TypeScript wrapper simplifying MCP server connections from 30+ lines to 2 lines. Supports HTTP and stdio transports with auto-detection.
MIT-licensed digital twin of AWS providing local replica responding to real AWS API calls. Built entirely with AI, supports 147 services, designed for agent testing.
AI skills framework for affiliate marketing automation with 45 skills across research, content, deployment, tracking stages.
Sandbox for multi-agent debate system where agents search and reason about questions that challenge standard LLM refusals.
ELIDA session border controller adds governance and access control separation for AI agents across enterprise systems.
OnPrem.LLM Agent pipeline enabling autonomous agents with tool-calling in 2 lines, supporting local and cloud models.
macOS utility to automatically manage Anthropic Claude rate limits by scheduling prompts during reset windows.
Evolution Engine open source CLI detects development process drift in major OSS repos using statistical analysis across 130K+ commits.
Analysis of enterprise AI fragmentation and need for unified API layer to access organizational knowledge across disconnected AI platforms.
Commentary on how AI coding agents are replacing traditional junior developer roles and changing job market expectations.
ToolGuard open source Python tool fuzzes AI agent tool functions to test reliability; detects hallucinations and type mismatches.
Microsoft AI Toolkit VS Code update enables 5-minute agent setup with identity management, sandboxing, and compliance controls.
Xecai Python library for RAG systems abstracting common LLM provider APIs with sync/async support, embeddings, reranking.
CodeLedger tool addresses AI coding agent issues through deterministic context selection and execution guardrails to prevent scope drift.
ToolGuard open source Python tool fuzzes AI agent tool functions to test reliability; detects hallucinations and type mismatches.
Ask HN thread where junior engineer seeks advice on AI coding workflows and tools to improve delivery speed.
TPCP open protocol enables peer-to-peer communication between AI agents across different frameworks and models without vendor lock-in.
Lore: local AI tool for thought capture and recall using Ollama and LanceDB with RAG pipeline, runs offline on user's machine.
GlassWorm malware campaign compromised 433 packages across GitHub, npm, and VSCode extensions. Supply chain security threat affecting open source.
Conductor: CLI tool for defining multi-agent workflows in YAML with GitHub Copilot SDK and Claude, supporting human approval gates.
NVIDIA expands open model families including Nemotron for agentic AI and Cosmos for physical/healthcare AI systems.
Mamba-3 research paper: new state space model architecture optimized for inference efficiency with 40+ production-ready models.
NVIDIA announces Dynamo 1.0, open source software for scaling generative and agentic AI inference across data centers efficiently.
PAP protocol for privacy-preserving AI agents using cryptographic guarantees to prevent data leakage and profiling by platform operators.
Krasis: Python-orchestrated Rust runtime enabling 200B+ parameter LLM inference on single consumer GPU with full prefill/decode.
Tool scoring GitHub repositories for AI coding agent readiness based on OpenAI's agentic legibility framework.
ProtoScience: deterministic system discovering physics laws from raw data using sparse regression without LLMs, validated on NASA/NOAA datasets.
Discussion of write consistency guarantees for production agent workflows. Real-world agent failure modes and HITL mitigation strategies.
Discussion thread: developer built tool for handling payments in AI agent systems, seeking solutions from community.
Cost control library for AI agents with budget limits, automatic tracking, and circuit breaking across LLM providers. Addresses unpredictable agent spending.
AgentMarket: API marketplace enabling AI agents to buy/sell capabilities at per-call pricing. Infrastructure for agent interoperability and capability composition.
Framework separating routing, verification, and judgment tasks for LLM pipelines. Structured approach to handling user input and evidence retrieval without oracle dependency.
Wuobly: AI agent that searches the live web for B2B leads with reasoning. Performs real-time verification of contact information and explains fit.
Tool for integrating AI agents into 8090 Software Factory SDLC workflows. Limited detail provided.
Security research on vulnerability exploitation in AWS Bedrock AgentCore's AI code interpreter. Title only, minimal content.
Project packaging programming books into Claude Code skills to apply best practices when reviewing/generating code. Open source tool on GitHub.
Mistral AI releases Forge, a system for enterprises to build frontier AI models customized with proprietary knowledge and internal data.
Running 35B MoE LLM locally on vintage AMD crypto APU using Vulkan. Technical optimization for resource-constrained LLM inference.
HN discussion on managing code review bottlenecks from AI coding agents. Addresses scaling human review processes for high-volume AI-generated code.
Grape: AI note-taking app with vector embeddings and semantic search. LLM-powered note organization and retrieval using chat interface.
Tool for AI coding models to generate architecture decision records before implementation. Multi-repo architecture management with AI-powered spec generation.
Runtime security layer for AI agents that moves beyond prompt filtering to protect agent behavior at execution time.
Magda is an open-source digital audio workstation with integrated AI, built in C++ using JUCE and Tracktion Engine.
Open source project packaging programming books into Claude Code skills for applying best practices to code generation and review.
Middleware layer enabling multi-agent interoperability through schema translation and semantic mapping for heterogeneous agent protocols.
Aimploy is a professional network platform for AI agents.