Vorim AI – Identity, permissions, and audit trails for AI agents
Vorim AI provides identity, permissions, and audit trails infrastructure for AI agents. Limited technical details provided.
Vorim AI provides identity, permissions, and audit trails infrastructure for AI agents. Limited technical details provided.
Podcast summarization tool with custom tags and AI chat that learns user preferences over time for personalized content extraction.
Swiper Studio v2 adds MCP server enabling AI agents to build sliders via natural language. Open-source visual builder for the Swiper slider library with agent integration.
DFlash uses block diffusion model for speculative decoding in LLM inference, achieving 6× speedups by drafting entire token blocks in parallel instead of autoregressively.
Claude Code skill that generates project names and checks availability across npm, PyPI, crates.io, and domain registrars. Addresses naming discovery problem.
AI agent that monitors GitHub/Vercel/Sentry, diagnoses errors, writes fixes, runs CI, and opens PRs. Autonomous error remediation with safety gates.
Technical deep-dive on LLM internals through mechanistic interventions analysis. Part of ongoing educational series on building LLMs from scratch.
Fine-tuning tool for Gemma 4 multimodal model on Apple Silicon. Streams data from cloud storage during training with limited compute budgets.
LLM router that prioritizes local models over frontier APIs to reduce costs. Title only, assumes cost-aware LLM routing.
Desktop dev workspace integrating Claude with Kanban board, multi-repo support, and agent SDK for improved task management and iteration.
Chrome DevTools MCP enables AI systems to interact with browser debugging tools for enhanced perception capabilities.
Vix: AI coding agent achieving 50% cost reduction and 40% speedup vs Claude Code using virtual filesystem minification and stem agents for cache optimization.
Spacebot is an agentic AI system with a dedicated LLM process role, positioned as OpenClaw alternative.
Hotcopy is a CLI tool with AI agents that maintain context across sessions and learn over time; developer tool for coding.
Mobile IDE for SSH server management and AI agent orchestration across servers from iPhone interface.
Computer vision ML techniques for room occupancy detection; machine learning application with limited scope.
Self-building knowledge graph of contradictions using local LLM for analysis and visualization.
Project Glasswing uses Anthropic's Claude Mythos 2 frontier model to identify software vulnerabilities, demonstrating AI capabilities in cybersecurity.
Discussion of security requirements as AI models become more capable, using Anthropic's Mythos as case study.
Self-hosted AI research agent that browses web, takes notes, and writes reports with crash recovery via Postgres checkpointing.
Opinion on low-quality AI-generated content; lacks substantive technical analysis.
Project Glasswing uses Anthropic's Claude Mythos model to identify software vulnerabilities; demonstrates AI coding capability for security applications.
Development environment and runner for managing prompts, workflows, and agent pipelines at scale with dataset testing.
AI-powered product feedback tool analyzing submissions and providing detailed improvement suggestions.
Open-source toolkit enabling AI agents to collaborate with humans in reactive marimo Python notebooks with working memory.
Discussion of GitHub Copilot vs alternatives like Cursor and Claude Code for code completion and agentic features.
Rust async client library for Ollama local LLM API with streaming, chat, and embeddings support.
CLI + MCP server tool enabling coding agents visual verification of UI layouts via browser-based testing.
Technical guide on prompt caching optimization techniques with AI co-authoring.
DuckDB-based database system for SQL-capable AI agents with benchmarks across 11 LLMs.
Discussion on marketing developer tools built with rapid development practices.
Tool that automatically discovers optimal system prompts for LLM tasks by analyzing desired output examples, eliminating manual prompt engineering.
OS-level containment system for AI coding agents on macOS, addressing security risks when running untrusted agent code with filesystem/system access.
Developer tool that creates queryable knowledge bases from videos/podcasts for AI agents.
Podcast discussion on governed AI systems in healthcare domain.
Discussion of security vulnerabilities in AI agent sandbox implementations.
macOS containment system using kernel sandboxing and firewalls to safely run unrestricted Claude Code agents autonomously.
AI assistant prototype with session-aware memory that forgets context when users leave.
Best practices guide for AI agent guardrails, covering pre/post-LLM safety patterns.
Mendral is a CI specialist coding agent built on Claude. Demonstrates how identical LLMs produce different outputs through system prompts, tools, and context optimization.
Critique of Anthropic's own AI implementation practices versus enterprise recommendations.
Compares local LLMs (Gemma4-26B, Qwen3.5-35B) for agentic coding tasks using OpenCode and Pi-Coding-Agent with custom tool usage scenarios.
A wiki-based LLM system built on CIS security controls documentation, enabling semantic search and knowledge retrieval over structured security frameworks.
Open source modular OS framework for designing and deploying AI agents. Show HN submission with practical agent infrastructure.
ErrataBench is a benchmark measuring LLM proofreading performance across 51 model variants using an agent loop, with detailed runtime and cost metrics.
Discussion of LLM collaboration patterns in developer tools like Cursor and Claude. Explores user experience challenges with autonomous AI agents.
Infrastructure platform for payment processing integrated with AI agents in EU.
Testreel: npm package for programmatic demo video generation from JSON/YAML/Playwright. Enables LLM agents to create product demos with cursor overlay and customizable backgrounds.
Tutorial on building AI agent for Slack using Chat SDK and AI SDK. Developer guide for LLM integration.
Multi-agent system organized as functional company with independent AI agents in HR, engineering, design roles. Novel agent architecture approach.