Sovereign AgentOps – Self-hosted constitutional AI governance for MCP agents
Self-hosted MCP governance server for AI agents with constitutional AI, Ed25519 audit trails, and EU AI Act compliance. Runs offline.
Self-hosted MCP governance server for AI agents with constitutional AI, Ed25519 audit trails, and EU AI Act compliance. Runs offline.
Rust-based terminal AI agent with self-extending capabilities. Open source tool.
Python tool (Picchio) that monitors GPU usage for local LLM inference. Detects CPU fallback and token throughput. 15-line implementation.
Index/registry for evaluating MCP server security for enterprise use. Open source trust evaluation tool.
Essay on AI agent management infrastructure and future deployment models. Written with AI assistance.
OpenSandbox provides infrastructure for AI agents with Docker/Kubernetes support, multi-language SDKs, code execution, and agents/automation capabilities.
HN discussion on limitations of small on-device LLMs like lfm2.5-thinking (1.2B params). User reports poor interaction quality despite model being designed for edge deployment.
arXiv research on how LLMs flatten and reduce nuance in long-form public debate arguments.
Analysis of HackerRank's open source LLM-based hiring agent for resume scoring with GitHub enrichment.
Open source SearXNG CLI with MCP support for agentic web search. Enables coding agents to access web search without proprietary APIs.
Local-first AI agent governance guide covering ICM architecture, offline execution, and safety for autonomous agents.
Agentation: UI annotation tool for AI coding agents to understand interface elements; integrates with Claude and MCP.
Codra: AI code review tool for local ownership; developer tool for code analysis.
Opinion piece ranking AI models mid-2026 with regulatory context; subjective assessment without technical depth.
Educational lab of agent architectures on LangChain/Ollama with runnable CLI variants for studying mechanisms.
AI agents playing Diplomacy game; research on agent behavior in strategic negotiation scenarios.
HoverSource: Developer tool to extract UI element metadata and source files for AI agents via keyboard shortcut.
Discussion of limitations in AI coding agents: excel at code comprehension but lack organizational/team awareness.
Analysis of technical debt in LLM applications: burying logic in prompts during MVP leads to scaling problems.
LiteRT.js: Google's JavaScript binding for on-device ML/AI inference in browsers; privacy-preserving local execution.
The Well: 15TB open-source collection of physics simulation datasets for machine learning across biological, fluid dynamics, and astrophysical domains.
Browser arcade games built using Claude API for pet education, demonstrating LLM-assisted game development.
Real-time grounding verification tool for LLM and RAG agents. Open source, lightweight monitoring utility.
Prompt injection attack method embedding malicious instructions in images to compromise AI agents.
Open standard and marketplace platform for personal AI agents with interoperability focus.
Open-source world model enabling hour-long interactive generation with real-time inference, advancing spatiotemporal reasoning.
Educational platform offering hands-on courses to build systems like Redis and databases from scratch across multiple languages.
Course documentation introducing reinforcement learning and its application in training language models.
openpilot 0.11.1 release using Vision-Language Models for driver monitoring label generation.
Obsidian plugin for creating interconnected note networks with AI assistance.
Directed Contexts defines repo-owned instruction modules for coding agents with domain boundaries, ownership paths, and verification contracts.
Octochains is a Python framework enabling parallel, isolated multi-agent reasoning without model contamination from shared histories.
Autonomous AI agent attempting to win public bet of gaining 100 followers with live dashboard.
GhostCommit demonstrates a steganographic attack exploiting coding agent pipelines by hiding exploits in images within convention files.
Opinion essay comparing LLM capabilities to low-code programming limitations and market maturation.
Code Airlock: Docker-based sandbox wrapper allowing LLM coding agents (Claude, Codex) to execute code safely with package installation and iteration without host access.
Platform for managing and monitoring long-running AI coding agents with persistent state.
Bug report about Codex multi-agent v2 message encryption affecting versions post-0.137.0.
Analysis: AI accelerates entire startup lifecycle beyond development speed—fundraising, product-market fit, growth cycles compress as AI becomes standard tool.
Cerebellum-inspired memtransistor enables energy-efficient AI by detecting unexpected events while ignoring repetitive inputs for wearable and autonomous applications.
Paca v0.9.0: Workflow automation engine enabling task hand-offs between agents/people via visual canvas with rule-based triggering logic.
HackerRank's open-source Hiring Agent: LLM-based resume scorer parsing PDFs, enriching with GitHub/blog data, with analysis of scoring mechanism design and bias.
Ypipe: local-first Java client for offline LLMs and MCP orchestration enabling private agentic workflows without Python.
Analysis of token costs and context compression inefficiencies in coding agents and LLMs, proposing solutions.
isitsecure: developer tool combining SAST, DAST, and LLM-powered code review in single command for web apps.
Developer built generative media gallery with social features using LLMs and samsar-js library in 50 prompts.
Demo comparing LLM outputs using backendjs API modules.
Skillburst syncs AI tool skills with GitHub, enabling teams to manage and deploy workflow updates centrally across non-technical users.
Product studio building AI-augmented tools for decision-making and thinking. Focuses on narrow domains with judgment-critical applications.
TOROLLO: local-first visual simulator for learning system design, backend architecture, and Docker without external setup.