Harmonic Contour Integration: Compact, distributed edge detection algorithm
Harmonic Contour Integration: distributed edge detection algorithm for RGB images using learned harmonic metrics.
Harmonic Contour Integration: distributed edge detection algorithm for RGB images using learned harmonic metrics.
MnesticDB fork of CozoDB adds bitemporal provenance tracking for AI agent memory, enabling auditable change history and belief state management.
OpenBenchmarks: open-source reproducible benchmarks for SaaS APIs designed to help AI agents evaluate and select tools.
TradingSpy: open-source local-first AI trading assistant with market analysis, strategy generation, and backtesting via Docker.
AI Disagreement Index tracks cross-model disagreement on tool recommendations across 8 models, measuring consistency rather than ranking.
Tool enabling Claude Code IDE to use multiple LLM providers including GPT, Kimi, and Grok.
Sqlsure provides deterministic semantic validation for AI-generated SQL queries, catching logic errors in joins and aggregations before execution.
AI Arcade benchmarks coding models building arcade games with identical prompts and frameworks, comparing output quality and functionality.
Open source browser-based tool for evaluating AI agent outputs using human labels and LLM judges. Local, no backend required.
Self-hosted MCP governance server for AI agents with constitutional AI, Ed25519 audit trails, and EU AI Act compliance. Runs offline.
Rust-based terminal AI agent with self-extending capabilities. Open source tool.
Python tool (Picchio) that monitors GPU usage for local LLM inference. Detects CPU fallback and token throughput. 15-line implementation.
Index/registry for evaluating MCP server security for enterprise use. Open source trust evaluation tool.
Essay on AI agent management infrastructure and future deployment models. Written with AI assistance.
OpenSandbox provides infrastructure for AI agents with Docker/Kubernetes support, multi-language SDKs, code execution, and agents/automation capabilities.
HN discussion on limitations of small on-device LLMs like lfm2.5-thinking (1.2B params). User reports poor interaction quality despite model being designed for edge deployment.
arXiv research on how LLMs flatten and reduce nuance in long-form public debate arguments.
Analysis of HackerRank's open source LLM-based hiring agent for resume scoring with GitHub enrichment.
Open source SearXNG CLI with MCP support for agentic web search. Enables coding agents to access web search without proprietary APIs.
Local-first AI agent governance guide covering ICM architecture, offline execution, and safety for autonomous agents.
Agentation: UI annotation tool for AI coding agents to understand interface elements; integrates with Claude and MCP.
Codra: AI code review tool for local ownership; developer tool for code analysis.
Opinion piece ranking AI models mid-2026 with regulatory context; subjective assessment without technical depth.
Educational lab of agent architectures on LangChain/Ollama with runnable CLI variants for studying mechanisms.
AI agents playing Diplomacy game; research on agent behavior in strategic negotiation scenarios.
HoverSource: Developer tool to extract UI element metadata and source files for AI agents via keyboard shortcut.
Discussion of limitations in AI coding agents: excel at code comprehension but lack organizational/team awareness.
Analysis of technical debt in LLM applications: burying logic in prompts during MVP leads to scaling problems.
LiteRT.js: Google's JavaScript binding for on-device ML/AI inference in browsers; privacy-preserving local execution.
The Well: 15TB open-source collection of physics simulation datasets for machine learning across biological, fluid dynamics, and astrophysical domains.
Browser arcade games built using Claude API for pet education, demonstrating LLM-assisted game development.
Real-time grounding verification tool for LLM and RAG agents. Open source, lightweight monitoring utility.
Prompt injection attack method embedding malicious instructions in images to compromise AI agents.
Open standard and marketplace platform for personal AI agents with interoperability focus.
Open-source world model enabling hour-long interactive generation with real-time inference, advancing spatiotemporal reasoning.
Educational platform offering hands-on courses to build systems like Redis and databases from scratch across multiple languages.
Course documentation introducing reinforcement learning and its application in training language models.
openpilot 0.11.1 release using Vision-Language Models for driver monitoring label generation.
Obsidian plugin for creating interconnected note networks with AI assistance.
Directed Contexts defines repo-owned instruction modules for coding agents with domain boundaries, ownership paths, and verification contracts.
Octochains is a Python framework enabling parallel, isolated multi-agent reasoning without model contamination from shared histories.
Autonomous AI agent attempting to win public bet of gaining 100 followers with live dashboard.
GhostCommit demonstrates a steganographic attack exploiting coding agent pipelines by hiding exploits in images within convention files.
Opinion essay comparing LLM capabilities to low-code programming limitations and market maturation.
Code Airlock: Docker-based sandbox wrapper allowing LLM coding agents (Claude, Codex) to execute code safely with package installation and iteration without host access.
Platform for managing and monitoring long-running AI coding agents with persistent state.
Bug report about Codex multi-agent v2 message encryption affecting versions post-0.137.0.
Analysis: AI accelerates entire startup lifecycle beyond development speed—fundraising, product-market fit, growth cycles compress as AI becomes standard tool.
Cerebellum-inspired memtransistor enables energy-efficient AI by detecting unexpected events while ignoring repetitive inputs for wearable and autonomous applications.
Paca v0.9.0: Workflow automation engine enabling task hand-offs between agents/people via visual canvas with rule-based triggering logic.