Dibi8 – Open-source AI tools directory, 4 languages
Dibi8: Open-source directory cataloging AI tools across four languages.
Dibi8: Open-source directory cataloging AI tools across four languages.
Hyperia 0.16.1 adds spoken summary generation feature for coding agents.
Analysis of LLM-generated unit test sprawl and strategies for managing excessive test coverage from AI tools.
Open source benchmark for evaluating LLM capability on tax filing tasks. Developer tool.
Discussion thread on patterns for managing long-lived research projects and LLM workflows.
DuckDB community extension enabling queries across multiple databases via ADBC drivers; supports Snowflake, BigQuery, PostgreSQL, MySQL.
Gap Map: Analysis of 24,626+ open-source AI projects identifying missing components across foundation models to inference backends.
Agent-driven news aggregator for AI filmmaking using agentic architecture for discovery and classification. Developer tool.
Crondex: open-source directory of pre-made cron jobs for AI agents to discover, customize, and schedule tasks.
Miles: Open-source PyTorch framework for large-scale LLM RL post-training combining SGLang, Megatron-LM, Ray, with MoE alignment and fault tolerance.
Operating system framework or concept for agent-based systems. Limited details provided.
Moo tool versions machine state for parallel AI agent development, solving isolation and collision issues. Developer tool.
Developer tool that logs AI agent actions and prevents agents from making false claims about execution. Agent reliability.
Framework or architecture pattern for building AI agents. Limited content detail available.
IBM and Red Hat release Lightwell tool protecting open-source code from AI attacks. Security tool.
SpaceX AI launches Grok 4.5 LLM trained for coding, agentic tasks, and reasoning. Competitive LLM release.
Discussion on building an agent sandbox platform with CLI interface and templates for deploying AI coding agents with Claude and other clients.
Foreman: open-source self-hosted LLM gateway with cost-aware routing across multiple models.
Open source formal verification system for requirements engineering that generates precise specifications and validation scenarios for AI coding agents.
Educational implementation of GPT-2 from scratch using JAX, breaking down components from bigrams onwards.
MailKite provides an alternative email API for autonomous AI agents, offering scoped addresses and parsed JSON push instead of Gmail API complexity.
Microsoft releases Flint, a visualization language designed to help AI agents generate reliable data visualizations by balancing simplicity and expressiveness.
Agent-zero-trust is a security scanner for code repositories before AI coding agents access them, preventing instruction injection attacks via README and config files.
Refusal training technique for LLM agents against masked MCP attacks. AI agent safety research.
Technical writeup on running small language models locally for AI-assisted coding tasks and agentic programming workflows.
Governance architecture patterns for autonomous AI security agents performing incident response and detection tasks.
vLLM transformers backend achieves native speed for LLM inference by optimizing transformers library integration.
Plugin identifying wasteful LLM API calls in code to reduce unnecessary spending.
SWE-1.7 model claims performance near GPT-4.5 and Claude Opus on software engineering tasks.
Technique to prevent AI agents from overfitting by implementing an 'oracle skill' constraint.
Quantum circuit compiler validated on IBM hardware outperforms Qiskit in speed and scales to 1M+ qubits.
Tool that loads and ranks GGUF-format LLMs on macOS directly from header metadata.
Native macOS image viewer with AI-powered OCR and annotation. Application tool with ML components.
Analysis of 1,080 YC startups on AI agent trends. Market research on agent adoption patterns.
Framework for preventing AI agents from executing irreversible actions without human approval.
Tool providing Obsidian vault access via Model Context Protocol for multi-device access.
Open-source memory/knowledge layer for AI agents. Addresses verification, evidence tracking, and semantic consistency for agent state across sessions.
Self-hosted AI agent platform supporting Telegram, Discord, Teams, terminal, desktop. Pluggable backends (Claude, OpenAI, etc.), MCP tool access, cross-platform binaries.
Tool tracking when AI crawlers (ChatGPT, Perplexity) send visitors. Analytics for AI-driven traffic.
Coinbase operates 1,200 agents and reduced AI costs 50%. Real-world agent deployment and cost optimization case study.
Open-source Node.js library providing multi-tenancy architecture for web applications.
Port of llama.cpp to Apple Watch enabling local 0.8B LLM inference on wearable hardware.
Framework for controlling AI agent behavior using session state and contextual policies. State-based agent governance system.
Declarative specification system for AI agents using Terraform-style syntax. Infrastructure-as-code approach for agent configuration.
Provider-agnostic agent skills router for LLMs. Dynamically injects relevant skills into prompts based on context, avoiding monolithic prompt issues.
Tool to audit website content readability and accessibility for AI agents.
AI agent architecture using NATS message bus for reactive node communication instead of conversational loops.
Open-source self-hosted investment research platform using bring-your-own-key LLM integration.
Tool generating LLM-readable codebase documentation with grep-friendly tags for better code understanding.
Tool enabling multiple AI coding agents to operate concurrently within a single Git repository.