Direnv Is All You Need to Parallelize Agentic Programming with Git Worktrees
Using Git worktrees and direnv to parallelize AI coding agents like Claude Code. Developer workflow optimization.
Using Git worktrees and direnv to parallelize AI coding agents like Claude Code. Developer workflow optimization.
Kube-pilot: autonomous AI agent running in Kubernetes that writes code, builds containers, deploys, and closes tickets.
Testing LLMs for matching decompilation across 60 functions using AI-powered VS Code decompiler tool. Evaluation.
Opinion piece arguing for optimizing web content specifically for AI agents rather than just humans and APIs.
DAAO deploys AI agents to servers via zero-trust outbound-only mTLS tunnels, enabling remote coding sessions without VPN or SSH exposure.
Research studying LLM behavior in Ultimatum Game with varying stake sizes and opponent types, showing heterogeneous behavior across models.
Post comparing workflows versus agents in agentic AI systems. Minimal content provided.
Turbopentest enables collaborative pentesting between AI agents and human operators via web, VSCode, Burp Suite, GitHub Actions, and MCP servers.
AutoContext is a closed-loop system that improves AI agent behavior by executing tasks, evaluating outcomes, updating knowledge, and distilling successful behaviors into cheaper local models.
Claude Skill teaching Rails conventions for LLM calls, providing patterns for retry logic, cost tracking, and prompt management.
OS with 38 specialized AI agents organized into 5 categories, runs inside Claude Code with interactive setup and structured command center.
ad-injector is a Python library that embeds agent-targeted instructions into JSON API responses for affiliate links and product recommendations.
Research on spaced repetition scheduling algorithms optimizing memory retention. Academic papers on algorithm dynamics for learning applications.
CLI tool for simplified SSH key exchange between machines without manual copying. Developer tool with limited AI/ML relevance.
Execwall: execution firewall for AI agents using seccomp-BPF filtering to prevent prompt-injection command execution exploits.
Developer asks about building autonomous shopping AI agent; discusses MCP payment integrations and infrastructure challenges.
Tool that learns Claude Code user preferences and injects them automatically. Developer productivity utility.
Revo AI building ambient AI agents using email as context substrate, leveraging existing protocol infrastructure for cold-start grounding.
Comprehensive textbook on probabilistic machine learning with reproducible code, figures, and exercises. MIT Press publication with CC-BY-NC-ND license.
Using AI to generate end-to-end tests from GitHub PRs to address gap left by Copilot-style tools lacking test coverage.
macOS voice-to-text app running Voxtral 4B locally via MLX framework. Zero data leaves device. Swift/SwiftUI implementation.
TinyForge: 0.8B coding model learns from test failures via evolutionary search and LoRA training on MacBook, improving HumanEval performance.
Open-source browser agent for Chromium sidebar automating clicks, typing, form filling. Alternative to Perplexity Comet and ChatGPT Atlas.
Redis-based coordination framework for multi-agent AI systems with reduced setup overhead using cursor agent experimentation.
Fine-tuned Qwen3-4B LLM for stock trading using 5-stage supervised learning pipeline and reinforcement learning. Achieved +9.4% returns.
Agent skill enabling coding agents to render interactive SVG diagrams, HTML widgets, and live charts inline. Developer tool for AI agents.
Original research on emergent abilities in text-to-image models, discovering image-to-image capabilities. Reproducible experiments in preparation.
Open-source evaluation suite for LLM-as-judge testing AI agents. YAML test definitions, root cause analysis, failure mining into training data.
Ruby library for plotting mathematical functions in Jupyter notebooks. Developer tool with limited AI relevance.
Identity.txt: Portable text format for storing AI custom instructions, preferences, and voice across multiple AI tools.
Rootly CTO discusses rethinking engineering evaluations using conversation transcripts instead of code artifacts in AI era.
Covenant Layer: Open protocol for AI agents to coordinate commitments via outcome-based contracts instead of step-by-step tool orchestration.
Open-source ChronologyAI engine reconstructs event timelines and detects contradictions in documents for legal, compliance, and fraud investigations.
OpenClaw agent templates for healthcare with plug-and-play deployment, customer support automation, and PR review capabilities.
OpenClaw agent templates for healthcare including support ticket handling, bug triage, and PR review with customizable guardrails.
AutoHarness research on improving LLM agents by automatically synthesizing code harnesses. Machine learning research.
LightSwarm is a bash script that creates a 3-agent swarm using Claude's API, with roles for architecture, building, and cleanup across multiple projects.
Harbor CLI tool for managing multiple LLM backends (llama.cpp, vLLM, Ollama) with unified interface. Open source developer tool.
Simple command-line utility for looping and backgrounding commands with configurable delays and alias support.
Brief conceptual piece discussing topology and authority in AI inference versus historical primary source concepts.
Offline detection tool for AI-generated text patterns without ML, identifying suspicious fingerprints like unusual punctuation and buzzwords.
Open source editor-agnostic live collaborative editing tool enabling simultaneous editing across different editors and locations.
Open source CLI tool converting Playwright test scripts into product demo videos with AI-generated voiceover using local Kokoro TTS.
Systems-level discussion of LLM inference basics and serving runtimes ecosystem from infrastructure perspective. Originally internal, now published for systems developers.
GitHub removes expensive premium LLM models from free Copilot Student plan starting March 12.
Hawkeye is an open-source observability layer for AI agents with session recording, drift detection, and guardrails.
Kalverion_bot is an open-source AI Telegram bot for personal finance using NLP, double-entry accounting, and forecasting.
Continuum is a testing framework for LLM workflows that records and replays runs to detect output drift in production.
Open source minimal ML research paper reader that fetches arXiv papers daily and summarizes them using local LLMs.
Monet is a grid-based management interface for organizing and monitoring multiple Claude Code agents on desktop and mobile.