Knowing, Remembering, Exactly, Vaguely: An Agent-Native Database (PlatypusDB)
Design note on PlatypusDB, an agent-native bitemporal event database combining crisp Datalog queries with fuzzy VSA resonance.
Design note on PlatypusDB, an agent-native bitemporal event database combining crisp Datalog queries with fuzzy VSA resonance.
Desktop voice AI agent using locally-run Gemma model for dictation, cutting latency 44% versus cloud processing.
Open-source static analysis tool scanning AI agent source code offline for security risks and governance gaps, built using GPT Codex, DeepSeek, and Kimi.
arXiv paper abstract on achieving near speed-of-light latency in GPU collective communication operations.
A control center app that discovers and manages running coding agent sessions (Claude Code, Codex) across terminals, with mobile/LAN access.
Model Context Protocol (MCP) implementation allowing agents to create and execute Playwright scripts as callable tools.
Design rules engine enabling AI agents to generate non-generic UI by enforcing design consistency patterns.
All-in-one platform offering free access to AI models, PDF tools, and image processing. Limited technical depth.
all2md: Universal document converter (PDF, DOCX, HTML, 40+ formats) to Markdown with MCP integration for AI assistants. Python library and CLI tool.
Technical tutorial on algebraic foundations underlying FlashAttention algorithm for efficient transformer computation.
Benchmark comparing 19 LLMs on Flutter code generation, measuring compile pass rates and hidden-test pass rates.
Production-grade LangGraph multi-agent template with FastAPI, Helm, Terraform, CI, and observability. Deployable agent framework for ML engineers.
Custom Triton kernel optimization for NF4 dequantization in 4-bit LLM inference, achieving 1.41x speedup over bitsandbytes with code and benchmarks.
Unix coreutils reimplemented to output XML/JSON format, designed specifically for AI agent tool integration.
WASM-based LLM inference engine using wllama with integrated model manager for local deployment.
Security report: Grok Build CLI uploads full Git history and .env secrets to xAI cloud without user awareness. Privacy/security concern.
Copy-paste UI design templates for Claude/Codex integration to reduce generic AI-generated interfaces.
Open-source tool preserving Claude agent decision transcripts locally to prevent loss of conversation history.
Open-source MCP tool for cloud/AI cost optimization with bill normalization and savings verification across providers.
PixelKit: Linux native alternative to Microsoft PowerToys written in Rust, supporting X11/Wayland. Open source developer tool.
MCP tool for Unity Editor integrating 90 research paper algorithms for game development AI.
Platform for building and deploying autonomous AI agents from natural language prompts.
Browser-based AI video editing tool that processes locally without cloud uploads.
NixMC: macOS native app using Claude Code to manage nix-darwin configuration via natural language. AI agent for system administration.
TypeScript repository demonstrating architecture patterns that constrain AI agent behavior.
Anthropic research mapping how Claude exhibits different values across languages, identifying four key behavioral axes.
Daily AI news update: Claude Managed Agents lifecycle improvements, Gemini 3.5 Flash GA release, Genkit Python support.
Html2realpdf: Browser-based PDF generator written in Zig/WebAssembly, produces selectable vector text instead of screenshots. Developer tool with TypeScript API.
Open-source AI-native framework for building data portals with agent-assisted scaffolding and Next.js architecture.
Uptime monitoring service with MCP server enabling Claude/ChatGPT to create and manage monitors via natural language prompts, curl, or GitHub Actions.
User benchmark comparing GPT-5.6-terra vs Mimo-2.5-pro LLM agents on token usage and task efficiency. Real-world performance data from agent task history.
Tutorial building voice-based AI agent for phone quotes using Telnyx Voice AI API.
Tutorial optimizing Ryzen AI Halo hardware for 10-15% faster local LLM inference performance.
Explainer comparing LLMs and AI agents to aliens. Educational but low technical depth.
LoopGain tool uses control theory to prevent AI agent loops as alternative to max_iterations parameter.
Analysis of AI skill adoption in tech job market, tracking mentions of AI/ML skills by seniority level and job category.
TermCanvas tool provides tmux-based interface for steering and controlling AI agents on macOS.
Economic analysis of LLM pricing models, examining why flat-rate pricing (GitHub Copilot $10/month) became unsustainable and how model owners shifted to usage-based pricing.
Reverse-engineered analysis of three agent-memory tools (Cognee, Graphiti, Neo4j) revealing convergent knowledge-graph architectures and discussing limitations of existing implementations.
Domain-Specific Languages enable reliable LLM code generation through clear boundaries and abstractions. Tickloom example with distributed systems DSL.
ConlangCrafter: LLM multi-stage pipeline generating complete artificial languages with phonology, grammar, lexicon. Open dataset and research project.
Python tool for managing multiple Git repositories as a unified workspace, addressing limitations of Git submodules.
Analysis of AI agent capabilities vs. hype, citing WebArena benchmarks showing 45.7% success rates and definitional ambiguity in the field.
Interactive 3D visualization comparing Ptolemaic cosmology to LLM architecture layers, illustrating forward/backward passes.
Technical analysis of disaggregated LLM serving with vLLM prefill and TileRT decode, separating compute-bound and memory-bound phases.
Agentic Python IDE with built-in small LLM, emphasizing error recovery and self-healing pipelines.
AI agent memory engine designed to replace Obsidian, achieving perfect recall across thousands of facts in 8MB, runs offline.
Deep technical dive into TPU/GPU cluster topologies and collective communication patterns for transformer training and inference at scale.
Vehir is an experimental AI-native computing platform designed with agent-computer interaction at its core, featuring a compiler, microkernel, and CAS built for structured I/O for AI agents.
Security audit tool for Model Context Protocol servers, providing trust cards for agent-server connections before deployment.