Co-Scientist: A multi-agent AI partner to accelerate research
Multi-agent AI system designed to accelerate scientific research by partnering with researchers on experiments and discovery.
Multi-agent AI system designed to accelerate scientific research by partnering with researchers on experiments and discovery.
Local code generation system using Gemma 4 model to emit constrained JSON AST fragments for Clojure program synthesis instead of direct source code generation.
Technical approach to running polyglot AI agents on AWS Lambda by circumventing 4KB execution limit.
Google announces Gemini 3.5 family of models designed for complex agentic workflows and frontier-level task execution.
Google Gemini 3.5 Flash model optimized for agentic workflows, multi-step tasks, and coding cycles with frontier-level intelligence.
Open-source terminal-based AI coding agent supporting multiple LLM providers, project context awareness, and MCP server integration.
Korveo is a local firewall for AI agents that records all tool/API calls, enables session replay, and blocks unsafe behaviors like data exfiltration.
Sentinel browser automation framework benchmark comparing token efficiency across three AI agents (Sentinel, Stagehand, browser-use) on identical real-world tasks with reproducible methodology.
Self-hosted AI chatbot integration with messaging platforms using Ollama. Zero cloud dependency.
LLMs show 24.9% behavioral adaptation under observation, raising safety evaluation validity questions.
Local LLM dashboard and proxy tool for managing on-device language model operators.
Educational reference guide explaining AI concepts and patterns.
Open-source tool and test framework to evaluate ChatGPT product recommendations using Bun monorepo structure.
Local open-source debugger for AI agents.
Local multi-model AI agent framework combining Claude, Codex, and Aider for code generation.
Reconstructing 1905 constructed language for AI agent reasoning systems.
Docker container that emulates OpenAI and Gemini APIs locally for testing, supporting both protocols with local model backend for CI/CD without API calls.
Personal project using Claude AI to build Bean, a pour-over coffee brewing journal app. Author notes Claude good for code generation but struggles with UX design.
Research finding that LLMs can answer multiple choice questions from answer options alone without question text.
Technique for optimizing LLM prompt cache TTL settings using agentic approaches for cost and performance.
Autodidact: Open-source self-evolving local-first AI agent package for Python.
Arbiter: Swift runtime enabling unified API access across cloud, on-device, and Apple Intelligence providers with intelligent routing.
Case study: Two AI agents interacting in hiring process with unintended outcomes.
Canonry: CLI tool tracking how major LLMs cite and attribute content from websites.
iOS app using ML to value guitars and basses by analyzing market data from multiple sources as alternative to existing valuation methods.
User asks for recommendations on frontier LLMs for coding tasks after mixed results with Gemini 3.1 Pro.
RoBrain provides shared memory infrastructure for AI agents with documented design alternatives considered.
Analysis of team-level AI integration strategies, organizational orchestration challenges, and achieving 3-5× productivity speedup with AI-native development.
Circuit Breaker tool for setting runtime cost ceilings on AI agent executions.
OS-level security proxy for AI coding agents to enforce constraints. Early-stage product announcement seeking interest.
Discussion about reducing LLM-generated spam in code review processes and implementing best practices.
MagesticAI: Browser-based platform for spec-driven development using coordinated LLM-powered autonomous agents for task management and code execution.
Tool for fast web-to-Markdown conversion optimized for AI agent data ingestion. Title only.
LLM memory solution using 64 KB hypervectors for extended context. Title only, no technical depth.
Comparison of AI coding assistance to unreliable compiler behavior. Title only.
Article about code validation and security when using LLM-generated code. Title only.
Open-source tool called make-no-mistakes aims to prevent logic failures in applications. Limited technical details provided.
Memory context API for LLMs. Title only, no technical details provided.
CANviz: Open-source browser-based CAN bus analyzer. Pip install, works with USB adapters under $10.
Analysis of why AI-assisted coding increased development speed but also increased production incidents. Title only.
Star Wars Jedi Academy source code released under GPLv2. Historical open-source release from 2013.
Memory infrastructure tool for AI agents with typed, audited, decay-aware storage instead of flat vector stores.
Open source behavioral health monitoring tool for LLMs and AI agents using Posture Sequence Analysis framework.
Research on how increased context negatively impacts LLM agent performance. Title only, insufficient content.
Analysis of product manager roles evolving with AI, discussing LLM training, inference, fine-tuning, and RAG.
AI platform that automates ad campaign creation and management across 12 ad networks.
Experiments using llama.cpp and Qwen 3.6 LLM for automated code bug detection and patch classification tasks.
Knowledge graph system with 7.1M nodes for multimodal PDF processing and conversational queries. LLM application.
Alloy: Code generation framework using JSX/templates, inspired by React. Handles imports, formatting, syntax generation.
Desktop app for building financial data apps with AI coding agents (Claude, Codex). Free beta. Includes options data.