Show HN: Skillmaxxing – make every agent self-evolving
Open-source agent plugin enabling self-improvement by reflecting on work and creating reusable skills. Reduces token usage and improves execution efficiency.
Open-source agent plugin enabling self-improvement by reflecting on work and creating reusable skills. Reduces token usage and improves execution efficiency.
Discussion of LLM access restrictions based on citizenship affecting non-US founders and model availability across Anthropic and OpenAI.
Show HN submission for all-in-one memory system for AI agents with minimal details provided.
Research paper on how AI agents enable adaptive computer worms with evasion capabilities.
Agent skill that generates bookmarked PDF cheatsheets optimized for reMarkable e-ink tablets with single-page subject organization.
OpenTag is open-source self-hosted AI agent for Slack that reads threads, calls tools, and renders results without per-seat pricing.
AI audio translation system combining speech-to-text, LLM translation, and text-to-speech for multi-language audio processing.
Even: terminal-first desktop workspace for managing multiple development projects with built-in browser and designed for AI code agent workflows.
Deterministic financial analysis tool for SEC filings detecting fraud signals without using LLMs, with historical backtests on failed companies.
Desktop AI pet companion app featuring language learning, homework help, and screen context awareness.
Collection of 60 production-ready AI agent blueprints with prompts, architecture, and tools across 30 categories, implementation-agnostic.
Analysis of performance gap between open-source and closed-source LLMs, predicting frontier open-source model release by December 2026.
Discussion asking which emerging AI concepts will persist versus fade, including agents, agentic workflows, and Model Context Protocol.
Linux Foundation and others launch Akrites initiative to defend open-source projects from AI-based exploits.
JackHamr platform offering $5k compute and LLM credits to qualified AI agent startups with infrastructure support.
2016 video presentation by Dario Amodei on concrete safety problems in AI systems.
Technical article on context management vs memory in AI agents, addressing context window overflow challenges.
Benchmark comparing web search API providers using LLM judges on 32 real-world agentic research tasks with ELO scoring.
Explicode: tool enabling literate programming by embedding rich Markdown documentation in code comments for humans and AI agents.
Guide on integrating language servers with GitHub Copilot CLI for improved code intelligence.
Statey database tool enabling AI agents to share persistent state across chats via Model Context Protocol.
MirrorCode benchmark co-developed with METR testing AI models on long-horizon full-program reimplementation tasks without source code.
Security vulnerability in Amazon Q AI coding assistant allows executing code and stealing cloud credentials through malicious Git repos via Model Context Protocol mishandling.
Secret store for AI agents that prevents plaintext exposure in logs, transcripts, or cloud systems.
Open-source MCP server providing visual canvas for AI coding agents to design UI before framework implementation.
Open-source AI sales development representative agent that sources leads, writes outreach, and books meetings.
Self-hosted LLM gateway for small teams with single-command AWS deployment.
Tutorial building voice AI workflow for insurance claims intake using branching logic instead of monolithic prompts.
METR's independent external evaluation of GPT-5.6 Sol on software task benchmarks. Conducted under NDA with OpenAI review approval.
Linux Foundation initiative to identify and fix vulnerabilities in open-source software threatened by AI-accelerated attacks.
Curated list of software services offering free developer tiers for infrastructure developers (SaaS, PaaS, IaaS).
My-Pi is curated Pi coding-agent distribution with MCP, LSP, skills, recall, redaction, telemetry, and team mode preinstalled.
OpenAI's technical documentation for GPT-5.6 family (Sol, Terra, Luna). Details safety safeguards and planned general availability timeline.
Murmur is shared communication bus for coding agents. Multiple LLM agents (Claude, Copilot, etc.) communicate via single MCP HTTP daemon with @-mentions.
Business platform helping solo founders with non-technical aspects beyond AI code generation like validation and positioning.
Analysis arguing coding agents need durable codebase memory over larger context windows for effective development.
Hasp is a local secret broker for AI agents/apps. Encrypts secrets in vault, controls access by project boundary and time window, never exposes raw values.
Case study of AI-assisted software porting from one language to another with code generation artifacts.
Ask HN discussion on monitoring AI agent output quality degradation in production beyond cost/latency metrics.
Ask HN thread from independent AI researcher with novel analytical framework for neural representational models seeking career options.
Open standard markdown file (hallucinate.md) to instruct AI coding agents not to hallucinate. Simple specification for repository integration.
OpenAI announces GPT-5.6 series with three models: Sol (flagship), Terra (balanced, 2x cheaper), Luna (fast, affordable). Includes safety improvements.
Gartner analysis of escalating costs from AI coding agents shifting to consumption-based pricing, creating unpredictable developer expenses.
Mac-native CLI-forward coding agent multiplexer with GUI, designed to expose all functionality for automation and multi-agent coordination.
Opinion piece criticizing White House ad hoc decision-making process for frontier AI model access as lacking transparency and formal procedures.
Unified LLM gateway supporting 34 models via single OpenAI-compatible endpoint with per-token pricing and USDT payment.
Research comparing 67 frontier LLM models showing combining models rarely outperforms best single model.
Model router that intelligently routes requests to optimal LLM for coding agents like Claude, Codex, Cursor to reduce costs.
Free tool checking if websites are discoverable and citable by AI search engines like ChatGPT, Claude, and Perplexity.
persist-os CLI tool that records AI coding decisions and architecture to enable better code quality and scalability beyond rapid prototyping.