Show HN: `uvx ptn` and expose any system to agents (dangerously)
Tool enabling agents remote access to Linux/Mac/PC systems via simple CLI with phone-based kill switch.
Tool enabling agents remote access to Linux/Mac/PC systems via simple CLI with phone-based kill switch.
Personal account of month-long experience building IntelliJ plugin using AI coding at reduced velocity.
Benchmark comparison of AI gateways (GoModel, LiteLLM, Portkey, Bifrost) for production performance and cost optimization.
AI agent design pattern for handling hold states by pausing runtime execution during call holds.
Notion shutting down Notion Mail, shifting focus to AI agents for inbox management. Product strategy pivot toward agent automation.
Video walkthrough of AI-assisted coding workflow by Matt Pocock. Title only, no technical details available.
Open-source AI coding agent supporting open-weight LLMs with 60+ tools, 7 modes, and specialist personas.
Comparison of 15 AI agent frameworks evaluated across 4 production deployment stacks.
Tutorial: building RAG application using Telnyx AI Inference service.
BetterDB: open-source Valkey-native context layer for AI agents with memory, semantic caching, typed retrieval. Available on npm/PyPi.
Case study: module decomposition reduced AI agent token usage 32.6%, execution time 22.9%, costs 24.6% on follow-up feature work.
OpenSpec for spec-driven development with AI. Open Code Review runs multiple review agents in parallel to verify specs and code.
CLI tool for finding container base images, displaying CVE counts, and pinning by digest without local scanning.
Subquadratic claims breakthrough on mathematical bottleneck limiting LLMs. Independent evaluation results shared but skepticism remains.
Open-source GPU orchestration gateway that scales cloud inference to zero when idle, eliminating cold starts.
ESI standard for AI agent memory systems to report confidence and freshness of recalled information.
Research paper examining shift toward agentic AI with evidence from Codex data.
Ornith-1.0 open source LLM with self-scaffolding for agentic coding claimed comparable to Claude.
Adobe acquires Topaz Labs for AI video/image enhancement tools. M&A news with product capabilities mentioned.
Open-source AI agent system for automated video production. Agents handle research, scripting, asset generation, editing.
Show HN about guardrails for offensive AI agents. Limited content visible.
Analysis of open-source LLM cost competitiveness matching closed-source benchmarks at 87% lower cost. Developer economics and model selection insights.
Framework for developing AI agent skills for UI/design tasks using Codex. Agent-focused development tool.
Tutorial on deploying OpenClaw AI agent on VPS with cloud infrastructure. Practical deployment guide.
Study on vulnerability of AI recommendation systems to manipulation via single fake webpage. ML security research.
Git-lazy-mount: tool for mounting large repos without full clone, useful for AI coding sessions requiring partial file access. Open-source developer tool.
Book on modern GPU programming for ML: attention kernels, LLM prefill/decode, fused operations. GPU optimization for AI workloads.
Analysis of shift from token consumption maximization to deliberate context curation in coding agents. Economics and architectural patterns of LLM applications.
Analysis of AI thought leaders' posts in Q2 covering Anthropic/OpenAI valuations and model performance.
Unofficial API wrapper accessing GPT-4/GPT-5 via Microsoft Copilot web interface without API keys. Violates ToS.
OpenAI announces GPT-5.6 series with three models: Sol (flagship), Terra (balanced, 2x cheaper), Luna (fast, affordable). Includes safety improvements.
llama.cpp's ggrun auto-tuning tool for GGUF model optimization across multiple GPUs without manual flags.
Security scanner for MCP server configurations used by AI agents. Detects dangerous permissions, hardcoded secrets, missing guardrails.
Exa raises $250M Series C at $2.2B valuation to power web search for AI agents and LLM applications.
Engineering case study: Claude AI helped fix flaky tests but required significant iteration to become production-ready.
Open source audit layer for LLM-as-judge systems that validates verdicts against supporting evidence.
Tracks AI capability predictions against actual progress; mentions Agent systems and research acceleration.
LLMs enable rapid Pandas-to-Polars migration for data engineers; increasingly used for code translation tasks.
Analysis of LLM inference cost models: kilowatt-hour billing vs token-based pricing for open-weight models.
Technical guide on building autonomous agents for penetration testing and security assessments.
Open-source home server OS for Docker apps with personal cloud aspirations. Limited technical depth in excerpt.
Performance evaluation of GitHub Copilot's agentic capabilities and efficiency metrics.
Analysis of LLM operational cost sustainability challenges.
Curated library of resources for building and evaluating AI agents including papers, tools, courses, and benchmarks. Maintained collection with verified annotations and dead links removed.
PatentScore introduces a multi-dimensional evaluation framework for assessing LLM-generated patent claims. Published at EMNLP 2025, presents research on evaluating LLM output quality.
Notion acquired Skiff email and is pivoting to AI agents for inbox management, shutting down Notion Mail. News about enterprise AI agent adoption.
HN discussion on gap between AI demo performance and real-world deployment challenges in AI voice applications for small businesses.
tuicr is a terminal UI for code review with vim keybindings, GitHub-style diffs, and line-level comments. Integrates with GitHub PRs and supports structured markdown output for coding agents.
Research on converting brain prediction models into testable explanations. LLMs predict human cortical activity from language but lack interpretability.
Developer shares SlimSnap tool that feeds JSON to coding agents instead of screenshots to reduce token costs and improve accuracy.