Don't just paste the AI at me
Opinion piece criticizing developers who paste unedited LLM outputs directly to users without adding value or context.
Opinion piece criticizing developers who paste unedited LLM outputs directly to users without adding value or context.
Google's AI agents built operating system using Gemini 3.5 Flash and Antigravity 2.0 agent framework with minimal cost and single prompt.
BonzAI enables local LLM inference in the browser with self-sovereign control. Open-source LLM application tool.
MCP-safeguard: Security scanner for Model Context Protocol servers with 52 detection rules for automated vulnerability scanning.
Vehicle routing optimization handling 1M stops on Raspberry Pi 400 in 10 hours, proving efficient algorithms beat cloud infrastructure.
AI-assisted reconnaissance toolkit for security researchers combining automation with LLM triage prompts for prioritizing findings.
Analysis of twelve flawed approaches to measuring AI-assisted coding tool effectiveness and productivity metrics.
Cloudflare CEO argues AI has made certain worker categories obsolete as company grows and expands customer base.
Microcodegen.py generates FastAPI applications from PRDs without LLM calls. Code generation developer tool.
Research measuring LLMs' capability to develop security exploits. Evaluates LLM abilities on malware/exploit generation.
Local-first macOS dictation app using Whisper transcription with searchable history and optional AI cleanup on Apple Silicon.
Integration of xAI's Grok model as coding agent within OpenCode editor for X Premium subscribers.
Open-protocol AI memory system storing encrypted data locally on device with cross-platform compatibility across Claude, ChatGPT, Gemini. Apache 2.0 on npm.
Open-source database cataloging AI model specifications, pricing, and capabilities for comparison.
Personal perspective on using Claude Code as development tool and importance of code review practices for effectiveness.
Open-source agent-first tool tracking how AI systems cite and attribute sources.
Technical overview of embeddings as foundation for semantic search systems.
Llmff v0.1.2 tool providing FFmpeg-style pipeline abstractions for LLM workflows. Developer tool for LLM applications.
Debugging tool for AI agents with event capture, replay, and cryptographic verification. EU AI Act compliant.
Research on domain-camouflaged injection attacks targeting multi-agent LLM systems. Security analysis of agent vulnerabilities.
OpenCode project: OpenAI-compatible proxy API for Cursor IDE, enabling standard SDK usage with Cursor's inference engine.
Analysis of frontier AI lab compute utilization trends, showing most available compute capacity remains unused despite massive investment.
Security vulnerability: Google API keys remain functional ~23 minutes after deletion, allowing unauthorized access to Gemini and other APIs.
Google announces Gemini Omni Flash: multimodal model generating video, images, and other content from any input type.
CoreMem: portable context management tool for AI agents. Share context via URL, Chrome extension, MCP, Cursor/VS Code plugins, enabling persistent memory across agent sessions.
Brief report: Microsoft cancels internal Claude licenses after unexpected token-based billing costs. Minimal details provided.
Talk on using simulation sandboxes for safe AI agent development and testing.
Analysis of why AI inference costs have dropped 70-90% annually, driven by software optimizations rather than hardware, enabling local model deployment on commodity hardware.
Technical analysis of architectural debt patterns in AI systems: feature shipping stability followed by sudden cascading failures requiring rewrites.
Career advice on whether machine learning is still worth learning given LLM advances. Opinion piece from ML book author.
Talk on specialized data infrastructure requirements for AI agents in production.
Video content explaining how AI chips function at a hardware level.
Talk on Flower SuperGrid Agents framework for distributed agent systems.
Talk on optimizing accuracy, cost, and latency tradeoffs for real-world AI agents.
Developer survey of 7,258 respondents on AI's impact on software development work, with 2026 findings on trends and adoption patterns.
Google I/O announcement of generative UI in Search with AI agents and natural language mini-apps.
WebGPU backend implementation added to llama.cpp/ggml for LLM inference acceleration in web environments.
Agents League: week-long esports-style hackathon for AI agent development with live coding battles and competitions.
Interactive 3D browser-based climate physics simulator built with AI coding assistance (Claude Opus); explains climate processes.
Open source Telegram-based personal AI agent with memory, skills, tools, and persistent workspace; model-agnostic.
Open-source FPGA-based 400G SmartNIC with LLM-generated testbenches for verification. Uses AI for cocotb Python workflows.
Mock OpenAI/Anthropic API server for testing LLM applications and agent flows without API costs.
Conference talk on building agent-native office software with multiple agents.
MCP server adding persistent, searchable long-term memory to Claude Code using local SurrealDB without APIs.
Conference talk on agentic engineering paradigm shift.
Personal reflection on using AI coding assistants in development workflow and cognitive impact on learning.
LLMs perform better with consistent, less fragmented programming languages and ecosystems; consistency compounds agentic output quality.
Open-source IDE for running multiple coding agents (Claude Code, Codex, OpenCode) in parallel. YC P26 company.
TypeScript rewrite of Prisma ORM with data contracts, migration graphs, and integration skills for AI agent tooling.
Visual workflow builder that Claude can control via MCP (Model Context Protocol) for agent automation.