Demystifying Security Risks of AI-Powered Applications on Pre-Trained Model Hubs
Security research examining vulnerabilities in AI applications using pre-trained models from model hubs.
Security research examining vulnerabilities in AI applications using pre-trained models from model hubs.
TRAINS is a formally verified leaderless total-order broadcast primitive in Rust for control-plane coordination, alternative to Paxos/Raft.
Technical optimization for LLM inference on AMD GPUs using low-latency GEMM implementations.
ZLUDA CUDA-compatible runtime for AMD GPUs adds 32-bit compatibility and PhysX/Blender support despite losing funding.
HN discussion thread asking about origins of AI terminology like 'system card', 'alignment', and 'safety'.
User reports that installing Cursor iOS app irreversibly changes privacy settings, downgrading from legacy 'Do not store code' mode.
Discussion of agentic AI systems replacing static information retrieval with dynamic autonomous decision-making.
Security technique embedding obfuscated text in malware to complicate AI-based analysis and detection.
Tool optimizing local LLM inference with 39% faster time-to-first-token and 46% faster agent execution through adaptive resource management.
Technical guide for implementing LLM training loops using JAX framework with code examples.
Minima agent harness with dynamic model routing and memory management for generalized AI task execution.
Self-hosted, model-agnostic multi-agent AI assistant supporting multiple LLM backends for deployment flexibility.
AgentWire is a self-hosted orchestration tool for managing multiple Claude Code agent sessions with tmux, command palette, voice control, and no telemetry.
Analysis of prompt injection vulnerabilities beyond chatbots, examining broader attack surfaces in LLM systems.
Plugin connecting self-hosted WordPress sites to LLM providers for content integration and AI features.
Analysis of Anthropic's coding agent control model architecture showing patterns for securing autonomous agents with delegation, per-action authorization, and attribution.
Collection of LLM prompts for editing technical documentation using Claude Code for improved writing quality.
Open-sourced internal infrastructure for managing coding agents with team collaboration, context sharing, and observability features.
Agent-QA: open source self-improving QA agent for software testing. 500M+ tokens processed, seeking community feedback.
Title-only stub about governing AI agents operating on sovereign data. No content provided.
Trajeckt: open source Rust firewall for AI agents. Enforces action plans with <5ms latency via trajectory tracking.
Anthropic launches Claude Science, an AI workbench for scientific research with MCPs and domain-specific capabilities.
Editorial on AI coworkers and workplace integration. Opinion piece from MIT Technology Review newsletter.
Analysis of 1,000 bugs from AI-generated code across 99 codebases, identifying patterns in agent-generated bugs and automated code review limitations.
Using LLMs to reverse-engineer security products and discover evasion techniques for EDR/AV systems.
fenic: dataframe API with LLMs as first-class operators for unified structured/unstructured data querying.
Benchmark results for running local LLMs on AMD Ryzen 8700G iGPU with throughput measurements.
Claude Code session URLs leak into Git history in v2.1.179, a security issue for users of the coding agent tool.
Performance comparison showing LLMs achieving 5x faster execution in sandbox environments.
Technical approach to reducing CI feedback latency for developers and AI agents using local testing.
Closed-loop controller system preventing GPU out-of-memory errors during large language model training.
Open-source omnichannel AI marketing platform alternative to ManyChat and Chatfuel, supporting Claude and other LLMs.
Case study using Claude AI for end-to-end testing on Airbnb clone, covering test planning, spec generation, and CI/CD integration.
Google releases Gemini Nano Banana 2 Lite and Gemini Omni Flash for faster, cost-efficient image and video generation.
ECP: Vendor-neutral protocol for portable agent evaluations across frameworks, models, and CI systems. Standardization tool.
Capacitor: Shared memory system enabling context persistence across Claude Code, Cursor and other coding agents. Open-source developer tool.
Security testing comparison of U.S. and Chinese LLMs by Booz Allen analyzing code vulnerabilities.
Technical assessment platform evaluating candidate skills in collaborating with AI tools—prompting, reading output, debugging. Developer tool for hiring.
AI agent-managed landing page builder for creators and small businesses with automated content updates.
FastAPI and React starter kit for AI SaaS applications with token metering and AI copilot demo.
Busabase is an approval-first database and knowledge base designed for AI agents with built-in human oversight.
ML research on training smaller models to learn when to defer to larger models instead of building explicit routers.
Tool management system for agent tools providing installation, versioning, discovery and execution via single configuration file for multi-agent projects.
Multi-agent system for stock research using Claude Code and Robinhood MCP with human-in-the-loop approval workflow.
Self-contained autonomous control loop binary for agent-driven work, integrating with GitHub/Linear/Grafana and spawning worker agents via LLM decisions.
CLI-based LLM REPL built with bash and Unix tools, demonstrating minimal-dependency approach to local language model interaction.
CLI tool for AI agents to scaffold and manage BigQuery, dbt, and Cube projects. Developer tool for data infrastructure.
Core ML port of Rampart PII token-classification model with Swift package. Open source ML tool for local data privacy.
Aitori: open-source traffic inspection tool for Claude and ChatGPT API interactions. Developer utility for debugging.
Research-Git tool captures code ideas as semantic units and regenerates them onto current codebases, compatible with Claude Code and MCP clients.