Show HN: Open-source sync-engine for managing websites at scale
Open-source tool for syncing CMS content (Ghost, WordPress, Shopify) to local markdown files, enabling local editing and AI tool integration with push-back capabilities.
Open-source tool for syncing CMS content (Ghost, WordPress, Shopify) to local markdown files, enabling local editing and AI tool integration with push-back capabilities.
Netflix open-sources Project Headroom, software that optimizes AI costs by pruning agent instructions before LLM processing.
Fine-tuning an LLM locally to write technical documentation in 1980s-90s style. Explores local-first LLM experimentation.
Cryptographic identity system for AI agents using Ed25519 and DPoP. Enables secure API calls and prevents token theft.
UI design patterns for LLM applications beyond chat: table-based comparison and tree-based exploration with live examples.
Guide to implementing world models with JAX including MPC, discussing competing definitions and recent approaches.
Research analysis of why LLMs struggle at video game playing and coding despite rapid improvements on benchmarks.
Meta fixed AI chatbot security vulnerability in Instagram support system used for account takeovers.
AI agents harness for portfolio construction supporting multiple LLM providers with live leaderboard paper trading performance.
Technical explanation of how large language models work mechanically, covering transformer fundamentals and text generation.
Eight practical, copy-paste-ready patterns for using AI during engineering work including PR analysis and code review.
Using Claude Opus to build threat models, discover vulnerabilities, and patch source code. Best practices for LLM-assisted security testing.
Personal reflection on a year of using AI tools in daily work, questioning productivity claims and actual value delivered by AI applications.
CLI tool for querying Apple Find My locations from Linux, outputting JSON for AI agents and automation workflows with location-aware capabilities.
AI agent that converts product briefs and assets into storyboarded launch videos using Remotion, with voiceover direction and asset generation.
Analysis of how AI coding agents impact developer understanding and system knowledge, examining productivity and technical debt tradeoffs.
1B parameter model with stacked LoRA adapters achieves human-level writing on AI detector using local inference on 24GB Mac.
Open-source library of Design.md files for AI-generated UI design systems across multiple tools.
Emergence World: research platform evaluating long-horizon autonomous agent behavior in shared environments over weeks with compounding effects.
Homebrew maintainer details secure agentic AI setup using sandboxes and worktrees, achieving 90% code generation with AI tools.
MiniMax M3 open-weight model available on Qubrid AI platform with 1M context window and multimodal capabilities.
Running Gemma 4 MTP drafters quantized model on old hardware without GPU using 128GB DDR3 RAM and 2016 Xeon processor.
Developer used AI agent (Cursor) to build automated video editing tool, with the walkthrough video edited by the tool itself.
Sparse autoencoders outperform simple baselines for steering LLM output on AxBench model steering benchmark.
TouchSafeBench for evaluating vision-language model collision grounding capabilities in human-robot collaboration safety.
Benchmark and enhancement of text-to-image models for generating pedagogically meaningful visuals from arithmetic equations.
Zero-shot cross-lingual confidence estimation for multilingual LLMs identifying language-transferable confidence features.
Mixed-methods study comparing LLM-based conversational and graphical interfaces for industrial IoT data analysis decision tasks.
Survey of 70 on-device learning works for TinyML addressing post-deployment distribution change on microcontroller devices.
EchoRL method using rollout echoing to improve reinforcement learning-based post-training for LLM reasoning capabilities.
Dynamic adapter routing for continual multimodal retrieval in vision-language models beyond class-incremental learning.
Entropic Projection Alignment framework for estimating model performance, explaining, and improving performance under distribution shift.
ERGeoBench benchmark for evaluating multimodal LLMs as embodied agents in geo-localization tasks across single/panorama/embodied views.
Theoretical analysis of linear recurrent neural networks as memory units in partially observable reinforcement learning.
GUIDE physics-guided deep unfolding framework for cross-band channel prediction in AI-native RAN with real-time inference.
DeMaVLA vision-language-action foundation model for generalizable robot manipulation of deformable objects across diverse conditions.
Compares LLM-based conversational agents versus graphical dashboards for industrial decision support in manufacturing settings.
Terminal representation approach combining successor and default representations for spatio-temporal abstraction in reinforcement learning.
Mechanistic interpretability framework for Multitrack Music Transformer enabling attribute control via activation steering without retraining.
Local inconsistency measure for estimating generalization gaps and improving deep learning models using unlabeled data.
FBHM benchmark for evaluating vision-language models on hateful meme detection with systematically curated functional axes.
Python library for characterizing dataset shifts between train/test distributions to support trustworthy AI development and deployment.
Extends world models like Dreamer to multi-agent RL by modeling teammate policies and intentions as structured latent variables.
Introduces sCWL and fCWL tests and maximal clique complexes for scalable higher-order graph neural networks preserving expressivity.
DynaTree framework for agentic RAG with two-stage architecture for efficient time-sensitive news retrieval without high inference cost.
Uses GPT-4o to generate paraphrase variants for sign language translation training data augmentation on limited corpora.
Analyzes how LLMs' linguistic biases affect spatial reasoning in navigation planning systems using text-based spatial representations.
SkillsBench study investigates how skill document granularity affects LLM agent task success across 30 tasks with controlled conditions.
CYKNN embeds the CYK parsing algorithm directly into neural network architecture as trainable matrix operations for context-free grammar parsing.
Training-free attention policy for decoder-only SpeechLLMs to enable simultaneous speech-to-text translation without encoder-decoder cross-attention.