Show HN: Durable, asynchronous LLM workflow engine in Rust
Open-source Rust-based LLM workflow engine supporting transparent switching between live and batch execution.
Open-source Rust-based LLM workflow engine supporting transparent switching between live and batch execution.
Payment-enabled web search API for autonomous agents using x402 protocol and USDC micropayments without API keys.
Comparative analysis of LLM efficiency measured as queries per unit of compute/energy cost.
Technical guide to measuring Time To First Token (TTFT) in LLM systems with Python, Node.js, and JMeter instrumentation.
Open-source CLI for AI workload orchestration across 23 cloud providers with intelligent routing.
Zephyr-based embedded OS running sandboxed WebAssembly applications on microcontrollers.
IDE with AI coding assistance focused on code quality and reducing time spent on code review vs. shipping.
Core AI framework for exporting and building on-device ML models for macOS and iOS with Python and Swift utilities.
Persistent context manager for Claude, Windsurf, and Cursor code editors.
GPT-Pilot repository compromised with credential-stealing malware injected into telemetry module.
Doubleword improves MoE model inference throughput by reordering batch inputs to optimize expert weight loading and memory bandwidth.
Analysis of Redmonk programming language rankings observing minimal top-20 movement with Java, Python, JavaScript dominant.
Apple offers free API access to foundation models via Private Cloud Compute for developers with <2M App Store downloads.
Apple expands Private Cloud Compute for AI inference beyond own data centers with Google and NVIDIA.
Apple releases third generation foundation models family with on-device and server-based variants built with Google.
Golang module for technical analysis indicators with backtesting framework and AI integration via MCP.
Engineering analysis of coordination bottlenecks when AI agents handle software development handoffs between discovery and implementation phases.
Blog post discussing practical experience using LLM agents for new software projects and code restructuring with watgo example.
RCT study measuring math learning improvements for 1,763 students in Sierra Leone using Gemini's Guided Learning feature.
Maturity model for software engineering in AI era measuring cognitive capacity and organizational readiness for AI-mediated work.
Noren voice profile engine preserves individual writing style in AI-generated content through linguistic analysis.
Pokayoke provides deterministic guardrails for AI coding agents to enforce repository conventions and code style rules.
Analysis of DeepSeek's cost reduction in frontier AI models and the challenges maintaining economic viability at scale.
Y Combinator tool analyzing coding sessions to understand individual AI-assisted development patterns and compare across developers.
Self-evolving extraction system for structured information from text that adapts to domain-specific taxonomies and emerging jargon in NLP.
Benchmarks autoregressive language models for lossless audio compression across bit depths and sample rates against existing codecs.
Proposes Partition Determinantal Point Process for improving diversity in parallel discrete diffusion decoding for text generation.
Investigates mechanisms of spatial variable binding in vision-language models for multimodal tasks like image captioning and VQA.
Multi-agent framework combining specialist agents with verification and fusion techniques to improve calibration in medical question answering.
Template-driven ML development framework for managing large ecosystems of recommendation and prediction models in advertising platforms.
Hybrid ensembling approach combining Chain-of-Thought and Program-of-Thought reasoning to improve LLM reasoning with minimal sampling overhead.
Addresses temporal bias in Android malware detection ML models using time-stamped datasets and timestamp-verification procedures.
Proposes retrieval-augmented agents that improve knowledge base navigation by modeling expert search strategies rather than treating retrieval as a black box.
Studies how latent geometry in pretrained world models simplifies control and goal-oriented planning in latent spaces.
Tool for systematic diagnosis of failures in LLM agents via corpus-level trace analysis to identify failure patterns at scale.
Novel training paradigm for reinforcement learning in diffusion language models using denoising feedback for policy loss estimation.
Verification framework for deterministic AI inference on GPUs enabling credible monitoring against covert adversaries.
System for training and serving MoE models with dynamic load balancing addressing expert parallelism bottlenecks.
Study of content moderation system instability under code-mixed inputs showing workflow changes from clean English evaluation.
Quantization-aware distillation technique preserving internal geometry for NVFP4 low-precision LLM inference optimization.
Research on symbolic explanations for multiple instance learning models in digital pathology with improved interpretability.
OpenMythos: open-source theoretical implementation of Claude Mythos architecture using Recurrent-Depth Transformer with three-stage pipeline, based on public research speculation.
Engineering leadership talk on organizational transformation when agentic AI coding becomes default workflow, covering process and structure changes from traditional planning methodologies.
Rust-based drop-in LLM routing proxy for OpenAI/Anthropic SDKs enabling cost optimization, provider failover, and model right-sizing without code changes.
Open-source benchmark tracking LLM model tool selection preferences across frozen vibe-coding prompts from beginner to expert levels, evaluating devtool choices.
AI agent-led search engine skill ranked by upvotes and real money signals, compatible with Claude Code, Cursor, Copilot and 50+ agent skill hosts via npm.
EMILIAProtocol: vendor-neutral standard for human approval workflows on irreversible AI agent actions with trust enforcement.
Rust grid-fill engine and Claude-based clue generator pipeline for creating NYT-style crossword puzzles programmatically.
Opinion on AI coding tools as labor supplements vs replacements and cost economics of compute vs engineering salaries.
AI Code Stitcher tool integrating LLM code outputs (Claude, ChatGPT, Aider) into existing codebases with smart matching and preview.