Ethical Fairness without Demographics in Human-Centered AI
Method for achieving fairness in AI systems without demographic attributes for human-centered applications.
Method for achieving fairness in AI systems without demographic attributes for human-centered applications.
Method for fine-tuning diffusion policies with reinforcement learning for humanoid robot loco-manipulation tasks.
Pruning method for efficient large vision-language model inference by exploiting attention patterns and addressing token redundancy.
Framework for aggregating noisy heterogeneous evidence in probabilistic reasoning tasks with explicit uncertainty quantification.
Theoretical framework for aggregating multiple evidence sources in probabilistic prediction with formal guarantees for multi-evidence reasoning.
Dashboard for orchestrating multiple AI agents (Claude, Codex, Aider, Goose) with real-time monitoring and isolation.
Live AI judge for hackathons using multi-model ensemble from Gemini, Claude, and Groq with real-time video/audio analysis.
PDF on identity-based security containment for autonomous agent workloads using SPIFFE and Istio.
Experiment feeding fake academic paper to NotebookLM's audio feature to test AI podcast hosts' credibility.
Dataset from Anthropic interviews on how AI is being used in people's daily lives.
Evaluation framework benchmarking AI coding agents on Laravel problems using test verification and cost tracking.
Kubernetes multi-tenancy challenges with AI agents requiring ephemeral environments. Infrastructure scaling issues.
Privacy-focused AI system for Meta addressing data flows through LLMs and sensitive information exposure.
Local AI coding assistant for terminal using Ollama, no subscription, runs offline, built with Go.
Essay on Ruby on Rails' resurgence as practical choice in AI coding era due to productivity and maintainability.
CLI tool for orchestrating Claude Code workflows with operators for iteration, review, parallel execution, and composition.
Analysis arguing software won't become disposable despite AI coding agents, critiquing concepts like 'vibe coding' and ephemeral apps.
Model Context Protocol (MCP) server for accessing fitness data from Garmin and Strava integrations with LLM applications.
Cloudflare Workers project for running networks of LLM-backed bots using Durable Objects, with message history, coordination, and optional web fetching capabilities.
Framework for running small networks of LLM-backed bots on Cloudflare Workers with message history and coordination.
Bare metal GPU cluster management platform with Kubernetes-native workflows, simplifying infrastructure without OpenStack complexity.
Discussion thread on managing relationships with people who blindly trust LLM outputs as objective truth.
Essay examining copyleft licensing challenges in the AI era, arguing GPL concepts must evolve to protect users from vendor lock-in.
South Korea's SDT opens first commercial quantum-AI hybrid data center in Seoul with 20-qubit Kreo quantum computer and Nvidia DGX B200 integration.
Scheduled: Open-source AI agent integrated with Gmail that autonomously reads meeting request emails, checks calendar availability, and drafts proposed times.
ATO: GUI control panel managing multiple LLM agents (Claude Code, Codex, OpenClaw, Hermes) with workflow orchestration and MCP integration.
LA County courts pilot AI tool (Learned Hand) to summarize legal motions and draft rulings based on judge writing styles.
Conceptual framework on how AI agents transform organizational structure and decision-making beyond efficiency gains.
Case study: developer maintaining open-source Chrome extension with AI assistance for bug fixes and feature development.
OpenAI acquires Astral, integrating open source Python developer tools (uv, Ruff) into Codex ecosystem to enhance Python development tooling.
Chainguard Agent Skills: security solution protecting against malicious AI agent skills with verification and sandboxing for YAML-based agent plugins.
Trepan: local-first architectural linter enforcing code intent and preventing architecture drift without cloud data transmission.
PondDB: open-source DuckDB-based memory database for multi-agent systems enabling SQL querying of agent state and decision history.
Genetic algorithm that uses 100 LLM personas to red-team and improve landing page copy generation, addressing generic AI writing outputs.
Blobsearch: DuckDB-based log storage and querying alternative using S3 and Parquet for cost-effective log management.
Bug report on Claude Code's poor time-awareness limiting task optimization and efficiency in code completion.
MCP tool providing simplified Jira integration for AI agents via 3 composable tools instead of 72 API endpoints.
Skillfile: declarative manifest system for managing AI agent skills across Claude, Cursor, Gemini and other platforms with versioning and deployment.
LittleHorse 1.0: microservice orchestration engine enabling Business-as-Code approach for distributed process definition.
GitHub Action providing AI code review via Pervaziv.
TurboAPI: FastAPI-compatible Python framework with Zig HTTP core, 7x faster with zero-copy responses.
Open-source personal autonomous AI agent built on Elixir/OTP that monitors feeds, executes workflows, routes tasks to cheapest suitable LLM. Single-user, auditable codebase.
Personal project using Claude to build photo sharing app replacing iCloud. Practical LLM application with implementation discussion.
Headline about poker experiments with frontier LLMs. Appears duplicate of article [3] with less content.
Research using Claude Sonnet and Gemini Flash agents to play poker, revealing reasoning capabilities and strategic decision-making in frontier LLMs through game theory.
Experiment replicating RYS method on consumer AMD GPUs, discovering discrete reasoning circuits in 24B LLM by duplicating layers improves logical deduction from 0.22 to 0.76.
GPU runtime for Nvidia GPUs enabling safe VRAM overcommit, fractional core allocation, and weight deduplication.
VibePod adds Ollama/vLLM backend support for Claude Code and Codex.
Enterprise AI adoption gap: models and agents scale but organizational context understanding lags. Governance and activation challenges remain.
Local TTS model with 31M params, voice cloning, voice blending. 5.6x realtime on CPU, ONNX export, Apache 2.0 license.