Integration of TinyML and LargeML: A Survey of 6G and Beyond
Survey on integrating TinyML and LargeML for 6G networks, covering deep learning applications in mobile systems, autonomous vehicles, and smart services.
Survey on integrating TinyML and LargeML for 6G networks, covering deep learning applications in mobile systems, autonomous vehicles, and smart services.
Attention-aware embedding initialization method for new tokens in LLMs without expensive retraining, addressing vocabulary limitations in specialized domains.
Conditional marked point processes for reliable object detection uncertainty quantification, addressing miscalibrated confidence scores in neural networks.
Self-supervised learning approach adapting joint embedding architecture from video to EEG signals for brain activity analysis with limited labeled data.
Quantum-informed ML framework combining quantum generative models with classical predictors for long-term spatiotemporal chaos prediction.
Supervised fine-tuning method to align LLM agents with rational and moral preferences in strategic economic games, addressing systematic behavioral deviations.
Object-centric representations for visual RL policies using dynamic tokens to improve generalization under visual condition changes without fixed-size slots.
Security evaluation of ML model sharing frameworks and hubs, assessing vulnerabilities in loading shared models and security awareness gaps among practitioners.
Neural quantum states impurity solver for quantum embedding problems. Graph transformer-based NQS for solving Hamiltonians in quantum chemistry.
Dynamic Aware: out-of-distribution detection for trajectory prediction in autonomous vehicles. Adaptive multi-mode approach for distribution shift in AVs.
AutoClimDS: agentic AI system for climate data science. Knowledge graph-based workflows for discovering climate patterns from fragmented data sources.
Formal language theory applied to statistical learning. Proves subregular language classes are linearly separable with simple models.
DataMind: scalable data-analytic AI agents for automated discovery. Open-source agent framework handling diverse-format data files and multi-step reasoning.
HoneyBee: data curation approaches for vision-language reasoning datasets. Analyzes impact of context, content, and format on VLM reasoning capabilities.
CBF-RL: integrates control barrier functions into reinforcement learning training. Enforces dynamic safety constraints during RL policy training, not just inference.
RobotArena ∞: scalable robot benchmarking via real-to-sim translation. Enables rigorous evaluation of robot policies across diverse tasks and environments.
Verifying LLM inference to detect model weight exfiltration via steganography. Defends inference servers against model theft and anomalous behavior.
AnatomiX: anatomy-aware multimodal LLM for chest X-ray interpretation. Improves spatial reasoning and anatomical understanding in medical imaging.
FARM framework for malware family classification under concept drift. Uses triplet autoencoder for few-shot adaptation to covariate and label drift.
LatentChem: latent reasoning interface for chemical LLMs. Decouples chemical computation from discrete tokens to improve efficiency and performance in chemical reasoning.
Pyramid MoA: probabilistic framework for cost-optimized LLM inference via cascading and routing. Balances inference cost and reasoning capability for large language models.
IROSA: framework combining foundation models with imitation learning for robot skill adaptation via natural language. LLM application to robotics.
Disentangled Safety Hypothesis: mechanistic study of LLM safety showing decoupling between harmfulness detection and refusal. ML interpretability research.
Benchmark evaluating frontier AI models on multi-step cyber attack scenarios. Agent capability measurement across extended action sequences.
Agentic framework for multimodal query processing with adaptive tool orchestration across text/image/audio/video. Research on agent coordination and tool selection.
Proof-Carrying Materials: falsifiable safety certificates for machine-learned interatomic potentials. ML research on reliability guarantees for scientific models.
HyperAI provides browser-based notebooks for running LLM-Course tutorials covering LLM fundamentals and production applications.
DeepSteve: hackable multi-terminal environment for AI coding agents with customizable interfaces.
On-Call Health: open-source tool detecting burnout patterns in incident responders using data from PagerDuty, GitHub, and other platforms.
Rust project documentation on community perspectives regarding AI use policy and governance in open-source development.
Hacker News discussion questioning company layoff rationale when AI increases developer productivity.
Desktop app converting podcasts into searchable knowledge base with transcription, translation, and archival capabilities.
Training-free infinite video generation using evolving memory tokens via MemRoPE without KV cache limitations.
Framework for effective multi-agent teams treating them as organizational units rather than parallel single agents for improved collaboration outcomes.
IdeaCred: tool scoring GitHub repositories using LLM analysis and GitHub API metrics across innovation, craft, traction, and scope.
Testing framework for LLM applications using call interception and behavior comparison. CI gates for prompt changes with baseline drift detection.
Jsse: JavaScript engine written in Rust entirely by AI agents, achieving 99.96% test262 compliance without manual code contributions.
Enterprise AI coding agents create unauthorized shadow IT systems, raising organizational governance concerns as autonomous tools bypass approval processes.
Gstack-auto tool for automated, parallelized builds using AI coding agents iteratively to refine software through testing and bugfixes rather than single prompts.
SDK for Next.js/React blog automation with AI-assisted integration. Generates blog metadata, schema, sitemaps, and scheduling via agent workflow.
Terminal emulator for iOS enabling remote agent supervision. Addresses Claude Code workflow by fixing mobile terminal UX limitations.
TypeScript port of Pydantic AI framework. Agent SDK with 50+ LLM provider support via Vercel AI SDK. Open-source with full documentation.
Personal essay on using LLMs for software development, emphasizing how LLMs enable creative making and shift focus from programming mechanics to creative output.
Voxlert is an open-source tool using LLM and TTS to give distinct character voices to concurrent Claude/Codex agent sessions for better notification tracking.
HN discussion about team collaboration practices when using AI coding agents like Claude Code and Codex, integration with project management tools.
Essay on transductive inference theory applied to AI systems, exploring how next-word prediction relates to broader AI reasoning tasks.
Overview of agentic engineering practices for developing software with AI coding agents like Claude Code and Gemini CLI, including definitions and patterns.
Self-hosted PaaS platform using GitHub Actions as control plane for deploying multiple applications on single Debian server.
Reference to open-source agentic AI physicist project (video format, minimal content provided).
Codex Security: AI agent for code security that analyzes repository architecture and trust boundaries before validating findings with humans.