Memory Inception: Latent-Space KV Cache Manipulation for Steering LLMs
Training-free method for steering LLMs via latent attention space manipulation. Enables structured memory insertion without prompting overhead.
Training-free method for steering LLMs via latent attention space manipulation. Enables structured memory insertion without prompting overhead.
Soft deterministic policy gradient method using Gaussian smoothing to handle sparse/discrete rewards in continuous control.
Defense against evasion attacks in multimodal recommender systems using adversarial training with cross-modal coordination.
Diagnostic framework to distinguish memorization from structural learning in graph language models using subgraph mining.
Analysis of free-riding behavior in Forward-Forward neural network training. Identifies layer-wise gradient decay and proposes solutions.
Research on Lagrangian Gaussian Processes for learning dynamical systems while preserving geometric structure. Original arXiv paper.
Study analyzing how node features and graph topology interact in graph pooling for classification tasks, examining conditions for effective pooling operators.
ArXiv paper proposing feature-centric framework showing weight Gram matrix captures sequential feature linearization in deep network training.
ArXiv paper proposing dual manifold calibration for graph federated learning to handle semantic and structural heterogeneity across distributed clients.
ArXiv paper on inference-time refinement for tabular diffusion models to close synthetic-real gap without retraining the generator backbone.
ArXiv paper presenting Function Projection for Flow Matching algorithm that conditions generative models on target distribution samples for adaptation.
ArXiv paper proposing Hierarchy-Aware Cross-Entropy loss that incorporates class hierarchy to penalize semantically distant misclassifications differently.
ArXiv paper presenting INEUS, a meshfree iterative neural solver for partial integro-differential equations using single-jump sampling and recursive regression.
ArXiv paper introducing meta-attributions framework to measure second-order interaction effects of model explanations using Shapley value game theory.
ArXiv paper analyzing region seeding in piecewise affine neural networks through pre-activation regularization to improve expressive capacity.
ArXiv paper on confound-aware representation learning in Transformer-VAE for molecular generation with chemical property steering in latent spaces.
ArXiv paper proposing Dynamic Pattern Recalibration for time series forecasting that adapts to shifting local temporal patterns instead of using fixed weights.
ArXiv paper analyzing benign overfitting in ℓ2-boosting under ℓ1 implicit bias of greedy ensembles with high-dimensional risk asymptotics.
ArXiv paper introducing Pro-KLShampoo optimizer combining Kronecker-factored preconditioning with orthogonalization for LLM pre-training efficiency.
53K-parameter transformer learning valid SMILES generation with 95% validity, outperforming larger models.
Neural routing decoder architecture computing explicit one-step consequences for constructive solvers.
Method to extract clinical variable associations from LLMs using structured comparison questions.
Decision-theoretic framework optimizing cost-quality tradeoffs in LLM cascades for model deferral.
Topological analysis of grokking phenomenon using persistent homology on embedding matrices.
Order-agnostic autoregressive models for generative modeling with incomplete and missing data handling.
Memory-efficient gradient attack framework for evaluating diffusion and Langevin-based adversarial defenses.
Analysis of Chronos foundation model's ability to process and represent frequency domain time-series data.
Flow matching generative framework supporting arbitrary auxiliary distributions for flexible trajectories.
Analysis of layer collapse phenomenon in diffusion language models affecting activation dynamics.
Pair-GRPO framework for stable LLM alignment via reinforcement learning from human preferences.
Data-driven local covariate selection method for causal effect estimation with latent confounding.
On-policy distillation method for LLM token-level training addressing variance and exploration issues.
Research on unified convolutional learning framework for infinite-dimensional signals on manifolds.
Post-training semi-structured sparsification framework for LLMs using Hessian-guided soft masking and annealing for efficiency.
Dimensionless control parameter predicting mixture-of-experts model stability and expert ecology collapse across vision and language tasks.
Interpretable concept bottleneck models using hyperbolic geometry to capture semantic hierarchies between concepts.
Federated learning optimization for Transformer models with attention kernel freezing to reduce client drift.
Continual learning approach for CSI-based activity recognition using mixture of experts to handle domain shifts.
Bayesian hyperparameter optimization addressing acquisition estimation noise and unstable candidate ranking decisions.
Geometric framework characterizing invariant semantic features in language model latent space under paraphrasing.
Multimodal retrieval method mining internal representations for efficient visual document retrieval with reduced index footprint.
Graph invariant diagnostic framework for benchmarking graph foundation models to separate structure and feature contributions.
Graph representation learning method using diversity curves to compare graphs of different sizes with structural awareness.
Benchmark evaluation framework for topological deep learning models on manifold data with triangulation operations.
arXiv paper on efficient LLM serving for dynamic agent workflows using prediction-based KV-Cache management.
arXiv paper presenting Q-MMR framework for off-policy evaluation in MDPs via recursive reweighting and moment matching.
arXiv paper on operator-guided invariance learning for continuous reinforcement learning under distribution shift.
arXiv paper interpreting transformer attention as Nadaraya-Watson regression, proposing Cubit token mixer alternative.
arXiv paper on PACZero, a PAC-private zeroth-order method for fine-tuning LLMs via sign quantization.
arXiv paper analyzing layerwise inference dynamics in transformer-based tabular foundation models.