300 canonical paper links on this archive page.
- InsightTok: Improving Text and Face Fidelity in Discrete Tokenization for Autoregressive Image Generationarxiv-2605.14333 Sparse Blocked context onlyMay 14, 2026
- Delta Attention Residualsarxiv-2605.18855 Sparse Blocked context onlyMay 13, 2026
- KIT-TIP-NLP at MultiPride: Continual Learning with Multilingual Foundation Modelarxiv-2605.13415 Sparse Blocked context onlyMay 13, 2026
- Probing Persona-Dependent Preferences in Language Modelsarxiv-2605.13339 Sparse Blocked context onlyMay 13, 2026
- IndicMedDialog: A Parallel Multi-Turn Medical Dialogue Dataset for Accessible Healthcare in Indic Languagesarxiv-2605.13292 Sparse Blocked context onlyMay 13, 2026
- F-GRPO: Factorized Group-Relative Policy Optimization for Unified Candidate Generation and Rankingarxiv-2605.12995 Sparse Blocked context onlyMay 13, 2026
- Asymmetric Flow Modelsarxiv-2605.12964 Direct Blocked context onlyMay 13, 2026
- EgoForce: Forearm-Guided Camera-Space 3D Hand Pose from a Monocular Egocentric Cameraarxiv-2605.12498 Sparse Blocked context onlyMay 12, 2026
- Pion: A Spectrum-Preserving Optimizer via Orthogonal Equivalence Transformationarxiv-2605.12492 Sparse Blocked context onlyMay 12, 2026
- A Causal Language Modeling Detour Improves Encoder Continued Pretrainingarxiv-2605.12438 Sparse Blocked context onlyMay 12, 2026
- Geometric Factual Recall in Transformersarxiv-2605.12426 Sparse Blocked context onlyMay 12, 2026
- Hölder Policy Optimisationarxiv-2605.12058 Sparse Blocked context onlyMay 12, 2026
- OmniHumanoid: Streaming Cross-Embodiment Video Generation with Paired-Free Adaptationarxiv-2605.12038 Sparse Blocked context onlyMay 12, 2026
- L2P: Unlocking Latent Potential for Pixel Generationarxiv-2605.12013 Sparse Blocked context onlyMay 12, 2026
- ELF: Embedded Language Flowsarxiv-2605.10938 Sparse Blocked context onlyMay 11, 2026
- Pixal3D: Pixel-Aligned 3D Generation from Imagesarxiv-2605.10922 Direct Blocked context onlyMay 11, 2026
- Prompt-Activation Duality: Improving Activation Steering via Attention-Level Interventionsarxiv-2605.10664 Sparse Blocked context onlyMay 11, 2026
- Mela: Test-Time Memory Consolidation based on Transformation Hypothesisarxiv-2605.10537 Sparse Blocked context onlyMay 11, 2026
- GLiNER-Relex: A Unified Framework for Joint Named Entity Recognition and Relation Extractionarxiv-2605.10108 Curated Related Blocked context onlyMay 11, 2026
- TD3B: Transition-Directed Discrete Diffusion for Allosteric Binder Generationarxiv-2605.09810 Direct Blocked context onlyMay 10, 2026
- Parameter-Efficient Neuroevolution for Diverse LLM Generation: Quality-Diversity Optimization via Prompt Embedding Evolutionarxiv-2605.09781 Curated Related Blocked context onlyMay 10, 2026
- Forcing-KV: Hybrid KV Cache Compression for Efficient Autoregressive Video Diffusion Modelsarxiv-2605.09681 Sparse Blocked context onlyMay 10, 2026
- Geometry Conflict: Explaining and Controlling Forgetting in LLM Continual Post-Trainingarxiv-2605.09608 Sparse Blocked context onlyMay 10, 2026
- Shaping Schema via Language Representation as the Next Frontier for LLM Intelligence Expandingarxiv-2605.09271 Sparse Blocked context onlyMay 10, 2026
- Lost in Translation? Exploring the Shift in Grammatical Gender from Latin to Occitanarxiv-2605.09156 Sparse Blocked context onlyMay 9, 2026
- LLiMba: Sardinian on a Single GPU -- Adapting a 3B Language Model to a Vanishing Romance Languagearxiv-2605.09015 Curated Related Blocked context onlyMay 9, 2026
- AdaPreLoRA: Adafactor Preconditioned Low-Rank Adaptationarxiv-2605.08734 Sparse Blocked context onlyMay 9, 2026
- DiagnosticIQ: A Benchmark for LLM-Based Industrial Maintenance Action Recommendation from Symbolic Rulesarxiv-2605.08614 Sparse Blocked context onlyMay 9, 2026
- PASA: A Principled Embedding-Space Watermarking Approach for LLM-Generated Text under Semantic-Invariant Attacksarxiv-2605.10977 Sparse Blocked context onlyMay 9, 2026
- Can Language Models Identify Side Effects of Breast Cancer Radiation Treatments?arxiv-2605.08439 Sparse Blocked context onlyMay 8, 2026
- Queryable LoRA: Instruction-Regularized Routing Over Shared Low-Rank Update Atomsarxiv-2605.08423 Sparse Blocked context onlyMay 8, 2026
- GLiGuard: Schema-Conditioned Classification for LLM Safeguardarxiv-2605.07982 Sparse Blocked context onlyMay 8, 2026
- How to Train Your Latent Diffusion Language Model Jointly With the Latent Spacearxiv-2605.07933 Sparse Blocked context onlyMay 8, 2026
- How Value Induction Reshapes LLM Behaviourarxiv-2605.07925 Sparse Blocked context onlyMay 8, 2026
- What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusionarxiv-2605.07915 Sparse Blocked context onlyMay 8, 2026
- A Comparative Analysis of Classical Machine Learning and Deep Learning Approaches for Sentiment Classification on IMDb Movie Reviewsarxiv-2605.07811 Sparse Blocked context onlyMay 8, 2026
- TRACE: Tourism Recommendation with Accountable Citation Evidencearxiv-2605.07677 Sparse Blocked context onlyMay 8, 2026
- Multi-Dimensional Evaluation of LLMs for Grammatical Error Correctionarxiv-2605.07635 Sparse Blocked context onlyMay 8, 2026
- TCMIIES: A Browser-Based LLM-Powered Intelligent Information Extraction System for Academic Literaturearxiv-2605.07507 Sparse Blocked context onlyMay 8, 2026
- SEIF: Self-Evolving Reinforcement Learning for Instruction Followingarxiv-2605.07465 Sparse Blocked context onlyMay 8, 2026
- GRaSp: Automatic Example Optimization for In-Context Learning in Low-Data Tasksarxiv-2605.07454 Sparse Blocked context onlyMay 8, 2026
- Data Contamination in Neural Hieroglyphic Translation: A Reproducibility Studyarxiv-2605.07453 Sparse Blocked context onlyMay 8, 2026
- MIPIAD: Multilingual Indirect Prompt Injection Attack Defense with Qwen -- TF-IDF Hybrid and Meta-Ensemble Learningarxiv-2605.07269 Sparse Blocked context onlyMay 8, 2026
- A Reproducible Multi-Architecture Baseline for Token-Level Chinese Metaphor Identification under the MIPVU Frameworkarxiv-2605.07170 Sparse Blocked context onlyMay 8, 2026
- Beyond LoRA vs. Full Fine-Tuning: Gradient-Guided Optimizer Routing for LLM Adaptationarxiv-2605.07111 Sparse Blocked context onlyMay 8, 2026
- IntentGrasp: A Comprehensive Benchmark for Intent Understandingarxiv-2605.06832 Sparse Blocked context onlyMay 7, 2026
- When No Benchmark Exists: Validating Comparative LLM Safety Scoring Without Ground-Truth Labelsarxiv-2605.06652 Sparse Blocked context onlyMay 7, 2026
- Efficient Pre-Training with Token Superpositionarxiv-2605.06546 Sparse Blocked context onlyMay 7, 2026
- Sparkle: Realizing Lively Instruction-Guided Video Background Replacement via Decoupled Guidancearxiv-2605.06535 Sparse Blocked context onlyMay 7, 2026
- Token Time Continuous Diffusion for Language Modelingarxiv-2607.14106 Sparse Blocked context onlyMay 7, 2026
- MARBLE: Multi-Aspect Reward Balance for Diffusion RLarxiv-2605.06507 Sparse Blocked context onlyMay 7, 2026
- Litespark Inference For CPUs: Ultra-Fast SIMD Framework for Ternary (1.58-bit) Language Modelsarxiv-2605.06485 Sparse Blocked context onlyMay 7, 2026
- Probabilistic Dating of Historical Manuscripts via Evidential Deep Regression on Visual Script Featuresarxiv-2605.06475 Sparse Blocked context onlyMay 7, 2026
- Empirical Evidence for Simply Connected Decision Regions in Image Classifiersarxiv-2605.06380 Sparse Blocked context onlyMay 7, 2026
- SEQUOR: A Multi-Turn Benchmark for Realistic Constraint Followingarxiv-2605.06353 Sparse Blocked context onlyMay 7, 2026
- Who and What? Using Linguistic Features and Annotator Characteristics to Analyze Annotation Variationarxiv-2605.06318 Sparse Blocked context onlyMay 7, 2026
- MultiLinguahah : A New Unsupervised Multilingual Acoustic Laughter Segmentation Methodarxiv-2605.06309 Sparse Blocked context onlyMay 7, 2026
- Log-Likelihood, Simpson's Paradox, and the Detection of Machine-Generated Textarxiv-2605.06294 Sparse Blocked context onlyMay 7, 2026
- Quantifying the Statistical Effect of Rubric Modifications on Human-Autorater Agreementarxiv-2605.06283 Sparse Blocked context onlyMay 7, 2026
- YEZE at SemEval-2026 Task 9: Detecting Multilingual, Multicultural and Multievent Online Polarization via Heterogeneous Ensemblingarxiv-2605.06231 Sparse Blocked context onlyMay 7, 2026
- When to Trust Imagination: Adaptive Action Execution for World Action Modelsarxiv-2605.06222 Sparse Blocked context onlyMay 7, 2026
- UniPrefill: Universal Long-Context Prefill Acceleration via Block-wise Dynamic Sparsificationarxiv-2605.06221 Sparse Blocked context onlyMay 7, 2026
- TIDE: Every Layer Knows the Token Beneath the Contextarxiv-2605.06216 Sparse Blocked context onlyMay 7, 2026
- Mean Mode Screaming: Mean--Variance Split Residuals for 1000-Layer Diffusion Transformersarxiv-2605.06169 Sparse Blocked context onlyMay 7, 2026
- Navigating by Old Maps: The Pitfalls of Static Mechanistic Localization in LLM Post-Trainingarxiv-2605.06076 Sparse Blocked context onlyMay 7, 2026
- More Aligned, Less Diverse? Analyzing the Grammar and Lexicon of Two Generations of LLMsarxiv-2605.06030 Sparse Blocked context onlyMay 7, 2026
- MDN: Parallelizing Stepwise Momentum for Delta Linear Attentionarxiv-2605.05838 Sparse Blocked context onlyMay 7, 2026
- A Few Good Clauses: Comparing LLMs vs Domain-Trained Small Language Models on Structured Contract Extractionarxiv-2605.05532 Sparse Blocked context onlyMay 7, 2026
- MRI-Eval: A Tiered Benchmark for Evaluating LLM Performance on MRI Physics and GE Scanner Operations Knowledgearxiv-2605.05175 Sparse Blocked context onlyMay 6, 2026
- PSK at SemEval-2026 Task 9: Multilingual Polarization Detection Using Ensemble Gemma Models with Synthetic Data Augmentationarxiv-2605.05159 Sparse Blocked context onlyMay 6, 2026
- Text Corpora as Concept Fields: Black-Box Hallucination and Novelty Measurementarxiv-2605.05103 Sparse Blocked context onlyMay 6, 2026
- The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differencesarxiv-2605.05080 Sparse Blocked context onlyMay 6, 2026
- Detecting Hallucinations in Large Language Models via Internal Attention Divergence Signalsarxiv-2605.05025 Sparse Blocked context onlyMay 6, 2026
- Conceptors for Semantic Steeringarxiv-2605.04980 Sparse Blocked context onlyMay 6, 2026
- TabEmbed: Benchmarking and Learning Generalist Embeddings for Tabular Understandingarxiv-2605.04962 Direct Blocked context onlyMay 6, 2026
- KernelBench-X: A Comprehensive Benchmark for Evaluating LLM-Generated GPU Kernelsarxiv-2605.04956 Sparse Blocked context onlyMay 6, 2026
- Adapting Large Language Models to a Low-Resource Agglutinative Language: A Comparative Study of LoRA and QLoRA for Bashkirarxiv-2605.04948 Sparse Blocked context onlyMay 6, 2026
- Self-Attention as Transport: Limits of Symmetric Spectral Diagnosticsarxiv-2605.04893 Sparse Blocked context onlyMay 6, 2026
- A Comparative Analysis of Machine Learning and Deep Learning Models for Tweet Sentiment Classification: A Case Study on the Sentiment140 Datasetarxiv-2605.04888 Curated Related Blocked context onlyMay 6, 2026
- Sentiment Analysis and Customer Satisfaction Prediction on E-Commerce Platforms Based on YouTube Comments Using the XGBoost Algorithmarxiv-2605.04887 Sparse Blocked context onlyMay 6, 2026
- Anticipating Innovation Using Large Language Modelsarxiv-2605.04875 Sparse Blocked context onlyMay 6, 2026
- StoryAlign: Evaluating and Training Reward Models for Story Generationarxiv-2605.04831 Sparse Blocked context onlyMay 6, 2026
- Paraphrase-Induced Output-Mode Collapse: When LLMs Break Character Under Semantically Equivalent Inputsarxiv-2605.04665 Sparse Blocked context onlyMay 6, 2026
- Gradients with Respect to Semantics Preserving Embeddings Tell the Uncertainty of Large Language Modelsarxiv-2605.04638 Sparse Blocked context onlyMay 6, 2026
- Benchmarking POS Tagging for the Tajik Language: A Comparative Study of Neural Architectures on the TajPersParallel Corpusarxiv-2605.04576 Sparse Blocked context onlyMay 6, 2026
- A Hybrid Method for Low-Resource Named Entity Recognitionarxiv-2605.04489 Sparse Blocked context onlyMay 6, 2026
- Stabilizing LLM Supervised Fine-Tuning via Explicit Distributional Controlarxiv-2605.04468 Curated Related Blocked context onlyMay 6, 2026
- StableI2I: Spotting Unintended Changes in Image-to-Image Transitionarxiv-2605.04453 Sparse Blocked context onlyMay 6, 2026
- Coral: Cost-Efficient Multi-LLM Serving over Heterogeneous Cloud GPUsarxiv-2605.04357 Sparse Blocked context onlyMay 5, 2026
- SWAN: Semantic Watermarking with Abstract Meaning Representationarxiv-2605.04305 Sparse Blocked context onlyMay 5, 2026
- Towards Self-Referential Analytic Assessment: A Profile-Based Approach to L2 Writing Evaluation with LLMsarxiv-2605.04298 Sparse Blocked context onlyMay 5, 2026
- Nsanku: Evaluating Zero-Shot Translation Performance of LLMs for Ghanaian Languagesarxiv-2605.04208 Sparse Blocked context onlyMay 5, 2026
- Not All That Is Fluent Is Factual: Investigating Hallucinations of Large Language Models in Academic Writingarxiv-2605.04171 Sparse Blocked context onlyMay 5, 2026
- EQUITRIAGE: A Fairness Audit of Gender Bias in LLM-Based Emergency Department Triagearxiv-2605.03998 Sparse Blocked context onlyMay 5, 2026
- Feature-Augmented Transformers for Robust AI-Text Detection Across Domains and Generatorsarxiv-2605.03969 Sparse Blocked context onlyMay 5, 2026
- Steer Like the LLM: Activation Steering that Mimics Promptingarxiv-2605.03907 Sparse Blocked context onlyMay 5, 2026
- Position: the Stochastic Parrot in the Coal Mine. Model Collapse is a Threat to Low-Resource Communitiesarxiv-2605.04127 Sparse Blocked context onlyMay 5, 2026
- TriBench-Ko: Evaluating LLM Risks in Judicial Workflowsarxiv-2605.03792 Sparse Blocked context onlyMay 5, 2026
- Benchmarking Parameter-Efficient Fine-Tuning of Large Language Models for Low-Resource Tajik Text Generation with the Tajik Web Corpusarxiv-2605.03742 Sparse Blocked context onlyMay 5, 2026
- Segmenting Human-LLM Co-authored Text via Change Point Detectionarxiv-2605.03723 Sparse Blocked context onlyMay 5, 2026
- Annotation Quality in Aspect-Based Sentiment Analysis: A Case Study Comparing Experts, Students, Crowdworkers, and Large Language Modelarxiv-2605.03624 Sparse Blocked context onlyMay 5, 2026
- AfriVox-v2: A Domain-Verticalized Benchmark for In-the-Wild African Speech Recognitionarxiv-2605.03590 Sparse Blocked context onlyMay 5, 2026
- Detecting Stealth Sycophancy in Mental-Health Dialogue with Dynamic Emotional Signature Graphsarxiv-2605.03472 Sparse Blocked context onlyMay 5, 2026
- Retrieving Floods without Floodlights: Topic Models as Binary Classifiers for Extreme Climate Events in German Newsarxiv-2605.03450 Sparse Blocked context onlyMay 5, 2026
- Benchmarking Logistic Regression, SVM, Naive Bayes, and IndoBERT Fine-Tuning for Sentiment Analysis on Indonesian Product Reviewsarxiv-2605.03439 Sparse Blocked context onlyMay 5, 2026
- Discovering Reinforcement Learning Interfaces with Large Language Modelsarxiv-2605.03408 Sparse Blocked context onlyMay 5, 2026
- Leveraging Argument Structure to Predict Content Hatefulnessarxiv-2605.02457 Sparse Blocked context onlyMay 4, 2026
- InfoLaw: Information Scaling Laws for Large Language Models with Quality-Weighted Mixture Data and Repetitionarxiv-2605.02364 Sparse Blocked context onlyMay 4, 2026
- A Hybrid Approach for Closing the Sim2real Appearance Gap in Game Engine Synthetic Datasetsarxiv-2605.02291 Sparse Blocked context onlyMay 4, 2026
- Zero-Shot Confidence Estimation for Small LLMs: When Supervised Baselines Aren't Worth Trainingarxiv-2605.02241 Sparse Blocked context onlyMay 4, 2026
- Recovering Hidden Reward in Diffusion-Based Policiesarxiv-2605.00623 Sparse Blocked context onlyMay 1, 2026
- Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balancearxiv-2605.00553 Sparse Blocked context onlyMay 1, 2026
- When Do Diffusion Models learn to Generate Multiple Objects?arxiv-2605.00273 Sparse Blocked context onlyApr 30, 2026
- RouteProfile: Elucidating the Design Space of LLM Profiles for Routingarxiv-2605.00180 Sparse Blocked context onlyApr 30, 2026
- Repetition over Diversity: High-Signal Data Filtering for Sample-Efficient German Language Modelingarxiv-2604.28075 Sparse Blocked context onlyApr 30, 2026
- TwinGate: Stateful Defense against Decompositional Jailbreaks in Untraceable Traffic via Asymmetric Contrastive Learningarxiv-2604.27861 Sparse Blocked context onlyApr 30, 2026
- Instruction-Guided Poetry Generation in Arabic and Its Dialectsarxiv-2604.27766 Sparse Blocked context onlyApr 30, 2026
- ExoActor: Exocentric Video Generation as Generalizable Interactive Humanoid Controlarxiv-2604.27711 Sparse Blocked context onlyApr 30, 2026
- Assessing Pancreatic Ductal Adenocarcinoma Vascular Invasion: the PDACVI Benchmarkarxiv-2604.27582 Sparse Blocked context onlyApr 30, 2026
- Exploring Applications of Transfer-State Large Language Models: Cognitive Profiling and Socratic AI Tutoringarxiv-2604.27454 Sparse Blocked context onlyApr 30, 2026
- Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulationarxiv-2604.27263 Sparse Blocked context onlyApr 29, 2026
- SafeReview: Defending LLM-based Review Systems Against Adversarial Hidden Promptsarxiv-2604.26506 Sparse Blocked context onlyApr 29, 2026
- Structural Generalization on SLOG without Hand-Written Rulesarxiv-2604.26157 Sparse Blocked context onlyApr 28, 2026
- Sample Selection Using Multi-Task Autoencoders in Federated Learning with Non-IID Dataarxiv-2604.26116 Sparse Blocked context onlyApr 28, 2026
- From Prompt Risk to Response Risk: Paired Analysis of Safety Behavior of Large Language Modelarxiv-2604.26052 Sparse Blocked context onlyApr 28, 2026
- A Survey on LLM-based Conversational User Simulationarxiv-2604.24977 Curated Related Blocked context onlyApr 27, 2026
- Less Is More: Engineering Challenges of On-Device Small Language Model Integration in a Mobile Applicationarxiv-2604.24636 Sparse Blocked context onlyApr 27, 2026
- Learning to Route Queries to Heads for Attention-based Re-ranking with Large Language Modelsarxiv-2604.24608 Sparse Blocked context onlyApr 27, 2026
- Diffusion Model as a Generalist Segmentation Learnerarxiv-2604.24575 Sparse Blocked context onlyApr 27, 2026
- Graph Memory Transformer (GMT)arxiv-2604.23862 Curated Related Blocked context onlyApr 26, 2026
- Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly Segmentationarxiv-2604.23604 Sparse Blocked context onlyApr 26, 2026
- Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigationarxiv-2604.23600 Sparse Blocked context onlyApr 26, 2026
- Talker-T2AV: Joint Talking Audio-Video Generation with Autoregressive Diffusion Modelingarxiv-2604.23586 Curated Related Blocked context onlyApr 26, 2026
- JudgeSense: A Benchmark for Prompt Sensitivity in LLM-as-a-Judge Systemsarxiv-2604.23478 Sparse Blocked context onlyApr 26, 2026
- V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Thinkarxiv-2604.23380 Sparse Blocked context onlyApr 25, 2026
- From Similarity to Structure: Training-free LLM Context Compression with Hybrid Graph Priorsarxiv-2604.23277 Sparse Blocked context onlyApr 25, 2026
- Neural Recovery of Historical Lexical Structure in Bantu Languages from Modern Dataarxiv-2604.22730 Sparse Blocked context onlyApr 24, 2026
- CRAFT: Clustered Regression for Adaptive Filtering of Training dataarxiv-2604.22693 Sparse Blocked context onlyApr 24, 2026
- Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelinesarxiv-2604.22661 Sparse Blocked context onlyApr 24, 2026
- Dharma, Data and Deception: An LLM-Powered Rhetorical Analysis of Cow-Urine Health Claims on YouTubearxiv-2604.22606 Sparse Blocked context onlyApr 24, 2026
- FeatEHR-LLM: Leveraging Large Language Models for Feature Engineering in Electronic Health Recordsarxiv-2604.22534 Sparse Blocked context onlyApr 24, 2026
- Dynamically Acquiring Text Content to Enable the Classification of Lesser-known Entities for Real-world Tasksarxiv-2604.22325 Sparse Blocked context onlyApr 24, 2026
- Protect the Brain When Treating the Heart: A Convolutional Neural Network for Detecting Emboliarxiv-2604.22258 Sparse Blocked context onlyApr 24, 2026
- Tell Me Why: Designing an Explainable LLM-based Dialogue System for Student Problem Behavior Diagnosisarxiv-2604.22237 Sparse Blocked context onlyApr 24, 2026
- Evaluating LLM-Based Goal Extraction in Requirements Engineering: Prompting Strategies and Their Limitationsarxiv-2604.22207 Sparse Blocked context onlyApr 24, 2026
- How Large Language Models Balance Internal Knowledge with User and Document Assertionsarxiv-2604.22193 Sparse Blocked context onlyApr 24, 2026
- ReCast: Recasting Learning Signals for Reinforcement Learning in Generative Recommendationarxiv-2604.22169 Sparse Blocked context onlyApr 24, 2026
- Where Should LoRA Go? Component-Type Placement in Hybrid Language Modelsarxiv-2604.22127 Sparse Blocked context onlyApr 24, 2026
- PermaFrost-Attack: Stealth Pretraining Seeding(SPS) for planting Logic Landmines During LLM Trainingarxiv-2604.22117 Sparse Blocked context onlyApr 23, 2026
- Spontaneous Persuasion: An Audit of Model Persuasiveness in Everyday Conversationsarxiv-2604.22109 Sparse Blocked context onlyApr 23, 2026
- Ethics Testing: Proactive Identification of Generative AI System Harmsarxiv-2604.22089 Sparse Blocked context onlyApr 23, 2026
- EgoMAGIC- An Egocentric Video Field Medicine Dataset for Training Perception Algorithmsarxiv-2604.22036 Sparse Blocked context onlyApr 23, 2026
- Shared Lexical Task Representations Explain Behavioral Variability In LLMsarxiv-2604.22027 Sparse Blocked context onlyApr 23, 2026
- When Cow Urine Cures Constipation on YouTube: Limits of LLMs in Detecting Culture-specific Health Misinformationarxiv-2604.22002 Sparse Blocked context onlyApr 23, 2026
- Evaluation of Automatic Speech Recognition Using Generative Large Language Modelsarxiv-2604.21928 Curated Related Blocked context onlyApr 23, 2026
- TingIS: Real-time Risk Event Discovery from Noisy Customer Incidents at Enterprise Scalearxiv-2604.21889 Sparse Blocked context onlyApr 23, 2026
- Misinformation Span Detection in Videos via Audio Transcriptsarxiv-2604.21767 Sparse Blocked context onlyApr 23, 2026
- Sapiens2arxiv-2604.21681 Sparse Blocked context onlyApr 23, 2026
- Sub-Token Routing in LoRA for Adaptation and Query-Aware KV Compressionarxiv-2604.21335 Sparse Blocked context onlyApr 23, 2026
- Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AIarxiv-2604.21300 Direct Blocked context onlyApr 23, 2026
- Hyperloop Transformersarxiv-2604.21254 Sparse Blocked context onlyApr 23, 2026
- Adaptive Instruction Composition for Automated LLM Red-Teamingarxiv-2604.21159 Sparse Blocked context onlyApr 22, 2026
- TabSHAParxiv-2604.21120 Sparse Blocked context onlyApr 22, 2026
- DWTSumm: Discrete Wavelet Transform for Document Summarizationarxiv-2604.21070 Sparse Blocked context onlyApr 22, 2026
- Convergent Evolution: How Different Language Models Learn Similar Number Representationsarxiv-2604.20817 Sparse Blocked context onlyApr 22, 2026
- Near-Future Policy Optimizationarxiv-2604.20733 Sparse Blocked context onlyApr 22, 2026
- CityRAG: Stepping Into a City via Spatially-Grounded Video Generationarxiv-2604.19741 Sparse Blocked context onlyApr 21, 2026
- Micro Language Models Enable Instant Responsesarxiv-2604.19642 Sparse Blocked context onlyApr 21, 2026
- LoopCTR: Unlocking the Loop Scaling Power for Click-Through Rate Predictionarxiv-2604.19550 Sparse Blocked context onlyApr 21, 2026
- Evaluation-driven Scaling for Scientific Discoveryarxiv-2604.19341 Sparse Blocked context onlyApr 21, 2026
- ShadowPEFT: Shadow Network for Parameter-Efficient Fine-Tuningarxiv-2604.19254 Sparse Blocked context onlyApr 21, 2026
- Dual-View Training for Instruction-Following Information Retrievalarxiv-2604.18845 Sparse Blocked context onlyApr 20, 2026
- Towards Understanding the Robustness of Sparse Autoencodersarxiv-2604.18756 Sparse Blocked context onlyApr 20, 2026
- LLM Safety From Within: Detecting Harmful Content with Internal Representationsarxiv-2604.18519 Sparse Blocked context onlyApr 20, 2026
- UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Modelsarxiv-2604.18518 Sparse Blocked context onlyApr 20, 2026
- Extending One-Step Image Generation from Class Labels to Text via Discriminative Text Representationarxiv-2604.18168 Sparse Blocked context onlyApr 20, 2026
- NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASRarxiv-2604.18105 Sparse Blocked context onlyApr 20, 2026
- Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarksarxiv-2604.17761 Sparse Blocked context onlyApr 20, 2026
- MoVE: Translating Laughter and Tears via Mixture of Vocalization Experts in Speech-to-Speech Translationarxiv-2604.17435 Sparse Blocked context onlyApr 19, 2026
- Back to Repair: A Minimal Denoising Network\ for Time Series Anomaly Detectionarxiv-2604.17388 Sparse Blocked context onlyApr 19, 2026
- Align Documents to Questions: Question-Oriented Document Rewriting for Retrieval-Augmented Generationarxiv-2604.17325 Sparse Blocked context onlyApr 19, 2026
- Beyond Text-Dominance: Understanding Modality Preference of Omni-modal Large Language Modelsarxiv-2604.16902 Sparse Blocked context onlyApr 18, 2026
- Stochasticity in Tokenisation Improves Robustnessarxiv-2604.16037 Sparse Blocked context onlyApr 17, 2026
- RAGognizer: Hallucination-Aware Fine-Tuning via Detection Head Integrationarxiv-2604.15945 Sparse Blocked context onlyApr 17, 2026
- UsefulBench: Towards Decision-Useful Information as a Target for Information Retrievalarxiv-2604.15827 Sparse Blocked context onlyApr 17, 2026
- CHOP: Chunkwise Context-Preserving Framework for RAG on Multi Documentsarxiv-2604.15802 Sparse Blocked context onlyApr 17, 2026
- Target-Oriented Pretraining Data Selection via Neuron-Activated Grapharxiv-2604.15706 Sparse Blocked context onlyApr 17, 2026
- How Do LLMs and VLMs Understand Viewpoint Rotation Without Vision? An Interpretability Studyarxiv-2604.15294 Sparse Blocked context onlyApr 16, 2026
- AdaSplash-2: Faster Differentiable Sparse Attentionarxiv-2604.15180 Sparse Blocked context onlyApr 16, 2026
- Fabricator or dynamic translator?arxiv-2604.15165 Sparse Blocked context onlyApr 16, 2026
- An Axiomatic Benchmark for Evaluation of Scientific Novelty Metricsarxiv-2604.15145 Sparse Blocked context onlyApr 16, 2026
- IUQ: Interrogative Uncertainty Quantification for Long-Form Large Language Model Generationarxiv-2604.15109 Sparse Blocked context onlyApr 16, 2026
- Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimizationarxiv-2604.15022 Sparse Blocked context onlyApr 16, 2026
- Calibration-Gated LLM Pseudo-Observations for Online Contextual Banditsarxiv-2604.14961 Sparse Blocked context onlyApr 16, 2026
- WavAlign: Enhancing Intelligence and Expressiveness in Spoken Dialogue Models via Adaptive Hybrid Post-Trainingarxiv-2604.14932 Sparse Blocked context onlyApr 16, 2026
- Retrieve, Then Classify: Corpus-Grounded Automation of Clinical Value Set Authoringarxiv-2604.14616 Sparse Blocked context onlyApr 16, 2026
- Rhetorical Questions in LLM Representations: A Linear Probing Studyarxiv-2604.14128 Sparse Blocked context onlyApr 15, 2026
- UI-Zoomer: Uncertainty-Driven Adaptive Zoom-In for GUI Groundingarxiv-2604.14113 Sparse Blocked context onlyApr 15, 2026
- Reinforcement Learning via Value Gradient Flowarxiv-2604.14265 Sparse Blocked context onlyApr 15, 2026
- OneHOI: Unifying Human-Object Interaction Generation and Editingarxiv-2604.14062 Sparse Blocked context onlyApr 15, 2026
- GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectificationarxiv-2604.14258 Sparse Blocked context onlyApr 15, 2026
- Robust Reward Modeling for Large Language Models via Causal Decompositionarxiv-2604.13833 Sparse Blocked context onlyApr 15, 2026
- Parcae: Scaling Laws For Stable Looped Language Modelsarxiv-2604.12946 Sparse Blocked context onlyApr 14, 2026
- Beyond Prompt: Fine-grained Simulation of Cognitively Impaired Standardized Patients via Stochastic Steeringarxiv-2604.12210 Sparse Blocked context onlyApr 14, 2026
- Robust Explanations for User Trust in Enterprise NLP Systemsarxiv-2604.12069 Sparse Blocked context onlyApr 13, 2026
- LangFlow: Continuous Diffusion Rivals Discrete in Language Modelingarxiv-2604.11748 Curated Related Blocked context onlyApr 13, 2026
- Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetryarxiv-2604.10101 Sparse Blocked context onlyApr 11, 2026
- Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Modelsarxiv-2604.10079 Sparse Blocked context onlyApr 11, 2026
- Steered LLM Activations are Non-Surjectivearxiv-2604.09839 Sparse Blocked context onlyApr 10, 2026
- BERT-as-a-Judge: A Robust Alternative to Lexical Methods for Efficient Reference-Based LLM Evaluationarxiv-2604.09497 Sparse Blocked context onlyApr 10, 2026
- Cram Less to Fit More: Training Data Pruning Improves Memorization of Factsarxiv-2604.08519 Sparse Blocked context onlyApr 9, 2026
- AI generates well-liked but templatic empathic responsesarxiv-2604.08479 Sparse Blocked context onlyApr 9, 2026
- Synthetic Data for any Differentiable Targetarxiv-2604.08423 Sparse Blocked context onlyApr 9, 2026
- Selective Attention System (SAS): Device-Addressed Speech Detection for Real-Time On-Device Voice AIarxiv-2604.08412 Sparse Blocked context onlyApr 9, 2026
- A GAN and LLM-Driven Data Augmentation Framework for Dynamic Linguistic Pattern Modeling in Chinese Sarcasm Detectionarxiv-2604.08381 Sparse Blocked context onlyApr 9, 2026
- Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Tracesarxiv-2604.08362 Sparse Blocked context onlyApr 9, 2026
- HistDiT: A Structure-Aware Latent Conditional Diffusion Model for High-Fidelity Virtual Staining in Histopathologyarxiv-2604.08305 Sparse Blocked context onlyApr 9, 2026
- CIAO - Code In Architecture Out - Automated Software Architecture Documentation with Large Language Modelsarxiv-2604.08293 Sparse Blocked context onlyApr 9, 2026
- AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Planarxiv-2604.08184 Sparse Blocked context onlyApr 9, 2026
- OceanMAE: A Foundation Model for Ocean Remote Sensingarxiv-2604.08171 Sparse Blocked context onlyApr 9, 2026
- Training Data Size Sensitivity in Unsupervised Rhyme Recognitionarxiv-2604.08156 Sparse Blocked context onlyApr 9, 2026
- Graph Neural Networks for Misinformation Detection: Performance-Efficiency Trade-offsarxiv-2604.08131 Sparse Blocked context onlyApr 9, 2026
- LLM-Based Data Generation and Clinical Skills Evaluation for Low-Resource French OSCEsarxiv-2604.08126 Curated Related Blocked context onlyApr 9, 2026
- Initialisation Determines the Basin: Efficient Codebook Optimisation for Extreme LLM Quantizationarxiv-2604.08118 Sparse Blocked context onlyApr 9, 2026
- Revise: A Framework for Revising OCRed text in Practical Information Systems with Data Contamination Strategyarxiv-2604.08115 Sparse Blocked context onlyApr 9, 2026
- Quantum Vision Theory Applied to Audio Classification for Deepfake Speech Detectionarxiv-2604.08104 Sparse Blocked context onlyApr 9, 2026
- Kathleen: Oscillator-Based Byte-Level Text Classification Without Tokenization or Attentionarxiv-2604.07969 Sparse Blocked context onlyApr 9, 2026
- HCRE: LLM-based Hierarchical Classification for Cross-Document Relation Extraction with a Prediction-then-Verification Strategyarxiv-2604.07937 Curated Related Blocked context onlyApr 9, 2026
- Tool Retrieval Bridge: Aligning Vague Instructions with Retriever Preferences via Bridge Modelarxiv-2604.07816 Sparse Blocked context onlyApr 9, 2026
- AsyncTLS: Efficient Generative LLM Inference with Asynchronous Two-level Sparse Attentionarxiv-2604.07815 Sparse Blocked context onlyApr 9, 2026
- GRASS: Gradient-based Adaptive Layer-wise Importance Sampling for Memory-efficient Large Language Model Fine-tuningarxiv-2604.07808 Sparse Blocked context onlyApr 9, 2026
- SepSeq: A Training-Free Framework for Long Numerical Sequence Processing in LLMsarxiv-2604.07737 Sparse Blocked context onlyApr 9, 2026
- IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measuresarxiv-2604.07709 Sparse Blocked context onlyApr 9, 2026
- ADAG: Automatically Describing Attribution Graphsarxiv-2604.07615 Sparse Blocked context onlyApr 8, 2026
- CAMO: A Class-Aware Minority-Optimized Ensemble for Robust Language Model Evaluation on Imbalanced Dataarxiv-2604.07583 Sparse Blocked context onlyApr 8, 2026
- Learning is Forgetting: LLM Training As Lossy Compressionarxiv-2604.07569 Sparse Blocked context onlyApr 8, 2026
- TR-EduVSum: A Turkish-Focused Dataset and Consensus Framework for Educational Video Summarizationarxiv-2604.07553 Sparse Blocked context onlyApr 8, 2026
- ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Trainingarxiv-2604.07484 Sparse Blocked context onlyApr 8, 2026
- Why teaching resists automation in an AI-inundated era: Human judgment, non-modular work, and the limits of delegationarxiv-2604.07285 Sparse Blocked context onlyApr 8, 2026
- On the Price of Privacy for Language Identification and Generationarxiv-2604.07238 Sparse Blocked context onlyApr 8, 2026
- Dynamic Context Evolution for Scalable Synthetic Data Generationarxiv-2604.07147 Sparse Blocked context onlyApr 8, 2026
- Language Bias under Conflicting Information in Multilingual LLMsarxiv-2604.07123 Sparse Blocked context onlyApr 8, 2026
- Are Non-English Papers Reviewed Fairly? Language-of-Study Bias in NLP Peer Reviewsarxiv-2604.07119 Sparse Blocked context onlyApr 8, 2026
- The Impact of Steering Large Language Models with Persona Vectors in Educational Applicationsarxiv-2604.07102 Sparse Blocked context onlyApr 8, 2026
- Selective Neuron Amplification for Training-Free Task Enhancementarxiv-2604.07098 Sparse Blocked context onlyApr 8, 2026
- Corpora deduplication or duplication in Natural Language Processing of few resourced languages ? A case of study: The Mexico's Nahuatlarxiv-2604.07015 Sparse Blocked context onlyApr 8, 2026
- Continuous Interpretive Steering for Scalar Diversityarxiv-2604.07006 Sparse Blocked context onlyApr 8, 2026
- ChunQiuTR: Time-Keyed Temporal Retrieval in Classical Chinese Annalsarxiv-2604.06997 Sparse Blocked context onlyApr 8, 2026
- The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Eraarxiv-2604.06906 Sparse Blocked context onlyApr 8, 2026
- Is Biomedical Specialization Still Worth It? Insights from Domain-Adaptive Language Modelling with a New French Health Corpusarxiv-2604.06903 Sparse Blocked context onlyApr 8, 2026
- To Adapt or not to Adapt, Rethinking the Value of Medical Knowledge-Aware Large Language Modelsarxiv-2604.06854 Sparse Blocked context onlyApr 8, 2026
- MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviorsarxiv-2604.06846 Sparse Blocked context onlyApr 8, 2026
- Environmental, Social and Governance Sentiment Analysis on Slovene News: A Novel Dataset and Modelsarxiv-2604.06826 Sparse Blocked context onlyApr 8, 2026
- AGSC: Adaptive Granularity and Semantic Clustering for Uncertainty Quantification in Long-text Generationarxiv-2604.06812 Sparse Blocked context onlyApr 8, 2026
- Multilingual Cognitive Impairment Detection in the Era of Foundation Modelsarxiv-2604.06758 Sparse Blocked context onlyApr 8, 2026
- SQLStructEval: Structural Evaluation of LLM Text-to-SQL Generationarxiv-2604.06736 Curated Related Blocked context onlyApr 8, 2026
- A Graph-Enhanced Defense Framework for Explainable Fake News Detection with LLMarxiv-2604.06666 Sparse Blocked context onlyApr 8, 2026
- Team Fusion@ SU@ BC8 SympTEMIST track: transformer-based approach for symptom recognition and linkingarxiv-2604.06424 Sparse Blocked context onlyApr 7, 2026
- FMI@SU ToxHabits: Evaluating LLMs Performance on Toxic Habit Extraction in Spanish Clinical Textsarxiv-2604.06403 Sparse Blocked context onlyApr 7, 2026
- ART: Attention Replacement Technique to Improve Factuality in LLMsarxiv-2604.06393 Sparse Blocked context onlyApr 7, 2026
- Severity-Aware Weighted Loss for Arabic Medical Text Generationarxiv-2604.06346 Sparse Blocked context onlyApr 7, 2026
- In-Place Test-Time Trainingarxiv-2604.06169 Direct Blocked context onlyApr 7, 2026
- Exclusive Unlearningarxiv-2604.06154 Sparse Blocked context onlyApr 7, 2026
- LAG-XAI: A Lie-Inspired Affine Geometric Framework for Interpretable Paraphrasing in Transformer Latent Spacesarxiv-2604.06086 Sparse Blocked context onlyApr 7, 2026
- The Model Agreed, But Didn't Learn: Diagnosing Surface Compliance in Large Language Modelsarxiv-2604.05995 Sparse Blocked context onlyApr 7, 2026
- BOSCH: Black-Box Binary Optimization for Short-Context Attention-Head Selection in LLMsarxiv-2604.05942 Sparse Blocked context onlyApr 7, 2026
- FrontierFinance: A Long-Horizon Computer-Use Benchmark of Real-World Financial Tasksarxiv-2604.05912 Sparse Blocked context onlyApr 7, 2026
- Swiss-Bench 003: Evaluating LLM Reliability and Adversarial Security for Swiss Regulatory Contextsarxiv-2604.05872 Sparse Blocked context onlyApr 7, 2026
- GenomeQA: Benchmarking General Large Language Models for Genome Sequence Understandingarxiv-2604.05774 Sparse Blocked context onlyApr 7, 2026
- SemLink: A Semantic-Aware Automated Test Oracle for Hyperlink Verification using Siamese Sentence-BERTarxiv-2604.05711 Sparse Blocked context onlyApr 7, 2026
- YoNER: A New Yorùbá Multi-domain Named Entity Recognition Datasetarxiv-2604.05624 Sparse Blocked context onlyApr 7, 2026
- AI-Driven Modular Services for Accessible Multilingual Education in Immersive Extended Reality Settings: Integrating Speech Processing, Translation, and Sign Language Renderingarxiv-2604.05591 Sparse Blocked context onlyApr 7, 2026
- Bridging Natural Language and Microgrid Dynamics: A Context-Aware Simulator and Datasetarxiv-2604.05429 Sparse Blocked context onlyApr 7, 2026
- Hierarchical SVG Tokenization: Learning Compact Visual Programs for Scalable Vector Graphics Modelingarxiv-2604.05072 Direct Blocked context onlyApr 6, 2026
- Beyond the Final Actor: Modeling the Dual Roles of Creator and Editor for Fine-Grained LLM-Generated Text Detectionarxiv-2604.04932 Sparse Blocked context onlyApr 6, 2026
- Data Attribution in Adaptive Learningarxiv-2604.04892 Sparse Blocked context onlyApr 6, 2026
- Noise Immunity in In-Context Tabular Learning: An Empirical Robustness Analysis of TabPFN's Attention Mechanismsarxiv-2604.04868 Sparse Blocked context onlyApr 6, 2026
- HUKUKBERT: Domain-Specific Language Model for Turkish Lawarxiv-2604.04790 Sparse Blocked context onlyApr 6, 2026
- Darkness Visible: Reading the Exception Handler of a Language Modelarxiv-2604.04756 Sparse Blocked context onlyApr 6, 2026
- Hallucination Basins: A Dynamic Framework for Understanding and Controlling LLM Hallucinationsarxiv-2604.04743 Sparse Blocked context onlyApr 6, 2026
- ZeD-MAP: Bundle Adjustment Guided Zero-Shot Depth Maps for Real-Time Aerial Imagingarxiv-2604.04667 Sparse Blocked context onlyApr 6, 2026
- SLSREC: Self-Supervised Contrastive Learning for Adaptive Fusion of Long- and Short-Term User Interestsarxiv-2604.04530 Sparse Blocked context onlyApr 6, 2026
- Compressible Softmax-Attended Language under Incompressible Attentionarxiv-2604.04384 Sparse Blocked context onlyApr 6, 2026
- Convolutional Neural Network and Adversarial Autoencoder in EEG images classificationarxiv-2604.04313 Sparse Blocked context onlyApr 5, 2026
- StoryScope: Investigating idiosyncrasies in AI fictionarxiv-2604.03136 Sparse Blocked context onlyApr 3, 2026
- Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharingarxiv-2604.22782 Sparse Blocked context onlyApr 3, 2026
- VISTA: Visualization of Token Attribution via Efficient Analysisarxiv-2604.02217 Sparse Blocked context onlyApr 2, 2026
- CV-18 NER: Augmented Common Voice for Named Entity Recognition from Arabic Speecharxiv-2604.02209 Direct Blocked context onlyApr 2, 2026
- LEO: Graph Attention Network based Hybrid Multi Sensor Extended Object Fusion and Tracking for Autonomous Driving Applicationsarxiv-2604.02206 Sparse Blocked context onlyApr 2, 2026
- AstroConcepts: A Large-Scale Multi-Label Classification Corpus for Astrophysicsarxiv-2604.02156 Sparse Blocked context onlyApr 2, 2026
- TRACE-Bot: Detecting Emerging LLM-Driven Social Bots via Implicit Semantic Representations and AIGC-Enhanced Behavioral Patternsarxiv-2604.02147 Sparse Blocked context onlyApr 2, 2026
- GaelEval: Benchmarking LLM Performance for Scottish Gaelicarxiv-2604.02135 Sparse Blocked context onlyApr 2, 2026
- BidirLM: From Text to Omnimodal Bidirectional Encoders by Adapting and Composing Causal LLMsarxiv-2604.02045 Sparse Blocked context onlyApr 2, 2026
- From Guessing to Placeholding: A Cost-Theoretic Framework for Uncertainty-Aware Code Completionarxiv-2604.01849 Sparse Blocked context onlyApr 2, 2026
- Development and multi-center evaluation of domain-adapted speech recognition for human-AI teaming in real-world gastrointestinal endoscopyarxiv-2604.01705 Sparse Blocked context onlyApr 2, 2026
- Coupled Query-Key Dynamics for Attentionarxiv-2604.01683 Sparse Blocked context onlyApr 2, 2026
- Swift-SVD: Theoretical Optimality Meets Practical Efficiency in Low-Rank LLM Compressionarxiv-2604.01609 Sparse Blocked context onlyApr 2, 2026
- Countering Catastrophic Forgetting of Large Language Models for Better Instruction Following via Weight-Space Model Mergingarxiv-2604.01538 Sparse Blocked context onlyApr 2, 2026
- Cost-Efficient Estimation of General Abilities Across Benchmarksarxiv-2604.01418 Sparse Blocked context onlyApr 1, 2026
- ReFormeR: Learning and Applying Explicit Query Reformulation Patternsarxiv-2604.01417 Sparse Blocked context onlyApr 1, 2026