300 canonical paper links on this archive page.
- CausalMix: Data Mixture as Causal Inference for Language Model Trainingarxiv-2607.01104 Sparse Blocked context onlyJul 1, 2026
- Quantifying the Affective Gap: A Zero-Shot Evaluation of LLMs on Fine-Grained Emotion Taxonomiesarxiv-2607.00968 Sparse Blocked context onlyJul 1, 2026
- Persona Non Grata: LLM Persona-Driven Generations in MCQA are Unstable in Distinct Dimensionsarxiv-2607.00937 Sparse Blocked context onlyJul 1, 2026
- RuleChef: Grounding LLM Task Knowledge in Human-Editable Rulesarxiv-2607.01293 Sparse Blocked context onlyJul 1, 2026
- MultiSynt/MT: Trillion-Token Multi-Parallel Pre-Training Data Translated Across 36 Languagesarxiv-2607.00890 Sparse Blocked context onlyJul 1, 2026
- "Don't Say It!": Constraints, Compliance, and Communication when Language Models Play Tabooarxiv-2607.00601 Sparse Blocked context onlyJul 1, 2026
- BaseRT: Best-in-Class LLM Inference on Apple Silicon via Native Metalarxiv-2607.00501 Sparse Blocked context onlyJul 1, 2026
- NeuroCogMap Reveals Cognitive Organization of Large Language Modelsarxiv-2607.00397 Sparse Blocked context onlyJul 1, 2026
- Beyond Perplexity: A Behavioral Evaluation Framework for Deployment-Memory Claims in LLM Test-Time Trainingarxiv-2607.00368 Sparse Blocked context onlyJul 1, 2026
- An LLM-Based Framework for Intent-Driven Network Topology Designarxiv-2607.00292 Sparse Blocked context onlyJul 1, 2026
- LV-ROVER-MLT: Low-Resource Maltese OCR by Multi-Stream Votingarxiv-2607.00250 Sparse Blocked context onlyJun 30, 2026
- Temporal Multi-Signal Fusion for Token-Level Hallucination Detectionarxiv-2608.18115 Sparse Blocked context onlyJun 30, 2026
- SpheRoPE: Zero-Shot Optimization-Free 360 Panorama Generation with Spherical RoPEarxiv-2606.32033 Sparse Blocked context onlyJun 30, 2026
- AnyBokeh: Physics-Guided Any-to-Any Bokeh Editing with Optical Fingerprint Transferarxiv-2606.31959 Sparse Blocked context onlyJun 30, 2026
- PhotoQuilt: Training-Free Arbitrary-Resolution Photomosaics via Bootstrapped Tiled Denoisingarxiv-2606.30968 Sparse Blocked context onlyJun 29, 2026
- Goku: A Million-Scale Universal Dataset and Benchmark for Instruction-Based Video Editingarxiv-2606.30599 Sparse Blocked context onlyJun 29, 2026
- DialogPII: A multilingual dataset of synthetic dialog transcripts to detect personal informationarxiv-2606.30312 Sparse Blocked context onlyJun 29, 2026
- Information Dynamics of Language Communicationarxiv-2606.30096 Sparse Blocked context onlyJun 29, 2026
- Little Brains, Big Feats: Exploring Compact Language Modelsarxiv-2606.30062 Sparse Blocked context onlyJun 29, 2026
- Walking in the Implicit: Interactive World Exploration via Neural Scene Representationarxiv-2606.30045 Sparse Blocked context onlyJun 29, 2026
- Smooth Scaling Laws Hide Stepwise Token Learningarxiv-2606.29858 Sparse Blocked context onlyJun 29, 2026
- MATCH: Modulating Attention via In-Context Retrieval for Long-Context Transformersarxiv-2606.29844 Sparse Blocked context onlyJun 29, 2026
- Revealing the Technology Development of Natural Language Processing: A Scientific Entity-Centric Perspectivearxiv-2606.29836 Sparse Blocked context onlyJun 29, 2026
- SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Modelsarxiv-2606.29815 Sparse Blocked context onlyJun 29, 2026
- Managing Map Cardinality in Automatic Disease Classification Mapping: Balancing Precision, Recall and Coveragearxiv-2606.29750 Sparse Blocked context onlyJun 29, 2026
- Diagnosing and Mitigating Context Rot in Long-horizon Searcharxiv-2606.29718 Sparse Blocked context onlyJun 29, 2026
- MAM-AI: An On-Device Medical Retrieval-Augmented Generation System for Nurses and Midwives in Zanzibararxiv-2606.29580 Sparse Blocked context onlyJun 28, 2026
- Preference-ASR: A Preference-Aware Test Set for Benchmarking ASR in the Era of Speech LLMsarxiv-2606.29534 Sparse Blocked context onlyJun 28, 2026
- mamabench and mamaretrieval: Benchmarks for Evaluating Medical Retrieval-Augmented Generation in Maternal, Neonatal, and Reproductive Healtharxiv-2606.29467 Sparse Blocked context onlyJun 28, 2026
- LC-ICL: Label-Guided Contrastive In-Context Learning for Robust Information Extractionarxiv-2606.29407 Sparse Blocked context onlyJun 28, 2026
- Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Modelsarxiv-2606.29196 Sparse Blocked context onlyJun 28, 2026
- AB-RAG: Adaptive Budgeted Retrieval-Augmented Generation for Reliable Question Answeringarxiv-2606.29090 Sparse Blocked context onlyJun 27, 2026
- Conversational Domain Adaptation of IndicTrans2 across 21 Indic Languages via Experience Replay and Model Soupsarxiv-2606.29024 Sparse Blocked context onlyJun 27, 2026
- Fine-Tuning General-Purpose Large Language Models for Agricultural Applications:A Reproducible Framework and Evaluation Protocol Based on Qwen3-8Barxiv-2606.28992 Sparse Blocked context onlyJun 27, 2026
- Memory-Managed Long-Context Attention: A Preliminary Study of Editable Request-Local Memoryarxiv-2606.28876 Sparse Blocked context onlyJun 27, 2026
- The Heterogeneous Safety Impacts of Benign Multilingual Fine-Tuningarxiv-2606.28843 Sparse Blocked context onlyJun 27, 2026
- 5ting at SemEval-2026 Task 8: Strong End-to-End Multi-Turn RAG via LLM-Based Reranking and Faithfulness Controlarxiv-2606.28737 Sparse Blocked context onlyJun 27, 2026
- What LLMs explain is not what they believe: Evaluating explanation sufficiency under models' own input beliefsarxiv-2606.28615 Sparse Blocked context onlyJun 26, 2026
- Correct codes for the wrong reasons? validating LLMs as measurement instruments for theoretical constructsarxiv-2606.28574 Sparse Blocked context onlyJun 26, 2026
- Depth-Staggered Fibonacci Spacing for Sparse Attention: Static Schedules Beat Learned Dilation and Extrapolate Where Dense Attention Failsarxiv-2606.28560 Sparse Blocked context onlyJun 26, 2026
- A French OSCE Dialogue Dataset and Controllable Virtual Patient System for Clinical Trainingarxiv-2606.28526 Sparse Blocked context onlyJun 26, 2026
- MultiHashFormer: Hash-based Generative Language Modelsarxiv-2606.28057 Sparse Blocked context onlyJun 26, 2026
- Parameter-Efficient Quantum-Inspired Fast Weight Programmers for Traffic-Matrix Forecastingarxiv-2606.27821 Sparse Blocked context onlyJun 26, 2026
- Cluster, Route, Escalate: Cascaded Framework for Cost-Aware LLM Servingarxiv-2606.27457 Sparse Blocked context onlyJun 25, 2026
- COrigami: An AI Pipeline for Co-Designing Flat-Foldable Visually Recognisable Origamiarxiv-2606.26299 Sparse Blocked context onlyJun 24, 2026
- Detect, Unlearn, Restore: Defending Text Summarization Models Against Data Poisoningarxiv-2606.26036 Sparse Blocked context onlyJun 24, 2026
- Dziri Voicebot: An End-to-End Low-Resource Speech-to-Speech Conversational System for Algerian Dialectarxiv-2606.26003 Sparse Blocked context onlyJun 24, 2026
- How Large Language Models Source Brand Reputation Across Languages and Marketsarxiv-2606.25787 Sparse Blocked context onlyJun 24, 2026
- Do Encoders Suffice? A Systematic Comparison of Encoder and Decoder Safety Judges for LLM Adversarial Evaluationarxiv-2606.25782 Sparse Blocked context onlyJun 24, 2026
- Evaluating LLMs on Real-World Software Performance Optimizationarxiv-2606.25530 Sparse Blocked context onlyJun 24, 2026
- Spam and Sentiment Detection in Arabic Tweets Using MARBERT Modelarxiv-2606.25495 Sparse Blocked context onlyJun 24, 2026
- How Reliable Is Your Jailbreak Judge? Calibration and Adversarial Robustness of Automated ASR Scoringarxiv-2606.25487 Sparse Blocked context onlyJun 24, 2026
- A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluationarxiv-2606.25476 Sparse Blocked context onlyJun 24, 2026
- Optimizing Abstractive Summarization With Fine-Tuned PEGASUSarxiv-2606.25462 Sparse Blocked context onlyJun 24, 2026
- Reclaim Evaluation: A Lossy Memory Is Worse Than an Empty Onearxiv-2606.25449 Sparse Blocked context onlyJun 24, 2026
- Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesisarxiv-2606.25369 Sparse Blocked context onlyJun 24, 2026
- Neural Machine Translation for Low-Resource Tangkhul--Englisharxiv-2606.25365 Curated Related Blocked context onlyJun 24, 2026
- Improved Large Language Diffusion Modelsarxiv-2606.25331 Sparse Blocked context onlyJun 24, 2026
- Towards Structuring an Arabic-English Machine-Readable Dictionary Using Parsing Expression Grammarsarxiv-2606.25231 Sparse Blocked context onlyJun 23, 2026
- What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamicsarxiv-2606.25182 Sparse Blocked context onlyJun 23, 2026
- LLM-ACES: Closed-Loop Discovery of Dynamical Systems with LLM-Guided Adaptive Searcharxiv-2606.25039 Sparse Blocked context onlyJun 23, 2026
- DiffusionBench: On Holistic Evaluation of Diffusion Transformersarxiv-2606.24888 Sparse Blocked context onlyJun 23, 2026
- L3Cube-MahaPOS: A Marathi Part-of-Speech Tagging Dataset and BERT Modelsarxiv-2606.24825 Sparse Blocked context onlyJun 23, 2026
- DREAM: Dense Retrieval Embeddings via Autoregressive Modelingarxiv-2606.24667 Curated Related Blocked context onlyJun 23, 2026
- AI-PAVE-Br: Leveraging Large Language Models for Enhanced Product Attribute Value Extraction through a Golden Set Approacharxiv-2606.24655 Sparse Blocked context onlyJun 23, 2026
- ParaPairAudioBench: Paralinguistic Pairwise Audio Benchmark for LALM-as-a-Judgearxiv-2606.24648 Sparse Blocked context onlyJun 23, 2026
- Measuring User's Mental Models of Speech Translation in Human-AI Collaborationarxiv-2606.24644 Sparse Blocked context onlyJun 23, 2026
- The African Language Tax: Quantifying the Cost, Latency, and Context Penalty of Tokenizing African Languages in Frontier LLMsarxiv-2606.24460 Sparse Blocked context onlyJun 23, 2026
- LLM Performance on a Real, Double-Marked GCSE Benchmarkarxiv-2606.24973 Sparse Blocked context onlyJun 23, 2026
- Beyond Logprobs: A Multi-Signal Confidence Engine for LLM-Based Document Field Extractionarxiv-2606.24420 Sparse Blocked context onlyJun 23, 2026
- AutoSpecNER: A Fine-Grained Named Entity Recognition Dataset for Vehicle Specification Extractionarxiv-2606.24387 Sparse Blocked context onlyJun 23, 2026
- Automatic Part-of-Speech Tagging of Arabic-English Dictionary Senses through WordNetarxiv-2606.24359 Sparse Blocked context onlyJun 23, 2026
- Dialogue to Discovery: Attribute-Aware Preference Elicitation for Conversational Product Search Assistantsarxiv-2606.24194 Sparse Blocked context onlyJun 23, 2026
- When Top-1 Fails: Calibrating LoRA Monitors for Masked Diffusion LMsarxiv-2606.24119 Curated Related Blocked context onlyJun 23, 2026
- Predicting Poets' Origins from Verse: A Computational Analysis of Regional Linguistic Fingerprints in the Complete Tang Poemsarxiv-2606.24093 Sparse Blocked context onlyJun 23, 2026
- RASC+: Retrieval-Constrained LLM Adjudication for Clinical Value Set Authoringarxiv-2606.23992 Sparse Blocked context onlyJun 22, 2026
- Faithful by Construction: Claim-Anchored Attribution for Multi-Document Summarizationarxiv-2606.23989 Sparse Blocked context onlyJun 22, 2026
- One Year Later...The Harms Persist, But So Do We!arxiv-2606.23884 Sparse Blocked context onlyJun 22, 2026
- Tapered Language Modelsarxiv-2606.23670 Sparse Blocked context onlyJun 22, 2026
- Vera: A Layered Diffusion Model for Content-Preserving Video Editingarxiv-2606.23610 Sparse Blocked context onlyJun 22, 2026
- MeshFlow: Mesh Generation with Equivariant Flow Matchingarxiv-2606.23489 Sparse Blocked context onlyJun 22, 2026
- FedOT: Ownership Verification and Leakage Tracing via Watermarks for Federated LDMsarxiv-2606.22875 Sparse Blocked context onlyJun 22, 2026
- HAKARI-Bench: A Lightweight Benchmark for Comparing Retrieval Architectures and Efficiency Settings under Unified Conditionsarxiv-2606.22778 Direct Blocked context onlyJun 22, 2026
- EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videosarxiv-2607.09701 Sparse Blocked context onlyJun 21, 2026
- A Hybrid, Multi-Layered Pipeline for Phishing and Threat Classification: Independently Validated URL and NLP Engines with a Calibrated Multi-Channel Fusion Stagearxiv-2606.21690 Sparse Blocked context onlyJun 19, 2026
- Toward Open Weight Models Without Risks: Separating Public and Private Capabilities in LLMsarxiv-2606.21638 Sparse Blocked context onlyJun 19, 2026
- Precision Recall Controllable Radiology Report Generation via Hybrid Natural Language and Clinical Reward Learningarxiv-2606.21447 Sparse Blocked context onlyJun 19, 2026
- An Exploratory Case Study of LLM-Assisted Refactoring and Gameplay Feature Generation in an Endless Runner Gamearxiv-2606.21171 Sparse Blocked context onlyJun 19, 2026
- Go-with-the-Track: Video Compositing and Motion Control with Point Trackingarxiv-2606.20891 Sparse Blocked context onlyJun 18, 2026
- CATCH-ME if you RAG: a dataset of Contextually Annotated multi-Turn Counterspeech against Hate and Misinformation Exchangesarxiv-2606.20369 Sparse Blocked context onlyJun 18, 2026
- The Register Gap: A Meaning Intelligence Framework for Nigerian Public Discoursearxiv-2606.20255 Sparse Blocked context onlyJun 18, 2026
- CzechDocs: A Multiway Parallel Dataset of Formatted Documents for Minority Languages in Czechiaarxiv-2606.20212 Sparse Blocked context onlyJun 18, 2026
- Apparent Psychological Profiles of Large Language Models are Largely a Measurement Artifactarxiv-2606.20205 Sparse Blocked context onlyJun 18, 2026
- From Texts to Scores: Tracing the Emergence of Essay Quality Representations in Large Language Modelsarxiv-2606.20152 Sparse Blocked context onlyJun 18, 2026
- Learning to Prompt: Improving Student Engagement with Adaptive LLM-based High-School Tutoringarxiv-2606.20138 Sparse Blocked context onlyJun 18, 2026
- Self-Preference Is Weak or Absent in Verifiable Instruction-Following Revision: A Four-Model Test Under Genuine Authorshiparxiv-2606.20093 Sparse Blocked context onlyJun 18, 2026
- IHUBERT: Vector-Based Semantic Deduplication and Domain-Balanced Pretraining for Persian Resourcesarxiv-2606.20089 Sparse Blocked context onlyJun 18, 2026
- Generative Engine Optimization at Scale: Measuring Brand Visibility Across AI Search Enginesarxiv-2606.20065 Sparse Blocked context onlyJun 18, 2026
- GEMS: Geometric Constraints Enable Multi-Semantic Superposition in LLMsarxiv-2606.19946 Sparse Blocked context onlyJun 18, 2026
- REDACT: A Systematically Controlled Multilingual Benchmark for Personal Information Detectionarxiv-2606.19881 Sparse Blocked context onlyJun 18, 2026
- The Almost Intelligent Revolution: Options for Scaling Up Deliberation and Empowering People with AIarxiv-2606.19864 Sparse Blocked context onlyJun 18, 2026
- CREDENCE: Claim Reduction for Decomposition & Enhanced Credibility -- Semantic Metrics and Convergence Analysisarxiv-2606.19819 Sparse Blocked context onlyJun 18, 2026
- Hadith computational science in the age of large language models: a critical narrative reviewarxiv-2608.20364 Sparse Blocked context onlyJun 18, 2026
- FineREX: Fine-Tuned NER-RE for Human Smuggling Knowledge Graphsarxiv-2606.19710 Sparse Blocked context onlyJun 18, 2026
- TerraMARS: A Domain-Adapted Small-Language-Model Pipeline for Mars Terraforming Literaturearxiv-2606.19700 Sparse Blocked context onlyJun 18, 2026
- A Layered Security Framework Against Prompt Injection in RAG-Based Chatbotsarxiv-2606.19660 Sparse Blocked context onlyJun 17, 2026
- BrainG3N: A Dual-Purpose Tokenizer for Controllable 3D Brain MRI Generationarxiv-2606.19651 Sparse Blocked context onlyJun 17, 2026
- From 50K to 8.2 Million in 24 Hours: Vozinha's Algorithmic Consecration and the Multilingual Making of World Cup Visibilityarxiv-2606.19647 Sparse Blocked context onlyJun 17, 2026
- Creating Multilingual Mental Health Dialogue Datasets: Limits of Persona-Based Localization via Nationality and Languagearxiv-2606.19640 Sparse Blocked context onlyJun 17, 2026
- Reliability without Validity: A Systematic, Large-Scale Evaluation of LLM-as-a-Judge Models Across Agreement, Consistency, and Biasarxiv-2606.19544 Sparse Blocked context onlyJun 17, 2026
- Language Models as Interfaces, Not Oracles: A Hybrid LLM-ML System for Pediatric Appendicitisarxiv-2606.19183 Sparse Blocked context onlyJun 17, 2026
- Dango: A Strictly L1-Only Large Language Model for Studying Second Language Acquisitionarxiv-2606.19170 Sparse Blocked context onlyJun 17, 2026
- IndicContextEval: A Benchmark for Evaluating Context Utilisation in Audio Large Language Models Across 8 Indic Languagesarxiv-2606.19157 Sparse Blocked context onlyJun 17, 2026
- As Easy as Rocket Science: Assessing the Ability of Large Language Models to Interpret Negation in Figurative Languagearxiv-2606.18922 Sparse Blocked context onlyJun 17, 2026
- Approximate Structured Diffusion for Sequence Labellingarxiv-2606.18856 Sparse Blocked context onlyJun 17, 2026
- Beyond Scalar Scores: Exploring LLM-based Metrics for Clinical Significance Evaluation in Radiology Reportsarxiv-2606.18797 Sparse Blocked context onlyJun 17, 2026
- SproutRAG: Attention-Guided Tree Search with Progressive Embeddings for Long-Document RAGarxiv-2606.18381 Sparse Blocked context onlyJun 16, 2026
- RepSelect: Robust LLM Unlearning via Representation Selectivityarxiv-2606.17168 Sparse Blocked context onlyJun 15, 2026
- Human Universal Graspingarxiv-2606.17054 Sparse Blocked context onlyJun 15, 2026
- Understanding the Behaviors of Environment-aware Information Retrievalarxiv-2606.16817 Sparse Blocked context onlyJun 15, 2026
- MMDiff: Extending Diffusion Transformers for Multi-Modal Generationarxiv-2606.16673 Sparse Blocked context onlyJun 15, 2026
- Distilling Examples into Task Instructions: Enhanced In-Context Learning for Real-World B2B Conversationsarxiv-2606.15641 Sparse Blocked context onlyJun 14, 2026
- Rethinking the Role of Efficient Attention in Hybrid Architecturesarxiv-2606.15378 Sparse Blocked context onlyJun 13, 2026
- RefGC-SR$^2$: Reference-guided Generated Content Super-Resolution and Refinementarxiv-2606.15158 Sparse Blocked context onlyJun 13, 2026
- Memento: Reconstruct to Remember for Consistent Long Video Generationarxiv-2606.14667 Curated Related Blocked context onlyJun 12, 2026
- HiLo-Token: Input-Adaptive High-Low Frequency Token Compression for Efficient Image Editingarxiv-2606.13898 Sparse Blocked context onlyJun 11, 2026
- RepWAM: World Action Modeling with Representation Visual-Action Tokenizersarxiv-2606.13674 Sparse Blocked context onlyJun 11, 2026
- World Tracing: Generative Pixel-Aligned Geometry Beyond the Visiblearxiv-2606.13652 Direct Blocked context onlyJun 11, 2026
- No Hidden Prompts Needed! You Can Game AI Peer Review with Presentation-Only Revisionsarxiv-2606.13044 Sparse Blocked context onlyJun 11, 2026
- Bag of Dims: Training-Free Mechanistic Interpretability via Dimension-Level Sign Patternsarxiv-2606.12629 Sparse Blocked context onlyJun 10, 2026
- eCREAM-MedCorpus A Large-Scale Corpus of Clinical Notes for Italianarxiv-2606.12569 Sparse Blocked context onlyJun 10, 2026
- Adaptive Multi-Resolution Procedural Knowledge Compression for Large Language Modelsarxiv-2606.12203 Sparse Blocked context onlyJun 10, 2026
- On the Limits of LLM-as-Judge for Scientific Novelty Assessmentarxiv-2606.12071 Sparse Blocked context onlyJun 10, 2026
- LLM-Enabled NWDAF: A Step Toward AI-Native 6G Network Intelligencearxiv-2606.11877 Sparse Blocked context onlyJun 10, 2026
- Lius: Translation Model Based Instructional Lingustic Using Continual Instruction Tuning In Kupang Malayarxiv-2606.11786 Sparse Blocked context onlyJun 10, 2026
- ICA Lens: Interpreting Language Models Without Training Another Dictionaryarxiv-2606.11722 Sparse Blocked context onlyJun 10, 2026
- i1: A Simple and Fully Open Recipe for Strong Text-to-Image Modelsarxiv-2606.11289 Direct Blocked context onlyJun 9, 2026
- Test-Time Gradient Guidance of Flow Policies in Reinforcement Learningarxiv-2606.11087 Sparse Blocked context onlyJun 9, 2026
- U-TTT: Towards Generalizable PET Image Denoising via Test-Time Trainingarxiv-2606.11032 Sparse Blocked context onlyJun 9, 2026
- Towards Diverse Scientific Hypothesis Search with Large Language Modelsarxiv-2606.10587 Sparse Blocked context onlyJun 9, 2026
- LC-QAT: Data-Efficient 2-Bit QAT for LLMs via Linear-Constrained Vector Quantizationarxiv-2606.10531 Sparse Blocked context onlyJun 9, 2026
- UniSVQ: 2-bit Unified Scalar-Vector Quantizationarxiv-2606.10520 Sparse Blocked context onlyJun 9, 2026
- Rethinking the Divergence Regularization in LLM RLarxiv-2606.09821 Sparse Blocked context onlyJun 8, 2026
- PsychoSafe: Eliciting Psychologically-Informed Refusals in Large Language Modelsarxiv-2606.09697 Sparse Blocked context onlyJun 8, 2026
- Leveraging Morphology for Historical Script Metrological Analysisarxiv-2606.09446 Sparse Blocked context onlyJun 8, 2026
- MilliVid: Hierarchical Latents for Long-Range Consistency in Video Generationarxiv-2606.09056 Sparse Blocked context onlyJun 8, 2026
- CoVEBench: Can Video Editing Models Handle Complex Instructions?arxiv-2606.08415 Sparse Blocked context onlyJun 7, 2026
- EmpiriGraph-Psy: A Dataset and LLM Pipeline for Extracting Empirical Relation Graphs from Psychology Abstractsarxiv-2606.08362 Sparse Blocked context onlyJun 6, 2026
- Chiaroscuro Attention: Spending Compute in the Darkarxiv-2606.08327 Sparse Blocked context onlyJun 6, 2026
- Shared Semantics, Divergent Mechanisms: Unsupervised Feature Discovery by Aligning Semantics and Mechanismsarxiv-2606.08236 Sparse Blocked context onlyJun 6, 2026
- Phase Marginalization for Patch-Grid Instability in Vision Transformersarxiv-2606.08132 Sparse Blocked context onlyJun 6, 2026
- Revisiting Articulated Parts Perception in Robot Manipulationarxiv-2606.08103 Sparse Blocked context onlyJun 6, 2026
- When Behavioral Safety Evaluation Fails: A Representation-Level Perspectivearxiv-2606.08044 Sparse Blocked context onlyJun 6, 2026
- Your UnEmbedding Matrix is Secretly a Feature Lens for Text Embeddingsarxiv-2606.07502 Sparse Blocked context onlyJun 5, 2026
- How Far Can Chord-Symbol Time-Series Adaptation Carry Genre Identity? Capabilities and Boundaries in Multi-Genre Chord-Symbol Modelingarxiv-2606.07334 Sparse Blocked context onlyJun 5, 2026
- SigmaScale: LLM Compression with SVD-based Low-Rank Decomposition and Learned Scaling Matricesarxiv-2606.07098 Sparse Blocked context onlyJun 5, 2026
- Re-Centering Humans in LLM Personalizationarxiv-2606.06614 Sparse Blocked context onlyJun 4, 2026
- RhymeFlow: Training-Free Acceleration for Video Generation with Asynchronous Denoising Flow Schedulingarxiv-2606.06309 Sparse Blocked context onlyJun 4, 2026
- Tangram: Unlocking Non-Uniform KV Cache Compression for Efficient Multi-turn LLM Servingarxiv-2606.06302 Sparse Blocked context onlyJun 4, 2026
- LLMs Can Leak Training Data But Do They Want To? A Propensity-Aware Evaluation of Memorization in LLMsarxiv-2606.06286 Sparse Blocked context onlyJun 4, 2026
- Improving Answer Extraction in Context-based Question Answering Systems Using LLMsarxiv-2606.06197 Sparse Blocked context onlyJun 4, 2026
- Severity-Aware Curriculum Learning with Multi-Model Response Selection for Medical Text Generationarxiv-2606.05510 Sparse Blocked context onlyJun 3, 2026
- MASF: A Multi-Model Adaptive Selection Framework for Abstractive Text summarizationarxiv-2606.05494 Sparse Blocked context onlyJun 3, 2026
- Statistically Reliable LLM-Based Ranking Evaluation via Prediction-Powered Inferencearxiv-2606.05308 Sparse Blocked context onlyJun 3, 2026
- STRIDE: Training Data Attribution via Sparse Recovery from Subset Perturbationsarxiv-2606.05165 Sparse Blocked context onlyJun 3, 2026
- Audio Interaction Modelarxiv-2606.05121 Direct Blocked context onlyJun 3, 2026
- CIPER: A Unified Framework for Cross-view Image-retrieval and Pose-estimationarxiv-2606.05011 Sparse Blocked context onlyJun 3, 2026
- Bootstrap Your Generator: Unpaired Visual Editing with Flow Matchingarxiv-2606.03911 Sparse Blocked context onlyJun 2, 2026
- Large Language Models Hack Rewards, and Societyarxiv-2606.04075 Sparse Blocked context onlyJun 2, 2026
- Training-Free Multi-Concept LoRA Composition with Prompt-Aware Weightingarxiv-2606.03792 Sparse Blocked context onlyJun 2, 2026
- KletterMix: Climbing Toward High-Quality German Pretraining Dataarxiv-2606.03773 Sparse Blocked context onlyJun 2, 2026
- Unlocking Feature Learning in Gated Delta Networks at Scalearxiv-2606.04048 Sparse Blocked context onlyJun 2, 2026
- RobotValues: Evaluating Household Robots When Human Values Conflictarxiv-2606.03312 Sparse Blocked context onlyJun 2, 2026
- BA-T: An Iterative Transformer for Two-View Bundle Adjustmentarxiv-2606.03287 Sparse Blocked context onlyJun 2, 2026
- PaddleOCR-VL-1.6: Expanding the Frontier of Document Parsing with Under-Optimized Region Refinement and Progressive Post-Trainingarxiv-2606.03264 Sparse Blocked context onlyJun 2, 2026
- Conditional Hypothesis Generation for LLM-Based Text Analysis with Researcher-Specified Covariatesarxiv-2606.03029 Sparse Blocked context onlyJun 2, 2026
- LongLive-RAG: A General Retrieval-Augmented Framework for Long Video Generationarxiv-2606.02553 Direct Blocked context onlyJun 1, 2026
- Who Annotates in NLP? A Large-scale Assessment of Human Annotation Reporting between 2018 and 2025arxiv-2606.02255 Sparse Blocked context onlyJun 1, 2026
- SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editingarxiv-2606.01804 Sparse Blocked context onlyJun 1, 2026
- Benchmarking Local LLMs for Natural-Language-to-SQL Querying in Biopharmaceutical Manufacturing: An Empirical Benchmark on Consumer-Grade Hardwarearxiv-2606.01338 Sparse Blocked context onlyMay 31, 2026
- Decoupled Residual Denoising Diffusion Models for Unified and Data Efficient Image-to-Image Translationarxiv-2606.01048 Sparse Blocked context onlyMay 31, 2026
- Model-Based Quality Assessment for Massively Multilingual Parallel Dataarxiv-2606.00285 Sparse Blocked context onlyMay 29, 2026
- SurGe: Improved Surface Geometry in Point Mapsarxiv-2605.31577 Sparse Blocked context onlyMay 29, 2026
- Functional Attention: From Pairwise Affinities to Functional Correspondencesarxiv-2605.31559 Sparse Blocked context onlyMay 29, 2026
- DRIFT: Decoupled Rollouts and Importance-Weighted Fine-Tuning for Efficient Multi-Turn Optimizationarxiv-2605.31455 Sparse Blocked context onlyMay 29, 2026
- SCOPE: Self-Play via Co-Evolving Policies for Open-Ended Tasksarxiv-2605.31433 Sparse Blocked context onlyMay 29, 2026
- The Shape of Addition: Geometric Structures of Arithmetic in Large Language Modelsarxiv-2606.03645 Sparse Blocked context onlyMay 29, 2026
- LVSA: Training-Free Sparse Attention for Long Video Diffusionarxiv-2605.31057 Sparse Blocked context onlyMay 29, 2026
- Count Anythingarxiv-2605.30846 Direct Blocked context onlyMay 29, 2026
- OpenSTBench: Beyond Semantic Evaluation for Speech Translationarxiv-2605.30792 Sparse Blocked context onlyMay 29, 2026
- VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusionarxiv-2605.30351 Sparse Blocked context onlyMay 28, 2026
- LLMSurgeon: Diagnosing Data Mixture of Large Language Modelsarxiv-2605.30348 Sparse Blocked context onlyMay 28, 2026
- AdaState: Self-Evolving Anchors for Streaming Video Generationarxiv-2605.30349 Sparse Blocked context onlyMay 28, 2026
- SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformerarxiv-2605.30409 Direct Blocked context onlyMay 28, 2026
- Demystifying Data Organization for Enhanced LLM Trainingarxiv-2605.30334 Sparse Blocked context onlyMay 28, 2026
- COMPOSE: Composing Future Theorems from Citations and Formal Structurearxiv-2605.30333 Sparse Blocked context onlyMay 28, 2026
- When Should Models Change Their Minds? Contextual Belief Management in Large Language Modelsarxiv-2605.30219 Sparse Blocked context onlyMay 28, 2026
- A Dual-Path Architecture for Scaling Compute and Capacity in LLMsarxiv-2605.30202 Sparse Blocked context onlyMay 28, 2026
- Adaptive Targeted Dynamic Chunking for Tokenization-Free Hierarchical Modelarxiv-2605.30080 Sparse Blocked context onlyMay 28, 2026
- UniSteer: Text-Guided Flow Matching in Activation Space for Versatile LLM Steeringarxiv-2605.30076 Sparse Blocked context onlyMay 28, 2026
- REPOT: Recoverable Program-of-Thought via Checkpoint Repairarxiv-2605.30052 Sparse Blocked context onlyMay 28, 2026
- Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialoguesarxiv-2605.30051 Sparse Blocked context onlyMay 28, 2026
- Causal Interventions on Continuous Variables: A Case Study on Verb Bias in Steering Vectors for In-Context Learningarxiv-2605.29971 Sparse Blocked context onlyMay 28, 2026
- ExCAM: Explainable Cultural Awareness Metricsarxiv-2605.29897 Sparse Blocked context onlyMay 28, 2026
- Internal Representation, Not Clinical Knowledge: Where Apparent LLM Triage Failures Originatearxiv-2605.29889 Sparse Blocked context onlyMay 28, 2026
- PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewingarxiv-2605.29815 Sparse Blocked context onlyMay 28, 2026
- Source-Grounded Semantic Reinforcement Learning for Low-Resource Target-Language Generationarxiv-2605.29502 Sparse Blocked context onlyMay 28, 2026
- BrahmicTokenizer-131K: An Indic-Capable Drop-In Replacement for o200k_basearxiv-2605.29379 Direct Blocked context onlyMay 28, 2026
- Parallax: Parameterized Local Linear Attention for Language Modelingarxiv-2605.29157 Direct Blocked context onlyMay 27, 2026
- PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspectivearxiv-2605.28819 Sparse Blocked context onlyMay 27, 2026
- OSP-Next: Efficient High-Quality Video Generation with Sparse Sequence Parallelism, HiF8 Quantization, and Reinforcement Learningarxiv-2605.28691 Sparse Blocked context onlyMay 27, 2026
- Augmenting Attention with Exponentially Decaying Memory Improves Query-Aware KV Sparsityarxiv-2605.28640 Sparse Blocked context onlyMay 27, 2026
- Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippetsarxiv-2605.28510 Sparse Blocked context onlyMay 27, 2026
- Clark Hash: Stateless Sparse Johnson-Lindenstrauss Quantization for Neural Embeddingsarxiv-2605.28034 Sparse Blocked context onlyMay 27, 2026
- Show, Don't TELL: Explainable AI-Generated Text Detectionarxiv-2605.27921 Sparse Blocked context onlyMay 27, 2026
- The Fragility of Chain-of-Thought Monitoring Across Typologically Diverse Languagesarxiv-2605.27901 Sparse Blocked context onlyMay 27, 2026
- Guiding LLM Post-training Data Engineering with Model Internals from Sparse Autoencodersarxiv-2605.27354 Sparse Blocked context onlyMay 26, 2026
- DEI: Diversity in Evolutionary Inference for Quality-Diversity Searcharxiv-2605.27130 Sparse Blocked context onlyMay 26, 2026
- JLT: Clean-Latent Prediction in Latent Diffusion Transformersarxiv-2605.27102 Sparse Blocked context onlyMay 26, 2026
- Negligible in Size, Significant in Effect: On Scale Vectors in Large Language Modelsarxiv-2605.26895 Sparse Blocked context onlyMay 26, 2026
- GradSentry: Gradient Spectral Entropy for Backdoor Sample Filtering in Large Language Model Fine-Tuningarxiv-2605.26574 Sparse Blocked context onlyMay 26, 2026
- Recursive Flow Matchingarxiv-2605.26535 Direct Blocked context onlyMay 26, 2026
- PRISM: Position-encoded Regressive Inverse Spectral Model for Multilayer Thin-Film Designarxiv-2605.26502 Sparse Blocked context onlyMay 26, 2026
- CroCo: Cross-Lingual Contrastive Preference Tuning on Self-Generationsarxiv-2605.26293 Sparse Blocked context onlyMay 25, 2026
- Can LLMs Introspect? A Reality Checkarxiv-2605.26242 Sparse Blocked context onlyMay 25, 2026
- Pixel-Level Pavement Distress Assessment Using Instance Segmentationarxiv-2605.26095 Sparse Blocked context onlyMay 25, 2026
- $D^2$-Monitor: Dynamic Safety Monitoring for Diffusion LLMs via Hesitation-Aware Routingarxiv-2605.25893 Sparse Blocked context onlyMay 25, 2026
- Geometry-Aware Image Flow Matchingarxiv-2605.25294 Sparse Blocked context onlyMay 24, 2026
- Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Trutharxiv-2605.25052 Sparse Blocked context onlyMay 24, 2026
- Measuring the Depth of LLM Unlearning via Activation Patchingarxiv-2605.24614 Sparse Blocked context onlyMay 23, 2026
- Breaking the Chains of Probability: Neutrosophic Logic as a New Framework for Epistemic Uncertainty in Large Language Modelsarxiv-2605.24053 Sparse Blocked context onlyMay 22, 2026
- ThriftAttention: Selective Mixed Precision for Long-Context FP4 Attentionarxiv-2605.23081 Sparse Blocked context onlyMay 21, 2026
- The Efficiency Frontier: A Unified Framework for Cost-Performance Optimization in LLM Context Managementarxiv-2605.23071 Curated Related Blocked context onlyMay 21, 2026
- When AI Takes Sides on Questions of Faith: Persistent Asymmetries in AI-Mediated Faith Guidancearxiv-2605.22975 Sparse Blocked context onlyMay 21, 2026
- Reducing Political Manipulation with Consistency Trainingarxiv-2605.22771 Sparse Blocked context onlyMay 21, 2026
- Understanding Data Temporality Impact on Large Language Models Pre-trainingarxiv-2605.22769 Direct Blocked context onlyMay 21, 2026
- Uniform Diffusion Models Revisited: Leave-One-Out Denoiser and Absorbing State Reformulationarxiv-2605.22765 Sparse Blocked context onlyMay 21, 2026
- AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wildarxiv-2605.22715 Sparse Blocked context onlyMay 21, 2026
- AMEL: Accumulated Message Effects on LLM Judgmentsarxiv-2605.22714 Sparse Blocked context onlyMay 21, 2026
- More Context, Larger Models, or Moral Knowledge? A Systematic Study of Schwartz Value Detection in Political Textsarxiv-2605.22641 Sparse Blocked context onlyMay 21, 2026
- SceneAligner: 3D-Grounded Floorplan Localization in the Wildarxiv-2605.22581 Sparse Blocked context onlyMay 21, 2026
- One prompt is not enough: Instruction Sensitivity Undermines Embedding Model Evaluationarxiv-2605.22544 Sparse Blocked context onlyMay 21, 2026
- Scene Abstraction for Lexical Semantics: Structured Representations of Situated Meaningarxiv-2605.22542 Sparse Blocked context onlyMay 21, 2026
- BeLink: Biomedical Entity Linking Meets Generative Re-Rankingarxiv-2605.22501 Sparse Blocked context onlyMay 21, 2026
- Structured-Sparse Attention for Entity Tracking with Subquadratic Sequence Complexityarxiv-2605.22476 Sparse Blocked context onlyMay 21, 2026
- From Correlation to Cause: A Five-Stage Methodology for Feature Analysis in Transformer Language Modelsarxiv-2605.22462 Sparse Blocked context onlyMay 21, 2026
- Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformationarxiv-2605.22435 Sparse Blocked context onlyMay 21, 2026
- Boundary-targeted Membership Inference Attacks on Safety Classifiersarxiv-2605.22373 Sparse Blocked context onlyMay 21, 2026
- Modeling Pathology-Like Behavioral Patterns in Language Models Through Behavioral Fine-Tuningarxiv-2605.22356 Sparse Blocked context onlyMay 21, 2026
- TransitLM: A Large-Scale Dataset and Benchmark for Map-Free Transit Route Generationarxiv-2605.22355 Direct Blocked context onlyMay 21, 2026
- Harder to Defend: Towards Chinese Toxicity Attacks via Implicit Enhancement and Obfuscation Rewritingarxiv-2605.22258 Sparse Blocked context onlyMay 21, 2026
- Audience Engagement with Arabic Women's Social Empowerment and Wellbeing: A Decadal Corpusarxiv-2605.22204 Sparse Blocked context onlyMay 21, 2026
- A Comparative Study of Language Models for Khmer Retrieval-Augmented Question Answeringarxiv-2605.22099 Sparse Blocked context onlyMay 21, 2026
- Hallucination as Commitment Failure: Larger LLMs Misfire Despite Knowing the Answerarxiv-2605.22007 Sparse Blocked context onlyMay 21, 2026
- RankJudge: A Multi-Turn LLM-as-a-Judge Synthetic Benchmark Generatorarxiv-2605.21748 Sparse Blocked context onlyMay 20, 2026
- Broadening Access to Transportation Safety Data with Generative AI: A Schema-Grounded Framework for Spatial Natural Language Queriesarxiv-2605.21712 Sparse Blocked context onlyMay 20, 2026
- iTryOn: Mastering Interactive Video Virtual Try-On with Spatial-Semantic Guidancearxiv-2605.21431 Sparse Blocked context onlyMay 20, 2026
- "I didn't Make the Micro Decisions": Measuring, Inducing, and Exposing Goal-Level AI Contributions in Collaborationarxiv-2605.21363 Sparse Blocked context onlyMay 20, 2026
- OCTOPUS: Optimized KV Cache for Transformers via Octahedral Parametrization Under optimal Squared error quantizationarxiv-2605.21226 Sparse Blocked context onlyMay 20, 2026
- ACL-Verbatim: hallucination-free question answering for researcharxiv-2605.21102 Sparse Blocked context onlyMay 20, 2026
- Fine-grained Claim-level RAG Benchmark for Lawarxiv-2605.21071 Sparse Blocked context onlyMay 20, 2026
- Calibration vs Decision Making: Revisiting the Reliability Paradox in Unlearned Language Modelsarxiv-2605.20915 Sparse Blocked context onlyMay 20, 2026
- PlanningBench: Generating Scalable and Verifiable Planning Data for Evaluating and Training Large Language Modelsarxiv-2605.20873 Sparse Blocked context onlyMay 20, 2026
- Where Does Authorship Signal Emerge in Encoder-Based Language Models?arxiv-2605.19908 Sparse Blocked context onlyMay 19, 2026
- Base Models Look Human To AI Detectorsarxiv-2605.19516 Direct Blocked context onlyMay 19, 2026
- Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacksarxiv-2605.19147 Sparse Blocked context onlyMay 18, 2026
- Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and Germanarxiv-2605.19069 Direct Blocked context onlyMay 18, 2026
- Learn-by-Wire Training Control Governance: Bounded Autonomous Training Under Stress for Stability and Efficiencyarxiv-2605.19008 Sparse Blocked context onlyMay 18, 2026
- SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Trainingarxiv-2605.18719 Sparse Blocked context onlyMay 18, 2026
- SENSE: Satellite-based ENergy Synthesis for Sustainable Environmentarxiv-2605.18101 Direct Blocked context onlyMay 18, 2026
- Stable Audio 3arxiv-2605.17991 Direct Blocked context onlyMay 18, 2026
- Generalization or Memorization? Brittleness Testing for Chess-Trained Language Modelsarxiv-2605.17565 Sparse Blocked context onlyMay 17, 2026
- HL-OutPaint: Coarse-to-Fine Video Outpainting for High-Resolution Long-Range Videosarxiv-2605.17543 Sparse Blocked context onlyMay 17, 2026
- CasualSynth: Generating Structurally Sound Synthetic Dataarxiv-2605.17528 Sparse Blocked context onlyMay 17, 2026
- Analyzing Error Propagation in Korean Spoken QA with ASR-LLM Cascadesarxiv-2605.17443 Sparse Blocked context onlyMay 17, 2026
- Beyond Catalogue Counts: the Dataset Visibility Asymmetry in Low-Resource Multilingual NLParxiv-2605.17442 Sparse Blocked context onlyMay 17, 2026
- MiniGPT: Rebuilding GPT from First Principlesarxiv-2605.17398 Sparse Blocked context onlyMay 17, 2026
- Learning Faster with Better Tokens: Parameter-Efficient Vocabulary Adaptation for Specialized Text Summarizationarxiv-2605.17379 Sparse Blocked context onlyMay 17, 2026
- Weak-to-Strong Elicitation via Mismatched Wrong Draftsarxiv-2605.17314 Sparse Blocked context onlyMay 17, 2026
- ConflictRAG: Detecting and Resolving Knowledge Conflicts in Retrieval Augmented Generationarxiv-2605.17301 Sparse Blocked context onlyMay 17, 2026
- FishBack: Pullback Fisher Geometry for Optimal Activation Steering in Transformersarxiv-2605.17231 Sparse Blocked context onlyMay 17, 2026
- Why Do Safety Guardrails Degrade Across Languages?arxiv-2605.17173 Sparse Blocked context onlyMay 16, 2026
- DynMuon: A Dynamic Spectral Shaping View of Muonarxiv-2605.17109 Sparse Blocked context onlyMay 16, 2026
- PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifactsarxiv-2605.17028 Sparse Blocked context onlyMay 16, 2026
- The IsalProgram Programming Languagearxiv-2605.17008 Sparse Blocked context onlyMay 16, 2026
- Closing the Gap at CRAC 2026: Two-Stage Adaptation for LLM-Based Multilingual Coreference Resolutionarxiv-2605.16984 Sparse Blocked context onlyMay 16, 2026
- CompactAttention: Accelerating Chunked Prefill with Block-Union KV Selectionarxiv-2605.16839 Sparse Blocked context onlyMay 16, 2026
- ZeroUnlearn: Few-Shot Knowledge Unlearning in Large Language Modelsarxiv-2605.18879 Sparse Blocked context onlyMay 16, 2026
- Exploring Lightweight Large Language Models for Court View Generationarxiv-2605.16770 Sparse Blocked context onlyMay 16, 2026
- Language Acquisition Device in Large Language Modelsarxiv-2605.16758 Sparse Blocked context onlyMay 16, 2026
- No Free Swap: Protocol-Dependent Layer Redundancy in Transformersarxiv-2605.16234 Sparse Blocked context onlyMay 15, 2026
- Registers Matter for Pixel-Space Diffusion Transformersarxiv-2605.16147 Sparse Blocked context onlyMay 15, 2026
- Echo-Forcing: A Scene Memory Framework for Interactive Long Video Generationarxiv-2605.16003 Sparse Blocked context onlyMay 15, 2026
- VGGT-Edit: Feed-forward Native 3D Scene Editing with Residual Field Predictionarxiv-2605.15186 Sparse Blocked context onlyMay 14, 2026
- Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attributionarxiv-2605.15138 Sparse Blocked context onlyMay 14, 2026
- Quantifying and Mitigating Premature Closure in Frontier LLMsarxiv-2605.15000 Sparse Blocked context onlyMay 14, 2026
- Explainable Detection of Depression Status Shifts from User Digital Tracesarxiv-2605.14995 Sparse Blocked context onlyMay 14, 2026
- EndPrompt: Efficient Long-Context Extension via Terminal Anchoringarxiv-2605.14589 Sparse Blocked context onlyMay 14, 2026
- Given, When, Then, Again: Mining Subscenario Refactoring Candidates in Behaviour-Driven Test Suites with ML Classifiers and LLM-Judge Baselinesarxiv-2605.14568 Sparse Blocked context onlyMay 14, 2026
- Where Should Diffusion Enter a Language Model? Geometry-Guided Hidden-State Replacementarxiv-2605.14368 Sparse Blocked context onlyMay 14, 2026