300 canonical paper links on this archive page.
- Projected Autoregression: Autoregressive Language Generation in Continuous State Spacearxiv-2601.04854 Sparse Blocked context onlyJan 8, 2026
- Interpreting Transformers Through Attention Head Interventionarxiv-2601.04398 Sparse Blocked context onlyJan 7, 2026
- RADAR: Retrieval-Augmented Detector with Adversarial Refinement for Robust Fake News Detectionarxiv-2601.03981 Sparse Blocked context onlyJan 7, 2026
- From Domains to Instances: Dual-Granularity Data Synthesis for LLM Unlearningarxiv-2601.04278 Sparse Blocked context onlyJan 7, 2026
- Compact Example-Based Explanations for Language Modelsarxiv-2601.03786 Sparse Blocked context onlyJan 7, 2026
- Embedding Retrofitting: Data Engineering for better RAGarxiv-2601.15298 Sparse Blocked context onlyJan 6, 2026
- The Invisible Hand of AI Libraries Shaping Open Source Projects and Communitiesarxiv-2601.01944 Sparse Blocked context onlyJan 5, 2026
- ARGUS: Adaptive Rotation-Invariant Geometric Unsupervised Systemarxiv-2601.01297 Sparse Blocked context onlyJan 3, 2026
- Improving Variational Autoencoder using Random Fourier Transformation: An Aviation Safety Anomaly Detection Case-Studyarxiv-2601.01016 Sparse Blocked context onlyJan 3, 2026
- ADOPT: Adaptive Dependency-Guided Joint Prompt Optimization for Multi-Step LLM Pipelinesarxiv-2512.24933 Sparse Blocked context onlyDec 31, 2025
- Intrinsic-Metric Physics-Informed Neural Networks (IM-PINN) for Reaction-Diffusion Dynamics on Complex Riemannian Manifoldsarxiv-2601.00834 Sparse Blocked context onlyDec 26, 2025
- Measuring all the noises of LLM Evalsarxiv-2512.21326 Sparse Blocked context onlyDec 24, 2025
- Semantic Refinement with LLMs for Graph Representationsarxiv-2512.21106 Sparse Blocked context onlyDec 24, 2025
- TS-Arena -- A Live Forecast Pre-Registration Platformarxiv-2512.20761 Sparse Blocked context onlyDec 23, 2025
- On the Existence and Behavior of Secondary Attention Sinksarxiv-2512.22213 Sparse Blocked context onlyDec 22, 2025
- NASTaR: NovaSAR Automated Ship Target Recognition Datasetarxiv-2512.18503 Sparse Blocked context onlyDec 20, 2025
- Exploration vs. Fixation: Scaffolding Divergent and Convergent Thinking for Human-AI Co-Creation with Generative Modelsarxiv-2512.18388 Sparse Blocked context onlyDec 20, 2025
- Affect, Body, Cognition, Demographics, and Emotion: The ABCDE of Text Features for Computational Affective Sciencearxiv-2512.17752 Sparse Blocked context onlyDec 19, 2025
- SFBD-OMNI: Bridge models for lossy measurement restoration with limited clean samplesarxiv-2512.17051 Sparse Blocked context onlyDec 18, 2025
- Learning continuous state of charge dependent thermal decomposition kinetics for Li-ion cathodes using Kolmogorov-Arnold Chemical Reaction Neural Networks (KA-CRNNs)arxiv-2512.15628 Sparse Blocked context onlyDec 17, 2025
- Physics-driven human-like working memory outperforms digital networks in dynamic visionarxiv-2512.15829 Sparse Blocked context onlyDec 17, 2025
- A Multicenter Benchmark of Multiple Instance Learning Models for Lymphoma Subtyping from HE-stained Whole Slide Imagesarxiv-2512.14640 Sparse Blocked context onlyDec 16, 2025
- Evaluating Music Context Preservation: A Multi-facet Framework for Music Editing Systemsarxiv-2512.14629 Sparse Blocked context onlyDec 16, 2025
- Global Sensitivity Analysis for Engineering Design Based on Individual Conditional Expectationsarxiv-2512.11946 Sparse Blocked context onlyDec 12, 2025
- Maximum Risk Minimization with Random Forestsarxiv-2512.10445 Sparse Blocked context onlyDec 11, 2025
- Near--Real-Time Conflict-Related Fire Detection in Sudan Using Unsupervised Deep Learningarxiv-2512.07925 Sparse Blocked context onlyDec 8, 2025
- GUMBridge: a Corpus for Varieties of Bridging Anaphoraarxiv-2512.07134 Sparse Blocked context onlyDec 8, 2025
- SETUP: Sentence-level English-To-Uniform Meaning Representation Parserarxiv-2512.07068 Sparse Blocked context onlyDec 8, 2025
- Arc Gradient Descent: A Geometrically Motivated Gradient Descent-based Optimiser with Phase-Aware, User-Controlled Step Dynamics (proof-of-concept)arxiv-2512.06737 Sparse Blocked context onlyDec 7, 2025
- SpaceControl: Introducing Test-Time Spatial Control to 3D Generative Modelingarxiv-2512.05343 Sparse Blocked context onlyDec 5, 2025
- BERnaT: Basque Encoders for Representing Natural Textual Diversityarxiv-2512.03903 Sparse Blocked context onlyDec 3, 2025
- Reconstructing KV Caches with Cross-layer Fusion For Enhanced Transformersarxiv-2512.03870 Sparse Blocked context onlyDec 3, 2025
- AITutor-EvalKit: Exploring the Capabilities of AI Tutorsarxiv-2512.03688 Sparse Blocked context onlyDec 3, 2025
- Constant-Time Motion Planning with Manipulation Behaviorsarxiv-2512.00939 Sparse Blocked context onlyNov 30, 2025
- ByteStorm: a multi-step data-driven approach for Tropical Cyclones detection and trackingarxiv-2512.07885 Sparse Blocked context onlyNov 28, 2025
- Real-Time Long Horizon Air Quality Forecasting via Group-Relative Policy Optimizationarxiv-2511.22169 Sparse Blocked context onlyNov 27, 2025
- Structured Prompts Improve Evaluation of Language Modelsarxiv-2511.20836 Sparse Blocked context onlyNov 25, 2025
- Human-computer interactions predict mental healtharxiv-2511.20179 Sparse Blocked context onlyNov 25, 2025
- NSTR: Neural Spectral Transport Representation for Space-Varying Frequency Fieldsarxiv-2511.18384 Sparse Blocked context onlyNov 23, 2025
- A Unified Stability Analysis of SAM vs SGD: Role of Data Coherence and Emergence of Simplicity Biasarxiv-2511.17378 Sparse Blocked context onlyNov 21, 2025
- DeepCoT: Deep Continual Transformers for Real-Time Inference on Data Streamsarxiv-2511.17693 Sparse Blocked context onlyNov 21, 2025
- STREAM-VAE: Dual-Path Routing for Slow and Fast Dynamics in Vehicle Telemetry Anomaly Detectionarxiv-2511.15339 Sparse Blocked context onlyNov 19, 2025
- Cheating Stereo Matching in Full-scale: Physical Adversarial Attack against Binocular Depth Estimation in Autonomous Drivingarxiv-2511.14386 Sparse Blocked context onlyNov 18, 2025
- D-GAP: Improving Out-of-Domain Robustness via Dataset-Agnostic and Gradient-Guided Augmentation in Frequency and Pixel Spacesarxiv-2511.11286 Sparse Blocked context onlyNov 14, 2025
- What We Don't C: Manifold Disentanglement for Structured Discoveryarxiv-2511.09433 Sparse Blocked context onlyNov 12, 2025
- Does Scientific Writing Converge to U.S. English? Evidence from Generative AI-Assisted Publicationsarxiv-2511.11687 Sparse Blocked context onlyNov 12, 2025
- A robust methodology for long-term sustainability evaluation of Machine Learning modelsarxiv-2511.08120 Sparse Blocked context onlyNov 11, 2025
- Categorical Emotions or Appraisals - Which Emotion Model Explains Argument Convincingness Better?arxiv-2511.07162 Sparse Blocked context onlyNov 10, 2025
- How AI Fails: An Interactive Pedagogical Tool for Demonstrating Dialectal Bias in Automated Toxicity Modelsarxiv-2511.06676 Sparse Blocked context onlyNov 10, 2025
- Sub-exponential Growth Dynamics in Complex Systems: A Piecewise Power-Law Model for the Diffusion of New Words and Namesarxiv-2511.04106 Sparse Blocked context onlyNov 6, 2025
- GRDD+: An Extended Greek Dialectal Dataset with Cross-Architecture Fine-tuning Evaluationarxiv-2511.03772 Sparse Blocked context onlyNov 5, 2025
- Structured Matrix Scaling for Multi-Class Calibrationarxiv-2511.03685 Sparse Blocked context onlyNov 5, 2025
- Complete asymptotic type-token relationship for growing complex systems with inverse power-law count rankingsarxiv-2511.02069 Sparse Blocked context onlyNov 3, 2025
- A Proof of Learning Rate Transfer under $μ$Parxiv-2511.01734 Sparse Blocked context onlyNov 3, 2025
- Addressing Longstanding Challenges in Cognitive Science with Language Modelsarxiv-2511.00206 Sparse Blocked context onlyOct 31, 2025
- Can SAEs reveal and mitigate racial biases of LLMs in healthcare?arxiv-2511.00177 Sparse Blocked context onlyOct 31, 2025
- Temporal Sparse Autoencoders: Leveraging the Sequential Nature of Language for Interpretabilityarxiv-2511.05541 Sparse Blocked context onlyOct 30, 2025
- LLMs Process Lists With General Filter Headsarxiv-2510.26784 Sparse Blocked context onlyOct 30, 2025
- GraphKeeper: Graph Domain-Incremental Learning via Knowledge Disentanglement and Preservationarxiv-2511.00097 Sparse Blocked context onlyOct 30, 2025
- Dark & Stormy: Modeling Humor in Sentences from the Bulwer-Lytton Fiction Contestarxiv-2510.24538 Sparse Blocked context onlyOct 28, 2025
- An Information-Theoretic Analysis of OOD Generalization in Meta-Reinforcement Learningarxiv-2510.23448 Sparse Blocked context onlyOct 27, 2025
- A Diagnostic Benchmark for Sweden-Related Factual Knowledgearxiv-2510.21360 Sparse Blocked context onlyOct 24, 2025
- Designing and Evaluating Chain-of-Hints for Scientific Question Answeringarxiv-2510.21087 Sparse Blocked context onlyOct 24, 2025
- Transferable Graph Learning for Transmission Congestion Management via Busbar Splittingarxiv-2510.20591 Sparse Blocked context onlyOct 23, 2025
- Modality Matching Matters: Calibrating Language Distances for Cross-Lingual Transfer in URIEL+arxiv-2510.19217 Sparse Blocked context onlyOct 22, 2025
- MoMaGen: Generating Demonstrations under Soft and Hard Constraints for Multi-Step Bimanual Mobile Manipulationarxiv-2510.18316 Sparse Blocked context onlyOct 21, 2025
- Latent-Augmented Discrete Diffusion Modelsarxiv-2510.18114 Sparse Blocked context onlyOct 20, 2025
- DELULU: Discriminative Embedding Learning Using Latent Units for Speaker-Aware Self-Trained Speech Foundational Modelarxiv-2510.17662 Sparse Blocked context onlyOct 20, 2025
- Towards a Practical Understanding of Lagrangian Methods in Safe Reinforcement Learningarxiv-2510.17564 Sparse Blocked context onlyOct 20, 2025
- OffSim: Offline Simulator for Model-based Offline Inverse Reinforcement Learningarxiv-2510.15495 Sparse Blocked context onlyOct 17, 2025
- Learning to Answer from Correct Demonstrationsarxiv-2510.15464 Sparse Blocked context onlyOct 17, 2025
- MNO: Multiscale Neural Operator for 3D Computational Fluid Dynamicsarxiv-2510.16071 Sparse Blocked context onlyOct 17, 2025
- AI-BAAM: AI-Driven Bank Statement Analytics as Alternative Data for Malaysian MSME Credit Scoringarxiv-2510.16066 Sparse Blocked context onlyOct 17, 2025
- CBF-RL: Safety Filtering Reinforcement Learning in Training with Control Barrier Functionsarxiv-2510.14959 Sparse Blocked context onlyOct 16, 2025
- Circuit Insights: Towards Interpretability Beyond Activationsarxiv-2510.14936 Sparse Blocked context onlyOct 16, 2025
- Detecting Early and Implicit Suicidal Ideation via Longitudinal and Information Environment Signals on Social Mediaarxiv-2510.14889 Sparse Blocked context onlyOct 16, 2025
- LUMI: Unsupervised Intent Clustering with Multiple Pseudo-Labelsarxiv-2510.14640 Sparse Blocked context onlyOct 16, 2025
- Readers Prefer Outputs of AI Trained on Copyrighted Books over Expert Human Writersarxiv-2510.13939 Sparse Blocked context onlyOct 15, 2025
- From Prompts to Packets: A View from the Network on ChatGPT, Copilot, and Geminiarxiv-2510.11269 Sparse Blocked context onlyOct 13, 2025
- FactAppeal: Identifying Epistemic Factual Appeals in News Mediaarxiv-2510.10627 Sparse Blocked context onlyOct 12, 2025
- Personalized Motion Guidance Framework for Athlete-Centric Coachingarxiv-2510.10496 Sparse Blocked context onlyOct 12, 2025
- CQA-Eval: Designing Reliable Evaluations of Multi-paragraph Clinical QA under Resource Constraintsarxiv-2510.10415 Sparse Blocked context onlyOct 12, 2025
- Mapping Semantic & Syntactic Relationships with Geometric Rotationarxiv-2510.09790 Sparse Blocked context onlyOct 10, 2025
- Chlorophyll-a Mapping and Prediction in the Mar Menor Lagoon Using C2RCC-Processed Sentinel 2 Imageryarxiv-2510.09736 Sparse Blocked context onlyOct 10, 2025
- Do LLMs Really Know What They Don't Know? Internal States Mainly Reflect Knowledge Recall Rather Than Truthfulnessarxiv-2510.09033 Sparse Blocked context onlyOct 10, 2025
- Counterfactual Identifiability via Dynamic Optimal Transportarxiv-2510.08294 Sparse Blocked context onlyOct 9, 2025
- Lossless Vocabulary Reduction for Auto-Regressive Language Modelsarxiv-2510.08102 Sparse Blocked context onlyOct 9, 2025
- Everything is Plausible: Investigating the Impact of LLM Rationales on Human Notions of Plausibilityarxiv-2510.08091 Sparse Blocked context onlyOct 9, 2025
- A Systematic Evaluation of Self-Supervised Learning for Label-Efficient Sleep Staging with Wearable EEGarxiv-2510.07960 Sparse Blocked context onlyOct 9, 2025
- Standard-to-Dialect Transfer Trends Differ across Text and Speech: A Case Study on Intent and Topic Classification in German Dialectsarxiv-2510.07890 Sparse Blocked context onlyOct 9, 2025
- PATCH: Mitigating PII Leakage in Language Models with Privacy-Aware Targeted Circuit PatcHingarxiv-2510.07452 Sparse Blocked context onlyOct 8, 2025
- Open ASR Leaderboard: Towards Reproducible and Transparent Multilingual and Long-Form Speech Recognition Evaluationarxiv-2510.06961 Sparse Blocked context onlyOct 8, 2025
- Robustness assessment of large audio language models in multiple-choice evaluationarxiv-2510.04584 Sparse Blocked context onlyOct 6, 2025
- Psychological Steering in LLMs: An Evaluation of Effectiveness and Trustworthinessarxiv-2510.04484 Sparse Blocked context onlyOct 6, 2025
- Finding Diamonds in Conversation Haystacks: A Benchmark for Conversational Data Retrievalarxiv-2510.02938 Sparse Blocked context onlyOct 3, 2025
- Dynamic Stress Detection: A Study of Temporal Progression Modelling of Stress in Speecharxiv-2510.08586 Sparse Blocked context onlyOct 2, 2025
- On Discovering Algorithms for Adversarial Imitation Learningarxiv-2510.00922 Sparse Blocked context onlyOct 1, 2025
- Removing Noise, not Finding Gold: Quality Filtering for Large-Scale Pretrainingarxiv-2510.00866 Sparse Blocked context onlyOct 1, 2025
- Family Matters: Language Transfer and Merging for Adapting Small LLMs to Faroesearxiv-2510.00810 Sparse Blocked context onlyOct 1, 2025
- On Deepfake Voice Detection -- It's All in the Presentationarxiv-2509.26471 Sparse Blocked context onlySep 30, 2025
- Vector sketch animation generation with differentiable motion trajectoriesarxiv-2509.25857 Sparse Blocked context onlySep 30, 2025
- Polychromic Objectives for Reinforcement Learningarxiv-2509.25424 Sparse Blocked context onlySep 29, 2025
- Predicting Training Re-evaluation Curves Enables Effective Data Curriculums for LLMsarxiv-2509.25380 Sparse Blocked context onlySep 29, 2025
- Collaboration of Fusion and Independence: Hypercomplex-driven Robust Multi-Modal Knowledge Graph Completionarxiv-2509.23714 Sparse Blocked context onlySep 28, 2025
- Characteristic Root Analysis and Regularization for Linear Time Series Forecastingarxiv-2509.23597 Sparse Blocked context onlySep 28, 2025
- Robust Fine-Tuning from Non-Robust Pretrained Models: Mitigating Suboptimal Transfer With Epsilon-Schedulingarxiv-2509.23325 Sparse Blocked context onlySep 27, 2025
- Induction Signatures Are Not Enough: A Matched-Compute Study of Load-Bearing Structure in In-Context Learningarxiv-2509.22947 Sparse Blocked context onlySep 26, 2025
- Compute-Optimal Quantization-Aware Trainingarxiv-2509.22935 Sparse Blocked context onlySep 26, 2025
- From Formal Language Theory to Statistical Learning: Finite Observability of Subregular Languagesarxiv-2509.22598 Sparse Blocked context onlySep 26, 2025
- From Parameters to Behaviors: Unsupervised Compression of the Policy Spacearxiv-2509.22566 Sparse Blocked context onlySep 26, 2025
- Bridging Kolmogorov Complexity and Deep Learning: Asymptotically Optimal Description Length Objectives for Transformersarxiv-2509.22445 Sparse Blocked context onlySep 26, 2025
- Leveraging Wireless Sensor Networks for Real-Time Monitoring and Control of Industrial Environmentsarxiv-2510.13820 Sparse Blocked context onlySep 26, 2025
- "I think this is fair": Uncovering the Complexities of Stakeholder Decision-Making in AI Fairness Assessmentarxiv-2509.17956 Sparse Blocked context onlySep 22, 2025
- SLAyiNG: A Diverse and Community-validated Dataset of Queer Slangarxiv-2509.17449 Sparse Blocked context onlySep 22, 2025
- KANO: Kolmogorov-Arnold Neural Operatorarxiv-2509.16825 Sparse Blocked context onlySep 20, 2025
- Quantifying Genuine Awareness in Hallucination Prediction Beyond Question-Side Shortcutsarxiv-2509.15339 Sparse Blocked context onlySep 18, 2025
- Frame Sampling Strategies Matter: A Benchmark for small vision language modelsarxiv-2509.14769 Sparse Blocked context onlySep 18, 2025
- Masked Diffusion Models as Energy Minimizationarxiv-2509.13866 Sparse Blocked context onlySep 17, 2025
- Similarity-Distance-Magnitude Activationsarxiv-2509.12760 Sparse Blocked context onlySep 16, 2025
- Collaborative Document Editing with Multiple Users and AI Agentsarxiv-2509.11826 Sparse Blocked context onlySep 15, 2025
- CogniAlign: Survivability-Grounded Multi-Agent Moral Reasoning for Safe and Transparent AIarxiv-2509.13356 Sparse Blocked context onlySep 14, 2025
- Index-Preserving Lightweight Token Pruning for Efficient Document Understanding in Vision-Language Modelsarxiv-2509.06415 Sparse Blocked context onlySep 8, 2025
- CausalARC: Abstract Reasoning with Causal World Modelsarxiv-2509.03636 Sparse Blocked context onlySep 3, 2025
- BioBlue: Systematic runaway-optimiser-like LLM failure modes on biologically and economically aligned AI safety benchmarks for LLMs with simplified observation formatarxiv-2509.02655 Sparse Blocked context onlySep 2, 2025
- CMRAG: Co-modality-based visual document retrieval and question answeringarxiv-2509.02123 Sparse Blocked context onlySep 2, 2025
- L-MARS: Legal Multi-Agent Workflow with Orchestrated Reasoning and Agentic Searcharxiv-2509.00761 Sparse Blocked context onlyAug 31, 2025
- Your AI Bosses Are Still Prejudiced: The Emergence of Stereotypes in LLM-Based Multi-Agent Systemsarxiv-2508.19919 Sparse Blocked context onlyAug 27, 2025
- Language and Experience: A Computational Model of Social Learning in Complex Tasksarxiv-2509.00074 Sparse Blocked context onlyAug 26, 2025
- Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoningarxiv-2508.19113 Sparse Blocked context onlyAug 26, 2025
- FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflectionarxiv-2508.18684 Sparse Blocked context onlyAug 26, 2025
- Why Synthetic Isn't Real Yet: A Diagnostic Framework for Contact Center Dialogue Generationarxiv-2508.18210 Sparse Blocked context onlyAug 25, 2025
- Vevo2: A Unified and Controllable Framework for Speech and Singing Voice Generationarxiv-2508.16332 Sparse Blocked context onlyAug 22, 2025
- Mapping the Course for Prompt-based Structured Predictionarxiv-2508.15090 Sparse Blocked context onlyAug 20, 2025
- Tokens with Meaning: A Hybrid Tokenization Approach for Turkisharxiv-2508.14292 Sparse Blocked context onlyAug 19, 2025
- The Yokai Learning Environment: Tracking Beliefs Over Space and Timearxiv-2508.12480 Sparse Blocked context onlyAug 17, 2025
- AVEX: What Matters for Animal Vocalization Encodingarxiv-2508.11845 Sparse Blocked context onlyAug 15, 2025
- The GPT-4o Shock Emotional Attachment to AI Models and Its Impact on Regulatory Acceptance: A Cross-Cultural Analysis of the Immediate Transition from GPT-4o to GPT-5arxiv-2508.16624 Sparse Blocked context onlyAug 14, 2025
- SQL-Exchange: Transforming SQL Queries Across Domainsarxiv-2508.07087 Sparse Blocked context onlyAug 9, 2025
- SEVADE: Self-Evolving Multi-Agent Analysis with Decoupled Evaluation for Hallucination-Resistant Irony Detectionarxiv-2508.06803 Curated Related Blocked context onlyAug 9, 2025
- Mixed-Initiative Dialog for Human-Robot Collaborative Manipulationarxiv-2508.05535 Sparse Blocked context onlyAug 7, 2025
- ReCode: Reinforcing Code Generation with Reasoning-Process Rewardsarxiv-2508.05170 Sparse Blocked context onlyAug 7, 2025
- LayerT2V: A Unified Multi-Layer Video Generation Frameworkarxiv-2508.04228 Sparse Blocked context onlyAug 6, 2025
- STEMTOX: From Collaborative Tags to Fine-Grained Toxic Meme Detection via Entropy-Guided Multi-Task Learningarxiv-2508.04166 Sparse Blocked context onlyAug 6, 2025
- SpeechRole: A Large-Scale Dataset and Benchmark for Evaluating Speech Role-Playing Agentsarxiv-2508.02013 Sparse Blocked context onlyAug 4, 2025
- RoboMemory: A Brain-inspired Multi-memory Agentic Framework for Interactive Environmental Learning in Physical Embodied Systemsarxiv-2508.01415 Sparse Blocked context onlyAug 2, 2025
- WebDS: An End-to-End Benchmark for Web-based Data Sciencearxiv-2508.01222 Sparse Blocked context onlyAug 2, 2025
- NeuralOS: Towards Simulating Operating Systems via Neural Generative Modelsarxiv-2507.08800 Sparse Blocked context onlyJul 11, 2025
- Traceable Evidence Enhanced Visual Grounded Reasoning: Evaluation and Methodologyarxiv-2507.07999 Sparse Blocked context onlyJul 10, 2025
- From Fragments to Facts: A Curriculum-Driven DPO Approach for Generating Hindi News Veracity Explanationsarxiv-2507.05179 Sparse Blocked context onlyJul 7, 2025
- Critique of World Modelarxiv-2507.05169 Sparse Blocked context onlyJul 7, 2025
- XISM: an eXploratory and Interactive Graph Tool to Visualize and Evaluate Semantic Map Modelsarxiv-2507.04070 Sparse Blocked context onlyJul 5, 2025
- Beyond cognacyarxiv-2507.03005 Sparse Blocked context onlyJul 2, 2025
- Skywork-Reward-V2: Scaling Preference Data Curation via Human-AI Synergyarxiv-2507.01352 Sparse Blocked context onlyJul 2, 2025
- Cognitive models can reveal interpretable value trade-offs in language modelsarxiv-2506.20666 Sparse Blocked context onlyJun 25, 2025
- Context Biasing for Pronunciation-Orthography Mismatch in Automatic Speech Recognitionarxiv-2506.18703 Sparse Blocked context onlyJun 23, 2025
- A Simple "Motivation" Can Enhance Reinforcement Finetuning of Large Reasoning Modelsarxiv-2506.18485 Sparse Blocked context onlyJun 23, 2025
- Improving Black-Box Generative Attacks via Generator Semantic Consistencyarxiv-2506.18248 Sparse Blocked context onlyJun 23, 2025
- Towards AI Search Paradigmarxiv-2506.17188 Sparse Blocked context onlyJun 20, 2025
- Multimodal Fused Learning for Solving the Generalized Traveling Salesman Problem in Robotic Task Planningarxiv-2506.16931 Sparse Blocked context onlyJun 20, 2025
- Dynamic Reinsurance Treaty Bidding via Multi-Agent Reinforcement Learningarxiv-2506.13113 Sparse Blocked context onlyJun 16, 2025
- Recent Advances in Multi-Agent Human Trajectory Prediction: A Comprehensive Reviewarxiv-2506.14831 Sparse Blocked context onlyJun 13, 2025
- Learning The Minimum Action Distancearxiv-2506.09276 Sparse Blocked context onlyJun 10, 2025
- A Signal Contract for Online Language Grounding and Discovery in Decision-Makingarxiv-2506.07915 Sparse Blocked context onlyJun 9, 2025
- Generalized Incremental Learning under Concept Drift across Evolving Data Streamsarxiv-2506.05736 Sparse Blocked context onlyJun 6, 2025
- Toward Data Systems That Are Business Semantic Centric and AI Agents Assistedarxiv-2506.05520 Sparse Blocked context onlyJun 5, 2025
- EHR2Path: Scalable Modeling of Longitudinal Patient Pathways from Multimodal Electronic Health Recordsarxiv-2506.04831 Sparse Blocked context onlyJun 5, 2025
- Watermarking Degrades Alignment in Language Models: Analysis and Mitigationarxiv-2506.04462 Sparse Blocked context onlyJun 4, 2025
- CyclicReflex: Improving Reasoning Models via Cyclical Reflection Token Schedulingarxiv-2506.11077 Sparse Blocked context onlyJun 4, 2025
- OmniSpatial: Towards Comprehensive Spatial Reasoning Benchmark for Vision Language Modelsarxiv-2506.03135 Sparse Blocked context onlyJun 3, 2025
- Online Fair Division with Additional Informationarxiv-2505.24503 Sparse Blocked context onlyMay 30, 2025
- StressTest: Can YOUR Speech LM Handle the Stress?arxiv-2505.22765 Sparse Blocked context onlyMay 28, 2025
- How Does Alignment Enhance LLMs' Multilingual Capabilities? A Language Neurons Perspectivearxiv-2505.21505 Sparse Blocked context onlyMay 27, 2025
- When Slower Isn't Truer: Inverse Scaling Law of Truthfulness in Multimodal Reasoningarxiv-2505.20214 Sparse Blocked context onlyMay 26, 2025
- Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPOarxiv-2505.19770 Sparse Blocked context onlyMay 26, 2025
- Consistency-based Abductive Reasoning over Perceptual Errors of Multiple Pre-trained Models in Novel Environmentsarxiv-2505.19361 Sparse Blocked context onlyMay 25, 2025
- Evaluation Faking: Unveiling Observer Effects in Safety Evaluation of Frontier AI Systemsarxiv-2505.17815 Sparse Blocked context onlyMay 23, 2025
- Dynamic Token Reweighting for Robust Vision-Language Modelsarxiv-2505.17132 Sparse Blocked context onlyMay 22, 2025
- Efficient PRM Training Data Synthesis via Formal Verificationarxiv-2505.15960 Sparse Blocked context onlyMay 21, 2025
- ALIEN: Aligned Entropy Head for Improving Uncertainty Estimation of LLMsarxiv-2505.15443 Sparse Blocked context onlyMay 21, 2025
- Guided Policy Optimization under Partial Observabilityarxiv-2505.15418 Sparse Blocked context onlyMay 21, 2025
- A quantitative analysis of semantic information in deep representations of text and imagesarxiv-2505.17101 Sparse Blocked context onlyMay 21, 2025
- Language Models use Lookbacks to Track Beliefsarxiv-2505.14685 Sparse Blocked context onlyMay 20, 2025
- RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understandingarxiv-2505.14462 Sparse Blocked context onlyMay 20, 2025
- Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMsarxiv-2505.14226 Sparse Blocked context onlyMay 20, 2025
- Word length predicts word order: "Min-max"-ing drives language evolutionarxiv-2505.13913 Sparse Blocked context onlyMay 20, 2025
- Shorten After You're Right: Lazy Length Penalties for Reasoning RLarxiv-2505.12284 Sparse Blocked context onlyMay 18, 2025
- The Counting Power of Transformersarxiv-2505.11199 Sparse Blocked context onlyMay 16, 2025
- Why 1 + 1 < 1 in Visual Token Pruning: Beyond Naive Integration via Multi-Objective Balanced Coveringarxiv-2505.10118 Curated Related Blocked context onlyMay 15, 2025
- Multi-Domain Audio Question Answering Benchmark Toward Acoustic Content Reasoningarxiv-2505.07365 Sparse Blocked context onlyMay 12, 2025
- Mastering Multi-Drone Volleyball through Hierarchical Co-Self-Play Reinforcement Learningarxiv-2505.04317 Sparse Blocked context onlyMay 7, 2025
- Decoding Open-Ended Information Seeking Goals from Eye Movements in Readingarxiv-2505.02872 Sparse Blocked context onlyMay 4, 2025
- TopoPoint: Enhance Topology Reasoning via Endpoint Detection in Autonomous Drivingarxiv-2505.17771 Sparse Blocked context onlyMay 1, 2025
- Reshaping MOFs text mining with a dynamic multi-agents framework of large language modelarxiv-2504.18880 Sparse Blocked context onlyApr 26, 2025
- A closer look at how large language models trust humans: patterns and biasesarxiv-2504.15801 Sparse Blocked context onlyApr 22, 2025
- Leakage and Interpretability in Concept-Based Modelsarxiv-2504.14094 Sparse Blocked context onlyApr 18, 2025
- Scalable Multi-Task Learning through Spiking Neural Networks with Adaptive Task-Switching Policy for Intelligent Autonomous Agentsarxiv-2504.13541 Sparse Blocked context onlyApr 18, 2025
- Cost-of-Pass: An Economic Framework for Evaluating Language Modelsarxiv-2504.13359 Sparse Blocked context onlyApr 17, 2025
- Adaptive Insurance Reserving with CVaR-Constrained Reinforcement Learning under Macroeconomic Regimesarxiv-2504.09396 Sparse Blocked context onlyApr 13, 2025
- Med-R2: Perception and Reflection-driven Complex Reasoning for Medical Report Generationarxiv-2504.02885 Sparse Blocked context onlyApr 2, 2025
- Relevance Isn't All You Need: Scaling RAG Systems With Inference-Time Compute Via Multi-Criteria Rerankingarxiv-2504.07104 Sparse Blocked context onlyApr 1, 2025
- Planet as a Brain: Towards Internet of AgentSites based on AIOS Serverarxiv-2504.14411 Sparse Blocked context onlyApr 1, 2025
- Genius: A Generalizable and Purely Unsupervised Self-Training Framework For Advanced Reasoningarxiv-2504.08672 Sparse Blocked context onlyApr 1, 2025
- SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvementarxiv-2504.07934 Sparse Blocked context onlyApr 1, 2025
- When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?arxiv-2503.23137 Sparse Blocked context onlyMar 29, 2025
- More Bang for the Buck: Process Reward Modeling with Entropy-Driven Uncertaintyarxiv-2503.22233 Sparse Blocked context onlyMar 28, 2025
- Survey on Evaluation of LLM-based Agentsarxiv-2503.16416 Sparse Blocked context onlyMar 20, 2025
- Implicit Bias-Like Patterns in Reasoning Modelsarxiv-2503.11572 Sparse Blocked context onlyMar 14, 2025
- Unicorn: A Universal and Collaborative Reinforcement Learning Approach Towards Generalizable Network-Wide Traffic Signal Controlarxiv-2503.11488 Sparse Blocked context onlyMar 14, 2025
- Reasoning-Grounded Natural Language Explanations for Language Modelsarxiv-2503.11248 Sparse Blocked context onlyMar 14, 2025
- TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Toolsarxiv-2503.10970 Sparse Blocked context onlyMar 14, 2025
- VQEL: Enabling Self-Play in Emergent Language Games via Agent-Internal Vector Quantizationarxiv-2503.04940 Sparse Blocked context onlyMar 6, 2025
- Training-free Adjustable Polynomial Graph Filtering for Ultra-fast Multimodal Recommendationarxiv-2503.04406 Sparse Blocked context onlyMar 6, 2025
- LINGOLY-TOO: Disentangling Reasoning from Knowledge with Templatised Orthographic Obfuscationarxiv-2503.02972 Sparse Blocked context onlyMar 4, 2025
- Depth-Width tradeoffs in Algorithmic Reasoning of Graph Tasks with Transformersarxiv-2503.01805 Sparse Blocked context onlyMar 3, 2025
- Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcementarxiv-2503.06520 Sparse Blocked context onlyMar 1, 2025
- Cerebrum (AIOS SDK): A Platform for Agent Development, Deployment, Distribution, and Discoveryarxiv-2503.11444 Sparse Blocked context onlyMar 1, 2025
- Urban Emergency Rescue Based on Multi-Agent Collaborative Learning: Coordination Between Fire Engines and Traffic Lightsarxiv-2502.16131 Sparse Blocked context onlyFeb 22, 2025
- Moving Beyond Medical Exams: A Clinician-Annotated Fairness Dataset of Real-World Tasks and Ambiguity in Mental Healthcarearxiv-2502.16051 Sparse Blocked context onlyFeb 22, 2025
- Integrating Personality into Digital Humans: A Review of LLM-Driven Approaches for Virtual Realityarxiv-2503.16457 Sparse Blocked context onlyFeb 22, 2025
- From Restless to Contextual: A Thresholding Bandit Reformulation For Finite-horizon Improvementarxiv-2502.05145 Sparse Blocked context onlyFeb 7, 2025
- OpenSTARLab: Open Approach for Spatio-Temporal Agent Data Analysis in Soccerarxiv-2502.02785 Sparse Blocked context onlyFeb 5, 2025
- VolleyBots: A Testbed for Multi-Drone Volleyball Game Combining Motion Control and Strategic Playarxiv-2502.01932 Sparse Blocked context onlyFeb 4, 2025
- GRADIEND: Feature Learning within Neural Networks Exemplified through Biasesarxiv-2502.01406 Sparse Blocked context onlyFeb 3, 2025
- LIMO: Less is More for Reasoningarxiv-2502.03387 Sparse Blocked context onlyFeb 1, 2025
- ITBench: Evaluating AI Agents across Diverse Real-World IT Automation Tasksarxiv-2502.05352 Sparse Blocked context onlyFeb 1, 2025
- Dialogue is Better Than Monologue: Instructing Medical LLMs via Strategical Conversationsarxiv-2501.17860 Sparse Blocked context onlyJan 29, 2025
- Safe Reinforcement Learning for Real-World Engine Controlarxiv-2501.16613 Sparse Blocked context onlyJan 28, 2025
- CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigationarxiv-2501.16609 Sparse Blocked context onlyJan 28, 2025
- Object-Centric World Models from Few-Shot Annotations for Sample-Efficient Reinforcement Learningarxiv-2501.16443 Sparse Blocked context onlyJan 27, 2025
- Explainable Multimodal Depression Recognition in Clinical Interviews via PHQ-Aligned Symptom Summarizationarxiv-2501.16106 Sparse Blocked context onlyJan 27, 2025
- CoverM: Read alignment statistics for metagenomicsarxiv-2501.11217 Sparse Blocked context onlyJan 20, 2025
- Cosmos World Foundation Model Platform for Physical AIdoi-10.48550_arxiv.2501.03575 Sparse Blocked context onlyJan 7, 2025
- MapEval: A Map-Based Evaluation of Geo-Spatial Reasoning in Foundation Modelsdoi-10.48550_arxiv.2501.00316 Sparse Blocked context onlyDec 31, 2024
- Is Contrastive Distillation Enough for Learning Comprehensive 3D Representations?arxiv-2412.08973 Sparse Blocked context onlyDec 12, 2024
- Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Searcharxiv-2412.18319 Sparse Blocked context onlyDec 1, 2024
- TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generationarxiv-2412.03069 Sparse Blocked context onlyDec 1, 2024
- Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systemsarxiv-2412.09413 Sparse Blocked context onlyDec 1, 2024
- The Limits of Inference Scaling Through Resamplingarxiv-2411.17501 Sparse Blocked context onlyNov 26, 2024
- RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Roboticsarxiv-2411.16537 Sparse Blocked context onlyNov 25, 2024
- Phrase-Instance Alignment for Generalized Referring Segmentationarxiv-2411.15087 Sparse Blocked context onlyNov 22, 2024
- Personalized Help for Optimizing Low-Skilled Users' Strategyarxiv-2411.09109 Sparse Blocked context onlyNov 14, 2024
- Renaissance: Investigating the Pretraining of Vision-Language Encodersarxiv-2411.06657 Sparse Blocked context onlyNov 11, 2024
- Probing the limitations of multimodal language models for chemistry and materials researcharxiv-2411.16955 Sparse Blocked context onlyNov 1, 2024
- RealCQA-V2: A Diagnostic Benchmark for Structured Visual Entailment over Scientific Chartsarxiv-2410.22492 Sparse Blocked context onlyOct 29, 2024
- Distill Visual Chart Reasoning Ability from LLMs to MLLMsarxiv-2410.18798 Sparse Blocked context onlyOct 1, 2024
- AFlow: Automating Agentic Workflow Generationarxiv-2410.10762 Sparse Blocked context onlyOct 1, 2024
- GraphTeam: Facilitating Large Language Model-based Graph Analysis via Multi-Agent Collaborationarxiv-2410.18032 Sparse Blocked context onlyOct 1, 2024
- AutoDAN-Turbo: A Lifelong Agent for Strategy Self-Exploration to Jailbreak LLMsarxiv-2410.05295 Sparse Blocked context onlyOct 1, 2024
- R-CoT: Reverse Chain-of-Thought Problem Generation for Geometric Reasoning in Large Multimodal Modelsarxiv-2410.17885 Sparse Blocked context onlyOct 1, 2024
- Pantograph: A Machine-to-Machine Interaction Interface for Advanced Theorem Proving, High Level Reasoning, and Data Extraction in Lean 4arxiv-2410.16429 Sparse Blocked context onlyOct 1, 2024
- OS-ATLAS: A Foundation Action Model for Generalist GUI Agentsarxiv-2410.23218 Sparse Blocked context onlyOct 1, 2024
- Gödel Agent: A Self-Referential Agent Framework for Recursive Self-Improvementarxiv-2410.04444 Sparse Blocked context onlyOct 1, 2024
- PACE: Procedural Abstractions for Communicating Efficientlyarxiv-2409.20120 Sparse Blocked context onlySep 30, 2024
- Human-like Affective Cognition in Foundation Modelsarxiv-2409.11733 Sparse Blocked context onlySep 18, 2024
- LongGenBench: Benchmarking Long-Form Generation in Long Context LLMsarxiv-2409.02076 Sparse Blocked context onlySep 1, 2024
- OLMoE: Open Mixture-of-Experts Language Modelsarxiv-2409.02060 Sparse Blocked context onlySep 1, 2024
- Parallel AutoRegressive Models for Multi-Agent Combinatorial Optimizationarxiv-2409.03811 Sparse Blocked context onlySep 1, 2024
- MediConfusion: Can you trust your AI radiologist? Probing the reliability of multimodal medical foundation modelsarxiv-2409.15477 Sparse Blocked context onlySep 1, 2024
- Abstracted Gaussian Prototypes for True One-Shot Concept Learningarxiv-2408.17251 Sparse Blocked context onlyAug 30, 2024
- LiCoEval: Evaluating LLMs on License Compliance in Code Generationarxiv-2408.02487 Curated Related Blocked context onlyAug 1, 2024
- LLMs as Zero-shot Graph Learners: Alignment of GNN Representations with LLM Token Embeddingsarxiv-2408.14512 Sparse Blocked context onlyAug 1, 2024
- LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMsarxiv-2408.07055 Curated Related Blocked context onlyAug 1, 2024
- Bidirectional Decoding: Improving Action Chunking via Guided Test-Time Samplingarxiv-2408.17355 Sparse Blocked context onlyAug 1, 2024
- SigmaRL: A Sample-Efficient and Generalizable Multi-Agent Reinforcement Learning Framework for Motion Planningarxiv-2408.07644 Sparse Blocked context onlyAug 1, 2024
- FinCon: A Synthesized LLM Multi-Agent System with Conceptual Verbal Reinforcement for Enhanced Financial Decision Makingarxiv-2407.06567 Sparse Blocked context onlyJul 1, 2024
- Measuring the Measurers: Quality Evaluation of Hallucination Benchmarks for Large Vision-Language Modelsarxiv-2406.17115 Sparse Blocked context onlyJun 24, 2024
- Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothingarxiv-2406.08464 Sparse Blocked context onlyJun 1, 2024
- AutoHallusion: Automatic Generation of Hallucination Benchmarks for Vision-Language Modelsarxiv-2406.10900 Sparse Blocked context onlyJun 1, 2024
- CoAct: A Global-Local Hierarchy for Autonomous Agent Collaborationarxiv-2406.13381 Sparse Blocked context onlyJun 1, 2024
- TorchSpatial: A Location Encoding Framework and Benchmark for Spatial Representation Learningarxiv-2406.15658 Sparse Blocked context onlyJun 1, 2024
- The CLRS-Text Algorithmic Reasoning Language Benchmarkarxiv-2406.04229 Sparse Blocked context onlyJun 1, 2024
- Are We Done with MMLU?arxiv-2406.04127 Sparse Blocked context onlyJun 1, 2024
- Augmenting Lateral Thinking in Language Models with Humor and Riddle Data for the BRAINTEASER Taskarxiv-2405.10385 Sparse Blocked context onlyMay 16, 2024
- Group Robust Preference Optimization in Reward-free RLHFarxiv-2405.20304 Sparse Blocked context onlyMay 1, 2024
- USP: A Unified Sequence Parallelism Approach for Long Context Generative AIarxiv-2405.07719 Sparse Blocked context onlyMay 1, 2024
- Xmodel-VLM: A Simple Baseline for Multimodal Vision Language Modelarxiv-2405.09215 Sparse Blocked context onlyMay 1, 2024
- Cooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM Agentsarxiv-2404.16698 Sparse Blocked context onlyApr 1, 2024
- LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Modelarxiv-2404.01331 Sparse Blocked context onlyApr 1, 2024
- Photo-Realistic Image Restoration in the Wild with Controlled Vision-Language Modelsarxiv-2404.09732 Curated Related Blocked context onlyApr 1, 2024
- Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answeringarxiv-2403.04890 Sparse Blocked context onlyMar 7, 2024
- Multi-agent deep reinforcement learning with centralized training and decentralized execution for transportation infrastructure managementarxiv-2401.12455 Sparse Blocked context onlyJan 23, 2024
- Moonshot: Towards Controllable Video Generation and Editing with Multimodal Conditionsarxiv-2401.01827 Sparse Blocked context onlyJan 1, 2024
- RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedbackarxiv-2312.00849 Sparse Blocked context onlyDec 1, 2023
- Mamba: Linear-Time Sequence Modeling with Selective State Spacesdoi-10.48550_arxiv.2312.00752 Sparse Blocked context onlyDec 1, 2023
- MetaTrinity: Enabling Fast Metagenomic Classification via Seed Counting and Edit Distance Approximationdoi-10.48550_arxiv.2311.02029 Sparse Blocked context onlyNov 3, 2023
- FinMem: A Performance-Enhanced LLM Trading Agent with Layered Memory and Character Designarxiv-2311.13743 Sparse Blocked context onlyNov 1, 2023
- Llemma: An Open Language Model For Mathematicsarxiv-2310.10631 Sparse Blocked context onlyOct 16, 2023
- Llemma: An Open Language Model For Mathematicsdoi-10.48550_arxiv.2310.10631 Sparse Blocked context onlyOct 16, 2023
- Efficient Streaming Language Models with Attention Sinksdoi-10.48550_arxiv.2309.17453 Sparse Blocked context onlySep 29, 2023
- Event Stream-based Visual Object Tracking: A High-Resolution Benchmark Dataset and A Novel Baselinearxiv-2309.14611 Sparse Blocked context onlySep 26, 2023
- Bespoke Nanoparticle Synthesis and Chemical Knowledge Discovery Via Autonomous Experimentationsdoi-10.48550_arxiv.2309.00349 Sparse Blocked context onlySep 1, 2023
- SynJax: Structured Probability Distributions for JAXarxiv-2308.03291 Sparse Blocked context onlyAug 7, 2023
- SwinJSCC: Taming Swin Transformer for Deep Joint Source-Channel Codingarxiv-2308.09361 Sparse Blocked context onlyAug 1, 2023
- DNABERT-2: Efficient Foundation Model and Benchmark For Multi-Species Genomedoi-10.48550_arxiv.2306.15006 Sparse Blocked context onlyJun 26, 2023
- RedMotion: Motion Prediction via Redundancy Reductiondoi-10.48550_arxiv.2306.10840 Sparse Blocked context onlyJun 19, 2023
- SugarCrepe: Fixing Hackable Benchmarks for Vision-Language Compositionalityarxiv-2306.14610 Sparse Blocked context onlyJun 1, 2023
- ChemCrow: Augmenting large-language models with chemistry toolsdoi-10.48550_arxiv.2304.05376 Sparse Blocked context onlyApr 11, 2023
- Generative Agents: Interactive Simulacra of Human Behaviorarxiv-2304.03442 Sparse Blocked context onlyApr 1, 2023
- LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attentiondoi-10.48550_arxiv.2303.16199 Sparse Blocked context onlyMar 28, 2023
- Neuro-symbolic Commonsense Social Reasoningdoi-10.48550_arxiv.2303.08264 Sparse Blocked context onlyMar 14, 2023