300 canonical paper links on this archive page.
- HACHIMI: Scalable and Controllable Student Persona Generation via Orchestrated Agentsarxiv-2603.04855 Sparse Blocked context onlyMar 5, 2026
- Breaking Contextual Inertia: Reinforcement Learning with Single-Turn Anchors for Stable Multi-Turn Interactionarxiv-2603.04783 Sparse Blocked context onlyMar 5, 2026
- HiMAP-Travel: Hierarchical Multi-Agent Planning for Long-Horizon Constrained Travelarxiv-2603.04750 Sparse Blocked context onlyMar 5, 2026
- Interactive Benchmarksarxiv-2603.04737 Sparse Blocked context onlyMar 5, 2026
- Solving an Open Problem in Theoretical Physics using AI-Assisted Discoveryarxiv-2603.04735 Sparse Blocked context onlyMar 5, 2026
- Model Medicine: A Clinical Framework for Understanding, Diagnosing, and Treating AI Modelsarxiv-2603.04722 Sparse Blocked context onlyMar 5, 2026
- AI-Assisted Moot Courts: Simulating Justice-Specific Questioning in Oral Argumentsarxiv-2603.04718 Sparse Blocked context onlyMar 5, 2026
- iAgentBench: Benchmarking Sensemaking Capabilities of Information-Seeking Agents on High-Traffic Topicsarxiv-2603.04656 Sparse Blocked context onlyMar 4, 2026
- $τ$-Knowledge: Evaluating Conversational Agents over Unstructured Knowledgearxiv-2603.04370 Sparse Blocked context onlyMar 4, 2026
- Dual-Modality Multi-Stage Adversarial Safety Training: Robustifying Multimodal Web Agents Against Cross-Modal Attacksarxiv-2603.04364 Sparse Blocked context onlyMar 4, 2026
- AILS-NTUA at SemEval-2026 Task 12: Graph-Based Retrieval and Reflective Prompting for Abductive Event Reasoningarxiv-2603.04319 Sparse Blocked context onlyMar 4, 2026
- $V_1$: Unifying Generation and Self-Verification for Parallel Reasonersarxiv-2603.04304 Sparse Blocked context onlyMar 4, 2026
- BD-Merging: Bias-Aware Dynamic Model Merging with Evidence-Guided Contrastive Learningarxiv-2603.03920 Sparse Blocked context onlyMar 4, 2026
- On the Suitability of LLM-Driven Agents for Dark Pattern Auditsarxiv-2603.03881 Sparse Blocked context onlyMar 4, 2026
- Narrative Weaver: Towards Controllable Long-Range Visual Consistency with Multi-Modal Conditioningarxiv-2603.06688 Sparse Blocked context onlyMar 4, 2026
- T2S-Bench & Structure-of-Thought: Benchmarking and Prompting Comprehensive Text-to-Structure Reasoningarxiv-2603.03790 Sparse Blocked context onlyMar 4, 2026
- A Neural Topic Method Using a Large-Language-Model-in-the-Loop for Business Researcharxiv-2603.03623 Sparse Blocked context onlyMar 4, 2026
- Expectation and Acoustic Neural Network Representations Enhance Music Identification from Brain Activityarxiv-2603.03190 Sparse Blocked context onlyMar 3, 2026
- Nodes Are Early, Edges Are Late: Probing Diagram Representations in Large Vision-Language Modelsarxiv-2603.02865 Sparse Blocked context onlyMar 3, 2026
- BrandFusion: A Multi-Agent Framework for Seamless Brand Integration in Text-to-Video Generationarxiv-2603.02816 Sparse Blocked context onlyMar 3, 2026
- Guideline-Grounded Evidence Accumulation for High-Stakes Agent Verificationarxiv-2603.02798 Sparse Blocked context onlyMar 3, 2026
- Agentified Assessment of Logical Reasoning Agentsarxiv-2603.02788 Sparse Blocked context onlyMar 3, 2026
- HateMirage: An Explainable Multi-Dimensional Dataset for Decoding Faux Hate and Subtle Online Abusearxiv-2603.02684 Sparse Blocked context onlyMar 3, 2026
- Credibility Governance: A Social Mechanism for Collective Self-Correction under Weak Truth Signalsarxiv-2603.02640 Sparse Blocked context onlyMar 3, 2026
- Tool Verification for Test-Time Reinforcement Learningarxiv-2603.02203 Sparse Blocked context onlyMar 2, 2026
- Organizing, Orchestrating, and Benchmarking Agent Skills at Ecosystem Scalearxiv-2603.02176 Sparse Blocked context onlyMar 2, 2026
- Boltzmann-based Exploration for Robust Decentralized Multi-Agent Planning (Extended Version)arxiv-2603.02154 Sparse Blocked context onlyMar 2, 2026
- Recursive Models for Long-Horizon Reasoningarxiv-2603.02112 Sparse Blocked context onlyMar 2, 2026
- Modeling Grammatical Hypothesis Testing in Young Learners: A Sequence-Based Learning Analytics Study of Morphosyntactic Reasoning in an Interactive Gamearxiv-2603.02084 Sparse Blocked context onlyMar 2, 2026
- According to Me: Long-Term Personalized Referential Memory QAarxiv-2603.01990 Sparse Blocked context onlyMar 2, 2026
- When Numbers Tell Half the Story: Human-Metric Alignment in Topic Model Evaluationarxiv-2603.01945 Sparse Blocked context onlyMar 2, 2026
- Demonstrating ViviDoc: Generating Interactive Documents through Human-Agent Collaborationarxiv-2603.01912 Sparse Blocked context onlyMar 2, 2026
- CRoCoDiL: Continuous and Robust Conditioned Diffusion for Languagearxiv-2603.20210 Sparse Blocked context onlyMar 2, 2026
- Legal RAG Bench: an end-to-end benchmark for legal RAGarxiv-2603.01710 Sparse Blocked context onlyMar 2, 2026
- Beyond the Grid: Layout-Informed Multi-Vector Retrieval with Parsed Visual Document Representationsarxiv-2603.01666 Sparse Blocked context onlyMar 2, 2026
- SEED-SET: Scalable Evolving Experimental Design for System-level Ethical Testingarxiv-2603.01630 Sparse Blocked context onlyMar 2, 2026
- Measuring What VLMs Don't Say: Validation Metrics Hide Clinical Terminology Erasure in Radiology Report Generationarxiv-2603.01625 Sparse Blocked context onlyMar 2, 2026
- Markovian ODE-guided scoring can assess the quality of offline reasoning traces in language modelsarxiv-2603.01580 Sparse Blocked context onlyMar 2, 2026
- End-to-End Simultaneous Dysarthric Speech Reconstruction with Frame-Level Adaptor and Multiple Wait-k Knowledge Distillationarxiv-2603.01382 Sparse Blocked context onlyMar 2, 2026
- Catalyst-Agent: Autonomous heterogeneous catalyst screening with an LLM Agentarxiv-2603.01311 Sparse Blocked context onlyMar 1, 2026
- Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended Impactarxiv-2603.00883 Sparse Blocked context onlyMar 1, 2026
- A Typologically Grounded Evaluation Framework for Word Order and Morphology Sensitivity in Multilingual Masked LMsarxiv-2603.00432 Sparse Blocked context onlyFeb 28, 2026
- Stepwise Penalization for Length-Efficient Chain-of-Thought Reasoningarxiv-2603.00296 Sparse Blocked context onlyFeb 27, 2026
- MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Gamesarxiv-2602.24188 Sparse Blocked context onlyFeb 27, 2026
- AgenticOCR: Parsing Only What You Need for Efficient Retrieval-Augmented Generationarxiv-2602.24134 Sparse Blocked context onlyFeb 27, 2026
- MemEmo: Evaluating Emotion in Memory Systems of Agentsarxiv-2602.23944 Sparse Blocked context onlyFeb 27, 2026
- Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasksarxiv-2602.23898 Sparse Blocked context onlyFeb 27, 2026
- CLFEC: A New Task for Unified Linguistic and Factual Error Correction in paragraph-level Chinese Professional Writingarxiv-2602.23845 Sparse Blocked context onlyFeb 27, 2026
- RLShield: Practical Multi-Agent RL for Financial Cyber Defense with Attack-Surface MDPs and Real-Time Response Orchestrationarxiv-2603.00186 Sparse Blocked context onlyFeb 26, 2026
- Humans and LLMs Diverge on Probabilistic Inferencesarxiv-2602.23546 Sparse Blocked context onlyFeb 26, 2026
- SeeThrough3D: Occlusion Aware 3D Control in Text-to-Image Generationarxiv-2602.23359 Sparse Blocked context onlyFeb 26, 2026
- Conformalized Neural Networks for Federated Uncertainty Quantification under Dual Heterogeneityarxiv-2602.23296 Sparse Blocked context onlyFeb 26, 2026
- SPARTA: Scalable and Principled Benchmark of Tree-Structured Multi-hop QA over Text and Tablesarxiv-2602.23286 Sparse Blocked context onlyFeb 26, 2026
- Evaluating Stochasticity in Deep Research Agentsarxiv-2602.23271 Sparse Blocked context onlyFeb 26, 2026
- Risk-Aware World Model Predictive Control for Generalizable End-to-End Autonomous Drivingarxiv-2602.23259 Sparse Blocked context onlyFeb 26, 2026
- A Model-Free Universal AIarxiv-2602.23242 Sparse Blocked context onlyFeb 26, 2026
- Spatio-Temporal Token Pruning for Efficient High-Resolution GUI Agentsarxiv-2602.23235 Sparse Blocked context onlyFeb 26, 2026
- ReCoN-Ipsundrum: An Inspectable Recurrent Persistence Loop Agent with Affect-Coupled Control and Mechanism-Linked Consciousness Indicator Assaysarxiv-2602.23232 Sparse Blocked context onlyFeb 26, 2026
- MovieTeller: Tool-augmented Movie Synopsis with ID Consistent Progressive Abstractionarxiv-2602.23228 Sparse Blocked context onlyFeb 26, 2026
- PATRA: Pattern-Aware Alignment and Balanced Reasoning for Time Series Question Answeringarxiv-2602.23161 Sparse Blocked context onlyFeb 26, 2026
- Efficient Encoder-Free Fourier-based 3D Large Multimodal Modelarxiv-2602.23153 Sparse Blocked context onlyFeb 26, 2026
- The Trinity of Consistency as a Defining Principle for General World Modelsarxiv-2602.23152 Sparse Blocked context onlyFeb 26, 2026
- On Sample-Efficient Generalized Planning via Learned Transition Modelsarxiv-2602.23148 Sparse Blocked context onlyFeb 26, 2026
- Modality Collapse as Mismatched Decoding: Information-Theoretic Limits of Multimodal LLMsarxiv-2602.23136 Sparse Blocked context onlyFeb 26, 2026
- DyGnROLE: Modeling Asymmetry in Dynamic Graphs with Node-Role-Oriented Latent Encodingarxiv-2602.23135 Sparse Blocked context onlyFeb 26, 2026
- Three AI-agents walk into a bar . . . . `Lord of the Flies' tribalism emerges among smart AI-Agentsarxiv-2602.23093 Sparse Blocked context onlyFeb 26, 2026
- Accelerated Online Risk-Averse Policy Evaluation in POMDPs with Theoretical Guarantees and Novel CVaR Boundsarxiv-2602.23073 Sparse Blocked context onlyFeb 26, 2026
- Make It Hard to Hear, Easy to Learn: Long-Form Bengali ASR and Speaker Diarization via Extreme Augmentation and Perfect Alignmentarxiv-2602.23070 Sparse Blocked context onlyFeb 26, 2026
- MoDora: Tree-Based Semi-Structured Document Analysis Systemarxiv-2602.23061 Sparse Blocked context onlyFeb 26, 2026
- Learning-based Multi-agent Race Strategies in Formula 1arxiv-2602.23056 Sparse Blocked context onlyFeb 26, 2026
- SkillNet: Create, Evaluate, and Connect AI Skillsarxiv-2603.04448 Sparse Blocked context onlyFeb 26, 2026
- Modeling Expert AI Diagnostic Alignment via Immutable Inference Snapshotsarxiv-2602.22973 Sparse Blocked context onlyFeb 26, 2026
- A Holistic Framework for Robust Bangla ASR and Speaker Diarization with Optimized VAD and CTC Alignmentarxiv-2602.22935 Sparse Blocked context onlyFeb 26, 2026
- Where Vision Becomes Text: Locating the OCR Routing Bottleneck in Vision-Language Modelsarxiv-2602.22918 Sparse Blocked context onlyFeb 26, 2026
- CeRA: Breaking the Linear Ceiling of Low-Rank Adaptation via Manifold Expansionarxiv-2602.22911 Sparse Blocked context onlyFeb 26, 2026
- MEDNA-DFM: A Dual-View FiLM-MoE Model for Explainable DNA Methylation Predictionarxiv-2602.22850 Sparse Blocked context onlyFeb 26, 2026
- Decentralized Ranking Aggregation: Gossip Algorithms for Borda and Copeland Consensusarxiv-2602.22847 Sparse Blocked context onlyFeb 26, 2026
- Improving Neural Argumentative Stance Classification in Controversial Topics with Emotion-Lexicon Featuresarxiv-2602.22846 Sparse Blocked context onlyFeb 26, 2026
- DeepPresenter: Environment-Grounded Reflection for Agentic Presentation Generationarxiv-2602.22839 Sparse Blocked context onlyFeb 26, 2026
- Moral Preferences of LLMs Under Directed Contextual Influencearxiv-2602.22831 Curated Related Blocked context onlyFeb 26, 2026
- When Should an AI Act? A Human-Centered Model of Scene, Context, and Behavior for Agentic AI Designarxiv-2602.22814 Sparse Blocked context onlyFeb 26, 2026
- TherapyProbe: Generating Design Knowledge for Relational Safety in Mental Health Chatbots Through Adversarial Simulationarxiv-2602.22775 Sparse Blocked context onlyFeb 26, 2026
- AuditBench: Evaluating Alignment Auditing Techniques on Models with Hidden Behaviorsarxiv-2602.22755 Sparse Blocked context onlyFeb 26, 2026
- Generative Data Transformation: From Mixed to Unified Dataarxiv-2602.22743 Sparse Blocked context onlyFeb 26, 2026
- Enhancing Persuasive Dialogue Agents by Synthesizing Cross-Disciplinary Communication Strategiesarxiv-2602.22696 Sparse Blocked context onlyFeb 26, 2026
- ViCLIP-OT: The First Foundation Vision-Language Model for Vietnamese Image-Text Retrieval with Optimal Transportarxiv-2602.22678 Curated Related Blocked context onlyFeb 26, 2026
- Search More, Think Less: Rethinking Long-Horizon Agentic Search for Efficiency and Generalizationarxiv-2602.22675 Sparse Blocked context onlyFeb 26, 2026
- ContextRL: Enhancing MLLM's Knowledge Discovery Efficiency with Context-Augmented RLarxiv-2602.22623 Sparse Blocked context onlyFeb 26, 2026
- Strategy Executability in Mathematical Reasoning: Leveraging Human-Model Differences for Effective Guidancearxiv-2602.22583 Sparse Blocked context onlyFeb 26, 2026
- VeRO: An Evaluation Harness for Agents to Optimize Agentsarxiv-2602.22480 Sparse Blocked context onlyFeb 25, 2026
- How Do Latent Reasoning Methods Perform Under Weak and Strong Supervision?arxiv-2602.22441 Sparse Blocked context onlyFeb 25, 2026
- Detecting Hate and Inflammatory Content in Bengali Memes: A New Multimodal Dataset and Co-Attention Frameworkarxiv-2602.22391 Sparse Blocked context onlyFeb 25, 2026
- NoLan: Mitigating Object Hallucinations in Large Vision-Language Models via Dynamic Suppression of Language Priorsarxiv-2602.22144 Sparse Blocked context onlyFeb 25, 2026
- fEDM+: A Risk-Based Fuzzy Ethical Decision Making Framework with Principle-Level Explainability and Pluralistic Validationarxiv-2602.21746 Sparse Blocked context onlyFeb 25, 2026
- The ASIR Courage Model: A Phase-Dynamic Framework for Truth Transitions in Human and AI Systemsarxiv-2602.21745 Sparse Blocked context onlyFeb 25, 2026
- PPCR-IM: A System for Multi-layer DAG-based Public Policy Consequence Reasoning and Social Indicator Mappingarxiv-2602.21650 Sparse Blocked context onlyFeb 25, 2026
- Self-Correcting VLA: Online Action Refinement via Sparse World Imaginationarxiv-2602.21633 Sparse Blocked context onlyFeb 25, 2026
- Enhancing Multilingual Embeddings via Multi-Way Parallel Text Alignmentarxiv-2602.21543 Sparse Blocked context onlyFeb 25, 2026
- LiLo-VLA: Compositional Long-Horizon Manipulation via Linked Object-Centric Policiesarxiv-2602.21531 Sparse Blocked context onlyFeb 25, 2026
- Adversarial Robustness of Deep Learning-Based Thyroid Nodule Segmentation in Ultrasoundarxiv-2602.21452 Sparse Blocked context onlyFeb 25, 2026
- Adversarial Intent is a Latent Variable: Stateful Trust Inference for Securing Multimodal Agentic RAGarxiv-2602.21447 Sparse Blocked context onlyFeb 24, 2026
- On the Structural Non-Preservation of Epistemic Behaviour under Policy Transformationarxiv-2602.21424 Sparse Blocked context onlyFeb 24, 2026
- NeuroNarrator: A Generalist EEG-to-Text Foundation Model for Clinical Interpretation via Spectro-Spatial Grounding and Temporal State-Space Reasoningarxiv-2603.16880 Sparse Blocked context onlyFeb 24, 2026
- Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanizationarxiv-2604.09574 Sparse Blocked context onlyFeb 24, 2026
- Agile V: A Compliance-Ready Framework for AI-Augmented Engineering -- From Concept to Audit-Ready Deliveryarxiv-2602.20684 Sparse Blocked context onlyFeb 24, 2026
- OrthoAI: A Neurosymbolic Framework for Evidence-Grounded Biomechanical Reasoning in Clear Aligner Orthodonticsarxiv-2603.00124 Sparse Blocked context onlyFeb 23, 2026
- CodeCompass: Navigating the Navigation Paradox in Agentic Code Intelligencearxiv-2602.20048 Sparse Blocked context onlyFeb 23, 2026
- NovaPlan: Zero-Shot Long-Horizon Manipulation via Closed-Loop Video Language Planningarxiv-2602.20119 Sparse Blocked context onlyFeb 23, 2026
- A Very Big Video Reasoning Suitearxiv-2602.20159 Sparse Blocked context onlyFeb 23, 2026
- Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Usearxiv-2602.20426 Sparse Blocked context onlyFeb 23, 2026
- Quantifying the Expectation-Realisation Gap for Agentic AI Systemsarxiv-2602.20292 Sparse Blocked context onlyFeb 23, 2026
- CTC-TTS: LLM-based dual-streaming text-to-speech with CTC alignmentarxiv-2602.19574 Sparse Blocked context onlyFeb 23, 2026
- Pixels Don't Lie (But Your Detector Might): Bootstrapping MLLM-as-a-Judge for Trustworthy Deepfake Detection and Reasoning Supervisionarxiv-2602.19715 Sparse Blocked context onlyFeb 23, 2026
- Training-Free Cross-Architecture Merging for Graph Neural Networksarxiv-2602.19332 Sparse Blocked context onlyFeb 22, 2026
- Rethinking Preference Alignment for Diffusion Models with Classifier-Free Guidancearxiv-2602.18799 Sparse Blocked context onlyFeb 21, 2026
- Neural Fields as World Modelsarxiv-2602.18690 Sparse Blocked context onlyFeb 21, 2026
- Rethinking Retrieval-Augmented Generation as a Cooperative Decision-Making Problemarxiv-2602.18734 Sparse Blocked context onlyFeb 21, 2026
- Comparative Assessment of Multimodal Earth Observation Data for Soil Moisture Estimationarxiv-2602.18083 Sparse Blocked context onlyFeb 20, 2026
- Decoding ML Decision: An Agentic Reasoning Framework for Large-Scale Ranking Systemarxiv-2602.18640 Sparse Blocked context onlyFeb 20, 2026
- From Lossy to Verified: A Provenance-Aware Tiered Memory for Agentsarxiv-2602.17913 Sparse Blocked context onlyFeb 20, 2026
- Condition-Gated Reasoning for Context-Dependent Biomedical Question Answeringarxiv-2602.17911 Sparse Blocked context onlyFeb 20, 2026
- Promptable segmentation with region exploration enables minimal-effort expert-level prostate cancer delineationarxiv-2602.17813 Sparse Blocked context onlyFeb 19, 2026
- WS-GRPO: Weakly-Supervised Group-Relative Policy Optimization for Rollout-Efficient Reasoningarxiv-2602.17025 Sparse Blocked context onlyFeb 19, 2026
- FAMOSE: A ReAct Approach to Automated Feature Discoveryarxiv-2602.17641 Sparse Blocked context onlyFeb 19, 2026
- MeDUET: Disentangled Unified Pretraining for 3D Medical Image Synthesis and Analysisarxiv-2602.17901 Sparse Blocked context onlyFeb 19, 2026
- The Bots of Persuasion: Examining How Conversational Agents' Linguistic Expressions of Personality Affect User Perceptions and Decisionsarxiv-2602.17185 Sparse Blocked context onlyFeb 19, 2026
- Suppression or Deletion: A Restoration-Based Representation-Level Analysis of Machine Unlearningarxiv-2602.18505 Sparse Blocked context onlyFeb 18, 2026
- Language Statistics and False Belief Reasoning: Evidence from 41 Open-Weight LMsarxiv-2602.16085 Sparse Blocked context onlyFeb 17, 2026
- A Curious Class of Adpositional Multiword Expressions in Koreanarxiv-2602.16023 Sparse Blocked context onlyFeb 17, 2026
- Enhancing Building Semantics Preservation in AI Model Training with Large Language Model Encodingsarxiv-2602.15791 Sparse Blocked context onlyFeb 17, 2026
- *-PLUIE: Personalisable metric with Llm Used for Improved Evaluationarxiv-2602.15778 Sparse Blocked context onlyFeb 17, 2026
- GLM-5: from Vibe Coding to Agentic Engineeringarxiv-2602.15763 Sparse Blocked context onlyFeb 17, 2026
- Beyond Binary Classification: Detecting Fine-Grained Sexism in Social Media Videosarxiv-2602.15757 Sparse Blocked context onlyFeb 17, 2026
- Beyond Static Pipelines: Learning Dynamic Workflows for Text-to-SQLarxiv-2602.15564 Sparse Blocked context onlyFeb 17, 2026
- RUVA: Personalized Transparent On-Device Graph Reasoningarxiv-2602.15553 Sparse Blocked context onlyFeb 17, 2026
- jina-embeddings-v5-text: Task-Targeted Embedding Distillationarxiv-2602.15547 Sparse Blocked context onlyFeb 17, 2026
- ExpertWeaver: Unlocking the Inherent MoE in Dense LLMs with GLU Activation Patternsarxiv-2602.15521 Sparse Blocked context onlyFeb 17, 2026
- NeuroSymActive: Differentiable Neural-Symbolic Reasoning with Active Exploration for Knowledge Graph Question Answeringarxiv-2602.15353 Sparse Blocked context onlyFeb 17, 2026
- Prescriptive Scaling Reveals the Evolution of Language Model Capabilitiesarxiv-2602.15327 Sparse Blocked context onlyFeb 17, 2026
- The Information Geometry of Softmax: Probing and Steeringarxiv-2602.15293 Sparse Blocked context onlyFeb 17, 2026
- FrameRef: A Framing Dataset and Simulation Testbed for Modeling Bounded Rational Information Healtharxiv-2602.15273 Sparse Blocked context onlyFeb 17, 2026
- How to Train Your Long-Context Visual Document Modelarxiv-2602.15257 Sparse Blocked context onlyFeb 16, 2026
- Colosseum: Auditing Collusion in Cooperative Multi-Agent Systemsarxiv-2602.15198 Sparse Blocked context onlyFeb 16, 2026
- CGRA-DeBERTa Concept Guided Residual Augmentation Transformer for Theologically Islamic Understandingarxiv-2602.15139 Sparse Blocked context onlyFeb 16, 2026
- Cold-Start Personalization via Training-Free Priors from Structured World Modelsarxiv-2602.15012 Sparse Blocked context onlyFeb 16, 2026
- Tool-Aware Planning in Contact Center AI: Evaluating LLMs through Lineage-Guided Query Decompositionarxiv-2602.14955 Sparse Blocked context onlyFeb 16, 2026
- GOT-JEPA: Generic Object Tracking with Model Adaptation and Occlusion Handling using Joint-Embedding Predictive Architecturearxiv-2602.14771 Sparse Blocked context onlyFeb 16, 2026
- Multi-Agent Comedy Club: Investigating Community Discussion Effects on LLM Humor Generationarxiv-2602.14770 Sparse Blocked context onlyFeb 16, 2026
- Breaking Data Efficiency Dilemma: A Federated and Augmented Learning Framework For Alzheimer's Disease Detection via Speecharxiv-2602.14655 Sparse Blocked context onlyFeb 16, 2026
- Is Information Density Uniform when Utterances are Grounded on Perception and Discourse?arxiv-2602.14653 Sparse Blocked context onlyFeb 16, 2026
- Measuring and Mitigating Post-hoc Rationalization in Reverse Chain-of-Thought Generationarxiv-2602.14469 Sparse Blocked context onlyFeb 16, 2026
- Synthetic Reader Panels: Tournament-Based Ideation with LLM Personas for Autonomous Publishingarxiv-2602.14433 Sparse Blocked context onlyFeb 16, 2026
- Does Socialization Emerge in AI Agent Society? A Case Study of Moltbookarxiv-2602.14299 Sparse Blocked context onlyFeb 15, 2026
- FMMD: A multimodal open peer review dataset based on F1000Researcharxiv-2602.14285 Sparse Blocked context onlyFeb 15, 2026
- Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factualityarxiv-2602.14080 Sparse Blocked context onlyFeb 15, 2026
- GTS: Inference-Time Scaling of Latent Reasoning with a Learnable Gaussian Thought Samplerarxiv-2602.14077 Sparse Blocked context onlyFeb 15, 2026
- LogitsCoder: Towards Efficient Chain-of-Thought Path Search via Logits Preference Decoding for Code Generationarxiv-2602.14054 Sparse Blocked context onlyFeb 15, 2026
- Geometry-Preserving Aggregation for Mixture-of-Experts Embedding Modelsarxiv-2602.14039 Sparse Blocked context onlyFeb 15, 2026
- Neuromem: A Granular Decomposition of the Streaming Lifecycle in External Memory for LLMsarxiv-2602.13967 Sparse Blocked context onlyFeb 15, 2026
- Mobile-Agent-v3.5: Multi-platform Fundamental GUI Agentsarxiv-2602.16855 Sparse Blocked context onlyFeb 15, 2026
- Why Code, Why Now: Learnability, Computability, and the Real Limits of Machine Learningarxiv-2602.13934 Sparse Blocked context onlyFeb 15, 2026
- A Comparative Analysis of Social Network Topology in Reddit and Moltbookarxiv-2602.13920 Sparse Blocked context onlyFeb 14, 2026
- DeepXiv-SDK: An Agentic Data Interface for Scientific Literaturearxiv-2603.00084 Sparse Blocked context onlyFeb 14, 2026
- Speculative Decoding with a Speculative Vocabularyarxiv-2602.13836 Sparse Blocked context onlyFeb 14, 2026
- OMGs: A multi-agent system supporting MDT decision-making across the ovarian tumour care continuumarxiv-2602.13793 Sparse Blocked context onlyFeb 14, 2026
- StackingNet: Collective Inference Across Independent AI Foundation Modelsarxiv-2602.13792 Sparse Blocked context onlyFeb 14, 2026
- OR-Agent: Bridging Evolutionary Search and Structured Research for Automated Algorithm Discoveryarxiv-2602.13769 Sparse Blocked context onlyFeb 14, 2026
- On Theoretically-Driven LLM Agents for Multi-Dimensional Discourse Analysisarxiv-2602.13713 Sparse Blocked context onlyFeb 14, 2026
- Beyond Normalization: Rethinking the Partition Function as a Difficulty Scheduler for RLVRarxiv-2602.12642 Sparse Blocked context onlyFeb 13, 2026
- Evolving Beyond Snapshots: Harmonizing Structure and Sequence via Entity State Tuning for Temporal Knowledge Graph Forecastingarxiv-2602.12389 Sparse Blocked context onlyFeb 12, 2026
- Variation-aware Flexible 3D Gaussian Editingarxiv-2602.11638 Sparse Blocked context onlyFeb 12, 2026
- Advancing AI Trustworthiness Through Patient Simulation: Risk Assessment of Conversational Agents for Antidepressant Selectionarxiv-2602.11391 Sparse Blocked context onlyFeb 11, 2026
- GraphSeek: Next-Generation Graph Analytics with LLMsarxiv-2602.11052 Sparse Blocked context onlyFeb 11, 2026
- Learning Adaptive Distribution Alignment with Neural Characteristic Function for Graph Domain Adaptationarxiv-2602.10489 Sparse Blocked context onlyFeb 11, 2026
- The Subjectivity of Respect in Police Traffic Stops: Modeling Community Perspectives in Body-Worn Camera Footagearxiv-2602.10339 Sparse Blocked context onlyFeb 10, 2026
- Towards Autonomous Mathematics Researcharxiv-2602.10177 Sparse Blocked context onlyFeb 10, 2026
- LLMs Encode Their Failures: Predicting Success from Pre-Generation Activationsarxiv-2602.09924 Sparse Blocked context onlyFeb 10, 2026
- How effective are VLMs in assisting humans in inferring the quality of mental models from Multimodal short answers?arxiv-2603.00056 Sparse Blocked context onlyFeb 10, 2026
- LingxiDiagBench: A Multi-Agent Framework for Benchmarking LLMs in Chinese Psychiatric Consultation and Diagnosisarxiv-2602.09379 Sparse Blocked context onlyFeb 10, 2026
- Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structurearxiv-2602.08783 Sparse Blocked context onlyFeb 9, 2026
- LLMs Know More About Numbers than They Can Sayarxiv-2602.07812 Sparse Blocked context onlyFeb 8, 2026
- Faithful Bi-Directional Model Steering via Distribution Matching and Distributed Interchange Interventionsarxiv-2602.05234 Sparse Blocked context onlyFeb 5, 2026
- Bagpiper: Solving Open-Ended Audio Tasks via Rich Captionsarxiv-2602.05220 Sparse Blocked context onlyFeb 5, 2026
- The Single-Multi Evolution Loop for Self-Improving Model Collaboration Systemsarxiv-2602.05182 Sparse Blocked context onlyFeb 5, 2026
- LABBench2: An Improved Benchmark for AI Systems Performing Biology Researcharxiv-2604.09554 Direct Blocked context onlyFeb 4, 2026
- VILLAIN at AVerImaTeC: Verifying Image-Text Claims via Multi-Agent Collaborationarxiv-2602.04587 Sparse Blocked context onlyFeb 4, 2026
- Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Expertsarxiv-2602.03473 Sparse Blocked context onlyFeb 3, 2026
- UNIKIE-BENCH: Benchmarking Large Multimodal Models for Key Information Extraction in Visual Documentsarxiv-2602.07038 Sparse Blocked context onlyFeb 3, 2026
- Aligning Language Model Benchmarks with Pairwise Preferencesarxiv-2602.02898 Sparse Blocked context onlyFeb 2, 2026
- Scaling Small Agents Through Strategy Auctionsarxiv-2602.02751 Sparse Blocked context onlyFeb 2, 2026
- From Sycophancy to Sensemaking: Premise Governance for Human-AI Decision Makingarxiv-2602.02378 Sparse Blocked context onlyFeb 2, 2026
- VQ-Style: Disentangling Style and Content in Motion with Residual Quantized Representationsarxiv-2602.02334 Sparse Blocked context onlyFeb 2, 2026
- CryoLVM: Self-supervised Learning from Cryo-EM Density Maps with Large Vision Modelsarxiv-2602.02620 Sparse Blocked context onlyFeb 2, 2026
- Beyond RAG for Agent Memory: Retrieval by Decoupling and Aggregationarxiv-2602.02007 Sparse Blocked context onlyFeb 2, 2026
- OpenVTON-Bench: A Large-Scale High-Resolution Benchmark for Controllable Virtual Try-On Evaluationarxiv-2601.22725 Sparse Blocked context onlyJan 30, 2026
- AI and My Values: User Perceptions of LLMs' Ability to Extract, Embody, and Explain Human Values from Casual Conversationsarxiv-2601.22440 Sparse Blocked context onlyJan 30, 2026
- MemOCR: Layout-Aware Visual Memory for Efficient Long-Horizon Reasoningarxiv-2601.21468 Sparse Blocked context onlyJan 29, 2026
- INSURE-Dial: A Phase-Aware Conversational Dataset & Benchmark for Compliance Verification and Phase Detectionarxiv-2602.18448 Sparse Blocked context onlyJan 28, 2026
- When Looks Do Not Lie: Discourse Structure Guided In-Context Learning for Faithful Diagram Generationarxiv-2601.20476 Sparse Blocked context onlyJan 28, 2026
- Meta-Cognitive Reinforcement Learning with Self-Doubt and Recoveryarxiv-2601.20193 Sparse Blocked context onlyJan 28, 2026
- Formula-One Prompting: Equation-First Reasoning For Applied Mathematicsarxiv-2601.19302 Sparse Blocked context onlyJan 27, 2026
- FROST: Filtering Reasoning Outliers with Attention for Efficient Reasoningarxiv-2601.19001 Sparse Blocked context onlyJan 26, 2026
- MalURLBench: A Benchmark Evaluating Agents' Vulnerabilities When Processing Web URLsarxiv-2601.18113 Sparse Blocked context onlyJan 26, 2026
- SWE-Pruner: Self-Adaptive Context Pruning for Coding Agentsarxiv-2601.16746 Curated Related Blocked context onlyJan 23, 2026
- Where is the multimodal goal post? On the Ability of Foundation Models to Recognize Contextually Important Momentsarxiv-2601.16333 Sparse Blocked context onlyJan 22, 2026
- RebuttalAgent: Strategic Persuasion in Academic Rebuttal via Theory of Mindarxiv-2601.15715 Sparse Blocked context onlyJan 22, 2026
- Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScorearxiv-2601.15050 Sparse Blocked context onlyJan 21, 2026
- MAS-Orchestra: Understanding and Improving Multi-Agent Reasoning Through Holistic Orchestration and Controlled Benchmarksarxiv-2601.14652 Sparse Blocked context onlyJan 21, 2026
- From Toil to Thought: Designing for Strategic Exploration and Responsible AI in Systematic Literature Reviewsarxiv-2603.05514 Sparse Blocked context onlyJan 21, 2026
- APEX-Agentsarxiv-2601.14242 Sparse Blocked context onlyJan 20, 2026
- Anonpsy: A Graph-Based Framework for Structure-Preserving De-identification of Psychiatric Narrativesarxiv-2601.13503 Sparse Blocked context onlyJan 20, 2026
- SciCoQA: Quality Assurance for Scientific Paper--Code Alignmentarxiv-2601.12910 Sparse Blocked context onlyJan 19, 2026
- Multimodal Multi-Agent Empowered Legal Judgment Predictionarxiv-2601.12815 Sparse Blocked context onlyJan 19, 2026
- Empowering All-in-Loop Health Management of Spacecraft Power System in the Mega-Constellation Era via Human-AI Collaborationarxiv-2601.12667 Sparse Blocked context onlyJan 19, 2026
- Orthogonalized Policy Optimization:Policy Optimization as Orthogonal Projection in Hilbert Spacearxiv-2601.12415 Sparse Blocked context onlyJan 18, 2026
- Vision-as-Inverse-Graphics Agent via Interleaved Multimodal Reasoningarxiv-2601.11109 Sparse Blocked context onlyJan 16, 2026
- Representation-Aware Unlearning via Activation Signatures: From Suppression to Knowledge-Signature Erasurearxiv-2601.10566 Sparse Blocked context onlyJan 15, 2026
- Handling Missing Modalities in Multimodal Survival Prediction for Non-Small Cell Lung Cancerarxiv-2601.10386 Sparse Blocked context onlyJan 15, 2026
- Training-Trajectory-Aware Token Selectionarxiv-2601.10348 Sparse Blocked context onlyJan 15, 2026
- AWED-FiNER: Agents, Web applications, and Expert Detectors for Fine-grained Named Entity Recognition across 36 Languages for 6.6 Billion Speakersarxiv-2601.10161 Sparse Blocked context onlyJan 15, 2026
- Alignment Pretraining: AI Discourse Causes Self-Fulfilling (Mis)alignmentarxiv-2601.10160 Sparse Blocked context onlyJan 15, 2026
- TopoDIM: One-shot Topology Generation of Diverse Interaction Modes for Multi-Agent Systemsarxiv-2601.10120 Sparse Blocked context onlyJan 15, 2026
- SocraticKG: Knowledge Graph Construction via QA-Driven Fact Extractionarxiv-2601.10003 Sparse Blocked context onlyJan 15, 2026
- Frame of Reference: Addressing the Challenges of Common Ground Representation in Situational Dialogsarxiv-2601.09365 Sparse Blocked context onlyJan 14, 2026
- CAST: Character-and-Scene Episodic Memory for Agentsarxiv-2602.06051 Sparse Blocked context onlyJan 14, 2026
- Imagine-then-Plan: Agent Learning from Adaptive Lookahead with World Modelsarxiv-2601.08955 Sparse Blocked context onlyJan 13, 2026
- Auditing Student-AI Collaboration: A Case Study of Online Graduate CS Studentsarxiv-2601.08697 Sparse Blocked context onlyJan 13, 2026
- Rewriting Video: Text-Driven Reauthoring of Video Footagearxiv-2601.08565 Sparse Blocked context onlyJan 13, 2026
- †DAGGER: Distractor-Aware Graph Generation for Executable Reasoning in Math Problemsarxiv-2601.06853 Sparse Blocked context onlyJan 11, 2026
- Mixture-of-Experts as Soft Clustering: A Dual Jacobian-PCA Spectral Geometry Perspectivearxiv-2601.11616 Sparse Blocked context onlyJan 9, 2026
- Distilling Feedback into Memory-as-a-Toolarxiv-2601.05960 Sparse Blocked context onlyJan 9, 2026
- FACTUM: Mechanistic Detection of Citation Hallucination in Long-Form RAGarxiv-2601.05866 Sparse Blocked context onlyJan 9, 2026
- HAG: Hierarchical Demographic Tree-based Agent Generation for Topic-Adaptive Simulationarxiv-2601.05656 Sparse Blocked context onlyJan 9, 2026
- Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelismarxiv-2601.05524 Sparse Blocked context onlyJan 9, 2026
- Miner:Mining Intrinsic Mastery for Data-Efficient RL in Large Reasoning Modelsarxiv-2601.04731 Sparse Blocked context onlyJan 8, 2026
- Neurosymbolic Retrievers for Retrieval-augmented Generationarxiv-2601.04568 Sparse Blocked context onlyJan 8, 2026
- What Matters For Safety Alignment?arxiv-2601.03868 Sparse Blocked context onlyJan 7, 2026
- From Intuition to Calibrated Judgment: A Rubric-Based Expert-Panel Study of Human Detection of LLM-Generated Korean Textarxiv-2601.19913 Sparse Blocked context onlyJan 6, 2026
- Understanding Pure Textual Reasoning for Blind Image Quality Assessmentarxiv-2601.02441 Sparse Blocked context onlyJan 5, 2026
- Supracompetitive Pricing Under AI Monoculturearxiv-2601.01279 Sparse Blocked context onlyJan 3, 2026
- EmoLoom-2B: Fast Base-Model Screening for Emotion Classification and VAD with Lexicon-Weak Supervision and KV-Off Evaluationarxiv-2601.01112 Sparse Blocked context onlyJan 3, 2026
- A Language-Agnostic Hierarchical LoRA-MoE Architecture for CTC-based Multilingual ASRarxiv-2601.00557 Sparse Blocked context onlyJan 2, 2026
- From Evidence-Based Medicine to Knowledge Graph: Retrieval-Augmented Generation for Sports Rehabilitation and a Domain Benchmarkarxiv-2601.00216 Sparse Blocked context onlyJan 1, 2026
- Let It Flow: Agentic Crafting on Rock and Roll, Building the ROME Model within an Open Agentic Learning Ecosystemarxiv-2512.24873 Sparse Blocked context onlyDec 31, 2025
- VL-RouterBench: A Benchmark for Vision-Language Model Routingarxiv-2512.23562 Sparse Blocked context onlyDec 29, 2025
- Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Lossarxiv-2512.23447 Sparse Blocked context onlyDec 29, 2025
- Is Chain-of-Thought Really Not Explainability? Chain-of-Thought Can Be Faithful without Hint Verbalizationarxiv-2512.23032 Sparse Blocked context onlyDec 28, 2025
- Schrödinger's Navigator: Imagining an Ensemble of Futures for Zero-Shot Object Navigationarxiv-2512.21201 Sparse Blocked context onlyDec 24, 2025
- DIAL: Direct Iterative Adversarial Learning for Realistic Multi-Turn Dialogue Simulationarxiv-2512.20773 Sparse Blocked context onlyDec 23, 2025
- Machine Unlearning in the Era of Quantum Machine Learning: An Empirical Studyarxiv-2512.19253 Sparse Blocked context onlyDec 22, 2025
- CycleChart: A Unified Consistency-Based Learning Framework for Bidirectional Chart Understanding and Generationarxiv-2512.19173 Sparse Blocked context onlyDec 22, 2025
- Adaptive Accountability in Networked MAS: Tracing and Mitigating Emergent Norms at Scalearxiv-2512.18561 Sparse Blocked context onlyDec 21, 2025
- Value Under Ignorance in Universal Artificial Intelligencearxiv-2512.17086 Sparse Blocked context onlyDec 18, 2025
- In-Context Algebraarxiv-2512.16902 Sparse Blocked context onlyDec 18, 2025
- Social Story Frames: Contextual Reasoning about Narrative Intent and Receptionarxiv-2512.15925 Sparse Blocked context onlyDec 17, 2025
- Imitation Game: Reproducing Deep Learning Bugs Leveraging an Intelligent Agentarxiv-2512.14990 Sparse Blocked context onlyDec 17, 2025
- CompanionCast: Toward Social Collaboration with Multi-Agent Systems in Shared Experiencesarxiv-2512.10918 Sparse Blocked context onlyDec 11, 2025
- KD-OCT: Efficient Knowledge Distillation for Clinical-Grade Retinal OCT Classificationarxiv-2512.09069 Sparse Blocked context onlyDec 9, 2025
- Group Representational Position Encodingarxiv-2512.07805 Sparse Blocked context onlyDec 8, 2025
- Echo-CoPilot: A Multiple-Perspective Agentic Framework for Reliable Echocardiography Interpretationarxiv-2512.09944 Sparse Blocked context onlyDec 6, 2025
- BOOM: Beyond Only One Modality KIT's Multimodal Multilingual Lecture Companionarxiv-2512.02817 Sparse Blocked context onlyDec 2, 2025
- See, Think, Learn: A Self-Taught Multimodal Reasonerarxiv-2512.02456 Sparse Blocked context onlyDec 2, 2025
- InnoGym: Benchmarking the Innovation Potential of AI Agentsarxiv-2512.01822 Sparse Blocked context onlyDec 1, 2025
- STELLAR: Structure-guided LLM Assertion Retrieval and Generation for Formal Verificationarxiv-2601.19903 Sparse Blocked context onlyNov 28, 2025
- CostNav: A Navigation Benchmark for Real-World Economic-Cost Evaluation of Physical AI Agentsarxiv-2511.20216 Sparse Blocked context onlyNov 25, 2025
- The Metaphysics We Train: A Heideggerian Reading of Machine Learningarxiv-2602.19028 Sparse Blocked context onlyNov 25, 2025
- Foundry: Distilling 3D Foundation Models for the Edgearxiv-2511.20721 Sparse Blocked context onlyNov 25, 2025
- Pedestrian Crossing Intention Prediction Using Multimodal Fusion Networkarxiv-2511.20008 Sparse Blocked context onlyNov 25, 2025
- SAGE: Shape-Adapting Gated Experts for Adaptive Histopathology Image Segmentationarxiv-2511.18493 Sparse Blocked context onlyNov 23, 2025
- Uni-DAD: Unified Distillation and Adaptation of Diffusion Models for Few-step Few-shot Image Generationarxiv-2511.18281 Sparse Blocked context onlyNov 23, 2025
- Bias Is a Subspace, Not a Coordinate: A Geometric Rethinking of Post-hoc Debiasing in Vision-Language Modelsarxiv-2511.18123 Sparse Blocked context onlyNov 22, 2025
- HEAD-QA v2: Expanding a Healthcare Benchmark for Reasoningarxiv-2511.15355 Sparse Blocked context onlyNov 19, 2025
- Stealth Fine-Tuning: Efficiently Breaking Alignment in RVLMs Using Self-Generated CoTarxiv-2511.14106 Sparse Blocked context onlyNov 18, 2025
- AISAC: An Integrated multi-agent System for Transparent, Retrieval-Grounded Scientific Assistancearxiv-2511.14043 Sparse Blocked context onlyNov 18, 2025
- Auditing Google's AI Overviews and Featured Snippets: A Case Study on Baby Care and Pregnancyarxiv-2511.12920 Sparse Blocked context onlyNov 17, 2025
- PRISM of Opinions: A Persona-Reasoned Multimodal Framework for User-centric Conversational Stance Detectionarxiv-2511.12130 Sparse Blocked context onlyNov 15, 2025
- MediRound: Multi-Round Entity-Level Reasoning Segmentation in Medical Imagesarxiv-2511.12110 Sparse Blocked context onlyNov 15, 2025
- From Synthetic Scenes to Real Performance: Enhancing Spatial Reasoning in VLMsarxiv-2511.11440 Sparse Blocked context onlyNov 14, 2025
- On-Device Fine-Tuning via Backprop-Free Zeroth-Order Optimizationarxiv-2511.11362 Sparse Blocked context onlyNov 14, 2025
- Mastering Olympiad-Level Physics with Artificial Intelligencearxiv-2511.10515 Sparse Blocked context onlyNov 13, 2025
- ViPRA: Video Prediction for Robot Actionsarxiv-2511.07732 Sparse Blocked context onlyNov 11, 2025
- Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agentsarxiv-2511.07397 Sparse Blocked context onlyNov 10, 2025
- More Agents Improve Math Problem Solving but Adversarial Robustness Gap Persistsarxiv-2511.07112 Sparse Blocked context onlyNov 10, 2025
- RPTS: Tree-Structured Reasoning Process Scoring for Faithful Multimodal Evaluationarxiv-2511.06899 Sparse Blocked context onlyNov 10, 2025
- VLAD-Grasp: Zero-shot Grasp Detection via Vision-Language Modelsarxiv-2511.05791 Sparse Blocked context onlyNov 8, 2025
- DeepEyesV2: Toward Agentic Multimodal Modelarxiv-2511.05271 Sparse Blocked context onlyNov 7, 2025
- Jr. AI Scientist and Its Risk Report: Autonomous Scientific Exploration from a Baseline Paperarxiv-2511.04583 Sparse Blocked context onlyNov 6, 2025
- Seeing Straight: Document Orientation Detection for Efficient OCRarxiv-2511.04161 Sparse Blocked context onlyNov 6, 2025
- Batch Prompting Suppresses Overthinking Reasoning Under Constraint: How Batch Prompting Suppresses Overthinking in Reasoning Modelsarxiv-2511.04108 Sparse Blocked context onlyNov 6, 2025
- T-FIX: Text-Based Explanations with Features Interpretable to eXpertsarxiv-2511.04070 Sparse Blocked context onlyNov 6, 2025
- Error-Aware Knowledge Distillation via Targeted Revision for Customer-Service Summarizationarxiv-2511.03005 Sparse Blocked context onlyNov 4, 2025
- Efficient Vector Symbolic Architectures from Histogram Recoveryarxiv-2511.01838 Sparse Blocked context onlyNov 3, 2025
- ParlaSpeech 3.0: Richly Annotated Spoken Parliamentary Corpora of Croatian, Czech, Polish, and Serbianarxiv-2511.01619 Sparse Blocked context onlyNov 3, 2025
- Patent Representation Learning via Self-supervisionarxiv-2511.10657 Sparse Blocked context onlyNov 3, 2025
- Self-Harmony: Learning to Harmonize Self-Supervision and Self-Play in Test-Time Reinforcement Learningarxiv-2511.01191 Sparse Blocked context onlyNov 3, 2025
- BEAT: Visual Backdoor Attacks on VLM-based Embodied Agents via Contrastive Trigger Learningarxiv-2510.27623 Sparse Blocked context onlyOct 31, 2025
- Atlas-Alignment: Making Interpretability Transferable Across Language Modelsarxiv-2510.27413 Sparse Blocked context onlyOct 31, 2025
- Simple Additions, Substantial Gains: Expanding Scripts, Languages, and Lineage Coverage in URIEL+arxiv-2510.27183 Sparse Blocked context onlyOct 31, 2025
- From Medical Records to Diagnostic Dialogues: A Clinical-Grounded Approach and Dataset for Psychiatric Comorbidityarxiv-2510.25232 Sparse Blocked context onlyOct 29, 2025
- Repurposing Synthetic Data for Fine-grained Search Agent Supervisionarxiv-2510.24694 Sparse Blocked context onlyOct 28, 2025