300 canonical paper links on this archive page.
- Dataset Distillation via Vision-Language Category Prototypearxiv-2506.23580 Sparse Blocked context onlyJun 1, 2025
- FlashDMoE: Fast Distributed MoE in a Single Kernelarxiv-2506.04667 Sparse Blocked context onlyJun 1, 2025
- Stream-Omni: Simultaneous Multimodal Interactions with Large Language-Vision-Speech Modelarxiv-2506.13642 Sparse Blocked context onlyJun 1, 2025
- Online Fair Division with Additional Informationarxiv-2505.24503 Sparse Blocked context onlyMay 30, 2025
- StressTest: Can YOUR Speech LM Handle the Stress?arxiv-2505.22765 Sparse Blocked context onlyMay 28, 2025
- RedTeamCUA: Realistic Adversarial Testing of Computer-Use Agents in Hybrid Web-OS Environmentsarxiv-2505.21936 Sparse Blocked context onlyMay 28, 2025
- How Does Alignment Enhance LLMs' Multilingual Capabilities? A Language Neurons Perspectivearxiv-2505.21505 Sparse Blocked context onlyMay 27, 2025
- When Slower Isn't Truer: Inverse Scaling Law of Truthfulness in Multimodal Reasoningarxiv-2505.20214 Sparse Blocked context onlyMay 26, 2025
- Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPOarxiv-2505.19770 Sparse Blocked context onlyMay 26, 2025
- Consistency-based Abductive Reasoning over Perceptual Errors of Multiple Pre-trained Models in Novel Environmentsarxiv-2505.19361 Sparse Blocked context onlyMay 25, 2025
- Evaluation Faking: Unveiling Observer Effects in Safety Evaluation of Frontier AI Systemsarxiv-2505.17815 Sparse Blocked context onlyMay 23, 2025
- Dynamic Token Reweighting for Robust Vision-Language Modelsarxiv-2505.17132 Sparse Blocked context onlyMay 22, 2025
- Efficient PRM Training Data Synthesis via Formal Verificationarxiv-2505.15960 Sparse Blocked context onlyMay 21, 2025
- ALIEN: Aligned Entropy Head for Improving Uncertainty Estimation of LLMsarxiv-2505.15443 Sparse Blocked context onlyMay 21, 2025
- Guided Policy Optimization under Partial Observabilityarxiv-2505.15418 Sparse Blocked context onlyMay 21, 2025
- A quantitative analysis of semantic information in deep representations of text and imagesarxiv-2505.17101 Sparse Blocked context onlyMay 21, 2025
- Language Models use Lookbacks to Track Beliefsarxiv-2505.14685 Sparse Blocked context onlyMay 20, 2025
- RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understandingarxiv-2505.14462 Sparse Blocked context onlyMay 20, 2025
- Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMsarxiv-2505.14226 Sparse Blocked context onlyMay 20, 2025
- Word length predicts word order: "Min-max"-ing drives language evolutionarxiv-2505.13913 Sparse Blocked context onlyMay 20, 2025
- Shorten After You're Right: Lazy Length Penalties for Reasoning RLarxiv-2505.12284 Sparse Blocked context onlyMay 18, 2025
- BARREL: Boundary-Aware Reasoning for Factual and Reliable LRMsarxiv-2505.13529 Sparse Blocked context onlyMay 18, 2025
- The Counting Power of Transformersarxiv-2505.11199 Sparse Blocked context onlyMay 16, 2025
- Why 1 + 1 < 1 in Visual Token Pruning: Beyond Naive Integration via Multi-Objective Balanced Coveringarxiv-2505.10118 Sparse Blocked context onlyMay 15, 2025
- Multi-Domain Audio Question Answering Benchmark Toward Acoustic Content Reasoningarxiv-2505.07365 Sparse Blocked context onlyMay 12, 2025
- Mastering Multi-Drone Volleyball through Hierarchical Co-Self-Play Reinforcement Learningarxiv-2505.04317 Sparse Blocked context onlyMay 7, 2025
- Adaptive Social Learning via Mode Policy Optimization for Language Agentsarxiv-2505.02156 Sparse Blocked context onlyMay 4, 2025
- Decoding Open-Ended Information Seeking Goals from Eye Movements in Readingarxiv-2505.02872 Sparse Blocked context onlyMay 4, 2025
- R&D-Agent-Quant: A Multi-Agent Framework for Data-Centric Factors and Model Joint Optimizationarxiv-2505.15155 Sparse Blocked context onlyMay 1, 2025
- LiteCUA: Computer as MCP Server for Computer-Use Agent on AIOSarxiv-2505.18829 Sparse Blocked context onlyMay 1, 2025
- MM-PRM: Enhancing Multimodal Mathematical Reasoning with Scalable Step-Level Supervisionarxiv-2505.13427 Sparse Blocked context onlyMay 1, 2025
- Efficient Multivariate Time Series Forecasting via Calibrated Language Models with Privileged Knowledge Distillationarxiv-2505.02138 Sparse Blocked context onlyMay 1, 2025
- The P$^3$ dataset: Pixels, Points and Polygons for Multimodal Building Vectorizationarxiv-2505.15379 Curated Related Blocked context onlyMay 1, 2025
- PaTH Attention: Position Encoding via Accumulating Householder Transformationsarxiv-2505.16381 Sparse Blocked context onlyMay 1, 2025
- Reshaping MOFs text mining with a dynamic multi-agents framework of large language modelarxiv-2504.18880 Sparse Blocked context onlyApr 26, 2025
- A closer look at how large language models trust humans: patterns and biasesarxiv-2504.15801 Sparse Blocked context onlyApr 22, 2025
- Leakage and Interpretability in Concept-Based Modelsarxiv-2504.14094 Sparse Blocked context onlyApr 18, 2025
- Scalable Multi-Task Learning through Spiking Neural Networks with Adaptive Task-Switching Policy for Intelligent Autonomous Agentsarxiv-2504.13541 Sparse Blocked context onlyApr 18, 2025
- Cost-of-Pass: An Economic Framework for Evaluating Language Modelsarxiv-2504.13359 Sparse Blocked context onlyApr 17, 2025
- Adaptive Insurance Reserving with CVaR-Constrained Reinforcement Learning under Macroeconomic Regimesarxiv-2504.09396 Sparse Blocked context onlyApr 13, 2025
- Med-R2: Perception and Reflection-driven Complex Reasoning for Medical Report Generationarxiv-2504.02885 Sparse Blocked context onlyApr 2, 2025
- REAL: Benchmarking Autonomous Agents on Deterministic Simulations of Real Websitesarxiv-2504.11543 Sparse Blocked context onlyApr 1, 2025
- InfiGUI-R1: Advancing Multimodal GUI Agents from Reactive Actors to Deliberative Reasonersarxiv-2504.14239 Sparse Blocked context onlyApr 1, 2025
- Genius: A Generalizable and Purely Unsupervised Self-Training Framework For Advanced Reasoningarxiv-2504.08672 Sparse Blocked context onlyApr 1, 2025
- Data-efficient LLM Fine-tuning for Code Generationarxiv-2504.12687 Curated Related Blocked context onlyApr 1, 2025
- SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvementarxiv-2504.07934 Sparse Blocked context onlyApr 1, 2025
- When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?arxiv-2503.23137 Sparse Blocked context onlyMar 29, 2025
- More Bang for the Buck: Process Reward Modeling with Entropy-Driven Uncertaintyarxiv-2503.22233 Sparse Blocked context onlyMar 28, 2025
- Survey on Evaluation of LLM-based Agentsarxiv-2503.16416 Sparse Blocked context onlyMar 20, 2025
- Measuring AI Ability to Complete Long Software Tasksarxiv-2503.14499 Sparse Blocked context onlyMar 18, 2025
- Implicit Bias-Like Patterns in Reasoning Modelsarxiv-2503.11572 Sparse Blocked context onlyMar 14, 2025
- Unicorn: A Universal and Collaborative Reinforcement Learning Approach Towards Generalizable Network-Wide Traffic Signal Controlarxiv-2503.11488 Sparse Blocked context onlyMar 14, 2025
- Reasoning-Grounded Natural Language Explanations for Language Modelsarxiv-2503.11248 Sparse Blocked context onlyMar 14, 2025
- TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Toolsarxiv-2503.10970 Sparse Blocked context onlyMar 14, 2025
- VQEL: Enabling Self-Play in Emergent Language Games via Agent-Internal Vector Quantizationarxiv-2503.04940 Sparse Blocked context onlyMar 6, 2025
- Training-free Adjustable Polynomial Graph Filtering for Ultra-fast Multimodal Recommendationarxiv-2503.04406 Sparse Blocked context onlyMar 6, 2025
- LINGOLY-TOO: Disentangling Reasoning from Knowledge with Templatised Orthographic Obfuscationarxiv-2503.02972 Sparse Blocked context onlyMar 4, 2025
- Depth-Width tradeoffs in Algorithmic Reasoning of Graph Tasks with Transformersarxiv-2503.01805 Sparse Blocked context onlyMar 3, 2025
- Cerebrum (AIOS SDK): A Platform for Agent Development, Deployment, Distribution, and Discoveryarxiv-2503.11444 Sparse Blocked context onlyMar 1, 2025
- Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcementarxiv-2503.06520 Sparse Blocked context onlyMar 1, 2025
- Cosmos-Transfer1: Conditional World Generation with Adaptive Multimodal Controlarxiv-2503.14492 Sparse Blocked context onlyMar 1, 2025
- Derm1M: A Million-scale Vision-Language Dataset Aligned with Clinical Ontology Knowledge for Dermatologyarxiv-2503.14911 Sparse Blocked context onlyMar 1, 2025
- Urban Emergency Rescue Based on Multi-Agent Collaborative Learning: Coordination Between Fire Engines and Traffic Lightsarxiv-2502.16131 Sparse Blocked context onlyFeb 22, 2025
- Moving Beyond Medical Exams: A Clinician-Annotated Fairness Dataset of Real-World Tasks and Ambiguity in Mental Healthcarearxiv-2502.16051 Sparse Blocked context onlyFeb 22, 2025
- Integrating Personality into Digital Humans: A Review of LLM-Driven Approaches for Virtual Realityarxiv-2503.16457 Sparse Blocked context onlyFeb 22, 2025
- From Restless to Contextual: A Thresholding Bandit Reformulation For Finite-horizon Improvementarxiv-2502.05145 Sparse Blocked context onlyFeb 7, 2025
- OpenSTARLab: Open Approach for Spatio-Temporal Agent Data Analysis in Soccerarxiv-2502.02785 Sparse Blocked context onlyFeb 5, 2025
- VolleyBots: A Testbed for Multi-Drone Volleyball Game Combining Motion Control and Strategic Playarxiv-2502.01932 Sparse Blocked context onlyFeb 4, 2025
- GRADIEND: Feature Learning within Neural Networks Exemplified through Biasesarxiv-2502.01406 Sparse Blocked context onlyFeb 3, 2025
- Dialogue is Better Than Monologue: Instructing Medical LLMs via Strategical Conversationsarxiv-2501.17860 Sparse Blocked context onlyJan 29, 2025
- Safe Reinforcement Learning for Real-World Engine Controlarxiv-2501.16613 Sparse Blocked context onlyJan 28, 2025
- CowPilot: A Framework for Autonomous and Human-Agent Collaborative Web Navigationarxiv-2501.16609 Sparse Blocked context onlyJan 28, 2025
- Object-Centric World Models from Few-Shot Annotations for Sample-Efficient Reinforcement Learningarxiv-2501.16443 Sparse Blocked context onlyJan 27, 2025
- Explainable Multimodal Depression Recognition in Clinical Interviews via PHQ-Aligned Symptom Summarizationarxiv-2501.16106 Sparse Blocked context onlyJan 27, 2025
- CoverM: Read alignment statistics for metagenomicsarxiv-2501.11217 Sparse Blocked context onlyJan 20, 2025
- Cosmos World Foundation Model Platform for Physical AIdoi-10.48550_arxiv.2501.03575 Sparse Blocked context onlyJan 7, 2025
- A Survey of State of the Art Large Vision Language Models: Alignment, Benchmark, Evaluations and Challengesarxiv-2501.02189 Sparse Blocked context onlyJan 1, 2025
- BIOMEDICA: An Open Biomedical Image-Caption Archive, Dataset, and Vision-Language Models Derived from Scientific Literaturearxiv-2501.07171 Sparse Blocked context onlyJan 1, 2025
- MapEval: A Map-Based Evaluation of Geo-Spatial Reasoning in Foundation Modelsdoi-10.48550_arxiv.2501.00316 Sparse Blocked context onlyDec 31, 2024
- Is Contrastive Distillation Enough for Learning Comprehensive 3D Representations?arxiv-2412.08973 Sparse Blocked context onlyDec 12, 2024
- VisionZip: Longer is Better but Not Necessary in Vision Language Modelsarxiv-2412.04467 Sparse Blocked context onlyDec 5, 2024
- Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Searcharxiv-2412.18319 Sparse Blocked context onlyDec 1, 2024
- Training Software Engineering Agents and Verifiers with SWE-Gymarxiv-2412.21139 Sparse Blocked context onlyDec 1, 2024
- Context Clues: Evaluating Long Context Models for Clinical Prediction Tasks on EHRsarxiv-2412.16178 Sparse Blocked context onlyDec 1, 2024
- Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systemsarxiv-2412.09413 Sparse Blocked context onlyDec 1, 2024
- The Limits of Inference Scaling Through Resamplingarxiv-2411.17501 Sparse Blocked context onlyNov 26, 2024
- RoboSpatial: Teaching Spatial Understanding to 2D and 3D Vision-Language Models for Roboticsarxiv-2411.16537 Sparse Blocked context onlyNov 25, 2024
- Phrase-Instance Alignment for Generalized Referring Segmentationarxiv-2411.15087 Sparse Blocked context onlyNov 22, 2024
- Personalized Help for Optimizing Low-Skilled Users' Strategyarxiv-2411.09109 Sparse Blocked context onlyNov 14, 2024
- Renaissance: Investigating the Pretraining of Vision-Language Encodersarxiv-2411.06657 Sparse Blocked context onlyNov 11, 2024
- Probing the limitations of multimodal language models for chemistry and materials researcharxiv-2411.16955 Sparse Blocked context onlyNov 1, 2024
- Multimodal Whole Slide Foundation Model for Pathologyarxiv-2411.19666 Sparse Blocked context onlyNov 1, 2024
- Generalizable Person Re-identification via Balancing Alignment and Uniformityarxiv-2411.11471 Sparse Blocked context onlyNov 1, 2024
- MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMsarxiv-2411.15296 Sparse Blocked context onlyNov 1, 2024
- RealCQA-V2: A Diagnostic Benchmark for Structured Visual Entailment over Scientific Chartsarxiv-2410.22492 Sparse Blocked context onlyOct 29, 2024
- AFlow: Automating Agentic Workflow Generationarxiv-2410.10762 Sparse Blocked context onlyOct 1, 2024
- Distill Visual Chart Reasoning Ability from LLMs to MLLMsarxiv-2410.18798 Sparse Blocked context onlyOct 1, 2024
- Pantograph: A Machine-to-Machine Interaction Interface for Advanced Theorem Proving, High Level Reasoning, and Data Extraction in Lean 4arxiv-2410.16429 Sparse Blocked context onlyOct 1, 2024
- Gödel Agent: A Self-Referential Agent Framework for Recursive Self-Improvementarxiv-2410.04444 Sparse Blocked context onlyOct 1, 2024
- MLE-bench: Evaluating Machine Learning Agents on Machine Learning Engineeringarxiv-2410.07095 Sparse Blocked context onlyOct 1, 2024
- Multimodality Helps Few-Shot 3D Point Cloud Semantic Segmentationarxiv-2410.22489 Sparse Blocked context onlyOct 1, 2024
- OS-ATLAS: A Foundation Action Model for Generalist GUI Agentsarxiv-2410.23218 Sparse Blocked context onlyOct 1, 2024
- Toward Guidance-Free AR Visual Generation via Condition Contrastive Alignmentarxiv-2410.09347 Sparse Blocked context onlyOct 1, 2024
- Many Heads Are Better Than One: Improved Scientific Idea Generation by A LLM-Based Multi-Agent Systemarxiv-2410.09403 Sparse Blocked context onlyOct 1, 2024
- SelfCodeAlign: Self-Alignment for Code Generationarxiv-2410.24198 Sparse Blocked context onlyOct 1, 2024
- AutoPenBench: Benchmarking Generative Agents for Penetration Testingarxiv-2410.03225 Sparse Blocked context onlyOct 1, 2024
- One-Step Diffusion Distillation through Score Implicit Matchingarxiv-2410.16794 Sparse Blocked context onlyOct 1, 2024
- PACE: Procedural Abstractions for Communicating Efficientlyarxiv-2409.20120 Sparse Blocked context onlySep 30, 2024
- Human-like Affective Cognition in Foundation Modelsarxiv-2409.11733 Sparse Blocked context onlySep 18, 2024
- IGEV++: Iterative Multi-range Geometry Encoding Volumes for Stereo Matchingarxiv-2409.00638 Sparse Blocked context onlySep 1, 2024
- Parallel AutoRegressive Models for Multi-Agent Combinatorial Optimizationarxiv-2409.03811 Sparse Blocked context onlySep 1, 2024
- Maia-2: A Unified Model for Human-AI Alignment in Chessarxiv-2409.20553 Sparse Blocked context onlySep 1, 2024
- EnIGMA: Enhanced Interactive Generative Model Agent for CTF Challengesarxiv-2409.16165 Sparse Blocked context onlySep 1, 2024
- MediConfusion: Can you trust your AI radiologist? Probing the reliability of multimodal medical foundation modelsarxiv-2409.15477 Sparse Blocked context onlySep 1, 2024
- Abstracted Gaussian Prototypes for True One-Shot Concept Learningarxiv-2408.17251 Sparse Blocked context onlyAug 30, 2024
- Bidirectional Decoding: Improving Action Chunking via Guided Test-Time Samplingarxiv-2408.17355 Sparse Blocked context onlyAug 1, 2024
- LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMsarxiv-2408.07055 Curated Related Blocked context onlyAug 1, 2024
- Lumina-mGPT: Illuminate Flexible Photorealistic Text-to-Image Generation with Multimodal Generative Pretrainingarxiv-2408.02657 Sparse Blocked context onlyAug 1, 2024
- LLMs as Zero-shot Graph Learners: Alignment of GNN Representations with LLM Token Embeddingsarxiv-2408.14512 Sparse Blocked context onlyAug 1, 2024
- LiCoEval: Evaluating LLMs on License Compliance in Code Generationarxiv-2408.02487 Curated Related Blocked context onlyAug 1, 2024
- Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encodersarxiv-2408.15998 Sparse Blocked context onlyAug 1, 2024
- SigmaRL: A Sample-Efficient and Generalizable Multi-Agent Reinforcement Learning Framework for Motion Planningarxiv-2408.07644 Sparse Blocked context onlyAug 1, 2024
- Mutual Reasoning Makes Smaller LLMs Stronger Problem-Solversarxiv-2408.06195 Sparse Blocked context onlyAug 1, 2024
- An Autonomous GIS Agent Framework for Geospatial Data Retrievalarxiv-2407.21024 Sparse Blocked context onlyJul 1, 2024
- OpenHands: An Open Platform for AI Software Developers as Generalist Agentsarxiv-2407.16741 Sparse Blocked context onlyJul 1, 2024
- NeedleBench: Can LLMs Do Retrieval and Reasoning in Information-Dense Context?arxiv-2407.11963 Sparse Blocked context onlyJul 1, 2024
- FinCon: A Synthesized LLM Multi-Agent System with Conceptual Verbal Reinforcement for Enhanced Financial Decision Makingarxiv-2407.06567 Sparse Blocked context onlyJul 1, 2024
- Agentless: Demystifying LLM-based Software Engineering Agentsarxiv-2407.01489 Sparse Blocked context onlyJul 1, 2024
- Measuring the Measurers: Quality Evaluation of Hallucination Benchmarks for Large Vision-Language Modelsarxiv-2406.17115 Sparse Blocked context onlyJun 24, 2024
- Improving Alignment and Robustness with Circuit Breakersarxiv-2406.04313 Sparse Blocked context onlyJun 6, 2024
- Unpacking DPO and PPO: Disentangling Best Practices for Learning from Preference Feedbackarxiv-2406.09279 Sparse Blocked context onlyJun 1, 2024
- BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystackarxiv-2406.10149 Sparse Blocked context onlyJun 1, 2024
- TRACE the Evidence: Constructing Knowledge-Grounded Reasoning Chains for Retrieval-Augmented Generationarxiv-2406.11460 Sparse Blocked context onlyJun 1, 2024
- MARS: Benchmarking the Metaphysical Reasoning Abilities of Language Models with a Multi-task Evaluation Datasetarxiv-2406.02106 Sparse Blocked context onlyJun 1, 2024
- AutoHallusion: Automatic Generation of Hallucination Benchmarks for Vision-Language Modelsarxiv-2406.10900 Sparse Blocked context onlyJun 1, 2024
- SEACrowd: A Multilingual Multimodal Data Hub and Benchmark Suite for Southeast Asian Languagesarxiv-2406.10118 Sparse Blocked context onlyJun 1, 2024
- Self-Distillation Prototypes Network: Learning Robust Speaker Representations without Supervisionarxiv-2406.11169 Sparse Blocked context onlyJun 1, 2024
- Mobile-Agent-v2: Mobile Device Operation Assistant with Effective Navigation via Multi-Agent Collaborationarxiv-2406.01014 Sparse Blocked context onlyJun 1, 2024
- MediQ: Question-Asking LLMs and a Benchmark for Reliable Interactive Clinical Reasoningarxiv-2406.00922 Sparse Blocked context onlyJun 1, 2024
- CoAct: A Global-Local Hierarchy for Autonomous Agent Collaborationarxiv-2406.13381 Sparse Blocked context onlyJun 1, 2024
- Are We Done with MMLU?arxiv-2406.04127 Sparse Blocked context onlyJun 1, 2024
- LLaMA-MoE: Building Mixture-of-Experts from LLaMA with Continual Pre-trainingarxiv-2406.16554 Sparse Blocked context onlyJun 1, 2024
- Unified Auto-Encoding with Masked Diffusionarxiv-2406.17688 Sparse Blocked context onlyJun 1, 2024
- The CLRS-Text Algorithmic Reasoning Language Benchmarkarxiv-2406.04229 Sparse Blocked context onlyJun 1, 2024
- Augmenting Lateral Thinking in Language Models with Humor and Riddle Data for the BRAINTEASER Taskarxiv-2405.10385 Sparse Blocked context onlyMay 16, 2024
- USP: A Unified Sequence Parallelism Approach for Long Context Generative AIarxiv-2405.07719 Sparse Blocked context onlyMay 1, 2024
- ACEGEN: Reinforcement learning of generative chemical agents for drug discoveryarxiv-2405.04657 Sparse Blocked context onlyMay 1, 2024
- Vista: A Generalizable Driving World Model with High Fidelity and Versatile Controllabilityarxiv-2405.17398 Sparse Blocked context onlyMay 1, 2024
- Group Robust Preference Optimization in Reward-free RLHFarxiv-2405.20304 Sparse Blocked context onlyMay 1, 2024
- Monte Carlo Tree Search Boosts Reasoning via Iterative Preference Learningarxiv-2405.00451 Sparse Blocked context onlyMay 1, 2024
- Xmodel-VLM: A Simple Baseline for Multimodal Vision Language Modelarxiv-2405.09215 Sparse Blocked context onlyMay 1, 2024
- DAPE: Data-Adaptive Positional Encoding for Length Extrapolationarxiv-2405.14722 Sparse Blocked context onlyMay 1, 2024
- XTrack: Multimodal Training Boosts RGB-X Video Object Trackersarxiv-2405.17773 Sparse Blocked context onlyMay 1, 2024
- MiniGPT4-Video: Advancing Multimodal LLMs for Video Understanding with Interleaved Visual-Textual Tokensarxiv-2404.03413 Sparse Blocked context onlyApr 1, 2024
- DeiT-LT Distillation Strikes Back for Vision Transformer Training on Long-Tailed Datasetsarxiv-2404.02900 Sparse Blocked context onlyApr 1, 2024
- InternLM-XComposer2-4KHD: A Pioneering Large Vision-Language Model Handling Resolutions from 336 Pixels to 4K HDarxiv-2404.06512 Sparse Blocked context onlyApr 1, 2024
- Photo-Realistic Image Restoration in the Wild with Controlled Vision-Language Modelsarxiv-2404.09732 Curated Related Blocked context onlyApr 1, 2024
- MER 2024: Semi-Supervised Learning, Noise Robustness, and Open-Vocabulary Multimodal Emotion Recognitionarxiv-2404.17113 Sparse Blocked context onlyApr 1, 2024
- LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Modelarxiv-2404.01331 Sparse Blocked context onlyApr 1, 2024
- Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answeringarxiv-2403.04890 Sparse Blocked context onlyMar 7, 2024
- Visual Decoding and Reconstruction via EEG Embeddings with Guided Diffusionarxiv-2403.07721 Sparse Blocked context onlyMar 1, 2024
- Breaking the HISCO Barrier: Automatic Occupational Standardization with OccCANINEarxiv-2402.13604 Sparse Blocked context onlyFeb 21, 2024
- Multi-agent deep reinforcement learning with centralized training and decentralized execution for transportation infrastructure managementarxiv-2401.12455 Sparse Blocked context onlyJan 23, 2024
- MLLM-Tool: A Multimodal Large Language Model For Tool Agent Learningarxiv-2401.10727 Sparse Blocked context onlyJan 1, 2024
- RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedbackarxiv-2312.00849 Sparse Blocked context onlyDec 1, 2023
- Mamba: Linear-Time Sequence Modeling with Selective State Spacesdoi-10.48550_arxiv.2312.00752 Sparse Blocked context onlyDec 1, 2023
- CogAgent: A Visual Language Model for GUI Agentsarxiv-2312.08914 Sparse Blocked context onlyDec 1, 2023
- MetaTrinity: Enabling Fast Metagenomic Classification via Seed Counting and Edit Distance Approximationdoi-10.48550_arxiv.2311.02029 Sparse Blocked context onlyNov 3, 2023
- GPQA: A Graduate-Level Google-Proof Q&A Benchmarkarxiv-2311.12022 Sparse Blocked context onlyNov 1, 2023
- Adversarial Diffusion Distillationarxiv-2311.17042 Sparse Blocked context onlyNov 1, 2023
- Mitigating Object Hallucinations in Large Vision-Language Models through Visual Contrastive Decodingarxiv-2311.16922 Sparse Blocked context onlyNov 1, 2023
- FinMem: A Performance-Enhanced LLM Trading Agent with Layered Memory and Character Designarxiv-2311.13743 Sparse Blocked context onlyNov 1, 2023
- Llemma: An Open Language Model For Mathematicsarxiv-2310.10631 Sparse Blocked context onlyOct 16, 2023
- Llemma: An Open Language Model For Mathematicsdoi-10.48550_arxiv.2310.10631 Sparse Blocked context onlyOct 16, 2023
- HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Modelsarxiv-2310.14566 Sparse Blocked context onlyOct 1, 2023
- L2MAC: Large Language Model Automatic Computer for Extensive Code Generationarxiv-2310.02003 Sparse Blocked context onlyOct 1, 2023
- Efficient Streaming Language Models with Attention Sinksdoi-10.48550_arxiv.2309.17453 Sparse Blocked context onlySep 29, 2023
- Event Stream-based Visual Object Tracking: A High-Resolution Benchmark Dataset and A Novel Baselinearxiv-2309.14611 Sparse Blocked context onlySep 26, 2023
- Bespoke Nanoparticle Synthesis and Chemical Knowledge Discovery Via Autonomous Experimentationsdoi-10.48550_arxiv.2309.00349 Sparse Blocked context onlySep 1, 2023
- A Survey on Large Language Model based Autonomous Agentsarxiv-2308.11432 Sparse Blocked context onlyAug 1, 2023
- DNABERT-2: Efficient Foundation Model and Benchmark For Multi-Species Genomedoi-10.48550_arxiv.2306.15006 Sparse Blocked context onlyJun 26, 2023
- RedMotion: Motion Prediction via Redundancy Reductionarxiv-2306.10840 Sparse Blocked context onlyJun 19, 2023
- RedMotion: Motion Prediction via Redundancy Reductiondoi-10.48550_arxiv.2306.10840 Sparse Blocked context onlyJun 19, 2023
- SugarCrepe: Fixing Hackable Benchmarks for Vision-Language Compositionalityarxiv-2306.14610 Sparse Blocked context onlyJun 1, 2023
- Language Models Can Improve Event Prediction by Few-Shot Abductive Reasoningarxiv-2305.16646 Sparse Blocked context onlyMay 1, 2023
- ChemCrow: Augmenting large-language models with chemistry toolsdoi-10.48550_arxiv.2304.05376 Sparse Blocked context onlyApr 11, 2023
- Generative Agents: Interactive Simulacra of Human Behaviorarxiv-2304.03442 Sparse Blocked context onlyApr 1, 2023
- LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attentiondoi-10.48550_arxiv.2303.16199 Sparse Blocked context onlyMar 28, 2023
- Neuro-symbolic Commonsense Social Reasoningdoi-10.48550_arxiv.2303.08264 Sparse Blocked context onlyMar 14, 2023
- Tag2Text: Guiding Vision-Language Model via Image Taggingarxiv-2303.05657 Sparse Blocked context onlyMar 1, 2023
- Controllability-Aware Unsupervised Skill Discoverydoi-10.48550_arxiv.2302.05103 Sparse Blocked context onlyFeb 10, 2023
- SantaCoder: don't reach for the stars!doi-10.48550_arxiv.2301.03988 Sparse Blocked context onlyJan 9, 2023
- Training language models to summarize narratives improves brain alignmentarxiv-2212.10898 Sparse Blocked context onlyDec 1, 2022
- An Optimal Transport-driven Approach for Cultivating Latent Space in Online Incremental Learningarxiv-2211.16780 Sparse Blocked context onlyNov 30, 2022
- Red Teaming with Mind Reading: White-Box Adversarial Policies Against RL Agentsarxiv-2209.02167 Sparse Blocked context onlySep 1, 2022
- Transformers are Sample-Efficient World Modelsarxiv-2209.00588 Sparse Blocked context onlySep 1, 2022
- Lossless Acceleration for Seq2seq Generation with Aggressive Decodingarxiv-2205.10350 Sparse Blocked context onlyMay 1, 2022
- HDGT: Heterogeneous Driving Graph Transformer for Multi-Agent Trajectory Prediction via Scene Encodingarxiv-2205.09753 Sparse Blocked context onlyMay 1, 2022
- Vision-Language Pre-Training for Boosting Scene Text Detectorsarxiv-2204.13867 Sparse Blocked context onlyApr 1, 2022
- Instant Neural Graphics Primitives with a Multiresolution Hash Encodingarxiv-2201.05989 Sparse Blocked context onlyJan 16, 2022
- Image Captioning via Compact Bidirectional Architecturearxiv-2201.01984 Sparse Blocked context onlyJan 6, 2022
- BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generationarxiv-2201.12086 Sparse Blocked context onlyJan 1, 2022
- DeepSpeed-MoE: Advancing Mixture-of-Experts Inference and Training to Power Next-Generation AI Scalearxiv-2201.05596 Sparse Blocked context onlyJan 1, 2022
- Screen2Words: Automatic Mobile UI Summarization with Multimodal Learningarxiv-2108.03353 Sparse Blocked context onlyAug 1, 2021
- Align before Fuse: Vision and Language Representation Learning with Momentum Distillationarxiv-2107.07651 Sparse Blocked context onlyJul 1, 2021
- Spatial Graph Attention and Curiosity-driven Policy for Antiviral Drug Discoverydoi-10.48550_arxiv.2106.02190 Sparse Blocked context onlyJun 4, 2021
- Deep Multi-agent Reinforcement Learning for Highway On-Ramp Merging in Mixed Trafficarxiv-2105.05701 Sparse Blocked context onlyMay 12, 2021
- Towards mental time travel: a hierarchical memory for reinforcement learning agentsarxiv-2105.14039 Sparse Blocked context onlyMay 1, 2021
- LayoutXLM: Multimodal Pre-training for Multilingual Visually-rich Document Understandingarxiv-2104.08836 Sparse Blocked context onlyApr 1, 2021
- QA-GNN: Reasoning with Language Models and Knowledge Graphs for Question Answeringarxiv-2104.06378 Sparse Blocked context onlyApr 1, 2021
- Class-Balanced Distillation for Long-Tailed Visual Recognitionarxiv-2104.05279 Sparse Blocked context onlyApr 1, 2021
- Combining Visual and Textual Features for Semantic Segmentation of Historical Newspapersdoi-10.5167_uzh-200811 Sparse Blocked context onlyJan 1, 2021
- Algorithms for Causal Reasoning in Probability Treesarxiv-2010.12237 Sparse Blocked context onlyOct 1, 2020
- Benchmarking Multi-Agent Deep Reinforcement Learning Algorithms in Cooperative Tasksdoi-10.48550_arxiv.2006.07869 Sparse Blocked context onlyJun 14, 2020
- Shared Experience Actor-Critic for Multi-Agent Reinforcement Learningarxiv-2006.07169 Sparse Blocked context onlyJun 1, 2020
- Improving Policies via Search in Cooperative Partially Observable Gamesdoi-10.1609_aaai.v34i05.6208 Sparse Blocked context onlyApr 3, 2020
- Soft-Label Dataset Distillation and Text Dataset Distillationarxiv-1910.02551 Sparse Blocked context onlyOct 1, 2019
- Machine Learning and System Identification for Estimation in Physical Systemsdoi-10.48550_arxiv.1906.02003 Sparse Blocked context onlyJun 5, 2019
- Stacked Dense U-Nets with Dual Transformers for Robust Face Alignmentarxiv-1812.01936 Sparse Blocked context onlyDec 1, 2018
- Unity: A General Platform for Intelligent Agentsarxiv-1809.02627 Sparse Blocked context onlySep 1, 2018
- Analogical Reasoning on Chinese Morphological and Semantic Relationsarxiv-1805.06504 Sparse Blocked context onlyMay 1, 2018
- Automatic Chemical Design Using a Data-Driven Continuous Representation of Moleculesdoi-10.1021_acscentsci.7b00572 Sparse Blocked context onlyJan 12, 2018
- DeepPath: A Reinforcement Learning Method for Knowledge Graph Reasoningarxiv-1707.06690 Sparse Blocked context onlyJul 1, 2017
- The StarCraft Multi-Agent Challengeopenalex-W4288594419 Sparse Blocked context onlyAug 22, 2026
- HFX: Joint Design of Algorithms and Systems for Multi-SLO Serving and Fast Scalingarxiv-2508.15919 Sparse Blocked context onlyAug 21, 2025
- Classification errors distort findings in automated speech processing: examples and solutions from child-development researcharxiv-2508.15637 Sparse Blocked context onlyAug 21, 2025
- HebID: Detecting Social Identities in Hebrew-language Political Textarxiv-2508.15483 Sparse Blocked context onlyAug 21, 2025
- Can synthetic data reproduce real-world findings in epidemiology? A replication study using adversarial random forestsarxiv-2508.14936 Sparse Blocked context onlyAug 19, 2025
- Interactive Query Answering on Knowledge Graphs with Soft Entity Constraintsarxiv-2508.13663 Sparse Blocked context onlyAug 19, 2025
- SNAP-UQ: Self-supervised Next-Activation Prediction for Single-Pass Uncertainty in TinyMLarxiv-2508.12907 Sparse Blocked context onlyAug 18, 2025
- SEA-BED: How Do Embedding Models Represent Southeast Asian Languages?arxiv-2508.12243 Sparse Blocked context onlyAug 17, 2025
- Singing Syllabi with Virtual Avatars: Enhancing Student Engagement Through AI-Generated Music and Digital Embodimentarxiv-2508.11872 Sparse Blocked context onlyAug 16, 2025
- Role-Augmented Intent-Driven Generative Search Engine Optimizationarxiv-2508.11158 Sparse Blocked context onlyAug 15, 2025
- UniPrompt-CL: Sustainable Continual Learning in Medical AI with Unified Prompt Poolsarxiv-2508.10954 Sparse Blocked context onlyAug 14, 2025
- CATNet: A geometric deep learning approach for CAT bond spread prediction in the primary marketarxiv-2508.10208 Sparse Blocked context onlyAug 13, 2025
- UbiQTree: Uncertainty Quantification in XAI with Tree Ensemblesarxiv-2508.09639 Sparse Blocked context onlyAug 13, 2025
- Shadow in the Cache: Unveiling and Mitigating Privacy Risks of KV-cache in LLM Inferencearxiv-2508.09442 Sparse Blocked context onlyAug 13, 2025
- Link Prediction for Event Logs in the Process Industryarxiv-2508.09096 Sparse Blocked context onlyAug 12, 2025
- Position: Beyond Sensitive Attributes, ML Fairness Should Quantify Structural Injustice via Social Determinantsarxiv-2508.08337 Sparse Blocked context onlyAug 10, 2025
- From Product Hilbert Spaces to the Generalized Koopman Operator and the Nonlinear Fundamental Lemmaarxiv-2508.07494 Sparse Blocked context onlyAug 10, 2025
- ObfusQAte: A Proposed Framework to Evaluate LLM Robustness on Obfuscated Factual Question Answeringarxiv-2508.07321 Sparse Blocked context onlyAug 10, 2025
- IntrinsicWeather: Controllable Weather Editing in Intrinsic Spacearxiv-2508.06982 Sparse Blocked context onlyAug 9, 2025
- Seeing Through the Noise: Improving Infrared Small Target Detection and Segmentation from Noise Suppression Perspectivearxiv-2508.06878 Sparse Blocked context onlyAug 9, 2025
- Personalized Feature Translation for Expression Recognition: An Efficient Source-Free Domain Adaptation Methodarxiv-2508.09202 Sparse Blocked context onlyAug 8, 2025
- Post-training for Efficient Communication via Convention Formationarxiv-2508.06482 Sparse Blocked context onlyAug 8, 2025
- LLMEval-Fair: A Large-Scale Longitudinal Study on Robust and Fair Evaluation of Large Language Modelsarxiv-2508.05452 Sparse Blocked context onlyAug 7, 2025
- Unsupervised Learning for Inverse Problems in Computed Tomographyarxiv-2508.05321 Sparse Blocked context onlyAug 7, 2025
- Hidden Dynamics of Massive Activations in Transformer Trainingarxiv-2508.03616 Sparse Blocked context onlyAug 5, 2025
- Cropping outperforms dropout as an augmentation strategy for self-supervised training of text embeddingsarxiv-2508.03453 Sparse Blocked context onlyAug 5, 2025
- RooseBERT: A New Deal For Political Language Modellingarxiv-2508.03250 Sparse Blocked context onlyAug 5, 2025
- When Algorithms Meet Artists: Semantic Compression of Artists' Concerns in the Public AI-Art Debatearxiv-2508.03037 Sparse Blocked context onlyAug 5, 2025
- PoeTone: A Framework for Constrained Generation of Structured Chinese Songci with LLMsarxiv-2508.02515 Sparse Blocked context onlyAug 4, 2025
- Harnessing Temporal Databases for Systematic Evaluation of Factual Time-Sensitive Question-Answering in Large Language Modelsarxiv-2508.02045 Sparse Blocked context onlyAug 4, 2025
- Decomposing Representation Space into Interpretable Subspaces with Unsupervised Learningarxiv-2508.01916 Sparse Blocked context onlyAug 3, 2025
- MLP Memory: A Retriever-Pretrained Memory for Large Language Modelsarxiv-2508.01832 Sparse Blocked context onlyAug 3, 2025
- GHTM: A Graph-based Hybrid Topic Modeling Approach with a Benchmark Dataset for the Low-Resource Bengali Languagearxiv-2508.00605 Sparse Blocked context onlyAug 1, 2025
- Activation-Guided Local Editing for Jailbreaking Attacksarxiv-2508.00555 Sparse Blocked context onlyAug 1, 2025
- Acoustic Imaging for UAV Detection: Dense Beamformed Energy Maps and U-Net SELDarxiv-2508.00307 Sparse Blocked context onlyAug 1, 2025
- Role-Aware Language Models for Secure and Contextualized Access Control in Organizationsarxiv-2507.23465 Sparse Blocked context onlyJul 31, 2025
- Model Directions, Not Words: Mechanistic Topic Models Using Sparse Autoencodersarxiv-2507.23220 Sparse Blocked context onlyJul 31, 2025
- Better Together: Cross and Joint Covariances Enhance Signal Detectability in Undersampled Dataarxiv-2507.22207 Sparse Blocked context onlyJul 29, 2025
- Culinary Crossroads: A RAG Framework for Enhancing Diversity in Cross-Cultural Recipe Adaptationarxiv-2507.21934 Sparse Blocked context onlyJul 29, 2025
- Who's important? -- SUnSET: Synergistic Understanding of Stakeholder, Events and Time for Timeline Generationarxiv-2507.21903 Sparse Blocked context onlyJul 29, 2025
- Estimating Object Physical Properties from RGB-D Vision and Depth Robot Sensors Using Deep Learningdoi-10.1007_978-3-031-99565-1_8 Sparse Blocked context onlyJul 29, 2025
- A survey of diversity quantification in natural language processing: The why, what, where and howarxiv-2507.20858 Sparse Blocked context onlyJul 28, 2025
- Enhancing Jailbreak Attacks on LLMs via Persona Promptsarxiv-2507.22171 Sparse Blocked context onlyJul 28, 2025
- Packet-Level DDoS Data Augmentation Using Dual-Stream Temporal-Field Diffusionarxiv-2507.20115 Sparse Blocked context onlyJul 27, 2025
- FeynTune: Large Language Models for High-Energy Theoryarxiv-2508.03716 Sparse Blocked context onlyJul 24, 2025
- Efficient Compositional Multi-tasking for On-device Large Language Modelsarxiv-2507.16083 Sparse Blocked context onlyJul 21, 2025
- Commonsense on Demand: Generating and Selectively Integrating Commonsense Knowledge for Natural Language Inferencearxiv-2507.15100 Sparse Blocked context onlyJul 20, 2025
- CPC-CMS: Cognitive Pairwise Comparison Classification Model Selection Framework for Document-level Sentiment Analysisarxiv-2507.14022 Sparse Blocked context onlyJul 18, 2025
- Anthropomimetic Uncertainty: What Verbalized Uncertainty in Language Models is Missingarxiv-2507.10587 Sparse Blocked context onlyJul 11, 2025
- A Third Paradigm for LLM Evaluation: Dialogue Game-Based Evaluation using clembencharxiv-2507.08491 Sparse Blocked context onlyJul 11, 2025
- Page image classification for content-specific data processingarxiv-2507.21114 Sparse Blocked context onlyJul 11, 2025
- Towards Robust Sensor-Fusion Ground SLAM: A Comprehensive Benchmark and A Resilient Frameworkarxiv-2507.08364 Direct Blocked context onlyJul 11, 2025
- From Ambiguity to Accuracy: The Transformative Effect of Coreference Resolution on Retrieval-Augmented Generation systemsarxiv-2507.07847 Sparse Blocked context onlyJul 10, 2025
- COALA: Numerically Stable and Efficient Framework for Context-Aware Low-Rank Approximationarxiv-2507.07580 Sparse Blocked context onlyJul 10, 2025
- A Systematic Analysis of Hybrid Linear Attentionarxiv-2507.06457 Direct Blocked context onlyJul 8, 2025
- Psychometric Item Validation Using Virtual Respondents with Trait-Response Mediatorsarxiv-2507.05890 Sparse Blocked context onlyJul 8, 2025
- GPTKB v1.5: A Massive Knowledge Base for Exploring Factual LLM Knowledgearxiv-2507.05740 Sparse Blocked context onlyJul 8, 2025
- Mechanistic Indicators of Understanding in Large Language Modelsarxiv-2507.08017 Sparse Blocked context onlyJul 7, 2025
- K-Function: Joint Pronunciation Transcription and Feedback for Evaluating Kids Language Functionarxiv-2507.03043 Sparse Blocked context onlyJul 3, 2025
- On the Inference (In-)Security of Vertical Federated Learning: Efficient Auditing against Inference Tampering Attackarxiv-2507.02376 Sparse Blocked context onlyJul 3, 2025
- MuRating: A High Quality Data Selecting Approach to Multilingual Large Language Model Pretrainingarxiv-2507.01785 Sparse Blocked context onlyJul 2, 2025
- A Comparative Study of Competency Question Elicitation Methods from Ontology Requirementsarxiv-2507.02989 Sparse Blocked context onlyJul 1, 2025
- Skywork-R1V3 Technical Reportarxiv-2507.06167 Direct Blocked context onlyJul 1, 2025
- TaP: A Taxonomy-Guided Framework for Automated and Scalable Preference Data Generationarxiv-2506.23979 Sparse Blocked context onlyJun 30, 2025
- Unveiling Decision-Making in LLMs for Text Classification : Extraction of influential and interpretable concepts with Sparse Autoencodersarxiv-2506.23951 Sparse Blocked context onlyJun 30, 2025
- TTSDS2: Resources and Benchmark for Evaluating Human-Quality Text to Speech Systemsarxiv-2506.19441 Sparse Blocked context onlyJun 24, 2025
- Programming by Backprop: An Instruction is Worth 100 Examples When Finetuning LLMsarxiv-2506.18777 Sparse Blocked context onlyJun 23, 2025
- Accelerating Residual Reinforcement Learning with Uncertainty Estimationarxiv-2506.17564 Sparse Blocked context onlyJun 21, 2025
- Long-Context Generalization with Sparse Attentionarxiv-2506.16640 Curated Related Blocked context onlyJun 19, 2025
- Measuring Intent Comprehension in LLMsarxiv-2506.16584 Sparse Blocked context onlyJun 19, 2025
- ConLID: Supervised Contrastive Learning for Low-Resource Language Identificationarxiv-2506.15304 Sparse Blocked context onlyJun 18, 2025
- Sysformer: Safeguarding Frozen Large Language Models with Adaptive System Promptsarxiv-2506.15751 Sparse Blocked context onlyJun 18, 2025
- Accurate and scalable exchange-correlation with deep learningarxiv-2506.14665 Sparse Blocked context onlyJun 17, 2025
- Hope Speech Detection in code-mixed Roman Urdu tweets: A Positive Turn in Natural Language Processingarxiv-2506.21583 Sparse Blocked context onlyJun 17, 2025
- Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMsarxiv-2506.13727 Sparse Blocked context onlyJun 16, 2025
- Strategic Scaling of Test-Time Compute: A Bandit Learning Approacharxiv-2506.12721 Sparse Blocked context onlyJun 15, 2025
- Persona-driven Simulation of Voting Behavior in the European Parliament with Large Language Modelsarxiv-2506.11798 Sparse Blocked context onlyJun 13, 2025