300 canonical paper links on this archive page.
- GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0)arxiv-2604.17091 Sparse Blocked context onlyApr 18, 2026
- Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RLarxiv-2604.17073 Sparse Blocked context onlyApr 18, 2026
- MNAFT: modality neuron-aware fine-tuning of multimodal large language models for image translationarxiv-2604.16943 Sparse Blocked context onlyApr 18, 2026
- EasyVideoR1: Easier RL for Video Understandingarxiv-2604.16893 Sparse Blocked context onlyApr 18, 2026
- FlowEvo: Self-Evolving Agents through the Co-Evolution of Workflows and Executable Skillsarxiv-2607.21596 Sparse Blocked context onlyApr 18, 2026
- The Illusion of Certainty: Decoupling Capability and Calibration in On-Policy Distillationarxiv-2604.16830 Sparse Blocked context onlyApr 18, 2026
- Repurposing 3D Generative Model for Autoregressive Layout Generationarxiv-2604.16299 Sparse Blocked context onlyApr 17, 2026
- VEFX-Bench: A Holistic Benchmark for Generic Video Editing and Visual Effectsarxiv-2604.16272 Sparse Blocked context onlyApr 17, 2026
- AtManRL: Towards Faithful Reasoning via Differentiable Attention Saliencyarxiv-2604.16158 Sparse Blocked context onlyApr 17, 2026
- Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMsarxiv-2604.16054 Sparse Blocked context onlyApr 17, 2026
- On the Robustness of LLM-Based Dense Retrievers: A Systematic Analysis of Generalizability and Stabilityarxiv-2604.16576 Sparse Blocked context onlyApr 17, 2026
- Neurosymbolic Repo-level Code Localizationarxiv-2604.16021 Sparse Blocked context onlyApr 17, 2026
- Hierarchical Codec Diffusion for Video-to-Speech Generationarxiv-2604.15923 Sparse Blocked context onlyApr 17, 2026
- CoEvolve: Training LLM Agents via Agent-Data Mutual Evolutionarxiv-2604.15840 Sparse Blocked context onlyApr 17, 2026
- Qwen3.5-Omni Technical Reportarxiv-2604.15804 Sparse Blocked context onlyApr 17, 2026
- KWBench: Measuring Unprompted Problem Recognition in Knowledge Workarxiv-2604.15760 Sparse Blocked context onlyApr 17, 2026
- GTA-2: Benchmarking General Tool Agents from Atomic Tool-Use to Open-Ended Workflowsarxiv-2604.15715 Sparse Blocked context onlyApr 17, 2026
- Faster LLM Inference via Sequential Monte Carloarxiv-2604.15672 Curated Related Blocked context onlyApr 17, 2026
- Why Fine-Tuning Encourages Hallucinations and How to Fix Itarxiv-2604.15574 Sparse Blocked context onlyApr 16, 2026
- (1D) Ordered Tokens Enable Efficient Test-Time Searcharxiv-2604.15453 Sparse Blocked context onlyApr 16, 2026
- MM-WebAgent: A Hierarchical Multimodal Web Agent for Webpage Generationarxiv-2604.15309 Sparse Blocked context onlyApr 16, 2026
- RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Frameworkarxiv-2604.15308 Direct Blocked context onlyApr 16, 2026
- Generalization in LLM Problem Solving: The Case of the Shortest Patharxiv-2604.15306 Sparse Blocked context onlyApr 16, 2026
- Why Do Vision Language Models Struggle To Recognize Human Emotions?arxiv-2604.15280 Sparse Blocked context onlyApr 16, 2026
- Scaling Test-Time Compute for Agentic Codingarxiv-2604.16529 Sparse Blocked context onlyApr 16, 2026
- From Tokens to Steps: Verification-Aware Speculative Decoding for Efficient Multi-Step Reasoningarxiv-2604.15244 Sparse Blocked context onlyApr 16, 2026
- MADE: A Living Benchmark for Multi-Label Text Classification with Uncertainty Quantification of Medical Device Adverse Eventsarxiv-2604.15203 Sparse Blocked context onlyApr 16, 2026
- Scepsy: Serving Agentic Workflows Using Aggregate LLM Pipelinesarxiv-2604.15186 Sparse Blocked context onlyApr 16, 2026
- Compressing Sequences in the Latent Embedding Space: $K$-Token Merging for Large Language Modelsarxiv-2604.15153 Sparse Blocked context onlyApr 16, 2026
- QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading Strategiesarxiv-2604.15151 Sparse Blocked context onlyApr 16, 2026
- IG-Search: Step-Level Information Gain Rewards for Search-Augmented Reasoningarxiv-2604.15148 Sparse Blocked context onlyApr 16, 2026
- Blinded Multi-Rater Comparative Evaluation of a Large Language Model and Clinician-Authored Responses in CGM-Informed Diabetes Counselingarxiv-2604.15124 Sparse Blocked context onlyApr 16, 2026
- OpenMobile: Building Open Mobile Agents with Task and Trajectory Synthesisarxiv-2604.15093 Sparse Blocked context onlyApr 16, 2026
- Autonomous Evolution of EDA Tools: Multi-Agent Self-Evolved ABCarxiv-2604.15082 Sparse Blocked context onlyApr 16, 2026
- Towards Faster Language Model Inference Using Mixture-of-Experts Flow Matchingarxiv-2604.15009 Sparse Blocked context onlyApr 16, 2026
- UniDoc-RL: Coarse-to-Fine Visual RAG with Hierarchical Actions and Dense Rewardsarxiv-2604.14967 Sparse Blocked context onlyApr 16, 2026
- RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Modelsarxiv-2604.14951 Sparse Blocked context onlyApr 16, 2026
- LongAct: Harnessing Intrinsic Activation Patterns for Long-Context Reinforcement Learningarxiv-2604.14922 Sparse Blocked context onlyApr 16, 2026
- RACER: Retrieval-Augmented Contextual Rapid Speculative Decodingarxiv-2604.14885 Sparse Blocked context onlyApr 16, 2026
- Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problemarxiv-2604.14808 Sparse Blocked context onlyApr 16, 2026
- Is This Edit Correct? A Multi-Dimensional Benchmark for Reasoning-Aware Image Editingarxiv-2606.05172 Sparse Blocked context onlyApr 16, 2026
- DR$^{3}$-Eval: Towards Realistic and Reproducible Deep Research Evaluationarxiv-2604.14683 Sparse Blocked context onlyApr 16, 2026
- StoryCoder: Narrative Reformulation for Structured Reasoning in LLM Code Generationarxiv-2604.14631 Curated Related Blocked context onlyApr 16, 2026
- TRACER: Trace-Based Adaptive Cost-Efficient Routing for LLM Classificationarxiv-2604.14531 Direct Blocked context onlyApr 16, 2026
- EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluationarxiv-2604.14306 Sparse Blocked context onlyApr 15, 2026
- ROSE: Retrieval-Oriented Segmentation Enhancementarxiv-2604.14147 Sparse Blocked context onlyApr 15, 2026
- HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worldsarxiv-2604.14268 Direct Blocked context onlyApr 15, 2026
- SpatialEvo: Self-Evolving Spatial Intelligence via Deterministic Geometric Environmentsarxiv-2604.14144 Sparse Blocked context onlyApr 15, 2026
- From $P(y|x)$ to $P(y)$: Investigating Reinforcement Learning in Pre-train Spacearxiv-2604.14142 Sparse Blocked context onlyApr 15, 2026
- HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation Systemarxiv-2604.14125 Sparse Blocked context onlyApr 15, 2026
- TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Explorationarxiv-2604.14116 Curated Related Blocked context onlyApr 15, 2026
- TIP: Token Importance in On-Policy Distillationarxiv-2604.14084 Sparse Blocked context onlyApr 15, 2026
- Feed-Forward 3D Scene Modeling: A Problem-Driven Perspectivearxiv-2604.14025 Sparse Blocked context onlyApr 15, 2026
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agentsarxiv-2604.14004 Sparse Blocked context onlyApr 15, 2026
- CollabCoder: Plan-Code Co-Evolution via Collaborative Decision-Making for Efficient Code Generationarxiv-2604.13946 Sparse Blocked context onlyApr 15, 2026
- DiPO: Disentangled Perplexity Policy Optimization for Fine-grained Exploration-Exploitation Trade-Offarxiv-2604.13902 Sparse Blocked context onlyApr 15, 2026
- SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attentionarxiv-2604.13847 Sparse Blocked context onlyApr 15, 2026
- An Empirical Investigation of Practical LLM-as-a-Judge Improvement Techniques on RewardBench 2arxiv-2604.13717 Sparse Blocked context onlyApr 15, 2026
- IndicDB -- Benchmarking Multilingual Text-to-SQL Capabilities in Indian Languagesarxiv-2604.13686 Sparse Blocked context onlyApr 15, 2026
- Calibrated Speculative Decoding: Frequency-Guided Candidate Selection for Efficient Inferencearxiv-2604.13634 Sparse Blocked context onlyApr 15, 2026
- Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challengesarxiv-2604.13602 Sparse Blocked context onlyApr 15, 2026
- ToolSpec: Accelerating Tool Calling via Schema-Aware and Retrieval-Augmented Speculative Decodingarxiv-2604.13519 Sparse Blocked context onlyApr 15, 2026
- Representation over Routing: Overcoming Surrogate Hacking in Multi-Timescale PPOarxiv-2604.13517 Sparse Blocked context onlyApr 15, 2026
- Peer-Predictive Self-Training for Language Model Reasoningarxiv-2604.13356 Sparse Blocked context onlyApr 14, 2026
- AgentSPEX: An Agent SPecification and EXecution Languagearxiv-2604.13346 Sparse Blocked context onlyApr 14, 2026
- WebXSkill: Skill Learning for Autonomous Web Agentsarxiv-2604.13318 Sparse Blocked context onlyApr 14, 2026
- KV Packet: Recomputation-Free Context-Independent KV Caching for LLMsarxiv-2604.13226 Sparse Blocked context onlyApr 14, 2026
- InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysisarxiv-2604.13201 Sparse Blocked context onlyApr 14, 2026
- Toward Autonomous Long-Horizon Engineering for ML Researcharxiv-2604.13018 Sparse Blocked context onlyApr 14, 2026
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipearxiv-2604.13016 Sparse Blocked context onlyApr 14, 2026
- Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillationarxiv-2604.13010 Sparse Blocked context onlyApr 14, 2026
- One Token Away from Collapse: The Fragility of Instruction-Tuned Helpfulnessarxiv-2604.13006 Sparse Blocked context onlyApr 14, 2026
- Boosting Visual Instruction Tuning with Self-Supervised Guidancearxiv-2604.12966 Sparse Blocked context onlyApr 14, 2026
- Towards Long-horizon Agentic Multimodal Searcharxiv-2604.12890 Sparse Blocked context onlyApr 14, 2026
- KnowRL: Boosting LLM Reasoning via Reinforcement Learning with Minimal-Sufficient Knowledge Guidancearxiv-2604.12627 Sparse Blocked context onlyApr 14, 2026
- Habitat-GS: A High-Fidelity Navigation Simulator with Dynamic Gaussian Splattingarxiv-2604.12626 Direct Blocked context onlyApr 14, 2026
- Balanced Aggregation: Understanding and Fixing Aggregation Bias in GRPOarxiv-2605.04077 Sparse Blocked context onlyApr 14, 2026
- Latent-Condensed Transformer for Efficient Long Context Modelingarxiv-2604.12452 Sparse Blocked context onlyApr 14, 2026
- Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoningarxiv-2604.12374 Curated Related Blocked context onlyApr 14, 2026
- Self-Adversarial One Step Generation via Condition Shiftingarxiv-2604.12322 Sparse Blocked context onlyApr 14, 2026
- Self-Distillation Zero: Self-Revision Turns Binary Rewards into Dense Supervisionarxiv-2604.12002 Sparse Blocked context onlyApr 13, 2026
- Towards Autonomous Mechanistic Reasoning in Virtual Cellsarxiv-2604.11661 Sparse Blocked context onlyApr 13, 2026
- RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Timearxiv-2604.11626 Sparse Blocked context onlyApr 13, 2026
- Self-Evolving LLM Memory Extraction Across Heterogeneous Tasksarxiv-2604.11610 Curated Related Blocked context onlyApr 13, 2026
- A Triadic Suffix Tokenization Scheme for Numerical Reasoningarxiv-2604.11582 Sparse Blocked context onlyApr 13, 2026
- Hidden Measurement Error in LLM Pipelines Distorts Annotation, Evaluation, and Benchmarkingarxiv-2604.11581 Sparse Blocked context onlyApr 13, 2026
- Seeing Through Touch: Tactile-Driven Visual Localization of Material Regionsarxiv-2604.11579 Sparse Blocked context onlyApr 13, 2026
- Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scalearxiv-2604.11554 Direct Blocked context onlyApr 13, 2026
- Time is Not a Label: Continuous Phase Rotation for Temporal Knowledge Graphs and Agentic Memoryarxiv-2604.11544 Sparse Blocked context onlyApr 13, 2026
- METER: Evaluating Multi-Level Contextual Causal Reasoning in Large Language Modelsarxiv-2604.11502 Sparse Blocked context onlyApr 13, 2026
- Revisiting Compositionality in Dual-Encoder Vision-Language Models: The Role of Inferencearxiv-2604.11496 Sparse Blocked context onlyApr 13, 2026
- Anthropogenic Regional Adaptation in Multimodal Vision-Language Modelarxiv-2604.11490 Sparse Blocked context onlyApr 13, 2026
- METRO: Towards Strategy Induction from Expert Dialogue Transcripts for Non-collaborative Dialoguesarxiv-2604.11427 Sparse Blocked context onlyApr 13, 2026
- OmniScript: Towards Audio-Visual Script Generation for Long-Form Cinematic Videoarxiv-2604.11102 Sparse Blocked context onlyApr 13, 2026
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Modelsarxiv-2604.10866 Sparse Blocked context onlyApr 13, 2026
- Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editingarxiv-2604.10708 Direct Blocked context onlyApr 12, 2026
- Critical-CoT: A Robust Defense Framework against Reasoning-Level Backdoor Attacks in Large Language Modelsarxiv-2604.10681 Sparse Blocked context onlyApr 12, 2026
- The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agentsarxiv-2604.10577 Sparse Blocked context onlyApr 12, 2026
- IceCache: Memory-efficient KV-cache Management for Long-Sequence LLMsarxiv-2604.10539 Sparse Blocked context onlyApr 12, 2026
- Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Policy Evaluationarxiv-2604.10511 Sparse Blocked context onlyApr 12, 2026
- A Temporally Augmented Graph Attention Network for Affordance Classificationarxiv-2604.10149 Sparse Blocked context onlyApr 11, 2026
- Think in Sentences: Explicit Sentence Boundaries Enhance Language Model's Capabilitiesarxiv-2604.10135 Sparse Blocked context onlyApr 11, 2026
- Many-Tier Instruction Hierarchy in LLM Agentsarxiv-2604.09443 Curated Related Blocked context onlyApr 10, 2026
- Do AI Coding Agents Log Like Humans? An Empirical Studyarxiv-2604.09409 Sparse Blocked context onlyApr 10, 2026
- HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?arxiv-2604.09408 Sparse Blocked context onlyApr 10, 2026
- CONSCIENTIA: Can LLM Agents Learn to Strategize? Emergent Deception and Trust in a Multi-Agent NYC Simulationarxiv-2604.09746 Sparse Blocked context onlyApr 10, 2026
- MAB-DQA: Addressing Query Aspect Importance in Document Question Answering with Multi-Armed Banditsarxiv-2604.08952 Sparse Blocked context onlyApr 10, 2026
- SPPO: Sequence-Level PPO for Long-Horizon Reasoning Tasksarxiv-2604.08865 Sparse Blocked context onlyApr 10, 2026
- Seeing but Not Thinking: Routing Distraction in Multimodal Mixture-of-Expertsarxiv-2604.08541 Sparse Blocked context onlyApr 9, 2026
- AVGen-Bench: A Task-Driven Benchmark for Multi-Granular Evaluation of Text-to-Audio-Video Generationarxiv-2604.08540 Sparse Blocked context onlyApr 9, 2026
- OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasksarxiv-2604.08539 Sparse Blocked context onlyApr 9, 2026
- RewardFlow: Generate Images by Optimizing What You Rewardarxiv-2604.08536 Sparse Blocked context onlyApr 9, 2026
- Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Modelsarxiv-2604.08527 Sparse Blocked context onlyApr 9, 2026
- Ads in AI Chatbots? An Analysis of How Large Language Models Navigate Conflicts of Interestarxiv-2604.08525 Sparse Blocked context onlyApr 9, 2026
- What Drives Representation Steering? A Mechanistic Case Study on Steering Refusalarxiv-2604.08524 Sparse Blocked context onlyApr 9, 2026
- What do Language Models Learn and When? The Implicit Curriculum Hypothesisarxiv-2604.08510 Sparse Blocked context onlyApr 9, 2026
- Quantifying Explanation Consistency: The C-Score Metric for CAM-Based Explainability in Medical Image Classificationarxiv-2604.08502 Sparse Blocked context onlyApr 9, 2026
- SUPERNOVA: Eliciting General Reasoning in LLMs with Reinforcement Learning on Natural Instructionsarxiv-2604.08477 Sparse Blocked context onlyApr 9, 2026
- Faithful GRPO: Improving Visual Spatial Reasoning in Multimodal Language Models via Constrained Policy Optimizationarxiv-2604.08476 Sparse Blocked context onlyApr 9, 2026
- From Safety Risk to Design Principle: Peer-Preservation in Multi-Agent LLM Systems and Its Implications for Orchestrated Democratic Discourse Analysisarxiv-2604.08465 Sparse Blocked context onlyApr 9, 2026
- OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidancearxiv-2604.08461 Sparse Blocked context onlyApr 9, 2026
- Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Modelsarxiv-2604.08456 Sparse Blocked context onlyApr 9, 2026
- KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluationarxiv-2604.08455 Direct Blocked context onlyApr 9, 2026
- Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEMarxiv-2604.08425 Sparse Blocked context onlyApr 9, 2026
- Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditingarxiv-2604.08401 Sparse Blocked context onlyApr 9, 2026
- Phantasia: Context-Adaptive Backdoors in Vision Language Modelsarxiv-2604.08395 Sparse Blocked context onlyApr 9, 2026
- Awakening the Sleeping Agent: Lean-Specific Agentic Data Reactivates General Tool Use in Goedel Proverarxiv-2604.08388 Sparse Blocked context onlyApr 9, 2026
- SkillClaw: Let Skills Evolve Collectively with Agentic Evolverarxiv-2604.08377 Sparse Blocked context onlyApr 9, 2026
- Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agentsarxiv-2604.08369 Sparse Blocked context onlyApr 9, 2026
- ASPECT:Analogical Semantic Policy Execution via Language Conditioned Transferarxiv-2604.08355 Sparse Blocked context onlyApr 9, 2026
- PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Modelsarxiv-2604.08340 Sparse Blocked context onlyApr 9, 2026
- InstAP: Instance-Aware Vision-Language Pre-Train for Spatial-Temporal Understandingarxiv-2604.08337 Sparse Blocked context onlyApr 9, 2026
- Dead Weights, Live Signals: Feedforward Graphs of Frozen Language Modelsarxiv-2604.08335 Sparse Blocked context onlyApr 9, 2026
- Lost in the Hype: Revealing and Dissecting the Performance Degradation of Medical Multimodal Large Language Models in Image Classificationarxiv-2604.08333 Sparse Blocked context onlyApr 9, 2026
- ProMedical: Hierarchical Fine-Grained Criteria Modeling for Medical LLM Alignment via Explicit Injectionarxiv-2604.08326 Sparse Blocked context onlyApr 9, 2026
- DMax: Aggressive Parallel Decoding for dLLMsarxiv-2604.08302 Sparse Blocked context onlyApr 9, 2026
- Distributed Multi-Layer Editing for Rule-Level Knowledge in Large Language Modelsarxiv-2604.08284 Sparse Blocked context onlyApr 9, 2026
- When to Trust Tools? Adaptive Tool Trust Calibration For Tool-Integrated Math Reasoningarxiv-2604.08281 Sparse Blocked context onlyApr 9, 2026
- DBMF: A Dual-Branch Multimodal Framework for Out-of-Distribution Detectionarxiv-2604.08261 Sparse Blocked context onlyApr 9, 2026
- Self-Debias: Self-correcting for Debiasing Large Language Modelsarxiv-2604.08243 Sparse Blocked context onlyApr 9, 2026
- EditCaption: Human-Aligned Instruction Synthesis for Image Editing via Supervised Fine-Tuning and Direct Preference Optimizationarxiv-2604.08213 Sparse Blocked context onlyApr 9, 2026
- Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inferencearxiv-2604.08133 Sparse Blocked context onlyApr 9, 2026
- Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generatorarxiv-2604.08121 Sparse Blocked context onlyApr 9, 2026
- Small Vision-Language Models are Smart Compressors for Long Video Understandingarxiv-2604.08120 Curated Related Blocked context onlyApr 9, 2026
- Dual-Pool Token-Budget Routing for Cost-Efficient and Reliable LLM Servingarxiv-2604.08075 Sparse Blocked context onlyApr 9, 2026
- Guaranteeing Knowledge Integration with Joint Decoding for Retrieval-Augmented Generationarxiv-2604.08046 Sparse Blocked context onlyApr 9, 2026
- Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMsarxiv-2604.08003 Sparse Blocked context onlyApr 9, 2026
- PASK: Toward Intent-Aware Proactive Agents with Long-Term Memoryarxiv-2604.08000 Sparse Blocked context onlyApr 9, 2026
- A Decomposition Perspective to Long-context Reasoning for LLMsarxiv-2604.07981 Sparse Blocked context onlyApr 9, 2026
- TOOLCAD: Exploring Tool-Using Large Language Models in Text-to-CAD Generation with Reinforcement Learningarxiv-2604.07960 Sparse Blocked context onlyApr 9, 2026
- Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learningarxiv-2604.07941 Sparse Blocked context onlyApr 9, 2026
- TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillationarxiv-2604.07894 Sparse Blocked context onlyApr 9, 2026
- Data Selection for Multi-turn Dialogue Instruction Tuningarxiv-2604.07892 Sparse Blocked context onlyApr 9, 2026
- Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injectionarxiv-2604.07831 Sparse Blocked context onlyApr 9, 2026
- TEMPER: Testing Emotional Perturbation in Quantitative Reasoningarxiv-2604.07801 Sparse Blocked context onlyApr 9, 2026
- Sensitivity-Positional Co-Localization in GQA Transformersarxiv-2604.07766 Sparse Blocked context onlyApr 9, 2026
- Beyond Social Pressure: Benchmarking Epistemic Attack in Large Language Modelsarxiv-2604.07749 Sparse Blocked context onlyApr 9, 2026
- Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealingarxiv-2604.07747 Sparse Blocked context onlyApr 9, 2026
- Emotion Concepts and their Function in a Large Language Modelarxiv-2604.07729 Sparse Blocked context onlyApr 9, 2026
- Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Predictionarxiv-2604.07659 Sparse Blocked context onlyApr 8, 2026
- Optimal Decay Spectra for Linear Recurrencesarxiv-2604.07658 Sparse Blocked context onlyApr 8, 2026
- How Independent are Large Language Models? A Statistical Framework for Auditing Behavioral Entanglement and Reweighting Verifier Ensemblesarxiv-2604.07650 Sparse Blocked context onlyApr 8, 2026
- Reasoning-Based Refinement of Unsupervised Text Clusters with LLMsarxiv-2604.07562 Sparse Blocked context onlyApr 8, 2026
- EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agentsarxiv-2604.07549 Sparse Blocked context onlyApr 8, 2026
- Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMsarxiv-2604.07518 Sparse Blocked context onlyApr 8, 2026
- SYN-DIGITS: A Synthetic Control Framework for Calibrated Digital Twin Simulationarxiv-2604.07513 Sparse Blocked context onlyApr 8, 2026
- ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Frameworkarxiv-2604.07506 Sparse Blocked context onlyApr 8, 2026
- Enabling Intrinsic Reasoning over Dense Geospatial Embeddings with DFR-Gemmaarxiv-2604.07490 Sparse Blocked context onlyApr 8, 2026
- Cross-Tokenizer LLM Distillation through a Byte-Level Interfacearxiv-2604.07466 Sparse Blocked context onlyApr 8, 2026
- Personalized RewardBench: Evaluating Reward Models with Human Aligned Personalizationarxiv-2604.07343 Sparse Blocked context onlyApr 8, 2026
- GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agentsarxiv-2604.07429 Sparse Blocked context onlyApr 8, 2026
- OpenSpatial: A Principled Data Engine for Empowering Spatial Intelligencearxiv-2604.07296 Sparse Blocked context onlyApr 8, 2026
- A Systematic Study of Retrieval Pipeline Design for Retrieval-Augmented Medical Question Answeringarxiv-2604.07274 Sparse Blocked context onlyApr 8, 2026
- Joint Optimization of Reasoning and Dual-Memory for Self-Learning Diagnostic Agentarxiv-2604.07269 Sparse Blocked context onlyApr 8, 2026
- How Much LLM Does a Self-Revising Agent Actually Need?arxiv-2604.07236 Sparse Blocked context onlyApr 8, 2026
- TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectoriesarxiv-2604.07223 Sparse Blocked context onlyApr 8, 2026
- MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulationarxiv-2605.13857 Sparse Blocked context onlyApr 8, 2026
- Agent-Driven Corpus Linguistics: A Framework for Autonomous Linguistic Discoveryarxiv-2604.07189 Sparse Blocked context onlyApr 8, 2026
- STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systemsarxiv-2604.07100 Sparse Blocked context onlyApr 8, 2026
- Is Cross-Lingual Transfer in Bilingual Models Human-Like? A Study with Overlapping Word Forms in Dutch and Englisharxiv-2604.07067 Sparse Blocked context onlyApr 8, 2026
- Sell More, Play Less: Benchmarking LLM Realistic Selling Skillarxiv-2604.07054 Sparse Blocked context onlyApr 8, 2026
- ReDAct: Uncertainty-Aware Deferral for LLM Agentsarxiv-2604.07036 Curated Related Blocked context onlyApr 8, 2026
- Gemma 4, Phi-4, and Qwen3: Accuracy-Efficiency Tradeoffs in Dense and MoE Reasoning Language Modelsarxiv-2604.07035 Sparse Blocked context onlyApr 8, 2026
- MARS: Enabling Autoregressive Models Multi-Token Generationarxiv-2604.07023 Sparse Blocked context onlyApr 8, 2026
- DTCRS: Dynamic Tree Construction for Recursive Summarizationarxiv-2604.07012 Sparse Blocked context onlyApr 8, 2026
- iTAG: Inverse Design for Natural Text Generation with Accurate Causal Graph Annotationsarxiv-2604.06902 Sparse Blocked context onlyApr 8, 2026
- On the Step Length Confounding in LLM Reasoning Data Selectionarxiv-2604.06834 Sparse Blocked context onlyApr 8, 2026
- Fast-dVLM: Efficient Block-Diffusion VLM via Direct Conversion from Autoregressive VLMarxiv-2604.06832 Sparse Blocked context onlyApr 8, 2026
- WRAP++: Web discoveRy Amplified Pretrainingarxiv-2604.06829 Sparse Blocked context onlyApr 8, 2026
- Beyond Accuracy: Diagnosing Algebraic Reasoning Failures in LLMs Across Nine Complexity Dimensionsarxiv-2604.06799 Sparse Blocked context onlyApr 8, 2026
- GCoT-Decoding: Unlocking Deep Reasoning Paths for Universal Question Answeringarxiv-2604.06794 Sparse Blocked context onlyApr 8, 2026
- From Perception to Autonomous Computational Modeling: A Multi-Agent Approacharxiv-2604.06788 Sparse Blocked context onlyApr 8, 2026
- Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Modelsarxiv-2604.06767 Sparse Blocked context onlyApr 8, 2026
- How Long Reasoning Chains Influence LLMs' Judgment of Answer Factualityarxiv-2604.06756 Sparse Blocked context onlyApr 8, 2026
- Select-then-Solve: Paradigm Routing as Inference-Time Optimization for LLM Agentsarxiv-2604.06753 Sparse Blocked context onlyApr 8, 2026
- StructKV: Preserving the Structural Skeleton for Scalable Long-Context Inferencearxiv-2604.06746 Sparse Blocked context onlyApr 8, 2026
- WisdomInterrogatory (LuWen): An Open-Source Legal Large Language Model Technical Reportarxiv-2604.06737 Sparse Blocked context onlyApr 8, 2026
- Steering the Verifiability of Multimodal AI Hallucinationsarxiv-2604.06714 Sparse Blocked context onlyApr 8, 2026
- Specializing Large Models for Oracle Bone Script Interpretation via Component-Grounded Multimodal Knowledge Augmentationarxiv-2604.06711 Sparse Blocked context onlyApr 8, 2026
- Adaptive Prompt Structure Factorization: A Framework for Self-Discovering and Optimizing Compositional Prompt Programsarxiv-2604.06699 Sparse Blocked context onlyApr 8, 2026
- ChemVLR: Prioritizing Reasoning in Perception for Chemical Vision-Language Understandingarxiv-2604.06685 Sparse Blocked context onlyApr 8, 2026
- Argus: Reorchestrating Static Analysis via a Multi-Agent Ensemble for Full-Chain Security Vulnerability Detectionarxiv-2604.06633 Sparse Blocked context onlyApr 8, 2026
- DiffuMask: Diffusion Language Model for Token-level Prompt Pruningarxiv-2604.06627 Sparse Blocked context onlyApr 8, 2026
- Scientific Knowledge-driven Decoding Constraints Improving the Reliability of LLMsarxiv-2604.06603 Sparse Blocked context onlyApr 8, 2026
- LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sourcesarxiv-2604.06571 Sparse Blocked context onlyApr 8, 2026
- CCD-CBT: Multi-Agent Therapeutic Interaction for CBT Guided by Cognitive Conceptualization Diagramarxiv-2604.06551 Sparse Blocked context onlyApr 8, 2026
- Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learningarxiv-2605.02913 Sparse Blocked context onlyApr 8, 2026
- MedConclusion: A Benchmark for Biomedical Conclusion Generation from Structured Abstractsarxiv-2604.06505 Sparse Blocked context onlyApr 7, 2026
- Closing the Speech-Text Gap with Limited Audio for Effective Domain Adaptation in LLM-Based ASRarxiv-2604.06487 Sparse Blocked context onlyApr 7, 2026
- ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMsarxiv-2604.06484 Sparse Blocked context onlyApr 7, 2026
- DataSTORM: Deep Research on Large-Scale Databases using Exploratory Data Analysis and Data Storytellingarxiv-2604.06474 Sparse Blocked context onlyApr 7, 2026
- Multi-objective Evolutionary Merging Enables Efficient Reasoning Modelsarxiv-2604.06465 Sparse Blocked context onlyApr 7, 2026
- Context-Aware Dialectal Arabic Machine Translation with Interactive Region and Register Selectionarxiv-2604.06456 Sparse Blocked context onlyApr 7, 2026
- Learning to Interrupt in Language-based Multi-agent Communicationarxiv-2604.06452 Sparse Blocked context onlyApr 7, 2026
- The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planningarxiv-2604.06427 Sparse Blocked context onlyApr 7, 2026
- State-of-the-Art Arabic Language Modeling with Sparse MoE Fine-Tuning and Chain-of-Thought Distillationarxiv-2604.06421 Sparse Blocked context onlyApr 7, 2026
- Attention Flows: Tracing LLM Conceptual Engagement via Story Summariesarxiv-2604.06416 Sparse Blocked context onlyApr 7, 2026
- Application-Driven Pedagogical Knowledge Optimization of Open-Source LLMs via Reinforcement Learning and Supervised Fine-Tuningarxiv-2604.06385 Sparse Blocked context onlyApr 7, 2026
- STDec: Spatio-Temporal Stability Guided Decoding for dLLMsarxiv-2604.06330 Sparse Blocked context onlyApr 7, 2026
- Paper Circle: An Open-source Multi-agent Research Discovery and Analysis Frameworkarxiv-2604.06170 Sparse Blocked context onlyApr 7, 2026
- Toward Consistent World Models with Multi-Token Prediction and Latent Semantic Enhancementarxiv-2604.06155 Sparse Blocked context onlyApr 7, 2026
- Social Dynamics as Critical Vulnerabilities that Undermine Objective Decision-Making in LLM Collectivesarxiv-2604.06091 Sparse Blocked context onlyApr 7, 2026
- Stories of Your Life as Others: A Round-Trip Evaluation of LLM-Generated Life Stories Conditioned on Rich Psychometric Profilesarxiv-2604.06071 Sparse Blocked context onlyApr 7, 2026
- Short Data, Long Context: Distilling Positional Knowledge in Transformersarxiv-2604.06070 Sparse Blocked context onlyApr 7, 2026
- From Hallucination to Structure Snowballing: The Alignment Tax of Constrained Decoding in LLM Reflectionarxiv-2604.06066 Sparse Blocked context onlyApr 7, 2026
- BiMind: A Dual-Head Reasoning Model with Attention-Geometry Adapter for Incorrect Information Detectionarxiv-2604.06022 Sparse Blocked context onlyApr 7, 2026
- Is CLIP Cross-Eyed? Revealing and Mitigating Center Bias in the CLIP Familyarxiv-2604.05971 Sparse Blocked context onlyApr 7, 2026
- FinReporting: An Agentic Workflow for Localized Reporting of Cross-Jurisdiction Financial Disclosuresarxiv-2604.05966 Direct Blocked context onlyApr 7, 2026
- "I See What You Did There": Can Large Vision-Language Models Understand Multimodal Puns?arxiv-2604.05930 Sparse Blocked context onlyApr 7, 2026
- Mechanistic Circuit-Based Knowledge Editing in Large Language Modelsarxiv-2604.05876 Sparse Blocked context onlyApr 7, 2026
- Evaluating Learner Representations for Differentiation Prior to Instructional Outcomesarxiv-2604.05848 Sparse Blocked context onlyApr 7, 2026
- AgentGL: Towards Agentic Graph Learning with LLMs via Reinforcement Learningarxiv-2604.05846 Sparse Blocked context onlyApr 7, 2026
- WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answeringarxiv-2604.05818 Sparse Blocked context onlyApr 7, 2026
- Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversationarxiv-2604.05795 Sparse Blocked context onlyApr 7, 2026
- What Models Know, How Well They Know It: Knowledge-Weighted Fine-Tuning for Learning When to Say "I Don't Know"arxiv-2604.05779 Sparse Blocked context onlyApr 7, 2026
- Beyond the Beep: Scalable Collision Anticipation and Real-Time Explainability with BADAS-2.0arxiv-2604.05767 Sparse Blocked context onlyApr 7, 2026
- Identifying Influential N-grams in Confidence Calibration via Regression Analysisarxiv-2604.05757 Sparse Blocked context onlyApr 7, 2026
- Controlling Distributional Bias in Multi-Round LLM Generation via KL-Optimized Fine-Tuningarxiv-2604.05756 Sparse Blocked context onlyApr 7, 2026
- Can Large Language Models Reinvent Foundational Algorithms?arxiv-2604.05716 Sparse Blocked context onlyApr 7, 2026
- Attention Editing: A Versatile Framework for Cross-Architecture Attention Conversionarxiv-2604.05688 Sparse Blocked context onlyApr 7, 2026
- LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signalsarxiv-2604.05655 Sparse Blocked context onlyApr 7, 2026
- See the Forest for the Trees: Loosely Speculative Decoding via Visual-Semantic Guidance for Efficient Inference of Video LLMsarxiv-2604.05650 Sparse Blocked context onlyApr 7, 2026
- Label Effects: Shared Heuristic Reliance in Trust Assessment by Humans and LLM-as-a-Judgearxiv-2604.05593 Sparse Blocked context onlyApr 7, 2026
- Weakly Supervised Distillation of Hallucination Signals into Transformer Representationsarxiv-2604.06277 Sparse Blocked context onlyApr 7, 2026
- Stress-Testing the Reasoning Competence of LLMs With Proofs Under Minimal Formalismarxiv-2605.12524 Sparse Blocked context onlyApr 7, 2026
- Spec Kit Agents: Context-Grounded Agentic Workflowsarxiv-2604.05278 Sparse Blocked context onlyApr 7, 2026
- XMark: Reliable Multi-Bit Watermarking for LLM-Generated Textsarxiv-2604.05242 Sparse Blocked context onlyApr 6, 2026
- Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NERarxiv-2604.05158 Sparse Blocked context onlyApr 6, 2026
- RAG or Learning? Understanding the Limits of LLM Adaptation under Continuous Knowledge Drift in the Real Worldarxiv-2604.05096 Sparse Blocked context onlyApr 6, 2026
- TriAttention: Efficient Long Reasoning with Trigonometric KV Compressionarxiv-2604.04921 Direct Blocked context onlyApr 6, 2026
- Vero: An Open RL Recipe for General Visual Reasoningarxiv-2604.04917 Sparse Blocked context onlyApr 6, 2026
- HI-MoE: Hierarchical Instance-Conditioned Mixture-of-Experts for Object Detectionarxiv-2604.04908 Sparse Blocked context onlyApr 6, 2026
- Rethinking Exploration in RLVR: From Entropy Regularization to Refinement via Bidirectional Entropy Modulationarxiv-2604.04894 Sparse Blocked context onlyApr 6, 2026
- Synthetic Sandbox for Training Machine Learning Engineering Agentsarxiv-2604.04872 Sparse Blocked context onlyApr 6, 2026
- Optimizing LLM Prompt Engineering with DSPy Based Declarative Learningarxiv-2604.04869 Sparse Blocked context onlyApr 6, 2026
- MemMachine: A Ground-Truth-Preserving Memory System for Personalized AI Agentsarxiv-2604.04853 Sparse Blocked context onlyApr 6, 2026
- Full-Duplex-Bench-v3: Benchmarking Tool Use for Full-Duplex Voice Agents Under Real-World Disfluencyarxiv-2604.04847 Direct Blocked context onlyApr 6, 2026
- InfBaGel: Human-Object-Scene Interaction Generation with Dynamic Perception and Iterative Refinementarxiv-2604.04843 Sparse Blocked context onlyApr 6, 2026
- Do No Harm: Exposing Hidden Vulnerabilities of LLMs via Persona-based Client Simulation Attack in Psychological Counselingarxiv-2604.04842 Sparse Blocked context onlyApr 6, 2026
- MERIT: Multilingual Expert-Reward Informed Tuning for Chinese-Centric Low-Resource Machine Translationarxiv-2604.04839 Sparse Blocked context onlyApr 6, 2026
- CreativityBench: Evaluating Agent Creative Reasoning via Affordance-Based Tool Repurposingarxiv-2605.02910 Sparse Blocked context onlyApr 6, 2026
- Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do notarxiv-2604.04825 Sparse Blocked context onlyApr 6, 2026
- ANX: Protocol-First Design for AI Agent Interaction with a Supporting 3EX Decoupled Architecturearxiv-2604.04820 Sparse Blocked context onlyApr 6, 2026
- LiveFact: A Dynamic, Time-Aware Benchmark for LLM-Driven Fake News Detectionarxiv-2604.04815 Sparse Blocked context onlyApr 6, 2026
- SkillX: Automatically Constructing Skill Knowledge Bases for Agentsarxiv-2604.04804 Direct Blocked context onlyApr 6, 2026
- How Far Are We? Systematic Evaluation of LLMs vs. Human Experts in Mathematical Contest in Modelingarxiv-2604.04791 Sparse Blocked context onlyApr 6, 2026
- Cog-DRIFT: Exploration on Adaptively Reformulated Instances Enables Learning from Hard Reasoning Problemsarxiv-2604.04767 Sparse Blocked context onlyApr 6, 2026
- Your Agent, Their Asset: A Real-World Safety Analysis of OpenClawarxiv-2604.04759 Sparse Blocked context onlyApr 6, 2026
- AI Trust OS -- A Continuous Governance Framework for Autonomous AI Observability and Zero-Trust Compliance in Enterprise Environmentsarxiv-2604.04749 Sparse Blocked context onlyApr 6, 2026
- Lighting Up or Dimming Down? Exploring Dark Patterns of LLMs in Co-Creativityarxiv-2604.04735 Sparse Blocked context onlyApr 6, 2026
- Discovering Failure Modes in Vision-Language Models using RLarxiv-2604.04733 Sparse Blocked context onlyApr 6, 2026
- Metaphors We Compute By: A Computational Audit of Cultural Translation vs. Thinking in LLMsarxiv-2604.04732 Sparse Blocked context onlyApr 6, 2026
- Individual and Combined Effects of English as a Second Language and Typos on LLM Performancearxiv-2604.04723 Sparse Blocked context onlyApr 6, 2026
- Is a Picture Worth a Thousand Words? Adaptive Multimodal Fact-Checking with Visual Evidence Necessityarxiv-2604.04692 Sparse Blocked context onlyApr 6, 2026
- ROSClaw: A Hierarchical Semantic-Physical Framework for Heterogeneous Multi-Agent Collaborationarxiv-2604.04664 Sparse Blocked context onlyApr 6, 2026
- Search, Do not Guess: Teaching Small Language Models to Be Effective Search Agentsarxiv-2604.04651 Sparse Blocked context onlyApr 6, 2026
- Temporal Inversion for Learning Interval Change in Chest X-Raysarxiv-2604.04563 Sparse Blocked context onlyApr 6, 2026
- Multilingual Prompt Localization for Agent-as-a-Judge: Language and Backbone Sensitivity in Requirement-Level Evaluationarxiv-2604.04532 Sparse Blocked context onlyApr 6, 2026
- One Model for All: Multi-Objective Controllable Language Modelsarxiv-2604.04497 Sparse Blocked context onlyApr 6, 2026
- A Patch-based Cross-view Regularized Framework for Backdoor Defense in Multimodal Large Language Modelsarxiv-2604.04488 Sparse Blocked context onlyApr 6, 2026
- Grid2Matrix: Revealing Digital Agnosia in Vision-Language Modelsarxiv-2604.09687 Sparse Blocked context onlyApr 6, 2026
- Explainable Autonomous Cyber Defense using Adversarial Multi-Agent Reinforcement Learningarxiv-2604.04442 Sparse Blocked context onlyApr 6, 2026
- STEER: Structured Event Evidence for Video Reasoning via Multi-Objective Reinforcement Learningarxiv-2604.04415 Sparse Blocked context onlyApr 6, 2026
- How Alignment Routes: Localizing, Scaling, and Controlling Policy Circuits in Language Modelsarxiv-2604.04385 Sparse Blocked context onlyApr 6, 2026
- REAM: Merging Improves Pruning of Experts in LLMsarxiv-2604.04356 Sparse Blocked context onlyApr 6, 2026
- Self-Distilled RLVRarxiv-2604.03128 Sparse Blocked context onlyApr 3, 2026
- JoyAI-LLM Flash: Advancing Mid-Scale LLMs with Token Efficiencyarxiv-2604.03044 Curated Related Blocked context onlyApr 3, 2026
- Council Mode: A Heterogeneous Multi-Agent Consensus Framework for Reducing LLM Hallucination and Biasarxiv-2604.02923 Sparse Blocked context onlyApr 3, 2026
- OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environmentsarxiv-2605.18758 Sparse Blocked context onlyApr 3, 2026
- Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchyarxiv-2604.02709 Sparse Blocked context onlyApr 3, 2026
- Mitigating LLM biases toward spurious social contexts using direct preference optimizationarxiv-2604.02585 Sparse Blocked context onlyApr 2, 2026
- Social Meaning in Large Language Models: Structure, Magnitude, and Pragmatic Promptingarxiv-2604.02512 Sparse Blocked context onlyApr 2, 2026
- ActionParty: Multi-Subject Action Binding in Generative Video Gamesarxiv-2604.02330 Sparse Blocked context onlyApr 2, 2026
- Steerable Visual Representationsarxiv-2604.02327 Sparse Blocked context onlyApr 2, 2026
- Batched Contextual Reinforcement: A Task-Scaling Law for Efficient Reasoningarxiv-2604.02322 Sparse Blocked context onlyApr 2, 2026
- Beyond the Assistant Turn: User Turn Generation as a Probe of Interaction Awareness in Language Modelsarxiv-2604.02315 Sparse Blocked context onlyApr 2, 2026
- VOID: Video Object and Interaction Deletionarxiv-2604.02296 Sparse Blocked context onlyApr 2, 2026
- Omni123: Exploring 3D Native Foundation Models with Limited 3D Data by Unifying Text to 2D and 3D Generationarxiv-2604.02289 Sparse Blocked context onlyApr 2, 2026
- Unifying Group-Relative and Self-Distillation Policy Optimization via Sample Routingarxiv-2604.02288 Sparse Blocked context onlyApr 2, 2026
- Novel Memory Forgetting Techniques for Autonomous AI Agents: Balancing Relevance and Efficiencyarxiv-2604.02280 Sparse Blocked context onlyApr 2, 2026