Explorar

Noticias de IA

21270 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
HiSpec: Hierarchical Speculative Decoding for LLMs
arXiv cs.AI Research & Papers
EconCausal: A Context-Aware Economic Reasoning Benchmark for Large Language Models
arXiv cs.AI Research & Papers
Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights
arXiv cs.AI Research & Papers
Mechanistic Interpretability of Antibody Language Models Using SAEs
arXiv cs.AI Research & Papers
SEAL: Self-Evolving Agentic Learning for Conversational Question Answering over Knowledge…
arXiv cs.AI Research & Papers
Seeing vs. Believing: Evaluating the Language Bias of Open-Source MLLMs in Counter-Intuit…
arXiv cs.AI Research & Papers
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic
arXiv cs.AI Research & Papers
2-ASP(Q) programs with weak constraints: Complexity and efficient implementation
arXiv cs.AI Research & Papers
Drive-P2D: A Progressive Perception-to-Decision Benchmark for VLMs in Autonomous Driving
arXiv cs.AI Research & Papers
The AI Cognitive Trojan Horse: How Large Language Models May Bypass Human Epistemic Vigil…
arXiv cs.AI Research & Papers
Left-Right Symmetry Breaking in CLIP-style Vision-Language Models Trained on Synthetic Sp…
arXiv cs.AI Research & Papers
GraphDancer: Training LLMs to Explore and Reason over Graphs via Two-Stage Curriculum Pos…
arXiv cs.AI Research & Papers
Constructing Industrial-Scale Optimization Modeling Benchmark
arXiv cs.AI Research & Papers
Adapting Actively on the Fly: Relevance-Guided Online Meta-Learning with Latent Concepts …
arXiv cs.AI Research & Papers
Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User Hist…
arXiv cs.AI Research & Papers
Natural Language Query to Configuration for Retrieval Agents
arXiv cs.AI Research & Papers
PaTAS: A Framework for Trust Propagation in Neural Networks Using Subjective Logic
arXiv cs.AI Research & Papers
GICDM: Mitigating Hubness for Reliable Distance-Based Generative Model Evaluation
arXiv cs.AI Research & Papers
Geometrically Constrained Outlier Synthesis
arXiv cs.AI Research & Papers
AI-Driven Contribution Evaluation and Conflict Resolution: A Framework & Design for Group…
arXiv cs.AI Research & Papers
Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility
arXiv cs.AI Research & Papers
Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Reliability
arXiv cs.AI Research & Papers
FedTreeLoRA: Reconciling Statistical and Functional Heterogeneity in Federated LoRA Fine-…
arXiv cs.AI Research & Papers
Omanic: Towards Step-wise Evaluation of Multi-hop Reasoning in Large Language Models
arXiv cs.AI Research & Papers
APEX-Searcher: Refining Credit Assignment with Subgoaling for Agentic Retrieval-Augmented…
arXiv cs.AI Research & Papers
Detached Skip-Links and $R$-Probe: Decoupling Feature Aggregation from Gradient Propagati…
arXiv cs.AI Research & Papers
Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argum…
arXiv cs.AI Research & Papers
From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Ques…
arXiv cs.AI Research & Papers
Understanding the Challenges in Iterative Generative Optimization with LLMs
arXiv cs.AI Research & Papers
SenBen: Sensitive Scene Graphs for Explainable Content Moderation