Explorar

Noticias de IA

37834 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
Knowing When to Stop: Bayesian Optimal Stopping for LLM Evaluations
arXiv cs.AI Research & Papers
AI Research Preference Models
arXiv cs.AI Research & Papers
Residual Dominance as a Structural Account of Last-Item Reliance in Causal Self-Attention…
arXiv cs.AI Research & Papers
A Four-Axis Trustworthiness Benchmark for LLM-as-Judge in Principle-Based Regulation
arXiv cs.AI Research & Papers
When Personal Memory Has No Single Answer: Evaluating LLM Agents under Irreducible Confli…
arXiv cs.AI Research & Papers
The Dynamics of Intelligence Explosions
arXiv cs.AI Research & Papers
Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages
arXiv cs.AI Research & Papers
LLMs Don't Pay for the Jump
arXiv cs.AI Security & Safety
Tripwire: Triggering Aligned Refusal via Statistically Certified Safety Neurons
arXiv cs.AI Research & Papers
The Past and Future of AI Scientists
arXiv cs.AI Research & Papers
AnchorBench: A Multi-Pathway Benchmark for the Anchoring Effect in LLMs
arXiv cs.AI Research & Papers
Sensor-Driven Mission Synthesis for UAV/UGV Swarms: A TB-CSPN Coordination Architecture w…
arXiv cs.AI Research & Papers
Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning
arXiv cs.AI Research & Papers
Clearing the Fog: Towards Installing and Refining Proactive Exploration Capabilities in L…
arXiv cs.AI Research & Papers
BCMT: Blockwise Causal Memory Transformer
arXiv cs.AI Research & Papers
ScienceFlow: A long-horizon agent for ML research, scientific discovery and beyond
arXiv cs.AI Research & Papers
Grounding Without Corrective Control: Truth-Tracking Profiles for Large Language Models
arXiv cs.AI Research & Papers
Disentangled Shared Representations Improve Morpho-Transcriptomic Integration
arXiv cs.AI Research & Papers
From Style Replication to Style Exploration: Enabling Art Style Exploration with Analyze-…
arXiv cs.AI Research & Papers
Attributing Preprocessing Invariance in Spectral Foundation Models
arXiv cs.AI Research & Papers
MathForm: Scaling Mathematical Autoformalization with Knowledge Retrieval and Verificatio…
arXiv cs.AI Research & Papers
Can Language Models Understand mmWave Data? Benchmarking Large Language Models for mmWave…
arXiv cs.AI Research & Papers
BiasTrace: Linking Reasoning Behaviours to Biased Outputs in LLMs
arXiv cs.AI Research & Papers
Towards Efficient Multimodal and Multilingual Opinion Extraction for STI: A QLoRA-Based F…
arXiv cs.AI Research & Papers
Removing Temporal Note Redundancy Improves Multimodal Reinforcement Learning for Medicine
arXiv cs.AI Research & Papers
Polaris : Multi Agentic System for Conversational Enterprise Analytics
arXiv cs.AI Research & Papers
GBU-Palm: A Multimodal Video Dataset and Benchmark for Palm Presentation Attack Detection
arXiv cs.AI Research & Papers
A Hybrid LLM-Based Framework for Automated Security Annotation Generation in Business Pro…
arXiv cs.AI Research & Papers
ASSERT: A Measurement Pipeline for GenAI Audits
arXiv cs.AI Research & Papers
QuaSAR: Quantization Compensation via Stable Activation-Aware Rank Truncation