Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

arXiv cs.AI Research & Papers
GrepSeek: Training Search Agents for Direct Corpus Interaction
arXiv cs.AI Research & Papers
OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajecto…
arXiv cs.AI Research & Papers
Surfacing Isolated Learners with Outcome-Independent Mediation of Feedback between Teache…
arXiv cs.AI Research & Papers
AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Cryst…
arXiv cs.AI Research & Papers
Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting
arXiv cs.AI Research & Papers
CoHyDE: Iterative Co-Training of LLM Rewriter & Dense Encoder for Tool Retrieval
arXiv cs.AI Research & Papers
Better Later Than Sooner: Neuro-Symbolic Knowledge Graph Construction via Ontology-ground…
arXiv cs.AI Research & Papers
Scaling Monosemanticity: Extracting Interpretable Features from Claude 3 Sonnet
arXiv cs.AI Research & Papers
GroundAct: Can LLM Agents Ground Actions in Environmental States?
arXiv cs.AI Research & Papers
MENTOR: Efficient Multimodal-Conditioned Tuning for Autoregressive Vision Generation Mode…
arXiv cs.AI Research & Papers
HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous S…
arXiv cs.AI Research & Papers
Position: Text Embeddings Should Capture Implicit Semantics, Not Just Surface Meaning
arXiv cs.AI Research & Papers
EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance
arXiv cs.AI Research & Papers
GPIC: A Giant Permissive Image Corpus for Visual Generation
arXiv cs.AI Research & Papers
Steering Language Models Before They Speak: Logit-Level Interventions
arXiv cs.AI Research & Papers
The Confidence Shortcut: A Reasoning Failure Mode of Masked Diffusion Models
arXiv cs.AI Research & Papers
Dataset-Driven Channel Masks in Transformers for Multivariate Time Series
arXiv cs.AI Research & Papers
Aligned but Fragile: Enhancing LLM Safety Robustness via Zeroth-Order Optimization
arXiv cs.AI Research & Papers
Towards Human-Like Interactive Speech Recognition With Agentic Correction and Semantic Ev…
arXiv cs.AI Research & Papers
DeepSurvey: Enhancing Analytical Depth and Citation Reliability in Automated Survey Gener…
arXiv cs.AI Research & Papers
A Survey on Recent Advances in Conversational Data Generation
arXiv cs.AI Research & Papers
A Language-Guided Bayesian Optimization for Efficient LoRA Hyperparameter Search
arXiv cs.AI Research & Papers
ParaTool: Shifting Tool Representations from Context to Parameters
arXiv cs.AI Research & Papers
GPS-Enhanced Tourist Mobility Modeling with Seasonal Spatial Priors and LLM-Based Activit…
arXiv cs.AI Research & Papers
Beyond Normalization: Rethinking the Partition Function as a Difficulty Scheduler for RLVR
arXiv cs.AI Research & Papers
Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought
arXiv cs.AI Research & Papers
MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing
arXiv cs.AI Research & Papers
When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodi…
arXiv cs.AI Research & Papers
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With N…
arXiv cs.AI Research & Papers
Revisiting the Effectiveness of LLM Pruning for Test-Time Scaling