Causal-JEPA: Learning World Models through Object-Level Latent Masking
Explorar
Noticias de IA
21272 elementos — filtrados, clasificados y sin duplicados
Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure
RewardFlow: Topology-Aware Reward Propagation on State Graphs for Agentic RL with Large L…
AgentDropoutV2: Optimizing Information Flow in Multi-Agent Systems via Test-Time Rectify-…
SCoOP: Semantic Consistent Opinion Pooling for Uncertainty Quantification in Multiple Vis…
MemCollab: Cross-Model Memory Collaboration via Contrastive Trajectory Distillation
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in M…
Guardrails Beat Guidance: A Large-Scale Study of Rules, Skills, and Persistent Configurat…
SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning
SciHorizon-DataEVA: An Agentic System for AI-Readiness Evaluation of Heterogeneous Scient…
Human-Guided Harm Recovery for Computer Use Agents
AttuneBench: A Conversation-Based Benchmark for LLM Emotional Intelligence
NOVA: Fundamental Limits of Knowledge Discovery Through AI
Are LLMs Socially Adaptive? Contrasting Belief Evolution in Large Language Models and Hum…
Crafting Desirable Climate Trajectories with RL Explored Socio-Environmental Simulations
Benchmarking Open-Source Safety Guard Models: A Comprehensive Evaluation
MATNet: Multi-Level Fusion Transformer-Based Model for Day-Ahead PV Generation Forecasting
Weakly Supervised Detection and Temporal Localization of Whale Calls in Long-Duration Bio…
Agent4Edu: Generating Learner Response Data by Generative Agents for Intelligent Educatio…
SchGen: PCB Schematic Generation with Semantic-Grounded Code Representations
Locally Coherent, Globally Incoherent: Bounding Compositional Incoherence in Multi-Compon…
VRAG: Learning World Models for Interactive Video Generation
Is Your LLM Overcharging You? Tokenization, Transparency, and Incentives
ProjectionBench: Evaluating Scientific Hypothesis Generation in LLMs Under Progressive In…
When Should Models Change Their Minds? Contextual Belief Management in Large Language Mod…
Taming Data Challenges in ML-based Security Tasks Using Generative AI
Approximate Proportionality in Online Fair Division
Towards Foundation Models for Zero-Shot Time Series Anomaly Detection: Leveraging Synthet…
VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior T…
Less Is More: Elevating RAG via Performance-Driven Context Compression