Zero-Fi: Zero-Shot Wi-Fi-Based Human Activity Recognition via Contrastive Signal-Language…
Explorar
Noticias de IA
22065 elementos — filtrados, clasificados y sin duplicados
The Age of AI Agents Demands A New Scientific Paradigm To Sustain Trustworthy Science
A Picture Says Thousands of Words - Harnessing Dermal Exposure Data from Images through H…
GuidedRAG: Semantic Steering of Retrieval-Augmented Generation
SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search
Emergent Sparsity in Frozen Random CNN Feature Extractors for Deep Reinforcement Learning
Sim2Win: A Team-Agnostic, Event-Based Pre-Match Outcome Prediction and Tactical Profiling…
Do Methods Support the Claims? Intra-Paper Verification for Peer Review
Forensic Reproducibility Audit of a Radiology Vision-Language Model Benchmark: From Inten…
Identifying Implicit Bias in LLM-based Chat AI Toward People with Intellectual Disabiliti…
Archetypes or ability? Clustering for modelling student mathematical competence
IDP AutoOpt: Agent-Driven Optimization of Document Processing Pipeline Configurations
Can AI agents conduct open-ended AI research? Early evidence from two case studies
Large-Scale ChatBot Validation Through Customer Digital Twin Simulations
On-Policy Distillation for LLM Safety: A Routing Approach to Template-Robust Realignment
What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection un…
Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating …
AgenticCANN: Automated Ascend C Operator Generation via Knowledge-Augmented Agentic Evolu…
Do Latent Channels Actually Communicate? A Causal Audit of Latent Multi-Agent LLM
Evidence-Ledger Adjudication for Claim-Evidence Traceability
Optimizing Sensor Placement for Hydrogen Leak Detection in Enclosed Infrastructure: A Com…
MultivationBench: A Benchmark for Multimodal Sequential Motivation Reasoning
Fewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adap…
From Passive Video to Editable Experience: Physically Grounded Experience Synthesis for E…
Position: Evaluation Scores Are Perishable Knowledge Claims
When benchmark inferences do not compose: Projectibility in AI evaluation
Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems
GuideSkill: Evolving Executable LLM Agent Skills for Guideline-Grounded Clinical Reasoning
ClinLens: Towards Long-Horizon Coding Agents for Longitudinal Multimodal Clinical Data Sc…
CaM-Wolf: Causal-Aware Multimodal Agents for Social Deduction Games