Explorar

Noticias de IA

27413 elementos — filtrados, clasificados y sin duplicados

Apple Machine Learning Research Research & Papers
Agent Seer: Synthesizing Scenarios from Specification Understanding
Hugging Face Blog Research & Papers
The Open ASR Leaderboard Adds Its First Global South Language
Google DeepMind Research & Papers
Piloting the world's first double-blind AI evaluations
OpenAI News Research & Papers
Better answers, broader thinking: What students gain from ChatGPT and critical-thinking t…
Apple Machine Learning Research Research & Papers
From Preferences to Principles: Rubric-Based Alignment for Grounded Knowledge Answers
Amazon Science Research & Papers
When LLM judges agree, should we believe them?
AWS Machine Learning Blog Research & Papers
Preparing data for supervised fine-tuning Part 2: Advanced data strategies
AWS Machine Learning Blog Research & Papers
Preparing data for supervised fine-tuning Part 1: Formatting and quality
Latent Space (swyx) Research & Papers
🔬“We have foundation models for language, not for physics” — Anima Anandkumar, Bren Profe…
OpenAI News Research & Papers
Learning never stops: How AI makes learning continuous
MIT Technology Review AI Research & Papers
AI models flub these intelligence tests. Can you fare any better?
arXiv cs.AI Research & Papers
A Human-Factors Guided Cognitive Model of Visuospatial Complexity in Embodied Active Visi…
arXiv cs.AI Research & Papers
Confidently Wrong, Silently So: Auditing Undetectable Failures of a Deployed On-Device La…
arXiv cs.AI Research & Papers
CoMMa: Contribution-Aware Medical Multi-Agents for Decentralized Oncology Decision Support
arXiv cs.AI Research & Papers
Fidelity Preference, Not Demographic Preference: A Pixel-Level Attribute-Sensitivity Audi…
arXiv cs.AI Research & Papers
Ensemble of Convolutional Neural Networks for StrokePrediction: Towards Improved Diagnost…
arXiv cs.AI Research & Papers
Method, Mind, and Morality: How People Make Sense of Artificial Intelligence
arXiv cs.AI Research & Papers
FedV-KGQA: Multi-Hop Question Answering over Vertically Partitioned Knowledge Graphs
arXiv cs.AI Research & Papers
Benchmarking LLM Judges for Voice-Agent Evaluation: Reliability, Calibration, and Human O…
arXiv cs.AI Research & Papers
ToolRobustBench: Stage-Wise Perturbation Evaluation and Failure Diagnosis for Tool-Callin…
arXiv cs.AI Research & Papers
ExPhy: A Benchmark for Explicit Physical Property Learning in Multi-Object Trajectory For…
arXiv cs.AI Research & Papers
ChorusTIC: Training-Free Multivariate Time Series Classification via Chorus In-Context Le…
arXiv cs.AI Research & Papers
PARTAB: Partition-Aware Reasoning with Structured Evidence for Scalable Table Understandi…
arXiv cs.AI Research & Papers
StepGuard: Learning Step-Level Guardrails with Scalable Supervision and Safety-Utility Ba…
arXiv cs.AI Research & Papers
When Less Is More: An Empirical Study of Minimal Responses in Counseling Dialogues and th…
arXiv cs.AI Research & Papers
Identifying Latent Declarative Representations of Code for Assisting Repository Migration
arXiv cs.AI Research & Papers
Quasar: A Programming Language Specialized for LLM Code Actions
arXiv cs.AI Research & Papers
Robust Motion Generation using Part-level Reliable Data from Videos
arXiv cs.AI Research & Papers
PatientHub: A Unified Framework for Patient Simulation
arXiv cs.AI Research & Papers
CoWorld-VLA: Thinking in a Multi-Expert World Model for Autonomous Driving