From Collaboration to Capability: Internalizing Routed LLM Experts into Compact Reasoners
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Beyond Generation and Accuracy: Diagnosing and Enhancing Visual Chain-of-Thought for Geom…
MPT: Missing Prototype Tracking via Barycentric Reconstruction in Vehicular Federated Lea…
Do LLMs Trust the Accuser or the Accusation? Measuring Belief Shifts in Werewolf
How Good Are Frontier Models at Physics? Expert Re-Grading Reveals Broken Evaluations and…
Beyond Vector Similarity: Hierarchical Context-Aware Graph RAG vs Standard RAG in Enterpr…
Calibrated Ambiguity in Multimodal Language Models: Humans reach for cultural references,…
Niching Agents in The Core
Decentralized Evolution of Hexapod Gaits with Independent Leg Controllers
BlueLM-GUI Technical Report: A Real-Device-Centric Flywheel for Self-Improving Mobile GUI…
Beyond ID Embeddings: Process-Grounded Language Modeling for Cognitive Diagnosis
TileNet: Tile-Based CNN-SVM Architecture for Autonomous Unmanned Aerial Systems Inspectio…
Hierarchical Belief Modeling for Zero-Shot Opponent Adaptation in Partially Observable Mu…
Can We Trust LLM Judges: A Study of Capability-Dependent Biases and Multi-Judge Ensemble …
TripPattern: A Pattern-based Text Watermarking Method for Large Language Models
Do Influence-Derived Data Perturbations Enable Machine Unlearning? A Controlled Study of …
When Rubrics Fail: Hallucinations Reveal Blind Spots in Medical AI Evaluation
SoK: Rethinking Jailbreaking in the Era of Agentic AI: Attacks, Defenses, and Practical C…
Assisted Spatial Cognition Through Vision-Language Models
Diffusion Models and Concept Formation
Can LLMs in Draft-Verify-Revise Pipelines Resolve Deictic Ambiguity?
GLARE: Generative Learning via Adversarial Reward Estimation For Social Dynamics Forecast…
WinSyn: An Automated Pipeline for Realistic Enterprise Question-Answering Evaluation
MInTRL: Off-policy Intervention can boost On-policy RL
Learning Symbolic Constraint Representations from Examples: A Neuro-Symbolic Approach
Debiasing as a Measurement Intervention: Calibrated Ties and Resolution Loss in LLM-as-a-…
Harness or Model? Isolating the Harness Effect in Agentic Coding with a Contamination-Con…
DU-NO: A Parameter-Efficient Double U-Shaped Neural Operator for Phase-Resolving Wave Mod…
Automated Detection and Structuring of Social Tipping Point Evidence in Climate related D…
DenseTRF: Texture-Aware Unsupervised Representation Adaptation for Surgical Scene Dense P…