Toward AI-Resilient Assessment in Computer Science Courses in an AI-Native World
Explorar
Noticias de IA
30934 elementos — filtrados, clasificados y sin duplicados
Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization
TreeAgent: A Generalizable Multi-Agent Framework for Automated Bias Labeling in Forestry …
Creating Intelligence: A Computational Foundation for AGI
Large Databases Need Small, Open-Weight Language Models
Spatial Reasoning via Modality Switching Between Language and Symbolic Representation
AgentBound: Verifiable Behavioral Governance for Autonomous AI Agents
Delta-JEPA: Learning Action-Sensitive World Models via Latent Difference Decoding
Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation…
Towards Inclusive Mobility Modeling: Characterizing and Evaluating Elderly Trajectory Pat…
AI-Assisted Discovery of Convex Relaxations via Dual Agents
Cross-Domain Feature Expansion for Tabular Medical Data via Knowledge Graphs Injection
The Past Is Prologue: A Plug-in Controller for Selective Updates in Sequentially Evolving…
Beyond expert users: agents should help users construct preferences, not just elicit them
How Can AI Find My Model? A Model-Finding Experimental Study Considering Data Formats, Em…
What Drives Interactive Improvement from Feedback?
A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimizati…
Understanding and Evaluating Claw-like Agent Security Through a Computer-Systems Lens
Learning Structurally Consistent Representations for Multi-View Radar Semantic Segmentati…
A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks…
Improving Certified Robustness via Adversarial Distillation
Histogram-constrained Image Generation
MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments
LUNA: Learning Universal 3D Human Animation Beyond Skinning
FLORA: A deep learning approach to predict forest attributes from heterogeneous LiDAR data
Freeform Preference Learning for Robotic Manipulation
Deductive Logic in Language Models: Horizontal vs Vertical Reasoning
QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents
Quantum Flow Matching
Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking