Pramana: A Composable, Domain-Specific Backend for Empirical Networking Research
Explorar
Noticias de IA
30329 elementos — filtrados, clasificados y sin duplicados
One Run Is Not an Idea: The Implementation Lottery in Automated Research
(EC)2: Event-Centric Explainability for Cybersecurity Through Multi-Agent LLM Investigati…
PUDA: An AI-Native Hardware Harness for Self-Driving Laboratories
Towards Trustworthy Embodied Intelligence: A Systems Framework and Graded Trustworthiness…
Guarding Organizations Against Malware Risk: A Novel Graph-Based Malware Detection Method
SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response
AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control
Recover, Decode, Reguard: Guard-Agnostic Defense Amplification againstEncoded VLM Jailbre…
WhisperRec: Latent Reasoning for Efficient Foundation Recommendation Models
Fewer Clarifications, Better Code: Benchmarking Cross-Session Personalized Ambiguity Adap…
Do Latent Channels Actually Communicate? A Causal Audit of Latent Multi-Agent LLM
IDP AutoOpt: Agent-Driven Optimization of Document Processing Pipeline Configurations
Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems
Reinforcement Learning on Cost-Constrained Quadrupedal Hardware
Balancing Centralized Learning and Distributed Self-Organization: A Hybrid Model for Embo…
Emergent Sparsity in Frozen Random CNN Feature Extractors for Deep Reinforcement Learning
Identifying Implicit Bias in LLM-based Chat AI Toward People with Intellectual Disabiliti…
Voice Memory for Agentic Speech Recognition
When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Respons…
Human diversity fuels collective creativity that large language models cannot simulate or…
Collaborative Weighting with Pessimistic Critic for Mitigating Overestimation in Off-Poli…
Archetypes or ability? Clustering for modelling student mathematical competence
From Representations to Behaviors: Exploring the Person-Situation-Behavior Triad in LLMs
BioVLN: A Simulation Platform for Visual Language Navigation in Biomedical Laboratories
MMAC: A Massive Multi-dimensional Benchmark for Audio Captioning
Measure what Matters: Psychometric Evaluation of AI with Situational Judgment Tests
Visual Credit Audit for Multimodal Spatial Reasoning
AI as Friction for Reflection Support in Ideation
Property-driven Causal Abstractions for Markov Decision Processes