Agent Security Needs Redefinition through a Holistic Framework
Explorar
Noticias de IA
30641 elementos — filtrados, clasificados y sin duplicados
EVL-MCoT: Enhanced Vision-Language Multi-CoT for Harmful Meme Detection
Practical Graph Optimisation and AI-Driven Models for Active Directory Security Hardening
J-CoT: Chain-of-Thought in J-Space
ACME: A Multi-Cultural, Multi-Embodiment Social-Navigation Dataset
MA-DAR: Manifold-Aligned Dynamic Adaptive Routing for Continual Temporal Knowledge Graph …
A Defense of the Quadratic Model
Enhancing SLMs for Sustainable Code Optimization in Radio-Astronomy
Ordered Action Tokens for Visuomotor Policy Learning
Generative and multimodal AI for materials prediction and design: Progress, challenges, a…
Cross-Model LLM Code Review: Should you use Claude to review Codex or vice versa?
Data Quality over Capacity: Internalizing Documents into LoRA Adapters for Closed-Book QA
Probing Latent Colombian Identity Inferences in Qwen2.5-7B with Natural Language Autoenco…
A Roadmap to Impactful Pluralistic Alignment Research
AI4PLE: A Methodology for Integrating AI into Product Line Engineering
Deconstructing Off-Policy Ratios: Entropy-Scaled Trust Regions for Asynchronous Reinforce…
Learning on the Job: Continual Learning from Deployment Feedback for Frozen-Weights Agents
Industrial Tokenization for LLM-Based Health Intelligence: A Federated Architecture for I…
Reasoning Denoiser: Denoising Reasoning Traces for Hallucination Detection in Large Reaso…
Learning as Reasoning Unfolds: Progressive Rollout Allocation for Efficient Reinforcement…
Semiotic logical hexagon theory for LLM logical reasoning
TRW: TRACE-RealWorld---An Auditable Consistency Contract for World Models as Materialized…
Multi-Agent System-driven Digital Twins for predictive maintenance: architectures, techno…
When Is a Learned Command Adapter Worth It? Closed-Loop Identification and Counterfactual…
DAGForge: Auditable Causal DAG Authoring with Biomedical Literature
Co-design of LLM-based preference agents: participation may drive overtrust
QLPO: Quadrant-weighted Sampling for Length-aware Policy Optimization
From Seasonality to Semantics: Benchmarking a Hybrid Probabilistic Forecasting System for…
What AI Red-Team Evaluations Can and Cannot Prove
Persistent Computational State: A Session-Centric Runtime for Generative World Models