Incomplete Prompt Jailbreaks in Large Language Models
Explorar
Noticias de IA
30666 elementos — filtrados, clasificados y sin duplicados
CMI-Mem: Toward Generalizable Long-Term Memory Management via CMI-Augmented Reinforcement…
OpenForgeRL: Train Harness-native Agents in Any Environment
Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle a…
Agentic coding without the cloud: evaluating open-weight large language models on longitu…
Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry
AREX: Towards a Recursively Self-Improving Agent for Deep Research
The Boundaries of Automation: A Theory of Persistent Human Participation
From Scalars to Time Series: Rethinking Implicit Neural Representations for Time-Varying …
Unsupervised Consensus-Based Anomaly Detection for Spatiotemporal Malaria Incidence in Gh…
From Noise to Diversity: Random Embedding Injection in LLM Reasoning
Bridging the Gap Between Plausibility and Admissibility: Constraint-Aware Flow Maps for D…
Error Certificates for KV-Cache Eviction via Randomized Design
Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment
Bayesian uncertainty estimation improves clinical decision making in medical AI agents
MKEvolve: A Modular Multi-Agent Framework for Kernel Code Generation
M$^3$-Gen: Interpretable Multimodal Generation of Gene Expression Profiles Using Clinical…
Self-Supervised Bio-Inspired Robotic Trajectory Planning with Obstacle Avoidance
Routing Without Training: Controllable-Ratio LLM Offloading via Reliability Gating
NVIDIA-labs OO Agents: Native Python Object-Oriented Agents
Autonomous disproofs of the sum-product conjecture over $\mathbb R$ with GPT-5.5 Pro
SPORD: A Simulation-Propose-then-OR-Dispose Approach for Supply Chain Planning
StackingNet: Collective Inference Across Independent AI Foundation Models
Diagnosing Pathological Chain-of-Thought in Reasoning Models
WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance
Multimodal Pretraining for Generalizable EEG Representation Learning
Logic Programming Semantics for Causal Processes
BasketEvent: Understanding Who Did What and When in Basketball Videos
Fragile Preferences: A Deep Dive Into Order Effects in Large Language Models
Expert Behavior Prior Reinforcement Learning