VibeGame: Exploring Vibe Coding Games
Explorar
Noticias de IA
21010 elementos — filtrados, clasificados y sin duplicados
Accelerating Qwen3-8B Agent on Intel® Core™ Ultra with Depth-Pruned Draft Models
Nemotron-Personas-Japan: ソブリン AI のための合成データセット
Measuring the performance of our models on real-world tasks
GDPVal finding: Claude Opus 4.1 within 95% of AGI (human experts in top 44 white collar j…
Smol2Operator: Post-Training GUI Agents for Computer Use
Gaia2 and ARE: Empowering the community to study agents
Detecting and reducing scheming in AI models
How people are using ChatGPT
Why language models hallucinate
SAIR: Accelerating Pharma R&D with AI-Powered Structural Intelligence
It’s the Humidity: How International Researchers in Poland, Deep Learning and NVIDIA GPUs…
Accelerating life sciences research
NVIDIA Releases 6 Million Multi-Lingual Reasoning Dataset
Kimina-Prover-RL
Neural Super Sampling is here!
TextQuests: How Good are LLMs at Text-Based Video Games?
🇵🇭 FilBench - Can LLMs Understand and Generate Filipino?
NVIDIA Research Shapes Physical AI
From GPT-2 to gpt-oss: Analyzing the Architectural Advances
Vision Language Model Alignment in TRL ⚡️
Estimating worst case frontier risks of open weight LLMs
Measuring Open-Source Llama Nemotron Models on DeepResearch Bench
📚 3LM: A Benchmark for Arabic LLMs in STEM and Code
TimeScope: How Long Can Your Video Large Multimodal Model Go?
OpenAI’s new economic analysis
OAI and GDM announce IMO Gold-level results with natural language reasoning, no specializ…
The Big LLM Architecture Comparison
ChatGPT agent System Card
Consilium: When Multiple LLMs Collaborate