From GPT-2 to gpt-oss: Analyzing the Architectural Advances
Explorar
Noticias de IA
21861 elementos — filtrados, clasificados y sin duplicados
Vision Language Model Alignment in TRL ⚡️
Estimating worst case frontier risks of open weight LLMs
Measuring Open-Source Llama Nemotron Models on DeepResearch Bench
📚 3LM: A Benchmark for Arabic LLMs in STEM and Code
TimeScope: How Long Can Your Video Large Multimodal Model Go?
OpenAI’s new economic analysis
OAI and GDM announce IMO Gold-level results with natural language reasoning, no specializ…
The Big LLM Architecture Comparison
ChatGPT agent System Card
Consilium: When Multiple LLMs Collaborate
Back to The Future: Evaluating AI Agents on Predicting Future Events
Ettin Suite: SoTA Paired Encoders and Decoders
Kimina-Prover: Applying Test-time RL Search on Large Formal Reasoning Models
Asynchronous Robot Inference: Decoupling Action Prediction and Execution
Announcing NeurIPS 2025 E2LM Competition: Early Training Evaluation of Language Models
LLM Research Papers: The 2025 List (January to June)
AlphaGenome: AI for better understanding the genome
Context Engineering: Much More than Prompts
(LoRA) Fine-Tuning FLUX.1-dev on Consumer Hardware
Toward understanding and preventing misalignment generalization
Understanding and Coding the KV Cache in LLMs from Scratch
Cognition vs Anthropic: Don't Build Multi-Agents/How to Build Multi-Agents
KV Cache from scratch in nanoVLM
CodeAgents + Structure: A Better Way to Execute Actions
🐯 Liger GRPO meets TRL
Gemini's AlphaEvolve agent uses Gemini 2.0 to find new Math and cuts Gemini cost 1% — wit…
Introducing HealthBench
LeRobot Community Datasets: The “ImageNet” of Robotics — When and How?
Coding LLMs from the Ground Up: A Complete Course