Introducing AutoRound: Intel’s Advanced Quantization for LLMs and VLMs
Explorar
Noticias de IA
37834 elementos — filtrados, clasificados y sin duplicados
Qwen 3: 0.6B to 235B MoE full+base models that beat R1 and o1
PipelineRL
Cognition's DeepWiki, a free encyclopedia of all GitHub repos
Tiny Agents: an MCP-powered agent in 50 lines of code
not much happened today
New in ChatGPT for Business: April 2025
Introducing our latest image generation model in the API
gpt-image-1 - ChatGPT's imagegen model, confusingly NOT 4o, now available in API
Finetuning olmOCR to be a faithful OCR-Engine
Speak is personalizing language learning with AI
The Washington Post partners with OpenAI on search content
not much happened today
not much happened today; New email provider for AINews
The State of Reinforcement Learning for LLM Reasoning
Grok 3 & 3-mini now API Available
Gemini 2.5 Flash completes the total domination of the Pareto Frontier
OpenAI o3, o4-mini, and Codex CLI
QwQ-32B claims to match DeepSeek R1-671B
Prefill and Decode for Concurrent Requests - Optimizing LLM Performance
Thinking with images
OpenAI o3 and o4-mini System Card
Introducing OpenAI o3 and o4-mini
SOTA Video Gen: Veo 2 and Kling 2 are GA for developers
17 Reasons Why Gradio Isn't Just Another UI Library
Cohere on Hugging Face Inference Providers 🔥
Introducing HELMET: Holistically Evaluating Long-context Language Models
OpenAI announces nonprofit commission advisors
GPT 4.1: The New OpenAI Workhorse
Our updated Preparedness Framework