← Todas las noticias

not much happened today

**xAI released Grok 4.3**, improving cost/performance with a **53 Intelligence Index score**, 4 points higher than Grok 4.20, and significant gains on **GDPval-AA** and **τ²-Bench Telecom**. However, accuracy tradeoffs raised reliability concerns. Community opinions are mixed, with some praising token-efficiency and others noting regressions and pricing concerns. **DeepSeek V4 Pro** emerges as a leading open-weight coding/agent model, comparable to **Codex** and **Claude Code**, featuring a 1M context window and efficient attention mechanisms. Benchmarking shows open-weight models like **Kimi K2.6**, **MiMo V2.5 Pro**, and **DeepSeek V4 Pro** closing the gap with closed models such as **Gemini 3.1 Pro Preview**, **Claude Opus 4.7**, and **GPT-5.5**. DeepSeek's multimodal efforts focus on explicit spatial grounding with a novel "point while thinking" approach using **DeepSeek-ViT** and CSA compression.
Leer el original en AINews / smol.ai →