← Все новости

not much happened today

**Alibaba** released the **Qwen 3.5** series with models ranging from **0.8B to 9B** parameters, featuring **native multimodality**, **scaled reinforcement learning**, and targeting **edge and lightweight agent** deployments. The models support very long context windows up to **262K tokens** (extendable to 1M) and use a novel **Gated DeltaNet hybrid attention** architecture combining linear and full attention layers. Deployment examples include **Ollama** and **LM Studio**, with a notable **6-bit on-device demo on iPhone 17 Pro**. Evaluators are cautioned that reasoning is disabled by default on smaller models. In coding agents, **Codex 5.3** shows promising benchmark results on **WeirdML** with **79.3%** accuracy, though availability and downtime remain critical challenges, especially highlighted by **Claude** outages. Agent reliability and observability are emphasized as cross-functional problems requiring clear success criteria and practical evaluation strategies. Studies show that using **AGENTS.md** and **SKILL.md** guardrails can significantly reduce runtime and token usage by mitigating worst-case thrashing in coding workflows.
Читать оригинал на AINews / smol.ai →