← Todas las noticias

not much happened today

**Helium-1 Preview** by **kyutai_labs** is a **2B-parameter multilingual base LLM** outperforming **Qwen 2.5**, trained on **2.5T tokens** with a **4096 context size** using token-level distillation from a **7B model**. **Phi-4 (4-bit)** was released in **lmstudio** on an **M4 max**, noted for speed and performance. **Sky-T1-32B-Preview** is a **$450 open-source reasoning model** matching **o1's performance** with strong benchmark scores. **Codestral 25.01** by **mistralai** is a new SOTA coding model supporting **80+ programming languages** and offering **2x speed**. Innovations include **AutoRAG** for optimizing retrieval-augmented generation pipelines, **Agentic RAG** for autonomous query reformulation and critique, **Multiagent Finetuning** using societies of models like **Phi-3**, **Mistral**, **LLaMA-3**, and **GPT-3.5** for reasoning improvements, and **VideoRAG** incorporating video content into RAG with LVLMs. Applications include a dynamic UI AI chat app by **skirano** on **Replit**, **LangChain** tools like **DocTalk** for voice PDF conversations, AI travel agent tutorials, and news summarization agents. **Hyperbolic Labs** offers competitive GPU rentals including **H100**, **A100**, and **RTX 4090**. **LLMQuoter** enhances RAG accuracy by identifying key quotes. Infrastructure updates include **MLX export** for LLM inference from Python to C++ by **fchollet** and **SemHash** semantic text deduplication by **philschmid**.
Leer el original en AINews / smol.ai →