Nemotron-4-340B: NVIDIA's new large open models, built on syndata, great for syndata
Explorar
Noticias de IA
975 elementos — filtrados, clasificados y sin duplicados
The Last Hurrah of Stable Diffusion?
Talaria: Apple's new MLOps Superweapon
HippoRAG: First, do know(ledge) Graph
Expanding on how Voice Engine works and our safety research
Qwen 2 beats Llama 3 (and we don't know how)
Mamba-2: State Space Duality
1 TRILLION token context, real time, on device?
Falcon 2: An 11B parameter pretrained language model and VLM, trained on over 5000B token…
Skyfall
Chameleon: Meta's (unreleased) GPT4o-like Omnimodal Model
Cursor reaches >1000 tok/s finetuning Llama3-70b for fast file editing
Google I/O in 60 seconds
PaliGemma – Google's Cutting-Edge Open Vision Language Model
GPT-4o: the new SOTA-EVERYTHING Frontier model (GPT4T version)
GPT-4o: the new SOTA-EVERYTHING Frontier model (GPT4O version)
Hello GPT-4o
Introducing GPT-4o and more tools to ChatGPT free users
Introducing the Model Spec
DeepSeek-V2 beats Mixtral 8x22B with >160 experts at HALF the cost
Introducing the Open Leaderboard for Hebrew LLMs!
A quiet weekend
StarCoder2-Instruct: Fully Transparent and Permissive Self-Alignment for Code Generation
Apple's OpenELM beats OLMo with 50% of its dataset, using DeLighT
Snowflake Arctic: Fully Open 10B+128x4B Dense-MoE Hybrid LLM
Llama-3-70b is GPT-4-level Open Model
Meta Llama 3 (8B, 70B)
Welcome Llama 3 - Meta's new open LLM
Mixtral 8x22B Instruct sparks efficiency memes
Multi-modal, Multi-Aspect, Multi-Form-Factor AI