Finetuning olmOCR to be a faithful OCR-Engine
Explorar
Noticias de IA
676 elementos — filtrados, clasificados y sin duplicados
Prefill and Decode for Concurrent Requests - Optimizing LLM Performance
17 Reasons Why Gradio Isn't Just Another UI Library
Cohere on Hugging Face Inference Providers 🔥
Google's Agent2Agent Protocol (A2A)
Journey to 1 Million Gradio Users!
Efficient Request Queueing – Optimizing LLM Performance
How Hugging Face Scaled Secrets Management for AI Infrastructure
🚀 Accelerating LLM Inference with TGI on Intel Gaudi
Training and Finetuning Reranker Models with Sentence Transformers
Introducing Gradio's new Dataframe!
Open R1: How to use OlympicCoder locally for coding
The new OpenAI Agents Platform
New tools for building agents
LLM Inference on Edge: A Fun and Easy Guide to run LLMs via React Native on your Phone!
Trace & Evaluate your Agent with Arize Phoenix
Remote VAEs for decoding with Inference Endpoints 🤗
Introducing Three New Serverless Inference Providers: Hyperbolic, Nebius AI Studio, and N…
Fixing Open LLM Leaderboard with Math-Verify
Welcome Fireworks.ai on the Hub 🎆
From Chunks to Blocks: Accelerating Uploads and Downloads on the Hub
Build awesome datasets for video generation
Open-source DeepResearch – Freeing our search agents
The AI tools for Art Newsletter - Issue 1
How to deploy and fine-tune DeepSeek models on AWS
Welcome to Inference Providers on the Hub 🔥
We now support VLMs in smolagents!
Introducing multi-backends (TRT-LLM, vLLM) support for Text Generation Inference
Timm ❤️ Transformers: Use any timm model with transformers
Train 400x faster Static Embedding Models with Sentence Transformers