Jupyter Agents: training LLMs to reason with notebooks
Explorar
Noticias de IA
773 elementos — filtrados, clasificados y sin duplicados
Make your ZeroGPU Spaces go brrr with ahead-of-time compilation
From Zero to GPU: A Guide to Building and Scaling Production-Ready CUDA Kernels
MCP for Research: How to Connect AI to Research Tools
Arm & ExecuTorch 0.7: Bringing Generative AI to the masses
Accelerate ND-Parallel: A guide to Efficient Multi-GPU Training
Implementing MCP Servers in Python: An AI Shopping Assistant with Gradio
Introducing Trackio: A Lightweight Experiment Tracking Library from Hugging Face
Say hello to `hf`: a faster, friendlier Hugging Face CLI ✨
Parquet Content-Defined Chunking
Fast LoRA inference for Flux with Diffusers and PEFT
Accelerate a World of LLMs on Hugging Face with NVIDIA NIM
Five Big Improvements to Gradio MCP Servers
ScreenEnv: Deploy your full stack Desktop Agent
Building the Hugging Face MCP Server
Upskill your LLMs With Gradio MCP Servers
Creating custom kernels for the AMD MI300
Efficient MultiModal Data Pipeline
Three Mighty Alerts Supporting Hugging Face’s Production Infrastructure
Transformers backend integration in SGLang
Groq on Hugging Face Inference Providers 🔥
How Long Prompts Block Other Requests - Optimizing LLM Performance
Learn the Hugging Face Kernel Hub in 5 Minutes
Featherless AI on Hugging Face Inference Providers 🔥
Introducing Training Cluster as a Service - a new collaboration with NVIDIA
ScreenSuite - The most comprehensive evaluation suite for GUI Agents!
Real-Time AI Sound Generation on Arm: A Personal Tool for Creative Freedom
No GPU left behind: Unlocking Efficiency with Co-located vLLM in TRL
Tiny Agents in Python: a MCP-powered agent in ~70 lines of code
Exploring Quantization Backends in Diffusers