Gradio-Lite: Serverless Gradio Running Entirely in Your Browser
Explorar
Noticias de IA
676 elementos — filtrados, clasificados y sin duplicados
Accelerating over 130,000 Hugging Face models with ONNX Runtime
🧨 Accelerating Stable Diffusion XL Inference with JAX on Cloud TPU v5e
Chat Templates: An End to the Silent Performance Killer
Non-engineers guide: Train a LLaMA 2 chatbot
Optimizing your LLM in production
Overview of natively supported quantization schemes in 🤗 Transformers
Fetch Cuts ML Processing Latency by 50% Using Amazon SageMaker & Hugging Face
Making LLMs lighter with AutoGPTQ and transformers
Optimizing Bark using 🤗 Transformers
Deploying Hugging Face Models with BentoML: DeepFloyd IF in Action
Releasing Swift Transformers: Run On-Device LLMs in Apple Devices
Deploy MusicGen in no time with Inference Endpoints
Huggy Lingo: Using Machine Learning to Improve Language Metadata on the Hugging Face Hub
Introducing Agents.js: Give tools to your LLMs using JavaScript
Happy 1st anniversary 🤗 Diffusers!
Open-Source Text Generation & LLM Ecosystem at Hugging Face
Making ML-powered web games with Transformers.js
Deploy LLMs with Hugging Face Inference Endpoints
Making a web app generator with open ML models
Leveraging Hugging Face for complex generative AI use cases
Deploy Livebook notebooks as apps to Hugging Face Spaces
Faster Stable Diffusion with Core ML on iPhone, iPad, and Mac
DuckDB: analyze 50,000+ datasets stored on the Hugging Face Hub
Welcome fastText to the Hugging Face Hub
AI Speech Recognition in Unity
Introducing the Hugging Face LLM Inference Container for Amazon SageMaker
Introducing BERTopic Integration with the Hugging Face Hub
Optimizing Stable Diffusion for Intel CPUs with NNCF and 🤗 Optimum
Making LLMs even more accessible with bitsandbytes, 4-bit quantization and QLoRA