12/12/2023: Towards LangChain 0.1
Explorar
Noticias de IA
773 elementos — filtrados, clasificados y sin duplicados
Optimum-NVIDIA Unlocking blazingly fast LLM inference in just 1 line of code
AMD + 🤗: Large Language Models Out-of-the-Box Acceleration with AMD GPU
Introducing Prodigy-HF: a direct integration with Hugging Face
Make your llama generation time fly with AWS Inferentia2
Interactively explore your Huggingface dataset with one line of code
Deploy Embedding Models with Hugging Face Inference Endpoints
Gradio-Lite: Serverless Gradio Running Entirely in Your Browser
Accelerating over 130,000 Hugging Face models with ONNX Runtime
🧨 Accelerating Stable Diffusion XL Inference with JAX on Cloud TPU v5e
Chat Templates: An End to the Silent Performance Killer
Non-engineers guide: Train a LLaMA 2 chatbot
Optimizing your LLM in production
Overview of natively supported quantization schemes in 🤗 Transformers
Fetch Cuts ML Processing Latency by 50% Using Amazon SageMaker & Hugging Face
Making LLMs lighter with AutoGPTQ and transformers
Deploying Hugging Face Models with BentoML: DeepFloyd IF in Action
Optimizing Bark using 🤗 Transformers
Releasing Swift Transformers: Run On-Device LLMs in Apple Devices
Deploy MusicGen in no time with Inference Endpoints
Huggy Lingo: Using Machine Learning to Improve Language Metadata on the Hugging Face Hub
Introducing Agents.js: Give tools to your LLMs using JavaScript
Happy 1st anniversary 🤗 Diffusers!
Open-Source Text Generation & LLM Ecosystem at Hugging Face
Making ML-powered web games with Transformers.js
Deploy LLMs with Hugging Face Inference Endpoints
Making a web app generator with open ML models
Leveraging Hugging Face for complex generative AI use cases
Faster Stable Diffusion with Core ML on iPhone, iPad, and Mac
Deploy Livebook notebooks as apps to Hugging Face Spaces