CPU Optimized Embeddings with 🤗 Optimum Intel and fastRAG
Explorar
Noticias de IA
676 elementos — filtrados, clasificados y sin duplicados
FSDP+QLoRA: the Answer to 70b-scale AI for desktop class GPUs
Data is better together: Enabling communities to collectively build better datasets toget…
Text-Generation Pipeline on Intel® Gaudi® 2 AI Accelerator
AI Watermarking 101: Tools and Techniques
Fine-Tuning Gemma Models in Hugging Face
🤗 PEFT welcomes new merging methods
From OpenAI to Open LLMs with Messages API on Hugging Face
Less Lazy AI
Hugging Face Text Generation Inference available for AWS Inferentia2
Patch Time Series Transformer in Hugging Face
Introducing the Enterprise Scenarios Leaderboard: a Leaderboard for Real World Use Cases
Accelerate StarCoder with 🤗 Optimum Intel on Xeon: Q8/Q4 and Speculative Decoding
Open-source LLMs as LangChain Agents
Fine-Tune W2V2-Bert for low-resource ASR with 🤗 Transformers
1/16/2024: ArtificialAnalysis - a new model/host benchmark site
Accelerating SD Turbo and SDXL Turbo Inference with ONNX Runtime and Olive
Run ComfyUI workflows for free with Gradio on Hugging Face Spaces
A guide to setting up your own Hugging Face leaderboard: an end-to-end example with Vecta…
Make LLM Fine-tuning 2x faster with Unsloth and 🤗 TRL
LoRA training scripts of the world, unite!
12/26/2023: not much happened today
12/19/2023: Everybody Loves OpenRouter
12/12/2023: Towards LangChain 0.1
Optimum-NVIDIA Unlocking blazingly fast LLM inference in just 1 line of code
AMD + 🤗: Large Language Models Out-of-the-Box Acceleration with AMD GPU
Make your llama generation time fly with AWS Inferentia2
Introducing Prodigy-HF: a direct integration with Hugging Face
Interactively explore your Huggingface dataset with one line of code
Deploy Embedding Models with Hugging Face Inference Endpoints