The DSPy Roadmap
Explorar
Noticias de IA
773 elementos — filtrados, clasificados y sin duplicados
Deploy Meta Llama 3.1 405B on Google Cloud Vertex AI
Introduction to ggml
GPT4o August + 100% Structured Outputs for All (GPT4o mini edition)
Serverless Inference with Hugging Face and NVIDIA NIM
WWDC 24: Running Mistral 7B with Core ML
TGI Multi-LoRA: Deploy Once, Serve 30 Models
How we leveraged distilabel to create an Argilla 2.0 Chatbot
Experimenting with Automatic PII Detection on the Hub using Presidio
Announcing New Hugging Face and KerasHub integration
Google Cloud TPUs made available to Hugging Face users
From DeepSpeed to FSDP and Back Again with Hugging Face Accelerate
Diffusers welcomes Stable Diffusion 3
Introducing the Hugging Face Embedding Container for Amazon SageMaker
Faster assisted generation support for Intel Gaudi
Benchmarking Text Generation Inference
Training and Finetuning Embedding Models with Sentence Transformers
Ten Commandments for Deploying Fine-Tuned Models
Deploy models on AWS Inferentia2 from Hugging Face
Hugging Face on AMD Instinct MI300 GPU
Introducing Spaces Dev Mode for a seamless developer experience
Hugging Face x LangChain : A new partner package
License to Call: Introducing Transformers Agents 2.0
Building Cost-Efficient Enterprise RAG applications with Intel Gaudi 2 and Intel Xeon
Bringing the Artificial Analysis LLM Performance Leaderboard to Hugging Face
Powerful ASR + diarization + speculative decoding with Hugging Face Inference Endpoints
Improving Prompt Consistency with Structured Generations
AI Apps in a Flash with Gradio's Reload Mode
Running Privacy-Preserving Inferences on Hugging Face Endpoints
Making thousands of open LLMs bloom in the Vertex AI Model Garden