1/13-14/2024: Don't sleep on #prompt-engineering
Explorar
Noticias de IA
21813 elementos — filtrados, clasificados y sin duplicados
1/12/2024: Anthropic coins Sleeper Agents
1/11/2024: Mixing Experts vs Merging Models
1/8/2024: The Four Wars of the AI Stack
1/6-7/2024: LlaMA Pro - an alternative to PEFT/RAG??
12/30/2023: Mega List of all LLMs
12/23/2023: NeurIPS Best Papers of 2023
Speculative Decoding for 2x Faster Whisper Inference
Practices for Governing Agentic AI Systems
Weak-to-strong generalization
Mixture of Experts Explained
SetFitABSA: Few-Shot Aspect Based Sentiment Analysis using SetFit
Goodbye cold boot - how we made LoRA Inference 300% faster
Open LLM Leaderboard: DROP deep dive
SDXL in 4 steps with Latent Consistency LoRAs
Comparing the Performance of LLMs: A Deep Dive into Roberta, Llama 2, and Mistral for Dis…
Exploring simple optimizations for SDXL
The N Implementation Details of RLHF with PPO
DALL·E 3 system card
Finetune Stable Diffusion Models with DDPO via TRL
Object Detection Leaderboard
Introduction to 3D Gaussian Splatting
Fine-tuning Llama 2 70B using PyTorch FSDP
Efficient Controllable Generation for SDXL with T2I-Adapters
Fine-tune Llama 2 with DPO
Towards Encrypted Large Language Models with FHE
Practical 3D Asset Generation: A Step-by-Step Guide
Results of the Open Source AI Game Jam
Fine-tuning Stable Diffusion models on Intel CPUs
Ethics and Society Newsletter #4: Bias in Text-to-Image Models