Consistency Models
Explorar
Noticias de IA
27413 elementos — filtrados, clasificados y sin duplicados
Improved Techniques for Training Consistency Models
Is this... OpenQ*?
BigCodeBench: The Next Generation of HumanEval
Hybrid SSM/Transformers > Pure SSMs/Pure Transformers
Putting RL back in RLHF
Francois Chollet launches $1m ARC Prize
Extracting Concepts from GPT-4
Contextual Position Encoding (CoPE)
Somebody give Andrej some H100s already
Clémentine Fourrier on LLM evals
Anthropic's "LLM Genome Project": learning & clamping 34m features on Claude Sonnet
Unlocking Longer Generation with Key-Value Cache Quantization
Introducing the Open Arabic LLM Leaderboard
LMSys advances Llama 3 eval analysis
Kolmogorov-Arnold Networks: MLP killers or just spicy MLPs?
$100k to predict LMSYS human preferences in a Kaggle contest
Evals: The Next Generation
OpenAI's Instruction Hierarchy for the LLM OS
FineWeb: 15T Tokens, 12 years of CommonCrawl (deduped and filtered, you're welcome)
Introducing the Open Chain of Thought Leaderboard
Jack of All Trades, Master of Some, a Multi-Purpose Transformer Agent
The Open Medical-LLM Leaderboard: Benchmarking Large Language Models in Healthcare
Introducing the LiveCodeBench Leaderboard - Holistic and Contamination-Free Evaluation of…
Vision Language Models Explained
Anime pfp anon eclipses $10k A::B prompting challenge
Mixture of Depths: Dynamically allocating compute in transformer-based language models
ReALM: Reference Resolution As Language Modeling
AdamW -> AaronD?
Evals-based AI Engineering