DeepSeek-R1 and exploring DeepSeek-R1-Distill-Llama-8B

from blog Simon Willison's Weblog, | ↗ original
DeepSeek are the Chinese AI lab who dropped the best currently available open weights LLM on Christmas day, DeepSeek v3. That model was trained in part using their unreleased R1 "reasoning" model. Today they've released R1 itself, along with a whole family of new models derived from that base. There's a whole lot of stuff in the new release....