deepseek-ai/DeepSeek-R1-Distill-Qwen-32B Text Generation β’ 33B β’ Updated Feb 24, 2025 β’ 727k β’ β’ 1.59k
Running on Zero Agents 61 MInference π 61 Chat with a fast LLaMAβ3 AI using dynamic sparse attention
gradientai/Llama-3-8B-Instruct-Gradient-1048k Text Generation β’ 8B β’ Updated Oct 29, 2024 β’ 23.8k β’ β’ 681