Knowledge Distillation - Students meta-llama/Llama-3.1-8B-Instruct Text Generation • 8B • Updated Sep 25, 2024 • 9.31M • • 5.74k Qwen/Qwen2.5-7B-Instruct Text Generation • 8B • Updated Jan 12, 2025 • 12.3M • • 1.22k
Knowledge Distillation - Teachers meta-llama/Llama-3.1-70B-Instruct Text Generation • 71B • Updated Dec 15, 2024 • 737k • • 907 Qwen/Qwen2.5-72B-Instruct Text Generation • 73B • Updated Jan 12, 2025 • 458k • • 929
Knowledge Distillation - Students meta-llama/Llama-3.1-8B-Instruct Text Generation • 8B • Updated Sep 25, 2024 • 9.31M • • 5.74k Qwen/Qwen2.5-7B-Instruct Text Generation • 8B • Updated Jan 12, 2025 • 12.3M • • 1.22k
Knowledge Distillation - Teachers meta-llama/Llama-3.1-70B-Instruct Text Generation • 71B • Updated Dec 15, 2024 • 737k • • 907 Qwen/Qwen2.5-72B-Instruct Text Generation • 73B • Updated Jan 12, 2025 • 458k • • 929