Readpodcast AI

Qwen3.8-27B & How to Serve it Fast

Sam Witteveen
준비됨2026. 8. 18.18:00228,370 조회수YouTube에서 보기

In this video, I look at the long awaited Qwen3.8-27B model. Both what it can do and how to serve it at the maximum tokens per second Thanks to Dell for Sponsoring the Compute #DellProPrecision #DellProMax #DellTech #NVIDIA 📖 Website: https://qwen.ai/ 🤗 HF: https://huggingface.co/collections/Qwen/qwen38 SGLang: https://lmsysorg.mintlify.app/cookbook/autoregressive/Qwen/Qwen3.8-27B Twitter: https://x.com/Sam_Witteveen 🕵️ Interested in building LLM Agents? Fill out the form below Building LLM Agents Form: https://drp.li/dIMes 👨‍💻Github: https://github.com/samwit/llm-tutorials ⏱️Time Stamps: 00:00 Intro 00:50 ThinkingCap 01:25 Qwen3.8 - 27B 01:59 Different Versions on Hugging Face 02:12 Benchmarks 03:14 Artificial Analysis Benchmark 04:10 Qwen3.8-27B on Hugging Face 06:40 Demo 14:17 SGLang

동영상 읽기 · 대본 및 인사이트

이 에피소드에는 전체 대본과 AI 인사이트가 있습니다

무료 계정 · 카드 불필요 · 가입 시 크레딧 150개, 이 에피소드를 잠금 해제하기에 충분합니다

  • 📄 타임스탬프가 포함된 전체 대본
  • ✨ AI 요약, 키워드와 마인드맵
  • 💡 주요 결론과 인용문

대본과 인사이트는 AI가 생성하므로 오류가 있을 수 있습니다. 정확도는 오디오 품질과 화자의 명료도에 따라 달라지며, 이상한 부분이 있다면 원본 오디오가 항상 기준입니다.

도움이 필요하거나 의견을 공유하고 싶으신가요?support@readpodcast.ai

바로 읽을 수 있는 에피소드와 동영상