Qwen3 is a fantastic open-source model
Try Zapier MCP for free today: https://bit.ly/42Babll Join My Newsletter for Regular AI Updates 👇🏼 https://forwardfuture.ai My Links 🔗 👉🏻 Subscribe: https://www.youtube.com/@matthew_berman 👉🏻 Twitter: https://twitter.com/matthewberman 👉🏻 Discord: https://discord.gg/xxysSXBxFW 👉🏻 Patreon: https://patreon.com/MatthewBerman 👉🏻 Instagram: https://www.instagram.com/matthewberman_ai 👉🏻 Threads: https://www.threads.net/@matthewberman_ai 👉🏻 LinkedIn: https://www.linkedin.com/company/forward-future-ai Media/Sponsorship Inquiries ✅ https://bit.ly/44TC45V Disclosure: I am a small investor in LM Studio. Links: https://x.com/Alibaba_Qwen/status/1916962087676612998 https://qwenlm.github.io/blog/qwen3/ https://x.com/ArtificialAnlys/status/1917246369510879280
동영상 읽기 · 대본 및 인사이트
이 에피소드에는 전체 대본과 AI 인사이트가 있습니다
무료 계정 · 카드 불필요 · 가입 시 크레딧 150개, 이 에피소드를 잠금 해제하기에 충분합니다
- 📄 타임스탬프가 포함된 전체 대본
- ✨ AI 요약, 키워드와 마인드맵
- 💡 주요 결론과 인용문
에피소드 타임라인
Qwen3 launches as an open-source model family rivaling Gemini 2.5 Pro on major benchmarks.
- Qwen3's flagship 235B model with 22B active parameters scores 85.7 on ArenaHard versus Gemini 2.5 Pro's 92, and beats Gemini 2.5 Pro on LiveCodeBench 70.7 to 70.4 and CodeForces ELO 2056 to 2001.
- The 30B mixture-of-experts model with only 3B active parameters is highlighted as the best of the batch, scoring 91 on ArenaHard versus GPT-4o's 85 and 62 versus 32 on LiveCodeBench.
- Qwen3 is optimized for agents and coding, with the 235B model scoring 70.8 on BFCL function calling versus 62.9 for Gemini 2.5 Pro, and even the 32B dense model scores 70.3.
Qwen3 introduces adjustable thinking budgets and a hybrid thinking/non-thinking mode.
- The hybrid model lets users control how many tokens the model spends reasoning, with performance increasing smoothly as more thinking tokens are allocated across benchmarks like AIME24, AIME25, LiveCodeBench, and GPQA Diamond.
- Non-thinking mode gives quick near-instant responses for simple tasks, while thinking mode reasons step by step for complex problems, enabling task-specific budget configuration.
- This flexibility is especially useful for vibe coding, where hard tasks like building features warrant deep thinking but routine commands like running tests and deploying do not.
Sponsor segment: Zapier's MCP server connects AI agents to thousands of apps.
- Zapier released an MCP server giving agents access to over 7,000 apps, configured by simply selecting apps and plugging the generated URL into tools like Windsurf, Claude Desktop, or Cursor.
- Zapier has over a decade of work automation experience and offers a free plan, making it easy to connect Qwen3 to real-world tools without writing code.
핵심 개념
- Qwen3— The newly launched open-source model family that is the main subject of the episode.
- open-source model— Qwen3 is fully open-source with open weights, making it accessible to everyone.
- hybrid thinking mode— Qwen3 can switch between step-by-step reasoning and quick responses, with adjustable thinking budget.
주목할 인용문
We now have a completely open-source and open weights model that is comparable to Gemini 2.5 Pro.
🤯— Highlights the surprising fact that an open-source model can now match a leading proprietary frontier model.
Code Forces actually scored a higher ELO rating 2056 as compared to 2001 for Gemini 2.5 Pro.
🤯— Reveals that Qwen3 outperforms Gemini 2.5 Pro on a competitive coding benchmark, a counterintuitive result for an open-source model.
실행 가능한 주요 결론
📊AI Model Evaluation
Open-source models like Qwen3 now rival proprietary frontier models on key benchmarks.
Download Qwen3 30B MoE from LM Studio and run the LiveCodeBench or BFCL benchmark to compare with Gemini 2.5 Pro.
Tool calling during chain of thought is a rare capability that enables complex agentic workflows.
Set up an MCP server with Zapier and test Qwen3's tool calling by asking it to fetch data and create a chart.
🛠️Practical AI Usage
Adjustable thinking budget allows balancing speed and quality for different tasks.
In your next coding session, use Qwen3's non-thinking mode for simple commands and thinking mode for complex features.
Qwen3 is optimized for agentic use cases and can perform computer use tasks like organizing files.
Try Qwen3's computer use demo by asking it to organize a folder on your desktop by file type.
대본과 인사이트는 AI가 생성하므로 오류가 있을 수 있습니다. 정확도는 오디오 품질과 화자의 명료도에 따라 달라지며, 이상한 부분이 있다면 원본 오디오가 항상 기준입니다.
바로 읽을 수 있는 에피소드와 동영상
팟캐스트 에피소드

He Couldn't Walk Away || How Jetha Devapura Built Sri Lanka's Biggest Crisis Line
The Giving Habit
2026년 7월 15일55:53EN
The shape-shifting sounds of the accordion | Maria Telesheva
TED Talks Daily
2026년 8월 27일13:22EN
Bad Maps and Good Intentions; Sophie Radice on the trials and tribulations of life beyond the comfort zone S5 E11
How to have Extraordinary Relationships
2026년 5월 26일57:25EN
商业小样43 | AI时代,谁在给服务器“降温”
商业就是这样
2026년 6월 21일12:16ZH
#71 Die Deutschen im Amerikanischen Unabhängigkeitskrieg
Wer wir sind und warum das nicht klappte ...
2026년 8월 19일44:55DE
SÉRIE: UMA VEZ SALVO, SALVO PARA SEMPRE - A CERTEZA DA SALVAÇÃO: PARTE 2| PR.PEDRO ESTRELLA
Minha Igreja Na Cidade
2026년 8월 31일59:49PT
동영상

Neuroscientist: You Will NEVER Feel Stressed Again | Andrew Huberman
RESPIRE
2023년 2월 27일11:05EN
No Priors Ep. 80 | With Andrej Karpathy from OpenAI and Tesla
No Priors: AI, Machine Learning, Tech, & Startups
2024년 9월 5일44:15EN
Steve Jobs' 2005 Stanford Commencement Address (with intro by President John Hennessy)
Stanford
2008년 5월 14일22:10EN
INDIA’S GOT LATENT S2 EP4 ft. Karan Aujla, Tanmay Bhat, Gurleen Pannu, Rahul Dua
Samay Raina
2026년 8월 2일53:54HI
#32 Die Schlacht von Worringen - Der Freiheitskampf der Kölner
99 mal Geschichte
2025년 11월 12일53:29DE