Qwen3 is a fantastic open-source model
Try Zapier MCP for free today: https://bit.ly/42Babll Join My Newsletter for Regular AI Updates 👇🏼 https://forwardfuture.ai My Links 🔗 👉🏻 Subscribe: https://www.youtube.com/@matthew_berman 👉🏻 Twitter: https://twitter.com/matthewberman 👉🏻 Discord: https://discord.gg/xxysSXBxFW 👉🏻 Patreon: https://patreon.com/MatthewBerman 👉🏻 Instagram: https://www.instagram.com/matthewberman_ai 👉🏻 Threads: https://www.threads.net/@matthewberman_ai 👉🏻 LinkedIn: https://www.linkedin.com/company/forward-future-ai Media/Sponsorship Inquiries ✅ https://bit.ly/44TC45V Disclosure: I am a small investor in LM Studio. Links: https://x.com/Alibaba_Qwen/status/1916962087676612998 https://qwenlm.github.io/blog/qwen3/ https://x.com/ArtificialAnlys/status/1917246369510879280
Read Video · 文字稿与深度分析
本集已有完整文字稿 + AI 深度分析
免费注册 · 无需信用卡 · 注册即获 150 积分,足够解锁本集
- 📄 完整文字稿含时间戳
- ✨ AI 摘要、关键词与思维导图
- 💡 核心要点与精彩引言
节目时间轴
Qwen3 launches as an open-source model family rivaling Gemini 2.5 Pro on major benchmarks.
- Qwen3's flagship 235B model with 22B active parameters scores 85.7 on ArenaHard versus Gemini 2.5 Pro's 92, and beats Gemini 2.5 Pro on LiveCodeBench 70.7 to 70.4 and CodeForces ELO 2056 to 2001.
- The 30B mixture-of-experts model with only 3B active parameters is highlighted as the best of the batch, scoring 91 on ArenaHard versus GPT-4o's 85 and 62 versus 32 on LiveCodeBench.
- Qwen3 is optimized for agents and coding, with the 235B model scoring 70.8 on BFCL function calling versus 62.9 for Gemini 2.5 Pro, and even the 32B dense model scores 70.3.
Qwen3 introduces adjustable thinking budgets and a hybrid thinking/non-thinking mode.
- The hybrid model lets users control how many tokens the model spends reasoning, with performance increasing smoothly as more thinking tokens are allocated across benchmarks like AIME24, AIME25, LiveCodeBench, and GPQA Diamond.
- Non-thinking mode gives quick near-instant responses for simple tasks, while thinking mode reasons step by step for complex problems, enabling task-specific budget configuration.
- This flexibility is especially useful for vibe coding, where hard tasks like building features warrant deep thinking but routine commands like running tests and deploying do not.
Sponsor segment: Zapier's MCP server connects AI agents to thousands of apps.
- Zapier released an MCP server giving agents access to over 7,000 apps, configured by simply selecting apps and plugging the generated URL into tools like Windsurf, Claude Desktop, or Cursor.
- Zapier has over a decade of work automation experience and offers a free plan, making it easy to connect Qwen3 to real-world tools without writing code.
关键概念
- Qwen3— The newly launched open-source model family that is the main subject of the episode.
- open-source model— Qwen3 is fully open-source with open weights, making it accessible to everyone.
- hybrid thinking mode— Qwen3 can switch between step-by-step reasoning and quick responses, with adjustable thinking budget.
精选金句
We now have a completely open-source and open weights model that is comparable to Gemini 2.5 Pro.
🤯— Highlights the surprising fact that an open-source model can now match a leading proprietary frontier model.
Code Forces actually scored a higher ELO rating 2056 as compared to 2001 for Gemini 2.5 Pro.
🤯— Reveals that Qwen3 outperforms Gemini 2.5 Pro on a competitive coding benchmark, a counterintuitive result for an open-source model.
可执行的洞察
📊AI Model Evaluation
Open-source models like Qwen3 now rival proprietary frontier models on key benchmarks.
Download Qwen3 30B MoE from LM Studio and run the LiveCodeBench or BFCL benchmark to compare with Gemini 2.5 Pro.
Tool calling during chain of thought is a rare capability that enables complex agentic workflows.
Set up an MCP server with Zapier and test Qwen3's tool calling by asking it to fetch data and create a chart.
🛠️Practical AI Usage
Adjustable thinking budget allows balancing speed and quality for different tasks.
In your next coding session, use Qwen3's non-thinking mode for simple commands and thinking mode for complex features.
Qwen3 is optimized for agentic use cases and can perform computer use tasks like organizing files.
Try Qwen3's computer use demo by asking it to organize a folder on your desktop by file type.
转录文字与 AI 洞察均由模型自动生成,可能存在少量误差。识别效果与音频质量、语速和发音清晰度相关——如有内容看起来不对,以原始音频为准。
播客与视频,已可阅读
音频播客

3 steps to turn everyday get-togethers into transformative gatherings | Priya Parker
TED Talks Daily
2026年8月15日12:38ENPulling Back the Curtain on Sportsbooks & VIP Programs w/ Dillon Borgida | Ep 63
The Risk Takers Podcast
2024年3月14日1:38:10EN
Stock Market EMERGENCY: Sell Your Stocks Now, The Collapse Is Weeks Away!
The Diary Of A CEO with Steven Bartlett
2026年6月25日1:45:21EN
S9E5 鲁豫对话陈玉亭 | 失去一切之前,我必须成为「白眼狼」
岩中花述
2026年7月8日2:19:41ZH-Hans
[EP30] Banalidade do mal
QOHÉLET, podcast de Ed René Kivitz
2024年1月18日38:31PT
Martin Luther - Wie wird man Reformator?
Wer wir sind und warum das nicht klappte ...
2026年4月15日1:09:23DE
视频

Stock Picking in Difficult Times with George Noble | The Real Eisman Playbook Ep 76
Steve Eisman
2026年9月21日50:53EN
Gemini Robotics – AI for the Physical World, with Keerthana and Ted of Google DeepMind
Cognitive Revolution "How AI Changes Everything"
2025年5月17日1:48:11EN
Jev: ChatGPT Co-Creator’s Answer to RLHF
AI Council
2026年6月19日36:39EN
从「上瘾模型」到「专注力训练」,如何在被算法理解的世界里重新找回主动?| 英文访谈 S9E33
声动活泼
2025年10月16日50:48ZH-Hans
Firma bez šéfov: Funguje to? - Money Talk 113 s Ferom Baníkom
Milan Dubec
2026年8月4日54:10SK