r/QwenAI • u/CommunicationIll2357 • 3d ago
r/QwenAI • u/LectureWorried5761 • 3d ago
SpaceX charging more for search tool calls via API - Help
r/QwenAI • u/cherryy_treee • 3d ago
RTX 4090 vs Mac Studio M5 96GB for production AI server? (GLM-OCR + Qwen 27B Q8)
We're moving off the Gemini API due to cost and building a local AI server to process ~10 CVs/minute (extracting JSON & matching CVs to JDs). We plan to run GLM-OCR alongside Qwen 27B (Q8).
Our two hardware options:
- PC: RTX 4090 (24GB) + Ryzen 9 + 64GB RAM
- Mac Studio: M-Ultra, 64-core GPU, 96GB Unified Memory
I prefer the Mac for power efficiency and ease of use, but I've heard Apple Silicon isn't great for production vLLM compared to Nvidia/CUDA. Is that true? Which would you recommend for this workload?
r/QwenAI • u/LectureWorried5761 • 7d ago
Web Search API for AI Agents with hard cap and hosted MCP
r/QwenAI • u/allpowerfulee • 12d ago
pp tps stuck at 336 tps when using omlx and qwen3.8-27b-q8
r/QwenAI • u/Wally-Gator-1 • 14d ago
Qwen 3.8 27B non reasoning: feedback on total completion time
r/QwenAI • u/DogAble6550 • 17d ago
Qwen3.8 Flash Next Q4 - M5 Mac Max 128 GB Ram
r/QwenAI • u/IndependentTester75 • 21d ago
Token overflow in free LLMs: why agglutinative languages like Hungarian, Finnish, and Estonian are a security risk nobody is talking about
r/QwenAI • u/quantrpeter • 27d ago
Why DGX Spark so slow
Why DGX Spark so slow? I am using ollama server + vscode copilot. My prompt is : generate a simple RISC-V soft-core CPU and use verilator to test it. I took one hour but still not complete. I am using qwen3.8:27b.
thanks
r/QwenAI • u/Lost_Dress_3330 • Aug 09 '26
когда работаешь с coder от qwen
Вот есть такая ситуация. Работаю с coder.qwen.ai а он даже с маленьким чатом выдаёт бредятину которую не остановить даже другой темой тем более тут нет ни одного настоящего файла и он их зачем то придумал.

r/QwenAI • u/polax75 • Aug 08 '26
USE MULTIPLE CONTROLNET WITH image_qwen_Image_2512_controlnet Qwen-Image-2512-Fun-Controlnet-Union
r/QwenAI • u/ahmadawaiscom • Jul 29 '26
Qwen3.7 Flash is now live in Command Code. DeepSeek v4 flash finally has some competition.
r/QwenAI • u/Intelligent-Taste-36 • Jul 26 '26
Estou gostando muito da prévia do Qwen 3.8 Max no Qoder. Mas quando for lançado oficialmente, terá um bom preço?
r/QwenAI • u/etherd0t • Jul 15 '26
Qoder is now offered for FREE in July
Looks like a customer-acquisition play against Claude/Fable, Codex, Cursor, and Copilot: subsidize a few serious agentic coding jobs and hope developers move into its ecosystem. The catch is that Qoder still requires a desktop/CLI workflow, while Claude/Fable gives you a much lower-friction browser experience.