
IT之家 8 月 15 日消息,据彭博社今日报道,阿里旗下开源 AI 模型千问家族在过去六个月内全球累计下载量突破 30 亿次,超越 Meta、Alphabet 等一众竞争对手,成为全球下载量第一的 AI 模型。 根据开源 AI 社区 Hugging Face 于当地时间 8 月 14 日发布的《开放模型现状》报告,2026 年谷歌模型下载量约 4.18 亿…
AI 点评 · 开源生态反超闭源巨头,中国AI出海新标杆。
共 57 条相关资讯 · 来自历史归档

IT之家 8 月 15 日消息,据彭博社今日报道,阿里旗下开源 AI 模型千问家族在过去六个月内全球累计下载量突破 30 亿次,超越 Meta、Alphabet 等一众竞争对手,成为全球下载量第一的 AI 模型。 根据开源 AI 社区 Hugging Face 于当地时间 8 月 14 日发布的《开放模型现状》报告,2026 年谷歌模型下载量约 4.18 亿…
AI 点评 · 开源生态反超闭源巨头,中国AI出海新标杆。
Tool: CORS Chat I built this today ( with GPT-5.6-Sol xhigh ) to help test Qwen 3.8 27B running in LM Studio on both my M5 MacBook Pro and an NVIDIA DGX Spark. It provides a web UI…
Implement an end-to-end fine-tuning pipeline for tool-calling language models. This tutorial covers parsing trajectories, structured tool-call extraction, Qwen-compatible ChatML re…

Alibaba's AI team Qwen has released new open model weights under the Apache 2.0 license with Qwen 3.8. The dense 27-billion-parameter model is designed to outperform the larger Qwe…
AI 点评 · 开源权重加Apache2.0许可,27B小模型逆袭大参数,性价比看点十足。

IT之家 8 月 13 日消息,阿里云“魔搭 ModelScope 社区”公众号昨天(12 日)深夜宣布,阿里 Qwen 团队正式开放 Qwen3.8-2.4T-A95B 模型权重。 此次是 Qwen-Max 级别的模型首次开源权重:模型总参数为 2.4T,每个 Token 激活 95B 参数, 原生支持 262,144 Token 上下文,并可扩展至 1,…
AI 点评 · 首次开源旗舰级模型,超长上下文与高性价比推理或重塑行业格局。
从面向个人的AI创作工具,扩展为可供组织使用的生产力平台

Qwen is so back!

Alibaba is marketing its new AI model Qwen 3.8 with a video that shows the AI working while a person enjoys their hobbies. It's a deliberate contrast to the job loss warnings from…
AI 点评 · 阿里云用反焦虑营销展示AI解放人力,直击职场痛点,商业化叙事值得玩味。
**Alibaba** launched **Qwen3.8-Max**, a **2.4T-parameter** open-weight model emphasizing autonomous coding, long-horizon execution, and multimodal feedback, with aggressive pricing…
GUI agents have the potential to become a general purpose executor over existing digital devices. To advance them toward real-world use, we envision agents that operate reliably on real devices, execu…
AI 点评 · 图像生成能力飞跃,长文本理解更精准,多模态创作门槛再降低。
A podcast with Florian Brand.
细粒度标签+ 20 种方言

IT之家 7 月 17 日消息,今日,阿里通义实验室发布 Wan-Streamer v0.2 模型,让 AI 像真人一样边听、边看、边回应。 IT之家从官方介绍获悉,它是面向实时双工交互的端到端全模态理解与生成模型,把“听、看、说、演”统一进单个 Transformer 中。 极致低延迟: 端到端响应延迟 550ms (200ms 模型延迟 + 350ms…
AI 点评 · 端到端延迟仅550ms,让AI视频通话逼近真人实时交互体验,技术突破极具商用价值。
阿里发布实时语音模型 Qwen-Audio-3.0-Realtime、腾龙发布 12-20mm F2.8 镜头等。 查看全文
AI 点评 · 苹果AI入华合规,阿里实时语音模型发布,硬件与AI双线突破值得关注。
The deal, which was rumored to be in the works last year, marks an important step for Apple's AI ambitions in a key market.

IT之家 7 月 15 日消息,@PrismML 官方账号今天(7 月 15 日)发布博文,宣布推出 Bonsai 27B 模型,基于 Qwen 3.6 27B 模型微调, 在保留 90% 智能水平的情况下,可以在 12GB 内存的 iPhone 上原生运行。 Qwen 3.6 27B 模型进一步提升本地 AI 的能力,重点涵盖多步推理、结构化工具使用、长上…
AI 点评 · 端侧大模型突破内存瓶颈,iPhone变身AI终端,本地智能体验迎来质变。
In this report, we introduce Qwen-Music, a powerful music generation model capable of producing highly musical and high-fidelity songs with complete vocal singing. Qwen-Music supports two core tasks:…
AI 点评 · Qwen模型线性注意力优化,或成高效大模型推理新突破口。

IT之家 7 月 10 日消息,科技媒体 The Information 昨日(7 月 9 日)发布博文,报道称苹果公司正接洽 PrismML 初创公司, 评估在 iPhone 上直接运行更大规模 AI 模型的可行性。 IT之家查询公开资料,PrismML 是加州理工衍生出来的 AI 初创公司,核心突破为原生 1-bit 模型压缩技术,模型体积可压缩至全精度…
AI 点评 · 苹果布局端侧大模型,小公司技术或成关键突破。
While text-to-image (T2I) models have achieved remarkable progress, they struggle with real-world requests that are often underspecified, implicit, or dependent on up-to-date knowledge. We identify th…
We present Qwen-Image-2.0-RL, a post-training pipeline that applies reinforcement learning from human feedback (RLHF) and on-policy distillation (OPD) to improve both the visual quality and instructio…
Speech conveys information through both words and vocal delivery. We evaluate four leading production realtime voice systems-OpenAI's GPT Realtime 2, Google's Gemini 3.1 Flash Live, and Alibaba's Qwen…
A world model predicts environment dynamics based on current observations and actions, serving as a core cognitive mechanism for reasoning and planning. In this work, we investigate how world modeling…
Agentic navigation systems require a base navigation model whose observation strategy can be externally reconfigured at inference time, because instruction following, object search, target tracking, a…
Foundation models in language and multimodality achieve strong generalization by aligning heterogeneous data under a unified formulation and training at scale. In this report, we investigate whether t…
边走、边看、边思考
We introduce Qwen-RobotWorld, a language-conditioned video world model for embodied intelligence. With natural language as a unified action interface, it predicts physically grounded future visual tra…
甩开视觉内卷
Few-step distillation has become an effective strategy for accelerating advanced visual generative models, yet prior work has largely focused on distillation objectives. In this work, we revisit few-s…
下一代CUA训练范式
AI 点评 · 解决Agent工具选择难题,复旦与通义提出全新训练思路,推动智能体实用化。
Tech Report GitHub Hugging Face ModelScope DISCORD Introduction We are excited to introduce Qwen3Guard, the first safety guardrail model in the Qwen family. Built upon the powerful…
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD We are excited to introduce Qwen-Image-Edit, the image editing version of Qwen-Image. Built upon our 20B Qwen-Image model, Qwen-Ima…
GITHUB HUGGING FACE MODELSCOPE DEMO DISCORD We are thrilled to release Qwen-Image, a 20B MMDiT image foundation model that achieves significant advances in complex text rendering a…
DEMO API DISCORD Introduction Here we introduce the latest update of Qwen-MT (qwen-mt-turbo) via Qwen API. This update builds upon the powerful Qwen3, leveraging trillions multilin…
API DISCORD Introduction Here we introduce the latest update of Qwen-TTS (qwen-tts-latest or qwen-tts-2025-05-22) through Qwen API . Trained on a large-scale dataset encompassing o…
QWEN CHAT DISCORD Introduction The evolution of multimodal large models is continually pushing the boundaries of what we believe technology can achieve. From the initial QwenVL to…
GITHUB HUGGING FACE MODELSCOPE DISCORD We release Qwen3 Embedding series, a new proprietary model of the Qwen model family. These models are specifically designed for text embeddin…
QWEN CHAT GitHub Hugging Face ModelScope Kaggle DEMO DISCORD Introduction Today, we are excited to announce the release of Qwen3, the latest addition to the Qwen family of large la…
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD Introduction Last December, we launched QVQ-72B-Preview as an exploratory model, but it had many issues. Today, we are officially r…
QWEN CHAT HUGGING FACE MODELSCOPE DASHSCOPE GITHUB PAPER DEMO DISCORD We release Qwen2.5-Omni, the new flagship end-to-end multimodal model in the Qwen series. Designed for compreh…
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD Introduction At the end of January this year, we launched the Qwen2.5-VL series of models, which received widespread attention and…
QWEN CHAT Hugging Face ModelScope DEMO DISCORD Scaling Reinforcement Learning (RL) has the potential to enhance model performance beyond conventional pretraining and post-training…
QWEN CHAT DISCORD This is a blog created by QwQ-Max-Preview. We hope you enjoy it! Introduction Okay, the user wants me to create a title and introduction for their blog an…
QWEN CHAT API DEMO DISCORD It is widely recognized that continuously scaling both data size and model size can lead to significant improvements in model intelligence. However, the…
Tech Report HuggingFace ModelScope Qwen Chat HuggingFace Demo ModelScope Demo DISCORD Introduction Two months after upgrading Qwen2.5-Turbo to support context length up to one mill…
QWEN CHAT GITHUB HUGGING FACE MODELSCOPE DISCORD We release Qwen2.5-VL, the new flagship vision-language model of Qwen and also a significant leap from the previous Qwen2-VL. To tr…