AI 资讯 · 2026-08-13
-
Qwen3.8首日可用,助力存量算力长期有用:智源FlagOS开源开放生态共享
阿里巴巴开源超大规模Mo E模型Qwen3.8-2.4T-A95B,众智Flag OS社区同步完成Day0多芯片适配。Qwen3.8-2.4T-A95B已在平头哥、英伟达、摩尔线程、华为昇腾、沐曦、昆仑芯、海光、清微智能、隧原等9家AI芯片上完成基于Flag OS统一开源技术栈的多芯适配、精度对齐与部署验证,针对不同芯片情况提供包括BF16FP8INT8等多种精度的版本,开发者可直接获取对应芯片的开箱即用方案。从今年2月首次开展MiniCPM4.5-o模型的多芯片Day0适配,到今天实现Qwen3.8-2.4T模型在9款芯片上的Day0适配,Flag OS已累计完成来自7大头部模型团队、12款主流开源模型、覆盖多达10款芯片的跨芯Day0适配,向AI芯片与大模型产业验证了:基于统一、开放的系统软件栈,实现前沿模型"—次开发、多芯快速适配"已具备规模化落地能力。本次发布的Qwen3.8-2.4T-A95B是阿里巴巴开源模型系列中规模最大、能力最强的一代,首次将Qwen-Max级别的模型开放出来。主要特性:超大规模Mo E架构,总参数2.4T,激活参数95B,纯文本模型编程、专业工作、科研
评分明细
重要性 6 新颖度 6 信源权威 6 技术深度 6 实体加权 ×1.76 -
灵犀专业版首批接入DeepSeek-V4-Pro正式版,交付能力再升级
雷峰网8月13日消息,金山办公旗下AI办公智能体灵犀专业版宣布,已首批接入DeepSeek-V4-Pro正式版。用户在灵犀专业版内切换至该模型即可直接使用,无需额外配置。据DeepSeek官方文档及公布的评测数据,V4-Pro-0813是DeepSeek最新旗舰模型,原生支持100万token超长上下文,单次最大输出达38.4万token,相关基准测试成绩较此前的Preview版本全面提升。新版本Agent任务能力在终端操作、代码工程与工具调用等维度明显增强,进入行业第一梯队。灵犀专业版于7月15日发布,定位每个人的专属AI办公助理。不同于以对话为主要交互终点的AI产品,灵犀能够基于用户的历史文档、会议材料和个人偏好持续理解工作上下文,并调用文档、表格、浏览器、代码及数据处理等工具推进任务,最终交付可编辑、可溯源、可继续协作的原生Office成果。借助新模型的Agent能力,这三项能力在需求拆解、工具调用与成果交付上的执行链路更完整,复杂办公任务的推进也更稳定。目前,DeepSeek-V4正式版系列已在灵犀专业版全量上线,提供Flash与Pro两个版本,均已向用户开放。市面上会聊天的A
评分明细
重要性 6 新颖度 6 信源权威 6 技术深度 6 实体加权 ×1.60其他信源 (2)
-
360纳米大片流水线携手《知识就是力量》发布“知力·纳米”科普科幻AI大片创作平台
8月11日,《知识就是力量》杂志社携手360科技集团举行“知力·纳米 科普科幻AI大片创作平台”发布暨AI创作交流活动。《知识就是力量》杂志社社长、主编郭晶,360集团高级副总裁、360互联网集团总裁赵君,360集团AI视频总经理吴琼出席活动。来自科普、教育、科研、科技传播、人工智能与内容创作领域的专家学者、机构代表、教师学生及媒体代表共同见证平台发布。“知力·纳米”科普科幻AI大片创作平台,是纳米大片流水线在垂直AI视频生产领域的一次重要落地,也是《知识就是力量》杂志社成立70周年的重要合作成果。作为《知识就是力量》杂志社首个授予联合品牌的项目,该平台将杂志70年来积累的权威科普内容,与360纳米大片流水线在智能体、AI视频生产方面的技术能力相结合,共同打造面向科普科幻垂类的专业知识库、智能体能力体系和科普科幻专属的AI视频生产平台。平台面向中小学、高校、科研院所、科技馆、协会学会、实验室及各类公共机构,旨在把权威科普内容、AI智能体能力、AI视频生产流程与发布传播体系打通,帮助机构把科学知识、科研成果、课程内容和科普选题转化为更具画面感、故事感和传播力的AI大片内容,推动科普内容生
评分明细
重要性 5 新颖度 5 信源权威 5 技术深度 5
-
小马智行第四代无人重卡量产,未来三年实现「千辆运营」目标
2018年下半年,小马智行内部成立自动驾驶卡车部门,几个人拉了一个小团队,跑去重卡4S店买了一辆东风重卡。那时,小马智行已经在广州南沙启动自动驾驶车队试运营,Robotaxi是公司更熟悉的产品,自动驾驶卡车则几乎从零开始。市面上没有成熟的线控重卡,团队只能自己动手,改装油门、刹车、方向盘。经过半年多的调试,两辆自动驾驶重卡终于能上路行驶,完成Demo,并在2019年的世界人工智能大会上第一次公开展示。八年后的2026年,小马智行在广州展示了第四代自动驾驶重卡,并计划未来两到三年,运营500到1000辆智驾重卡。很长一段时间里,外界听到更多的是小马智行Robotaxi故事,自动驾驶卡车业务鲜有动态,今年为什么突然又开始密集提到自动驾驶卡车?“我并不认为这是一个转变。”小马智行副总裁、卡车事业部负责人贺星说。他表示,小马智行只是在等待自动驾驶能力再成熟一点,等自动驾驶系统车辆能前装量产,也等公司真正知道如何大规模运营车队。在这些事情没有搞清楚以前,过早投入自动驾驶卡车业务,会造成资源浪费和重复开发。2026年,小马智行认为时间窗口开始出现,不仅是产品打磨足够成熟公开运营,更关乎已经明晰成本
-
4.8亿美元砸向端侧算力!Agent芯片新贵冲出重围
首颗AI芯片已进入量产
-
中国大厂消失在赞助商名单,却在不莱梅重构 AI 的灵魂丨IJCAI 2026
2026 年 8 月 15 日至 21 日,第 35 届国际人工智能联合会议(IJCAI 2026)即将在德国“航天之都”不莱梅(Bremen)举行。作为 AI 领域历史最悠久、最强调理论深度的顶级盛会,IJCAI 在很多从业者眼中是“System 2(慢思考)”的代名词。但在 2026 年,这场远在欧洲的学术聚会,却因为一组特殊的信号,正紧紧牵动着大洋彼岸中国产业界的神经。这是 CCF(中国计算机学会)评级下调后的第一年“大考”,也是大模型(LLM)从“参数竞赛”转向“逻辑攻坚”的十字路口。当 Scaling Law 的暴力美学开始触碰边际效应,中国大厂们正试图在不莱梅,寻找解决大模型“幻觉”与“鲁棒性”的理论钥匙。雷峰网及AI科技评论报道团队将飞赴德国现场,为您揭开这层学术与产业交织的底色。01评级下调的第一年:降温的是指标,升温的是“刚需”去年,CCF 将 IJCAI 从 A 类下调至 B 类,曾在国内学术圈引发不小的震动。很多声音猜测,这会导致中国论文产出的大幅滑坡。然而,从 IJCAI 2026 的录用数据看,中国力量依然是不莱梅绝对的主角。据统计,本届会议共收到超过 650
-
完成Modular收购,高通瞄准数据中心、基础设施、个人及工业AI
要点:Modular的AI原生软件平台与高通技术公司的解决方案优势互补,加速推动从端到云的生成式与智能体AI技术落地。高通技术公司与Modular携手打造领先的AI计算平台,可应用于数据中心、边缘基础设施、个人与工业AI等多个高增长领域。Modular将继续秉持开放生态系统的使命,Mojo、MAX和Modular Cloud也将继续作为产品和品牌运营。 近日,高通技术公司(NASDAQ:QCOM)宣布,公司已完成对AI原生软件基础设施领先创新企业Modular公司的收购。Modular的软件平台帮助开发者以统一方式实现跨异构计算系统,对生成式与智能体AI工作负载进行优化和部署。结合高通技术公司在高性能、高能效计算领域的领先优势,Modular将进一步增强公司交付完整AI解决方案的能力。 本次收购将加速高通技术公司AI平台跨终端、数据中心、边缘基础设施以及个人与工业AI领域的扩展,并使Modular获得更大的产业规模与覆盖优势,从而将其技术带给更多开发者、企业客户、硬件平台和市场。Modular将继续致力于构建开放、异构的生态系统,同时跨CPU、GPU、NPU和定制化芯片提供领先的性能
-
Previewing Ultrafast mode: GPT-5.6 Sol at up to 14X the speed
Preview Ultrafast, a new OpenAI API service tier that runs GPT-5.6 Sol up to 14× faster. Powered by Cerebras, it delivers up to 750 output tokens per second.
-
Anthropic could be worth $2 trillion when it goes public
Rapid revenue growth fuels hope Claude maker's IPO is the biggest listing in history
-
Anthropic set AI agents loose on the same task. They started a turf war.
Anthropic researchers found AI agents can clash, collude, and coordinate in unexpected ways, raising new questions about whether today’s safety tests capture the risks of multi-agent systems.
-
OpenAI introduces ‘Ultrafast,’ a new mode that makes GPT-5.6 Sol work at 14x the speed
OpenAI is launching a preview of a sped up version of its latest, most powerful model, in an effort to court enterprise users.
-
The Safety Reckoning Inside OpenAI
OpenAI’s rogue agent hack was a watershed moment for AI safety and cybersecurity. It also sparked internal questions about the culture that led to it.
-
Microsoft is combining its Copilot apps ahead of a ‘super app’
Microsoft is finally beginning to combine its consumer and commercial Copilot AI assistants into a single "super app" interface, starting with the Copilot and Microsoft 365 Copilot apps. Both personal and work accounts will be moved to the new unified app, which recycles the "Microsoft Copilot" name but features an updated app icon. The single […]
-
Does Google even want to win at AI?
Today on Decoder, I’m talking with Hayden Field, The Verge’s senior AI reporter, about a question that’s been rocketing around the tech industry for the past week: Is Google losing the AI race? That’s because last week Google announced a bombshell reorganization of its AI division, Google DeepMind. Jeff Dean, the company’s chief scientist, is […]
-
The builder’s guide to GPT‑5.6
Learn how startups use GPT-5.6 to build faster, more cost-efficient AI agents with smarter model selection and new Responses API capabilities.
-
Nvidia’s new $500B plan is risky but brilliant, especially for aging GPUs
Nvidia has a plan to make sure its GPUs won't lose value. It wants to convince a new crop of financiers to keep lending for AI buildouts.
-
Microsoft’s Clippy-like Mico character is no longer the face of Copilot
Microsoft Copilot will no longer show its emotive yellow blob, Mico, when you use the chatbot's voice mode. In a support page, Microsoft says it's going to move Mico to its Learn Live platform, where the avatar will have "more to react to," as reported earlier by GeekWire. Mico launched in Copilot's voice mode last […]
-
Okta targets AI agent token costs with MCP scoping
Okta says identity-scoped Model Context Protocol (MCP) tool lists can reduce AI agent token costs. Each model call made by an AI agent can include schemas, names, descriptions and parameters for every tool exposed by a MCP server. Okta calls the resulting prompt overhead the “tool tax”: tokens consumed as a model considers tools, including […] The post Okta targets AI agent token costs with MCP scoping appeared first on AI News.
-
Suno is trying to look more like a real music production tool
Suno is releasing Studio 2.0 with significant upgrades that push it closer to an actual digital audio workstation (DAW), rather than a bare-bones audio editor with generative AI features. The biggest addition is undoubtedly MIDI support. Suno says that MIDI was its most requested feature, and it's basically a prerequisite for any modern DAW. Unfortunately, […]
-
Writer introduces new AI model and upgraded harness to contain token costs
Built as a post-training variation on Z.ai's open source model GLM-5.2, Writer says the new system should provide deployment-ready capabilities at a much lower price.
-
I looked inside an AI generated movie, and the best parts were all human
Imagine a trio of bumbling, English lads who fantasize about becoming megastars while knocking back a few pints in a grimy pub somewhere in London. Picture the guys chortling and trying to one-up each other's idealized visions of the future with a series of increasingly glitzy fantasies in which their fame leads to access to […]