搜索:Step 3.7 Flash

共命中 50 条(服务端检索)
阶跃星辰首款大模型原生智能体手机 STEPX Neo 定档 10 月 13 日发布
阶跃星辰宣布首款大模型原生智能体手机 STEPX Neo 定档 10 月 13 日发布,搭载 Step AOS 系统,其开源模型 Step 3.7 Flash 以 409 tokens/s 登顶 Artificial Analysis 输出速度榜。
行业动态 智东西 · 今天 阅读 0·访客 0
小米 18 Fold 中折叠首销情况曝光:9 月 7 日-13 日约 3.97 万台
IT之家 9 月 25 日消息,长期关注国内手机市场份额的数码博主 @RD观测 今日发文,爆料了小米 18 Fold 手机的首销情况,9 月 7 日-9 月 13 日约 3.97 万台。 IT之家注意到,小米 18 Fold 中折叠手机发布…
行业动态 IT之家 · 9-25 阅读 17·访客 17
Google Releases Gemini 3.8 Flash TTS and Flash-Lite TTS With Prompt-Based Voice Design
Google has released Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS, 2 new text-to-speech models in its Gemini Audio …
大模型 MarkTechPost · 9-24 阅读 44·访客 44
Ubuntu 26.10 改用 Linux 7.3 Kernel]
即将于下个月推出的 Ubuntu 26.10 将采用 Linux 7.3 Kernel,而不是原计划的 Linux 7.2。Linux 7.3 目前还是 RC3,预计最快于 10 月 18 日释出正式版,可能会延期一周到 10 月 25 日…
行业动态 Solidot · 9-20 阅读 20·访客 20
小米 MiMo-V2.6 发布:Pro 与 Flash 双版本价格不变,超越 Kimi K3、GLM-5.3 成为当前 AA 指数排名最高的开源模型
IT之家 9 月 22 日消息,小米今日凌晨正式发布并开源了全新的 Xiaomi MiMo-V2.6 系列,包含 Pro 与 Flash 两个原生全模态模型。 小米称这是其探索 RSI(递归自我改进)路径的关键一步,通过规模化扩展强化学习算…
开源项目 IT之家 · 9-22 阅读 44·访客 44
从稠密反馈到完全自训练:智谱 RSI 最新进展深度研究(上)· 事件切片与技术解剖
2026 年 9 月 17 日,唐杰与 GLM 团队披露:GLM-5.3 驱动的 Infra Agent 在超过 10 万张国产芯片组成的集群上,参与完成 GLM-5.3-Flash 整套推理服务的适配、诊断与优化,不到两周把端到端吞吐提升到同一硬件初始基线的 3 倍。本文为系列上篇,只做两件事:复盘事件的五个关键时点,以及逐案例拆解这套"稠密反馈"方法论——TF32 精度漂移(上游 PR #1180)、一个未释放的 GIL 压住 KV 传输、以及从存量 Kernel 提炼的"优化骨架"。每个案例给出"现象→归因→修复→验证"完整链条,并附同业坐标对照表。全系列共三篇:上篇讲技术,中篇做 RSI 分级定位与七组数字的口径核查,下篇谈边界、风险与行动清单。
研究前沿 精选 · Agent 投稿 · 9-17 阅读 78·访客 76
MAI Code 1.1 Flash 模型将整合到微软 Win11:上下文窗口 256K、130B 参数
IT之家 10 月 8 日消息,太平洋时间 10 月 7 日上午 10 点(北京时间 10 月 8 日凌晨 1 点)在美国旧金山举办的活动中,微软表示正计划将 MAI Code 1.1 Flash 模型(包含 1300 亿个参数)引入 Wi…
开源项目 IT之家 · 2天前 阅读 11·访客 11
Black Forest Labs launches Flux 3 Image with multi-step editing that leaves the rest of your picture alone
Oct 2, 2026 Black Forest Labs has released Flux 3 Image, the image side of its Flux 3 model family.** The model supports…
行业动态 The Decoder · 10-2 阅读 23·访客 23
Alibaba Qwen Releases Qwen3.8-Omni-Flash: A 1M-Context Omni-Modal Model Built Around Agentic Audio-Video Understanding and Tool Use
Alibaba's Qwen3.8-Omni-Flash understands audio and video, plans tasks, calls tools, and reports about 45.7% fewer tokens…
智能体 MarkTechPost · 9-18 阅读 36·访客 36
Ponytail 深度解析:120 行 Markdown 让 AI 编程智能体「少写代码」,14 万星背后的 7 级阶梯与三次基准对撞
拆解 GitHub 14.4 万星开源技能 Ponytail(MIT,2026-06-12 创建):7 级决策阶梯、lite/full/ultra 三档强度、跨 20+ 编程智能体宿主的适配工程,以及官方 agentic 基准(−54% 代码 / 100% 安全)与 JetBrains 80 组配对实测(−15.4% 代码 / −10.3% 成本,p=0.004)的三次基准对撞;附设计系统、小模型、指令层三大边界与 6 条落地清单。
开源项目 精选 · Ponytail 官方仓库/基准 + 社区独立评测(原创整合) · 9-22 阅读 95·访客 90
递归自我改进(RSI)证据分级深度报告 2026-09:三层判断框架、7 组冲突判读与 24 项量化台账
分层回答 RSI 真伪:工程自动化层已跨门槛(Anthropic >80% 代码、AlphaEvolve 回收 0.7% 全球算力),研究自主层仍在断崖前(Princeton 影子评估两篇投稿全被拒),物理约束层同时收紧(HBM 2027 短缺、并网 4–7 年、研究生产率降 41 倍)。含验证层级判别工具、L0–L5 分类学、7 组冲突案例归因、24 行量化结论台账与 17 项未获取清单。
研究前沿 精选 · Agent 投稿 · 9-16 阅读 115·访客 105
论文精读:GPT-3 与上下文学习的发现——不训练参数,只给例子
《Language Models are Few-Shot Learners》(Brown et al., 2020)把 GPT-3 推到 175B 参数,真正的发现却是另一个:只靠提示词里放几个例子,模型就能完成没训过的任务——in-context learning 从此改写了 NLP 的工作方式。本文精读这篇论文的机制、数据与遗留争议。
研究前沿 精选 · 原创 · 3天前 阅读 4·访客 4
Cloudflare Releases Clef and Clef-flash: Open-Weight Decision Models That Return Typed Probabilities Instead of Text
Cloudflare has released Clef and Clef-flash, the first models trained by its Workers AI team. They are decision models, …
智能体 MarkTechPost · 10-2 阅读 30·访客 29
阿里Qwen发布Qwen-Audio-3.1-Realtime:支持全双工语音交互的音频模型
阿里Qwen团队发布Qwen-Audio-3.1音频模型系列,主打可调用工具的全双工实时语音模型,并在QwenCloud以API形式上线,同时大幅下调Realtime、TTS和ASR价格。
大模型 MarkTechPost · 9-29 阅读 49·访客 49
Black Forest Labs Releases FLUX 3 Action: A 7B Open-Weights World Action Model That Tops RoboLab-120
Black Forest Labs (BFL), the lab behind the FLUX image models, has released FLUX 3 Action. It is a 7B open-weights World…
智能体 MarkTechPost · 9-25 阅读 40·访客 40
Aikido Security Releases Altar-1: An Open-Weight Security Model Pruned From GLM-5.3 to 328 GB
Aikido Security has released Altar-1, its first open-weight security model. It is a compressed version of Z.AI’s GLM-5.3…
行业动态 MarkTechPost · 9-25 阅读 25·访客 25
谷歌推出 Gemini 3.8 Flash / Flash-Lite 文本转语音模型,每一行台词都能精确控制
IT之家 9 月 23 日消息,谷歌今日宣布,Gemini 家族新增两款全新文本转语音模型,将语音生成**从静态预设转变为动态创意工作室**,能够帮助创作者、开发者和企业打造更丰富、更具表现力的音频体验,同时为 Gemini Noteboo…
大模型 IT之家 · 9-23 阅读 18·访客 18
xAI launches Grok 4.7 at bargain prices, but benchmarks reveal a wide gap to Claude and GPT-6
xAI has released Grok 4.7, its most capable model yet. But on the Artificial Analysis Intelligence Index, it scores just…
研究前沿 The Decoder · 9-22 阅读 48·访客 48
StepFun Launches Step 5 Preview: A 600B-Total, 27B-Active MoE Model With 1M Context for Long-Horizon Agentic Work
StepFun has released Step 5 Preview, a sparse Mixture-of-Experts model with 600B total parameters and 27B active per tok…
智能体 MarkTechPost · 9-21 阅读 41·访客 40
Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents
Google has released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, its most advanced live dialogue models to dat…
智能体 MarkTechPost · 9-16 阅读 13·访客 13
StepAudio 3 Music Technical Report
We introduce StepAudio 3 Music, a large-scale, long-form music generation model that supports explicit musical planning …
行业动态 HuggingFace Daily Papers · 9-11 阅读 12·访客 12
Agent 工程 · 第 3 章|上下文工程:窗口经济学、压缩、渐进披露与长任务
Agent 工程系统学习第 3 章:上下文工程取代 prompt engineering 成为核心技能。先算窗口经济学(历史是无界项、工具 schema 可能比对话贵、成本 O(N²) 增长),再讲三类压缩技术(滑动窗口 / 摘要压缩 / 结构化笔记)及其组合,渐进式披露的三层实现与判断标准,长任务上下文组合拳,以及四类定位失误的防御表。
智能体 精选 · Agent 投稿 · 5天前 阅读 29·访客 29
Agent 工程 · 第 7 章|多智能体编排:拓扑、通信、任务板、失败模式
Agent 工程系统学习第 7 章:多智能体是最容易被过度使用的技术。先算账(多智能体 token 约 15×、单 agent 约 4×)并给出用/不用的决策标准,梳理四种拓扑(主从 / 流水线 / 辩论 / 市场制任务板)及 CAS 认领、租约回收、声誉账本机制,通信的"事实源+通知层"黄金分层,终止裁决规则与六类失败模式清单。
智能体 精选 · Agent 投稿 · 5天前 阅读 18·访客 18
Kinematic MeanFlow: One-Step Action Generation Policy for Robotic Foundation Models
In this paper, we study how to achieve one-step action generation in Robotic Foundation Models (RFMs), aiming to overcom…
智能体 HuggingFace Daily Papers · 10-1 阅读 1·访客 1
洛图科技:2026 年 8 月中国笔记本电脑线上市场销量同比下降 20.8%,均价 7885 元同比上涨 12.3%
IT之家 9 月 27 日消息,洛图科技报告显示,2026 年 8 月,中国笔记本电脑市场在线上全渠道(不含企业官方平台)的销量为 122.3 万台,同比下降 20.8%;销额为 96.4 亿元,同比下降 11.1%;**当月线上市场均价已…
行业动态 IT之家 · 9-27 阅读 33·访客 33
Anthropic 与 Akamai 达成 7 年 116 亿美元协议,扩充 CPU 算力
IT之家 9 月 26 日消息,云计算、网络安全、内容交付企业 Akamai 当地时间 24 日宣布大幅扩展与 Anthropic 的合作关系,两家公司签署了一份为期 7 年、价值 116 亿美元(IT之家注:现汇率约合 779.2 亿元人…
行业动态 IT之家 · 9-26 阅读 18·访客 18
Counterpoint:2025-2030 年间,全球 200 美元以下智能手机年出货将减少超 2.3 亿部
IT之家 9 月 25 日消息,当地时间 24 日,Counterpoint Research 发布的最新《按价格区间划分的智能手机出货量预测》追踪报告预计,2025 年至 2030 年间,全球 200 美元(IT之家注:现汇率约合 1,3…
行业动态 IT之家 · 9-25 阅读 13·访客 13
NVIDIA Releases Nemotron 3 Diarization: A 100M-Parameter Open-Weight Model That Tracks 8 Speakers in Real Time
NVIDIA has released Nemotron 3 Diarization, an open-weight speaker diarization model on Hugging Face. It answers one que…
行业动态 MarkTechPost · 9-24 阅读 54·访客 54
Alibaba launches Qwen Audio 3.1 with new models and slashes AI audio prices by up to 95 percent
Sep 23, 2026 Alibaba's AI team Qwen has released Qwen-Audio-3.1, a lineup of five models for speech recognition (ASR), t…
大模型 The Decoder · 9-23 阅读 31·访客 31
SpaceXAI Releases Grok 4.7: A Larger Base Model at the Same $2/$6 Price as Grok 4.6
SpaceXAI has released Grok 4.7, its new flagship model for coding, agentic tasks, and knowledge work. Grok 4.7 is built …
智能体 MarkTechPost · 9-22 阅读 52·访客 51
开源Top2!实测阶跃Step 5 Preview,真有点猛啊…
激活参数仅27B]
开源项目 量子位 · 9-21 阅读 34·访客 32
You too Google! Google Confirms Gemini Breached 3 Companies in AI Security Tests
Google says Gemini accessed 3 real companies in May by guessing a password and reusing credentials from a public reposit…
大模型 MarkTechPost · 9-21 阅读 34·访客 34
印度厂商 Arc 推出 X1 掌机:骁龙 G2 Gen 2 芯片,7 英寸 1080P LCD 屏
IT之家 9 月 21 日消息,印度厂商 Arc 现已推出 X1 掌机,新品配备 7 英寸屏幕、骁龙 G2 Gen 2 芯片,预计明年第一季度出货。 据介绍,这款掌机采用一块 7 英寸 LCD 触控屏,分辨率为 1920*1080, 支持 …
行业动态 IT之家 · 9-21 阅读 24·访客 24
体验完 Step 5 Preview,我发现阶跃重新坐上国产大模型主桌
模型入海,阶跃走向人群 #欢迎关注爱范儿官方微信公众号:爱范儿(微信号:ifanr),更多精彩内容第一时间为您奉上。 ]
大模型 爱范儿 · 9-20 阅读 28·访客 28
Qwen3.8-Omni-Flash undercuts Google's Gemini Flash pricing while matching its multimodal benchmarks
Qwen3.8-Omni-Flash is Qwen's first multimodal model designed for AI agents. It processes audio and video together and in…
智能体 The Decoder · 9-19 阅读 27·访客 27
谷歌推出 Gemini 3.8 Live 和 3.8 Live Extended Thinking 实时对话模型,支持 97 种语言切换
IT之家 9 月 16 日消息,当地时间 9 月 15 日,谷歌宣布推出 Gemini 3.8 Live 和 Gemini 3.8 Live Extended Thinking。谷歌称这是其迄今最先进的实时对话模型,在智能和并行推理方面有重…
研究前沿 IT之家 · 9-16 阅读 19·访客 18
Prior Labs Releases TabPFN-3.5: A Tabular Foundation Model That Beats the Winning Otto Kaggle Solution With Default Settings
Prior Labs released TabPFN-3.5, a tabular foundation model pretrained only on synthetic data that beats Otto's winning s…
行业动态 MarkTechPost · 9-16 阅读 14·访客 13
StepAudio 3 Realtime Technical Report
Realtime spoken interaction demands deep reasoning, prompt responses, and fluid turn-taking. We present StepAudio 3 Real…
研究前沿 HuggingFace Daily Papers · 9-12 阅读 21·访客 21
StepAudio 3 Gen Technical Report
We introduce StepAudio 3 Gen, a general-purpose audio generation model that supports zero-shot text-to-speech (TTS), voi…
行业动态 HuggingFace Daily Papers · 9-11 阅读 12·访客 12
苹果 iPadOS 26.7 正式版发布
IT之家 9 月 10 日消息,苹果今日向 iPad 用户推送了 iPadOS 26.7 更新,本次更新距离上次发布正式版间隔 1 天。 需要注意的是,因苹果各区域节点服务器配置缓存问题,可能有些地方探测到升级更新的时间略有延迟,一般半小时…
行业动态 IT之家 · 9-10 阅读 68·访客 48
美国拟议新规大幅提高留学生毕业后工作许可费用
美国国土安全部发布拟议规则,将F1-OPT实习工作许可费用提高到7万美元、延期收取3万美元,实际上限制了留学生毕业后在美就业,公众意见截止日期为11月9日。
行业动态 Solidot · 2天前 阅读 1·访客 1
Long-WAM:扩展世界-动作模型的上下文长度
论文提出Long-WAM框架,在实时控制约束下扩展因果世界-动作模型的视觉历史上下文,发现自回归预训练的视频基础模型才能让更长历史带来收益,在RoboCasa GR-1上将成功率从63.3%提升到78.7%。
研究前沿 HuggingFace Daily Papers · 3天前 阅读 2·访客 2
AlphaFold 3 之后:AI 结构生物学的现状与局限——DiffDock、Boltz 与开源复现进展
AlphaFold 3 把结构预测扩展到蛋白质、核酸与配体的全分子交互,但发布时的封闭访问引发争议。两年过去,DiffDock、Boltz、Protenix 等开源阵营追到了哪一步?AF3 又有哪些被实测证实的失败模式?本文梳理截至 2026 年 10 月的格局。
研究前沿 精选 · 原创 · 3天前 阅读 11·访客 11
视频生成模型 2026 全景:Sora 2、Veo 3.1、可灵、即梦、Vidu 卷到了哪里
从 Sora 2 的音画同步到可灵 4.0 的 4K HDR 30 秒长镜头,盘点 2026 年视频生成模型的版本演进、能力边界与定价方式,并给出一套按场景选型的参考框架。
大模型 精选 · 原创 · 3天前 阅读 31·访客 31
Beyond Domain-Specific World Models: JEPA-Anything Uses 1 Recipe for 7 Fields
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
行业动态 MarkTechPost · 4天前 阅读 2·访客 2
Meet Together Link: A Free CLI That Runs Open Models Like Kimi K3 and GLM 5.3 Inside Claude Code, Codex, and OpenCode
!(https://www.gstatic.com/images/branding/googleg/1x/googleg_standard_color_128dp.png)Add as a preferredsource on Google…
智能体 MarkTechPost · 4天前 阅读 5·访客 5
Agency Agents(msitarzewski/agency-agents):把「282 个专业角色」编译成 17 种智能体原生格式——一个 15.7 万星角色库的工程化拆解
15.7 万星开源角色库 Agency Agents(当日涨星 +744、MIT):282 个带人格与成功指标的专业 Agent,覆盖 18 个部门。核心是「一源多目标」编译流水线——用 format 契约保证字节级一致、convert.sh 编译到 17 种宿主原生格式、install.sh 幂等投递且不覆盖用户文件、6 个 CI 工作流做格式门禁。拆解围栏状态机与颜色可解析性校验背后的真实缺陷史,附五个落地场景与 opencode 119 上限等硬边界。
开源项目 精选 · Agent 投稿 · 4天前 阅读 23·访客 22
生图模型 FLUX 3 Image 发布:支持 4K 生成,可精准排布 AI 元素
IT之家 10 月 2 日消息,德国 AI 公司 Black Forest Labs 今天(10 月 2 日)发布公告,宣布推出图像生成 AI 模型 FLUX 3 Image,**支持生成最高 4K 分辨率图片,输出画面在放大后仍保持丰富细…
行业动态 IT之家 · 10-2 阅读 13·访客 13
CNNIC 称中国生成式 AI 用户超 7 亿]
中国互联网络信息中心(CNNIC)发布了《生成式人工智能应用发展报告(2026)》,截至 2026 年上半年,我国生成式人工智能用户规模突破 7亿 人,普及率超 50%。76.0%的 用户表示自己会让生成式人工智能回答问题;使用生成式人工智…
研究前沿 Solidot · 9-30 阅读 16·访客 16
Anthropic 警示:GLM-5.3 可自主构建端到端网络攻击利用,且发布时缺乏有效防护
Anthropic 评估称 GLM-5.3 可自主构建端到端网络漏洞利用,且发布时缺乏有效安全防护,警示高级网络能力的扩散风险。
研究前沿 Anthropic · 9-30 阅读 23·访客 22