红帽家 记录技术与生活
📡
AI 行业看板
数据每 60 分钟自动同步 · 健康 8/16
最后同步 2026-07-20 11:16
下次更新 2026-07-20 12:16 ·
🏆 大模型综合矩阵
OpenRouter 最近发布 Top 12 🔗 OpenRouter · 2026-07-20 11:15
#模型 / 厂商 综合分趋势 TTFT吞吐 $/M (in/out)性价比
1 Thinking Machines: Inkling · Thinkingmachines 1427 ▲ 12 $1.00 / $4.05 100
2 Kimi K3 · 月之暗面 1422 ▲ 6 $3.00 / $15 67
3 Muse Spark 1.1 · Meta 1417 ▲ 4 $1.25 / $4.25 90
4 KAT-Coder-Air V2.5 · Kwaipilot 1412 — 0 $0.15 / $0.60 100
5 KAT-Coder-Pro V2.5 · Kwaipilot 1407 — 0 $0.74 / $2.96 100
6 GPT-5.6 Luna Pro · OpenAI 1402 — 0 $1.00 / $6.00 100
7 GPT-5.6 Luna · OpenAI 1397 — 0 $1.00 / $6.00 100
8 GPT-5.6 Terra Pro · OpenAI 1392 — 0 $2.50 / $15 70
9 GPT-5.6 Terra · OpenAI 1387 — 0 $2.50 / $15 70
10 GPT-5.6 Sol Pro · OpenAI 1382 — 0 $5.00 / $30 60
11 GPT-5.6 Sol · OpenAI 1377 — 0 $5.00 / $30 60
12 Grok 4.5 · xAI 1372 — 0 $2.00 / $6.00 75
📰 中文 AI 新闻
36氪 2026-07-20 11:15
资讯 在WAIC地下一层找机会的年轻人:光鲜是过去,眼下是生存
6个年轻创业者在WAIC的故事 文|温丽虹 王欣逸 编辑|张雨忻 在上海世博的WAIC现场,想找初创公司的主场H4得颇费一番功夫。 经安检进入主会场,远远只看到会场墙壁上H1到H3的标识。跟随汹涌人群挤进一楼主展馆,几名忙着操纵具身智能机器人的知名厂商工程师翻了翻他们群聊里的展会地图,抱歉地告诉我,他们也不清楚:“这里只有H1到H3,没有H4。你是不是去分会场找找?” 属于初创团队的场地在主展馆地下一层偏居一隅,一道亮黄色的大门默不作声地支在场馆西北侧的角落。穿过这道大门,就是遍布初创公司的一方天地。 来之前,我们知道并非所有的AI创业者与科创公司都对参
资讯 谁还在卷参数?WAIC2026全是能干活的实体AI!
7月17日-20日,一起在WAIC2026现场,看见人工智能真正进入产业深处。 过去一年,围绕AI行业的讨论正在变得更具体。大模型能力仍在持续迭代,但外界关注的重点,已经不再只停留在模型参数、模型发布和单点能力展示上。随着智能体、具身智能、空间智能、AI基础设施等方向不断演进,行业开始更频繁地追问:AI如何进入真实流程,如何完成复杂任务,又如何在产业场景中形成可持续价值。 这种变化背后,是AI竞争维度的继续扩展:模型能力仍是基础,但企业真正要面对的,已经包括数据质量、工程系统、场景理解、交付能力和商业闭环等更复杂的问题。智能体需要从工具调用走向任务协同,
资讯 从“连得上”到“算得懂”,天基通算融合初创公司押注无人系统,首轮融资数千万元|36氪首发
文 | 阿至 封面来源|Pexels 中国低轨卫星星座的规模化部署已经进入关键时间窗口。 一组直观的数据进展是:中国星网已完成第一代组网,垣信卫星在轨运行卫星数量突破200颗,长征十号乙型火箭成功完成全球首次海上网系回收,头部商业公司的可回收火箭也将迎来关键首飞节点。 当“上天”这件事的成本有望迅速下降,产业链的注意力也在从“怎么把卫星打上去”转向一个更务实的问题: 卫星组网之后,要给谁用、怎么用? 这个问题目前还没有标准答案,但市场上不乏想要给出解法的玩家。 一家成立不到半年、聚焦“无人系统通算融合”的初创企业 「星联天枢」,近日正式宣布完成数千万元天
资讯 腾讯云ADP 4.0海外版发布,要把企业级智能体带到全球市场 | 最前线
腾讯云的企业级智能体平台,正式出海了。 7月18日,在2026世界人工智能大会上,腾讯云正式发布了智能体开发平台 ADP 4.0海外版,同步升级智能工作台、Claw 模式、Skill 广场三大核心模块,围绕触达、交互、生态、连接四大能力做了全面国际化适配。 ADP 的全称是 Agent Development Platform,定位为企业级 AgentOps 平台,覆盖智能体的构建、分发和治理全生命周期。 简单来说,它解决的是企业智能体从建立、运行到管理的问题。 智能体开发平台 ADP 4.0海外版 这次企业级智能体海外版的核心升级集中在四个方面:一是渠
资讯 从烤披萨到拿快递,满场跑的机器人终于要进你家了|WAIC 2026全面探展
史上最热的WAIC都整了哪些活?看这一篇就够了。 文|邓咏仪 周鑫雨 王欣逸 温丽虹 编辑|张雨忻 如果你想知道今年的AI圈第一盛事热度如何,只要来上海感受逼近40度的高温,就能同频共振。 没有很热,只有更热。 7月17日,2026年世界人工智能大会在上海世博展览馆正式开幕。展览面积首次突破10万平方米,1100余家企业参展,3000余项展品集中亮相,超300款产品全球首发。从展商数量来看,和2023年400家、2025年800家相比——两年翻了一倍还多。 从网传的消息来看,本届WAIC门票售罄的日子比以往也要更早—— 原价168元的单日票已经被黄牛炒到
资讯 AI软件开发平台Emergent以15亿美元估值完成C轮融资
7月20日,AI软件开发平台Emergent宣布完成1.3亿美元C轮融资。 本轮融资由Creaegis领投,MNI Ventures - Claypond Capital和Sentinel Global联合领投,Khosla Ventures、SoftBank Vision Fund 2、Lightspeed和Y Combinator参与投资。经过最新一轮融资,Emergent估值达到15亿美元,在公开发布后一年内跻身独角兽行列。(新浪财经)
查看全部 60 条 →
📚 国际技术论文
arXiv · cs.AI / cs.CL / cs.LG 2026-07-20 11:15
论文 PagedWeight: Efficient MoE LLM Serving with Dynamic Quality-Aware Weight Quantization
Mixture-of-Experts (MoE) is a popular class of large language models (LLMs), offering high efficiency and accuracy. However, in KV-cache-intensive serving scenarios, MoEs often exhibit a tension between the GPU memory requirements of the model weights and the growing KV cache. We propose PagedWeight, a novel management method for MoE LLM serving that dynamically quantizes MoE model's weights at ru
论文 A Blueprint for Equilibrium-Based Differentiable Continuous-Variable Thermodynamic Computing
To address the escalating energy and latency demands of machine-learning workloads, we introduce a blueprint for an energy-efficient and fast thermodynamic computing stack that leverages stochastic analog processes in physical hardware. In this work, we focus on energy-based thermodynamic computing where the stochastic process is well described by Langevin dynamics with tunable energy potentials.
论文 Cluster-Aware Matching via Laplacian Optimal Transport
In many applications of matching, the point clouds to be matched are not merely unstructured sets of points but rather samples from distributions with an intrinsic cluster structure. In such cases, as individual points are often interchangeable within a coherent region, finding a robust region-to-region alignment is more desirable than establishing a precise point-to-point correspondence. To this
论文 Physics-enhanced reinforcement learning for real-time optimal control of dynamical systems
Reinforcement learning (RL) has recently emerged as a promising feedback control strategy for nonlinear and complex dynamical systems. However, RL algorithms are sample inefficient and require a large number of interaction with the environment to synthesize optimal control strategies. Consequently, applications of RL are typically limited to sparse sensors and actuators due to the curse of dimensi
论文 Evaluating Open-Weight LLMs for Generating Structured Threat Information for Autonomous Vehicle Vulnerabilities
Connected and Autonomous Vehicles (CAVs) rely on interconnected software and hardware components, including sensors, Electronic Control Units, in-vehicle infotainment systems, and telematics units, where vulnerabilities can compromise assets, users, and vehicle operations. These vulnerabilities are commonly documented as plain text in the Common Vulnerabilities and Exposures (CVE) database; howeve
论文 When Does Muon Help Agentic Reinforcement Learning?
Muon is competitive with AdamW in large-scale pre-training, but its value for reinforcement-learning (RL) post-training remains unclear. We study vanilla Muon in sparse-reward agentic RL through matched single-seed comparisons with AdamW on ALFWorld using Qwen2.5-0.5B-Instruct. Under Group-in-Group Policy Optimization (GiGPO), applying Muon only to hidden weight matrices raises final-window valida
查看全部 60 条 →