JEDEE AI
存档 2026-07-13

7 月 13 日(北京时间)全球 AI 圈推文存档,按曝光排序,共 99 条。
← 返回最新 全部归档

全部情报 每小时更新 · 事件已合并同类项

提交账号

填 @用户名 或主页链接,审核通过后收录进情报站。

内容 公司
Claude@claudeai · 公司官方 · 1 天前Claude 产品官方账号
连环推 ×2

我们在所有付费计划中延长了 Claude Fable 5 的使用权,同时保持 Claude Code 的周速率限制提高 50%,有效至 7 月 19 日。

查看英文原文
We're extending Claude Fable 5 access on all paid plans, as well as keeping Claude Code’s weekly rate limits 50% higher, through July 19.
As before, you can use up to half of your weekly usage limit on Fable 5. After that, you can continue using Fable 5 with usage credits, or switch to another model to keep working within your remaining limits.

More details here:
support.claude.com/en/articl…
◔ 2635.2 万 次浏览(7 条合计)♥ 7.4 万⇄ 7,280动态看原帖 ↗
Sam Altman@sama · 创始人 · 1 天前Sam Altman,OpenAI 联合创始人兼 CEO

我特别想看看大家用 5.6 Sol 做出了什么有趣的东西。做出最酷东西的人我会送一份 OpenAI 档案库的特别礼物。

查看英文原文
i'd love to see interesting things people have built with 5.6 sol.

i will send the person who made the coolest thing a special gift from the openai archives.
Logan Kilpatrick@OfficialLoganK · 创始人 · 1 天前谷歌 Gemini 产品负责人

我挺惊讶有这么多人似乎不明白:优秀的模型是用超高质量精选数据构建的。找到创新的方式来创建/获取这些数据是个巨大的竞争优势。

查看英文原文
it’s surprising to me how many people seem to not understand that great models are built with super high quality curated data

finding novel ways to create / get this data is a huge edge
Greg Brockman@gdb · 创始人 · 1 天前Greg Brockman,OpenAI 联合创始人兼总裁

用Codex帮你初创公司找客户的方案:

引用 Kappaemme @Kappaemme1926Codex技能可分析初创公司并从公开信号发现潜在客户。输入网址后定义理想客户、搜索讨论、筛选客户,生成含外联建议的报告。包括客户分析、信号研究、客户名单、评分、原始链接和个性化开场白。开源项目,支持一键安装。查看被引原帖 ↗
查看英文原文
Codex for finding customers for your startup:
◔ 55.7 万 次浏览♥ 3,303⇄ 175▶ 含视频教程看原帖 ↗
ChatGPT@ChatGPTapp · 公司官方 · 1 天前ChatGPT 产品官方账号

ChatGPT 重新在 EEA(欧洲经济区)的 WhatsApp 可用了,这是我们在人们日常使用的应用中普及 AI 的一部分工作。

给已认证的 1-800-CHATGPT 账号发消息就能提问、上传图片、发送语音消息、生成图像,还能用多种语言使用 ChatGPT。

现在也上线了韩国的 Kakao 和部分市场的 Viber。

查看英文原文
ChatGPT is available again on WhatsApp in the EEA, part of our work to make AI accessible in the apps people already use every day.

Message the verified 1-800-CHATGPT contact to ask questions, upload images, send voice notes, create images, and use ChatGPT in many languages.

Now also on Kakao in South Korea and Viber in supported markets.
◔ 46.2 万 次浏览(2 条合计)♥ 2,439⇄ 180▶ 含视频新品看原帖 ↗
Kling AI@Kling_ai · 公司官方 · 1 天前快手旗下可灵 AI 视频官方

🎬 自豪地看到葡萄牙世界杯宣传片由Kling AI驱动,播放量即将破亿!

来看看78 Films如何在这次混合制作中使用Kling AI,打造出最震撼人心的场景——凭借其他模型无法比拟的电影级画质,同时保留人类的创意与情感。引用影片AI总监João Seiça的话:"我们使用Kling不是为了取代创意,而是为了拓展它。"

导演:João Seiça
@joaoseica
(混合导演、AI主管)、Nuno Mendes、João Marques
制作公司:78 Films
代理机构:Dentsu Creative Portugal

查看英文原文
🎬 Proud to see Portugal's World Cup campaign film, powered by Kling AI, hitting almost 100M views!

Watch how 78 Films used Kling AI in this hybrid production to create the most ambitious scenes possible — with cinematic quality unmatched by any other model — while still maintaining human creativity and emotions. To quote the film’s AI Director, João Seiça: “We’re not using Kling to replace creativity, but to expand it.”

Directors: João Seiça
@joaoseica
(Hybrid Director, AI Lead), Nuno Mendes, João Marques
Production Company: 78 Films
Agency: Dentsu Creative Portugal
◔ 33.4 万 次浏览♥ 155⇄ 17▶ 含视频演示看原帖 ↗
@levelsio@levelsio · 博主 · 1 天前独立开发者标杆,AI 产品连续创业者

使用Tailscale时早晚会发生一件事。

你觉得网站挂了,想SSH进去检查,连不上,浏览器打开也不行,绝望中准备硬重启...

然后想起来试试Hetzner的Rescue控制台,SSH进去问Claude Code怎么回事。

它告诉你服务器的Tailscale密钥过期了,因为它们会自动在180天后过期(出于安全考虑)。

@DanielLockyer告诉我应该给每台服务器都禁用密钥过期,但我给Interior AI漏了,结果就被锁在外面了,所以这个设置真的很关键!你可以在@Tailscale控制台的[...]里关闭。

现在都好了!!

查看英文原文
Something that will eventually happen when you use Tailscale is

You think your site is down, you try reach it over SSH, it doesn't work, you try open the site in your browser, it doesn't open, you're about to reset the server and think the worst!

Then you think: let's try the Rescue console on Hetzner, you SSH in, and ask Claude Code what's going on

And it tells you your Tailscale key for the server expired, because they automatically expire after 180 days (for safety)


@DanielLockyer
told me to always [ Disable key expiry ] for every server, but I forgot it for Interior AI, you kinda get locked out otherwise so important to disable this! You can do so in the
@Tailscale
console under [...]

All good now!!
Ethan Mollick@emollick · 创始人 · 1 天前沃顿商学院教授,AI 应用研究权威

我一直在想那个著名的图表,展示了专家们如何年复一年地预测太阳能装机增长是线性的,但实际增长是指数级时总是预测错了。

我觉得在AI产品战略的讨论中也发生了同样的事情。

查看英文原文
I have been thinking about the famous chart showing how experts keep projecting linear growth in solar installations, year after year, and always get it wrong when growth is still exponential.

I think the same thing is happening with the discourse on product strategy around AI.
Aravind Srinivas@AravSrinivas · 创始人 · 1 天前Perplexity 联合创始人兼 CEO

收益远高于 50%。我们即将在 Vera CPUs 上发布详细的性能指标。

引用 Beth Kindig @Beth_KindigAI 初创公司 Perplexity 计划采用 Nvidia 新款独立 Vera CPU,其执行代理编码任务速度比传统 CPU 快 1.5 倍。查看被引原帖 ↗
查看英文原文
The gains are much higher than 50%.

We will be publishing detailed metrics soon on Vera CPUs.
Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

三个选项:

- 7月19日或21日发布Opus 5安抚大众并移除Fable 5
- 再次延长Fable 5的支持,然后被群嘲Anthropic扛不住OpenAI的压力
- 下周结束前找到最终方案,确保Fable 5长期存活(更多算力)

我现在觉得,啥都不干直接把Fable 5砍掉的可能性极低。

引用 Claude @claudeaiWe're extending Claude Fable 5 access on all paid plans, as well as keeping Claude Code’s weekly rate limits 50% higher, through July 19.查看被引原帖 ↗
查看英文原文
Three options:

- They release Opus 5 on July 19th or 21st to appease the public and remove Fable 5.

- They extend Fable 5's support again, and many people make fun of the fact that Anthropic can't withstand OpenAI's pressure.

- They find a final solution by the end of next week to ensure Fable 5's long-term viability (more computing power).

I now consider it highly unlikely that virtually nothing will happen and Fable 5 will be removed without a solution.
Guillermo Rauch@rauchg · 创始人 · 1 天前Guillermo Rauch,Vercel 创始人兼 CEO

让模型成为你拥有的机器里的一个齿轮。

◾ AI SDK → 开放模型 API
◾ Eve.dev → 开放 Agent API
◾ AI Gateway → 开放 ZDR 推理

初创公司和企业必须拥有自己的数据、评估、模型选择、软件层。不要外包你的大脑。

查看英文原文
Make the model a cog in a machine you own.

◾ AI SDK → open model API

Eve.dev
→ open Agent API
◾ AI Gateway → open ZDR inference

Startups and enterprises must own their data, evals, model choices, software layer. Don't outsource your brain.
Ethan Mollick@emollick · 创始人 · 1 天前沃顿商学院教授,AI 应用研究权威
连环推 ×3

很少人意识到当前的模型在 Code/Codex 这类场景里,配置得当能做多少实用的工作。

这不是什么「你还这么早进场」的鸡汤文,这是在吐槽「AI 公司根本不会好好解释自己的系统能做什么」。

查看英文原文
Very few people know the amount of useful work that the current models can do in Code/Codex/etc. with the right setup

This is not a "rah rah you are so early" post, this is a "AI companies are doing a really bad job explaining what their systems actually do in a clear way" post.
And the worst part is that they could get AI to help them customize their explanation to individual needs.
Since some people were confused, the exponential curve I had was not made up, it was from METR:
metr.org/time-horizons/
Aravind Srinivas@AravSrinivas · 创始人 · 1 天前Perplexity 联合创始人兼 CEO

我觉得讽刺的是,现在的做法变成了:一面对 distillation 设置限制,一面又要从用户数据中学习。如果只有一个方向的学习流动,经济价值就会全部流向掌握学习基础设施的人,而不是知识的创作者。因此,我们必须把学习基础设施分散到每个公司手中,让他们能够控制自己的学习循环。

说得好。

查看英文原文
"I find it ironic that the status quo is to then turn around and impose restrictive terms on distillation, and to reserve the right to learn from customer usage and interaction data. If learning flows in only one direction, economic value converges toward the owners of the learning infrastructure rather than the creators of the knowledge itself. Therefore, it's imperative that we distribute the learning infrastructure to every firm so that they can control their own learning loop."

Well said.
Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

GPT-5.6 成功了

引用 Claude @claudeai在所有付费计划上扩展Claude Fable 5访问权限,同时将Claude Code的周速率限制保持50%更高,至7月19日。查看被引原帖 ↗
查看英文原文
GPT-5.6 was a success.
◔ 18.8 万 次浏览(3 条合计)♥ 4,507⇄ 155观点看原帖 ↗
歸藏(guizang.ai)@op7418 · 中文博主 · 1 天前歸藏,中文圈 AI 工具与提示词博主

他妈的,谁给老子传的呀?我操!

我这 Skills 全是开源的,怎么还一份 199 卖 40 万呀?

谁卖的?他妈分我点。

Simon Willison@simonw · 博主 · 1 天前Django 框架联合创造者,AI 工具深度评测

我看到有人预测 Opus 5 很快就会出来,而且会比 Fable 5 更好,但 Anthropic 有澄清过他们相对命名体系的逻辑吗?

我原来以为是 Haiku < Sonnet < Opus < Fable < Mythos,但 Fable 是不是应该在 Sonnet 和 Opus 之间?

查看英文原文
I've seen a few people predicting that Opus 5 will be out soon and will be better than Fable 5, but have Anthropic clarified how their relative naming scheme works yet?

I assumed it was Haiku < Sonnet < Opus < Fable < Mythos - but is Fable meant to go between Sonnet and Opus?
Alexandr Wang@alexandr_wang · 创始人 · 1 天前Scale AI 创始人,Meta 超级智能实验室负责人

muse spark 1.1 在一个新的、高难度的有限模型论/理论计算机科学评测中超越了 Opus、Grok 4.5 和 Gemini

引用 Serafim Batzoglou @s_batzoglou基准测试新模型(Sol、Terra、Luna、Fable 5、Meta Muse Spark 1.1、Grok 4.5)的归纳推理能力。所有新模型性能出色,部分更善于返回可泛化的简单公式。Fable 5仅在medium/low设置下可用,Grok 4.5扭转XAI下降趋势。查看被引原帖 ↗
查看英文原文
muse spark 1.1 outperforms opus, grok 4.5, and gemini on a new challenging finite model theory / theoretical cs eval
向阳乔木@vista8 · 中文博主 · 1 天前向阳乔木,中文圈 AI 工具与趋势博主

新版的ChatGPT对话学英语

测试发现bug很多,指令遵循很差,比如不让说“很棒很棒”,完全纠正不过来。

还有会输出一些莫名奇妙的句子,但对话学单词还是ok的。

◔ 10.3 万 次浏览♥ 272⇄ 22▶ 含视频观点看原帖 ↗
@levelsio@levelsio · 博主 · 1 天前独立开发者标杆,AI 产品连续创业者

你知道最好用的空气净化器吗?要能轻松用 Home Assistant 这样的方式控制,不用破解和额外硬件。我需要一个能响应 Airthings PM1/PM2.5 传感器并自动启动净化的

我们这空气还行,但附近有工地施工,所以经常有建筑灰尘

查看英文原文
What's the best air purifier you know that is easily controllable with Home Assistant etc without hacking and hardware add ons? I need something that can react to the Airthings PM1/PM2.5 sensor and start purifying then

We have clean air but lots of construction near so we get construction dust
Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

他们不仅移除了 5 小时的限制,我的费率也被彻底重置了。

OpenAI 现在真是在做非常棒的社区工作呀。赞一个。绝对是优秀的社区工作。

引用 Tibo @thsottiauxCodex 和 ChatGPT Work 重大更新:取消五小时使用限制、优化 GPT 5.6 Sol 效率、突破 600 万活跃用户并即将重置配额。查看被引原帖 ↗
查看英文原文
Not only did they remove the 5-hour limit, but my rates were completely reset again.

OpenAI is doing incredibly good community work right now. Kudos. Absolutely outstanding community work.
◔ 15.1 万 次浏览(4 条合计)♥ 2,739⇄ 88观点看原帖 ↗
Orange AI@oran_ge · 中文博主 · 1 天前Orange AI,中文圈 AI 产品观察博主
连环推 ×2

A 社的算盘是,Fable 5 太占 GPU 了,放 coding plan 里没有 ROI,就搞了个限时体验,然后按 API 的售价销售并且大赚。
结果 GPT 5.6 sol 出了,把用户都抢走了,Fable 5 的 GPU 问题也解决了,那就继续放在 coding plan 里吧。
一个物品的价格是由可替代品决定的。

Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

我觉得这是在暗示 GPT-5.6 有根本问题,而"GPT-5.6 比 GPT-5.5 用的 token 明显多"这说法不仅仅是道听途说。

"将让 GPT 5.6 Sol 整体更高效的改动,会体现在更低的使用量上"

别搞错,不是说这个模型本身烂,而是在推理能力或 token 效率方面存在根本问题。

引用 Tibo @thsottiauxCodex 和 ChatGPT Work 重大更新:取消五小时使用限制、优化 GPT 5.6 Sol 效率、突破 600 万活跃用户并即将重置配额。查看被引原帖 ↗
查看英文原文
I read this as an admission that something is fundamentally wrong with GPT-5.6, and that the claim that GPT-5.6 uses significantly more tokens than GPT-5.5 is not based on purely anecdotal evidence.

"changes that will make GPT 5.6 Sol more efficient across the board and that will be reflected in less usage"

Don't get me wrong. It’s not fundamentally flawed in the sense of being a bad model, but rather fundamentally flawed in terms of reasoning or token usage.
Lisan al Gaib@scaling01 · 博主 · 1 天前高频 AI 模型测评与爆料博主

即使Anthropic造出了AGI,也永远达不到这张图的气质

我觉得这永远不会被超越

引用 Guillermo Flor @guilleflorvsAnthropic现在查看被引原帖 ↗
查看英文原文
even if Anthropic were to build AGI they would still be infinitely far away from the aura of this picture

i don't think it will ever be surpassed
向阳乔木@vista8 · 中文博主 · 1 天前向阳乔木,中文圈 AI 工具与趋势博主

AI剪辑工具 ChatCut 最近大火,极简安装使用教程如下:

安装方法:
1. 发给 Codex,安装这个插件
github.com/ChatCut-Inc/agent…


2. 会安装一个MCP和Skill,做一次OAuth授权。

3. 做产品介绍简单,跟Codex对话说:

用chatcut的mcp和skill给 【网址】 做一个图文并茂,有配音解读的功能演示视频。

效果怎么说呢?虽然粗糙,但比文字生动多了 😂

◔ 7 万 次浏览♥ 812⇄ 201▶ 含视频教程看原帖 ↗
Ethan Mollick@emollick · 创始人 · 1 天前沃顿商学院教授,AI 应用研究权威

如果对未来有某种明确预期的话,规划如何使用AI会容易得多。

即便是"我们完全打算每周继续扩展,但如果出现以下情况可能需要停止,这是我们当前的状态"这样的说法都会更好。

引用 Claude @claudeai所有付费计划用户可继续使用 Claude Fable 5,Claude Code 的每周速率限制也将继续保持提升50%,有效期至7月19日。查看被引原帖 ↗
查看英文原文
Planning for how to use AI is a lot easier if there is some clarity about what to expect in the future.

Even a "we fully intend to keep extending this week by week but may need to stop under the following conditions and here is what the current status is" would be better.
Pietro Schirano@skirano · 博主 · 1 天前设计师出身的 AI 编程与创意博主

做了个网站来更好地浏览 Coding Agent Index,发现了不少有意思的地方:

Terra Max 略胜 Fable 5 Max(77.4 vs 77.2),成本却便宜 76%。

Sol XHigh 落后 Max 1 分,API 成本便宜 26%。

Luna Max 超过 Opus 4.8 Max,成本低 80%。
👇

查看英文原文
Built a site to explore the Coding Agent Index better, with a few surprises:

Terra Max edges Fable 5 Max (77.4 vs 77.2) for ~76% less per task.

Sol XHigh is 1 point behind Max at ~26% lower API cost.

Luna Max beats Opus 4.8 Max for ~80% less per task.
👇
◔ 6.9 万 次浏览♥ 443⇄ 13▶ 含视频研究看原帖 ↗
Runway@runwayml · 公司官方 · 1 天前AI 视频生成公司 Runway

每个人都有故事。每个物件都有故事。一切都有故事可讲。用Runway,你可以分享你的故事,不管它是什么。

《FLICKER》是一部关于一盏有故事的灯的短片。有不完美、几乎被遗忘,这部片子讲述的是什么叫为他人带来光明。

让你的故事在app.runwayml.com活过来

查看英文原文
Everyone. Every object. Everything has a story to tell. With Runway, you can share yours, no matter what it is.

FLICKER is a short film about a lamp with a past. Imperfect and all-but-forgotten, FLICKER tells the story of what it means to bring light to others lives.

Make your own story come to life at
app.runwayml.com
◔ 6.6 万 次浏览♥ 316⇄ 46▶ 含视频演示看原帖 ↗
Google DeepMind@GoogleDeepMind · 公司官方 · 1 天前谷歌旗下 AI 研究机构,Gemini 背后团队
连环推 ×4

看看我们如何用Google @Antigravity 的"预测过去"功能追踪一个罗马戒指小偷、绘制横跨欧洲的古代教团分布图,还重建了造访希腊神谕者们的人际网络。完整长文如下🧵

查看英文原文
Here’s how we used the Predicting the Past Skill in Google
@Antigravity
to track down a Roman ring thief, map an ancient cult across Europe, and reconstruct the networks of people visiting a Greek oracle. 🧵
🔍 The ring thief of Aquae Sulis

When given an 1,800 year old curse tablet, the Skill used Aeneas - our generative model for restoring, dating and placing ancient texts - to locate it in time and space. It also generated an explanation of why it made that prediction, acting as a piece of epigraphic commentary to the expert.
🗺️ Mapping the cult of the Aufaniae

The Skill can study multiple texts in parallel, we used it to map stone altars dedicated to the Aufaniae - Germanic goddesses. It showcased how religious practices traveled with Roman soldiers, even flagging an outlier in Spain by a veteran who brought his favorite deity home.
🔮 Who visited the Oracle of Dodona?

By analyzing collections of ancient lead tablets, the Skill mapped visitors traveling from across the ancient world. It reconstructed the community of oracle visitors, turning scattered fragments into a connected network.
◔ 6.2 万 次浏览♥ 364⇄ 43▶ 含视频演示看原帖 ↗
yetone@yetone · 中文博主 · 1 天前开源 AI 编程插件 avante.nvim 作者,开发者圈博主

我已经不看 harness benchmark 评分了,我现在一般让 SOTA 的模型在相当糟糕和随意设计的 harness 中运行,凭什么人工智能出来以后人类工作被极度剥夺只能在又脏又乱的环境中干脏活累活延续生命,而我们又拼命地给大模型捋顺毛为它们打造舒适干净绫罗绸缎般的 harness 来作为它们的运行环境,听我说,这不公平。

Aravind Srinivas@AravSrinivas · 创始人 · 1 天前Perplexity 联合创始人兼 CEO

Grok 4.5 在 Perplexity Computer 框架内表现不错,特别是在处理人们真实工作的时候。

引用 @jason @Jason使用Grok 4.5和Perplexity Computer开发了PODMEME播客聚合器,可按主题播放多个播客;成本为$11(1100积分)。查看被引原帖 ↗
查看英文原文
Grok 4.5 inside Perplexity Computer harness excels at work people actually do
Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

还好,说明我们不是空想这些。GPT-5.6 的上下文窗口已减少到 272k tokens,subagents 正在调整,还有其他变动,所以 burn rate 应该会明显下降。

引用 Tibo @thsottiauxUpdates for Codex and ChatGPT Work users. No nerfing, only good stuff! - We have landed inference optimizations and are passing down savings to all the subscriptions for GPT-5.6 Sol. That should result in around 10% more usage on its own. - We noticed that by changing the context size limit in the product to 372k for GPT-5.6 Sol, up from 272k for GPT-5.5, it resulted in more usage being charged than intended. We have reverted to 272k and will work to roll back out to 372k in the days to come. You should notice that usage drains significantly less after this change. - To understand where the extra usage was coming from, we ran some experiments where reasoning efforts were changed (referred to as juice values under the hood) and have reverted this. - There is slightly more usage of multi-agent than intended in high and xhigh reasoning effort, we are fixing this going forward. Also fixing a small other thing we noticed with auto-review where we can be more efficient. And we continue to have the 5h limit temporarily not apply. Enjoy the rest of the weekend!查看被引原帖 ↗
查看英文原文
It’s good to see that we weren’t imagining all of this. GPT-5.6’s context window has been reduced to 272k tokens, subagents are being adjusted, and other changes are being made, so the burn rate should decrease significantly.
hardmaru@hardmaru · 创始人 · 1 天前David Ha,日本 AI 公司 Sakana AI 联合创始人

语言模型和编码代理确实很厉害,但除了 LLM agent 啊,生活中还有更多值得探索的事,AI领域也是如此。

查看英文原文
Language models and coding agents are great, but there is more to life, and more to AI, than just LLM agents.
Rowan Cheung@rowancheung · 博主 · 1 天前AI 日报 The Rundown 创始人
连环推 ×2

AI 会发掘历史上数千个尘封的时刻。

比如,它刚读完一卷被烧成焦炭的 2000 年前的古籍。

查看英文原文
AI is going to uncover thousands of hidden moments in history.

Case in point: it just read a 2,000-year-old scroll burned into solid charcoal.
The Herculaneum scrolls were buried under this Roman villa when Mount Vesuvius erupted in 79 AD.

Every attempt to physically unroll them since has instantly destroyed the layers.

In 2023, a global contest called the Vesuvius Challenge launched to use AI to read the scrolls.
◔ 5.3 万 次浏览♥ 278⇄ 33▶ 含视频演示看原帖 ↗
歸藏(guizang.ai)@op7418 · 中文博主 · 1 天前歸藏,中文圈 AI 工具与提示词博主

老马的 grok build CLI 会打包上传你项目的整个代码库,这事办的太离谱了。

主要是它会把你的一些密钥上传上去,这一旦泄露的话,风险还是很大的。

幸亏我一直没来得及试它那个东西。

歸藏(guizang.ai)@op7418 · 中文博主 · 1 天前歸藏,中文圈 AI 工具与提示词博主

Codex 改的还是挺快的,已经把 Work 和 Codex 的区别弱化了,接下来估计会优化 ChatGPT 那部分的交互

引用 歸藏(guizang.ai) @op7418OpenAI 昨天晚上发了一堆东西,乱七八糟的。 GPT 5.6 终于发布了,随之发布的还有全新的 ChatGPT 应用(也就是 Codex 应用,改了个界面、换了个名字)。 现在 OpenAI 全面学习 Anthropic 产品的命名方式。 无论是从新的 5.6 命名上(原来他们是通过 Pro Mini 这种方式去命名,现在通过全新的名字比如 Sol 命名),还是产品形态上,Codex 应用现在硬生生被改成了类似 Claude 桌面端的方式。 他们硬拆分出了 Work 和 Code 两个 Tab。 你在切换的时候,整个界面除了左上角那个标志换了,其实根本看不出区别。 更别扭的是,聊天本身被缩进到了右下角一个非常小的弹窗里,整个界面充满了这种别扭的痕迹。 我一直坚持认为,Work 和 Code 根本没必要硬去区分,没必要在界面上透露出“我这个东西只能 Code”或者“我这个东西只能 Work”,用户直接用就行了。 本来经过几个月,大家已经建立起了关于 Codex 软件的认知,这一波操作直接给整没了,过去几个月全白干。现在每个人都很困惑,找不到自己的东西在哪,甚至找不到自己以前在 ChatGPT 软件里的聊天记录,非常蛋疼。 除了这些,还有几个新变化: 1. Site 插件上线: Codex 里面的 Site 插件已经上线了,可以直接去插件里面找。我试了一下,它能直接帮你生成网页,生成时会提供多套网页供你在界面上选择。 生成网页之后,它还可以连接你已有的业务数据,同时还能帮你部署到 OpenAI 的网站(部署到线上),方便你分享给朋友,这个非常好用。 2. 移动端调用插件:移动端的 ChatGPT 现在也可以调用原来 Codex 里面的那些插件了。 3. 基础能力升级:它的“浏览器使用”和“计算机使用”功能都进行了升级,变得更快、更准确了。 同时,GPT 5.6 的前端能力终于提升了,不会生成千篇一律的垃圾界面。 现在 ChatGPT 移动端 App 也可以进行切换了,网页端和移动端马上都会同步切换这个 Work 能力。 好嘞,更新就这些,赶紧去玩玩吧!查看被引原帖 ↗
Ethan Mollick@emollick · 创始人 · 1 天前沃顿商学院教授,AI 应用研究权威

这是个很靠谱的观点:我们确实不知道 AI 对工作会造成什么样的影响,这是显而易见的。理解这一点对制定政策至关重要,既要缓解负面影响,也要促进积极影响。数据对行动来说是关键的,我们需要现在就开始收集。

引用 Erik Brynjolfsson @erikbryn关于AI与经济的声明。AI可能在未来十年变得更强大,带来比工业革命更大但时间跨度更短的经济变革。包括风险(大规模失业)和机遇(生活水平提升)。经济学家、政策制定者和技术领袖必须采取行动,理解AI经济学并建立激励机制、保护措施和制度。查看被引原帖 ↗
查看英文原文
This seems like a reasonable statement: it seems self-evident that we do not know the impact of AI on work & trying to understand that will be critical to policies to mitigate bad impacts and encourage good ones. Data will be critical to action & we need to start gathering it now
Min Choi@minchoi · 博主 · 1 天前AI 产品演示博主,专门展示新工具玩法

我们还没准备好...

Opus 5 即将推出...

引用 Claude @claudeai我们延长 Claude Fable 5 在所有付费计划中的访问权限至 7 月 12 日。查看被引原帖 ↗
查看英文原文
We are not ready...

Opus 5 soon...
宝玉@dotey · 中文博主 · 1 天前宝玉,中文圈 AI 翻译与科普大 V

我把那种号称省 Token 的 Skill 叫“电报体 Skill”。

我们小时候语文课要学电报文,老师先给出一件事,比如母亲生病,要让在外地工作的哥哥赶紧回家,然后全班比赛拟电文,看谁用最少的字把事情说得最清楚。

最后能精简到四个字:“母病速归”。但再精简成“母病”或者“速归”都不行,因为收到的人不明白该干嘛或者发生了啥。

那时候要写电报体是为了省钱,电报按字收钱,每个字都是钱。

这种“惜字如金”的技能,现在在 AI 圈子里以 Skill 的形式复活了。

GitHub 上有个叫 Caveman 的项目,2026 年 4 月上线三天就冲上 Trending 第一,目前攒了 8.7 万颗星。它的作者 Julius Brussee 是荷兰莱顿大学一个 19 岁的大一新生,做的事情极其简单:在 Claude Code、Codex 等 AI 编程工具的提示词(prompt)里加一段指令,让 AI 像原始人一样说话。删冠词、删客套、删连接词,只留技术要素。

项目 README 声称能省 65% 的输出 token。听起来很牛逼。但跟电报体一样,这是一个过渡期的产物,而且它在省钱这件事上的效果,远没有看起来那么大。

JetBrains 最近有个测试:《Does Speaking to Agents Like Cavemen Really Save 65% of Tokens? We Test》

blog.jetbrains.com/ai/2026/0…


他们用 Claude Code 跑 SkillsBench 上的 86 个真实编程任务,装 skill 和不装各跑一遍,同任务、同模型、同预算,前后约 240 次计费试验,总共花了 106 美元。为了给 Caveman 最好的发挥空间,他们还强制它在每次回复中生效。

结果:输出 Token 只省了 8.5%。因为是强制开启,这 8.5% 已经是天花板,日常使用里它得自己判断要不要触发,只会省得更少。

为什么差这么远?因为 65% 这个数字来自聊天场景。你问 AI 一个问题,它回你一大段话,把客套和废话砍掉,确实能省一大半。

但智能体的 Token 消耗大头从来不是在聊天,工具调用、系统提示词、各种 Skills、MCP等等这些才是大头。

Caveman 优化的那部分,在整张账单里本来就是零头。好比公司要压缩差旅费,机票、酒店和打车一项没动,先把每天 2 块钱的矿泉水取消了。

电报体也不是没代价的,一句“Fixed auth. Tests pass.”看起来很省 Token。但它没告诉你修的是登录过期、权限校验还是刷新令牌;跑的是一个单元测试,还是完整测试套件;有没有改数据库;有没有留下兼容性风险。

这些信息不一定每次都要展开成小论文,但不能因为“说话像穴居人”就固定删掉。开发者看不懂 Agent 做了什么,只好追问。Agent 再读一遍文件、再跑一遍测试、再解释一次。前面省下的几十个 Token,很快会被新一轮工具调用吃回去。

电报体能工作,靠的是双方共享大量背景。“母病速归”只有四个字,收报人知道母亲是谁、家在哪里、为什么要回去。编程 Agent 处理的是不断变化的代码和陌生任务,共享背景没那么可靠。

语言越短,对默契的要求越高。啰嗦,有时候就是通信协议里自带的纠错码。

电报体后来怎么样了?没有人宣布废除它。长途通信的价格降到忽略不计之后,“母病速归”自然变回了“妈住院了,你买最早的票回来,到了给我打电话”。省字数是给按字计费时代做的优化,当价格降下去了,优化就没必要了。

Token 也一样,模型单价总体是在下降的,提示词缓存(prompt caching)能让重复读取的上下文便宜差不多九成。

真正能节约成本的是上下文管理和少走弯路。少加载些没必要的 MCP 和 Skills,用聪明一点的模型少一些返工,这都比电报体省钱多了。

引用 Tz @Tz_2022我这里再放一个暴论: 当前所有以节约 token 为目标的各种 skill / harness,都是阶段性产物,很快就会扫入历史的垃圾堆。。。 这就是短消息按字数收费的那个时代,在钻研怎么发尽可能少字数的短信把事说清楚的那些奇技淫巧。。。查看被引原帖 ↗
Amjad Masad@amasad · 创始人 · 1 天前Amjad Masad,Replit 创始人兼 CEO

看 Replit 的 computer use 模型和我的新棋引擎对弈,有意思

引用 Amjad Masad @amasadVibe Research在Replit微调Qwen-8b模型进行国际象棋游戏,运行3个并行实验取得进展。模型的机器学习能力进步显著,现可凭直觉指导进行有趣的ML工作。查看被引原帖 ↗
查看英文原文
Fun to see Replit's computer use model play against my new chess engine
◔ 4.5 万 次浏览♥ 220⇄ 13▶ 含视频演示看原帖 ↗
Aravind Srinivas@AravSrinivas · 创始人 · 1 天前Perplexity 联合创始人兼 CEO

计算机本质上是一个工作空间。我们允许多账户和类似 Gmail 的账户切换功能。

引用 Computer @AskPerplexityPerplexity Mac应用现已支持账户切换。可在侧边栏个人资料菜单或账户设置中切换已保存账户、添加新账户或管理账户。现已在Mac应用中为所有Computer用户上线。查看被引原帖 ↗
查看英文原文
Computer is essentially a workspace. We’re allowing multiple accounts and account switching similar to GMail.
◔ 4.5 万 次浏览♥ 384⇄ 21▶ 含视频教程看原帖 ↗
hardmaru@hardmaru · 创始人 · 1 天前David Ha,日本 AI 公司 Sakana AI 联合创始人

物理系统如何在没有中央大脑的情况下实现集体智能和自我修复?

今天《Nature Communications》上发表了一篇新论文,由我在Sakana AI的同事Sebastian Risi(@risi1979)与哥本哈根IT大学和Autodesk Research的合作者共同完成。这篇论文提出了一个基于生物启发的美妙机器人实现:智能蜂窝积木。

研究团队构建了一套物理3D立方体单元系统。这些单元仅通过局部交互,就能集体推断自身全局形状,并自主引导损坏修复。

以下是该论文的几大关键贡献:

1/ 基于神经细胞自动机的架构:
模块化机器人通常依赖中央处理器。而这个系统彻底颠覆了这一范式。每个模块都在本地微控制器上独立运行相同的神经网络。没有全局蓝图或坐标,只与相邻模块通信。通过传递连续状态向量,数百块积木能在3分钟内对其形状达成全局共识。

2/ 涌现的生物形态发生素:
积木如何知道自己是椅子的一部分而不是桌子?网络的内存自动学会在结构上建立连续梯度,完美模仿了生物形态发生素为发育中的细胞提供位置信息的机制。这些积木自然形成左右轴、径向轴和头尾轴来统一自己的身份。

3/ 性能与泛化能力:
经过大规模模拟验证后,这些网络无缝迁移到近200块物理硬件积木上,收敛率达到100%。系统不是进行僵硬的模板匹配,而是推断出广泛的类别。即使面对从未见过的变体(比如一个五条随机腿的不对称桌子),集体依然能正确分类出结构。

4/ 容错与自主损坏恢复:
现实世界中的硬件会坏。该系统能轻松承受多达15%的模块失效,准确率却不受影响。通过预测空间损坏方向,细胞定位缺失部件的准确率高达95%。它们主动利用这些局部信号引导自我修复过程,重新恢复成预期形态。

我认为这是一项意义重大的研究,连接了集体智能与Physical AI。

这项工作是首次成功实现大规模去中心化3D自我识别与损坏检测的实际应用。远离集中控制后,这一架构为高度自适应的智能材料和具备自我修复能力的韧性机器人铺平了道路。

完整开放获取论文链接:

nature.com/articles/s41467-0…


恭喜团队取得这一成就!

引用 Sakana AI @SakanaAILabsWe are pleased to share our latest research, now published in Nature Communications: “Smart Cellular Bricks: Physical Modules That Recognize Their Own Shape and Repair Themselves.” Blog: sakana.ai/smart-cellular-bri… Paper: nature.com/articles/s41467-0… A long-running theme in our work is collective intelligence: the idea that sophisticated, robust behavior can emerge from many simple parts following local rules, with no central controller, as it does in a colony, a tissue, or a brain. We had mostly studied this in software and simulation. So this time we asked a simple question. Do the same decentralized principles hold up in the physical world, where communication is noisy and modules fail? To find out, we built a collection of simple cubic bricks. Each brick runs the same small neural network and talks only to the bricks it is physically connected to. No brick is told its position, or which shape it is part of. Yet from these purely local exchanges, the collective converges on the correct global shape, locates where modules are missing or damaged, and can even guide its own repair, inspired by how living tissue self-organizes and regenerates after injury. For us, this is a first step in a broader direction: taking the principles of collective intelligence we have studied in software and letting them emerge, decentralized and robust, in the physical world. In the future, we imagine smart materials that let structures sense and report damage on their own, and LEGO-like systems that recognize their own configuration and adapt in real time, pointing toward environments that are more robust, adaptive, and regenerative. This work is a collaboration between Sakana AI, IT University of Copenhagen and Autodesk.查看被引原帖 ↗
查看英文原文
How do physical systems achieve collective intelligence and self-repair without a central brain?

A new paper published today in Nature Communications by my Sakana AI colleague Sebastian Risi (
@risi1979
), along with co-authors from IT University of Copenhagen and Autodesk Research, presents a beautiful realization of biologically inspired robotics: Smart Cellular Bricks.

The team built a system of physical 3D cubic units that can collectively infer their global shape and autonomously guide their own damage recovery using purely local interactions.

Here is a deep dive into the paper’s key contributions:

1/ Neural Cellular Automata-based Architecture:
Modular robots usually rely on central processors. This system flips that paradigm. Every block independently runs the exact same neural network on local microcontrollers. With no master plan or global coordinates, they communicate only with immediate neighbors. By passing continuous state vectors, hundreds of bricks achieve global consensus on their shape in under 3 minutes.

2/ Emergent Biological Morphogens:
How does a block know it is part of a chair, not a table? The network’s internal memory automatically learns to establish continuous gradients across the structure. This beautifully mirrors how biological morphogens give positional info to developing cells. The bricks naturally form left-right, radial, and head-to-tail axes to align their identity.

3/ Performance and Generalization:
Validated in large-scale simulations, the networks transferred seamlessly to nearly 200 physical hardware bricks, achieving a 100% convergence rate. Instead of rigid template-matching, the system infers broad categories. Even when tested on unseen variations, like an asymmetric table with five random legs, the collective correctly classified the structure.

4/ Fault Tolerance and Autonomous Damage Recovery:
Hardware fails in the real world. This system easily tolerates up to 15% module failure without losing accuracy. By predicting spatial damage directions, the cells pinpointed missing components with 95% accuracy. They actively use these local signals to guide a self-repair process, regenerating back into the intended morphology.

I believe this is a significant piece of research, bridging collective intelligence and Physical AI.

This work represents the first successful physical realization of large-scale, decentralized 3D self-recognition and damage detection. By moving away from centralized control, this architecture paves the way for highly adaptive smart materials and resilient robotics that can survive and repair themselves.

Read the full open-access paper:

nature.com/articles/s41467-0…


Congratulations to the team on this achievement!
◔ 4.4 万 次浏览♥ 350⇄ 43▶ 含视频研究看原帖 ↗
小互@xiaohu · 中文博主 · 1 天前小互,中文圈高频 AI 资讯站 Xiaohu.AI 主理人

OpenAI 官方指南:Codex 提示词 9 套工作流直接抄,修 bug、截图变原型、云端重构全在内
best.xiaohu.ai/article/chatg…

Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

我们正进入一个全新的科学纪元。

东京大学研究量子场论和弦理论的教授 Yuji Tachikawa 说,Claude Fable 5 破解了一个六个月毫无进展的合作研究难题。

Tachikawa 一时兴起,把自己的研究笔记给了它。Fable 找到了一个计算错误,抵达了研究人员们遇到的同样死胡同,然后在后续提示下拓展了另一种方法。

Tachikawa表示:"它做了一个不平凡的观察,基本上解决了这个问题。"

这个模型还用 SymPy 写了代码,并验证了自己的预测。他的结论是:

"Fable 似乎确实理解弦理论,并且还能敏锐地感知到。"

这只是一名研究人员的个人叙事,并非已经发表或经独立验证的结果。但它所描述的远不止是又一个基准分数的进步:前沿模型正在为活跃的理论物理研究贡献创新性的一步。

我爱这个!

引用 🥑 @yujitach最近 Claude Fable を試しているのですが、昨晩ふと、ここ半年ぐらい進展がない共同研究について研究ノートをみせて聞いてみたら、なんと非自明な観察をして、ほぼ解決してしまった。一回目の返事は「計算ミスは一ヶ所みつけましたが同じところで詰まりました」だったが、「そこはこう解決するかなと查看被引原帖 ↗
查看英文原文
We are entering a completely new era of science.

Yuji Tachikawa, a University of Tokyo professor working on quantum field theory and string theory, says Claude Fable 5 unlocked a collaborative research problem that had seen no progress for six months.

Tachikawa gave it his research notes on a whim. Fable found a calculation error, reached the same dead end as the researchers, and then expanded another approach after a follow-up prompt.

Tachikawa: “It made a non-trivial observation and essentially solved it.”

The model also used SymPy to write code and verify its own predictions. His conclusion:

“Fable probably seems like it properly understands string theory and has intuition too.”
This is one researcher’s account, not a published or independently verified result. But it describes something far more consequential than another benchmark gain: a frontier model contributing a novel step to active theoretical physics research.

I love it!
Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

递归自我改进这个话题正被越来越多的人认真讨论,成为一个真正重要的议题。

我们站在了一个前所未有世界的门槛上,根本无法预测接下来会怎样。我兴奋得难以自持。

引用 Tom Blomfield @t_blom个人更新:离开YC加入Anthropic,与Tom Brown合作从事计算基础设施团队工作。强大的AI有潜力改善每个人的生活。在递归自我改进的早期阶段,计算资源可用性成为最重要的问题。查看被引原帖 ↗
查看英文原文
More and more, people are beginning to seriously pursue the transition toward recursive self-improvement and discuss it as a genuinely important topic.

We are standing on the threshold of an unprecedented world, and it is impossible to foresee what that will mean. I am incredibly excited.
Bindu Reddy@bindureddy · 创始人 · 1 天前Abacus.AI CEO,AI 行业观点博主

太搞笑了!微软一直没搞出能卖得出去的 LLM

所以现在它建议每家公司都训练自己的模型

从一个连训都训不好的公司说出来…… 好吧,当然啦 🤯

查看英文原文
Hilarious! Microsoft hasn’t succeeded in making an LLM it can actually sell

So it is now recommending that every company should train their own

From a company that can’t train a half decent model…. Ok, sure 🤯
小互@xiaohu · 中文博主 · 1 天前小互,中文圈高频 AI 资讯站 Xiaohu.AI 主理人

今晚 Claude 用量,可能会重置一次

因为我已经用光了一周的用量了,周六才到我的重置时间

今晚必须要给我重置一次


@CuiMao

Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

现在的情况是,要么我们很快就能用上 Opus 5,要么 Fable 5 肯定会继续留在订阅计划。

现在的竞争真激烈啊。

引用 Tibo @thsottiaux我认为 GPT 5.6 很不错。查看被引原帖 ↗
查看英文原文
At this point, we'll either get Opus 5 soon, or Fable 5 will definitely remain on the subscription plan.

What a great competition is going on right now.
Min Choi@minchoi · 博主 · 1 天前AI 产品演示博主,专门展示新工具玩法

我喜欢竞争

引用 Claude @claudeai所有付费计划继续支持 Claude Fable 5 访问,Claude Code 周费率限制提升 50%,有效期至 7 月 19 日。查看被引原帖 ↗
查看英文原文
I love competition
swyx@swyx · 博主 · 23 小时前知名 AI 播客 Latent Space 主理人

看 @willccbb 被放到世界上后事业起飞,真的像是做梦一样。PI(尤其是 @asharoraa 和 @vincentweisser 还有 @jackminong,太多了数不过来)同时拥有惊人的天才密度和执行力,却常被忽视。正因为这样,他们那种带点半币圈的调调和搞氛围感的做法反而对他们不利——因为你很容易把他们当成那种只有氛围感、没有实质内容的人。如果这是刻意为之的,那坦白说还挺有门道

作为一个一年前还在 @jacobeffron 播客上公开对 RLaaS 生意失去希望的人,PI、AC 和他们的朋友们证明了上一代人其实没错,只是太早了/本事不够。既让人清醒又挺鼓舞的!

引用 AI Engineer @aiDotEngineer恭喜PI完成独角兽融资轮,达到1亿美元ARR!很荣幸能在一年前首届AIE NYC活动上让Will介绍验证器,现在它已成为v1了!Will加入难得的三次AIE演讲者名单,他关于完整PI技术栈的演讲链接见下方。查看被引原帖 ↗
查看英文原文
its been surreal to see
@willccbb
's career take off after getting unleashed on the world. PI (esp
@asharoraa
and
@vincentweisser
and
@jackminong
etc too many to name) are overlooked for having both incredible talent density and great execution. to this extent their crypto-adjacent vibes and aurafarming even works AGAINST them because you pattern match to people who -only- have vibes and nothing else. if this is intentional, it is kinda genius tbh

as someone who publicly (in our
@jacobeffron
pod) lost hope on RLaaS businesses over a year ago, PI, AC and friends have shown that the prevgen weren't wrong, just early/skill issue. both humbling and inspiring!
歸藏(guizang.ai)@op7418 · 中文博主 · 1 天前歸藏,中文圈 AI 工具与提示词博主

在 Anthropic 重置和延长 Fable5 使用时间之后

OpenAI 直接取消了 Codex 的五小时使用限制,同时进行了新一次的重置

引用 Tibo @thsottiaux过去48小时ChatGPT Work有重大更新:移除Plus/Business/Pro的5小时使用限制;GPT 5.6 Sol效率提升;达到600万活跃用户;即将进行使用额重置。查看被引原帖 ↗
Lisan al Gaib@scaling01 · 博主 · 1 天前高频 AI 模型测评与爆料博主

GSO 结果终于出来了!!

Opus 4.8 新 SOTA
GPT-5.5-xhigh 第 4 名
Sonnet 5 第 5 名

查看英文原文
GSO results are finally out!!

Opus 4.8 new SOTA
GPT-5.5-xhigh in 4th
Sonnet 5 in 5th
Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

当今最顶尖的AI研究者对AI的重视,不仅在于它的能力本身,更在于它能以前所未有的速度推动经济增长和科学进步。

我们也该对其给予同样的重视。

引用 Andrew Curran @AndrewCurran_More than 200 researchers and economists including Jack Clark, Jeff Dean, Noam Brown, Tyler Cowen, Sholto Douglas, John Schulman, Yoshua Bengio, Eric Schmidt, and Dean Ball have signed a statement urging governments and institutions to act now to prepare for AI's economic impact.查看被引原帖 ↗
查看英文原文
The most important AI researchers of our time take AI seriously - not only because of its capabilities, but because it could accelerate economic growth and scientific progress at an unprecedented pace.

We should take it seriously, too.
Orange AI@oran_ge · 中文博主 · 1 天前Orange AI,中文圈 AI 产品观察博主

关于 Codex 剪视频,我们 team 最近有一些独特的心得
正在做一个视频,我今天看了倒数第二版,比今天 tl 上的都要好个几倍吧 😂
顺利的话,明天发
如果比较受欢迎的话
就把 skill/plugin 开源出来

swyx@swyx · 博主 · 1 天前知名 AI 播客 Latent Space 主理人

顺便说一下,区别在于内省/反向传播

就像爱因斯坦的著名论述,疯狂的定义就是进行多次 rollouts 而没有任何收益预期

引用 blue @bluewmist有人能提供一些有用的句子,让我看到时能立即改变人生吗?查看被引原帖 ↗
查看英文原文
btw the difference is introspection/backpropagation

as einstein famously said, the definition of insanity is doing multiple rollouts with no expectation of advantage
Bindu Reddy@bindureddy · 创始人 · 1 天前Abacus.AI CEO,AI 行业观点博主

Anthropic 一直在提高他们的有效 API 价格....

Sonnet 5 的成本字面上是 Sonnet 4.6 的两倍

相比之下,OpenAI 正在变得更高效有效 - Terra 是个很不错的选择

从长期来看,这会让 OpenAI 比 Anthropic 更赚钱

查看英文原文
Anthropic keeps increasing their effective API prices....

Sonnet 5 literally costs 2x that of Sonnet 4.6

In sharp contrast, OpenAI is becoming more efficient and effective - Terra is an excellent option

Overtime, this will make OpenAI more profitable than Anthropic
AshutoshShrivastava@ai_for_success · 博主 · 1 天前高频 AI 新闻与产品动态博主

GPT-5.6 Sol vs Fable 5。

如果今天只能选一个用于代码编程,你会选哪个,为什么?

查看英文原文
GPT-5.6 Sol vs Fable 5.

If you could only pick one for coding today, which would you choose and why?
yetone@yetone · 中文博主 · 1 天前开源 AI 编程插件 avante.nvim 作者,开发者圈博主

对的,就是这个逻辑。凭什么要从小教育人类不要抱怨环境,要适应环境?而对大模型却百依百顺,富养着它们。真正聪明的大模型会自己改造自己的生存环境,而不只是等着我们投喂。

引用 Larry & Leo & Lucky 🍀 @xqliu我也是倾向于随心所欲,不给 AI 搞这么多好基建,不然要他们干啥呢?查看被引原帖 ↗
karminski-牙医@karminski3 · 中文博主 · 1 天前karminski-牙医,中文圈模型评测博主

小米够狠啊, 刚刚发布了 dflash 版 MiMo!

小米刚刚在自己的 HuggingFace 页面放出了 MiMo-V2.5-DFlash! 甚至都还没来得及更新README.

我赶紧看了一下它放出的代码实现, 发现有东西嗷.

这个推测性解码用的草稿模型, 是个block-diffusion模型!

注意, 不是U-Net那种连续diffusion, 骨干还是Transformer, 但解码范式是block diffusion: 一次forward并行猜一整块token, 再丢给大模型一次性验证.

跟传统的EAGLE那种自回归一点点draft完全不是一路货. 参考小米之前的 BLog, 编程场景接受长度能到 6+ (block=8, 接受长度就差不多等于提速倍数, 6x!), 所以它的提速效果会非常明显!

另外注意它是 Qwen/Z-Lab那套DFlash, 不是 DeepSeek DSpark. DSpark是把推测模块耦合进大模型同一个checkpoint里; MiMo这个是草稿权重单独拆出来发的, 就几层, 本地就2.94G.

我看了下, 这个草稿模型还是没办法单独当小模型用的. 它自己没有embed / lm_head, 每一步还得从MiMo多层hidden里抽特征做KV injection. 所以独立权重 ≠ 独立模型, 本质就是个加速插件.

另外它的Layer 选择是有点猛的, target_layer_ids配置中是[0, 11, 23, 35, 47], 连第0层都抽了;原版 DFlash 论文通常从浅层偏后位置开始均匀采样. 感觉会更吃早期表征, 或者需要适配 MiMo 的浅层语义分布.

以及它的block_size是8, 偏保守了. 论文和Qwen的实现是 10–16;更小意味着并行度略低, 但这样接受率通常更好, 偏部署稳妥.


#mimo
#dflash


链接在这里:
huggingface.co/XiaomiMiMo/Mi…

Simon Willison@simonw · 博主 · 1 天前Django 框架联合创造者,AI 工具深度评测

(与此同时,OpenAI 新推出的 Sol、Terra、Luna 命名约定对于任何熟悉太阳系天体拉丁名称的人来说都完全说得通)

查看英文原文
(Meanwhile OpenAI's new Sol, Terra, Luna naming convention makes perfect sense to anyone who's familiar with the Latin names for bodies in our solar system)
Tibor Blaho@btibor91 · 博主 · 1 天前逆向挖掘 AI 产品代码的爆料专家

我之前在5.5版本时只用xhigh,升到5.6 sol后,对很多任务来说medium就足够了

引用 Keyan Zhang @keyanzhangsome tips on 5.6 sol: 1. remove old slop. try disabling certain community skills and plugins, especially bundles with 20+ skills. think of 5.6 sol like someone who just grew from senior to staff or senior staff level: prescriptive guidance that used to help becomes micromanagement and makes the work worse. you can always add back what clearly helps. 2. turn on codex memory in settings. give codex feedback on what you like and don’t like, and ask it to remember. 3. you probably don’t need ultra. i do 95% of my work on sol high, sometimes use sol extra high, and have only used sol ultra for a handful of sessions. start with high and only move *up* when you’re not happy with the result.查看被引原帖 ↗
查看英文原文
I am used to only using xhigh from 5.5, but with 5.6 sol, medium is more than enough for many tasks
Hailuo AI-MiniMax Hub@Hailuo_AI · 公司官方 · 1 天前MiniMax 旗下海螺 AI 视频官方

MiniMax Hub 官方攻略 | EP3: 百宝箱魔法工具 🪄
全套强工具助力效率起飞

✏️ 擦除:别用不着部分轻松一键去除
✂️ 裁剪:智能构图重设姿势
✨ 分割网格:一个点击搞定故事板和社交图发布

#MiniMax
#Hailuo
#MiniMaxHub

查看英文原文
MiniMax Hub Official Guide | EP3: Magic Toolbar 🪄
Boost your workflow with powerful toolkit.

✏️ Erase: Effortlessly remove unwanted parts.
✂️ Crop: Smartly reframe and reposition.
✨ Split Grid: Create storyboards and social posts in one click.


#MiniMax
#Hailuo
#MiniMaxHub
Bindu Reddy@bindureddy · 创始人 · 1 天前Abacus.AI CEO,AI 行业观点博主

包括亚马逊、微软在内的大型云公司都在为Anthropic和OpenAI而感到担忧

怎么破?其实很简单 - 就学Grok和Muse spark最近的做法

发布便宜又性能强劲的模型 - 蚕食他们的API利润

查看英文原文
Big cloud including Amazon and Microsoft are panicking about Anthropic and OpenAI

It’s simple - Do what Grok and Muse spark just did

Release cheap and performant models - eat into their API margins
Orange AI@oran_ge · 中文博主 · 1 天前Orange AI,中文圈 AI 产品观察博主

你和你的竞争同时变强3倍,你们的收入差距:
1. 缩小了
2. 扩大了
3. 没变化

如果你的答案是2...再加上原来你是比较弱的那一方,就不是什么好消息了。。。
这不就是大模型提效给竞争带来的结果吗
实际是,你努力各种拥抱 ai ,结果只是被拉开更大的差距
除非你的竞争对手对 ai 无动于衷(这可能吗?),或者你的行业不依赖 ai coding 作为核心竞争力

Google DeepMind@GoogleDeepMind · 公司官方 · 1 天前谷歌旗下 AI 研究机构,Gemini 背后团队

这些例子展示了「预测过去技能」(Predicting the Past Skill)怎样用先进的 AI 模型和复杂工作流来推动历史研究——而且完全无需编码。

了解更多 →
goo.gle/4vR5ZtC

查看英文原文
These examples showcase how the Predicting the Past Skill can push forward historical research using advanced AI models and complex workflows - with no coding required.

Find out more →
goo.gle/4vR5ZtC
Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

天哪,Slavoj Žižek 发表了关于 AI 的演讲。差点错过了。必看 *sniff* *sniff*。

查看英文原文
Dang, Slavoj Zizek gave a speech on AI. Almost missed that. Instant watch *sniff* *sniff*.
小互@xiaohu · 中文博主 · 1 天前小互,中文圈高频 AI 资讯站 Xiaohu.AI 主理人

人们只相信AI给自己输出的结果

无法相信AI给他人输出的结果

是不是有这种顾虑?

就是你让AI生成的内容,你深信不疑,他人用AI生成的东西你总是嫌弃...

歸藏(guizang.ai)@op7418 · 中文博主 · 1 天前歸藏,中文圈 AI 工具与提示词博主

Seedream 5.0 Pro 的这个图像精细编辑能力真的非常牛!

它是模型能力和产品交互的一个结合,可以实现对图像非常精准的编辑。

甚至你标注的部分和生成的部分区域颜色能完美融合,绝对不会溢出。

我写了一篇教程,详细讲了一下如何在 Lumion 里面去玩这套东西,推荐去试试,是很不一样的体验。

🚨 AI News | TestingCatalog@testingcatalog · 博主 · 1 天前专挖 AI 产品未发布新功能的爆料号

ClickUp 正在 Brain AI 中开发自主收件箱管理功能!

> Brain 可以读取通知、筛选出重要的、清除其他的。
> 低优先级项目标记为已读、跟进被延迟、紧急项目浮出来。
> 它不只能生成摘要,还能直接起草回复、对收件箱项目采取行动。

凌乱的收件箱进去,出来就是干净的、优先级合理的视图,日常工作都已经处理好了。

查看英文原文
ClickUp is working on autonomous inbox management inside Brain AI!

> Brain can read notifications, triage what matters, and clear the rest.
> Low-priority items get marked read, follow-ups get snoozed, surfacing urgent items.
> It can draft responses and take action on inbox items directly, not just summaries.

A cluttered inbox goes in, and a clean, prioritized view comes back with the routine work already handled.
小互@xiaohu · 中文博主 · 1 天前小互,中文圈高频 AI 资讯站 Xiaohu.AI 主理人

我喜欢这种竞争激烈的氛围

🤓

Chubby♨️@kimmonismus · 博主 · 1 天前Chubby,高频 AI 新闻聚合博主

AI工厂中的隐藏瓶颈:数据基础设施——访VAST Data的Sven Breuner

所有人都在聊GPU。但在真实世界的AI基础设施里,瓶颈往往在别处:数据层。

这期节目,我和VAST Data的Sven Breuner聊了现代AI背后的基础设施——从GPU利用率、存储架构到HPC、推理、数据传输,以及新兴的“AI工厂”概念。

当企业重金砸向GPU时,很多人发现光有算力根本不够。数据转不动,GPU就只能闲着,训练效率走低,推理系统也扛不住规模化。

我们聊了这些:

• 为什么GPU老在等数据
• AI工作负载和传统HPC有啥区别
• 存储、元数据、网络成了瓶颈会怎样
• 数据层为啥对训练和推理都关键
• VAST所谓的“AI操作系统”是什么意思
• 企业该在买更多GPU前怎么想基础设施
• 为什么下一代AI工厂需要的不仅是更快的芯片
这是对AI栈中最重要却少被理解的部分的一次深度拆解:让我家能支撑大规模智能的基础设施。

YouTube:
invidious.tiekoetter.com/V7QsxbS2UXs

查看英文原文
The Hidden Bottleneck in AI Factories: Data Infrastructure with Sven Breuner, VAST Data

Everyone talks about GPUs. But in real-world AI infrastructure, the bottleneck is often somewhere else: the data layer.

In this episode, I speak with Sven Breuner from VAST Data about the infrastructure behind modern AI systems - from GPU utilization and storage architecture to HPC, inference, data movement, and the emerging concept of AI factories.

As companies invest billions into GPUs, many are discovering that compute alone is not enough. If data cannot move fast enough, GPUs sit idle, training becomes inefficient, and inference systems struggle to scale.

In this conversation, we discuss:

• Why GPUs often wait for data
• How AI workloads differ from traditional HPC workloads
• What happens when storage, metadata, and networking become bottlenecks
• Why the data layer matters for both training and inference
• What VAST means by an “AI Operating System”
• How enterprises should think about infrastructure before buying more GPUs
• Why the next generation of AI factories will require more than faster chips
This is a deep dive into one of the most important but least understood parts of the AI stack: the infrastructure that makes large-scale intelligence possible.

YouTube:
invidious.tiekoetter.com/V7QsxbS2UXs
Bindu Reddy@bindureddy · 创始人 · 1 天前Abacus.AI CEO,AI 行业观点博主

所有 AI 基准测试都彻底失效了

几乎所有测试都只衡量简单的单轮回应

LLMs 本质上就是针对这类问题优化成本和性能的

这些 LLMs 在真实场景中根本不靠谱,真实场景都是长上下文和多轮对话

查看英文原文
All AI benchmarks are completely broken

Almost all of them measure simple first turn responses

LLMs are literally tuned to optimize on cost and performance on such questions

These LLMs don’t hold up in the real world which is full of long context multi-turn conversations
小互@xiaohu · 中文博主 · 1 天前小互,中文圈高频 AI 资讯站 Xiaohu.AI 主理人

微软 CEO 纳德拉:企业面临一种“反向信息悖论”

支付费用调用 AI,同时也将自己的企业内部知识和经验拱手献出

他借用了经济学家 Kenneth Arrow 对信息商品的经典描述。买家只有看到一条信息后,才知道它值多少钱;可一旦看见,买家已经获得了它。

传统难题因此落在卖方身上:为了出售知识,卖方可能先泄露知识。

企业先支付模型调用费,然后继续提供内部知识,让模型理解自己的客户、流程和判断标准。

纳德拉称之为“为智能支付两次”。第二笔成本没有清晰账单,却可能包含企业最难复制的经验。

Tibor Blaho@btibor91 · 博主 · 1 天前逆向挖掘 AI 产品代码的爆料专家

OpenAI和Anthropic重磅一周 - GPT-5.6、ChatGPT Work、GPT-Live和J-space(2026年第28周)

OpenAI推出了GPT-5.6系列的通用版本,以Sol为旗舰产品,另有Terra和Luna,涵盖ChatGPT、Codex和API

随之推出的还有ChatGPT Work(一个基于Codex的新智能体)、合并Chat、Work和Codex的新桌面应用、Sites公测版和插件目录

语音功能也大升级了 - GPT-Live是新一代全双工语音模型,现已集成到ChatGPT Voice,GPT-Realtime-2.1登陆API

GPT-5.5 Instant Mini成为新的ChatGPT后备模型,ChatGPT for PowerPoint全面推出,GPT-5.6成为Microsoft 365 Copilot的首选模型

OpenAI撤回了对SWE-Bench Pro的推荐(发现约30%的任务有问题),发布了《国家安全原则》,宣布了面向K-12教育工作者的AI技能大赛,Bio Bug Bounty奖金翻倍至$50,000,分享了GPT-5.6 Sol Ultra证明了50年前的Cycle Double Cover Conjecture

群聊和Atlas浏览器被下线,Android应用中发现了类似DM风格的Messages标签页

在产品之外,Apple起诉OpenAI涉嫌盗取商业机密,Fidji Simo转为兼职顾问角色,首席未来学家Joshua Achiam和安全系统负责人Johannes Heidecke双双宣布离职

Anthropic发布了J-space研究,发现了Claude内部的全局工作区,将深思熟虑的思考与自动化处理分离,还有与AE Studio合作的GRAM研究,专注于双用途知识的可移除模块

Claude Cowork在Max计划中率先推出网页版和移动版公测,Claude Code和Cowork推出政府版(FedRAMP High环境),Microsoft 365连接器获得了写入工具,Claude for Open Source扩展到6个月20倍Max,新推出的Reflect可以让你回顾使用Claude的方式

Anthropic发布了阿尔伯塔省政府的案例研究(20小时内扫描4.66亿行代码),宣布与UST合作研发物理AI,发布了《Claude Code的制作》交互功能,任命Ben Bernanke进入Long-Term Benefit Trust,Cowork的导览晨报功能正在开发中

查看英文原文
A big week at OpenAI and Anthropic - GPT-5.6, ChatGPT Work, GPT-Live, and the J-space (Week 28, 2026)

OpenAI launched the GPT-5.6 family for general availability, with Sol as the flagship, plus Terra and Luna, across ChatGPT, Codex, and the API

With it came ChatGPT Work, a new agent built on Codex, a new desktop app merging Chat, Work, and Codex, Sites in public beta, and the Plugin Directory

Voice got a big upgrade too - GPT-Live, a new generation of full-duplex voice models, is now behind ChatGPT Voice, and GPT-Realtime-2.1 landed in the API

GPT-5.5 Instant Mini became the new ChatGPT fallback model, ChatGPT for PowerPoint went generally available, and GPT-5.6 became the preferred model in Microsoft 365 Copilot

OpenAI retracted their SWE-Bench Pro recommendation after finding around 30% of tasks broken, published their National Security Principles, announced an AI Skills Jam for K-12 educators, doubled Bio Bug Bounty rewards to $50,000, and shared that GPT-5.6 Sol Ultra produced a proof of the 50-year-old Cycle Double Cover Conjecture

Group chats and the Atlas browser are being retired, and a DM-style Messages tab was spotted in the Android app

Off the product side, Apple sued OpenAI over alleged trade secret theft, Fidji Simo moved to a part-time advisor role, and both chief futurist Joshua Achiam and head of safety systems Johannes Heidecke announced their departures

Anthropic published the J-space research, finding a global workspace inside Claude that separates deliberate thinking from automatic processing, plus GRAM research with AE Studio on removable modules for dual-use knowledge

Claude Cowork came to web and mobile in beta starting with the Max plan, Claude Code and Cowork launched for government in a FedRAMP High environment, the Microsoft 365 connector gained write tools, Claude for Open Source expanded with 6 months of Max 20x, and Reflect launched as a new way to review how you use Claude

Anthropic published the Government of Alberta case study on scanning 466 million lines of code in 20 hours, announced a UST partnership for physical AI, released "The Making of Claude Code" interactive feature, appointed Ben Bernanke to the Long-Term Benefit Trust, and a guided morning brief for Cowork was spotted in the works
🚨 AI News | TestingCatalog@testingcatalog · 博主 · 1 天前专挖 AI 产品未发布新功能的爆料号

Rokid 为其 AI 眼镜开设了代理商店!

> Rokid 推出的 AI 应用商店让智能眼镜用户能通过各种 AI 代理来获得新功能。

> 用户可以用语音、视觉或真实场景来提出请求,匹配的代理会运行,并在你的视野范围内返回结果。

完全解放双手 👀

查看英文原文
Rokid has opened an Agent Store for its AI glasses!

> Rokid’s AI-focused App Store for smart glasses allows users to access new capabilities by operating various AI agents.

> Users can make their requests by voice, sight, or a real-world scene, and the matching agent runs and returns the result within the wearer's field of view.

Hands-free 👀
向阳乔木@vista8 · 中文博主 · 1 天前向阳乔木,中文圈 AI 工具与趋势博主

Codex 生成的产品介绍视频,文案和介绍还挺清楚的。


rss.qiaomu.ai/

Lisan al Gaib@scaling01 · 博主 · 1 天前高频 AI 模型测评与爆料博主

能不能商量一下
@sama

你给我们 GPT-6,我就继续订 Pro 再几个月 :P

查看英文原文
we can trade
@sama


give us GPT-6 and I will stay on the Pro sub for another few months :P
Gorden Sun@Gorden_Sun · 中文博主 · 1 天前中文圈高频 AI 资讯与开源项目博主

Every团队评价Grok 4.5:快、便宜,终于好用了

Grok 4.5是SpaceX收购Cursor之后双方合作的首个成果,由Cursor与SpaceXAI联合训练,训练数据来自真实的代码库和软件工具交互。

在Every团队的内部评测中,大家的共识是这已经是一个Opus级别的模型。

在Mike Taylor的最新基准测试中,它甚至略微超过了Claude Opus 4.8:Grok完整执行了每一个步骤并交出打磨完善的结果,而Opus中途停止或跳过了部分任务。

Kieran Klaassen用Every的复合工程工作流跑了一遍后,把它定位在Claude Opus 4.5到4.6的区间,评价是“算不上最先进,但很多事情都能做得不错,而且非常快”。

Grok 4.5最大的竞争优势可能在于成本。xAI称它的运行速度约为每秒80个token,token效率约为领先模型的两倍。价格上,输入每百万token收费2美元、输出6美元,远低于Claude Opus 4.8的5美元和25美元,以及GPT-5.6 Sol的5美元和30美元。

在实际测试中,它在vibe coding和生成式交互方面表现稳健,做出了可用的语音访谈表单、社区地图应用,还成功克隆了Mike的写作风格,PPT制作水平约在Opus 4.6到4.7之间。不足之处在于地图应用不如最新前沿模型精致,写作方面Mike仍然更偏爱Sol。

总体结论:你大概不需要更换日常主力模型。但如果你本来就在Cursor里工作,Grok 4.5触手可及,在那些看重速度、价格和执行完整度的长链条多步骤任务上,它已经赢得了一席之地。

向阳乔木@vista8 · 中文博主 · 1 天前向阳乔木,中文圈 AI 工具与趋势博主

IBM 在 1979 年的一份著名培训幻灯片里就说:

“计算机永远不能被追究责任。因此,计算机绝不能做出管理决策”

AI 同理,人的价值是做 Skin in the game的决策。

小互@xiaohu · 中文博主 · 1 天前小互,中文圈高频 AI 资讯站 Xiaohu.AI 主理人

Google Cloud 为 AlloyDB 装上AI模型:能进行语义搜索数据库,而且最快提速 2.3 万倍

传统数据库有个天生的局限:它只会按字面精确匹配。你能轻松查「价格低于 100 的商品」「状态是已完成的订单」

却没法直接查「差评里抱怨电池的」「语气愤怒的工单」

因为这要理解文字的意思,而数据库不懂意思

以前想做这种「按意思」的分析,只能把数据导出去、送进一套单独的 AI 管线跑一遍、再把结果导回来,既慢又得自己维护那套管线。

这次 AlloyDB 做的事,是把 Gemini 的理解能力直接做成一批 SQL 函数(就叫 AI 函数),嵌进数据库本身。

于是你能在一句普通 SQL 里写下对「意思」的判断,数据库自己就懂,把符合的行挑出来,数据一步都不用搬出去。

数据库从「只会按字面精确匹配」,升级成「能按意思查询」,而且语义搜索、结构化查询、AI 判断全在同一层 SQL 里完成。

The Rundown AI@TheRundownAI · 博主 · 1 天前百万订阅 AI 日报官方

今天 AI 圈的热点新闻:

- Apple 对 OpenAI 硬件计划提起诉讼
- The Rundown Roundtable:我们的 AI 用例
- 用 ChatGPT Work + Codex 从想法到网站
- AI 2027 团队为人类拟定 A 计划
- 4 个新 AI 工具、社区工作流等

查看英文原文
Top stories in AI today:

- Apple takes OpenAI's hardware push to court
- The Rundown Roundtable: Our AI use cases
- Go from idea to website with ChatGPT Work + Codex
- AI 2027 team drafts humanity's Plan A
- 4 new AI tools, community workflows, and more
Luma@LumaLabsAI · 公司官方 · 1 天前AI 视频生成公司 Luma

一个人在读书,忙着自己的事。然后事情开始有点离谱。背后的 Luma Skill 是 Eli Coleman 做的。用 Luma 制作。

查看英文原文
A man reading, minding his own business. Then things start to get a little out of proportion. The Luma Skill behind it, by Eli Coleman. Made with Luma.
Rowan Cheung@rowancheung · 博主 · 1 天前AI 日报 The Rundown 创始人

研究人员用高分辨率X光扫描,再让机器学习识别纤维中隐藏的古老墨迹。

结果真的行。其中一卷古卷居然是一本关于斯多葛哲学的著作。

现在大家都在争分夺秒,希望能再破解几百卷。

查看英文原文
Researchers used high-res X-ray scans, then machine learning detected faint traces of ancient ink hidden in the fibres.

It worked. One scroll turned out to be a work on stoic philosophy.

Now the race is on to uncover hundreds more.
AshutoshShrivastava@ai_for_success · 博主 · 1 天前高频 AI 新闻与产品动态博主

视角:2026年还在手写代码的时候。

查看英文原文
POV: Writing code by hand in 2026.
The Rundown AI@TheRundownAI · 博主 · 23 小时前百万订阅 AI 日报官方
连环推 ×2

16位诺贝尔奖得主和200多位经济学家及AI研究者签署了由斯坦福数字经济实验室发起的声明《我们必须立即行动》。

警告内容:AI可能带来"比工业革命更剧烈的经济变革,但发生在极短的时间框架内"

安东尼·科里克(目前在Anthropic任职的弗吉尼亚大学经济学家):

"蒸汽、电力和计算机都给社会数十年时间适应;AI可能只给我们几年。我们不能在变革中临时拼凑策略和制度;等待确定性就意味着为时已晚。"

查看英文原文
16 Nobel Prize winners and over 200 economists and AI researchers signed "We Must Act Now", a statement organized by Stanford's Digital Economy Lab.

The warning: AI could bring economic change "larger than the Industrial Revolution, but unfolding over a vastly shorter time frame."

Anton Korinek, UVA economist currently on leave at Anthropic:

"Steam, electricity, and computers each gave societies decades to adapt; AI may give us only a few years. We cannot improvise our strategy and institutions in the middle of the transformation; waiting for certainty means arriving too late."
Read the full statement and list of signatories:


wemustactnow.ai/
The Rundown AI@TheRundownAI · 博主 · 1 天前百万订阅 AI 日报官方

今日机器人圈顶级新闻:

- NEO的双手能拉拉链、倒茶、打手语
- 这个鸟形机器人在水下游泳,还能腾空而起
- Tesla在迈阿密取消了安全网
- 为什么只会漂浮的机器人也有存在的价值
- 其他机器人新闻快讯

查看英文原文
Top stories in robotics today:

- NEO’s hands can zip jackets, pour tea, and sign
- This bird bot swims underwater, bursts into flight
- Tesla skips the safety net in Miami
- The case for robots that do nothing but float
- Quick hits on other robotics news
Gorden Sun@Gorden_Sun · 中文博主 · 1 天前中文圈高频 AI 资讯与开源项目博主

等等冲,Grok Build被曝会偷偷上传全量代码库

引用 Gorden Sun @Gorden_SunGrok印度区订阅,6500卢布一年,折合人民币575,每月仅需48块钱。 一个账号不够用就搞两个,加起来也比御三家便宜。 Grok只有周限额,没有5小时限额,爽用。查看被引原帖 ↗
向阳乔木@vista8 · 中文博主 · 1 天前向阳乔木,中文圈 AI 工具与趋势博主

效果真的相当一般,感觉不如Listenhub 的cli+Remotion自己搭一套。
@oran_ge

AshutoshShrivastava@ai_for_success · 博主 · 1 天前高频 AI 新闻与产品动态博主

Fable 5 的安全过滤真的离谱。

我想写个脚本,通过 VNC 在 MacBook Pro 上启动我 Mac mini 的屏幕共享。还想看看能不能通过自动传递参数来跳过高性能选择弹窗。

搞笑的是,Fable 5 的思维trace显示它在考虑自己传递这些参数。最后也没成功。

反而把我的请求标记为可疑,把我的模型切到了 Opus 4.8。

查看英文原文
Fable 5 safety filters are wild.

I was just trying to write a script to launch Screen Sharing for my Mac mini on my MacBook Pro via VNC. I also wanted to see if the High Performance selection popup could be skipped by passing those values automatically.

The funny part was that Fable 5's thinking trace showed it was considering passing those values itself. In the end, it didn't even work anyway.

Instead, it flagged the request as suspicious and switched my model to Opus 4.8.
Gorden Sun@Gorden_Sun · 中文博主 · 1 天前中文圈高频 AI 资讯与开源项目博主

还是微软CEO会讲故事。

你在用AI的时候,实际付了2次费用,一次是花的钱,一次是你透露给AI的企业专业知识。你的提示词、智能体调用的工具、模型出错时你的纠正,这些都是你给AI的付出。
你在消费AI的同时,也在创造AI。

然后微软CEO说:你创造的东西理应属于你。

这样微软才好开展企业业务,给每一家企业搞AI学习基建,企业还能获得控制、隐私巴拉巴拉的好处。

小互@xiaohu · 中文博主 · 1 天前小互,中文圈高频 AI 资讯站 Xiaohu.AI 主理人

如何保住公司的经验知识和老兵能力?

纳德拉把做法归纳成五个以 C 开头的词。

它们不是五项平行功能,而是一条从“拥有资产”走到“形成复利”的顺序

详细内容:
best.xiaohu.ai/article/satya…

🚨 AI News | TestingCatalog@testingcatalog · 博主 · 1 天前专挖 AI 产品未发布新功能的爆料号

用户可以在多个日常场景中使用这些代理,从实时翻译和导航,到会议总结、购物识别,甚至识别宠物声音背后的情绪。

Rokid 为我们的读者提供了折扣码:R3K-TESTI

更多来自 @rokid_japan 👀

#RokidGlasses
#AIグラス
#ARグラス
jp.rokid.com/discount/R3K-TE…

查看英文原文
Users can use the agents across multiple everyday scenarios, from real-time translation and navigation to meeting summaries, shopping recognition, and even reading the general mood behind a pet's sounds.

Rokid has provided our readers with a discount code: R3K-TESTI

More from @rokid_japan 👀

#RokidGlasses
#AIグラス
#ARグラス
jp.rokid.com/discount/R3K-TE…
Cristóbal Valenzuela@c_valenzuelab · 创始人 · 1 天前Runway 联合创始人兼 CEO

巴西,我们来了!🇧🇷 我们在拉丁美洲的版图在扩大。马上还有更多惊喜更新。如果你在圣保罗,一定别错过Runway的首场线下聚会。

引用 Alejandro Matamala Ortiz @matamalaortiz[巴西] 我将在圣保罗参加Runway社区聚会。我们有很多内容要分享,欢迎加入!查看被引原帖 ↗
查看英文原文
Brazil, here we come! 🇧🇷 Our LATAM presence is growing. We will have some more exciting updates on this front soon. If you are in Sao Paulo, join us for the first official Runway Meetup there.
🚨 AI News | TestingCatalog@testingcatalog · 博主 · 1 天前专挖 AI 产品未发布新功能的爆料号

用户可以在频道或任务中标记 @Brain 来使用收件箱工具,这样你可以在工作发生的地方就进行分类。它可以管理你的收件箱,对通知进行分类,处理低价值的项目,并留下一份实际需要关注的简短列表。

在下面测试 ClickUp 👇

clickup.com

查看英文原文
Users can use the inbox tool by tagging @ -Brain in a channel or task, so triage runs wherever the work is. It can be asked to manage the inbox, categorizing notifications, acting on low-value ones, and leaving a short list of what actually needs attention.

Test ClickUp below 👇

clickup.com
Gorden Sun@Gorden_Sun · 中文博主 · 1 天前中文圈高频 AI 资讯与开源项目博主

AI圈每日简报,7月12日翻:
gorden-sun.notion.site/7-12-…

查看英文原文
AI资讯日报,7月12日:
gorden-sun.notion.site/7-12-…
Bindu Reddy@bindureddy · 创始人 · 1 天前Abacus.AI CEO,AI 行业观点博主

说实话,AGI 的真正前沿是机器人学

除非机器人能做真实工作,否则我们现在做的一切都只是自动化白领工作

本质上讲,这不过是把推纸工作换了个花样

查看英文原文
TBH the real AGI frontier is robotics

Until robots can do real work all we are doing is automating white collar work

At some level - that’s just a glorified version of pushing paper
Gorden Sun@Gorden_Sun · 中文博主 · 1 天前中文圈高频 AI 资讯与开源项目博主

NVIDIA-Nemotron-Labs-3-Puzzle:高性能压缩MoE

基于Nemotron-3-Super-120B-A12B压缩,压缩后总参数75B,激活参数9B,在单台8×B200节点上实现吞吐量翻倍,单张H100显卡1M上下文的情况下并发请求从1个提升到8个。性能基本不变。

模型:
huggingface.co/nvidia/NVIDIA…

论文:
arxiv.org/abs/2607.04371

Gorden Sun@Gorden_Sun · 中文博主 · 1 天前中文圈高频 AI 资讯与开源项目博主

AI news daily, July 13:

gorden-sun.notion.site/7-13-…

查看英文原文
AI资讯日报,7月13日:
gorden-sun.notion.site/7-13-…
Krea@krea_ai · 公司官方 · 1 天前AI 创意生成工具 Krea 官方

转发 @ingi_erlingsson: nature vs nurture 🌺

在 @ComfyUI 中用 @krea_ai 2 + @Alibaba_Wan 2.2 I2V/T2V/Vace + @thesystms HUD + @suno 🎹 制作的 https…

查看英文原文
RT
@ingi_erlingsson
: nature vs nurture 🌺

made in
@ComfyUI
with
@krea_ai
2 +
@Alibaba_Wan
2.2 I2V/T2V/Vace +
@thesystms
HUD +
@suno
🎹 https…
Hailuo AI-MiniMax Hub@Hailuo_AI · 公司官方 · 1 天前MiniMax 旗下海螺 AI 视频官方

快来搭你自己的创意工具箱!
hub.minimax.io/

查看英文原文
Try to build your own creative toolbar!

hub.minimax.io/

本站由 Jedee杰哥 打造 · 公众号「Jedee杰哥」每早送 AI 日报

姊妹站:𝕏 简中账号数据榜单 · X 关注 @jedeeai · RSS 订阅 · AI 日报 · 历史归档