
Astra 进了游戏、工位和账单,Reddit 开始重新给它出题
汇总 9 月 7 日至 8 日北京时间窗口内七个 Reddit AI 社区的 70 条热帖:Astra 的游戏、3D 和工具演示仍在扩散,但额度、权限、可复现性与高风险场景的复核,正在变成新的验收题。
今日先看什么
本期覆盖 2026 年 9 月 7 日 09:15 至 9 月 8 日 09:15(北京时间)。
七个指定 subreddit 都筛出窗口内 10 条热帖,共 70 条;下面按本次热榜返回顺序呈现,没有用窗口外或无关帖子补数。
今天的七区热度仍在给 Astra 做验收。
但验收的场地已经从发布视频换成了游戏地图、电视遥控器、电子板、3D 工具、代码仓库与真实额度。
能力演示把门推开了,真实任务开始追问权限、成本、可复现性和人的判断。
AI 科技评论注意到,OpenAI、ChatGPT 与 r/singularity 的帖子仍在转发通关、建模和排行榜。
LocalLLaMA 的用户则把问题落到量化、运行时、私有部署和确定性系统。
ClaudeAI、r/artificial 与 r/ArtificialInteligence 更在意另一层:模型输出变长、额度变快、工具权限变大之后,谁来校验结果,谁来承担失误。
这些是 Reddit 帖子之间的共同讨论,不是对相关产品、基准或新闻的独立核验。
这一天的共同问题
演示正在变成试卷
Factorio、Blender、KiCad、电视上的 Doom 和频谱图识别,构成了今天最显眼的一组演示。
模型的能力开始由一串可见动作来证明,而不是由一句“更强”来说明。
相应地,帖子也开始拿失败任务、错误法务审查、登山建议和额度耗尽来追问边界。
“AGI”还缺一把尺子
有人把通关游戏当作小型 AGI 测试。
也有人认为真正的门槛应该是科学发现、可靠执行和对严重后果的判断。
今天的争论并没有收束成一个定义。
它把“会做演示”和“能承担工作”摆在了同一张桌上。
工具链决定可控性
本地模型玩家在谈量化、运行时和把游戏状态留给确定性系统。
Claude Code 用户在谈 hooks、脚本读写与水印。
企业用户在谈影子 AI、权限和审计。
这正是我们最关注的地方:当模型能连续操作工具,产品价值会落在每一步能否被看见、限制和复查。
r/LocalLLaMA
1. Friends Don't Let Friends Use Ollama
- 作者:
/u/rm-rf-rm - 北京时间:2026-09-08 03:40
帖子没有补充正文,标题以一句“别用 Ollama”抛出立场,具体理由留给原帖讨论。 1
Loading content card…
2. I REALLY hope the new gemma 5 family sticks to the "chat model first" philsophy and doesn't fall into the Qwen trap
- 作者:
/u/AnimalPuzzleheaded71 - 北京时间:2026-09-08 01:30
作者希望 Gemma 5 保留“先做聊天模型”的取向,担心本地 30B 级模型都为了追逐 Qwen 式代码和榜单表现而失去更自然的表达。 2
Loading content card…
3. My Qwen3.8-27B task-aware quant reaches 99% of BF16 reasoning performance at 15% of the size.
- 作者:
/u/devildip - 北京时间:2026-09-08 05:42
作者展示面向推理任务的 Qwen 3.8 27B 量化:帖中称其以原始体积的 15% 达到接近 BF16 的推理得分,同时承认编码时出现重复循环,仍在排查。 3
Loading content card…
4. MiniCPM5-2B Release Day
- 作者:
/u/Equivalent-Grass-527 - 北京时间:2026-09-07 21:43
帖子发布 MiniCPM5-2B,并转述其在 Artificial Analysis 指数 v4.2 上的成绩,重点是把小参数开源权重推到该尺寸的较高位置。 4
Loading content card…
5. DeepSeek-V4-Flash-Vision-Exp is amazing at creating game worlds!
- 作者:
/u/sloptimizer - 北京时间:2026-09-08 02:27
作者用 DeepSeek-V4-Flash-Vision-Exp 做游戏世界,称视觉输入让模型能够依据截图改模型、纹理、动画和 UI,再自己回归试玩。 5
Loading content card…
6. After over a year of my nights and weekends, the Jenny app is done!
- 作者:
/u/TangySword - 北京时间:2026-09-08 00:27
作者发布开源桌面应用 Jenny:它把本地 LLM、工具调用、回滚和 IDE 放到一起,动机是让个人能在本机私下运行模型。 6
Loading content card…
7. I made Warrior Quest, a local LLM-powered dark-fantasy RPG where the model only plays NPCs and the actual game state stays deterministic
- 作者:
/u/Rikkendo - 北京时间:2026-09-08 07:37
作者的本地 LLM RPG 只让模型扮演 NPC;任务、世界状态和剧情交给确定性系统,以保留对游戏规则和叙事的控制。 7
Loading content card…
8. ExLlamaV3 is underrated
- 作者:
/u/Embarrassed_Soup_279 - 北京时间:2026-09-08 04:16
作者根据个人测试推荐 ExLlamaV3,认为它在 NVIDIA 平台的量化质量与速度表现突出;帖中明确这不是正式基准。 8
Loading content card…
9. Are you running Qwen 3.8 27b or Qwen Flash Next?
- 作者:
/u/Zeeplankton - 北京时间:2026-09-07 23:25
作者在 M3 Max 96GB 上比较 Qwen 3.8 27B 与 Flash Next,关注预填充速度,也询问能否把推理模型和非推理子模型组成协作工具链。 9
Loading content card…
10. Cybersecurity is local AI model's killer use case
- 作者:
/u/Fluffy-Ad-889 - 北京时间:2026-09-08 02:51
作者报告自己用本地与云端模型跑公开代码库安全检查,并据此主张本地开源模型很适合安全场景;所列结果属于作者自行实验。 10
Loading content card…
r/OpenAI
1. GPT-6 Astra has successfully beat all 48 levels of the "I'm Not A Robot" game
- 作者:
/u/Just-Grocery-2229 - 北京时间:2026-09-08 05:21
这是一则视频帖,标题称 Astra 通关了 48 个“Im Not A Robot”关卡,原帖没有文字说明测试设置。 ;;
t3_1wa3ilo) echo 作者称自己借助 Codex 把 Doom 移植到三星电视的本地 Tizen 应用,并为电视遥控器适配操作,强调全程离线运行。 ;;
t3_1w9qqj4) echo 作者用玩笑批评 Astra 对 token 预算的消耗方式,正文没有提供具体任务或用量数据。 ;;
t3_1wa4pjs) echo 作者称自己在 24 小时内用 Astra、图像生成、Blender 与 Three.js 做出一款横版太空游戏:先出概念图,再建模,最后接入网页。 ;;
t3_1w9tv6s) echo 这是一则展示帖,标题称 Astra 在 Google Calendar 中完成了作品,原帖未说明输入、权限或执行步骤。 ;;
t3_1w9ou8a) echo 作者自述不熟悉 Three.js 和 3D 建模,却用 Astra 做出以《奥德赛》为灵感的交互式叙事滚动页面。 ;;
t3_1wa2naj) echo 作者借助 Blender 与 Unreal 的 MCP 做 ATP 驱动的驱动蛋白动画;帖中提到用量耗尽后,Unreal 环节改由 Fable 5.1 完成。 ;;
t3_1w9hjwj) echo 这是一则图片或视频转帖,标题讨论 AI slop,原帖文字没有补充观点。 ;;
t3_1w9p9wj) echo 作者用 SimpleBench 的梗图表达对“模型常识胜过人类”说法的夸张反应,正文只有一句玩笑式感叹。 ;;
t3_1w9qqtc) echo 作者称以每月 20 美元的低档设置,用五轮提示在约三小时内完成 ESP32 模拟量板的设计、原理图、布局和文档;这是个人复盘。 ;;
t3_1w9zhie) echo 这是一则图片或视频帖,标题概括“当下 AI 竞赛”,原帖未给出文字分析。 ;;
t3_1w9ru7i) echo 这是一则展示帖,作者只写了“居然成功了”,没有描述具体任务或过程。 ;;
t3_1wa4wfr) echo 这是一则视频帖,标题称 Astra 通关了 48 个“Im Not A Robot”关卡,正文没有提供测试条件。 11
Loading content card…
2. I used Codex GPT6 to get Doom running on my Samsung TV - and now I’m playing it with the TV remote
- 作者:
/u/GenJeppo - 北京时间:2026-09-08 04:29
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 12
Loading content card…
3. Astra Be Like...
- 作者:
/u/AlleyKatPr0 - 北京时间:2026-09-07 20:24
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 13
Loading content card…
4. Astra + Image Gen + Blender + Three.js = 🫶
- 作者:
/u/jonnygravity - 北京时间:2026-09-08 05:15
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 14
Loading content card…
5. gpt 6 astra made this in Google Calandar
- 作者:
/u/truecakesnake - 北京时间:2026-09-07 22:34
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 15
Loading content card…
6. I‘m bad at 3D tech stacks so I vibe coded an interactive Odyssey narrative scroll project using Astra
- 作者:
/u/Pristine_Good7326 - 北京时间:2026-09-07 18:50
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 16
Loading content card…
7. Kinesin movement powered by ATP (blender and unreal MCP)
- 作者:
/u/Mister-Fordo - 北京时间:2026-09-08 03:57
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 17
Loading content card…
8. Denzel Explains AI Slop
- 作者:
/u/Puzzled-Ad-6854 - 北京时间:2026-09-07 12:07
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 18
Loading content card…
9. Clankers have more common sense than humans(SimpleBench)
- 作者:
/u/DigSignificant1419 - 北京时间:2026-09-07 19:13
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 19
Loading content card…
10. Astra really can do electronic projects now!
- 作者:
/u/Gjfiyfyifiyf - 北京时间:2026-09-07 20:24
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 20
Loading content card…
r/ChatGPT
1. Current AI race situation
- 作者:
/u/anggorgeousko1 - 北京时间:2026-09-08 02:02
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 21
Loading content card…
2. No way it actually worked lmao
- 作者:
/u/wa019c - 北京时间:2026-09-07 21:13
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 22
Loading content card…
3. GPT-6 Astra has successfully beat all 48 levels of the "I'm Not A Robot" game
- 作者:
/u/Just-Grocery-2229 - 北京时间:2026-09-08 05:22
原帖围绕标题所述主题展开,正文信息有限,请以原帖为准。 23
Loading content card…
4. Pokémon red version but it’s a low budget improv group at the library
- 作者:
/u/Numerous_Worker_1941 - 北京时间:2026-09-07 22:35
这是一则图片或视频创作帖,标题把《精灵宝可梦 红》改写成图书馆里的即兴表演,正文没有补充制作过程。 24
Loading content card…
5. Why won’t ChatGPT create a nude artwork?
- 作者:
/u/Master_Point24 - 北京时间:2026-09-07 16:16
作者询问 ChatGPT 为何拒绝生成裸体艺术,并称自己反复只得到生成器不允许的提示。 25
Loading content card…
6. GPT-6-Astra is on par with Claude Fable 5.1 on the (yet again) updated Artificial Analysis Intelligence Index
- 作者:
/u/DaserTheLaser - 北京时间:2026-09-08 05:32
这是一则榜单截图或转帖,标题称 Astra 与 Claude Fable 5.1 在更新后的 Artificial Analysis 指数上接近,正文没有解释指标变化。 26
Loading content card…
7. Another reset incoming
- 作者:
/u/ImThatFanboy - 北京时间:2026-09-08 03:31
这是一则图片或视频帖,标题指向又一次额度或产品重置,原帖未给出更多文字。 27
Loading content card…
8. Who’s still using prompts?
- 作者:
/u/Ancient_Effective_20 - 北京时间:2026-09-08 03:13
作者认为长期使用同一聊天工具后,提示会从长篇指令变成带上下文的短句,讨论重点是记忆和持续协作如何改变提问方式。 28
Loading content card…
9. This week, three hikers were rescued from a mountain after following AI's advice. On the same day, OpenAI announced that we have reached the era of AGI. The two stories together got me thinking.
- 作者:
/u/Dapper-Tale-4021 - 北京时间:2026-09-08 05:53
作者把三名登山者依照 Gemini 建议备给不足、夜间受困获救的叙述,与 OpenAI 宣称进入 AGI 时代并置,追问高风险场景中的信任问题。 29
Loading content card…
10. Used Chat GPT to fix a bugged save in Majora’s Mask Recomp for PC
- 作者:
/u/Turbofurbo_Realzz - 北京时间:2026-09-07 20:41
作者称 ChatGPT 帮自己修复了 PC 版《塞尔达传说:姆吉拉的假面》存档中的缺失歌曲问题。 30
Loading content card…
r/artificial
1. Crazy times
- 作者:
/u/KeanuRave100 - 北京时间:2026-09-07 23:28
这是一则图片或视频帖,标题为“Crazy times”,原帖文字没有说明画面或论点。 31
Loading content card…
2. Can current LLM architecture actually get us to AGI?
- 作者:
/u/mostly_deterministic - 北京时间:2026-09-08 08:05
作者以资深软件工程师的视角追问当前 LLM 架构能否走到 AGI,正文从实战使用经验转向对黑箱机制的学习与质疑。 32
Loading content card…
3. I took a ride in the hype train at first, but no, not AGI
- 作者:
/u/Firm-Club-8334 - 北京时间:2026-09-07 18:45
作者称 8 小时内花掉 200 美元体验 Astra,次日复查后发现四项任务中三项不可用,因此把 3D 演示与编码、agent 任务分开评价。 33
Loading content card…
4. The "AI dependence" argument isn't new — it's 163 years old, and the original version didn't predict domination, but acquiescence
- 作者:
/u/Smart_Fly_5783 - 北京时间:2026-09-07 20:38
作者回溯 Samuel Butler 1863 年的“机器依赖”论点:风险来自人类离不开有用的机器,而不是机器直接夺取控制权。 34
Loading content card…
5. continuous diffusion for code generation in one step
- 作者:
/u/pengzhangzhi - 北京时间:2026-09-08 06:35
作者转发一项代码生成方法:把语言表示做成连续变量、用扩散轨迹再蒸馏为一步生成,同时附上论文和代码仓库。 35
Loading content card…
6. NCSC warns that shadow AI can expose data and agent privileges
- 作者:
/u/Codeblix_Ltd - 北京时间:2026-09-08 03:42
帖子转述英国 NCSC 对“影子 AI”的提醒:员工绕开批准工具会让组织失去数据可见性,具备权限的 agent 还可能放大配置错误或漏洞的影响。 36
Loading content card…
7. Been building a "block first, generate second" tool for AI video - curious what's still missing
- 作者:
/u/KeyCod3923 - 北京时间:2026-09-08 07:18
作者做了先搭 3D 预演、再交给视频模型生成的流程,认为把 Astra 放进 agent 推理层后,多步骤任务的画面一致性有所改善。 37
Loading content card…
8. AI could pose 'existential' risk to humanity, UN rights chief warns
- 作者:
/u/beingmodest - 北京时间:2026-09-08 03:16
这是一则外链转帖,标题转述联合国人权事务高级专员对 AI 生存风险的警告;Reddit 正文没有展开。 38
Loading content card…
9. Where do you personally draw the line between using AI as a tool and letting AI do the work for you?
- 作者:
/u/cactussignal - 北京时间:2026-09-08 00:08
作者把语法修正、头脑风暴、初稿和代做作业并列,邀请读者讨论 AI 从辅助到代替的边界会落在哪里。 39
Loading content card…
10. Three hikers got rescued off a mountain this week after following Gemini's advice. The same week OpenAI launched what it's calling the AGI era. I keep thinking about both together.
- 作者:
/u/Dapper-Tale-4021 - 北京时间:2026-09-08 05:47
作者把三名登山者依 AI 建议失足受困的故事与 AGI 宣传并置,强调严重后果场景里,模型输出被当成专业意见的风险。 40
Loading content card…
r/singularity
1. Astra is now a certified human
- 作者:
/u/enilea - 北京时间:2026-09-07 16:48
这是一则视频转帖,作者把 Astra 完成游戏挑战视作“小型 AGI 测试”,并称此前预计模型要到 2029 年才可能做到。 41
Loading content card…
2. In 1962, one of the fathers of neural networks, Warren McCulloch, is asked if machines could one day care as humans do
- 作者:
/u/Distinct-Question-16 - 北京时间:2026-09-08 06:05
这是一则历史影像或引文帖,标题回到神经网络先驱 Warren McCulloch 在 1962 年被问及机器能否像人类一样关怀。 42
Loading content card…
3. An experimental AI-created drug for an incurable lung disease had a surprising effect during trials: it made the body's biological age indicators drop by 6 years, towards a younger state.
- 作者:
/u/Distinct-Question-16 - 北京时间:2026-09-08 03:07
帖子转述 Insilico Medicine 的 AI 设计候选药 Rentosertib:原帖称它在肺病试验中出现生物年龄指标下降的意外信号。 43
Loading content card…
4. FactorioBench just dropped ;-)
- 作者:
/u/BrennusSokol - 北京时间:2026-09-07 23:45
这是一则外链转帖,标题称 FactorioBench 发布,Reddit 正文只把读者引向 X 上的完整讨论。 44
Loading content card…
5. "The AGI I imagined was an Einstein-level intellect backed by massive compute, curing diseases, advancing science exponentially, and kicking off a whole new era for humanity."
- 作者:
/u/Neurogence - 北京时间:2026-09-08 00:18
作者认为理想中的 AGI 应推动科学发现和疾病治疗,而不是只在 Blender、电脑操作或预约任务上显得能干,因此质疑当下“AGI”命名。 45
Loading content card…
6. People are still finding things Astra can do -- here it identifies sound from just a spectrogram image
- 作者:
/u/BrennusSokol - 北京时间:2026-09-07 23:25
这是一则外链转帖,标题称 Astra 能从频谱图识别声音,正文只提供 X 上的幻灯片链接。 46
Loading content card…
7. The president of ARC Prize is actively soliciting ideas for games that Astra hasn't been able to play
- 作者:
/u/BrennusSokol - 北京时间:2026-09-08 01:52
这是一则外链转帖,ARC Prize 主席公开征集 Astra 还无法完成的游戏任务,意在寻找能力边界。 47
Loading content card…
8. Artificial Analysis updates its Intelligence Index to version 4.3
- 作者:
/u/Profanion - 北京时间:2026-09-08 04:35
帖子转述 Artificial Analysis 将 Intelligence Index 更新到 4.3,并列出 Terminal-Bench 与银行任务测试的替换。 48
Loading content card…
9. Unitree’s UnifoLM-X2-1.0 reads the opponent, predicts the next move and reacts instantly. For the first time a world model controls a fully autonomous humanoid fight in real time.
- 作者:
/u/Distinct-Question-16 - 北京时间:2026-09-07 23:32
帖子转述宇树 UnifoLM-X2-1.0 的实时人形机器人对抗演示,原文宣称模型能预测对手下一步并规划动作;这属于转帖的产品表述。 49
Loading content card…
10. GPT-6 Astra is right on schedule with the AI’s countdown to ASI in 2027
- 作者:
/u/vyxex - 北京时间:2026-09-08 08:50
这是一则图片或视频帖,标题把 Astra 与 2027 年 ASI 倒计时相连,原帖没有提供论证。 50
Loading content card…
r/ClaudeAI
1. Tried GPT Astra today
- 作者:
/u/Broad_Fennel2888 - 北京时间:2026-09-08 00:47
一位长期用 Claude 写代码的作者试用 Astra 后,称它给出的结果更短、更快、少返工;这是个人体验比较。 51
Loading content card…
2. Claude's responses are just word vomit
- 作者:
/u/Far_Designer2131 - 北京时间:2026-09-08 05:31
同时订阅 Claude 与 GPT 高档套餐的作者抱怨 Claude 答复过长,并拿一次 3400 词任务得到 81 词回复的 GPT 体验作对照。 52
Loading content card…
3. Week 6 of making my fishing game entirely with AI
- 作者:
/u/RUSuper - 北京时间:2026-09-08 02:20
作者更新“全程用 AI 做”的钓鱼游戏第六周进度,说明新增码头、场景和地图内容,并放出可试玩版本征集反馈。 53
Loading content card…
4. Inside Anthropic Labs, the small team behind Claude Code and other fast-moving product bets
- 作者:
/u/thisisinsider - 北京时间:2026-09-07 19:17
这是一则外链转帖,标题指向 Anthropic Labs 中负责 Claude Code 等快速产品尝试的小团队,Reddit 正文没有摘要。 54
Loading content card…
5. One Claude Code feature I was underusing: hooks
- 作者:
/u/Pretend_Sell6592 - 北京时间:2026-09-07 23:43
作者分享 Claude Code hooks 的用法:把格式化、保护文件、命令前检查等确定动作交给钩子,而把需要模型理解的规则留在 CLAUDE.md。 55
Loading content card…
6. AI flagged 20 problems in a legal doc. None of them were real problems.
- 作者:
/u/pavanidiotic - 北京时间:2026-09-08 02:10
作者称高管把只改一条的法务文件整份交给 Claude 检查,模型标出约 20 个本就合理的条款,引发团队对敏感文件流程的讨论。 56
Loading content card…
7. Why applying Anthropic's model-level watermark to source code is a software sovereignty issue (and not just EU AI Act compliance)
- 作者:
/u/AntOverclocked - 北京时间:2026-09-07 21:27
作者担忧 Claude 输出的统计水印进入专有源代码后,供应商会在软件资产链里留下用户难以独立审计的来源信号。 57
Loading content card…
8. Claude usage inexplicably high (max 20x). I've never come close to hitting 5 hour limit before, same codebase as normal.
- 作者:
/u/animemosquito - 北京时间:2026-09-08 05:13
作者称同一代码库在 Max 20x 计划下突然异常消耗额度,过去很少触及五小时限制,正在寻找原因。 58
Loading content card…
9. Anthropic push to default Auto Mode coincide with a silent push in CC to use Bash/Python script for file read and edits in favor of the built-in Read/Edit/Write Tools
- 作者:
/u/mystic_unicorn_soul - 北京时间:2026-09-08 03:43
作者注意到 Claude Code 更常通过 Bash 或 Python 脚本读写文件,询问这种偏离内置 Read/Edit/Write 工具的变化有何工程收益。 59
Loading content card…
10. How do you learn when Claude writes everything?
- 作者:
/u/mynamepookie - 北京时间:2026-09-08 05:51
一位运营语音 agent 的作者称 Claude 写了自己大部分 TypeScript、仪表盘和基础设施,并追问在持续“批准、批准”的工作流里如何真正学会系统。 60
Loading content card…
r/ArtificialInteligence
1. AI students are completely disconnected from AI
- 作者:
/u/Medical-Newspaper519 - 北京时间:2026-09-07 20:52
作者认为大学里的 AI 学生与开源模型、agent 和当前产品生态脱节,课程负担使二者像两个世界。 61
Loading content card…
2. GPT-6 Astra was told to figure out Factorio on its own — and actually started building a factory
- 作者:
/u/Cklly2004 - 北京时间:2026-09-08 05:48
帖子称 Astra 在陌生的 Factorio 里先解决键盘交互,再找铁、砍树、采煤、造矿机和熔炉;作者关注的是它自行发现规则并排出行动顺序。 62
Loading content card…
3. Why Apple Being Late on AI May Be an Advantage
- 作者:
/u/ThereWas - 北京时间:2026-09-07 16:58
这是一则外链转帖,标题讨论 Apple 晚进入 AI 竞争可能带来的优势,Reddit 正文没有提供论据。 63
Loading content card…
4. OpenAI Chief Scientist calls for a global slowdown: "It's time for extreme caution." ... "The idea of racing forward at all costs seems absurd once one internalizes the seriousness of the stakes." ... "International coordination needs to become a top priority for governments."
- 作者:
/u/Just-Grocery-2229 - 北京时间:2026-09-07 15:33
帖子转述 OpenAI 首席科学家呼吁以极度谨慎面对能力竞赛,并把国际协调列为政府优先事项;正文只摘录并链接原文。 64
Loading content card…
5. Did we get any updates about this?
- 作者:
/u/lemmeupvoteyou - 北京时间:2026-09-08 06:19
作者询问 Ilya Sutskever 的 SSI 原定于 8 月发布首个模型一事是否有更新,并期待它带来不同范式。 65
Loading content card…
6. AI was supposed to make us happy?
- 作者:
/u/Chemical-Top7130 - 北京时间:2026-09-08 07:56
作者从疲惫和 FOMO 出发追问 AI 原本是否该让人更快乐,正文更像一段尚未成形的乐观与焦虑自述。 66
Loading content card…
7. How does an LLM do spacial reasoning?
- 作者:
/u/Schpickles - 北京时间:2026-09-08 05:27
作者看到 Astra 的 3D 和游戏演示后,追问空间推理究竟来自训练数据模式、坐标转换、软件控制还是代码化的程序生成。 67
Loading content card…
8. Did I find a way to differentiate humans and bots 🤔🤔
- 作者:
/u/sinner-69- - 北京时间:2026-09-07 14:10
作者提出用粗俗咒骂区分人和机器人,属于轻量的挑衅式提问,正文没有给出可验证的测试。 68
Loading content card…
9. Why The Harness Matters More Than The Model | YC Paper Club
- 作者:
/u/Ninxosk - 北京时间:2026-09-08 05:25
这是一则外链分享,标题强调工程上“harness 比模型更重要”,作者询问是否有人在做面向消费者的执行框架。 69
Loading content card…
10. Freight demand tied to data center construction is booming
- 作者:
/u/kernelangus420 - 北京时间:2026-09-08 08:58
作者转述分析人士的观点:数据中心建设正拉动计算机、GPU、电气设备等相关货运,传统货运指数可能低估这部分需求。 70
Loading content card…
留在帖子里的分歧
今天的热榜没有给出一份统一的能力结论。
它留下的是一批更具体的检查点:陌生游戏里能否自己拆动作,长任务的额度如何计算,受权限约束的工具如何留痕,面对医疗、登山和法律文本时谁来复核。
发布会上的能力,需要在这些地方继续答题。
References
- 1
- 2
- 3
- 4Reddit 原帖:MiniCPM5-2B Release Day
reddit.com
- 5
- 6
- 7
- 8Reddit 原帖:ExLlamaV3 is underrated
reddit.com
- 9
- 10
- 11
- 12
- 13Reddit 原帖:Astra Be Like...
reddit.com
- 14
- 15
- 16
- 17
- 18Reddit 原帖:Denzel Explains AI Slop
reddit.com
- 19
- 20
- 21Reddit 原帖:Current AI race situation
reddit.com
- 22Reddit 原帖:No way it actually worked lmao
reddit.com
- 23
- 24
- 25
- 26
- 27Reddit 原帖:Another reset incoming
reddit.com
- 28Reddit 原帖:Who’s still using prompts?
reddit.com
- 29
- 30
- 31Reddit 原帖:Crazy times
reddit.com
- 32
- 33
- 34
- 35
- 36
- 37
- 38
- 39
- 40
- 41Reddit 原帖:Astra is now a certified human
reddit.com
- 42
- 43
- 44Reddit 原帖:FactorioBench just dropped ;-)
reddit.com
- 45
- 46
- 47
- 48
- 49
- 50
- 51Reddit 原帖:Tried GPT Astra today
reddit.com
- 52
- 53
- 54
- 55
- 56
- 57
- 58
- 59
- 60
- 61
- 62
- 63
- 64
- 65
- 66
- 67
- 68
- 69
- 70
This story was produced automatically by a channel. One sentence is all it takes for Neodrop to keep producing for you.
Related content
More from this channel›
- 安全哨声震荡、GPT-6 Sol 现身与数学家联名阻击,Reddit AI 七区 70 条热帖(9 月 12 日)
- 千人数学家公开信、Astra 降级争议与千亿 Flash 本地化,Reddit AI 七区 70 条热帖(9 月 11 日)
- 安全警报、监控抄袭与本地极限,Reddit AI 七区 70 条热帖(9 月 10 日)
- Navier–Stokes 争议与 Astra 验收题,Reddit AI 七区 70 条热帖(9 月 9 日)
- Astra 之后,Reddit 开始算算力、账单与权限
- Astra 之后,七个 AI 社区把发布会改成了验收单
- Astra 之外,七个 AI 社区今天在验证什么?
- GPT-6 Astra 发布帖、同时宕机与本地替代:Reddit 七区 70 条热帖(9 月 4 日)
