Agent 从聊天框迁入工作现场。 Perplexity 进法律工作系统、Claude 进 Slack、豆包进办公任务,入口正在贴近真实流程和权限边界。
今日重点项目雷达
baidu/Unlimited-OCR
是什么:Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.;判断标签:趋势增强
Source ↗QwenLM/Qwen-AgentWorld
是什么:Qwen-AgentWorld: Language World Models for General Agents;判断标签:新变量
Source ↗omnigent-ai/omnigent
是什么:Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.;判断标签:趋势增强
Source ↗ksimback/looper
是什么:Design visual, review-gated agent loops for Claude Code before you run them.;判断标签:新变量
Source ↗StarTrail-org/PixelRAG
是什么:The end of web parsing. The beginning of scalable pixel-native search.;判断标签:新变量
Source ↗nexu-io/open-design
是什么:🎨 Local-first, open-source Claude Design alternative. 🖥️ Native desktop app. ⚡ 259+ Skills · ✨ 142+ Design Systems 🖼️ Web · desktop · mobile prototypes · slides · images · videos · HyperFrames 📦 Sandboxed preview · HTML/PDF/PPTX/MP4 export 🤖 Claude Code / OpenClaw / Codex / Cursor / OpenCode / Qwen / Copilot / Hermes / Kimi & 17+ CLIs.;判断标签:趋势增强
Source ↗safishamsi/graphify
是什么:AI coding assistant skill (Claude Code, Codex, OpenCode, Cursor, Gemini CLI, and more). Turn any folder of code, SQL schemas, R scripts, shell scripts, docs, papers, images, or videos into a queryable knowledge graph. App code + database schema + infrastructure in one graph.;判断标签:趋势增强
Source ↗MemPalace/mempalace
是什么:The best-benchmarked open-source AI memory system. And it's free.;判断标签:待验证
Source ↗数据来源:AI HOT + GitHub 近期爆发项目 + BestBlogs 今日精选 + Follow Builders 建造者 feed + 必要核实。筛选标准:实时性、增速、AI/Agent 相关度、对 ABU9/OpenClaw 的启发。
一、今日主线判断
Agent 从聊天框迁入工作现场。 Perplexity 进法律工作系统、Claude 进 Slack、豆包进办公任务,入口正在贴近真实流程和权限边界。
企业 Agent 的基础设施竞争变清晰。 Identity、Runtime、Sandbox、Evaluation、Connectors Debugger 成为平台必答题,不再只是模型能力竞赛。
可复用 Skill 正在成为 AI 交付资产。 Codex 手册、行业投研 Skill、BuilderIO skills、OpenClaw/open-design 类项目都在把一次经验沉淀为可安装能力。
Benchmark 和官方效果声明必须降权处理。 Qwen-AgentWorld、NVIDIA、DFlash、OpenAI 新模型都有量化说法,但今天只把发布事实视为高可信,效果要等第三方或本地复测。
ABU9 的机会不是追热点模型,而是做行业 Agent 工作台。 汽车销售、售后、培训、合同和车间知识可以先从文档理解、权限审计、可溯源 Skill 三件事落地。
二、GitHub 近期爆发项目
baidu/Unlimited-OCR
- 是什么:Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.
- 元数据:创建 2026-06-18;stars 6176; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=6;stars/day≈1029.33;repeat_7d=2。
- 判断标签:趋势增强
- 意味着什么:长文档一次性解析正在变成企业知识入口,适合拿维修手册、合同扫描件、工单票据做 ABU9 文档理解 POC。
- 链接:https://github.com/baidu/Unlimited-OCR
QwenLM/Qwen-AgentWorld
- 是什么:Qwen-AgentWorld: Language World Models for General Agents
- 元数据:创建 2026-06-22;stars 257; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=2;stars/day≈128.5;repeat_7d=0。
- 判断标签:新变量
- 意味着什么:“先预测环境再行动”的语言世界模型把 Agent 训练从只靠真实环境推进到可模拟、可控的方向,值得跟踪其基准是否可复现。
- 链接:https://github.com/QwenLM/Qwen-AgentWorld
HKUDS/AgentSpace
- 是什么:"AgentSpace: Human + Agents. One Team. One Workspace"
- 元数据:创建 2026-06-22;stars 347; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=2;stars/day≈173.5;repeat_7d=1。
- 判断标签:趋势增强
- 意味着什么:人和多个 Agent 共处同一工作区的产品形态继续升温,OpenClaw 需要把权限、频道记忆、任务状态做成一等能力。
- 链接:https://github.com/HKUDS/AgentSpace
omnigent-ai/omnigent
- 是什么:Omnigent is an open-source AI agent framework and meta-harness: orchestrate Claude Code, Codex, Cursor, Pi, and custom agents — swap harnesses without rewriting, enforce policies and sandboxing, and collaborate in real time from any device.
- 元数据:创建 2026-06-11;stars 4695; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=13;stars/day≈361.15;repeat_7d=3。
- 判断标签:趋势增强
- 意味着什么:跨 Claude Code、Codex、Cursor 的 meta-harness 说明开发者正在追求可切换执行底座;但 7 天重复出现 3 次,今天只保留为持续跟踪。
- 链接:https://github.com/omnigent-ai/omnigent
ksimback/looper
- 是什么:Design visual, review-gated agent loops for Claude Code before you run them.
- 元数据:创建 2026-06-18;stars 312; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=6;stars/day≈52.0;repeat_7d=0。
- 判断标签:新变量
- 意味着什么:把 Agent loop 先可视化、评审再运行,正好对应企业长任务的“先审计划、再授权执行”。
- 链接:https://github.com/ksimback/looper
BuilderIO/skills
- 是什么:Skills for coding agents
- 元数据:创建 2026-06-10;stars 2568; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=14;stars/day≈183.43;repeat_7d=2。
- 判断标签:趋势增强
- 意味着什么:coding agent 的能力沉淀正在从 prompt 走向可安装 skill,和 OpenClaw 的可复用技能体系高度同构。
- 链接:https://github.com/BuilderIO/skills
StarTrail-org/PixelRAG
- 是什么:The end of web parsing. The beginning of scalable pixel-native search.
- 元数据:创建 2026-05-29;stars 5129; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=26;stars/day≈197.27;repeat_7d=2。
- 判断标签:新变量
- 意味着什么:pixel-native 检索绕开传统网页解析,对复杂页面、报表、车机截图和售后系统界面有潜在价值。
- 链接:https://github.com/StarTrail-org/PixelRAG
nexu-io/open-design
- 是什么:🎨 Local-first, open-source Claude Design alternative. 🖥️ Native desktop app. ⚡ 259+ Skills · ✨ 142+ Design Systems 🖼️ Web · desktop · mobile prototypes · slides · images · videos · HyperFrames 📦 Sandboxed preview · HTML/PDF/PPTX/MP4 export 🤖 Claude Code / OpenClaw / Codex / Cursor / OpenCode / Qwen / Copilot / Hermes / Kimi & 17+ CLIs.
- 元数据:创建 2026-04-28;stars 70589; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=57;stars/day≈1238.4;repeat_7d=1。
- 判断标签:趋势增强
- 意味着什么:本地优先、沙箱预览、HTML/PDF/PPTX/MP4 导出的设计工作台,继续验证“售前 artifact 工厂”方向。
- 链接:https://github.com/nexu-io/open-design
safishamsi/graphify
- 是什么:AI coding assistant skill (Claude Code, Codex, OpenCode, Cursor, Gemini CLI, and more). Turn any folder of code, SQL schemas, R scripts, shell scripts, docs, papers, images, or videos into a queryable knowledge graph. App code + database schema + infrastructure in one graph.
- 元数据:创建 2026-04-03;stars 71563; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=82;stars/day≈872.72;repeat_7d=1。
- 判断标签:趋势增强
- 意味着什么:把代码、SQL、文档、图片、视频转知识图谱,适合企业项目交付审计和遗留系统理解。
- 链接:https://github.com/safishamsi/graphify
MemPalace/mempalace
- 是什么:The best-benchmarked open-source AI memory system. And it's free.
- 元数据:创建 2026-04-05;stars 56289; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=80;stars/day≈703.61;repeat_7d=1。
- 判断标签:待验证
- 意味着什么:开源 AI memory 高 star 仍需复核 benchmark 与真实接口,但长期记忆是 OpenClaw 必补底座。
- 链接:https://github.com/MemPalace/mempalace
三、最近 7 天新项目:值得点开
bozhouDev/codex-orange-book
- 是什么:Codex 橙皮书:从安装到实战案例的全链路 Codex 使用指南(非官方开源,含可下载 PDF)
- 元数据:创建 2026-06-23;stars 1310; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=1;stars/day≈1310.0;repeat_7d=2。
- 判断标签:趋势增强
- 意味着什么:Codex 使用经验正在被中文社区手册化,可转成企业 AI Coding 培训材料和内部规范。
- 链接:https://github.com/bozhouDev/codex-orange-book
goehou/tabbit-toy
- 是什么:这是一个基于tabbit的研究包,可以转化成OAI格式出来,同时增加了会员认证功能和一键提取cookie的浏览器拓展,方便快速本地快速使用claude gpt等模型
- 元数据:创建 2026-06-23;stars 269; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=1;stars/day≈269.0;repeat_7d=0。
- 判断标签:待验证
- 意味着什么:OAI 格式转换和本地使用便利性有需求,但包含 cookie 提取和会员认证,必须复核合规边界。
- 链接:https://github.com/goehou/tabbit-toy
lyra81604/zhengxi-views
- 是什么:可溯源的郑希(易方达基金经理)投研 Agent Skill——基于他全部公开观点原文 + 有原话佐证的投资方法 + 全市场基金真实数据,能溯源问答、按他框架给基金打分,绝不杜撰。⚠️仅研究学习辅助,不构成投资建议‼️website是郑希主页!
- 元数据:创建 2026-06-20;stars 987; 最近更新 2026-06-23;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=4;stars/day≈246.75;repeat_7d=2。
- 判断标签:新变量
- 意味着什么:可溯源投研 Agent Skill 展示了“观点原文 + 方法论 + 真实数据”的行业专家封装方式,可迁移到汽车销售/售后专家。
- 链接:https://github.com/lyra81604/zhengxi-views
yo-WASSUP/Good-Badminton
- 是什么:🏸 AI Badminton Hawk-Eye System
- 元数据:创建 2026-06-20;stars 476; 最近更新 2026-06-23;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=4;stars/day≈119.0;repeat_7d=1。
- 判断标签:新变量
- 意味着什么:AI Hawk-Eye 从体育视觉切入,说明低成本多摄像头/边缘视觉仍在扩散;对门店行为识别仅作低优先观察。
- 链接:https://github.com/yo-WASSUP/Good-Badminton
raiyanyahya/recall
- 是什么:Stop wasting tokens and re-explaining your project every session. Recall gives Claude Code durable memory — entirely offline.
- 元数据:创建 2026-06-19;stars 463; 最近更新 2026-06-22;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=4;stars/day≈115.75;repeat_7d=2。
- 判断标签:趋势增强
- 意味着什么:离线长期记忆解决“每次重新解释项目”的痛点,和 OpenClaw wiki/memory 的产品价值一致。
- 链接:https://github.com/raiyanyahya/recall
sums001/Windows-Copilot-API
- 是什么:Reverse engineered Windows Copilot into an OpenAI-compatible API. Access GPT-4 and GPT-5 models through a simple REST interface without API keys or billing.
- 元数据:创建 2026-06-19;stars 646; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=5;stars/day≈129.2;repeat_7d=2。
- 判断标签:噪音降权
- 意味着什么:逆向 Windows Copilot 到兼容 API,技术上说明“统一 API”需求强,但企业方案不能依赖非官方接口。
- 链接:https://github.com/sums001/Windows-Copilot-API
shadcn-labs/agentcn
- 是什么:shadcn/ui, but for building agents. 🤖
- 元数据:创建 2026-06-17;stars 246; 最近更新 2026-06-24;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=7;stars/day≈35.14;repeat_7d=0。
- 判断标签:新变量
- 意味着什么:shadcn/ui 的 agent 组件化思路值得看:企业 Agent UI 需要标准工具、状态、权限和回放组件。
- 链接:https://github.com/shadcn-labs/agentcn
eooce/transfer-api
- 是什么:Cloudflare Worker adapter for unlimited.surf OpenAI/Anthropic-compatible API routes.
- 元数据:创建 2026-06-19;stars 347; 最近更新 2026-06-21;采集窗口/来源查询 2026-06-23T20:00:20Z;采集时间 2026-06-24T20:00:20.079350+00:00;github.ranked_recent(综合 new_7d_100 / new_14d_200_ai_agent / new_30d_500_ai_agent / fresh_90d_active_ai_agent);age_days=5;stars/day≈69.4;repeat_7d=1。
- 判断标签:噪音降权
- 意味着什么:Cloudflare Worker 转接兼容 API 反映灰色代理需求,不进入 ABU9/OpenClaw 正式路线,只作为供应链风险样本。
- 链接:https://github.com/eooce/transfer-api
四、AI 热点新闻
Perplexity 推出 Computer for Counsel
- 事件:Perplexity 将 Computer 接入律师常用研究数据库、文档和案件管理系统,面向 Pro/Max 用户开放。
- 可信度:一手发布
- 意味着什么:垂直行业 Agent 已从“问答”进入“连接行业工作系统”,ABU9 的售后/法务/合同场景也应按系统连接而非聊天框设计。
- 链接:https://x.com/perplexity_ai/status/2069866668671766804
OpenAI 推送 GPT-5.5 Instant 新版本
- 事件:OpenAI 称新版 GPT-5.5 Instant 更会理解意图和复杂约束,先向付费用户推送。
- 可信度:一手发布
- 意味着什么:这是正式发布,但“更有趣/更可靠”属于官方效果描述,仍要以真实任务回归集复测。
- 链接:https://x.com/OpenAI/status/2069843083701915755
Mistral Connectors 加强企业控制面
- 事件:Mistral 发布连接器权限、API key scope、多账号、Debugger、Vibe Code/Workflows 集成。
- 可信度:一手发布
- 意味着什么:企业 AI 的胜负点正在变成连接器权限、调试和长任务不中断,OpenClaw 的工具权限与审计要跟上。
- 链接:https://mistral.ai/news/more-control-over-connectors
OpenAI 与 Broadcom 发布 Jalapeno 推理芯片
- 事件:OpenAI 与 Broadcom 宣布面向 LLM 推理的定制芯片 Jalapeno。
- 可信度:一手发布
- 意味着什么:模型公司继续向基础设施下钻,后续推理成本和供应链会影响企业 AI 价格;性能收益待第三方验证。
- 链接:https://openai.com/index/openai-broadcom-jalapeno-inference-chip
火山引擎推出 Agent Ready 基础设施
- 事件:火山引擎发布 AgentKit 与 ArkClaw 企业版升级,强调 Identity、Runtime、Sandbox、Evaluation。
- 可信度:一手发布
- 意味着什么:国内企业 Agent 平台也开始把身份、运行时、沙箱、评估摆到同一层,OpenClaw 可以借此校准产品语言。
- 链接:https://mp.weixin.qq.com/s/83mrPAPgQRKhxLkoSvRgBQ
豆包正式推出专业版
- 事件:豆包专业版面向复杂办公与生产力场景,包含本地电脑/浏览器/Skills/定时任务与 Office 套件。
- 可信度:一手发布
- 意味着什么:“办公任务模式 + Skills + 定时任务”正在成为 C 端到 B 端的过渡形态,ABU9 应关注办公 Agent 如何接企业流程。
- 链接:https://mp.weixin.qq.com/s/Sb-NMXTrWFQES1EDO_Gr2g
Qwen-AgentWorld 开源
- 事件:通义发布语言世界模型和 AgentWorldBench,覆盖 MCP、Search、Terminal、SWE、Web、OS、Android。
- 可信度:一手发布
- 意味着什么:开源是确定事件;榜单超过 GPT/Claude 的量化结论需要第三方复核。对 OpenClaw 最有价值的是环境模拟和轨迹数据。
- 链接:https://mp.weixin.qq.com/s/NV9WGpGsfFz35jww5agM9g
Figma Config 2026 扩展 AI 画布能力
- 事件:Figma 将设计画布扩展到代码、动画、3D、Shader 与生成式插件,但 AI 能力依赖外部模型。
- 可信度:媒体报道
- 意味着什么:设计工具正在变成 artifact 操作系统;但模型依赖会压缩利润,也给 OpenClaw 这类多模型工作流留下空间。
- 链接:https://the-decoder.com/figma-bets-on-human-judgment-at-config-2026-while-the-ai-powering-its-canvas-belongs-to-someone-else
ChatGPT Bidi 1 双向语音测试
- 事件:部分用户反馈 ChatGPT 出现 Bidi 1,可边听边说并支持打断。
- 可信度:待核实
- 意味着什么:若属实,语音 Agent 会从回合制走向实时协作;但 OpenAI 未官宣,暂不作为产品承诺依据。
- 链接:https://www.ithome.com/0/967/852.htm
Oracle 因 AI 应用大规模裁员并加码云基建
- 事件:Ars Technica 报道 Oracle 财年裁员 21000 人,同时计划继续融资扩建 OCI。
- 可信度:媒体报道
- 意味着什么:企业 AI 不是只增效,也会触发组织再分配;ABU9 做 AI 项目时要把岗位变化和流程责任讲清楚。
- 链接:https://arstechnica.com/ai/2026/06/oracles-21000-layoffs-help-drive-its-debt-fueled-ai-investments
五、论文 / 研究 / 观点
Google Research:推理如何解锁参数化知识
- 核心观点:启用推理 token 可能让模型更好调用内部事实记忆,说明“多想一步”有时不是推理而是检索自身参数。
- 意味着什么:优先把它转成可复测任务,而不是转发结论;对 ABU9/OpenClaw 相关的,进入选型或质量门。
- 链接:https://research.google/blog/thinking-to-recall-how-reasoning-unlocks-parametric-knowledge-in-llms
NVIDIA NeMo AutoModel:MoE 微调加速
- 核心观点:Hugging Face 文章显示一行 import 可让 MoE 微调吞吐提升、显存下降;官方/厂商 benchmark 仍需复测。
- 意味着什么:优先把它转成可复测任务,而不是转发结论;对 ABU9/OpenClaw 相关的,进入选型或质量门。
- 链接:https://huggingface.co/blog/nvidia/accelerating-fine-tuning-nvidia-nemo-automodel
DFlash:块扩散草稿模型做投机解码
- 核心观点:一次生成 token block 再并行验证,目标是把推理吞吐从模型层继续压成本。
- 意味着什么:优先把它转成可复测任务,而不是转发结论;对 ABU9/OpenClaw 相关的,进入选型或质量门。
- 链接:https://www.marktechpost.com/2026/06/24/dflash-speculative-decoding-drafts-whole-token-blocks-in-parallel-for-up-to-15x-higher-throughput-on-nvidia-blackwell
FFASR:真实远场 ASR 评测
- 核心观点:远场、噪声、混响条件比近场榜单更接近门店/车间语音场景,适合 ABU9 语音助手选型。
- 意味着什么:优先把它转成可复测任务,而不是转发结论;对 ABU9/OpenClaw 相关的,进入选型或质量门。
- 链接:https://huggingface.co/blog/ffasr-leaderboard
Stanford HAI:AI 招聘工具偏见研究
- 核心观点:大规模实地研究提示算法单一文化会放大系统性排斥;企业 AI 项目要保留独立审计和人工申诉通道。
- 意味着什么:优先把它转成可复测任务,而不是转发结论;对 ABU9/OpenClaw 相关的,进入选型或质量门。
- 链接:https://hai.stanford.edu/news/ai-hiring-tools-can-yield-racial-bias-and-systemic-rejection
字节 AI Coding 实践:代码贡献率不是最终指标
- 核心观点:TRAE 团队代码超 90% 由 AI 生成,但人均需求吞吐只提升 60%;可交付性和 harness 比生成率更重要。
- 意味着什么:优先把它转成可复测任务,而不是转发结论;对 ABU9/OpenClaw 相关的,进入选型或质量门。
- 链接:https://mp.weixin.qq.com/s/mdmaAyUIvxE8WT_GEbF2wQ
六、今天值得读 / 值得看(BestBlogs / Follow Builders 精选)
降级说明:BestBlogs public_today 本次返回 success=true 但 data=null,未取得可用条目;本栏目使用 Follow Builders normalized_signals 补位,并只收录带原始链接的内容。
标题/核心观点:Karpathy:Claude 交互范式更 inline
- 为什么值得看:把 Claude 嵌入人类正在做的活动,而不是让人切到单独聊天窗口;这正是 OpenClaw 要追的入口形态。
- 原帖链接:https://x.com/karpathy/status/2069547676849557725
标题/核心观点:Claude in Slack:频道里 @Claude 后独立沙箱工作
- 为什么值得看:Ben Cherny 展示 Claude 可在频道中 clone repo、写代码、测试编译,关键是“频道上下文 + 独立沙箱 + 多人协作”。
- 原帖链接:https://x.com/bcherny/status/2069474689819480394
标题/核心观点:Claude in Slack:给任务后让它持续工作
- 为什么值得看:这条补足了异步长任务形态:不是问答,而是指向频道和任务后持续推进。
- 原帖链接:https://x.com/bcherny/status/2069474691010707486
标题/核心观点:NotebookLM 校园使用反馈
- 为什么值得看:教育场景中的 NotebookLM 说明“把资料变成学习伙伴”仍是高频刚需,企业培训/维修手册可类比。
- 原帖链接:https://x.com/joshwoodward/status/2069406832523624696
标题/核心观点:OpenAI Codex 团队:反馈驱动修 Bug
- 为什么值得看:Codex 产品团队公开回应 bug 修复,说明 coding agent 的迭代速度会由真实任务反馈决定。
- 原帖链接:https://x.com/thsottiaux/status/2069579993588625574
标题/核心观点:swyx:Z.ai/GLM 使用度变化
- 为什么值得看:Z.ai 从少人使用到被更多人讨论,提示国内模型出海和开源生态仍有变量;但 IPO/股价相关信息只作观察。
- 原帖链接:https://x.com/swyx/status/2069598378191941835
标题/核心观点:No Priors:Biohub 与开源生物学
- 为什么值得看:Follow Builders 只有 1 条 podcast_signal;价值在“开放科学 + AI 工具链”趋势,不是 ABU9 主线,只保留低频观察。
- 链接:https://www.youtube.com/@NoPriorsPodcast
七、对我们有用的行动建议
做一个汽车行业可溯源 Skill 样板。 选 Harley 或 Hyundai,一个销售/售后主题,要求每个回答带原文证据、数据来源和不可回答边界。
给 OpenClaw 增加 Agent 运行时审计清单。 每个长任务记录身份、工具权限、沙箱、网络访问、验证命令、成本和人工确认点。
用 OCR/多模态做企业知识入口 POC。 拿维修手册、扫描合同、工单票据横评 Unlimited-OCR、Qwen/豆包/商业 OCR,指标含表格、版面、引用、成本。
把 AI Coding 培训从工具课升级为交付课。 参考字节实践,不考代码生成率,考需求吞吐、可交付性、缺陷率、回滚和 review 证据。
把“团队频道 Agent”作为 OpenClaw 展示场景。 模拟一个项目群:Agent 能读上下文、生成 PR/报告,但所有外发、删除、配置变更必须二次确认。
八、今日结论
今天的结构性变化是:AI Agent 的主战场从模型聊天迁到企业工作现场,真正值得下注的是可审计运行时、可溯源行业 Skill 和能进入流程的连接器。
降级项
- BestBlogs:
public_today.data=null,未纳入正式精选;已用 Follow Builders 补位。 - GitHub:API 正常;部分异常高 star 或非官方 API 项目已标
待验证/噪音降权,不作为确定趋势。 - AI HOT:采集正常;X 爆料、媒体报道、未官宣能力已按可信度降权。
- HTML 发布 / Wiki 写入:已完成;日报质量门、HTML 发布、Wiki 入库与 Wiki 质量门均通过。