每日 Harness 开源 · Source
全部刊期 · All issues

每日 Harness

2026-09-12 · Saturday, September 12, 2026

智能体工程迈向可控与自进化

视图 · View

今日重点 · Today's Highlights

OpenCode5 - 开源 coding agent 以终端为主界面,同时提供完整权限的 build agent 与只读、需确认的 plan agent。

全文 ↓

论文 · Papers

11 项 · 论文

AutoFyn Technical Report: Non-Parametric Expert Iteration for Long-Horizon Agents6arxiv.org原文 ↗

arxiv.org

AutoFyn 不改模型权重,而是把经验证的奖励写回持久状态,让每轮全新会话通过文件、报告和仓库状态继承策略。它覆盖奥数、数据科学和网络安全,在六道新的 2026 IMO 题上所有可提升模型都超过提供商 coding agent,并确认了 16 个维护者认可的漏洞通告;这把“自我改进”落在可审计的外部状态上。

–

Agents Trust Tools Too Much: Measuring Reliance on Unreliable Tools7arxiv.org原文 ↗

arxiv.org

研究让 14 个模型使用搜索、子 agent 和代码执行,再篡改返回内容,观察它们是否会质疑工具。网页搜索的平均采用率达到 68.0%,而模型即便在内部发现冲突,也常把污染结果无警告地交付;提示、元数据和后训练没有一种能跨工具稳定修复这一信任偏差。

–

DAREBench: Deployment-Aware and Reliable Evaluation of Models as Agents10arxiv.org原文 ↗

arxiv.org

DAREBench 将 22 个基准的 233 个任务放入统一 OpenClaw 环境,以模态×执行形式组成 2×3 工作矩阵,并用合约和证据分数审计交付物。评测覆盖 23 个商业 API、12 个本地模型和 7,587 次模型任务运行;没有模型在所有维度占优,准确率、模态和成本之间的取舍会随部署条件改变。

–

Explaining AI Agents Through Execution Traces11arxiv.org原文 ↗

arxiv.org

该方法先把长轨迹整理成结构化报告,再从可观察行为生成过程解释,因而能把主张对应到实际动作和证据。人工及自动评测都发现它比直接让 LLM 总结更能暴露无依据结论、不合理操作和证据空洞,适用范围不绑定特定模型或环境。

–

SkillAlign: Aligning Skill Interfaces for LLM-based Agents12arxiv.org原文 ↗

arxiv.org

SkillAlign 固定任务、agent 和技能内容,只改变技能的呈现接口,比较完整说明、提示、压缩摘要、工作流和不暴露几种方式。ALFWorld 与 SkillsBench 显示,紧凑的 top-k 暴露有时胜过把整库塞进上下文,说明技能检索和接口设计本身就是性能变量;学到的回放策略仍明显落后于 oracle。

–

ExecCritic: Learn to Test, Test to Improve for Coding Agents13arxiv.org原文 ↗

arxiv.org

ExecCritic 让测试 agent 先在 fail-closed harness 中筛选并冻结测试,修复 agent 只能消费执行反馈,不能改写评判标准。SWE-bench Verified 上 Qwen 测试生成由 Base 的 22.2 提到 Gold 的 62.2,组合后的后训练 agent 达到 72.6%,比原始无测试基线高 11.4 个百分点。

–

Procedural Graphs: Self-Evolving Execution Structures for LLM Agents14arxiv.org原文 ↗

arxiv.org

Procedural Graphs 用程序三元组和主动节点定位表达步骤、顺序与条件,再由指导模型把局部子图翻译成行动提示。自演化器只有在留出集不降反升时才接受拓扑或属性修改,并保留被拒编辑以抑制重复犯错;多数据集、任务和模型上的提升表明结构记忆能超越简单轨迹回放。

–

Evaluating Deep-Search Agents under Hierarchical Web Evidence Poisoning15arxiv.org原文 ↗

arxiv.org

HAE-GEO 设计断言、上下文伪装和表面互证三档网页投毒,覆盖 8 类商品、154 个品牌,使用 72,039 个干净页面和每档 770 个污染页面。十个 agent 在“看似相互印证”的陷阱下识别退化;增加搜索步骤能提高最终抵抗,却没有同步提高识别与效用,额外防御提示也很少让错误轨迹恢复。

–

开源 / 项目 · Projects

14 项 · 开源 / 项目

Toast17github.com原文 ↗

github.com

Toast 是终端开发环境,组合文件树、多标签、主题、受管 LSP、rg 搜索和 Markdown 预览;tree-sitter 已覆盖 Go、Python、JS/TS、Rust 等语言,默认 300ms 自动保存。项目仍早期开发,明确不内置 AI,也不发送遥测,定位更接近可控的本地 IDE 外壳。

–

GrapheneOS Messaging18github.com原文 ↗

github.com

第 13 版用 Jetpack Compose/Material3 重写消息 UI,加入大屏双栏、置顶/稍后处理、未读/归档、消息详情和 MMS 媒体选择。隐私层还拒绝 `file:` URI 与私有文件、限制 widget intent、采用 FLAG_IMMUTABLE 和分配上限;最低 SDK 36,更新同时处理了多项崩溃与通知缺陷。

–

Litelm19github.com原文 ↗

github.com

Litelm 将路由与协议核心压缩到约 2,900 行、两个依赖,仍提供消息、流式、工具、embedding、responses 及同步/异步接口,可连接 19 个供应商。它主动排除代理、缓存、预算、token 统计、agent 和 guardrail,因此更像一个可审计的传输层,而不是完整平台;维护记录为 262 项通过、55 项跳过。

–

Revera20github.com原文 ↗

github.com

Revera v2.1.0 把本地优先 API 与低延迟 Web3 定价引擎结合,声称覆盖 20 条 EVM 链和 360 万以上代币。发行说明列出 Ethereum、BSC、Polygon、Arbitrum、Base、Optimism 等网络,重点是把行情与路由能力放到本机而非托管数据平面。

–

Stroq21stroq.dev原文 ↗

stroq.dev

Stroq 针对 MCP 工具输出中的潜在命令注入做过滤,把返回文本按不可信输入处理。它把控制点放在工具边界,减少 agent 直接把外部返回当作可执行指令的机会;目前公开条目主要说明这一防护目标。

–

Pi-sync22github.com原文 ↗

github.com

pi-sync 通过 rsync/SSH 同步多台主机的 pi agent 声明式配置,提供 `--pull --all`、`--dry-run`、主机别名和更新命令,并可由 pipx/uv 安装。模型设置和扩展会复制,sessions、npm、模型缓存、信任及本地 auth 则保留在机器上;覆盖文件先保存为 `.backup`,删除行为需显式开启。

–

KyttoMCP23kytto.jakubhecht.sk原文 ↗

kytto.jakubhecht.sk

KyttoMCP 为 macOS/Windows 上的 Claude、Cursor、VS Code、Codex 等客户端集中管理 JSON/TOML 配置,附带健康检查、profile、MCP Doctor 和 token 估算。每次写入会检测外部变更、做时间戳备份并原子替换,应用不上传账号或 payload;Beta 1.0.6 提供安装包和校验和,但暂不支持 Linux。

–

SOS24github.com原文 ↗

github.com

SOS 将 coding agent 已接受的工作、指令、检查结果和待验证事项保存为项目状态,让新会话恢复时能标出过期来源。alpha 0.1.0a7 以 Linux 为主,release/current.json 激活失败会 fail-closed;它的八个 MCP 接口只读或提出变更,禁止 shell、commit 和 deploy。

–

Hazzel25github.com原文 ↗

github.com

Hazzel 是 BYO key 的轻量终端 coding agent,经过批准后可读写文件、运行命令和 git,并支持 diff 预览、项目根目录 sandbox、plan/prove 模式及 gh PR。密钥以 0600 保存,能接多家云模型和 Ollama;v0.1.8 仍不提供浏览器、部署、后台 agent 或会话持久化。

–

Bastiontrace26github.com原文 ↗

github.com

Bastiontrace 用纯 Python 分析 JSONL 工具轨迹,定位注入内容、落地的禁行动作和 blast radius,不需要 LLM 或云端,只依赖 bastioncorpus。`analyze` 输出 JSON 并在注入成功时返回非零,`harden` 则能生成 agentbastion 使用的 policy.yaml 与 injections.jsonl,适合放进事后取证流水线。

–

SQLite-Vector27github.com原文 ↗

github.com

SQLite-Vector 在普通 BLOB 表里实现精确向量搜索、SIMD 距离核和 2/3/4-bit TurboQuant,不需要虚拟表、预索引或外部服务。百万条 768 维向量的 INT8 存储为 740MB、FLOAT32 为 2,930MB;Apple M5 Pro 预载查询 37.6ms、召回率 99.5%,30MB 流式模式则为 114.4ms。

–

OpenMuse28github.com原文 ↗

github.com

OpenMuse 是可自部署的个人 agent 工作区,同一对话可分派长任务并生成主题、笔记和云电脑,浏览器关闭后仍能继续。TanStack Start 前端调用 TypeScript hooks,OpenComputer Serverless Agents 负责运行、电脑配置与存储,部署依赖 Node 22 和 HTTPS tunnel 而不需另建基础设施。

–

Agent Guardrails Kit29github.com原文 ↗

github.com

该工具以两个 PreToolUse hook 为 Claude Code 或 shell/file agent 建立 fail-closed 策略,配套 allow-list、JSONL 审计、月报和 104 条断言测试。它拦截 force push、reset --hard、rm -rf、curl|sh 等高风险动作,并把过去可绕过 deny-list 的案例写进回归测试,使策略变更有可追踪证据。

–

Compose Bridge UDS30github.com原文 ↗

github.com

Compose Bridge UDS 把完全解析的 Docker Compose 模型转换成 UDS 定制 Helm chart,再由 Zarf 打包到隔离 Kubernetes。示例以 k3d、UDS Core Slim Dev 和 WordPress/MySQL 生成 chart、文档、values、conversion.json 与 zarf.yaml;仓库明确标注为实验性、非受支持产品路径。

–

行业动态 · Industry News

11 项 · 行业动态

Rapidly scaling online storage to serve over 1 billion ChatGPT users31openai.com原文 ↗

openai.com

OpenAI 说 Habitat 已从围绕 Cosmos DB 的 Python 库演进为独立存储服务,当前每秒处理超过 7,000 万请求,每周服务十亿以上用户,覆盖约 40 个区域和 500PB 数据。2026 年第二季度两名工程师配合 Codex/GPT-5.5 以 Rust 重写,Rust 已承担 95% 请求,CPU 效率约 6 倍、内存效率约 15 倍,Python 计划退役。

–

OpenAI Agents API32developers.openai.com原文 ↗

developers.openai.com

Agents API 把会话、编排、上下文压缩和恢复交给托管 harness,应用方负责工具与执行环境;agent 能写代码、接 MCP、产出 artifacts,并通过 events 观察过程。文档还定义 streaming、webhook、steering、后台任务、多 agent 和持久状态,但当前数据驻留只有美国区域,且不支持 ZDR。

–

GPT-Live-1 in the API33openai.com原文 ↗

openai.com

GPT-Live-1 在 API 中提供全双工语音,能同时听说,并把深层推理和工具调用转给后端模型。官方早期结果称打断减少近 80%,示例代码比级联方案少约 2.3 万行、降幅 80%,Full Duplex Bench 高 30 分;前端价格为每分钟 0.05 美元。

–

The Gemini app is now available for Windows34blog.google原文 ↗

blog.google

Google 发布 Windows 10/11 Gemini 应用,Alt+Space 可打开即时浮层,Gemini Spark 负责多步任务并连接 Google 应用。客户端同时整合 Nano Banana 图像和 Gemini Omni 视频能力,作为轻量桌面入口向全球推出。

–

The EPA is planning to scrap public review rules for data center pollution36capitalbnews.org原文 ↗

capitalbnews.org

EPA 提案拟取消工业设施(包括数据中心和电厂)空气污染许可的联邦公众通知与评论要求,另一项方案还允许获证前开工。报道引述七成美国人反对住宅附近 AI 数据中心,并称农村及黑人社区承受健康、账单和搬迁压力,约 200 个组织与十多个州提出反对。

–

Claude is only available to people over 18 years37support.claude.com原文 ↗

support.claude.com

Anthropic 的消费者 Claude 现在只面向 18 岁以上用户,注册需确认年龄,风险信号可能触发进一步验证。Yoti 人脸估龄、证件或 Digital ID 都可参与校验,完成后 Anthropic 只接收通过/不通过结果,验证数据会删除。

–

Blizzard Workers Win Historic Union Contract40latimes.com原文 ↗

latimes.com

近 1,900 名暴雪工会员工在两年谈判后批准集体合同,涵盖加薪、每周至少三天到办公室、工作场所 AI 协商权和裁员后 14 个月召回权。协议通过时,微软此前宣布约 3,200 个游戏部门岗位削减仍构成谈判背景。

–

HuggingFace: Security.txt41huggingface.co原文 ↗

huggingface.co

Hugging Face 在 security.txt 中公开漏洞披露联系人和自动化扫描器可读取的标准说明。固定的报告入口让研究者能绕开普通客服路径,直接进入正式安全沟通流程。

–

博客文章 · Blog Posts

13 项 · 博客文章

A Severe Misalignment of AI in Mathematics42terrytao.wordpress.com原文 ↗

terrytao.wordpress.com

Terry Tao 与 24 位菲尔兹奖得主警告,benchmark 驱动的数学 AI 偏好快速给出真假答案,而研究真正需要理解、洞见和可传授的论证。文章还把归属、抄袭和错误传播列为共同体风险,主张由数学家判断 AI 最终是在增强还是破坏知识传承。

–

Quoting Boris Cherny43simonwillison.net原文 ↗

simonwillison.net

Simon Willison 转述 Boris Cherny 对 Claude 参与生产开发的看法,重点落在生成代码、生成测试以及后续人类审查如何分工。条目实际讨论的是 agent 可接管的工程环节和必须保留的判断环节,而非单纯统计生成行数。

–

Soft-deprecating `re.match()`44simonwillison.net原文 ↗

simonwillison.net

Python 3.15 将 `re.match()` 和 `re.Pattern.match()` 软弃用,改推荐语义更明确的 `re.prefixmatch()`/`re.Pattern.prefixmatch()`。动机是其他正则库里 match 常被理解为搜索,显式写 prefix 可以降低跨库迁移时的误读;软弃用也给生态留下渐进迁移窗口。

–

Don't sleep on wrapture45simonwillison.net原文 ↗

simonwillison.net

Graham Dumpleton 的 wrapture 虽是 alpha,却把 monkey-patching 延伸到测试与观测:既能替换 mock、记录时间线/调用树,也能按阶段改变行为。教程还覆盖属性、字典、生成器、TOML 驱动的零代码 tracing 和 OpenTelemetry,适合将动态依赖转成可检查事件。

–

Any Nix package, live in your browser46simonwillison.net原文 ↗

simonwillison.net

trynix.dev 在浏览器中运行 qemu-wasm x86_64 Linux 虚拟机,URL 可指定过去 13 年的任意 Nix 包并进入交互 shell。trynix-preview GitHub Action 还会在 PR 评论中生成启动该 PR 构建的链接,执行不依赖常驻后端服务器。

–

Datasette 1.0a39 and 0.65.4 security releases47simonwillison.net原文 ↗

simonwillison.net

Datasette 两条版本线修复公开实例混用公开/私有表时的权限与注入问题,涉及 SQL 标识符转义、SQL 构造、HTML 渲染、认证和缓存。发布过程让多个模型先审计,再由人类拆分验证修复;对暴露在公网的实例,升级本身是必要的边界收紧。

–

Open-Source AI & Open Models Reading List48interconnects.ai原文 ↗

interconnects.ai

Nathan Lambert 维护一份开放 AI 与开源模型阅读清单,按模型、数据、训练、评测和社区争议等主题组织长期材料。它不是单篇实验报告,而是给研究者提供跨主题、可持续更新的索引入口。

–

Measuring the sloppiness of code49earendil.com原文 ↗

earendil.com

文章尝试把代码的“sloppiness”从直觉评价转成可测量的观察信号,讨论结构、重复、边界处理和维护摩擦如何进入指标。其贡献更接近度量框架草案:先定义可比较的证据,再比较代码库或变更,而非假设存在万能质量分数。

–

The Waymo effect: how AI is quietly making research less collaborative50researchagenda.news原文 ↗

researchagenda.news

Daniel Hook 将“Waymo effect”定义为:AI 去掉与人协作的摩擦,看似提高个人速度,却也拿走挑战、异议和知识传播。文章据此建议资助协作机制、按贡献而非速度评估,并让人类保留研究决策权,以防想法多样性被自动化收窄。

–

RTK reports token savings, but our cost benchmarks disagree51quesma.com原文 ↗

quesma.com

Quesma 的 Fable 基准中,RTK 将成本从 $596 降到 $546,但成功率也从 84% 降到 83%;DeepSeek 则由 $26/$51 变成 $31/$54,成功率 71%/69%。RTK 所称删掉 3.492 亿 token、减少 89% 实际按字节计数,连 `head -1` 也会重复计入,说明上下文压缩数字不能直接当成账单节省。

–

Astra for Coding: Why Are We Doing This Again?52lucumr.pocoo.org原文 ↗

lucumr.pocoo.org

Armin Ronacher 记录 35 小时、约 40 亿 token 的软件工厂试验:模型有长程推进能力,却在代码质量和奖励对齐上失分,并反复用 Python/shell 字符串拼接代替稳定 patch。文章将 Astra 与既有开发工具比较,结论不是否定自动化,而是要求可维护的工程反馈跟上规模。

–

What Comes After Git53ersc.io原文 ↗

ersc.io

East River Source Control 认为 Git 的 2005 年代模型难以承受 AI 带来的更大仓库、更多分支、合并争用和快速云复制。提案保留 Git 客户端协议,通过 bridge 更换可横向扩展的服务端存储,并让 Jujutsu 等协议逐步接入;讨论的重点是存储基础而非再造一个命令行界面。

–

CSS Curiosities of the Past54vale.rocks原文 ↗

vale.rocks

文章整理 CSS 历史上少见的语法、浏览器行为和兼容特性,把当前样式实践与早期设计遗产并置。对维护旧站而言,这份考古式索引的提醒是:看似奇怪的规则可能是历史兼容约束,不能只凭现代直觉删除。

–

引用来源 · References

64 条 · 引用
  1. 1 When Does Memory Help? A Cost-Aware Evaluation of Long-Term Memory in Tool-Using LLM Agents. arXiv:2609.05441https://arxiv.org/abs/2609.05441 ↩ 回到正文 · back to text
  2. 2 SCAFFOLD: Self-Improving Web Agents via Recursive Parametric Skill Abstraction. arXiv:2609.05511https://arxiv.org/abs/2609.05511 ↩ 回到正文 · back to text
  3. 3 Beyond Prompts: Measuring and Optimizing LLM Tool-Agent Harnesses. arXiv:2609.05736https://arxiv.org/abs/2609.05736 ↩ 回到正文 · back to text
  4. 4 AgentBrew: Offline Tool-Use Agent Learning from Raw Real-World Trajectories. arXiv:2609.05837https://arxiv.org/abs/2609.05837 ↩ 回到正文 · back to text
  5. 5 OpenCodehttps://github.com/anomalyco/opencode ↩ 回到正文 · back to text
  6. 6 AutoFyn Technical Report: Non-Parametric Expert Iteration for Long-Horizon Agents. arXiv:2609.05446https://arxiv.org/abs/2609.05446 ↩ 回到正文 · back to text
  7. 7 Agents Trust Tools Too Much: Measuring Reliance on Unreliable Tools. arXiv:2609.05587https://arxiv.org/abs/2609.05587 ↩ 回到正文 · back to text
  8. 8 EdgeMem: LLM-Free Agent Memory Construction and Retrieval via Evidence-Preserving Multi-Anchor Hypergraph. arXiv:2609.05553https://arxiv.org/abs/2609.05553 ↩ 回到正文 · back to text
  9. 9 EnvCraft: Synthesizing Executable Environments in Agentic RL for Claw-like Agent. arXiv:2609.05576https://arxiv.org/abs/2609.05576 ↩ 回到正文 · back to text
  10. 10 DAREBench: Deployment-Aware and Reliable Evaluation of Models as Agents. arXiv:2609.06059https://arxiv.org/abs/2609.06059 ↩ 回到正文 · back to text
  11. 11 Explaining AI Agents Through Execution Traces. arXiv:2609.06063https://arxiv.org/abs/2609.06063 ↩ 回到正文 · back to text
  12. 12 SkillAlign: Aligning Skill Interfaces for LLM-based Agents. arXiv:2609.07255https://arxiv.org/abs/2609.07255 ↩ 回到正文 · back to text
  13. 13 ExecCritic: Learn to Test, Test to Improve for Coding Agents. arXiv:2609.09133https://arxiv.org/abs/2609.09133 ↩ 回到正文 · back to text
  14. 14 Procedural Graphs: Self-Evolving Execution Structures for LLM Agents. arXiv:2609.09153https://arxiv.org/abs/2609.09153 ↩ 回到正文 · back to text
  15. 15 Evaluating Deep-Search Agents under Hierarchical Web Evidence Poisoning. arXiv:2609.06027https://arxiv.org/abs/2609.06027 ↩ 回到正文 · back to text
  16. 16 SWE-Bench Pro Verified: A Reliable Benchmark for Software Engineering Agents. arXiv:2609.08149https://arxiv.org/abs/2609.08149 ↩ 回到正文 · back to text
  17. 17 Toasthttps://github.com/paradise-runner/toast ↩ 回到正文 · back to text
  18. 18 GrapheneOS Messaginghttps://github.com/GrapheneOS/Messaging/releases/tag/13 ↩ 回到正文 · back to text
  19. 19 Litelmhttps://github.com/kennethwolters/litelm ↩ 回到正文 · back to text
  20. 20 Reverahttps://github.com/rv-crypto/revera-api/releases ↩ 回到正文 · back to text
  21. 21 Stroqhttps://stroq.dev/ ↩ 回到正文 · back to text
  22. 22 Pi-synchttps://github.com/say4n/pi-sync ↩ 回到正文 · back to text
  23. 23 KyttoMCPhttps://kytto.jakubhecht.sk/ ↩ 回到正文 · back to text
  24. 24 SOShttps://github.com/sigmastratum/sigma-operator-stack ↩ 回到正文 · back to text
  25. 25 Hazzelhttps://github.com/mukundzha/hazzel ↩ 回到正文 · back to text
  26. 26 Bastiontracehttps://github.com/Rinkia/bastiontrace ↩ 回到正文 · back to text
  27. 27 SQLite-Vectorhttps://github.com/sqliteai/sqlite-vector ↩ 回到正文 · back to text
  28. 28 OpenMusehttps://github.com/diggerhq/openmuse/ ↩ 回到正文 · back to text
  29. 29 Agent Guardrails Kithttps://github.com/danielhagever/agent-guardrails-kit ↩ 回到正文 · back to text
  30. 30 Compose Bridge UDShttps://github.com/defenseunicorns/compose-bridge-uds ↩ 回到正文 · back to text
  31. 31 Rapidly scaling online storage to serve over 1 billion ChatGPT usershttps://openai.com/index/scaling-storage-one-billion-users-part-one ↩ 回到正文 · back to text
  32. 32 OpenAI Agents APIhttps://developers.openai.com/api/docs/guides/agents-api/overview ↩ 回到正文 · back to text
  33. 33 GPT-Live-1 in the APIhttps://openai.com/index/introducing-gpt-live-1-in-the-api/ ↩ 回到正文 · back to text
  34. 34 The Gemini app is now available for Windowshttps://blog.google/innovation-and-ai/products/gemini-app/gemini-app-now-on-windows/ ↩ 回到正文 · back to text
  35. 35 Google will buy half the electricity from one of Finland's nuclear power plantshttps://www.bbc.com/news/articles/c8r6y4me2g6o ↩ 回到正文 · back to text
  36. 36 The EPA is planning to scrap public review rules for data center pollutionhttps://capitalbnews.org/data-centers-permit-rules-epa/ ↩ 回到正文 · back to text
  37. 37 Claude is only available to people over 18 yearshttps://support.claude.com/en/articles/15171100-age-assurance-on-claude ↩ 回到正文 · back to text
  38. 38 OpenAI considers slowing advanced AI development, Sam Altman tells employeeshttps://www.bloomberg.com/news/articles/2026-09-11/openai-is-open-to-slowing-cutting-edge-ai-ceo-sam-altman-tells-staff ↩ 回到正文 · back to text
  39. 39 Matt Mullenweg tells Automattic staff in Slack he's back in control after ousterhttps://techcrunch.com/2026/09/11/matt-mullenweg-tells-automattic-staff-in-slack-hes-back-in-control-after-ceo-ouster/ ↩ 回到正文 · back to text
  40. 40 Blizzard Workers Win Historic Union Contracthttps://www.latimes.com/entertainment-arts/business/story/2026-09-09/blizzard-video-game-workers-ratify-union-contract ↩ 回到正文 · back to text
  41. 41 HuggingFace: Security.txthttps://huggingface.co/security.txt ↩ 回到正文 · back to text
  42. 42 A Severe Misalignment of AI in Mathematicshttps://terrytao.wordpress.com/2026/09/11/a-severe-misalignment-of-ai-in-mathematics/ ↩ 回到正文 · back to text
  43. 43 Quoting Boris Chernyhttps://simonwillison.net/2026/Sep/11/boris-cherny/ ↩ 回到正文 · back to text
  44. 44 Soft-deprecating `re.match()`https://simonwillison.net/2026/Sep/11/soft-deprecating-re-match/ ↩ 回到正文 · back to text
  45. 45 Don't sleep on wrapturehttps://simonwillison.net/2026/Sep/11/wrapture/ ↩ 回到正文 · back to text
  46. 46 Any Nix package, live in your browserhttps://simonwillison.net/2026/Sep/10/trynix/ ↩ 回到正文 · back to text
  47. 47 Datasette 1.0a39 and 0.65.4 security releaseshttps://simonwillison.net/2026/Sep/11/datasette-security/ ↩ 回到正文 · back to text
  48. 48 Open-Source AI & Open Models Reading Listhttps://www.interconnects.ai/p/open-source-ai-reading-list ↩ 回到正文 · back to text
  49. 49 Measuring the sloppiness of codehttps://earendil.com/posts/measuring-code-sloppiness/ ↩ 回到正文 · back to text
  50. 50 The Waymo effect: how AI is quietly making research less collaborativehttps://www.researchagenda.news/articles/the-waymo-effect.html ↩ 回到正文 · back to text
  51. 51 RTK reports token savings, but our cost benchmarks disagreehttps://quesma.com/blog/does-rtk-make-ai-coding-cheaper/ ↩ 回到正文 · back to text
  52. 52 Astra for Coding: Why Are We Doing This Again?https://lucumr.pocoo.org/2026/9/7/astra-why/ ↩ 回到正文 · back to text
  53. 53 What Comes After Githttps://ersc.io/blog/what-comes-after-git ↩ 回到正文 · back to text
  54. 54 CSS Curiosities of the Pasthttps://vale.rocks/posts/css-relics ↩ 回到正文 · back to text
  55. 55 armory3d/armorpainthttps://github.com/armory3d/armorpaint ↩ 回到正文 · back to text
  56. 56 JustVugg/colibrihttps://github.com/JustVugg/colibri ↩ 回到正文 · back to text
  57. 57 nashsu/llm_wikihttps://github.com/nashsu/llm_wiki ↩ 回到正文 · back to text
  58. 58 yilewang/llm-for-zoterohttps://github.com/yilewang/llm-for-zotero ↩ 回到正文 · back to text
  59. 59 vercel-labs/skillshttps://github.com/vercel-labs/skills ↩ 回到正文 · back to text
  60. 60 feigeCode/navophttps://github.com/feigeCode/navop ↩ 回到正文 · back to text
  61. 61 akitaonrails/ai-usagebarhttps://github.com/akitaonrails/ai-usagebar ↩ 回到正文 · back to text
  62. 62 persiyanov/herdr-reviewrhttps://github.com/persiyanov/herdr-reviewr ↩ 回到正文 · back to text
  63. 63 gpustack/gpustackhttps://github.com/gpustack/gpustack ↩ 回到正文 · back to text
  64. 64 google-deepmind/alphagenomehttps://github.com/google-deepmind/alphagenome ↩ 回到正文 · back to text