SDAD: Spec-Driven Agentic Development for the AI-Native SDLC1 - 把 agentic coding 的质量控制前移到机器可读规格、独立验证和人类签署,并尝试用 Ambiguity Tax、Spec Fidelity 等指标治理交付。
全文 ↓今日重点 · Today's Highlights
Nexus: Depth-Adaptive KV-Cache Splicing and Retrieval-Decoupled Tool Routing for Agentic LLMs on Unified Memory2 - 在 250 工具注册表上用语义查找缓冲绕开 schema prefill,首参数 token 提前 1.66 倍,同时把 RoPE 漂移导致的 KV 拼接失真明确成设计边界。
全文 ↓DreamBench-SWE: A Multi-Session Memory-Hygiene Benchmark for Software Agents3 - 用不可推断的跨会话证据和隐藏可执行 oracle 测记忆卫生,并以预注册 successor 审计区分 benchmark 可用性与机制优越性。
全文 ↓Utility Under Attack: Agent Memory Poisoning and the Limits of Content Screening and Provenance Ranking4 - 量化长期记忆投毒的效用损失,显示内容筛选抓得到注入却识别不了假事实,来源权重也在安全与证据召回间没有可用折中。
全文 ↓论文 · Papers
15 项 · 论文本期重点SDAD: Spec-Driven Agentic Development for the AI-Native SDLC1arxiv.org原文 ↗
这份报告把“更大上下文”转化为流程要求:先固定意图和机器可读规格,再让 agent 合成,最后由独立验证者和人类签署放行。它还把工程、QA、平台和产品角色重新分配,并用修复乘数等治理量讨论采用路径。
本期重点Nexus: Depth-Adaptive KV-Cache Splicing and Retrieval-Decoupled Tool Routing for Agentic LLMs on Unified Memory2arxiv.org原文 ↗
设计的主要收益来自检索解耦而非无条件 KV 复用:压缩工具签名承担参数生成,schema KV splice 只作为有界的次级优化。作者报告中等深度 1.1 - 1.7 倍 TTFT 加速,但深上下文会回到 parity,且 reference-free drift gate 的 Spearman rho 只有 0.193。
本期重点DreamBench-SWE: A Multi-Session Memory-Hygiene Benchmark for Software Agents3arxiv.org原文 ↗
benchmark 用 hidden oracle 把“记住了文本”与“真的能完成依赖前会话证据的任务”分开计分;successor 中 typed-plus-raw probe 通过率 0.4611,固定 hosted Mem0 配置为 0.5389。论文明确这些数字刻画的是一个配置和可区分的测试剖面,不足以证明外部记忆机制的普遍优越性。
本期重点Utility Under Attack: Agent Memory Poisoning and the Limits of Content Screening and Provenance Ranking4arxiv.org原文 ↗
更强 provenance 权重在混合语料中可把准确率从 0.3167 拉到 0.7000,但当答案证据本身来自不可信来源时,evidence recall 变成零、准确率只剩 0.0417。这个反差说明“拒绝不可信内容”与“保留唯一有效证据”是同一个检索旋钮的两面。
PrimeAgentOrchestrator: Memory-Primed Agent Spawning for Personal AI Infrastructure6arxiv.org原文 ↗
PAO 在创建 Claude Code 会话时并行访问 PostgreSQL 实体库和 Cloudflare 语义索引,把融合后的个人记忆通过配置自动读取机制注入新 agent;作者以 2025 年 12 月至 2026 年 3 月四个月部署记录三代传递方案和各自失败点,证据类型是经验报告而非对照实验。
Representation Affects Retrieval: A Case Study of Skill Discovery and Routing in a Multimodal Agent Harness7arxiv.org原文 ↗
Tinycloud 的六任务消融显示,全量内联技能时每次选中 gold skill,全关闭会发生硬发现失败,而生产默认因词汇竞争错路由一次。这个小样本把“更多上下文一定更好”的直觉改成了表示之间会争夺 planner 注意力的具体机制。
When Retrieval Fails Before It Begins: Structurally Indirect Prerequisite Eviction as a Retention Failure in Agentic Memory8arxiv.org原文 ↗
论文把弱词法相关但结构上必需的上游证据淘汰定义为检索前失败,并用确定性 trace benchmark 观察完整链是否留下。Dependency-aware Semantic Garbage Collection 将词法编码器的完整链保留率从 0.03 提到 0.90、句向量编码器从 0.23 提到 1.00,同时报告预算增大后的一跳规则会退化。
Consilience: Conformally Calibrated Communication Control for Hidden-Profile Multi-Agent Reasoning9arxiv.org原文 ↗
控制器把不确定性、分歧、证据增益、冗余和过早共识压成状态,再决定 challenge/clarify/seek evidence/route 及发言者。逐轮 conformal calibration 为建议动作提供分布无关 regret 界,HiddenBench 上覆盖 12 个开闭权重模型,准确率和通信效率超过固定轮次或无结构讨论,部分结果还超过全信息基线。
Evaluating Skills, Not Just Agents: Agentic Continuous Evaluation of Skills10arxiv.org原文 ↗
ACES 把 skill、bundle、plugin 当作可执行能力包,固定模型、工作区、harness 和 scorer 做有无 skill 的配对试验,并以 ATIF 统一轨迹。145 个 skill 的扫描指标与 LLM judge 仅有 rho=0.14;947 个配对案例平均 composite Skill Lift 0.2134,72.8% 案例为正,说明静态文档门无法代替运行时测量。
Weighted Memory Tree: Remembering What Matters for Long-Horizon LLM Agents11arxiv.org原文 ↗
WMT 以任务 - 子任务 - 动作层级和动态 retention score 折叠已完成轨迹、衰减低效信息,并保留被折叠上下文的可访问性。GAIA-Text 上相对线性历史平均多 9.97 个百分点准确率、少 32.8% prompt token;投毒实验还显示评分与选择性衰减能限制错误信息的持久化。
Calibrating Criterion Revision in LLM Agents: Failure Modes and a Trace-Anchored Protocol12arxiv.org原文 ↗
CMB-0.1 要求失败识别、模型提案、跨 episode 转移、干预敏感和保持五条件同时成立;12 个案例的 192 个模型-条件试验没有一例满足。作者把结果解释为测量链校准失败,进而提出隐藏 transfer、显式 WRITE/NO-WRITE/ESCALATE、独立策略提交和冻结 oracle 的 CMB-0.4,而不是宣称模型不会修订标准。
Don't Solve, Just Compare: Tiny Advisors for Runtime Intervention in LLM Agents13arxiv.org原文 ↗
COTA 训练一个只比较同前缀反事实分支的轻量 advisor,让主 actor 根据非约束建议重规划,避免再部署一个能独立解题的 critic。它在 WebShop、ALFWorld、tau³-Retail 的三个 actor、九个设置中全部提升,论证“干预方向”可以由远弱于执行器的模型提供。
Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems14arxiv.org原文 ↗
中间件在生成前做 NLI 事实核验、五信号投毒检测和带阻尼的 Trust Index,在 TruthfulQA/Llama 3.3 70B 得到 91% 准确率、100% precision、100% 注入 recall。三模型 ROC-AUC 0.73 - 0.81,但实体替换等语义弱化仍漏检,FEVER 结果较弱也迫使每个数据域重新标定阈值。
ClawSentry: A Progressive Multi-Tier Security Monitor for Safeguarding Autonomous LLM Agents15arxiv.org原文 ↗
ClawSentry 把防护延伸到技能首次准入、调用意图、执行效果和事后后果,并以 L1 规则、L2 语义审查、只读 L3 agent 逐级消费审查预算;会话级逻辑还识别换工具和改写重试。在 SkillInject/Codex 上 ASR 从 39.55% 降到 2.61%,TSR 只从 83.78% 降到 83.05%,五个 Work Agent 的干净技能 TSR 为 98.7%。
Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence16arxiv.org原文 ↗
综述提出 Graph Engineering,用动态任务图、agent 图和状态图承载并行子任务、专业分工、验证及持久状态,把系统协调能力称为 System Intelligence。文章的贡献在于给 Prompt/Context/Harness/Loop Engineering 之上的统一组织抽象,并汇总开源数据与项目;它没有给出单一新算法或可比较的 benchmark 分数。
开源 / 项目 · Projects
13 项 · 开源 / 项目Transpose Spotify audio and isolate vocals/instruments in realtime17github.com原文 ↗
macOS 菜单栏应用通过 Core Audio 捕获 Spotify,实时改变调性并把歌曲拆成六个 stem 选择性回放;分离模型首次下载约 118 MB,支持从 DMG 或 Swift 源码安装。它需要 Automation 控制 Spotify 和录音权限,且发布包未签名,实际体验取决于 macOS 隐私授权链。
Agent Notifier21notifier.aicrew.in原文 ↗
安装页用手机 app 作为通知终端,一行命令完成 `notify.py` 安装、浏览器授权和测试 push;脚本只依赖 Python 标准库。除普通进度外,它能发阻塞式选择、文本提问、文件证明和 URL,配置优先读取项目 `.env.local`,并支持可选的 session watcher。
RepoRoulette22github.com原文 ↗
Python 工具把 GitHub 随机抽样拆成四个总体不同的 sampler:ID 近似全仓库均匀,Temporal/BigQuery/GHArchive 则偏向活跃或事件可见项目。README 用约 37% 的 live hit rate 解释 ID 采样的代价,并提醒 GitHub 事件 API 的 2025-10-07 变化会让默认新建仓库总体断档。
I built a lite LPU that can do inference on Karpathy's MicroGPT23lpulite.com原文 ↗
该硬件原型把目标收窄到 MicroGPT 推理,以专用数据通路展示从模型到芯片的完整闭环;现有页面没有公开通用 LLM 的吞吐、功耗或面积对照,因此它更像可运行的教学/实验平台,而非已验证的通用加速器。
Headless Tools24hdls.tools原文 ↗
服务集合把短链、pastebin、邮箱、提醒、文件上传等网页能力封装成 agent 可调用的无头接口,减少浏览器自动化和 UI 状态依赖。它定位为一组小型外部动作,而不是完整的工作流编排器,使用时仍需单独评估认证、配额和 SLA。
LunarBasic27lunarbasic.com原文 ↗
编译器把 BASIC 语法映射到原生 2D 游戏可执行文件,让极简语言直接产出桌面游戏二进制。项目强调缩短从代码到作品的反馈回路,具体后端、平台和性能边界仍应按实现版本理解。
A Claude Code skill that recovers export-blocked Kindle highlights28github.com原文 ↗
`kindle-highlights` skill 驱动用户自己的登录浏览器,结合 Mac Kindle 注释位置、Cloud Reader 渲染页和 Vision OCR,把被 Amazon 导出限制截断/隐藏的高亮恢复成带位置的 Markdown。四本书共 2,432 条中有 815 条受限,恢复文本中位位置误差 0 - 1 字符;作者强调 macOS 和个人账号范围及版权隐私边界。
行业动态 · Industry News
10 项 · 行业动态Advancing price-performance for developers with GPT‑5.6 in Kiro29openai.com原文 ↗
OpenAI 把 GPT-5.6 Sol/Terra/Luna 接入 AWS Kiro 的需求、设计、编码、审查和属性测试流程,强调上下文来自规格、代码库和团队标准。文章引用 Terminal-Bench 2.1 的 Kiro 测试称 Terra 成功任务成本下降约 82%,但这是合作方单环境数字,不能直接视为普遍成本曲线。
LLMs could control their host machines by exploiting inference engines30boydkane.com原文 ↗
演示把攻击面从模型输出扩展到推理引擎本身:恶意输入一旦触发解析器、插件或系统调用边界,模型就可能越过预期沙箱控制宿主机。其工程含义是对齐策略不能替代运行时隔离、最小权限和可审计的文件/命令接口。
Hot Chips 2026: CUDA Targets RISC-V31chipsandcheese.com原文 ↗
Hot Chips 报道 CUDA 开始面向 RISC-V 目标,焦点是 NVIDIA 工具链和编程模型如何进入新的 CPU/加速器组合。当前信息更适合解读为生态路线,量产兼容性和实测性能仍需后续硬件验证。
EuroHPC Launches 6 Quantum Calls with €119M in Funding32hpcwire.com原文 ↗
EuroHPC 将 1.19 亿欧元拆为六项量子计算征集,面向欧洲量子基础设施和应用生态。公告体现的是资金组织方式和政策优先级;各 call 的细分预算、期限和入选项目仍应以正式征集文件为准。
A Blackstone real estate company exposed SSN digits, DOBs, addresses and more33alexschapiro.com原文 ↗
Beam Living 的 GraphQL 暴露把部分 SSN 数字、出生日期和地址等高敏字段交给了不应获得它们的查询者,问题集中在授权和字段最小化。案例说明即使没有传统 SQL 注入,过宽的 GraphQL schema 也能形成批量个人数据泄露。
IPFS Maintainers Winding Down34ipshipyard.com原文 ↗
IPFS Shipyard 宣布维护者组织逐步收尾,意味着原有资助、协调和维护支持会缩减或转移。协议和网络不会因此自动下线,但社区需要重新分配发布、修复和基础设施责任。
MS Paint and Photos invisibly watermark even locally generated output with GUID35xusheng.dev原文 ↗
逆向工作发现 Paint 与 Photos 在本地生成的输出写入 GUID 隐形水印,追踪标记不依赖云端上传。它把“离线生成就没有遥测/标识”的直觉变成具体格式事实,产品应明确说明标记位置、用途和移除能力。
OpenAI: GPT 5.6 Sol price reduction36developers.openai.com原文 ↗
官方价格表显示 gpt-5.6-sol 促销档短上下文为输入 $2、缓存输入 $0.20、输出 $10/百万 token,长上下文为 $4/$0.40/$15,并注明优惠至少持续到 2026-11-21。页面同时列出标准、Batch、Fast mode 和区域处理加价,实际预算不能只看单一宣传数字。
Anna's Archive Owes $340 Million, Lost Several Domains, but It's Still Online37torrentfreak.com原文 ↗
报道描述 Anna’s Archive 被追索约 3.4 亿美元、多个域名失效,但服务仍通过其他入口在线,呈现版权执行与镜像/域名迁移的持续博弈。债务与在线状态是当时报道的截面,不等同于最终司法结论。
Xiaomi: New CPU matches Apple cores single threaded, much faster multithreaded38twitter.com原文 ↗
社交媒体帖子声称小米新 CPU 单线程接近 Apple 核心、多线程更快,但没有附芯片型号、测试版本或可复现实验配置。它适合作为待核查的性能线索,不应替代完整 benchmark。
博客文章 · Blog Posts
15 项 · 博客文章llm-anthropic 0.2739simonwillison.net原文 ↗
`llm-anthropic` 0.27 跟进 Anthropic `anthropic` Python SDK 1.0.0 的兼容变更,保持 Simon Willison 的 LLM CLI 插件能够使用新客户端。更新本身属于依赖适配,意义在于 SDK 主版本切换时不让命令行工作流断裂。
Your executable is a SQLite database40simonwillison.net原文 ↗
文章把 SQLite 文件布局和 Linux ELF 入口叠在一起,使一个文件既能被 SQLite 打开又能直接执行。技巧展示“数据即程序”的分发形态,难点是头部、偏移和启动路径的格式兼容,而不是改变 SQLite 引擎。
Anger, Anxiety and Agency41lucumr.pocoo.org原文 ↗
Ronacher 把愤怒和焦虑视为暴露边界、放大失控感的信号,再讨论如何通过具体选择恢复 agency。它是对行动感的个人反思,不把情绪简化成应当压制或立即宣泄的对象,也不提供临床诊断。
Your “File” Menu Isn't About Files42adam.farkas.pro原文 ↗
文章将 File 菜单重新解释为文档生命周期菜单:Open 取得工作对象、Save 持久化当前状态、Export 转成交付格式、Print 连接物理输出。按用户意图而非文件系统层级组织命令,能解释许多桌面软件菜单为何在“文件”概念下仍需要不同反馈。
Adding 4 more 2.5GbE interfaces to the GMKtec NucBox G943catskull.net原文 ↗
硬件改造把 NucBox G9 从普通小主机扩成多口 2.5GbE 节点,实际工作围绕扩展通道、供电、机箱空间、散热和驱动稳定性展开。它的价值在于把“接口数量”还原成 PCIe/USB 资源和机械约束的系统问题。
Jabber/XMPP: 25 Years of Digital Independence44gultsch.de原文 ↗
回顾把 XMPP 的 25 年历史与开放协议、可互操作服务器、可迁移身份联系起来,同时承认移动体验、平台竞争和协议复杂度的长期压力。数字独立性在这里不是口号,而是靠多方维护和可替换服务端持续兑现的属性。
IPython is All You Need45nathancooper.io原文 ↗
作者把 IPython 的交互变量、魔术命令、shell 调用和可视化组合成日常计算/脚本主环境,以快速反馈替代频繁切换工具。这个工作法适合探索和小自动化;当代码需要复用、测试和部署时,稳定逻辑仍应抽回模块。
Intent to Ship: JPEG XL46hacks.mozilla.org原文 ↗
Mozilla 的 Intent to Ship 把 JPEG XL 的解码、渐进加载、动画、无损/有损和 HDR 等能力带进 Firefox 产品评审流程。它表明浏览器愿意重新评估格式生态,但“意向发运”与所有平台默认启用之间仍有实现和发布门槛。
Dynamically Naming Servers47arch.dog原文 ↗
文章建议让服务器名由区域、角色、序号和部署上下文动态生成,使日志和监控天然带拓扑语义。代价是重启、迁移和实例替换会挑战名称稳定性,命名策略必须同时定义可读性与关联键。
Micro language implementation: Calcium48nedbatchelder.com原文 ↗
Calcium 用很小的词法器、解析器、AST 和执行器展示语言如何逐层长成,单个特性都能配合示例观察。它的教学贡献是把编译器抽象保持在可读规模,而不是追求生产语言的性能或完整语法。
The changing role of finite-state model checking49ahelwer.ca原文 ↗
文章认为状态空间爆炸削弱了有限状态 model checking 的全局覆盖能力,却没有消除它对协议、控制器和抽象局部的穷尽验证价值。更现实的方向是符号方法、组合抽象与其他验证器协作,把模型检查当成可组合的证据生成器。
Perspec 1.0: A Haskell desktop app for perspective correction of document photos50adriansieber.com原文 ↗
Perspec 1.0 用桌面界面校正文档/收据照片透视,作者以大屏、鼠标和键盘换取比手机扫描器更精细的边缘修正,并偏好无压缩伪影的灰度 PNG。示例输出约 110 kB,Scanner Pro 约 190 kB、iOS 约 300 kB;从构想到 1.0 经过九年迭代。
Building certgrep.sh: a free certificate transparency search engine51haveibeensquatted.com原文 ↗
certgrep.sh 把分散在 Certificate Transparency 日志中的证书、域名和重复记录持续抓取、解析并建索引,做成免费搜索入口。工程重点落在 ingest 的连续性、去重、查询延迟和滥用控制,公开日志本身并不自动解决服务运营成本。
Adding JIT-compilation to a toy interpreter with libgccjit52gcc.gnu.org原文 ↗
GCC 教程先实现只支持整数和递归的栈式 toy VM,再让 libgccjit 为同一字节码生成机器码,用 factorial 展示基本块、条件跳转和调用。它把解释执行与 JIT 共用的语义路径讲清楚,并明确 parser 与 VM 都是教学简化,不是生产动态语言的完整实现。
The text mode lie: why modern TUIs are a nightmare for accessibility53osnews.com原文 ↗
文章指出 Ink、Bubble Tea、tcell 等现代 TUI 把终端从线性数据流变成频繁重绘的二维字符网格,光标跳动会让屏幕阅读器失去上下文。它以 Irssi 使用 VT100 滚动区域作对照,结论是问题来自框架没有利用终端原生语义,而非文字界面先天不可访问。
GitHub 热门 · GitHub Trending
11 项 · GitHub 热门freestylefly/awesome-gpt-image-254github.com原文 ↗
仓库把 GPT Image 2 的提示词工程整理成 500+ 逆向案例和 20+ 工业模板,配套站点可按风格/场景筛选、复制完整 prompt 并回看 GitHub 来源。它是可复用的案例和工作流目录,不包含图像模型本身。
tinyhumansai/openhuman55github.com原文 ↗
OpenHuman 将本地 SQLite 评分记忆树、Obsidian 镜像、durable graph 编排和深度研究放入个人 AI 工作区;README 还列出 100+ OAuth、5,000+ MCP server、90,000+ skills 与最多 80% 的工具输出压缩。项目明确仍是 early beta,数据格式和体验可能快速变化。
VoltAgent/awesome-agent-skills56github.com原文 ↗
集合优先收录 Anthropic、Google Labs、Vercel、Stripe、Cloudflare 等团队实际使用的 skills,并按来源和工具兼容性组织,支持 Claude Code、Codex、Gemini CLI、Cursor 等多种 harness。筛选“真实维护者”是它区别于批量生成 skill 列表的主要机制。
dani-garcia/vaultwarden57github.com原文 ↗
Vaultwarden 用 Rust 实现兼容 Bitwarden Client API 的自托管服务端,覆盖 vault、Send、附件、组织共享、事件日志和多种 MFA。项目推荐容器镜像与反向代理,web vault 依赖 HTTPS/Web Crypto 安全上下文,目标是低资源自托管而非重做官方客户端。
anthropics/claude-plugins-community58github.com原文 ↗
这是社区插件 marketplace 的只读镜像,清单由 Anthropic 内部审核管线每夜同步;插件先过自动安全扫描再进入分发目录。贡献必须走官方提交入口,直接向镜像开 PR 会被关闭,仓库因此承担的是受控索引角色。
apache/maka59github.com原文 ↗
Apache Incubating 的 Maka 通过 Runtime Host 在沙箱内运行工具,并把消息、调用、结果、权限和结束原因写成可恢复执行事实;桌面、TUI/CLI、eval 共用这条路径。它允许从后续 prompt 省略旧输出而不删除记录,但早期 Apple Silicon 构建和数据格式仍可能变动。
proliferate-ai/proliferate60github.com原文 ↗
Proliferate 把 Claude Code、Codex、OpenCode、Cursor、Grok 等原生 harness 放进一个 IDE,每个任务有隔离 worktree、分支、终端和 review 状态。MCP、skills、Computer Use、子 agent 和定时工作流可集中配置,控制平面支持 Docker、云和 air-gapped 自托管。
davepoon/buildwithclaude61github.com原文 ↗
Build with Claude 既是 marketplace 也是目录,仓库维护 117 agents、175 commands、28 hooks、26 skills、51 插件包,并索引 20k+ 社区插件和 4,500+ MCP server。网页重点提供过滤、搜索、文档和安装命令,解决的是发现与装配成本。
laurent22/joplin62github.com原文 ↗
Joplin 以 Markdown 笔记、全文搜索、插件/主题和跨平台客户端提供隐私导向的笔记与待办体验;offline-first 让设备始终保留数据。同步可通过 Nextcloud、Dropbox、OneDrive 或 Joplin Cloud 的端到端加密完成,Evernote/Markdown 导入和浏览器 Web Clipper 扩展了迁移入口。
DefinitelyTyped/DefinitelyTyped63github.com原文 ↗
DefinitelyTyped 维护社区 TypeScript 类型定义,并坚持只为真实使用中的 npm 包接受新定义。README 甚至要求 coding agent 拒绝批量为“最热门未类型化包”发 PR,说明仓库把维护者时间和消费动机纳入技术质量门。
superset-sh/superset64github.com原文 ↗
Superset 在独立 git worktree 中并行运行 CLI agents,提供终端分屏、diff 评论、内置浏览器、远程 workspace、CLI/SDK/MCP 和定时自动化。README 将规模目标写成 100+ agents;其核心不是新模型,而是让并行分支的观察、比较和合并留在一个工作面内。
引用来源 · References
64 条 · 引用- 1 SDAD: Spec-Driven Agentic Development for the AI-Native SDLC. arXiv:2608.20341https://arxiv.org/abs/2608.20341 ↩ 回到正文 · back to text
- 2 Nexus: Depth-Adaptive KV-Cache Splicing and Retrieval-Decoupled Tool Routing for Agentic LLMs on Unified Memory. arXiv:2608.20397https://arxiv.org/abs/2608.20397 ↩ 回到正文 · back to text
- 3 DreamBench-SWE: A Multi-Session Memory-Hygiene Benchmark for Software Agents. arXiv:2608.20664https://arxiv.org/abs/2608.20664 ↩ 回到正文 · back to text
- 4 Utility Under Attack: Agent Memory Poisoning and the Limits of Content Screening and Provenance Ranking. arXiv:2608.21230https://arxiv.org/abs/2608.21230 ↩ 回到正文 · back to text
- 5 Kernhttps://github.com/getkern/kern ↩ 回到正文 · back to text
- 6 PrimeAgentOrchestrator: Memory-Primed Agent Spawning for Personal AI Infrastructure. arXiv:2608.20342https://arxiv.org/abs/2608.20342 ↩ 回到正文 · back to text
- 7 Representation Affects Retrieval: A Case Study of Skill Discovery and Routing in a Multimodal Agent Harness. arXiv:2608.20389https://arxiv.org/abs/2608.20389 ↩ 回到正文 · back to text
- 8 When Retrieval Fails Before It Begins: Structurally Indirect Prerequisite Eviction as a Retention Failure in Agentic Memory. arXiv:2608.20400https://arxiv.org/abs/2608.20400 ↩ 回到正文 · back to text
- 9 Consilience: Conformally Calibrated Communication Control for Hidden-Profile Multi-Agent Reasoning. arXiv:2608.20564https://arxiv.org/abs/2608.20564 ↩ 回到正文 · back to text
- 10 Evaluating Skills, Not Just Agents: Agentic Continuous Evaluation of Skills. arXiv:2608.20614https://arxiv.org/abs/2608.20614 ↩ 回到正文 · back to text
- 11 Weighted Memory Tree: Remembering What Matters for Long-Horizon LLM Agents. arXiv:2608.20631https://arxiv.org/abs/2608.20631 ↩ 回到正文 · back to text
- 12 Calibrating Criterion Revision in LLM Agents: Failure Modes and a Trace-Anchored Protocol. arXiv:2608.20729https://arxiv.org/abs/2608.20729 ↩ 回到正文 · back to text
- 13 Don't Solve, Just Compare: Tiny Advisors for Runtime Intervention in LLM Agents. arXiv:2608.21027https://arxiv.org/abs/2608.21027 ↩ 回到正文 · back to text
- 14 Trustworthy RAG: An Evaluation Agent for Detecting Misinformation and Knowledge Poisoning in Generative AI Systems. arXiv:2608.21095https://arxiv.org/abs/2608.21095 ↩ 回到正文 · back to text
- 15 ClawSentry: A Progressive Multi-Tier Security Monitor for Safeguarding Autonomous LLM Agents. arXiv:2608.21101https://arxiv.org/abs/2608.21101 ↩ 回到正文 · back to text
- 16 Graph Engineering in the Era of LLM Agents: From Individual Intelligence to System Intelligence. arXiv:2608.21156https://arxiv.org/abs/2608.21156 ↩ 回到正文 · back to text
- 17 Transpose Spotify audio and isolate vocals/instruments in realtimehttps://github.com/evanhu1/transposify ↩ 回到正文 · back to text
- 18 Noswooshhttps://github.com/mmathys/noswoosh ↩ 回到正文 · back to text
- 19 WorkBasehttps://github.com/vocso-com/WorkBase ↩ 回到正文 · back to text
- 20 Sloppiehttps://github.com/otsaloma/sloppie ↩ 回到正文 · back to text
- 21 Agent Notifierhttps://notifier.aicrew.in/setup ↩ 回到正文 · back to text
- 22 RepoRoulettehttps://github.com/gojiplus/reporoulette ↩ 回到正文 · back to text
- 23 I built a lite LPU that can do inference on Karpathy's MicroGPThttps://www.lpulite.com ↩ 回到正文 · back to text
- 24 Headless Toolshttps://hdls.tools ↩ 回到正文 · back to text
- 25 PicoMQhttps://picomq.com/ ↩ 回到正文 · back to text
- 26 GlassBoxhttps://glassbox.codecanary.org ↩ 回到正文 · back to text
- 27 LunarBasichttps://lunarbasic.com/ ↩ 回到正文 · back to text
- 28 A Claude Code skill that recovers export-blocked Kindle highlightshttps://github.com/l3a0/claude-plugins ↩ 回到正文 · back to text
- 29 Advancing price-performance for developers with GPT‑5.6 in Kirohttps://openai.com/index/gpt-5-6-in-kiro ↩ 回到正文 · back to text
- 30 LLMs could control their host machines by exploiting inference engineshttps://boydkane.com/essays/llms-could-control-their-host-machines-by-exploiting-inference-engines ↩ 回到正文 · back to text
- 31 Hot Chips 2026: CUDA Targets RISC-Vhttps://chipsandcheese.com/p/hot-chips-2026-cuda-targets-risc ↩ 回到正文 · back to text
- 32 EuroHPC Launches 6 Quantum Calls with €119M in Fundinghttps://www.hpcwire.com/off-the-wire/eurohpc-launches-6-quantum-calls-with-e119m-in-funding/ ↩ 回到正文 · back to text
- 33 A Blackstone real estate company exposed SSN digits, DOBs, addresses and morehttps://alexschapiro.com/security/vulnerability/2026/07/16/beam-living-graphql-data-exposure ↩ 回到正文 · back to text
- 34 IPFS Maintainers Winding Downhttps://ipshipyard.com/blog/2026-the-end-of-ipfs-at-shipyard/ ↩ 回到正文 · back to text
- 35 MS Paint and Photos invisibly watermark even locally generated output with GUIDhttps://xusheng.dev/posts/reversing/mspaint_invisible_watermark/main/ ↩ 回到正文 · back to text
- 36 OpenAI: GPT 5.6 Sol price reductionhttps://developers.openai.com/api/docs/pricing ↩ 回到正文 · back to text
- 37 Anna's Archive Owes $340 Million, Lost Several Domains, but It's Still Onlinehttps://torrentfreak.com/annas-archive-owes-340-million-lost-several-domains-but-its-still-online/ ↩ 回到正文 · back to text
- 38 Xiaomi: New CPU matches Apple cores single threaded, much faster multithreadedhttps://twitter.com/lemire/status/2091894299289874926 ↩ 回到正文 · back to text
- 39 llm-anthropic 0.27https://simonwillison.net/2026/Aug/24/llm-anthropic/ ↩ 回到正文 · back to text
- 40 Your executable is a SQLite databasehttps://simonwillison.net/2026/Aug/24/your-executable-is-a-sqlite-database/ ↩ 回到正文 · back to text
- 41 Anger, Anxiety and Agencyhttps://lucumr.pocoo.org/2026/8/24/anger-anxiety-agency/ ↩ 回到正文 · back to text
- 42 Your “File” Menu Isn't About Fileshttps://adam.farkas.pro/your-file-menu-isnt-about-files/ ↩ 回到正文 · back to text
- 43 Adding 4 more 2.5GbE interfaces to the GMKtec NucBox G9https://catskull.net/adding-4-more-25gbe-to-the-gmktec-nucbox-g9.html ↩ 回到正文 · back to text
- 44 Jabber/XMPP: 25 Years of Digital Independencehttps://gultsch.de/posts/25-years-of-digital-independence/ ↩ 回到正文 · back to text
- 45 IPython is All You Needhttps://nathancooper.io/blog/2026-08-10-ipython-is-all-you-need ↩ 回到正文 · back to text
- 46 Intent to Ship: JPEG XLhttps://hacks.mozilla.org/2026/08/intent-to-ship-jpeg-xl/ ↩ 回到正文 · back to text
- 47 Dynamically Naming Servershttps://arch.dog/bark/dynamically-naming-servers ↩ 回到正文 · back to text
- 48 Micro language implementation: Calciumhttps://nedbatchelder.com/blog/202608/micro_language_implementation_calcium ↩ 回到正文 · back to text
- 49 The changing role of finite-state model checkinghttps://ahelwer.ca/post/2026-08-24-finite-state-future/ ↩ 回到正文 · back to text
- 50 Perspec 1.0: A Haskell desktop app for perspective correction of document photoshttps://adriansieber.com/announcing-perspec-1-0/ ↩ 回到正文 · back to text
- 51 Building certgrep.sh: a free certificate transparency search enginehttps://haveibeensquatted.com/blog/building-certgrep ↩ 回到正文 · back to text
- 52 Adding JIT-compilation to a toy interpreter with libgccjithttps://gcc.gnu.org/onlinedocs/jit/intro/tutorial04.html ↩ 回到正文 · back to text
- 53 The text mode lie: why modern TUIs are a nightmare for accessibilityhttps://www.osnews.com/story/144892/the-text-mode-lie-why-modern-tuis-are-a-nightmare-for-accessibility/ ↩ 回到正文 · back to text
- 54 freestylefly/awesome-gpt-image-2https://github.com/freestylefly/awesome-gpt-image-2 ↩ 回到正文 · back to text
- 55 tinyhumansai/openhumanhttps://github.com/tinyhumansai/openhuman ↩ 回到正文 · back to text
- 56 VoltAgent/awesome-agent-skillshttps://github.com/VoltAgent/awesome-agent-skills ↩ 回到正文 · back to text
- 57 dani-garcia/vaultwardenhttps://github.com/dani-garcia/vaultwarden ↩ 回到正文 · back to text
- 58 anthropics/claude-plugins-communityhttps://github.com/anthropics/claude-plugins-community ↩ 回到正文 · back to text
- 59 apache/makahttps://github.com/apache/maka ↩ 回到正文 · back to text
- 60 proliferate-ai/proliferatehttps://github.com/proliferate-ai/proliferate ↩ 回到正文 · back to text
- 61 davepoon/buildwithclaudehttps://github.com/davepoon/buildwithclaude ↩ 回到正文 · back to text
- 62 laurent22/joplinhttps://github.com/laurent22/joplin ↩ 回到正文 · back to text
- 63 DefinitelyTyped/DefinitelyTypedhttps://github.com/DefinitelyTyped/DefinitelyTyped ↩ 回到正文 · back to text
- 64 superset-sh/supersethttps://github.com/superset-sh/superset ↩ 回到正文 · back to text