跳到正文
原文
arXiv:cs.AI(全量分类)· Aubrey M. Brueckner, Darshil Patel, Yuhuan He, Timothy Kassis·· 5 小时前AI 评分56

K-Dense BYOK:本地运行的开源 AI 科研助手论文发布

K-Dense BYOK: An Open-Source AI Research Assistant That Runs Locally and Keeps a Hash-Chained Lab Notebook

AI 导读

作者团队发布 K-Dense BYOK,一个免费开源、在研究者本地电脑运行的 AI 科研助手,MIT 许可,代码见 https://arxiv.org/abs/2610.00074。

正文

View PDF HTML (experimental)

Abstract:K-Dense BYOK (bring your own keys) is a free, open-source AI research assistant for scientists in any field that runs on the researcher's own computer. The researcher supplies access to a model of their choice, hosted or running locally, and the application supplies everything else: a place for the work to run, a layer of scientific scaffolding, and a complete record. Each project is an ordinary folder, so the data, the code, the results, and the record stay on a machine the researcher administers and can be read years later without the application. Three things separate it from a chat assistant or a general-purpose coding agent. It ships a library of written scientific procedures, guided workflow templates, catalogs of where research data can be found, and reviewer and writer roles the agent can hand work to. It keeps a Living Lab Notebook whose entries link into an argument and are added to but never erased. And it records what happened by watching what the agent does rather than by taking the agent's word for it, in a log the agent has no tool that can write to. That choice targets the most common failure, model overclaiming, in our earlier benchmark of nine frontier models, by making claims checkable rather than preventing them. On twenty interdisciplinary research prompts, scored under a rubric fixed in advance, K-Dense BYOK led two managed platforms on both scientific quality and research execution. Its deliverables were the only ones that recorded the software they ran in, and the only ones that usually arrived with a command that regenerates the results. One of the managed platforms ran the same frontier model and supplied neither. Those environment records were files the agent wrote, not part of the observed log, which does not yet capture the software environment itself. The code is available under the MIT license at this https URL.
Comments: 38 pages, 8 figures plus a graphical abstract; includes benchmark prompts, scoring rubric, and per-prompt scores. Code: this https URL
Subjects: Artificial Intelligence (cs.AI)
Cite as: arXiv:2610.00074 [cs.AI]
  (or arXiv:2610.00074v1 [cs.AI] for this version)
  https://doi.org/10.48550/arXiv.2610.00074

arXiv-issued DOI via DataCite (pending registration)

Submission history

From: Timothy Kassis [view email]
[v1] Fri, 4 Sep 2026 23:26:54 UTC (688 KB)

来源:arXiv:cs.AI(全量分类) · arxiv.org