跳到正文

#RAG

今日 1 条
今天10月1日周四
  1. Aravind Srinivas48

    我们正在开源我们最先进的上下文嵌入模型,它在 turbopuffer 的 context-bench 中表现最佳。

    引用Perplexity@perplexity_ai

    We built a new way to train contextual embedding models, which encode each chunk of a document with the whole document in view. pplx-embed-v2-context-9b-preview sets a new state of the art on ConTEB and @turbopuffer's new, privately held context-bench. https://www.perplexity.ai/hub/blog/contextual-embedding-beyond-the-gold-passage

9月30日周三
  1. Liquid AI 模型与工程博客(网页)53

    Liquid AI 发布 LFM2-ColBERT-350M 多语言晚交互检索模型

    Liquid AI 发布 LFM2-ColBERT-350M,一个基于 LFM2 骨干的晚交互检索模型,支持用一种语言存储文档、用其他语言查询检索。在扩展的 NanoBEIR 多语言基准上,其跨语言检索能力显著优于 150M 参数的 GTE-ModernColBERT-v1,尤其在德语、阿拉伯语、韩语和日语上;尽管模型体积更大,查询和文档编码吞吐与对方相当。

  2. Baseten Base Labs:模型研究55

    Baseten 介绍 Google 300M 参数嵌入模型 EmbeddingGemma 及其调用示例

    EmbeddingGemma 是 Google 基于 Gemma3 架构的 3 亿参数文本嵌入模型,在 MTEB 基准上为不到 5 亿参数的嵌入模型中质量最高,支持 100 多种语言、输入最长 2k token。页面给出使用 baseten_performance_client 的完整代码示例,包括各任务的 prompt 前缀和通过 PerformanceClient 调用生成嵌入的方法。

9月3日周四
5月28日周三