SAVRN
Search Contact SAVRN

Open-weight model · Feature extraction

dragon-multiturn-context-encoder

by NVIDIA nvidia/dragon-multiturn-context-encoder

dragon-multiturn-context-encoder is an open-weight model for feature extraction from NVIDIA, released under other. It has 512-token context. Its published files total 438.2 MB. It draws 1M downloads a month.

We introduce Dragon-multiturn, a retriever specifically designed for the conversational QA scenario. It can handle conversational query which combine dialogue history with the current query. It is built on top of the Dragon retriever.

Parameters—
Context512
Weights438.0 MB
Licenseother
AccessOpen weights
Monthly Downloads1M

Model Card

We introduce Dragon-multiturn, a retriever specifically designed for the conversational QA scenario. It can handle conversational query which combine dialogue history with the current query. It is built on top of the Dragon retriever. The details of Dragon-multiturn can be found in here. Please note that Dragon-multiturn is a dual encoder consisting of a query encoder and a context encoder. This repository is only for the context encoder of Dragon-multiturn for getting the context embeddings, and you also need the query encoder to get query embeddings, which can be found here. Both query encoder and context encoder share the same tokenizer. Retrieval results across five multi-turn QA…

Excerpt from the card by NVIDIA, licensed other.

Configuration

Architecture
BertModel
Context length (tokens)
512
Layers
12
Hidden size
768
Feed-forward size
3,072
Attention heads
12
Vocabulary size
30,522
Stored precision
float32
Model type
bert

Identity and Version

Repository
nvidia/dragon-multiturn-context-encoder
Publisher
NVIDIA
Task
Feature extraction
Modality
Text
Library
transformers
Parameters
Not stated by the source
Languages
en
Revision
14ba90bb7472741e2a12ac4c0455110966c3f69d
First published
2024-04-30
Last updated
2024-05-24

Files and Weights

7 files, 438.2 MB in total. The weights are 1 file totalling 438.0 MB in bin.

Weights1 file · 438.0 MB
Configuration2 files · 789 B
Tokenizer2 files · 231.5 KB
Documentation1 file · 9.0 KB
Repository1 file · 1.5 KB
Every file
FileTypeSizeSHA-256
pytorch_model.binWeights438.0 MB dfae6270ccce
config.jsonConfiguration677 B —
special_tokens_map.jsonConfiguration112 B —
README.mdDocumentation9.0 KB —
.gitattributesRepository1.5 KB —
tokenizer_config.jsonTokenizer28 B —
vocab.txtTokenizer231.5 KB —

License and Download

License
other
Access
Open weights, no gate
Download size
438.0 MB
Download from NVIDIA

Released by NVIDIA through its official repository on Hugging Face.

Built From

  • Described by arXiv:2302.07452
  • Described by arXiv:2401.10225

Memory Requirements

PrecisionWeights in memory
As published438.0 MB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About dragon-multiturn-context-encoder

What license is dragon-multiturn-context-encoder released under?

other, as its publisher declares it. Read the license text before commercial use.

What is dragon-multiturn-context-encoder's context length?

512 tokens, from the maximum position embeddings in its published configuration.

Similar Models

Model · Feature extraction

all-MiniLM-L6-v2

Joshua

https://huggingface.co/sentence-transformers/all-MiniLM-L6-v2 with ONNX weights to be compatible with Transformers.js. If you haven't already, you can install the Transformers.js JavaScript library from NPM using: You can then use the model to compute embeddings like this: You can convert this Tensor to a nested JavaScript array using.tolist(): Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using Optimum and structuring your repo like this one (with ONNX weights located in a subfolder named onnx).

Open weights apache-2.0 512 tokens transformers.js

Model · Feature extraction

bge-base-en-v1.5

Joshua

https://huggingface.co/BAAI/bge-base-en-v1.5 with ONNX weights to be compatible with Transformers.js. If you haven't already, you can install the Transformers.js JavaScript library from NPM using: You can then use the model to compute embeddings, as follows: You can also use the model for retrieval. For example: Note: Having a separate repo for ONNX weights is intended to be a temporary solution until WebML gains more traction. If you would like to make your models web-ready, we recommend converting to ONNX using Optimum and structuring your repo like this one (with ONNX weights located in a subfolder named onnx).

Open weights mit 512 tokens transformers.js

Model · Feature extraction

clap-htsat-unfused

LAION eV

The abstract of the paper states that: You can use this model for zero shot audio classification or extracting audio and/or textual features. You can also get the audio and text embeddings using ClapModel If you are using this model for your work, please consider citing the original paper

Open weights apache-2.0 514 tokens transformers

For more details please refer to our Github: FlagEmbedding. If you are looking for a model that supports more languages, longer texts, and other retrieval methods, you can try using bge-m3. FlagEmbedding focuses on retrieval-augmented LLMs, consisting of the following projects currently: - 1/30/2024: Release BGE-M3, a new member to BGE model series! M3 stands for Multi-linguality (100+ languages), Multi-granularities (input length up to 8192), Multi-Functionality (unification of dense, lexical, multi-vec/colbert retrieval). It is the first embedding model which supports all three retrieval methods, achieving new SOTA on multi-lingual (MIRACL) and cross-lingual (MKQA) benchmarks. Technical…

Open weights mit 512 tokens sentence-transformers

Model · Feature extraction

wavlm-large

Microsoft

The large model pretrained on 16kHz sampled speech audio. When using the model, make sure that your speech input is also sampled at 16kHz. Note: This model does not have a tokenizer as it was pretrained on audio alone. In order to use this model speech recognition, a tokenizer should be created and the model should be fine-tuned on labeled text data. Check out this blog for more in-detail explanation of how to fine-tune the model. - 60,000 hours of Libri-Light - 10,000 hours of GigaSpeech - 24,000 hours of VoxPopuli Authors: Sanyuan Chen, Chengyi Wang, Zhengyang Chen, Yu Wu, Shujie Liu, Zhuo Chen, Jinyu Li, Naoyuki Kanda, Takuya Yoshioka, Xiong Xiao, Jian Wu, Long Zhou, Shuo Ren, Yanmin…

Open weights transformers