Add new SentenceTransformer model.

Browse files

Files changed (5) hide show

README.md +124 -38
config.json +3 -3
config_sentence_transformers.json +5 -4
model.safetensors +2 -2
modules.json +14 -0

README.md CHANGED Viewed

@@ -1,58 +1,144 @@
 ---
-license: mit
-base_model: WhereIsAI/UAE-Large-V1
 tags:
-- generated_from_trainer
-model-index:
-- name: pre-UAE-Medical-Large-V1
-  results: []
 ---
-<!-- This model card has been generated automatically according to the information the Trainer had access to. You
-should probably proofread and complete it, then remove this comment. -->
-[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>]()
-# pre-UAE-Medical-Large-V1
-This model is a fine-tuned version of [WhereIsAI/UAE-Large-V1](https://huggingface.co/WhereIsAI/UAE-Large-V1) on an unknown dataset.
-## Model description
-More information needed
-## Intended uses & limitations
-More information needed
-## Training and evaluation data
-More information needed
-## Training procedure
-### Training hyperparameters
-The following hyperparameters were used during training:
-- learning_rate: 1e-06
-- train_batch_size: 32
-- eval_batch_size: 8
-- seed: 42
-- distributed_type: multi-GPU
-- num_devices: 2
-- total_train_batch_size: 64
-- total_eval_batch_size: 16
-- optimizer: Adam with betas=(0.9,0.999) and epsilon=1e-08
-- lr_scheduler_type: linear
-- lr_scheduler_warmup_steps: 50
-- num_epochs: 1
-### Training results
-### Framework versions
-- Transformers 4.42.3
-- Pytorch 2.3.0+cu121
-- Datasets 2.19.1
-- Tokenizers 0.19.1

 ---
+base_model: WhereIsAI/pre-UAE-Medical-Large-V1
+datasets: []
+language: []
+library_name: sentence-transformers
+pipeline_tag: sentence-similarity
 tags:
+- sentence-transformers
+- sentence-similarity
+- feature-extraction
+widget: []
 ---
+# SentenceTransformer based on WhereIsAI/pre-UAE-Medical-Large-V1
+This is a [sentence-transformers](https://www.SBERT.net) model finetuned from [WhereIsAI/pre-UAE-Medical-Large-V1](https://huggingface.co/WhereIsAI/pre-UAE-Medical-Large-V1). It maps sentences & paragraphs to a 1024-dimensional dense vector space and can be used for semantic textual similarity, semantic search, paraphrase mining, text classification, clustering, and more.
+## Model Details
+### Model Description
+- **Model Type:** Sentence Transformer
+- **Base model:** [WhereIsAI/pre-UAE-Medical-Large-V1](https://huggingface.co/WhereIsAI/pre-UAE-Medical-Large-V1) <!-- at revision c989d8965d489e9a6e873eabce06e6ef6f2a0188 -->
+- **Maximum Sequence Length:** 512 tokens
+- **Output Dimensionality:** 1024 tokens
+- **Similarity Function:** Cosine Similarity
+<!-- - **Training Dataset:** Unknown -->
+<!-- - **Language:** Unknown -->
+<!-- - **License:** Unknown -->
+### Model Sources
+- **Documentation:** [Sentence Transformers Documentation](https://sbert.net)
+- **Repository:** [Sentence Transformers on GitHub](https://github.com/UKPLab/sentence-transformers)
+- **Hugging Face:** [Sentence Transformers on Hugging Face](https://huggingface.co/models?library=sentence-transformers)
+### Full Model Architecture
+```
+SentenceTransformer(
+  (0): Transformer({'max_seq_length': 512, 'do_lower_case': False}) with Transformer model: BertModel
+  (1): Pooling({'word_embedding_dimension': 1024, 'pooling_mode_cls_token': True, 'pooling_mode_mean_tokens': False, 'pooling_mode_max_tokens': False, 'pooling_mode_mean_sqrt_len_tokens': False, 'pooling_mode_weightedmean_tokens': False, 'pooling_mode_lasttoken': False, 'include_prompt': True})
+)
+```
+## Usage
+### Direct Usage (Sentence Transformers)
+First install the Sentence Transformers library:
+```bash
+pip install -U sentence-transformers
+```
+Then you can load this model and run inference.
+```python
+from sentence_transformers import SentenceTransformer
+# Download from the 🤗 Hub
+model = SentenceTransformer("WhereIsAI/pre-UAE-Medical-Large-V1")
+# Run inference
+sentences = [
+    'The weather is lovely today.',
+    "It's so sunny outside!",
+    'He drove to the stadium.',
+]
+embeddings = model.encode(sentences)
+print(embeddings.shape)
+# [3, 1024]
+# Get the similarity scores for the embeddings
+similarities = model.similarity(embeddings, embeddings)
+print(similarities.shape)
+# [3, 3]
+```
+<!--
+### Direct Usage (Transformers)
+<details><summary>Click to see the direct usage in Transformers</summary>
+</details>
+-->
+<!--
+### Downstream Usage (Sentence Transformers)
+You can finetune this model on your own dataset.
+<details><summary>Click to expand</summary>
+</details>
+-->
+<!--
+### Out-of-Scope Use
+*List how the model may foreseeably be misused and address what users ought not to do with the model.*
+-->
+<!--
+## Bias, Risks and Limitations
+*What are the known or foreseeable issues stemming from this model? You could also flag here known failure cases or weaknesses of the model.*
+-->
+<!--
+### Recommendations
+*What are recommendations with respect to the foreseeable issues? For example, filtering explicit content.*
+-->
+## Training Details
+### Framework Versions
+- Python: 3.10.12
+- Sentence Transformers: 3.0.1
+- Transformers: 4.42.3
+- PyTorch: 2.3.0+cu121
+- Accelerate: 0.30.1
+- Datasets: 2.19.1
+- Tokenizers: 0.19.1
+## Citation
+### BibTeX
+<!--
+## Glossary
+*Clearly define terms in order to be accessible across audiences.*
+-->
+<!--
+## Model Card Authors
+*Lists the people who create the model card, providing recognition and accountability for the detailed work that goes into its construction.*
+-->
+<!--
+## Model Card Contact
+*Provides a way for people who have updates to the Model Card, suggestions, or questions, to contact the Model Card authors.*
+-->

config.json CHANGED Viewed

@@ -1,7 +1,7 @@
 {
-  "_name_or_path": "WhereIsAI/UAE-Large-V1",
   "architectures": [
-    "BertForMaskedLM"
   ],
   "attention_probs_dropout_prob": 0.1,
   "classifier_dropout": null,
@@ -18,7 +18,7 @@
   "num_hidden_layers": 24,
   "pad_token_id": 0,
   "position_embedding_type": "absolute",
-  "torch_dtype": "float32",
   "transformers_version": "4.42.3",
   "type_vocab_size": 2,
   "use_cache": false,

 {
+  "_name_or_path": "WhereIsAI/pre-UAE-Medical-Large-V1",
   "architectures": [
+    "BertModel"
   ],
   "attention_probs_dropout_prob": 0.1,
   "classifier_dropout": null,
   "num_hidden_layers": 24,
   "pad_token_id": 0,
   "position_embedding_type": "absolute",
+  "torch_dtype": "float16",
   "transformers_version": "4.42.3",
   "type_vocab_size": 2,
   "use_cache": false,

config_sentence_transformers.json CHANGED Viewed

@@ -1,9 +1,10 @@
 {
   "__version__": {
-    "sentence_transformers": "2.5.1",
-    "transformers": "4.37.0",
-    "pytorch": "2.1.0+cu121"
   },
   "prompts": {},
-  "default_prompt_name": null
 }

 {
   "__version__": {
+    "sentence_transformers": "3.0.1",
+    "transformers": "4.42.3",
+    "pytorch": "2.3.0+cu121"
   },
   "prompts": {},
+  "default_prompt_name": null,
+  "similarity_fn_name": null
 }

model.safetensors CHANGED Viewed

@@ -1,3 +1,3 @@
 version https://git-lfs.github.com/spec/v1
-oid sha256:ddba6fbf557171bcf83b02d009ef0b12c68f8778f08b0ee545f2cfd556a2c334
-size 1340745016

 version https://git-lfs.github.com/spec/v1
+oid sha256:790f6e90679346a3155d57a008e8ac4d3efb79c43123f6573bbea0c8bad1cec2
+size 670328392

modules.json ADDED Viewed

	@@ -0,0 +1,14 @@

+[
+  {
+    "idx": 0,
+    "name": "0",
+    "path": "",
+    "type": "sentence_transformers.models.Transformer"
+  },
+  {
+    "idx": 1,
+    "name": "1",
+    "path": "1_Pooling",
+    "type": "sentence_transformers.models.Pooling"
+  }
+]