MTG Embeddings v2
Collection
5 items • Updated
How to use philipp-zettl/bge-micro-v2-mtg-v2 with sentence-transformers:
from sentence_transformers import SentenceTransformer
model = SentenceTransformer("philipp-zettl/bge-micro-v2-mtg-v2")
sentences = [
"howl from beyond",
"Title: Howl from Beyond\nCost: {X}{B}\nColors: B\nType: Instant\nDesc: Target creature gets +X/+0 until end of turn.",
"Title: Thundermane Dragon\nCost: {3}{R}\nColors: R\nType: Creature — Dragon\nDesc: Flying\nYou may look at the top card of your library any time.\nYou may cast creature spells with power 4 or greater from the top of your library. If you cast a creature spell this way, it gains haste until end of turn.",
"Title: Mantis Engine\nCost: {5}\nType: Artifact Creature — Insect\nDesc: {2}: This creature gains flying until end of turn. (It can't be blocked except by creatures with flying or reach.)\n{2}: This creature gains first strike until end of turn. (It deals combat damage before creatures without first strike.)"
]
embeddings = model.encode(sentences)
similarities = model.similarity(embeddings, embeddings)
print(similarities.shape)
# [4, 4]This is a sentence-transformers model finetuned from TaylorAI/bge-micro-v2. It maps sentences & paragraphs to a 384-dimensional dense vector space and can be used for retrieval.
SentenceTransformer(
(0): Transformer({'transformer_task': 'feature-extraction', 'modality_config': {'text': {'method': 'forward', 'method_output_name': 'last_hidden_state'}}, 'module_output_name': 'token_embeddings', 'architecture': 'BertModel'})
(1): Pooling({'embedding_dimension': 384, 'pooling_mode': 'mean', 'include_prompt': True})
)
First install the Sentence Transformers library:
pip install -U sentence-transformers
Then you can load this model and run inference.
from sentence_transformers import SentenceTransformer
# Download from the 🤗 Hub
model = SentenceTransformer("philipp-zettl/bge-micro-v2-mtg-v2")
# Run inference
sentences = [
')\nremove a charge counter from this artifact: add one mana of any color',
'Title: Pentad Prism\nCost: {2}\nType: Artifact\nDesc: Sunburst (This artifact enters with a charge counter on it for each color of mana spent to cast it.)\nRemove a charge counter from this artifact: Add one mana of any color.',
"Title: Blockade Runner\nCost: {3}{U}\nColors: U\nType: Creature — Merfolk\nDesc: {U}: This creature can't be blocked this turn.",
]
embeddings = model.encode(sentences)
print(embeddings.shape)
# [3, 384]
# Get the similarity scores for the embeddings
similarities = model.similarity(embeddings, embeddings)
print(similarities)
# tensor([[ 1.0000, 0.6128, -0.0944],
# [ 0.6128, 1.0000, -0.1033],
# [-0.0944, -0.1033, 1.0000]])
sentence_0 and sentence_1| sentence_0 | sentence_1 | |
|---|---|---|
| type | string | string |
| modality | text | text |
| details |
|
|
| sentence_0 | sentence_1 |
|---|---|
flying, vigilance |
Title: Eagle of the Watch |
−3: gain control of target creature or planeswalker until end of turn |
Title: Geyadrone Dihada |
lightning greaves |
Title: Lightning Greaves |
MultipleNegativesRankingLoss with these parameters:{
"scale": 20.0,
"similarity_fct": "cos_sim",
"gather_across_devices": false,
"directions": [
"query_to_doc"
],
"partition_mode": "joint",
"hardness_mode": null,
"hardness_strength": 0.0
}
per_device_train_batch_size: 96fp16: Trueper_device_eval_batch_size: 96multi_dataset_batch_sampler: round_robinper_device_train_batch_size: 96num_train_epochs: 3max_steps: -1learning_rate: 5e-05lr_scheduler_type: linearlr_scheduler_kwargs: Nonewarmup_steps: 0optim: adamw_torch_fusedoptim_args: Noneweight_decay: 0.0adam_beta1: 0.9adam_beta2: 0.999adam_epsilon: 1e-08optim_target_modules: Nonegradient_accumulation_steps: 1average_tokens_across_devices: Truemax_grad_norm: 1label_smoothing_factor: 0.0bf16: Falsefp16: Truebf16_full_eval: Falsefp16_full_eval: Falsetf32: Nonegradient_checkpointing: Falsegradient_checkpointing_kwargs: Nonetorch_compile: Falsetorch_compile_backend: Nonetorch_compile_mode: Noneuse_liger_kernel: Falseliger_kernel_config: Noneuse_cache: Falseneftune_noise_alpha: Nonetorch_empty_cache_steps: Noneauto_find_batch_size: Falselog_on_each_node: Truelogging_nan_inf_filter: Trueinclude_num_input_tokens_seen: nolog_level: passivelog_level_replica: warningdisable_tqdm: Falseproject: huggingfacetrackio_space_id: Nonetrackio_bucket_id: Nonetrackio_static_space_id: Noneper_device_eval_batch_size: 96prediction_loss_only: Trueeval_on_start: Falseeval_do_concat_batches: Trueeval_use_gather_object: Falseeval_accumulation_steps: Noneinclude_for_metrics: []batch_eval_metrics: Falsesave_only_model: Falsesave_on_each_node: Falseenable_jit_checkpoint: Falsepush_to_hub: Falsehub_private_repo: Nonehub_model_id: Nonehub_strategy: every_savehub_always_push: Falsehub_revision: Noneload_best_model_at_end: Falseignore_data_skip: Falserestore_callback_states_from_checkpoint: Falsefull_determinism: Falseseed: 42data_seed: Noneuse_cpu: Falseaccelerator_config: {'split_batches': False, 'dispatch_batches': None, 'even_batches': True, 'use_seedable_sampler': True, 'non_blocking': False, 'gradient_accumulation_kwargs': None}parallelism_config: Nonedataloader_drop_last: Falsedataloader_num_workers: 0dataloader_pin_memory: Truedataloader_persistent_workers: Falsedataloader_prefetch_factor: Noneremove_unused_columns: Truelabel_names: Nonetrain_sampling_strategy: randomlength_column_name: lengthddp_find_unused_parameters: Noneddp_bucket_cap_mb: Noneddp_broadcast_buffers: Falseddp_static_graph: Noneddp_backend: Noneddp_timeout: 1800fsdp: Nonefsdp_config: Nonedeepspeed: Nonedebug: []skip_memory_metrics: Truedo_predict: Falseresume_from_checkpoint: Nonewarmup_ratio: Nonelocal_rank: -1prompts: Nonebatch_sampler: batch_samplermulti_dataset_batch_sampler: round_robinrouter_mapping: {}learning_rate_mapping: {}| Epoch | Step | Training Loss |
|---|---|---|
| 0.0225 | 500 | 1.3234 |
| 0.0450 | 1000 | 0.6360 |
| 0.0675 | 1500 | 0.5297 |
| 0.0900 | 2000 | 0.4882 |
| 0.1125 | 2500 | 0.4683 |
| 0.1350 | 3000 | 0.4477 |
| 0.1575 | 3500 | 0.4327 |
| 0.1800 | 4000 | 0.4275 |
| 0.2025 | 4500 | 0.4095 |
| 0.2250 | 5000 | 0.3984 |
| 0.2475 | 5500 | 0.4051 |
| 0.2701 | 6000 | 0.3964 |
| 0.2926 | 6500 | 0.3913 |
| 0.3151 | 7000 | 0.3880 |
| 0.3376 | 7500 | 0.3858 |
| 0.3601 | 8000 | 0.3730 |
| 0.3826 | 8500 | 0.3770 |
| 0.4051 | 9000 | 0.3753 |
| 0.4276 | 9500 | 0.3696 |
| 0.4501 | 10000 | 0.3699 |
| 0.4726 | 10500 | 0.3642 |
| 0.4951 | 11000 | 0.3685 |
| 0.5176 | 11500 | 0.3569 |
| 0.5401 | 12000 | 0.3632 |
| 0.5626 | 12500 | 0.3624 |
| 0.5851 | 13000 | 0.3533 |
| 0.6076 | 13500 | 0.3567 |
| 0.6301 | 14000 | 0.3564 |
| 0.6526 | 14500 | 0.3590 |
| 0.6751 | 15000 | 0.3579 |
| 0.6976 | 15500 | 0.3477 |
| 0.7201 | 16000 | 0.3424 |
| 0.7426 | 16500 | 0.3395 |
| 0.7651 | 17000 | 0.3441 |
| 0.7876 | 17500 | 0.3472 |
| 0.8102 | 18000 | 0.3421 |
| 0.8327 | 18500 | 0.3458 |
| 0.8552 | 19000 | 0.3461 |
| 0.8777 | 19500 | 0.3410 |
| 0.9002 | 20000 | 0.3449 |
| 0.9227 | 20500 | 0.3422 |
| 0.9452 | 21000 | 0.3412 |
| 0.9677 | 21500 | 0.3417 |
| 0.9902 | 22000 | 0.3396 |
| 1.0127 | 22500 | 0.3304 |
| 1.0352 | 23000 | 0.3358 |
| 1.0577 | 23500 | 0.3400 |
| 1.0802 | 24000 | 0.3322 |
| 1.1027 | 24500 | 0.3390 |
| 1.1252 | 25000 | 0.3395 |
| 1.1477 | 25500 | 0.3359 |
| 1.1702 | 26000 | 0.3323 |
| 1.1927 | 26500 | 0.3309 |
| 1.2152 | 27000 | 0.3364 |
| 1.2377 | 27500 | 0.3331 |
| 1.2602 | 28000 | 0.3326 |
| 1.2827 | 28500 | 0.3340 |
| 1.3052 | 29000 | 0.3353 |
| 1.3278 | 29500 | 0.3297 |
| 1.3503 | 30000 | 0.3226 |
| 1.3728 | 30500 | 0.3280 |
| 1.3953 | 31000 | 0.3241 |
| 1.4178 | 31500 | 0.3305 |
| 1.4403 | 32000 | 0.3246 |
| 1.4628 | 32500 | 0.3265 |
| 1.4853 | 33000 | 0.3304 |
| 1.5078 | 33500 | 0.3198 |
| 1.5303 | 34000 | 0.3270 |
| 1.5528 | 34500 | 0.3231 |
| 1.5753 | 35000 | 0.3287 |
| 1.5978 | 35500 | 0.3262 |
| 1.6203 | 36000 | 0.3232 |
| 1.6428 | 36500 | 0.3303 |
| 1.6653 | 37000 | 0.3296 |
| 1.6878 | 37500 | 0.3250 |
| 1.7103 | 38000 | 0.3258 |
| 1.7328 | 38500 | 0.3220 |
| 1.7553 | 39000 | 0.3237 |
| 1.7778 | 39500 | 0.3267 |
| 1.8003 | 40000 | 0.3154 |
| 1.8228 | 40500 | 0.3265 |
| 1.8454 | 41000 | 0.3218 |
| 1.8679 | 41500 | 0.3236 |
| 1.8904 | 42000 | 0.3236 |
| 1.9129 | 42500 | 0.3200 |
| 1.9354 | 43000 | 0.3209 |
| 1.9579 | 43500 | 0.3205 |
| 1.9804 | 44000 | 0.3262 |
| 2.0029 | 44500 | 0.3141 |
| 2.0254 | 45000 | 0.3182 |
| 2.0479 | 45500 | 0.3170 |
| 2.0704 | 46000 | 0.3185 |
| 2.0929 | 46500 | 0.3192 |
| 2.1154 | 47000 | 0.3136 |
| 2.1379 | 47500 | 0.3155 |
| 2.1604 | 48000 | 0.3154 |
| 2.1829 | 48500 | 0.3193 |
| 2.2054 | 49000 | 0.3235 |
| 2.2279 | 49500 | 0.3145 |
| 2.2504 | 50000 | 0.3217 |
| 2.2729 | 50500 | 0.3157 |
| 2.2954 | 51000 | 0.3201 |
| 2.3179 | 51500 | 0.3171 |
| 2.3404 | 52000 | 0.3276 |
| 2.3629 | 52500 | 0.3179 |
| 2.3855 | 53000 | 0.3095 |
| 2.4080 | 53500 | 0.3125 |
| 2.4305 | 54000 | 0.3232 |
| 2.4530 | 54500 | 0.3165 |
| 2.4755 | 55000 | 0.3169 |
| 2.4980 | 55500 | 0.3151 |
| 2.5205 | 56000 | 0.3157 |
| 2.5430 | 56500 | 0.3197 |
| 2.5655 | 57000 | 0.3185 |
| 2.5880 | 57500 | 0.3143 |
| 2.6105 | 58000 | 0.3190 |
| 2.6330 | 58500 | 0.3187 |
| 2.6555 | 59000 | 0.3159 |
| 2.6780 | 59500 | 0.3120 |
| 2.7005 | 60000 | 0.3119 |
| 2.7230 | 60500 | 0.3172 |
| 2.7455 | 61000 | 0.3160 |
| 2.7680 | 61500 | 0.3232 |
| 2.7905 | 62000 | 0.3155 |
| 2.8130 | 62500 | 0.3191 |
| 2.8355 | 63000 | 0.3200 |
| 2.8580 | 63500 | 0.3184 |
| 2.8805 | 64000 | 0.3216 |
| 2.9031 | 64500 | 0.3127 |
| 2.9256 | 65000 | 0.3129 |
| 2.9481 | 65500 | 0.3125 |
| 2.9706 | 66000 | 0.3169 |
| 2.9931 | 66500 | 0.3146 |
@inproceedings{reimers-2019-sentence-bert,
title = "Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks",
author = "Reimers, Nils and Gurevych, Iryna",
booktitle = "Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing",
month = "11",
year = "2019",
publisher = "Association for Computational Linguistics",
url = "https://arxiv.org/abs/1908.10084",
}
@misc{oord2019representationlearningcontrastivepredictive,
title={Representation Learning with Contrastive Predictive Coding},
author={Aaron van den Oord and Yazhe Li and Oriol Vinyals},
year={2019},
eprint={1807.03748},
archivePrefix={arXiv},
primaryClass={cs.LG},
url={https://arxiv.org/abs/1807.03748},
}
Base model
TaylorAI/bge-micro-v2