BGE-Large Embeddings
BAAI/bge-large-en-v1.5, FP16, 512-token inputs, batch 32.
Primary metric: Embeddings / sec (emb/s)
Reference prompts
Representative prompts for this workload. Exact prompts and harness settings still depend on the cited source for each record.
- Prompt 1
512-token Wikipedia paragraph (English).
- Prompt 2
512-token code comment block (Python).
- Prompt 3
512-token product description (e-commerce).
Reference runtime command
A representative invocation for reproducing this workload class. Source-specific runs may use adjacent runtimes unless the record says otherwise.
text-embeddings-router --model-id BAAI/bge-large-en-v1.5 --dtype float16 --max-batch-tokens 16384Batch 32 of 512-token inputs. Sustained throughput, excluding cold-start. We report embeddings/sec.
Full leaderboard
Every record for Embedding throughput.
No comparable multi-device results
The current selection lacks two devices with matching workload, quantization, context, batch and documented runtime. No speed winner is assigned. Inspect the source-attributed records.
Cite this benchmark
Use this in your paper, blog post, or comparison table.
@misc{myaihardware_embedding-bge-large_2026,
title = {MyAI Bench: BGE-Large Embeddings},
author = {{MyAIHardware Contributors}},
year = {2026},
url = {https://www.myaihardware.com/benchmarks/workload/embedding-bge-large},
note = {Version 1.3, accessed 2026-09-09}
}MyAIHardware Contributors. (2026). MyAI Bench: BGE-Large Embeddings. MyAIHardware. Retrieved 2026-09-09, from https://www.myaihardware.com/benchmarks/workload/embedding-bge-large
MyAIHardware Contributors. "MyAI Bench: BGE-Large Embeddings." MyAIHardware, 2026, https://www.myaihardware.com/benchmarks/workload/embedding-bge-large. Accessed 2026-09-09.
MyAI Bench, BGE-Large Embeddings. MyAIHardware Contributors, 2026. Version 1.3. https://www.myaihardware.com/benchmarks/workload/embedding-bge-large (accessed 2026-09-09).