Quantized versions of Prometheus 2 - an alternative of GPT-4 evaluation when doing fine-grained evaluation of an underlying LLM.
Seva Leonov
vsevolodl
AI & ML interests
ML Ops, scalable Inference, quantization, long context, fine-tuning
Organizations
models 18
vsevolodl/qwen3-embed-v4-bashar_docs-192-hn
Sentence Similarity • 0.6B • Updated • 15
vsevolodl/qwen3-embed-v4-bashar_docs-192
Sentence Similarity • 0.6B • Updated • 19
vsevolodl/qwen3-embed-v3-bashar_docs-192
Sentence Similarity • 0.6B • Updated • 62
vsevolodl/qwen3-embed-v3-bashar_docs-192-hn
Sentence Similarity • 0.6B • Updated • 55
vsevolodl/Llama-3-8B-Instruct-Gradient-4194k-GGUF
8B • Updated • 129 • 1
vsevolodl/Yi-1.5-34B-Chat-GGUF
34B • Updated • 57 • 1
vsevolodl/falcon-11B-GGUF
11B • Updated • 144 • 1
vsevolodl/Yi-1.5-9B-Chat-GGUF
9B • Updated • 99 • 1
vsevolodl/Llama-3-70B-Instruct-Gradient-1048k-GGUF
Text Generation • 71B • Updated • 163 • 3
vsevolodl/Llama-3-8B-Instruct-Gradient-1048k-GGUF
Text Generation • 8B • Updated • 278 • 1