Text Generation
GGUF
Italian
Inference Endpoints
Edit model card

QuantFactory/Minerva-3B-base-RAG-GGUF

This is quantized version of DeepMount00/Minerva-3B-base-RAG created using llama.cpp

Model Card for Minerva-3B-base-QA-v1.0

Minerva-3B-base-RAG is a specialized question-answering (QA) model derived through the finetuning of Minerva-3B-base-v1.0. This finetuning was independently conducted to enhance the model's performance for QA tasks, making it ideally suited for use in Retrieval-Augmented Generation (RAG) applications.

Overview


Downloads last month
168
GGUF
Model size
2.89B params
Architecture
llama

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference Examples
Unable to determine this model's library. Check the docs .

Model tree for QuantFactory/Minerva-3B-base-RAG-GGUF

Quantized
(4)
this model

Dataset used to train QuantFactory/Minerva-3B-base-RAG-GGUF