Llama-3.2-1B-Instruct, with domain adapted pretraining (DAPT), also called Continuous Pre-training (CPT) on a Dutch medical corpus.

Training for one full epoch, with a 256 batch size, maximally 768 sequence length and a linear-cosine schedule (details follow..).

This model will be further pre-trained on 5 million cardiology records from the UMCU.

The perplexity was around 5 on the validation set.

Safetensors

Model size

1.5B params

Tensor type

BF16

Inference Providers NEW

This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for UMCU/CardioLlama.nl

Base model

Finetuned

(1056)

this model

Quantizations

UMCU
/

CardioLlama.nl