LLMs and models

SigLIP

Reference entryEmbedding and vision models
Reference entrySigLIP

SigLIP is Google’s vision-language representation model, aligning images and text with a sigmoid-based training objective for retrieval and classification.

Where it sits