shi-labs/probe_seg_llava-1.5-pt-0.5ift

🤗 Hugging Face sourceimage-text-to-textapache-2.09.4B params38 GBsafetensors✓ 8 checksumsupdated today
Submit in one command

Run it next to your model folder. It makes the torrent, checks your files against Hugging Face, and submits it. You just start seeding and paste your key from your account. It only reads your files and never changes them. Read the script first if you like.

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo shi-labs/probe_seg_llava-1.5-pt-0.5ift ./model-folder
Needs a seeder →

probe_seg_llava-1.5-pt-0.5ift

This model checkpoint contains the seg probes for CLIP-ConvNeXT-XXL Llama-3-8b based LLaVA-1.5 model after the PT stage and 50% of the IFT stage, i.e., trained on the LLaVA-558K and 50% of the LLaVA-665K datasets. Please refer to documentation for more details.

Citation

If you found our work useful in your research, please consider starring ⭐ us on GitHub and citing 📚 us in your research!

@article{jain2024ola_vlm,
    title={{OLA-VLM: Elevating Visual Perception in Multimodal LLMs with Auxiliary Embedding Distillation}},
    author={Jitesh Jain and Zhengyuan Yang and Humphrey Shi and Jianfeng Gao and Jianwei Yang},
    journal={arXiv},
    year={2024}
}