LLM Models
Browse the world's large language models. Compare parameters, benchmarks, VRAM and more.
This model is based on LLaVA1.5-7b. The model is finetuned with LoRA on OpenCOLE1.0 dataset to generate text layouts.
This model is a fine-tuned version of Qwen2.5-0.5B. It has been trained using TRL.
Overview rinna/youri-7b-instruction-gptq is the quantized model for rinna/youri-7b-instruction using AutoGPTQ. The quantized version is 4x smaller than the original model and thus requires less memory and provides faster inference.
Overview rinna/youri-7b-gptq is the quantized model for rinna/youri-7b using AutoGPTQ. The quantized version is 4x smaller than the original model and thus requires less memory and provides faster inference.
Overview rinna/youri-7b-chat-gptq is the quantized model for rinna/youri-7b-chat using AutoGPTQ. The quantized version is 4x smaller than the original model and thus requires less memory and provides faster inference.
This model is a fine-tuned version of Qwen3-0.6B-Base. It has been trained using TRL.
This model is a fine-tuned version of Llama-3.2-1B. It has been trained using TRL.
This model is a fine-tuned version of Qwen2.5-1.5B. It has been trained using TRL.
Overview This repository provides an English-Japanese bilingual multimodal conversational model like MiniGPT-4 by combining GPT-NeoX model of 3.8 billion parameters and BLIP-2.
Play with the instruction-tuned StarCoderPlus at StarChat-Beta.