MODELS.REALIGNSINC.COM

Local AI models for RINC AI Local V2

Browse practical GGUF models for local chat, vision and image-generation workflows — from compact 0.5B models to large 120B-class deployments.

✓ Original source repositories ✓ Hardware guidance ✓ Local-first workflows
RINC AI Local V2 model ecosystem
How downloads work

models.realignsinc.com is a curated catalog. Download buttons point directly to original Hugging Face model files or repositories; Realigns Inc. does not re-host third-party weights. RAM figures are practical planning ranges, not guarantees — context size, runtime, GPU offload and concurrent workloads change memory use.

FEATURED COMPACT MODELS

Five strong starting points for local AI

Compact models for 4–8 GB and entry-level local workstations, plus RINC's own 0.5B repository.

Realigns Inc.0.5B

Realigns Core V5 GGUF

Lightweight local chat, business assistance and compact offline deployments.

FORMATGGUF
PLANNING RAM4–8 GB

Realigns model repository. Select the quantization file that matches your runtime and hardware.

Alibaba Qwen3B

Qwen2.5 3B Instruct

General chat, multilingual work, drafting and compact local assistance.

FORMATGGUF Q4_K_M
PLANNING RAM6–8 GB

Direct Q4_K_M download.

Meta3B

Llama 3.2 3B Instruct

General assistant, summarization and local productivity.

FORMATGGUF Q4_K_M
PLANNING RAM6–8 GB

Approx. 2.02 GB model file.

Microsoft3.8B

Phi-3.5 Mini Instruct

Compact reasoning, code assistance and local productivity.

FORMATGGUF Q4_K_M
PLANNING RAM6–8 GB

Approx. 2.39 GB model file.

Google4B

Gemma 3 4B IT

Text plus image understanding with a compatible multimodal runtime.

FORMATGGUF Q4_K_M
PLANNING RAM8 GB+

Approx. 2.49 GB model file; multimodal use can require projector assets.

TEXT LLM

From 3B compact models to GPT-OSS 120B

Use the model size, quantization and context length that fit your actual machine. Larger is not automatically better for every workflow.

ModelBrandSizeQuantPlanning RAMUseLinks
Qwen2.5 3B Instruct Alibaba Qwen 3B Q4_K_M 6–8 GB Compact general assistant DownloadModel Page
Llama 3.2 3B Instruct Meta 3B Q4_K_M 6–8 GB General-purpose local assistant DownloadModel Page
Phi-3.5 Mini Instruct Microsoft 3.8B Q4_K_M 6–8 GB Compact reasoning and coding DownloadModel Page
Gemma 3 4B IT Google 4B Q4_K_M 8 GB+ Text and vision capable family DownloadModel Page
Mistral 7B Instruct v0.3 Mistral AI 7B Q4_K_M 8–12 GB Balanced local assistant DownloadModel Page
Gemma 3 12B IT Google 12B Q4_K_M 16 GB+ Higher-quality text / multimodal workflows DownloadModel Page
Qwen2.5 14B Instruct Alibaba Qwen 14B Q4_K_M 16 GB+ Stronger reasoning and business workflows Model Page
Qwen2.5 32B Instruct Alibaba Qwen 32B Q4_K_M 32 GB+ Large local reasoning model DownloadModel Page
GPT-OSS 120B OpenAI 120B MoE Q4_K_M 96 GB+ planning range Very large local model; repository uses split files for large quantizations Model Page
GPT-OSS 120B: the cited Q4_K_M conversion is roughly 62.8 GB and is split into multiple files. Use the repository instructions rather than expecting a single browser download file.
VISION LANGUAGE MODELS

Five multimodal models for image understanding

VLMs analyze images and produce text. They do not replace a dedicated image generator.

Hugging Face2.2B

SmolVLM2 2.2B Instruct

Lightweight image understanding and multimodal assistant workflows.

PLANNING RAM6–8 GB
TYPEVLM / GGUF

Use a compatible llama.cpp build; multimodal models may also require projector files.

Alibaba Qwen3B

Qwen2.5-VL 3B Instruct

Compact OCR, image understanding and document/visual analysis.

PLANNING RAM8 GB+
TYPEVLM / GGUF

Main Q4_K_M model file; projector assets may be required by your runtime.

Google4B

Gemma 3 4B IT

Image-to-text, question answering, summarization and visual understanding.

PLANNING RAM8 GB+
TYPEVLM / GGUF

Multimodal family; check the repository for the matching projector file.

Alibaba Qwen7B

Qwen2.5-VL 7B Instruct

Stronger visual reasoning, OCR and complex image/document analysis.

PLANNING RAM12–16 GB
TYPEVLM / GGUF

Use the Q4_K_M variant and matching multimodal projector supported by your runtime.

Google12B

Gemma 3 12B IT

Higher-quality multimodal understanding and enterprise knowledge workflows.

PLANNING RAM16 GB+
TYPEVLM / GGUF

Main model file only; check the model repository for multimodal projector requirements.

LOCAL IMAGE GENERATION

High-resolution image models from ~6 GB to high-memory builds

These are GGUF image-model conversions. Image generation usually requires a compatible image runtime plus supporting text encoders and VAE assets.

Black Forest Labs / city96 GGUF

FLUX.1 Schnell

Fast high-quality generation

  • Model12B
  • FileQ4_K_S ≈ 6.78 GB
  • Memory12–16 GB+ combined memory recommended

Core diffusion/transformer file. ComfyUI-GGUF or another compatible runtime may require text encoders and VAE separately.

Black Forest Labs / city96 GGUF

FLUX.1 Dev

Detailed photorealistic and design workflows

  • Model12B
  • FileQ4_K_S ≈ 6.81 GB
  • Memory16 GB+ combined memory recommended

FLUX.1-dev license restrictions apply. Additional runtime assets are normally required.

Black Forest Labs / city96 GGUF

FLUX.1 Dev High Precision

Higher-precision local image generation

  • Model12B
  • FileQ8_0 ≈ 12+ GB
  • Memory24 GB+ combined memory recommended

Choose Q8_0 from the repository when hardware allows. Image runtimes typically need supporting encoders/VAE.

Stability AI / city96 GGUF

Stable Diffusion 3.5 Large

High-quality text-to-image generation

  • ModelLarge
  • FileQ5_1 ≈ 6.27 GB
  • Memory12–16 GB+ combined memory recommended

Core model file; compatible runtime and supporting assets are required.

Stability AI / city96 GGUF

Stable Diffusion 3.5 Large Q8

Maximum local quantized precision in this catalog

  • ModelLarge
  • FileQ8_0 ≈ 16.3 GB
  • Memory24–32 GB+ combined memory recommended

Large memory requirement. Verify your image runtime and license before download.

USE MODELS RESPONSIBLY

Check license, hardware fit and model provenance before deployment.

Model names and trademarks belong to their respective owners. Inclusion in this library is informational and does not imply endorsement, partnership or ownership. Always verify the original Hugging Face repository before use.