How to use from
Lemonade
Pull the model
# Download Lemonade from https://lemonade-server.ai/
lemonade pull prithivMLmods/FaithEyes-7B-RL-GGUF:
Run and chat with the model
lemonade run user.FaithEyes-7B-RL-GGUF-
List all available models
lemonade list
Quick Links

FaithEyes-7B-RL-GGUF

FaithEyes-7B-RL is the final reinforcement-learning checkpoint of FaithEyes on Qwen2.5-VL-7B-Instruct, obtained by running GRPO on top of the FaithEyes-7B-SFT cold-start model, and is the primary model reported in the paper "FaithEyes: Towards Faithful Tool Use via Multi-Agent Process-Image Self-Verification." FaithEyes is a multi-agent self-judging framework in which a single VLM plays two roles: a main agent that solves visual questions by interleaving reasoning with executable code-based tool calls (like image crops), and a subagent — instantiated by the same model under a separate prompt, requiring no external model dependency — that judges whether each process image the main agent produces is actually helpful, with that verdict both injected into the tool observation to steer subsequent reasoning and used to scale the tool reward. The RL stage combines four reward terms (accuracy, format, consistency, and an accuracy-independent tool-faithfulness reward that credits helpful, executable calls and penalizes failed ones) to push the model from merely imitating demonstrations toward autonomously producing genuinely faithful rather than decorative tool calls, while keeping code-failure ratio and tool usage stable throughout training (an accuracy-gated variant was found to degenerate into tool avoidance). Training built on the verl framework alongside DeepEyes, Thyme, ms_swift, and VLMEvalKit, and the model is released under the Apache 2.0 license. FaithEyes-7B-RL on Hugging Face

Model Files

File Name Quant Type File Size File Link Description
FaithEyes-7B-RL.BF16.gguf BF16 15.2 GB Link Full BF16 weights. Highest quality, largest file size.
FaithEyes-7B-RL.Q3_K_L.gguf Q3_K_L 4.09 GB Link Lower quality but usable, good for low RAM availability.
FaithEyes-7B-RL.Q3_K_M.gguf Q3_K_M 3.81 GB Link Low quality.
FaithEyes-7B-RL.Q4_K_M.gguf Q4_K_M 4.68 GB Link Good quality, default size for most use cases, recommended.
FaithEyes-7B-RL.Q4_K_S.gguf Q4_K_S 4.46 GB Link Slightly lower quality with more space savings, recommended.
FaithEyes-7B-RL.Q5_K_M.gguf Q5_K_M 5.44 GB Link High quality, recommended.
FaithEyes-7B-RL.Q5_K_S.gguf Q5_K_S 5.32 GB Link High quality, recommended.
FaithEyes-7B-RL.mmproj-bf16.gguf mmproj-bf16 1.36 GB Link Multimodal projection file in BF16 format. Used for vision/language models.

llama.cpp

LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp

Downloads last month
608
GGUF
Model size
8B params
Architecture
qwen2vl
Hardware compatibility
Log In to add your hardware

3-bit

4-bit

5-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for prithivMLmods/FaithEyes-7B-RL-GGUF

Quantized
(1)
this model

Collections including prithivMLmods/FaithEyes-7B-RL-GGUF