MLX quantizations of
Hcompany/Holo4-35B-A3B
, a 35B mixture-of-experts (MoE, 3B active) vision-language model (VLM) for Computer Use, tool-driven work, and agentic workflows.
Each quantization includes the full multimodal vision tower (333 visual weights +
vision_config
), enabling native image and screenshot understanding in MLX-VLM, LM Studio, and Apple Silicon workflows.
Upstream benchmarks
Results reported by H Company from evaluations of the original Holo4-35B-A3B model across computer tasks, long workflows, and tool servers.
Vision tower weights and multimodal projectors are preserved in full precision (BF16) to ensure optimal visual comprehension, OCR, and GUI element grounding without degradation.
Usage with MLX-VLM
Installation
pip install -U mlx-vlm
Python API
from mlx_vlm import load, generate
from mlx_vlm.prompt_utils import apply_chat_template
from mlx_vlm.utils import load_config
# Load 4-bit (or subfolder="Holo4-35B-A3B-MLX-6bit", subfolder="Holo4-35B-A3B-MLX-8bit")
model_path = "abenzerps/Holo4-35B-A3B-MLX"
subfolder = "Holo4-35B-A3B-MLX-4bit"
model, processor = load(model_path, subfolder=subfolder)
config = load_config(model_path, subfolder=subfolder)
prompt = "Describe the user interface elements shown in this screenshot."
image = ["screenshot.png"]
formatted_prompt = apply_chat_template(
processor, config, prompt, num_images=len(image)
)
output = generate(model, processor, formatted_prompt, image, verbose=True)
print(output)
Command Line Interface
# 4-bit
python -m mlx_vlm.generate \
--model abenzerps/Holo4-35B-A3B-MLX --subfolder Holo4-35B-A3B-MLX-4bit \
--image screenshot.png \
--prompt "What action should be taken next to achieve the user goal?"# 6-bit
python -m mlx_vlm.generate \
--model abenzerps/Holo4-35B-A3B-MLX --subfolder Holo4-35B-A3B-MLX-6bit \
--image screenshot.png \
--prompt "What action should be taken next to achieve the user goal?"# 8-bit
python -m mlx_vlm.generate \
--model abenzerps/Holo4-35B-A3B-MLX --subfolder Holo4-35B-A3B-MLX-8bit \
--image screenshot.png \
--prompt "What action should be taken next to achieve the user goal?"
Holo4-35B-A3B-MLX huggingface.co is an AI model on huggingface.co that provides Holo4-35B-A3B-MLX's model effect (), which can be used instantly with this abenzerps Holo4-35B-A3B-MLX model. huggingface.co supports a free trial of the Holo4-35B-A3B-MLX model, and also provides paid use of the Holo4-35B-A3B-MLX. Support call Holo4-35B-A3B-MLX model through api, including Node.js, Python, http.
Holo4-35B-A3B-MLX huggingface.co is an online trial and call api platform, which integrates Holo4-35B-A3B-MLX's modeling effects, including api services, and provides a free online trial of Holo4-35B-A3B-MLX, you can try Holo4-35B-A3B-MLX online for free by clicking the link below.
abenzerps Holo4-35B-A3B-MLX online free url in huggingface.co:
Holo4-35B-A3B-MLX is an open source model from GitHub that offers a free installation service, and any user can find Holo4-35B-A3B-MLX on GitHub to install. At the same time, huggingface.co provides the effect of Holo4-35B-A3B-MLX install, users can directly use Holo4-35B-A3B-MLX installed effect in huggingface.co for debugging and trial. It also supports api for free installation.