How to use from
llama.cpp
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
# Run inference directly in the terminal:
llama cli -hf Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
Install from WinGet (Windows)
winget install llama.cpp
# Start a local OpenAI-compatible server with a web UI:
llama serve -hf Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
# Run inference directly in the terminal:
llama cli -hf Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
Use pre-built binary
# Download pre-built binary from:
# https://github.com/ggerganov/llama.cpp/releases
# Start a local OpenAI-compatible server with a web UI:
./llama-server -hf Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
# Run inference directly in the terminal:
./llama-cli -hf Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git
cd llama.cpp
cmake -B build
cmake --build build -j --target llama-server llama-cli
# Start a local OpenAI-compatible server with a web UI:
./build/bin/llama-server -hf Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
# Run inference directly in the terminal:
./build/bin/llama-cli -hf Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
Use Docker
docker model run hf.co/Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M:Q4_K_M
Quick Links

Qwen2.5-3B BuildEng GGUF Q4_K_M

Repository: Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M

This repository contains the Q4_K_M GGUF release of BuildEng V8 3B based on Qwen2.5-3B-Instruct.

BuildEng is a domain-specialized engineering language model project focused on civil engineering, structural reasoning, construction workflows, and conservative engineering-assistant behavior.

The Q4_K_M release is intended mainly for efficient local inference while still preserving strong engineering reasoning quality.

Model Information

Base model:

Qwen/Qwen2.5-3B-Instruct

Format:

GGUF

Release type:

Q4_K_M

Main focus areas include reinforced concrete, foundations, retaining walls, slabs, columns, structural diagnostics, settlement reasoning, temporary works, construction sequencing, renovation uncertainty, and inspection-first engineering workflows.

Related Repositories

Merged model:

https://huggingface.co/Irfanuruchi/qwen2.5-3b-buildeng

Q4_K_M GGUF:

https://huggingface.co/Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M

Q8_0 GGUF:

https://huggingface.co/Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q8_0

F16 GGUF:

https://huggingface.co/Irfanuruchi/qwen2.5-3b-buildeng-GGUF-F16

Dataset:

https://huggingface.co/datasets/Irfanuruchi/buildeng-v8-3b

License

This model is a fine-tune based on Qwen2.5-3B-Instruct. Due to the upstream licensing of the 3B parameter model, this specific variant is released under the Qwen Research License Agreement and is for non-commercial research and evaluation purposes only.

  • Notice: Qwen is licensed under the Qwen RESEARCH LICENSE AGREEMENT, Copyright (c) Alibaba Cloud. All Rights Reserved.
  • Branding: Built with Qwen.

(Note: If you require a BuildEng model for commercial deployment, please look at the variants based on Qwen2.5-1.5B or Qwen2.5-32B, which are Apache 2.0 licensed).

Important Notice

This model is intended for research and engineering-assistant workflows only.

It must not be used as final engineering approval, construction sign-off, or replacement for licensed engineering review.

Author

Irfan Uruchi

Downloads last month
133
GGUF
Model size
3B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Irfanuruchi/qwen2.5-3b-buildeng-GGUF-Q4_K_M

Base model

Qwen/Qwen2.5-3B
Quantized
(264)
this model