メインコンテンツへ移動

Pixal3D GGUF: No Official Build — Run Image-to-3D Online

Pixal3D ships no GGUF build, so local weights are not an option. The hosted route instead: one image to GLB with no install, plus the upstream multiview path.

Guides公開日 更新日
Pixal3D GGUF: No Official Build — Run Image-to-3D Online

Try Pixal3D AI 3D Generator

Generate one GLB first, then decide whether to finish the article.

Start Generating 3D Models →
TL;DR
Generate image-to-3D GLB models online in the Pixal3D workspace with no install. Pixal3D is Tencent ARC Lab and Tsinghua University's SIGGRAPH 2026 image-to-3D project. The current official main branch uses a TRELLIS.2-based implementation with pixel back-projection conditioning and exports GLB assets with geometry and PBR textures. The hosted steps below use this site's settings; the local sections follow the upstream repository and require your own GPU environment.

Generate your first GLB online

Pixal3D.ai is an independent hosted service running models through Fal, with no affiliation to the research team. Generation requires sign-in. New accounts receive 220 one-time credits valid for 30 days, enough for one default Pixal3D job. Used or expired credits do not reset.

1

Sign in and upload one image

Use an image you have permission to process, with the whole subject visible and a clear silhouette.

2

Confirm the model and credit cost

Choose Pixal3D, 1024 resolution, and the Balanced preset (2048 textures) for a 220-credit default job. Check the amount beside Generate after changing settings.

3

Generate, inspect, and download

Wait for success, download the textured GLB, and check hidden sides, thin structures, and destination import. Manual cleanup may still be needed.

Open the image-to-3D generator · Inspect a GLB for free · Compare Pixal3D and Trellis 2

Official local single-image defaults

1 image
Input
GLB
Output
1536
Local default
1024
Local low VRAM

What Pixal3D actually is

Pixal3D lifts image features into 3D with explicit pixel back-projection. That direct 2D-to-3D correspondence is the project's defining idea. The official repository separates two implementations: main is the current, improved TRELLIS.2-based version; paper is the Direct3D-S2-based version used for the published paper results.

Do not mix the two branches

Paper metrics describe the paper branch. Installation commands and current output behavior describe main. Treating them as one identical pipeline produces misleading comparisons and broken setup advice.

Before choosing local or hosted

QuestionLocal official workflowHosted Pixal3D.ai workflow
SetupLinux, CUDA toolchain, compiled dependenciesBrowser upload and managed runtime
HardwareStart from TRELLIS.2 official 24GB NVIDIA baselineNo local GPU required
ControlFull repository and local filesManaged settings and account task history
Best forResearch, reproducible local pipelinesEvaluating an image before local setup

Microsoft's official TRELLIS.2 instructions currently say the code is tested on Linux and requires an NVIDIA GPU with at least 24GB memory. Pixal3D adds an on-demand --low_vram mode, but its README does not publish a guaranteed 6GB or 12GB minimum. Measure peak memory instead of treating a community port as an upstream guarantee.


Official installation path

1

Prepare TRELLIS.2 first

Use its official Linux, CUDA 12.4, Conda, and dependency instructions. Confirm its example runs before adding Pixal3D.

2

Clone Pixal3D and install its requirements

Keep the environment isolated so NATTEN, CUDA, and renderer versions cannot collide with another ML project.

3

Compile NATTEN for your GPU architecture

Replace the placeholders with the CUDA architecture and a bounded worker count suitable for the machine.

4

Run one known sample before personal images

This separates environment failures from image-quality failures.

Terminal
git clone --recursive https://github.com/microsoft/TRELLIS.2.git
# Follow TRELLIS.2/setup.sh from the official repository, then:

git clone https://github.com/TencentARC/Pixal3D.git
cd Pixal3D
pip install -r requirements.txt
NATTEN_CUDA_ARCH="YOUR_ARCH" NATTEN_N_WORKERS=4 \
  pip install natten==0.21.0 --no-build-isolation
pip install https://github.com/LDYang694/Storages/releases/download/20260430/utils3d-0.0.2-py3-none-any.whl

python inference.py --image assets/images/0_img.png --output ./output.glb

Use low-VRAM mode deliberately

python inference.py --image input.png --output output.glb --low_vram loads models on demand and defaults to resolution 1024 instead of 1536. It lowers peak memory; it is not a promise that every small GPU will work.

Input image checklist

Prepare one image that exposes useful geometry
1

Keep the whole subject visible

Avoid cutting off legs, handles, antennas, or the base of the object.

2

Use a clean silhouette

Separate the subject from a busy or similarly colored background.

3

Prefer even lighting

Hard shadows and reflections can be mistaken for geometry or material changes.

4

Show depth

A three-quarter view usually reveals more shape than a perfectly flat front view.

5

Expect hidden-side uncertainty

A single image cannot prove the unseen back. Inspect it instead of assuming reconstruction accuracy.

Resolution and failure triage

SymptomLikely boundaryFirst bounded action
CUDA out of memoryResolution or simultaneous model residencyUse --low_vram, then try --resolution 1024
NATTEN build failureWrong CUDA architecture/toolchainVerify CUDA_HOME and rebuild for the installed GPU
Flash Attention unavailableUnsupported build or GPUUse ATTN_BACKEND=sdpa as documented upstream
Weak hidden sideSingle-view ambiguityUse a clearer three-quarter source or a verified multi-view workflow
Soft material detailInput/texture resolutionImprove the source before raising output resolution

Pixal3D GGUF: can you run the weights quantized?

There is no official GGUF build to download. The hosted workspace accepts an image, not GGUF weights, local models, or workflow imports, and the upstream project publishes its checkpoints as .safetensors files in the official model file listing. As of September 28, 2026, neither that listing nor the official repository README documents a supported GGUF workflow. Any GGUF file for Pixal3D is therefore a third-party conversion that needs its own compatibility, license, and output checks, and it is not a hosted option here.

Pixal3D multiview: multi-view inference in the official repository

Multi-view is a real upstream feature, and it is separate from anything this site runs. The repository released multi-view inference code in September 2026, and inference_mv.py conditions the cascade on several views of the same object at once instead of on one image.

  • Input: a directory of views plus a transforms.jsondescribing each view's camera — a 4×4 camera-to-world matrix and the horizontal FOV in radians, in the same Blender/NeRF convention the upstream training renders use. The world is Z-up.
  • Weights: a separate multi-view set (ckpts/*_mv, selected by pipeline_mv.json) from the same official model repository.
  • Flags: --views_dir selects the views, --num_views N restricts the run to the first N of them, and --low_vram, --resolution, and ATTN_BACKEND behave exactly as in the single-image path.
  • Framing: views are never cropped or rescaled, the first frame is the canonical front view, and every other view is placed relative to it.
  • Masks: a view with an alpha channel uses it as the object mask, and a view without one is segmented automatically.

None of that is available in the hosted workspace, which conditions on one source image per job: multi-view stays a local repository workflow that needs your own camera metadata.

Validate the exported GLB

Quality gate before shipping an asset
1

Rotate through 360 degrees

Check the hidden side, bottom, and thin structures.

2

Inspect topology

Look for holes, disconnected islands, self-intersections, and excessive triangle count.

3

Inspect PBR channels

Check base color, roughness, metallic, opacity, seams, and color-space settings.

4

Confirm scale and origin

Normalize units, pivot, orientation, and bounds for the target engine.

5

Test the destination

Open the final GLB in the actual web viewer, Blender, Unity, or Unreal path.

Where to go next

This page covers the whole project. For the shorter guides: the 3D AI comparison picks a lane, the AI 3D workflow walks the hosted steps, and the 3D pixel guide covers what one source image can and cannot carry into the GLB.

Primary sources

More from the Pixal3D Blog

Related 3D guides and workspaces