Commit History

Remove verbose= kwarg from ReduceLROnPlateau: removed in newer PyTorch (we are now on 2.8.0)
ab4dd9b
Running
verified

Daankular commited on

Reset torch default device to cpu after offload.profile(): fixes device mismatches in econdataset/pymaf/smplx code that assumes CPU-default tensor creation, without affecting mmgps explicit hook-based device management
181be20
verified

Daankular commited on

Fix device mismatch: maskrcnn detector lands on cuda under mmgps lingering default device, but torch.from_numpy() input stays on cpu
cf35ba7
verified

Daankular commited on

Bump GPU duration to 600s: memory issue is fixed (attention slicing), task now just needs more time for the full multiview diffusion + mesh reconstruction pipeline
cb5743a
verified

Daankular commited on

Implement attention slicing: chunk SDPA calls along the batch*heads dim since torch 2.8s fused kernels reject this shape on this GPU, forcing an O(seq_len^2 * batch) math fallback that OOMs even a 96GB card
63d710a
verified

Daankular commited on

Diagnose why SDPA auto-select avoids flash/efficient backends: try FLASH_ATTENTION explicitly and log the actual error, plus dtype/contiguity/mask info
4df1266
verified

Daankular commited on

Add memory diagnostics at each attention call to see if VRAM usage is growing unbounded or just spiking once
4ba1e6a
verified

Daankular commited on

Add GPU memory diagnostic print to confirm whether size=xlarge is actually being granted
33ea232
verified

Daankular commited on

Request ZeroGPU xlarge (full 96GB Blackwell card) instead of default large (48GB half-card): the multiview attention activation memory OOMs on the half-card even with mmgp weight offloading
8526e94
verified

Daankular commited on

Try mmgp VerylowRAM_LowVRAM profile: LowRAM_LowVRAM still OOMs during attention
432c498
verified

Daankular commited on

Fix remaining comment typo
3a0919b
verified

Daankular commited on

Fix comment typo (apostrophe escaping) from previous commit
97e5396
verified

Daankular commited on

Exclude image_normalizer from mmgp offload management (its .scale() method bypasses mmgps forward-hook, stranding buffers on CPU); pin it on GPU directly instead
a143759
verified

Daankular commited on

Fix create_mean_pose() to build its numpy array from CPU tensors explicitly, instead of masking it with a global torch default-device reset that broke mmgp's own device handling
7f3d882
verified

Daankular commited on

Reset torch default device to cpu after offload.profile(): mmgp leaves it as cuda, which broke SMPLDataset tensor creation later in the script
fc017ad
verified

Daankular commited on

Bump diffusers to 0.29.0: mmgp (via optimum-quanto) needs PixArtTransformer2DModel, absent from the old 0.26.0 pin; verified all of PSHumans own diffusers imports (incl. LoRACompatibleConv/AdaLayerNorm) still resolve at 0.29.0
6ce5fd0
verified

Daankular commited on

Use MMGP budget-based offloading/quantization instead of a blanket .to('cuda'): the pipeline genuinely OOMs on this ZeroGPU MIG slice's effective VRAM
9eacf98
verified

Daankular commited on

Revert forced EFFICIENT_ATTENTION backend (unsupported for these tensor shapes on this GPU); back to auto-selected SDPA
42c2660
verified

Daankular commited on

Force SDPA's memory-efficient backend for PSHuman's multiview attention: rule out a math-fallback OOM on Blackwell
ff5e65e
verified

Daankular commited on

Disable expandable_segments CUDA allocator: it hits an NVML assertion on MIG-sliced Blackwell GPUs
bb80551
verified

Daankular commited on

Keep PSHuman's custom multiview attention processors but swap their inner xformers.ops.memory_efficient_attention call for torch's native scaled_dot_product_attention (xformers' Hopper kernel crashes on Blackwell/sm_120)
2080275
verified

Daankular commited on

Drop xformers: its Hopper-specific flash-attention kernel crashes on Blackwell (sm_120) with 'CUDA error: invalid argument'. torch 2.8's native SDPA (used automatically by diffusers) replaces it.
4725c84
verified

Daankular commited on

Bump warp-lang to 1.17.0: kaolin 0.18.0 physics subsystem imports warp.fem.linalg.inverse_qr, missing from the old 1.4.2 pin
e4d389f
verified

Daankular commited on

Bump sympy to 1.13.3 and triton to 3.4.0: both hard-required by torch 2.8.0 (sympy>=1.13.3, triton==3.4.0 exactly on linux x86_64)
2cc7752
verified

Daankular commited on

Bump typing_extensions to 4.12.2: torch 2.8.0 requires >=4.10.0, conflicting with the old 4.9.0 pin
e7397e9
verified

Daankular commited on

Upgrade to torch 2.8.0+cu128 (adds Blackwell/sm_12x kernels) with matching kaolin 0.18.0, pytorch3d 0.7.8+pt2.8.0cu128, torch_scatter 2.1.2+pt28cu128, xformers 0.0.32.post2
c2448e9
verified

Daankular commited on

Add nvidia-cuda-runtime + LD_LIBRARY_PATH shim so nvdiffrast's JIT-compiled CUDA extension can find libcudart.so.13 on Blackwell/CUDA-13 ZeroGPU workers
f5c6af6
verified

Daankular commited on

Fix torch pin to 2.1.2 (xformers==0.0.23.post1 requires torch==2.1.2 exactly, conflicting with the 2.1.0 I pinned previously)
b908182
verified

Daankular commited on

Pin torch/torchvision/torchaudio to 2.1.0+cu121: unpinned torch was drifting to latest PyPI (now 2.14.0/cu13x), breaking ABI with the hard-pinned kaolin/pytorch3d/torch_scatter cu121 wheels (libcudart.so.13 not found)
f75e574
verified

Daankular commited on

Pin setuptools<82: pkg_resources was fully removed in setuptools 82, breaking nvdiffrast/torch.utils.cpp_extension import
606c14a
verified

Daankular commited on

Add setuptools so pkg_resources is available for nvdiffrast/torch.utils.cpp_extension import
f2c23c5
verified

Daankular commited on

Update requirements.txt
3e0afe2
verified

fffiloni commited on

Update requirements.txt
8561d1a
verified

fffiloni commited on

increase ZeroGPU needed time for inference
afb1563
verified

fffiloni commited on

includes ZeroGPU deco for main function
eb0fbff
verified

fffiloni commited on

import spaces for ZeroGPU
bd444a3
verified

fffiloni commited on

set xformers 0.0.23.post1
f79ea6f
verified

fffiloni commited on

Update requirements.txt
729a66a
verified

fffiloni commited on

Update requirements.txt
62b7383
verified

fffiloni commited on

Update requirements.txt
51191bd
verified

fffiloni commited on

downgradio 5.8.0
67e2a93
verified

fffiloni commited on

Update requirements.txt
ea3c3d6
verified

fffiloni commited on

back to working properly config, psguman currently not compatible with Zero
e1eb50a
verified

fffiloni commited on

Update requirements.txt
9f8cd50
verified

fffiloni commited on

Update requirements.txt
934dbb6
verified

fffiloni commited on

Update requirements.txt
0121406
verified

fffiloni commited on

upgrade xformers==0.0.29.post1
55f4540
verified

fffiloni commited on

back to torch modules 2.1.2 v
0dcb8cc
verified

fffiloni commited on

look for the right putorch3d
3a9651e
verified

fffiloni commited on

upgrade peft>=0.17.0
90fd7a6
verified

fffiloni commited on