Engineer working on generative media - diffusion models, LoRA fine-tuning, and squeezing image/vision models down to run fast and cheap (NVFP4 & low-bit quantization). Also fine-tune vision-language models (Qwen2-VL on ChartQA). Building tools that put creative AI in the hands of creators.
ππΊπ New Research Alert - CVPR 2024 (Avatars Collection)! πππ π Title: 3DGS-Avatar: Animatable Avatars via Deformable 3D Gaussian Splatting π
π Description: 3DGS-Avatar is a novel method for creating animatable human avatars from monocular videos using 3D Gaussian Splatting (3DGS). By using a non-rigid deformation network and as-isometric-as-possible regularizations, the method achieves comparable or better performance than SOTA methods while being 400x faster in training and 250x faster in inference, allowing real-time rendering at 50+ FPS.
π₯ Authors: Zhiyin Qian, Shaofei Wang, Marko Mihajlovic, Andreas Geiger, Siyu Tang
π Conference: CVPR, Jun 17-21, 2024 | Seattle WA, USA πΊπΈ