Han-Bit Kang

hbkang

AI & ML interests

Recent Activity

updated a collection 14 days ago

artistic rendering

upvoted a paper 14 days ago

SynthLight: Portrait Relighting with Diffusion Model by Learning to Re-render Synthetic Faces

upvoted a paper 14 days ago

Tensor Product Attention Is All You Need

View all activity

Organizations

None yet

hbkang's activity

upvoted 3 papers 14 days ago

SynthLight: Portrait Relighting with Diffusion Model by Learning to Re-render Synthetic Faces

Paper • 2501.09756 • Published 15 days ago • 19

Tensor Product Attention Is All You Need

Paper • 2501.06425 • Published 20 days ago • 79

MangaNinja: Line Art Colorization with Precise Reference Following

Paper • 2501.08332 • Published 17 days ago • 55

upvoted a paper 17 days ago

Infecting Generative AI With Viruses

Paper • 2501.05542 • Published 22 days ago • 13

upvoted a paper 21 days ago

The GAN is dead; long live the GAN! A Modern GAN Baseline

Paper • 2501.05441 • Published 22 days ago • 87

upvoted a paper 22 days ago

LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token

Paper • 2501.03895 • Published 24 days ago • 48

upvoted a paper 25 days ago

2.5 Years in Class: A Multimodal Textbook for Vision-Language Pretraining

Paper • 2501.00958 • Published 30 days ago • 99

upvoted a paper 29 days ago

PERSE: Personalized 3D Generative Avatars from A Single Portrait

Paper • 2412.21206 • Published Dec 30, 2024 • 17

upvoted a paper about 1 month ago

Byte Latent Transformer: Patches Scale Better Than Tokens

Paper • 2412.09871 • Published Dec 13, 2024 • 89

upvoted 4 papers about 2 months ago

upvoted 2 papers 2 months ago

Differential Transformer

Paper • 2410.05258 • Published Oct 7, 2024 • 169

BLIP3-KALE: Knowledge Augmented Large-Scale Dense Captions

Paper • 2411.07461 • Published Nov 12, 2024 • 22

upvoted 3 papers 3 months ago

Acoustic Volume Rendering for Neural Impulse Response Fields

Paper • 2411.06307 • Published Nov 9, 2024 • 5

Learning Video Representations without Natural Videos

Paper • 2410.24213 • Published Oct 31, 2024 • 15

DAWN: Dynamic Frame Avatar with Non-autoregressive Diffusion Framework for Talking Head Video Generation

Paper • 2410.13726 • Published Oct 17, 2024 • 11

upvoted 2 papers 4 months ago

Meissonic: Revitalizing Masked Generative Transformers for Efficient High-Resolution Text-to-Image Synthesis

Paper • 2410.08261 • Published Oct 10, 2024 • 50

FAN: Fourier Analysis Networks

Paper • 2410.02675 • Published Oct 3, 2024 • 25