Jaekoo Kang
jkang
AI & ML interests
Anything fun and interesting
Organizations
multimodal
-
MMICL: Empowering Vision-language Model with Multi-Modal In-Context Learning
Paper • 2309.07915 • Published • 4 -
Skywork: A More Open Bilingual Foundation Model
Paper • 2310.19341 • Published • 6 -
Multimodal ChatGPT for Medical Applications: an Experimental Study of GPT-4V
Paper • 2310.19061 • Published • 8 -
Lumiere: A Space-Time Diffusion Model for Video Generation
Paper • 2401.12945 • Published • 86
image
multimodal
-
MMICL: Empowering Vision-language Model with Multi-Modal In-Context Learning
Paper • 2309.07915 • Published • 4 -
Skywork: A More Open Bilingual Foundation Model
Paper • 2310.19341 • Published • 6 -
Multimodal ChatGPT for Medical Applications: an Experimental Study of GPT-4V
Paper • 2310.19061 • Published • 8 -
Lumiere: A Space-Time Diffusion Model for Video Generation
Paper • 2401.12945 • Published • 86
spaces 7
Runtime error
Agents
3
ESPNet2 ASR Librispeech word vs bpe tokens
💩
Runtime error
Agents
ESPNet2 ASR Librispeech Conformer (100h)
🐨
Runtime error
Agents
Featured
10
Artist Classifier
🎨
Runtime error
Agents
4
Demo Painttransformer
🏃
Runtime error
Agents
4
Demo GradCAM Imagenet
👀
Runtime error
Agents
3
Demo Image Pyxelate
👀
models 8
jkang/espnet2_an4_transformer
Automatic Speech Recognition • Updated
jkang/espnet2_librispeech_100_conformer_char
Automatic Speech Recognition • Updated • 1
jkang/espnet2_librispeech_100_conformer_word
Automatic Speech Recognition • Updated • 6 • 4
jkang/espnet2_librispeech_100_conformer
Automatic Speech Recognition • Updated • 6
jkang/espnet2_mini_librispeech_diar
Updated
jkang/espnet2_an4_asr
Automatic Speech Recognition • Updated • 4
jkang/drawing-artistic-trend-classifier
Updated • 11
jkang/drawing-artist-classifier
Updated • 3 • 2
datasets 0
None public yet