P
PaperPicks
Conferences
Tae-Hyun Oh
26 papers at tracked venues · 17 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0003-0468-1571 ↗
Google Scholar ↗
Homepage ↗
Venues
CVPR
×4
ICCV
×4
AAAI
×3
ICLR
×3
WACV
×3
ECCV
×2
InterSpeech
×2
ACL
×1
ACM MM
×1
BMVC
×1
NAACL
×1
NeurIPS
×1
Frequent coauthors
Sung-Bin Kim
DBLP profile ↗
ORCID search ↗
×4
JungMok Lee
DBLP profile ↗
ORCID search ↗
×2
Byung-Ki Kwon
DBLP profile ↗
ORCID search ↗
×2
Jaehun Bang
DBLP profile ↗
ORCID search ↗
×1
Won-Seok Choi
DBLP profile ↗
ORCID search ↗
×1
Kyeong Seon Kim
DBLP profile ↗
ORCID search ↗
×1
Jeongsoo Choi
DBLP profile ↗
ORCID search ↗
×1
Jungbin Cho
DBLP profile ↗
ORCID search ↗
×1
Kim Jun-Seong
DBLP profile ↗
ORCID search ↗
×1
Lee Chae-Yeon
DBLP profile ↗
ORCID search ↗
×1
Junhyeong Cho
DBLP profile ↗
ORCID search ↗
×1
Do Huu Dat
DBLP profile ↗
ORCID search ↗
×1
Papers
Beyond the Highlights: Video Retrieval with Salient and Surrounding Contexts
WACV 2026
·
Jaehun Bang
DBLP profile ↗
ORCID search ↗
Patch-wise Retrieval: A Bag of Practical Techniques for Instance-level Matching
WACV 2026
·
Won-Seok Choi
DBLP profile ↗
ORCID search ↗
SMILE-Next: Teaching Large Language Models to Detect, Classify, and Reason about Laughter
ACL 2026
·
JungMok Lee
DBLP profile ↗
ORCID search ↗
mEOL: Training-Free Instruction-Guided Multimodal Embedder for Vector Graphics and Image Retrieval
WACV 2026
·
Kyeong Seon Kim
DBLP profile ↗
ORCID search ↗
AVHBench: A Cross-Modal Hallucination Benchmark for Audio-Visual Large Language Models
ICLR 2025
·
Sung-Bin Kim
DBLP profile ↗
ORCID search ↗
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation
ACM MM 2025
·
Jeongsoo Choi
DBLP profile ↗
ORCID search ↗
Automated Model Discovery via Multi-modal & Multi-step Pipeline
NeurIPS 2025
·
JungMok Lee
DBLP profile ↗
ORCID search ↗
DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding
ICCV 2025
·
Jungbin Cho
DBLP profile ↗
ORCID search ↗
Dr. Splat: Directly Referring 3D Gaussian Splatting via Direct Language Embedding Registration
CVPR 2025
·
Kim Jun-Seong
DBLP profile ↗
ORCID search ↗
JointDiT: Enhancing RGB-Depth Joint Modeling with Diffusion Transformers
ICCV 2025
·
Byung-Ki Kwon
DBLP profile ↗
ORCID search ↗
Perceptually Accurate 3D Talking Head Generation: New Definitions, Speech-Mesh Representation, and Evaluation Metrics
CVPR 2025
·
Lee Chae-Yeon
DBLP profile ↗
ORCID search ↗
Robust 3D Shape Reconstruction in Zero-Shot from a Single Image in the Wild
CVPR 2025
·
Junhyeong Cho
DBLP profile ↗
ORCID search ↗
SoundBrush: Sound as a Brush for Visual Scene Editing
AAAI 2025
·
Sung-Bin Kim
DBLP profile ↗
ORCID search ↗
VSC: Visual Search Compositional Text-to-Image Diffusion Model
ICCV 2025
·
Do Huu Dat
DBLP profile ↗
ORCID search ↗
VoiceCraft-Dub: Automated Video Dubbing with Neural Codec Language Models
ICCV 2025
·
Sung-Bin Kim
DBLP profile ↗
ORCID search ↗
Zero-shot Depth Completion via Test-time Alignment with Affine-invariant Depth Prior
AAAI 2025
·
Lee Hyoseok
DBLP profile ↗
ORCID search ↗
BEAF: Observing BEfore-AFter Changes to Evaluate Hallucination in Vision-Language Models
ECCV 2024
·
Moon Ye-Bin
DBLP profile ↗
ORCID search ↗
CAS: A Probability-Based Approach for Universal Condition Alignment Score
ICLR 2024
·
Chunsan Hong
DBLP profile ↗
ORCID search ↗
Enhancing Speech-Driven 3D Facial Animation with Audio-Visual Guidance from Lip Reading Expert
InterSpeech 2024
·
Han EunGi
DBLP profile ↗
ORCID search ↗
FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance Fields
AAAI 2024
·
GeonU Kim
DBLP profile ↗
ORCID search ↗
Learning-based Axial Video Motion Magnification
ECCV 2024
·
Byung-Ki Kwon
DBLP profile ↗
ORCID search ↗
MeTTA: Single-View to 3D Textured Mesh Reconstruction with Test-Time Adaptation
BMVC 2024
·
Kim Yu-Ji
DBLP profile ↗
ORCID search ↗
MultiTalk: Enhancing 3D Talking Head Generation Across Languages with Multilingual Video Dataset
InterSpeech 2024
·
Sung-Bin Kim
DBLP profile ↗
ORCID search ↗
Noise Map Guidance: Inversion with Spatial Context for Real Image Editing
ICLR 2024
·
Hansam Cho
DBLP profile ↗
ORCID search ↗
Paint-it: Text-to-Texture Synthesis via Deep Convolutional Texture Map Optimization and Physically-Based Rendering
CVPR 2024
·
Kim Youwang
DBLP profile ↗
ORCID search ↗
SMILE: Multimodal Dataset for Understanding Laughter in Video with Language Models
NAACL 2024
·
Lee Hyun
DBLP profile ↗
ORCID search ↗