Research

Publications
Physics-aware Multi-Object 3D Scene Reconstruction (in Progress)
Publication · In Progress
Physics-aware Multi-Object 3D Scene Reconstruction (in Progress)
Kyaw Ye Thu
Recently, research in 3D reconstruction shifts from achieving consistency in mere appearance and geometry to attaining physically plausible models of the scene or the object. For this problem, while test-time optimization approaches takes hours to optimize reasonble physical parameters of even a single object, the generalizability of feed-forward approaches is too limited. We are currently striving to overcome the shortcomings of both approaches and provide a simulation pipeline that is easily generalizable.
When Tom Eats Kimchi: Evaluating Cultural Bias of Multimodal Large Language Models in Cultural Mixture Contexts
Publication · C3NLP Workshop @ NAACL 2025 · Outstanding Paper Award
When Tom Eats Kimchi: Evaluating Cultural Bias of Multimodal Large Language Models in Cultural Mixture Contexts
Jun Seong Kim, Kyaw Ye Thu, Javad Ismayilzada, Junyeong Park, Eunsu Kim, Huzama Ahmad, Na Min An, James Thorne, Alice Oh
MixCuBe — a cross-cultural VQA benchmark built via a novel image-augmentation pipeline, used to evaluate the cultural bias of SOTA multimodal LLMs in mixed-cultural settings.
Research Projects
RenderFormer with Linear Attention
Research Project
RenderFormer with Linear Attention
Bringing a transformer-based rendering pipeline (Microsoft's RenderFormer) from O(N²) to linear time complexity via Performer (FAVOR++) attention.
Dynamic Brain Connectome Learning
Research Project
Dynamic Brain Connectome Learning
A novel graph-ML architecture learning temporal and spatial patterns of brain activation from fMRI images, with two downstream tasks: brain-activation link prediction during language tasks, and performance prediction from neural patterns (graph regression).