FLAME
First Author · AAAI 2023 Oral
Summary
FLAME is a diffusion-based model that synthesizes high-fidelity human motion from free-form text. It can edit selected frames or joints without fine-tuning, using a Transformer designed for variable-length motion and language conditioning. Experiments show state-of-the-art generation on HumanML3D, BABEL, and KIT, with editing also extending to motion prediction and in-betweening. (Kim et al., 2023)
Contribution
During my KakaoBrain research internship, I contributed across FLAME’s full research lifecycle and published the work as first author at AAAI 2023, where it received an oral presentation.