Melvin Sevi

PhD Student @ Inria Grenoble

profile.png

Grenoble, France

I’m a PhD student at Inria Grenoble, advised by Stéphane Lathuilière. I work on RL finetuning for generative image models — mostly autoregressive and flow-matching architectures — trying to get them to follow instructions and preferences more reliably instead of just sampling whatever the pretraining objective happened to reward.

Before the PhD, I did the MVA master’s (Mathématiques, Vision, Apprentissage) at ENS Paris-Saclay, and before that Applied Mathematics and Computer Science at Sorbonne University. In between, I spent a few months at LMU Munich with the Computer Vision & Learning group, working with Björn Ommer’s team on controllable text-to-image diffusion — that turned into a CVPR 2025 paper.

I’m generally drawn to problems at the intersection of generative modeling and control: getting a model to do precisely what you ask, not just what looks plausible.

news

Oct 03, 2024 Published my first paper as a co author : Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions !
Apr 15, 2024 Started My Research Internship at LMU Munich under the supervision of Björn Ommer and Stefan Baumann !
Sep 22, 2023 Joined the MVA program at ENS Paris-Saclay !

selected publications

  1. CVPR
    Continuous, Subject-Specific Attribute Control in T2I Models by Identifying Semantic Directions
    Stefan Andreas Baumann ,  Felix Krause ,  Michael Neumayr , and 4 more authors
    In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2025