London, UK
Matan Ben Yosef
Generative-AI researcher and engineer at LTX (Lightricks). I enjoy training models and building cool stuff with them.
About
I'm a generative-AI researcher and engineer at LTX, Lightricks' video model team. I work on post-training, which mostly means taking a base video model and making it directable: control adapters, IC-LoRAs, and a lot of performance work on the training stack. I designed and built the LTX Trainer from scratch: the open-source trainer that ships with LTX-2, and what the community uses to fine-tune it. Earlier at Lightricks I built LoRA training infrastructure for SD and Flux, and the backend of a text-to-image product serving millions.
The research side of my work has followed one arc for a decade: generative models of people, from GANs to video diffusion. It started with my MSc at the Hebrew University of Jerusalem (2015–2018), on multi-modal GANs with Daphna Weinshall. I was at D-ID from 2018 to late 2021, ending as lead deep-learning and computer-vision researcher, where I built the face-animation models behind Live Portraits and AI Avatars and worked on adversarial face de-identification.
After D-ID I built FaceFX on my own: an iOS and Android app with around fifty face effects, each one its own GAN, served from GPU nodes on GCP, down to the ads and the subscriptions. Lightricks acquired it in mid-2022, and that is how I ended up here.
Before AI I spent about a decade as a software engineer, mostly on real-time and distributed systems. I've been writing code since I was ten, and I still care most about the same thing: simple, well-designed software.
Publications
Open source
One production training system, and two command-line tools I built because I kept needing them.
-
ltx-trainer
The open-source trainer for LTX-2, designed and built from scratch. It ships inside the model repo, 8.4k stars, and it is what the community uses to fine-tune the model. -
sft-cli
Inspect, diff and edit.safetensorscheckpoints, plus a full LoRA toolkit: extract an adapter from a fine-tune, resize its rank from the singular-value spectrum, convert between Kohya and PEFT. It also ships a skill for coding agents.uv tool install sft-cli -
vidio
ffmpeg without the flags: trim, crop, concatenate, convert, and build grids.uv tool install vidio-cli
Patents
Co-inventor on eight patents from my years at D-ID, in synthetic-media animation, anonymization, and face recognition.
Animation and reenactment
- A System and Method for Voice-Driven Lip Syncing and Head Reenactment
- System and Method for Artificial Neural-Network-Based Animation with Three-Dimensional Rendering
- System and a Method for Artificial Neural-Network Based Animation
Anonymization and identity
- Facial Anonymization with Consistent Facial Attribute Preservation in Video
- System and Method for Performing Facial Image Anonymization
- System and Method for Image De-Identification to Humans While Remaining Recognizable by Machines
- System and Method for Performing Face Recognition
- System and Method for Reconstruction of Faces from Anonymized Media Using Neural Network Based Steganography