Tian (Owen) Ye
  • Agent Harness
  • Blogs
  • Publications

Blogs

Blogs

Notes, experiments, and longer-form writing on foundation models and automated AI systems.
Mar 20 2026 8 min read

Building Real-Time Editing on FLUX2: Inference Acceleration and Distillation with Reinforcement Learning (Preview)

A systems-and-training walkthrough of how a FLUX2-based editor was pushed toward real-time interaction using cache-aware two-step inference, causal attention distillation, and reward-guided DMDR.

↗
© 2026 Tian (Owen) Ye · Powered by Hugo & PaperMod