Twitter/X

PRX Pixel is an open-source 7B text-to-image model that generates images directly…

Brief

PRX Pixel, an open-source 7B text-to-image model from @photoroom_ML announced 2026-06-12, produces images directly in pixel space after pretraining on hundreds of millions of images followed by supervised fine-tuning and preference alignment. The team released model weights and a demo (link in comments), is integrating the model into Hugging Face Diffusers, and plans technical blog posts on the training recipe.

Why it matters

PRX Pixel is an open-source 7B text-to-image model that generates images directly in pixel space, announced by @photoroom_ML on 2026-06-12.

Key details

  • The model was pretrained on "hundreds of millions" of images, then underwent supervised fine-tuning and preference alignment; the weights are publicly available and a demo link is posted in the comments.
  • Photoroom is working to integrate PRX Pixel into Hugging Face Diffusers and will publish a series of technical blog posts describing the full training recipe.
Source evidence

🚀 Meet PRX Pixel.
Our new open-source 7B text-to-image model that generates images directly in pixel space.
After months of pretraining on hundreds of millions of images, supervised fine-tuning, and preference alignment, we're excited to share a first public preview.
The weights are already available, and we're currently working on integrating the model directly into Diffusers 🤗to make the model even easier to use.
Test it yourself in the demo below. And as always, we'll be sharing the full story behind the model through a series of technical blog posts covering the entire training recipe.
Link in the comments 👇