thanks to the Thinking Machines team, we used Tinker to prototype our reward models and train the prompt expander via RL. for more information, read the full technical report on the data, architecture, and training behind Krea 2 👇 https://t.co/okdxlqJuzA
https://t.co/mNLs2srSIw
Krea trained the prompt expander for its image model with reinforcement learning and prototyped the reward models on Tinker, a worked example of applying an RL service inside an image generation pipeline.
Checking sign-in…
Loading comments…