pull down to refresh

That's pretty wild.
Image:
Prompt:
A wide image taken with a phone of a glass whiteboard, in a room overlooking the Bay Bridge. The field of view shows a woman writing, sporting a tshirt wiith a large OpenAI logo. The handwriting looks natural and a bit messy, and we see the photographer's reflection.

The text reads:

(left)
"Transfer between Modalities:

Suppose we directly model
p(text, pixels, sound) [equation]
with one big autoregressive transformer.

Pros:
* image generation augmented with vast world knowledge
* next-level text rendering
* native in-context learning
* unified post-training stack

Cons:
* varying bit-rate across modalities
* compute not adaptive"

(Right)
"Fixes:
* model compressed representations
* compose autoregressive prior with a powerful decoder"

On the bottom right of the board, she draws a diagram:
"tokens -> [transformer] -> [diffusion] -> pixels"
not sure, this is what I got with the same prompt....
reply
Ha, you are right, it was listed as additional option, I missed that. Thanks for pointing that out, here we go, pretty close to yours... Wow it is much better...
reply
68 sats \ 0 replies \ @gmd 26 Mar
these are all fuckin mindblowing to me
reply
Try logging out and back in, and make sure 4o is selected so that it says this:
reply
The perfect angle of the bridge in the background is the giveaway
reply
How about this one 😂
reply
Looks legit lol
reply