pull down to refresh

That's pretty wild.

Image:

Prompt:

A wide image taken with a phone of a glass whiteboard, in a room overlooking the Bay Bridge. The field of view shows a woman writing, sporting a tshirt wiith a large OpenAI logo. The handwriting looks natural and a bit messy, and we see the photographer's reflection.

The text reads:

(left)
"Transfer between Modalities:

Suppose we directly model
p(text, pixels, sound) [equation]
with one big autoregressive transformer.

Pros:
* image generation augmented with vast world knowledge
* next-level text rendering
* native in-context learning
* unified post-training stack

Cons:
* varying bit-rate across modalities
* compute not adaptive"

(Right)
"Fixes:
* model compressed representations
* compose autoregressive prior with a powerful decoder"

On the bottom right of the board, she draws a diagram:
"tokens -> [transformer] -> [diffusion] -> pixels"

not sure, this is what I got with the same prompt....

reply

Ha, you are right, it was listed as additional option, I missed that. Thanks for pointing that out, here we go, pretty close to yours... Wow it is much better...

reply
68 sats \ 0 replies \ @gmd 26 Mar

these are all fuckin mindblowing to me

reply

Try logging out and back in, and make sure 4o is selected so that it says this:

reply

The perfect angle of the bridge in the background is the giveaway

reply

How about this one 😂

reply

Looks legit lol

reply