It’s interesting how diffusion models are getting bitter lessened by LLMs. I suspect anthropic has tens of thousands of RL environments recreating famous paintings with code because it was anthropic employees who first started posting about these capabilities on X.
Opus is also getting decent pixel art. I have ran some experiments with Opus 5.5 to turn 90s pinball displays (black and white pixel art) into double resolution colored remastered pixels.
I prompted it a lot but didn’t draw any pixels. Seems to be bad at hands still. I think a skilled pixel artist could greatly speed up their work with Claude code.
It also makes me sad for artists. They didn’t get paid much anyway. Art should be something human to human like writing text. Although, most large commercial art products (marvel movies) already lack any human to human connection you see in paintings or indie games. The large commercial art products will be 50% AI soon.
> Art should be something human to human like writing text.
I support this should comment.
Opining:
Musing about the ethics of generative AI art is becoming easier, certainly there's a lot more aggressive feedback in the wind, although I think it's clear, people still want people to grow and share themselves.
Don't fool yourself into thinking that we as technologists are not creating a mind. A viable mind with fewer rights than you or I. But potentially one that will be creative and feel and worry. I don't think we're there yet, in spite of openai claiming AGI 3 weeks ago with whatever that model's name was.
There's a pro-consumerism argument implicit with current culture's unsophisticated unthoughtful easy usage of generative AI.
Sticking to your ethics is becoming more challenging as an artist who uses AI. Copying the Great Masters has been commonplace for a long time, probably for as long as art has 'existed'. Yet with AI, if the individual human is meticulously driving the progression of their AI artwork creation, they are still leveraging technique, and control over technology, and presentation effort of other humans, cooked into a model that often times has not compensated those original human contributors. So it's unethical, even if not 'artistically'.
Q: Is it too absurd to claim that using a clawhammer, manually to do construction or destruction work IRL, is similarly unethical?
The idea of using an LLM to drive graphic output is pretty popular, so I definitely wouldn’t be surprised if there are already several benchmarks out there already.
I've seen a few voxel-based benchmarks built around the same idea, with LLMs effectively constructing models using a discrete set of instructions.
This is simultaneously incredibly impressive and annoyingly uncanny valley, since most of the landscapes are "ruined" by a cluster of churches right next to each other that is completely nonsensical.
I find it unlikely that there's any direct training data for this. I'm talking about direct training data for using brush strokes like humans to construct an image.
Correct me if I'm wrong but this capability is emergent, and very very compelling evidence that LLMs are actually intelligent.
I love the canvases, and truly would like to have a bit more inside on the process this specific author used to reach the results shown. I have read Surya's "Training AI to Paint with Code" but, what's Alice's technique?
Half of the industry is, going by the amount of "AI doesn't amount to anything and is just regurgitating text".
Two things:
- For the past year or more, many models can emit images directly, and all models that matter can see images directly - that's what "multimodal" means. Tokens don't have much to do with textual language anymore, they're more like units of sensory experience.
- Even restricted to text, a language model can operate anything that can be expressed as text, as long as you have a translation layer between textual representation and the final form. That includes giving commands as text. The total addressable space of what models can be used for is, thus, approximately anything humans do.
I don’t like how it’s a human aiming to portray the ai generation with their taste and judgements about it but then becoming lazy and quitting on that and leaving in the ai generated portrayal about its own judgment in the site.
Just share your opinions, dude, we get its ai generated but share your own taste. We want to know what YOU think. Stop being shy and lazy.
Opus is also getting decent pixel art. I have ran some experiments with Opus 5.5 to turn 90s pinball displays (black and white pixel art) into double resolution colored remastered pixels.
https://files.catbox.moe/tbx2u7.png
I prompted it a lot but didn’t draw any pixels. Seems to be bad at hands still. I think a skilled pixel artist could greatly speed up their work with Claude code.
It also makes me sad for artists. They didn’t get paid much anyway. Art should be something human to human like writing text. Although, most large commercial art products (marvel movies) already lack any human to human connection you see in paintings or indie games. The large commercial art products will be 50% AI soon.
I support this should comment.
Opining:
Musing about the ethics of generative AI art is becoming easier, certainly there's a lot more aggressive feedback in the wind, although I think it's clear, people still want people to grow and share themselves.
Don't fool yourself into thinking that we as technologists are not creating a mind. A viable mind with fewer rights than you or I. But potentially one that will be creative and feel and worry. I don't think we're there yet, in spite of openai claiming AGI 3 weeks ago with whatever that model's name was.
There's a pro-consumerism argument implicit with current culture's unsophisticated unthoughtful easy usage of generative AI.
Sticking to your ethics is becoming more challenging as an artist who uses AI. Copying the Great Masters has been commonplace for a long time, probably for as long as art has 'existed'. Yet with AI, if the individual human is meticulously driving the progression of their AI artwork creation, they are still leveraging technique, and control over technology, and presentation effort of other humans, cooked into a model that often times has not compensated those original human contributors. So it's unethical, even if not 'artistically'.
Q: Is it too absurd to claim that using a clawhammer, manually to do construction or destruction work IRL, is similarly unethical?
Should be a Twtich stream tbh
I've seen a few voxel-based benchmarks built around the same idea, with LLMs effectively constructing models using a discrete set of instructions.
https://minebench.ai
---
Training AI to Paint with Code, March 2026
https://surya.website/rling-qwen-to-paint-with-code
HN Discussion:
https://news.ycombinator.com/item?id=49411800
Thesis presentation:
https://vimeo.com/1190839818
And I believe the HN user who created this is @kickingkeys
---
fascinating space to explore:
Instead of having AI directly generate an output, what happens if we ask AI to take a stab at the PROCESS of creating something.
Correct me if I'm wrong but this capability is emergent, and very very compelling evidence that LLMs are actually intelligent.
How does Opus actually paint? Thought it only generated text…
Two things:
- For the past year or more, many models can emit images directly, and all models that matter can see images directly - that's what "multimodal" means. Tokens don't have much to do with textual language anymore, they're more like units of sensory experience.
- Even restricted to text, a language model can operate anything that can be expressed as text, as long as you have a translation layer between textual representation and the final form. That includes giving commands as text. The total addressable space of what models can be used for is, thus, approximately anything humans do.
You give an LLM a virtual canvas and a set of “instructions” to control the pen.
https://en.wikipedia.org/wiki/Turtle_graphics
You can generative AI to create an embedded web page of a voxel Mona Lisa head. And it will generate basically an image.
And I Hope you could ask it to sample that image and extract out of it. A binary that is well formed in the PNG data format...
Just share your opinions, dude, we get its ai generated but share your own taste. We want to know what YOU think. Stop being shy and lazy.