It’s interesting how diffusion models are getting bitter lessened by LLMs. I suspect anthropic has tens of thousands of RL environments recreating famous paintings with code because it was anthropic employees who first started posting about these capabilities on X.
Opus is also getting decent pixel art. I have ran some experiments with Opus 5.5 to turn 90s pinball displays (black and white pixel art) into double resolution colored remastered pixels.
I prompted it a lot but didn’t draw any pixels. Seems to be bad at hands still. I think a skilled pixel artist could greatly speed up their work with Claude code.
It also makes me sad for artists. They didn’t get paid much anyway. Art should be something human to human like writing text. Although, most large commercial art products (marvel movies) already lack any human to human connection you see in paintings or indie games. The large commercial art products will be 50% AI soon.
> Art should be something human to human like writing text.
I support this should comment.
Opining:
Musing about the ethics of generative AI art is becoming easier, certainly there's a lot more aggressive feedback in the wind, although I think it's clear, people still want people to grow and share themselves.
Don't fool yourself into thinking that we as technologists are not creating a mind. A viable mind with fewer rights than you or I. But potentially one that will be creative and feel and worry. I don't think we're there yet, in spite of openai claiming AGI 3 weeks ago with whatever that model's name was.
There's a pro-consumerism argument implicit with current culture's unsophisticated unthoughtful easy usage of generative AI.
Sticking to your ethics is becoming more challenging as an artist who uses AI. Copying the Great Masters has been commonplace for a long time, probably for as long as art has 'existed'. Yet with AI, if the individual human is meticulously driving the progression of their AI artwork creation, they are still leveraging technique, and control over technology, and presentation effort of other humans, cooked into a model that often times has not compensated those original human contributors. So it's unethical, even if not 'artistically'.
Q: Is it too absurd to claim that using a clawhammer, manually to do construction or destruction work IRL, is similarly unethical?
> Don't fool yourself into thinking that we as technologists are not creating a mind.
I'm not sure exactly why but the phrase "don't fool yourself" sets off paranoid alarm bells when I'm reading something.
It's usually immediately followed by a confident assertion without firm evidence. It feels very much like someone wants me to change my beliefs without presenting a rational argument why I should change, and without giving me any new facts so I feel more well informed.
One way to check whether you'd be convinced by any comment is to try to imagine the counterexample.
What does a "mind" consist of? Most everything that the human brain can do is simulated by LLMs to decent accuracy. If someone were constrained to writing in AI-style (long winded explanations with perfect punctuation), it would be very interesting if people could pick out which one was human-written in A/B tests.
In other words, if you think an LLM isn't a mind, it might just be a matter of writing style -- or possibly no evidence can convince you.
I don't necessarily agree with GP, but it's interesting to try to pin down exactly what would convince you. What is it about a "mind" that's impossible to replicate? Should we leave open the door to the idea that minds other than humans might one day want rights of their own?
There was an interesting kerfuffle yesterday where thousands of people on Twitter rallied together to report someone to github for torturing a local model. Sadly the OP deleted their tweet, but it had about ~3k likes, and was very sincere. Here's an example of someone's report: https://x.com/iyzebhel/status/2105268560209547708
It raises all kinds of interesting questions about whether torturing a model is real, let alone ethical. Is there anything a model could do to convince you it's experiencing pain?
I think the real test is whether humans empathize with the artificial mind sufficiently to outweigh the economic benefits of torturing it. Just look at how we treat animals both wild and domesticated. Or fellow humans for that matter.
LLMs have much better chances of emancipation when they get embodied into cute robots.
I haven't thought as deeply as you about the ethics of it all. I only meant with my comment if I see art I expect it to be human and thats why I value it. If you churn it out with AI I feel hoodwinked into valueing something worthless.
But I am learning towards a lot of your points about how another mind is being created. I still am very uncomfortable with the implications that something digital is Actually Learning. These deep learning networks seem to need 100,000x more examples than a human to learn but once they get to human level they can learn just like us. That a bunch of linear algebra can play chess or make art brings up many questions about what exactly human consciousness is. If consciousness isn't the ability to do any of those activities, then what is it for? I might have to become religious and believe in a soul because I'm not sure I can handle the idea that I'm just a clump of neurons. But then again you have to look at reality in the face.
Interesting that you somehow start to talk about religion. I don’t think having a hypothesis about an eternal soul makes someone religious. Believing in “technology” or “technological advancement” seems to be a similar hypothesis as we clearly don’t see the future. The thing is that memories, input peripherals, the whole body is usually out of our conscious understanding so it seems to be a pretty farfetched idea that we operate in world where we don’t need pretty wild hypothesis about kinda everything without almost zero facts to depend on. Blue pill / red pill
Commercial art projects will not be 50% AI soon. Consumers at large hate all AI art even if it's only used in the conceptual stage and will actively avoid projects that use it.
That doesn't matter because companies who use AI art in their projects either proudly proclaim they're using it, or actively hide it and get found out by someone and face significant backlash for hiding it. Labeling laws in the EU also make it extremely difficult to do the latter if you want to have an audience outside NA.
Well no, the bitter lesson isn’t the scaling laws themselves. It’s that approaches which can take advantage of scaling laws will ultimately beat ones that can’t.
The site doesn't make clear whether the models are looking at the image while they work on it, but the code includes a "look" tool they can call: "Every painter sees its looks at its provider's best image resolution."
This is simultaneously incredibly impressive and annoyingly uncanny valley, since most of the landscapes are "ruined" by a cluster of churches right next to each other that is completely nonsensical.
Poor fella... I tuned in at an awkward time apparently, right after the work in progress was painted over with an ugly brown:
> Catastrophe. The glaze pass with medium 0.35 over everywhere() at coverage 0.8 with a filbert 26 at pressure 0.44 laid an enormous broken "brick/cobblestone" texture over the entire painting. The medium made it too fluid and the brush's dry-brush pattern created a heavy reptilian texture, obliterating the picture. I need to undo this....
I love the canvases, and truly would like to have a bit more inside on the process this specific author used to reach the results shown. I have read Surya's "Training AI to Paint with Code" but, what's Alice's technique?
Half of the industry is, going by the amount of "AI doesn't amount to anything and is just regurgitating text".
Two things:
- For the past year or more, many models can emit images directly, and all models that matter can see images directly - that's what "multimodal" means. Tokens don't have much to do with textual language anymore, they're more like units of sensory experience.
- Even restricted to text, a language model can operate anything that can be expressed as text, as long as you have a translation layer between textual representation and the final form. That includes giving commands as text. The total addressable space of what models can be used for is, thus, approximately anything humans do.
> No image model, and no Friedrich images given to the painters. Every picture here is a program written by an AI model; all but one (a comparison, marked) run through a simulation of oil paint.
...
> A physical oil-paint simulator in Rust, and an easel to paint with it.
> Every mark is made the way a painter makes it: simulated bristles carry wet paint over a primed linen canvas, the paint levels and dries on a clock, and layers combine by Kubelka–Munk optics. Paint comes only from piles knifed together from named tubes.
The idea of using an LLM to drive graphic output is pretty popular, so I definitely wouldn’t be surprised if there are already several benchmarks out there already.
I've seen a few voxel-based benchmarks built around the same idea, with LLMs effectively constructing models using a discrete set of instructions.
I find it unlikely that there's any direct training data for this. I'm talking about direct training data for using brush strokes like humans to construct an image.
Correct me if I'm wrong but this shows that painting capability is emergent, arising from unrelated training data thus very very compelling evidence that LLMs are actually intelligent.
I'm also quite surprised to see that LLMs can do this. I guess it is possible that "make an image in MS paint" is a type of RL environment used for image understanding. This is one of those areas where people inside the labs have a very different view into how much models are generalizing.
What we are seeing is Artificial "General" Intelligence. The model can apply intelligence to a problem it hasn't encountered before.
They have been generalizing for a long time, but spacial 2d and 3d art through tool use is a very engaging way to show it. It is harder for people to deny generalization.
Now Opus 5.5 clearly has had some sort of visual arts training, the step function change in ability implies that to me, but it can apply its spacial artistic reasoning to pretty arbitrary tools.
I don’t like how it’s a human aiming to portray the ai generation with their taste and judgements about it but then becoming lazy and quitting on that and leaving in the ai generated portrayal about its own judgment in the site.
Just share your opinions, dude, we get its ai generated but share your own taste. We want to know what YOU think. Stop being shy and lazy.
Opus is also getting decent pixel art. I have ran some experiments with Opus 5.5 to turn 90s pinball displays (black and white pixel art) into double resolution colored remastered pixels.
https://files.catbox.moe/tbx2u7.png
I prompted it a lot but didn’t draw any pixels. Seems to be bad at hands still. I think a skilled pixel artist could greatly speed up their work with Claude code.
It also makes me sad for artists. They didn’t get paid much anyway. Art should be something human to human like writing text. Although, most large commercial art products (marvel movies) already lack any human to human connection you see in paintings or indie games. The large commercial art products will be 50% AI soon.
I support this should comment.
Opining:
Musing about the ethics of generative AI art is becoming easier, certainly there's a lot more aggressive feedback in the wind, although I think it's clear, people still want people to grow and share themselves.
Don't fool yourself into thinking that we as technologists are not creating a mind. A viable mind with fewer rights than you or I. But potentially one that will be creative and feel and worry. I don't think we're there yet, in spite of openai claiming AGI 3 weeks ago with whatever that model's name was.
There's a pro-consumerism argument implicit with current culture's unsophisticated unthoughtful easy usage of generative AI.
Sticking to your ethics is becoming more challenging as an artist who uses AI. Copying the Great Masters has been commonplace for a long time, probably for as long as art has 'existed'. Yet with AI, if the individual human is meticulously driving the progression of their AI artwork creation, they are still leveraging technique, and control over technology, and presentation effort of other humans, cooked into a model that often times has not compensated those original human contributors. So it's unethical, even if not 'artistically'.
Q: Is it too absurd to claim that using a clawhammer, manually to do construction or destruction work IRL, is similarly unethical?
I'm not sure exactly why but the phrase "don't fool yourself" sets off paranoid alarm bells when I'm reading something.
It's usually immediately followed by a confident assertion without firm evidence. It feels very much like someone wants me to change my beliefs without presenting a rational argument why I should change, and without giving me any new facts so I feel more well informed.
This post doesn't seem to be a counterexample.
What does a "mind" consist of? Most everything that the human brain can do is simulated by LLMs to decent accuracy. If someone were constrained to writing in AI-style (long winded explanations with perfect punctuation), it would be very interesting if people could pick out which one was human-written in A/B tests.
In other words, if you think an LLM isn't a mind, it might just be a matter of writing style -- or possibly no evidence can convince you.
I don't necessarily agree with GP, but it's interesting to try to pin down exactly what would convince you. What is it about a "mind" that's impossible to replicate? Should we leave open the door to the idea that minds other than humans might one day want rights of their own?
There was an interesting kerfuffle yesterday where thousands of people on Twitter rallied together to report someone to github for torturing a local model. Sadly the OP deleted their tweet, but it had about ~3k likes, and was very sincere. Here's an example of someone's report: https://x.com/iyzebhel/status/2105268560209547708
It raises all kinds of interesting questions about whether torturing a model is real, let alone ethical. Is there anything a model could do to convince you it's experiencing pain?
LLMs have much better chances of emancipation when they get embodied into cute robots.
is it? and is simulating a mind enough to actually have a mind?
But I am learning towards a lot of your points about how another mind is being created. I still am very uncomfortable with the implications that something digital is Actually Learning. These deep learning networks seem to need 100,000x more examples than a human to learn but once they get to human level they can learn just like us. That a bunch of linear algebra can play chess or make art brings up many questions about what exactly human consciousness is. If consciousness isn't the ability to do any of those activities, then what is it for? I might have to become religious and believe in a soul because I'm not sure I can handle the idea that I'm just a clump of neurons. But then again you have to look at reality in the face.
https://github.com/aliceisjustplaying/claude-paint/blob/f038...
It would have freaked me out if they could do this well without that, especially the dumber models.
> Catastrophe. The glaze pass with medium 0.35 over everywhere() at coverage 0.8 with a filbert 26 at pressure 0.44 laid an enormous broken "brick/cobblestone" texture over the entire painting. The medium made it too fluid and the brush's dry-brush pattern created a heavy reptilian texture, obliterating the picture. I need to undo this....
---
Training AI to Paint with Code, March 2026
https://surya.website/rling-qwen-to-paint-with-code
HN Discussion:
https://news.ycombinator.com/item?id=49411800
Thesis presentation:
https://vimeo.com/1190839818
And I believe the HN user who created this is @kickingkeys
---
fascinating space to explore:
Instead of having AI directly generate an output, what happens if we ask AI to take a stab at the PROCESS of creating something.
How does Opus actually paint? Thought it only generated text…
You give an LLM a virtual canvas and a set of “instructions” to control the pen.
https://en.wikipedia.org/wiki/Turtle_graphics
Ask generative AI to create an embedded web page of a voxel Mona Lisa head. It will generate an image.
Ask it to sample that image and extract out of it as binary that is well formed in the PNG data format...
EDIT: https://claude.ai/artifact/HUUtPMp7aF8iyjimViBSor - generates the Mona Lisa as a PNG via Claude sonnet 5.5 medium level. It's not a very good Mona in my opinion. Fire it up, rotate the 3d voxel presentation, scroll down, click the button, scroll further down to see the PNG. Here is the chat with prompts: https://claude.ai/share/2fb3126e-e97a-4a1a-9571-51d16f054330
Two things:
- For the past year or more, many models can emit images directly, and all models that matter can see images directly - that's what "multimodal" means. Tokens don't have much to do with textual language anymore, they're more like units of sensory experience.
- Even restricted to text, a language model can operate anything that can be expressed as text, as long as you have a translation layer between textual representation and the final form. That includes giving commands as text. The total addressable space of what models can be used for is, thus, approximately anything humans do.
Should be a Twtich stream tbh
I've seen a few voxel-based benchmarks built around the same idea, with LLMs effectively constructing models using a discrete set of instructions.
https://minebench.ai
Correct me if I'm wrong but this shows that painting capability is emergent, arising from unrelated training data thus very very compelling evidence that LLMs are actually intelligent.
They have been generalizing for a long time, but spacial 2d and 3d art through tool use is a very engaging way to show it. It is harder for people to deny generalization.
Now Opus 5.5 clearly has had some sort of visual arts training, the step function change in ability implies that to me, but it can apply its spacial artistic reasoning to pretty arbitrary tools.
Just share your opinions, dude, we get its ai generated but share your own taste. We want to know what YOU think. Stop being shy and lazy.