How My AI Image Won a Major Photography Competition

2023-04-25
关注

In March the Sony World Photography Awards announced the winning entry in their creative photo category: a black-and-white image of an older woman embracing a younger one, titled “PSEUDOMNESIA: The Electrician.” The press release announcing the win describes the photograph as “haunting” and “reminiscent of the visual language of 1940s family portraits.”

But the artist, Berlin-based Boris Eldagsen, turned down the award. His photograph was not a photograph at all, he announced: he had crafted it through creative prompting of DALL-E 2, an artificial intelligence image generator.

“I applied as a cheeky monkey, to find out if the [competitions] are prepared for AI images to enter. They are not,” Eldagsen explained on his website. His stunt has sparked controversy and conversation about when AI-generated or assisted images should be considered art.

Scientific American spoke with Eldagsen about the image’s creation and the future of AI-aided “promptography.”

[An edited transcript of the interview follows.]

How did you get started with AI art?

I started with photography because drawing was a lonely job. I was always experimenting. So when AI generators started, I was hooked from the very beginning. For me, as an artist, AI generators are absolute freedom. It’s like the tool I have always wanted. I was always working from my imagination as a photographer, and now the material I work with is knowledge. And if you are older, it’s a plus, because you can put all your knowledge into prompting and creating images. If I were 15, I would have probably just generated Batman.

Where did the inspiration for “The Electrician” come from?

I did it for myself as an exercise, and I just love the result. It sparked off from a project that started years back. My father was born in 1924. So he went to war when he was 17 but, like most of that generation in Germany, never talked about it. After his death, I found some images from the forties my mom and I hadn’t seen before. I learned a lot about their time just looking at these images, and I started to collect images from the forties at flea markets, and also on eBay, but didn’t know what to do with them.

So my first experiment was: Can I recreate images of that time using AI? And then “The Electrician” just grew. The best images are those you didn’t have in your mind before. They came out of the process. You start, and it leads you somewhere—with AI it’s the same. You start somewhere and then you make many different decisions. You delete elements, you add frames. Sometimes the AI has very good suggestions. Sometimes it’s just crap. That takes time and patience, so it’s not finished in 20 seconds. It can take days.

So how did you actually make this image?

I used DALL-E 2, and it was all done by text prompts and inpainting and outpainting. For inpainting, you could say, “I don’t like his tie,” and you erase it and write, “I want him to have a white tie.” Then you get suggestions. And if you don’t like any of those suggestions, you start again. Outpainting [is what] you do when the frame is not large enough. You put in an additional frame so you can see his whole tie, his pants, the chair, the floor. It’s endless.

Why did you decide to submit it to a photography competition?

I’ve been very involved in AI and photography—I’ve become one of the experts in Germany—so it’s not just me poking fun. I wanted to test if a competition has taken into account that AI-generated images can be sent in. I applied to three different competitions, and the image always was a finalist. There’s something about the image. When I apply, I don’t say it’s AI-generated. I keep the information very short: just the image and the title. Then when it was selected, I said the art is AI-generated.

What I was hoping for has happened: the conversation has started, and it was basically with the help of the community. I did not expect it to be that big; I thought it would have been a conversation for a week in the European photo scene.

Do you think someone could have figured out that it was AI-generated just by looking at the image?

Of course. There’s a difference in color that comes through outpainting. On the left side it’s too yellow, and on the right side it changes into black and white. And then maybe the fingers, but also part of the arm on the right side, you can tell. If you work with it on a daily basis, like I do, you can tell.

Have you ever been fooled by an AI image?

There’s a German magazine called GEO; it’s something like National Geographic in Germany. They had an online test with images, asking “Is it generated or authentic?” I failed once.

I think with “The Electrician,” it’s very easy to tell because it’s an old image from early last September. But I think by the end of this year, we won’t be able to tell.

Does that alarm you?

As an artist, I just love it. As a citizen, I’m deeply concerned. Most kinds of photography can be augmented by AI but not the photojournalism part. The press needs to come up with a system to make it clear what is authentic, manipulated or generated. The Pope Balenciaga AI-generated photo should have always been notated. If you don’t do that, democracy will be manipulated and misinformed by anyone who can write five words.

But to fact-check is a lot of work. That takes time. So for the press to do that, to pay all the staff and to also have AI technology to help—who is going to pay for this? Now as the citizens, I say, we cannot let the press work alone. It’s very important for a democratic society [to be able to distinguish real photos from fakes]. So we have to think about the structure to co-finance that [fact-checking] as citizens, as a democratic state. But how can we co-finance it and still maintain the freedom of press? This is something we need to think about.

So the future of democracy and journalism aside—how will AI fit in the art world?

One thing I propose is to clean up the terminology and not call realistic AI art “AI photography” anymore, because it’s not photography. And one suggestion that came out of the community was “promptography,” and I just love it. It is large enough to encompass that the result can look like a drawing, like a painting, like a photo.

The next step would be to talk about the relationship between promptography and photography. Do they belong into one basket—one museum, festival, gallery, competition? It’s very complex. And I don’t have any answer for that. The only thing I can say is that the easy answers on both sides—those who want to go back to analog times and those who say promptography is photography—are nonsense. We need to think deeper than that.

参考译文
我的AI生成图像如何赢得一个重要摄影比赛
3月,索尼世界摄影奖宣布了创意摄影类别的获奖作品:一张黑白照片,描绘了一位年长女性拥抱一位年轻女性,标题为《PSEUDOMNESIA:电工》。在宣布获奖的新闻稿中,该照片被形容为“令人不安”,并“让人想起20世纪40年代家庭肖像的视觉语言”。但这位艺术家、柏林的博里斯·埃尔达根(Boris Eldagsen)拒绝了这个奖项。他宣布,这张照片根本不是照片,而是他通过创造性的提示生成的DALL-E 2人工智能图像。“我以一种调皮的方式申请,想知道这些摄影比赛是否为AI图像做好了准备。它们没有做好。”埃尔达根在他的网站上解释道。他的行为引发了关于AI生成或辅助图像是否应被视为艺术的争议和讨论。《科学美国人》杂志对埃尔达根进行了采访,探讨这幅图像的创作过程以及AI辅助“提示摄影”(promptography)的未来。[以下为采访的整理记录。]你是如何开始接触AI艺术的? 我最早从事的是摄影,因为画画是一件孤独的工作。我一直在进行各种实验。因此当AI生成器出现时,我立刻就被吸引住了。作为一名艺术家,AI生成器对我来说是绝对的自由,它正是我一直以来想要的工具。作为摄影师,我一直从自己的想象出发进行创作,而现在我所处理的材料是知识。如果你年纪较大,这反而是一种优势,因为你可以将你所有的知识都融入提示与图像的创作中。如果我15岁,我可能只会生成蝙蝠侠。《电工》的灵感来自哪里? 我做这个项目是为了自己练习,我非常喜欢这个结果。它源于多年前一个项目的启发。我的父亲出生于1924年,17岁就参军了,但像德国当时的大多数人一样,他从未谈论过那段经历。在他去世之后,我发现了一些我和妈妈之前从未见过的40年代的照片。通过这些照片,我了解到他们那个时代的故事,于是我开始在跳蚤市场以及eBay上搜集40年代的照片,但不知道该拿它们怎么办。所以我最初的实验就是:我能否用AI重现那个时代的照片?然后,“电工”就诞生了。最好的图像,是你原本根本没有想到的。它们是在过程中诞生的。你开始创作,它就将你引向某个方向——AI也是如此。你从某个起点开始,然后做出许多不同的决定。你删除元素,添加边框。有时AI的建议非常出色,有时只是垃圾。这需要时间和耐心,所以它不是20秒就能完成的。有时可能需要几天。你是如何创作这幅图像的? 我使用了DALL-E 2,完全是通过文本提示、图像填充(inpainting)和图像扩展(outpainting)完成的。图像填充时,你可以这样说:“我不喜欢他的领带。”然后你擦掉它,再写:“我希望他有一条白色领带。”接着你会得到一些建议。如果你对这些建议都不满意,就重新开始。图像扩展则是在画框不够大的时候进行的。你可以添加一个额外的边框,从而看到他的整条领带、裤子、椅子和地板。这几乎是无限的。你为什么决定将它提交给摄影比赛? 我一直在AI和摄影领域非常活跃,已经成为了德国这个领域的一个专家——所以这不仅仅是我在开玩笑。我想测试一下,摄影比赛是否已经意识到AI生成的图像可能会被提交。我向三个不同的比赛提交了这幅作品,它每次都进入了决赛。这幅图像似乎有一种特质。当我提交时,我并不说明这是AI生成的,我只是给出最简短的信息:图像和标题。当它被选中时,我才说明这幅艺术作品是AI生成的。正如我所期望的那样,这场讨论开始了,而且很大程度上是社区推动的。我没想到它会引起这么大的反响,我以为它只会是欧洲摄影圈里一周的讨论。你认为有人能通过看图像就判断出它是AI生成的吗? 当然可以。通过图像扩展,颜色会有差异。左边太黄,而右边则变成了黑白。然后可能还有手指,以及右边手臂的某些部分,你可以察觉出来。如果你像我一样每天使用它,你是可以看出来的。你有没有被AI图像欺骗过? 在德国有一本杂志叫GEO,有点像德国的《国家地理》。他们曾在线上进行过一个测试,展示图片并提问:“这是AI生成的还是真实的?”我有一次失败了。我觉得《电工》还是很容易分辨的,因为它是去年9月初生成的旧图像。但我觉得到今年年底,我们可能就分不出来了。这让你感到担忧吗? 作为一名艺术家,我非常热爱这一点。但作为一名公民,我却深深忧虑。大多数类型的摄影都可以通过AI进行增强,唯独新闻摄影除外。媒体需要建立一个系统,明确区分真实、经过处理和生成的图像。那张AI生成的教皇和Balenciaga的照片本应一直标注清楚。如果你不做这些,民主制度就会被任何能写出五个词的人所操控和误导。但要验证事实是非常繁重的工作,需要耗费大量时间。对于媒体来说,雇佣足够的人手并配备AI技术来协助验证——这一切谁来买单?作为公民,我认为我们不能让媒体独自承担这一责任。在一个民主社会中,区分真实照片与虚假照片至关重要。所以我们必须思考如何作为公民,作为民主国家,共同承担验证事实的成本。但如何在共同承担的同时仍保障新闻自由?这是我们迫切需要思考的问题。那么撇开民主和新闻的未来不谈——AI将如何融入艺术世界? 我想,首先应清理术语,不要再称逼真的AI艺术为“AI摄影”,因为它并不是摄影。一个来自社区的建议是“提示摄影”(promptography),我非常喜欢这个说法。它足够宽泛,能够涵盖结果看起来像素描、像绘画、像照片的各种形式。下一步就是讨论“提示摄影”与传统摄影之间的关系。它们是否应归入同一个篮子,比如同一个博物馆、同一个艺术节、画廊或比赛?这是一个非常复杂的问题,我对此没有现成的答案。我能说的只有一点,就是两边那些简单的答案——那些想回到模拟时代的,以及认为“提示摄影”就等同于摄影的人——都是没有意义的。我们必须思考得更深入。
您觉得本篇内容如何
评分

评论

您需要登录才可以回复|注册

提交评论

广告
广告
提取码
复制提取码
点击跳转至百度网盘