A couple of years ago, after reading the book, I sketched some of them on Procreate with an isometric grid, but not all 55. I only arrived at number 4, since each took multiple hour...
Since this project (both yours and OP's) seem more for personal fulfillment and enjoyment, I wonder who got more enjoyment out of the process, and also who will remember the results better in ten years.
One is a toy. Created on a whim, it will end up in the bin with other toys. They might pick it up and paw at it later, stir up some memories, and then put it back.
One is art. When completed, they can look at it and be reminded of the journey of creation, a labor of love. It could be put away, or displayed for all to enjoy.
> One is art. When completed, they can look at it and be reminded of the journey of creation, a labor of love. It could be put away, or displayed for all to enjoy.
How is this not applicable to both things? This just reads as gatekeeping to me.
The amount of time or effort put into something does not determine its value.
The amount of time and effort a human put into something very much does tend to determine its value, both to that human and others in their social circle.
> The amount of time or effort put into something does not determine its value.
It literally does, not only in art. Is not the only factor, yet is definitely one of them. In many objects, the version with 'more human effort' is more expensive.
It’s interesting, with a human made piece of art, I’m always drawn to « zoom in », to see how you did little details, if you have hidden something here and there, how you connected different parts.
With AI generated stuff, sure it looks amazing at first glance but I definitely don’t want to zoom in, since I know it’s all a cardboard facade with no passion. Look a bit too closely and you’ll see the bridges that make no sense, the bunch of nonsensical threejs cylinders etc.
I always consider AI-generated content to be a draft — or, in its parlance: something that takes the shape of the final product.
I don't dismiss the ability of models to create. I dismiss their ability to create finished goods. Humans should always be involved, especially with creative works. Models are great at modeling existing creations and solutions — head starts on long roads — but it's the person that needs to walk it and take the journey.
Are they really? Nothing of value is produced unless the exact artifact is transmuted into a professional good? There’s nothing to be gained from sketching preliminary drafts or notes because they’re never polished for publication?
Not what I meant. The point was work under a certain grain size of effort has very little (or negative) value to others.
We're in an environment awash in 'content' where every person has an infinite slop producing machine, and a lot of people are publishing nearly zero effort drafts, awed by the appearance of meaning.
Not much, because they are fundamentally different types of errors. A slightly off perspective doesn’t change what the artist was trying to convey.
In the original link the first city I checked (forgot the name) was one were supposedly you have 5 bridge, every one of them more magnificent than the other. The city looks great at a distance, with glow and particle effects. Zoom in and you’ll see the bridges are actually floating over the river, they are not connecting both banks they are literally in the river parallel to the banks with no way to get in or out of the bridge.
The bridges were the central element of the story. Would a human that spend 10h working on this piece just forget that the bridges, the very central element they are trying to depict, need to actually function as a bridge?
With a human, you may get ugly output but you’ll not get that nonsense.
The only reason a human will not give you a nonsense bridge is that a human is running a 2.5kWh spacial physics calculation engine in his/her head.
Replace 3.js with an engine that embeds physics, ask it to do test fly-through through each city and fix inconsistencies and let's then compare the results.
> Make a three.js (pnpm) visualization of all Invisible Cities by Italo Calvino. Don’t ask questions, it is a one-shot task. You have 6h of work, use it until it becomes a masterpiece.
> Leaving there and proceeding for three days toward the east, you reach Diomira, a city with sixty silver domes, bronze statues of all the gods, streets paved with lead, a crystal theatre, a golden cock that crowns each morning on a tower...
Text in the second person is causing my my D&D instincts to kick in.
I'd much rather see yours than the slop version. One of the nice things about the book is that everybody sees the cities different in their minds eye, this link above came from no mind.
If you haven’t read the book yet, I’m tempted to warn you against opening the visualization. Part of the joy of reading this book is letting your mind’s eye run wild, and I fear if I read it for the first time after toying with this site, I’d be remembering the cities rather than conjuring them.
That is the exact problem with the city of Dinipro.
Citizens of this city high in the sky imagine it so fully in their minds that no one there needs to open their eyes to walk the streets and live their normal lives.
But if one makes the mistake of opening their eyes and tries to see the city directly, they would immediately fall through the clouds to the ground deep below.
I heartily recommend the audiobook, narrated by John Lee.
The book is a trip and the dreamy sort of narration matched it perfectly. My wife and I would spend evenings listening to one city at a time, on a walk around our little park, one headphone in each ear. I don't know if it was just the environment or the time in our lives, but we both really look back fondly on this book.
I read the book decades ago, then I lost it in one of the many relocations. I always thought I should buy it again but never acted on the idea... A few years ago, I was walking by an antiquarian, and noticed a thin unmarked gray spine. I entered, picked it up, paid, and walked away.
PS I didn't open the gallery. Nice idea but I don't care about visualizing something I should feel.
How Opus 5.5 imagines these cities is a question for Opus. The real experience is how we imagine - and I believe there are no two people who have exactly the same mental image of an invisible city.
It's a techie creating the torment nexus but the imagination version. No need to picture in your mind the invisible cities! Time to make them visible!!!
Talk about missing the entire point of the piece of literature.
Maybe the interactive link in the article for the optic focus can lead you on a fun journey instead. And the bottom-right of every visualization is a link to yet another fun diorama to explore.
> I’d be remembering the cities rather than conjuring them
Having looked at some of them (and it being one of my favorite books) ... it won't do that. These visualizations are as terrible as they are unnecessary.
Aphantasia can be cured by deliberately engaging the imagination even when nothing is seen - you might not see anything at first but eventually you gain the ability through letting go of external awareness and developing the imagination pathways.
In fact, having a physicalized 3D representation of the cities is IMHO antithetical to the text. I don't think OP ever stopped to ask if there was more to Calvino's work than 'cool city descriptions'.
Invisible Cities is my favorite piece of literature hands down. The audiobook narrated by Richard Higgins is also done so beautifully and can't recommend it more.
As neat as this project is, this does no justice to whatsoever to the book. On the surface Invisible Cities seems to be a book about random cities' descriptions. In reality though, this book is about semiotics, meaning, language, and it's limits.
I recently saw a project that intended to turn House of Leaves into a game. The graphics were shoddy and the mechanics were dull, but by far the worst aspect was how it "paraphrased" passages from the book into Claudespeak (allegedly for copyright reasons).
Seeing the same thing done to Calvino (and Weaver's) sublime prose feels downright criminal.
Just as an example, here's the beautiful opening of the book:
Kublai Khan does not necessarily believe everything Marco Polo says when he describes the cities visited on his expeditions, but the emperor of the Tartars does continue listening to the young Venetian with greater attention and curiosity than he shows any other messenger or explorer of his. In the lives of emperors there is a moment which follows pride in the boundless extension of the territories we have conquered, and the melancholy and relief of knowing we shall soon give up any thought of knowing and understanding them. There is a sense of emptiness that comes over us at evening, with the odor of the elephants after the rain and the sandalwood growing cold in the braziers, a dizziness that makes rivers and mountains tremble on the fallow curves of the planispheres where they are portrayed, and rolls up, one after the other, the despatches announcing to us the collapse of the last enemy troops, from defeat to defeat, and flakes the wax of the seals of obscure kings who beseech our armies' protection, offering in exchange annual tributes of precious metals, tanned hides, and tortoise shell. It is the desperate moment when we discover that this empire, which had seemed to us the sum of all wonders, is an endless, formless ruin, that corruption's gangrene has spread too far to be healed by our scepter, that the triumph over enemy sovereigns has made us the heirs of their long undoing. Only in Marco Polo's accounts was Kublai Khan able to discern, through the walls and towers destined to crumble, the tracery of a pattern so subtle it could escape the termites' gnawing.
And here's the AI distillation:
Kublai Khan does not necessarily believe everything Marco Polo says of the cities he has visited on his missions, but the emperor of the Tartars listens to the young Venetian more attentively than to any other envoy. In the melancholy hour when an emperor learns that his boundless empire is only an endless, formless ruin, the reports of Marco Polo allow him to glimpse, through the walls and towers destined to crumble, the tracery of a pattern so fine it could escape the termites’ gnawing.
Feels like listening to music at 2x speed for the "efficiency."
It's wild how fast "I did [thing that sounds pretty cool] with [model] X.Z" loses me.
I always think that it will interest me, the overview sounds cool, I click around for a few seconds, and then... my eyes glaze over. I feel nothing. I clicked away in under 45 seconds.
It makes me sad, too, because I love a passion project.
I am also stunned, because 5 years ago if a dev built something like this I would be enthralled - "How did you choose the engine? What was the hardest part? Did you like the UI framework?"
I understand why that’s impressive. But I don’t feel much. I guess, I was just used to envy how much people can be absorbed by some idea to spend so much time implementing it. Now it feels like a single shotted nice-typography-and-all presentation at work that someone did 10 min before the meeting.
Weird way to look at it, from another view we can now print a week salary in 10 minutes. Might devalue the output slightly but we still have the product.
Undoubtedly a fundamental reconsideration of political economy is necessary if we are going to go all in with AI. As of now, this technology will exacerbate the existing 'singularities' and wealth disparity distortions in the socio-economic body. And now, basic fucking income is not the answer. Maybe pitchforks.
> If something that previously took weeks for a human to design, now takes 10 min, that means a whole week of salary went up in smoke.
To me, this is the depressing sentiment. The idea that, somehow, the goal isn't to produce useful things but to spread work out over a period of time in order to maximize payment.
At a high level,I conceptualize my job as "doing useful things." Whether that's design, writing code, debugging, sysadmin, CI/CD, writing doc, whatever. I just want to be useful. If there's some tool that makes that easier, I am happy about it. I don't think that tool will result in my getting paid less. But if it does, oh well.
>But if it does, oh well.
I'm glad our ancestors didn't take that line of reasoning we'd still be feudal serfs. "The lord has come and taken all our grain again? Oh well, at least I've been useful to him."
If our ancestors worried more about technology obsoleting human labor we'd still be subsistence farmers. Priorizing usefulness is why we're rich today.
There's also an eerie feeling of emptiness to the outputs. They lack real intent. Glossy and technically impressive, but they don't tickle your brain. I think it's because they don't really communicate or give you that buzz of new insights. They are pretty but not _helpful_.
For example, I didn't learn about lenses or earthquakes or fusion messing with the controls on these.[1] I didn't know what to look at.
However, there were interactive exhibits at the Children's Museum I spent a long time exploring and loved. That's because they were carefully designed to really teach the concept?
Yeah, the demos look really cool but you're thrown into them and there's nothing explained as to what you're looking at or why you should care. Funny, an article explaining something like how a camera focus works with a few illustrations would be far more useful than a fancy 3D interactive.
It turns out the more technology you throw at something does not equate to ease of learning.
Isn’t that at least in part because the intention here was to demo the functionality?
Put in the hands of the educator, they will be able to add those illustrations, design for the light bulb moment, etc, and the tools at their disposal just got much better.
If it makes you feel better.. I find this impressive but it doesn't pique my curiosity. I may get biased, finding a reason not to care.. but honestly I think if this were presented differently I would. And it's not so much about using AI either. I think you could use Ai to make something compelling but so far the one shot prompt outputs to me lack focus
It's low entropy. The amount of entropy is prompt + model but now that you have seen the model various times, the entropy comes down back to the prompt only.
You'd not be amazed by seeing a 2 sentences prompt and thus you are not amazed by this work either.
"What we obtain too cheap, we esteem too lightly: it is dearness only that gives every thing its value. Heaven knows how to put a proper price upon its goods; and it would be strange indeed if so celestial an article as freedom should not be highly rated." – Thomas Paine
I've been thinking of this line recently — really gets at the heart of the matter for me. There is no substitute, no shortcut, no alternative for effort.
That might be true but I don't think that's the problem here. I have seen plenty of AI generated images that evoke awe, nostalgia, other emotions - engaging me for a short time, and a few carefully prompted ones that work for a little longer. This fails to relative even these.
Another thing to keep in mind is that if you asked twelve different artists to render these cities from their descriptions, you would likely end up with twelve very different renditions.
Now imagine the same thing but each person is using Opus 5.5 - I’d wager there would be a lot of commonality among each of the outputs.
Depends, really. Ten different artists using Opus 5.5 and with some experience with LLMs, outputs may differ quite widely due to:
- Each phrasing the prompt(s) differently
- Each having different history accumulated in context
If you asked me to issue the same prompt as you to write some web artifact, chances are mine would be substantially stylistically different, simply because at some point in the past half-year, Clause started styling everything in LCARS form for some reason.
Whether the differences would be substantial or surface-only? Probably depends, but chances for substance increase with iteration. Real artists don't zero-shot their masterpieces either.
Sure, but I'm referring more to this specific post e.g. "They gave Opus 5.5 one prompt, so it was more like a one-shot".
I’m saying if you provide the same prompt to Opus 5.5 across 12 different people, you’re going to get some similarities that immediately arise.
I’ve mentioned this across many HN threads - I call it the LLM corollary to “garbage in, garbage out”: "generic in, generic out". If you prompt with more specificity, you’re obviously going to get different and potentially better results. But at that point, the artist is doing more of the work.
A while back, I tried a similar test asking for LLMs to come up with engaging D&D puzzles that would fit organically into an underground labyrinth environment. This is a pretty general prompt. Let's just say that the results were about as inspired as a tepid bowl of tapioca.
Yeah, I haven't read the book but these models don't evoke any particular feeling in my glance at them. And I have seen AI-generated stuff that's decent at creating a mood. I suspect that what AI can't do currently is produce a finely tuned mood, it just has the equivalent of "presets". I mean, Hollywood is hardly better than that today as well.
Arts and Crafts Movement 2.0. Bring back handcrafting starting with woodworking, pottery and oil painting. And only objects that cannot currently be made by robots, e.g. 3d printing or print illustrations.
The first city I clicked on at random had the opening phrase:
> Arriving, you rejoice at its bridges, each different
The model city had like five bridges total, three of which were didn't bridge anything at all but instead sat in the middle of the river and went along its flow, the "bridges" for some reason pretending to be small islands.
Kublai Khan received an ornate letter signed by Marco Polo: "In Madrid, City of Lost
Things, no item remains where it was set. If one drops his key in the dirt, he may
never re-enter his home, and, even if he manages to stoop and recover the key, he may
rise to find a tulip garden where his house once stood. In complementary fashion,
things lost by others are forever turning up: A pocket watch on a coffee table. A fond
memory in your recollection. I even know of a prince who turned up in a prison cell.
When he appealed to the guards for his release, he failed to find the crown on his
head, and when he was asked his name, he searched his thoughts, but could not find it.
Indeed, the only hope now for the release of this prince of Spain is for you to send
back 300 ducats for his release. Of course, he will reward you handsomely once he is
out. Yours truly, Marco." Kublai Khan cocked an eyebrow and declared before his court,
"Hey, everyone: Looks like we're about to get ripped off by the guy who traded gold
for paper!" The court erupted in booming laughter.
— Italo Calvino, Invisible Cities 2: This Time It's Visible.
Like many other people in this thread, I feel like passing what would essentially be a hobby project off to Claude robs the creator of most the fulfillment and gives the viewer less reason to appreciate it. But with the models being as good as they are today, I genuinely feel conflicted. If your end goal is to share something with others, why spend a weekend on a project like this when you could just hand it off to Claude?
Your appreciation should come from the one shot prompt. It's a direct comparison to Astra's one shot. I feel like you and many others are trying to "appreciate" the wrong thing.
If the creator truly wanted viewers appreciate the end product they'd start with the one shot output and spend their weekend iterating on each city either by hand or with Claude. But they didn't want to do that, they wanted to show off what Claude/GPT could do in 6hrs. That's the whole blog post.
This isn't a ceiling, though. This is the new floor: What you can get with more or less zero human effort and ~$100.
I expect we're going to see more and more amazing "hobby projects" that did involve both a few hundred dollars in tokens and hours of human creativity and effort in combination.
I get your point, and I think it's broadly true. But I'm not building things with AI to impress others, and it could be that GP or OP are doing the same.
Unasked-for-story-time: In the early aughts, there was a PC video game called Star Trek: Elite Force. It was a FPS genre game and included a feature called "Virtual Voyager" that at the time was revolutionary because it would let you just freely explore a subset of the Starship Voyager without any linear mission requirements, the closest any Trek fans could get to gaining access to the real Paramount Studios Stage 9-and-friends sets where the TV show was filmed, which for Trekkies with no life is quite the big deal.
Wasting the GPT-6 Astra capabilities over a single month of the $200 plan, I've built out multiple decks of an Excelsior-class starship based on a deck plan of my design, dealing with creative decisions on how to adapt the limited Star Trek canon knowledge about the interior deck sizes and materials aboard an Excelsior-class vessel, and at the end of this project, all I will have will be a graphically crappy but 3d explorable virtual starship I can walk around, look at the interior of, and mentally place my daydreamed Star Trek personal fanfic that I used to pencil in a notebook, and now I direct/dictate/produce using agents to write their crappy sloppy prose, but is tuned exactly to my preferences, and now will be able to have specific camera angles and relative details because I can screenshot the 3d ship explorer I made with the crappy AI.
I've got the pencil, I've got the nerd no social life time for creative projects, and I am still writing the no-AI no-name fantasy novel that nobody will read except my family and AI-data-gorgers, but neither of the two projects above, which have provided me hours of entertainment dopamine, would have happened without AI tooling making it easy enough to be creative without having to hand render every little detail.
Art is produced for arts own sake, right? That's the phrase? Well this vibe slop story and this vibe slop 3d viewer are digital goods / digital art to _me_, because I like them for their own sake. I can't copyright any of it, its all been infected by AI stuff. I can't get a job with it, I can barely orbit in Blender and don't find it worth the time to learn to render drawing in 2d and sculpting in 3d (I've tried both over the years in pre and post AI eras, I'm from a creative family whether I wish it or not), the costs to render skills manually make it not worth my time to do.
But AI makes it worth my time because it is an enzyme that reduces the input cost of me prototyping my creative desire out without having to laboriously create every last everything.
That makes sense. The difference would be if you were to put it out there as proof of your ability rather than as proof of your imagination. Now I'd like to see what you made :)
Why would that someone want to listen something you haven't created? What have you "shared"? Claude could have made the tools that let the creator build their cities from their own imagination, but that's not what this is.
One of my favorite things about the National Treasure Cinematic Universe (two movies and a TV show) is the existence of characters that I like to call "treasure grumps". Their role in the story is to warn the main characters that searching for treasure is a terrible way to spend your time, and can only lead to personal and familial ruin.
Meanwhile, here in the non-fictional world, an estimated 20-30 million dollars have been spent excavating Oak Island with absolutely zero treasure of any sort found.
To Someone who is passionate enough to write about it, do something to build about it, share about it?
unlike someone who is just whining about it?
Wtf are you to someone who has put effort to spread about the book? I never seen/heard of the book, and now I am iintersted in it thanks tot he author.
The Opus demo was a tad garish and overwrought. The Astra one was supposed to feel lighter but it was also over-encumbered by UI. Both were hard to navigate, despite the UI trying to seem navigable.
I don't mind the AI but the human needed to put more time into curation and getting the magic right.
Yeah, I was wondering why the two results were so similar, despite lack of detail in the prompt. Both have an array of "miniature" cities on a game-like map, all of them drawn surrounded by a circular border. Something sketchy is happening here. Either more prompt or intervention is going into it than the author claims, or it's copying some preexisting portrayal.
$74 vs $ 10 vs $25 for the same prompt,
The interesting number is not actually quality, its what we can get per dollar like 6 subagents runs in parallel.
Yeah for now, this week, on that model, all of which are constantly changing. I would love to know actual cost of running those inferences on such hardware, including electricity, training, etc. i.e. what we should be paying. I can imagine that job would cost multiple hundreds if they were charging enough to actually be profitable companies.
TBH only scanned through ~20, impressive technically but disservice to Invisible Cities to render it as... uninspired, bland, miniatures. All the wrong of details to focus on to capture even notional urban verisimilitude. Looks like a 8 year old drew scribbles and their dad finished the project.
I was finishing Invisible Cities yesterday, because the book was due at my local library today. I was wondering whether the cities could be adapted into some visual art form. And obviously I thought about feeding it to Opus. 24 hours later, I see this post…
Citizens there gather their tools at the bottom of a giant staircase and climb up countless stairs, only to find out that what they were planning to do has already been done by the time they reach the top.
Dissapointed, they descend to grab a new set of tools and make new plans for tomorrow, only for the same thing to happen again.
The whole point of the OP's article is: Opus 5.5 is so much better at one-shot design! Also, your link is not public - I'm curious to see your Opus 4.8 comparison
And shame on all the people shitting on it because AI made it. If you valued art for art's sake, you wouldn't be so bothered when somebody makes something art-adjacent using some technique you don't approve of. It in no way detracts from the kind of art you admire. Don't be type who just has to let others know that you find something they like to be beneath you.
This is probably one amazing positive these models have brought forward - the ability to break out of a single modality, and utilise others to help us learn and visualise. While the visuals here are amazing, my son for example prefers to learn by listening and talking, so we convert lot of his study materials into audio and real-time voice roleplay.
Wonder how would Fable 5.1 perform. I guess to expensive for such experiment. I guess you had Max 20x. Still it taking 3 hours is significant time for agent. What thinking level you had it on?
First off - I'm a senior dev with ~30 years of experience in graphics engines, game dev, web dev, WebGL stuff, blah blah. I am generally skeptical of AI, but... after reading the top of every thread in this comment section, I feel like people forgot that they could zoom in on the cities to see all the details.
The amount of _stuff_ crammed into each little city is almost unbelievable. Drawbridges open and close, glowing lines sketch patterns, camels walk across a tiny desert. Buildings fade into existence and then burst in a shower of sparks. Flags fly, smoke drifts, buckets cause ripples as they dip into underground pools. The city made of plumbing has tiny people in the bathtubs! There's a freakin' roller coaster with moving cars and a ferris wheel!
I absolutely understand the AI hate that's being commented here, but... in this instance, I just can't feel it. I keep finding more and more little details every time I click on a city, and I am astounded. I just now noticed there's a little guy in the city of strings hanging up new strings at the top of the hill. Damn. My favorite is Marozia, where dark buildings open like treasure chests and release shining towers and flocks of glowing birds.
Yes, there are also a lot of stupid AI errors - every spoked wheel has the spokes off-center for example. But holy crap, I just can't hate this. I don't care about who built it or the random bits of jank - this thing is _beautiful_ in a way that I haven't seen from AI before, and it warms my old GPU-powered heart.
Even as I was writing this I kept finding more things I didn't notice on my first few passes through the cities:
Phyllis fades from color to black and white except for two houses.
The fairground has a big top with tiny bleachers and a guy inside.
The happy city has people with their arms raised in a V.
People in the marketplace city wear hats and some are carrying bundles on their heads.
There are weird cactus plants in the hanging city.
In the half-fairgroud city, there's a statue on one of the trucks.
The city that copies the dead has a set of pallbearers carrying a wrapped corpse down the stairs.
The people on top of Argia are laying with their ears to the ground, listening for the animated sound waves coming from below.
There's a tiny cat on one of the roofs in Raissa!
The rats in Marozia hop along their paths!
The center of Theodora has bookshelves with individual books!
But who cares about those details if they're just random? Why are they there? What do they mean? What was the intention? Sure it's jaded, but just cramming in elements doesn't necessarily impress, other than technically it's better at not overlapping random stuff.
both the linked camera lens example as well as the last one the author linked runs at ~5fps on vivaldi and pegs CPU 100%. i have never seen any 3d visualization website have performance this bad, i assume they're not using webgl given my GPU is 0% utilization.
It reminds me of Total War. The soundtrack and way that the map is shown. Is Total War an inspiration? Anyways, I liked the visualization. But I never read the book, so I followed (partly) droidjj recommendation
While reading the book years ago, I did not visualize Isidora as a city with 9 houses, like a children’s book, or a little video game level intro, but so unlike Kublai Khan or Italo Calvino.
wow that's awesome! super cool :), fyi so inspiring that I did something similar for d&d forgotten realms, including a timeline: https://narfman0.github.io/realms-atlas/
fable drove it, opus agents did the work, took 30-45 minutes. (intend on extending to planescape, elaborating on significant historical events, and some text to speech narrating cool events)
It's one of my favorite books - and I think it deserves better; it's a conceptually very dense text and just slapping together a few 3D models misses the point
it's also not very astonishing that the current models are able to create this - it looks exactly as expected
and I actually think getting an interesting AI generated rendition of the text from a single prompt is a worthwhile endeavor
maybe even a good replacement for the pelicans: it would test if a model can engage with a text beyond the surface level
No, why would it be? Unless very detailed web-based microphone or camera manipulation is needed, a modern LLM can built a fallback using safer technology instead.
I'm a little torn by this sort of demonstration, because while many of the scenes have obvious markers of slop (impossible intersecting geometry, bridges to nowhere, and so on), the scenes largely do work to convey the intended concept/emotion, and the low-poly aesthetic is executed decently well.
I guess I shouldn't be surprised that another visual medium is starting to fall to LLMs after the success of diffusion models for image generation. Nevertheless it's wild that a mostly text-focussed model is building little worlds like these.
What is this for? It doesn't deepen the understanding of the book. The illustrations aren't attractive on their own. They're not a very good representation of the descriptions in the book.
Apologies for the long post in advance (maybe I should turn this into a blog post).
Invisible Cities is one of my favorite books, and I've re-read it several times. It's a great "travel" read and always puts me in the mood to explore. It's also a very romantic book, obviously the cities have all named after women, so it evokes conquest in the sexual sense as well. It's also a piece of surrealist/allegorical literature in the same vein as Daumal's Mount Analogue (which also tops my list).
Back to the website. While this is an impressive tech demo (as they all are), it's clear that AI misses the forest for the trees, and here's a tangible example. I randomly clicked on Theodora. Honestly, I didn't even remember which city this was until I re-read the Calvino passage:
> Recurrent invasions racked the city of Theodora in the centuries of its history; no sooner was one enemy routed than another gained strength and threatened the survival of the inhabitants. When the sky was cleared of condors, they had to face the propagation of serpents; the spiders' extermination allowed the flies to multiply into a black swarm; the victory over the termites left the city at the mercy of the woodworms. One by one the species incompatible to the city had to succumb and were extinguished. By dint of ripping away scales and carapaces, tearing off elytra and feathers, the people gave Theodora the exclusive image of human city that still distinguishes it.
So, what do we know about this city? We can make a short sensible list:
1. It's an imperialistic (or at least millitaristic) reflection of man's will to power (as Nietzsche might put it)
2. It's decidedly anti-nature, the natural flora and fauna being its primary antagonist
3. It's been destroyed and rebuilt many times
4. The city is built by literally ripping away bugs and birds; tearing their wings from their bodies like a cruel child might
5. Nature is slowly winning (and maybe it already has)
This[1] is what AI saw. Now, this isn't bad, really. But it's so embarassingly pedestrian, it hardly counts as a visualisation. It just heard "city with animals and stuff" and made a 3D model. But it's all, in that uncanny-valley kind of feeling, wrong. Why is the city itself a grid? Why is there a library in the center?
Calvino, again, on Theodora:
> But first, for many long years, it was uncertain whether or not the final victory would not go to the last species left to fight man's posession of the city: the rats.
The vibe here is an overrun rat's nest of a city. A disgusting bloated corpse; more:
> The city, great cemetery of the animal kingdom, was closed, aseptic, over the final buried corpses with their last fleas and their last germs.
I see an overrun New York. I see a city built on a decaying graveyard. AI sees animals and humans living in a kumbaya harmony, but that is not what Theodora is: it's an eternal conflict between man and nature (and nature won a long time ago). For the record, this isn't even that particularly deep. Many authors and poets write about this eternal human struggle: the taming of the wilds. But Claude just doesn't get it. I hope readers of the book do.
Anyway, I leave you with this. I found an artist, Rian Hotton[2], who put brush to canvas to show us what he thought Theodora looks like[3]. And this... this is more like it.
You should avoid add model name (or at least use Open weight model names) in your title, so it would be 90 (amazing) : 10 (lost the book value) comments. /s
A couple of years ago, after reading the book, I sketched some of them on Procreate with an isometric grid, but not all 55. I only arrived at number 4, since each took multiple hour...
https://camillovisini.com/drawing/fc4jc6-le-citta-invisibili...
https://camillovisini.com/drawing/p2yrdj-le-citta-invisibili...
That looks pretty nice, I didn't know Procreate could be so good at precise line art!
Since this project (both yours and OP's) seem more for personal fulfillment and enjoyment, I wonder who got more enjoyment out of the process, and also who will remember the results better in ten years.
One is a toy. Created on a whim, it will end up in the bin with other toys. They might pick it up and paw at it later, stir up some memories, and then put it back.
One is art. When completed, they can look at it and be reminded of the journey of creation, a labor of love. It could be put away, or displayed for all to enjoy.
> One is art. When completed, they can look at it and be reminded of the journey of creation, a labor of love. It could be put away, or displayed for all to enjoy.
How is this not applicable to both things? This just reads as gatekeeping to me.
The amount of time or effort put into something does not determine its value.
The amount of time and effort a human put into something very much does tend to determine its value, both to that human and others in their social circle.
> The amount of time or effort put into something does not determine its value.
It literally does, not only in art. Is not the only factor, yet is definitely one of them. In many objects, the version with 'more human effort' is more expensive.
These are amazing!
It’s interesting, with a human made piece of art, I’m always drawn to « zoom in », to see how you did little details, if you have hidden something here and there, how you connected different parts.
With AI generated stuff, sure it looks amazing at first glance but I definitely don’t want to zoom in, since I know it’s all a cardboard facade with no passion. Look a bit too closely and you’ll see the bridges that make no sense, the bunch of nonsensical threejs cylinders etc.
I always consider AI-generated content to be a draft — or, in its parlance: something that takes the shape of the final product.
I don't dismiss the ability of models to create. I dismiss their ability to create finished goods. Humans should always be involved, especially with creative works. Models are great at modeling existing creations and solutions — head starts on long roads — but it's the person that needs to walk it and take the journey.
This pairs well with the saying "art is never finished, it is abandoned."
Abandoned one-shot drafts are akin to speaking words but never finishing a sentence before moving onto the next thought. There's no there there.
Are they really? Nothing of value is produced unless the exact artifact is transmuted into a professional good? There’s nothing to be gained from sketching preliminary drafts or notes because they’re never polished for publication?
Not what I meant. The point was work under a certain grain size of effort has very little (or negative) value to others.
We're in an environment awash in 'content' where every person has an infinite slop producing machine, and a lot of people are publishing nearly zero effort drafts, awed by the appearance of meaning.
What happens when you zoom in to a human-made painting and notice an error?
Depends on whether it is a happy error.
Not much, because they are fundamentally different types of errors. A slightly off perspective doesn’t change what the artist was trying to convey.
In the original link the first city I checked (forgot the name) was one were supposedly you have 5 bridge, every one of them more magnificent than the other. The city looks great at a distance, with glow and particle effects. Zoom in and you’ll see the bridges are actually floating over the river, they are not connecting both banks they are literally in the river parallel to the banks with no way to get in or out of the bridge.
The bridges were the central element of the story. Would a human that spend 10h working on this piece just forget that the bridges, the very central element they are trying to depict, need to actually function as a bridge?
With a human, you may get ugly output but you’ll not get that nonsense.
The only reason a human will not give you a nonsense bridge is that a human is running a 2.5kWh spacial physics calculation engine in his/her head.
Replace 3.js with an engine that embeds physics, ask it to do test fly-through through each city and fix inconsistencies and let's then compare the results.
> Make a three.js (pnpm) visualization of all Invisible Cities by Italo Calvino. Don’t ask questions, it is a one-shot task. You have 6h of work, use it until it becomes a masterpiece.
Love it. I see isometric angles - I updoot.
> Leaving there and proceeding for three days toward the east, you reach Diomira, a city with sixty silver domes, bronze statues of all the gods, streets paved with lead, a crystal theatre, a golden cock that crowns each morning on a tower...
Text in the second person is causing my my D&D instincts to kick in.
SOUL
You draw what you saw in your minds' eye. I can guarantee the author of this didn't visualize the cities like they came out.
I'd much rather see yours than the slop version. One of the nice things about the book is that everybody sees the cities different in their minds eye, this link above came from no mind.
If you haven’t read the book yet, I’m tempted to warn you against opening the visualization. Part of the joy of reading this book is letting your mind’s eye run wild, and I fear if I read it for the first time after toying with this site, I’d be remembering the cities rather than conjuring them.
That is the exact problem with the city of Dinipro.
Citizens of this city high in the sky imagine it so fully in their minds that no one there needs to open their eyes to walk the streets and live their normal lives. But if one makes the mistake of opening their eyes and tries to see the city directly, they would immediately fall through the clouds to the ground deep below.
On that note, I'd quite recommend the Coyote vs. Acme movie, some great world building there.
Where is the city of Dinipro from? It's not one of the 55 from this book and I can't find anything about what you're describing on the internet.
I opened it because I had no idea what Invisible Cities meant here. I thought they might be some real cities that don't come up on maps.
The website gives you a good hint before opening up the visuals. So, I got an idea.
I don't really know about Marco Polo. My only knowledge comes from the Netflix series, which doesn't talk about these invisible cities.
Invisible Cities is a book.
I heartily recommend the audiobook, narrated by John Lee.
The book is a trip and the dreamy sort of narration matched it perfectly. My wife and I would spend evenings listening to one city at a time, on a walk around our little park, one headphone in each ear. I don't know if it was just the environment or the time in our lives, but we both really look back fondly on this book.
I read the book decades ago, then I lost it in one of the many relocations. I always thought I should buy it again but never acted on the idea... A few years ago, I was walking by an antiquarian, and noticed a thin unmarked gray spine. I entered, picked it up, paid, and walked away.
PS I didn't open the gallery. Nice idea but I don't care about visualizing something I should feel.
I fully agree, thank you for this warning.
How Opus 5.5 imagines these cities is a question for Opus. The real experience is how we imagine - and I believe there are no two people who have exactly the same mental image of an invisible city.
It's a techie creating the torment nexus but the imagination version. No need to picture in your mind the invisible cities! Time to make them visible!!!
Talk about missing the entire point of the piece of literature.
Maybe the interactive link in the article for the optic focus can lead you on a fun journey instead. And the bottom-right of every visualization is a link to yet another fun diorama to explore.
Don't worry about opening the GPT-6 Astra one though, those ones are so bad that your brain will immediately expunge them.
> I’d be remembering the cities rather than conjuring them
Having looked at some of them (and it being one of my favorite books) ... it won't do that. These visualizations are as terrible as they are unnecessary.
Checking Amazon, several of the available editions are illustrated, which I suppose would have the same problem.
Only applicable to those who have a "mind's eye". Some of us don't.
Aphantasia can be cured by deliberately engaging the imagination even when nothing is seen - you might not see anything at first but eventually you gain the ability through letting go of external awareness and developing the imagination pathways.
I did some hopeful googling, and curing aphantasia seems to not be a real thing, unless you believe in magic. I'm disappointed.
https://www.researchgate.net/publication/405434749_A_unifyin...
In fact, having a physicalized 3D representation of the cities is IMHO antithetical to the text. I don't think OP ever stopped to ask if there was more to Calvino's work than 'cool city descriptions'.
Invisible Cities is my favorite piece of literature hands down. The audiobook narrated by Richard Higgins is also done so beautifully and can't recommend it more.
As neat as this project is, this does no justice to whatsoever to the book. On the surface Invisible Cities seems to be a book about random cities' descriptions. In reality though, this book is about semiotics, meaning, language, and it's limits.
Calvino's other creations are good too and recommended. I really enjoyed Cosmicomics!
Seconding this! It's a great pairing.
We read this in my undergrad phenomenology class and it blew my mind. I didn't know art could be this rich or deep
I recently saw a project that intended to turn House of Leaves into a game. The graphics were shoddy and the mechanics were dull, but by far the worst aspect was how it "paraphrased" passages from the book into Claudespeak (allegedly for copyright reasons).
Seeing the same thing done to Calvino (and Weaver's) sublime prose feels downright criminal.
Just as an example, here's the beautiful opening of the book:
Kublai Khan does not necessarily believe everything Marco Polo says when he describes the cities visited on his expeditions, but the emperor of the Tartars does continue listening to the young Venetian with greater attention and curiosity than he shows any other messenger or explorer of his. In the lives of emperors there is a moment which follows pride in the boundless extension of the territories we have conquered, and the melancholy and relief of knowing we shall soon give up any thought of knowing and understanding them. There is a sense of emptiness that comes over us at evening, with the odor of the elephants after the rain and the sandalwood growing cold in the braziers, a dizziness that makes rivers and mountains tremble on the fallow curves of the planispheres where they are portrayed, and rolls up, one after the other, the despatches announcing to us the collapse of the last enemy troops, from defeat to defeat, and flakes the wax of the seals of obscure kings who beseech our armies' protection, offering in exchange annual tributes of precious metals, tanned hides, and tortoise shell. It is the desperate moment when we discover that this empire, which had seemed to us the sum of all wonders, is an endless, formless ruin, that corruption's gangrene has spread too far to be healed by our scepter, that the triumph over enemy sovereigns has made us the heirs of their long undoing. Only in Marco Polo's accounts was Kublai Khan able to discern, through the walls and towers destined to crumble, the tracery of a pattern so subtle it could escape the termites' gnawing.
And here's the AI distillation:
Kublai Khan does not necessarily believe everything Marco Polo says of the cities he has visited on his missions, but the emperor of the Tartars listens to the young Venetian more attentively than to any other envoy. In the melancholy hour when an emperor learns that his boundless empire is only an endless, formless ruin, the reports of Marco Polo allow him to glimpse, through the walls and towers destined to crumble, the tracery of a pattern so fine it could escape the termites’ gnawing.
Feels like listening to music at 2x speed for the "efficiency."
It's wild how fast "I did [thing that sounds pretty cool] with [model] X.Z" loses me.
I always think that it will interest me, the overview sounds cool, I click around for a few seconds, and then... my eyes glaze over. I feel nothing. I clicked away in under 45 seconds.
It makes me sad, too, because I love a passion project.
I am also stunned, because 5 years ago if a dev built something like this I would be enthralled - "How did you choose the engine? What was the hardest part? Did you like the UI framework?"
But the author didn't experience any of that...
I'm finding that one of the most difficult parts for me of the AI takeover is the loss of stories.
I understand why that’s impressive. But I don’t feel much. I guess, I was just used to envy how much people can be absorbed by some idea to spend so much time implementing it. Now it feels like a single shotted nice-typography-and-all presentation at work that someone did 10 min before the meeting.
Looking at the posts on HN AI topics lately, this seem to be the depressing sentiment.
If something that previously took weeks for a human to design, now takes 10 min, that means a whole week of salary went up in smoke.
When have we automated enough?
> If something that previously took weeks for a human to design, now takes 10 min, that means a whole week of salary went up in smoke.
The problem also is, if a human took weeks to design this, it would be much better. This is quick but also bad.
This really sounds like a bias selection filter. Humans spend a lot of time on stuff that's shit too.
Weird way to look at it, from another view we can now print a week salary in 10 minutes. Might devalue the output slightly but we still have the product.
Who is "we"?
Derrek Fox, Rupert Hunt & myself of course, who else could I possibly have meant?
This is so often the essential question, and equally often unanswered
Undoubtedly a fundamental reconsideration of political economy is necessary if we are going to go all in with AI. As of now, this technology will exacerbate the existing 'singularities' and wealth disparity distortions in the socio-economic body. And now, basic fucking income is not the answer. Maybe pitchforks.
If basic income were to be the answer, the question would likely require pitchforks.
The person controlling the salary
> If something that previously took weeks for a human to design, now takes 10 min, that means a whole week of salary went up in smoke.
To me, this is the depressing sentiment. The idea that, somehow, the goal isn't to produce useful things but to spread work out over a period of time in order to maximize payment.
At a high level,I conceptualize my job as "doing useful things." Whether that's design, writing code, debugging, sysadmin, CI/CD, writing doc, whatever. I just want to be useful. If there's some tool that makes that easier, I am happy about it. I don't think that tool will result in my getting paid less. But if it does, oh well.
>But if it does, oh well. I'm glad our ancestors didn't take that line of reasoning we'd still be feudal serfs. "The lord has come and taken all our grain again? Oh well, at least I've been useful to him."
If our ancestors worried more about technology obsoleting human labor we'd still be subsistence farmers. Priorizing usefulness is why we're rich today.
The same value was created, and a whole week of free time was made available?
just ignore what HN said, most people still in denial mode
There's also an eerie feeling of emptiness to the outputs. They lack real intent. Glossy and technically impressive, but they don't tickle your brain. I think it's because they don't really communicate or give you that buzz of new insights. They are pretty but not _helpful_.
For example, I didn't learn about lenses or earthquakes or fusion messing with the controls on these.[1] I didn't know what to look at.
However, there were interactive exhibits at the Children's Museum I spent a long time exploring and loved. That's because they were carefully designed to really teach the concept?
[1] https://sael.net/fusion-pulse/?ref=earthquake-tower
Yeah, the demos look really cool but you're thrown into them and there's nothing explained as to what you're looking at or why you should care. Funny, an article explaining something like how a camera focus works with a few illustrations would be far more useful than a fancy 3D interactive.
It turns out the more technology you throw at something does not equate to ease of learning.
Isn’t that at least in part because the intention here was to demo the functionality?
Put in the hands of the educator, they will be able to add those illustrations, design for the light bulb moment, etc, and the tools at their disposal just got much better.
If it makes you feel better.. I find this impressive but it doesn't pique my curiosity. I may get biased, finding a reason not to care.. but honestly I think if this were presented differently I would. And it's not so much about using AI either. I think you could use Ai to make something compelling but so far the one shot prompt outputs to me lack focus
It's low entropy. The amount of entropy is prompt + model but now that you have seen the model various times, the entropy comes down back to the prompt only.
You'd not be amazed by seeing a 2 sentences prompt and thus you are not amazed by this work either.
"What we obtain too cheap, we esteem too lightly: it is dearness only that gives every thing its value. Heaven knows how to put a proper price upon its goods; and it would be strange indeed if so celestial an article as freedom should not be highly rated." – Thomas Paine
I've been thinking of this line recently — really gets at the heart of the matter for me. There is no substitute, no shortcut, no alternative for effort.
That might be true but I don't think that's the problem here. I have seen plenty of AI generated images that evoke awe, nostalgia, other emotions - engaging me for a short time, and a few carefully prompted ones that work for a little longer. This fails to relative even these.
Another thing to keep in mind is that if you asked twelve different artists to render these cities from their descriptions, you would likely end up with twelve very different renditions.
Now imagine the same thing but each person is using Opus 5.5 - I’d wager there would be a lot of commonality among each of the outputs.
Depends, really. Ten different artists using Opus 5.5 and with some experience with LLMs, outputs may differ quite widely due to:
- Each phrasing the prompt(s) differently
- Each having different history accumulated in context
If you asked me to issue the same prompt as you to write some web artifact, chances are mine would be substantially stylistically different, simply because at some point in the past half-year, Clause started styling everything in LCARS form for some reason.
Whether the differences would be substantial or surface-only? Probably depends, but chances for substance increase with iteration. Real artists don't zero-shot their masterpieces either.
Sure, but I'm referring more to this specific post e.g. "They gave Opus 5.5 one prompt, so it was more like a one-shot".
I’m saying if you provide the same prompt to Opus 5.5 across 12 different people, you’re going to get some similarities that immediately arise.
I’ve mentioned this across many HN threads - I call it the LLM corollary to “garbage in, garbage out”: "generic in, generic out". If you prompt with more specificity, you’re obviously going to get different and potentially better results. But at that point, the artist is doing more of the work.
A while back, I tried a similar test asking for LLMs to come up with engaging D&D puzzles that would fit organically into an underground labyrinth environment. This is a pretty general prompt. Let's just say that the results were about as inspired as a tepid bowl of tapioca.
Every decision a human makes has some level of intent. AI cannot intend, it can only suppose.
The achievement and commitment of others inspires us to reach further and further. Now that's being eroded I wonder what the future will hold.
Yeah, I haven't read the book but these models don't evoke any particular feeling in my glance at them. And I have seen AI-generated stuff that's decent at creating a mood. I suspect that what AI can't do currently is produce a finely tuned mood, it just has the equivalent of "presets". I mean, Hollywood is hardly better than that today as well.
Arts and Crafts Movement 2.0. Bring back handcrafting starting with woodworking, pottery and oil painting. And only objects that cannot currently be made by robots, e.g. 3d printing or print illustrations.
The first city I clicked on at random had the opening phrase:
> Arriving, you rejoice at its bridges, each different
The model city had like five bridges total, three of which were didn't bridge anything at all but instead sat in the middle of the river and went along its flow, the "bridges" for some reason pretending to be small islands.
like bridge over slop and water, i will let you down.
Astra Penthesilea has towers that sometimes clip into the green... hedges? Pipes?
Opus Eutropia has a few isekai circle cities with one of them... ON the river, giving no space for ship traffic.
Slop has never been this beautiful before!
Not to mention that the slop site claims that the cities are Invisible, when I can clearly see them.
You made the mistake of looking too closely
Like many other people in this thread, I feel like passing what would essentially be a hobby project off to Claude robs the creator of most the fulfillment and gives the viewer less reason to appreciate it. But with the models being as good as they are today, I genuinely feel conflicted. If your end goal is to share something with others, why spend a weekend on a project like this when you could just hand it off to Claude?
Your appreciation should come from the one shot prompt. It's a direct comparison to Astra's one shot. I feel like you and many others are trying to "appreciate" the wrong thing.
If the creator truly wanted viewers appreciate the end product they'd start with the one shot output and spend their weekend iterating on each city either by hand or with Claude. But they didn't want to do that, they wanted to show off what Claude/GPT could do in 6hrs. That's the whole blog post.
Well, I don't know - the author writes "And I’m mesmerized!" towards the end.
I suppose I was expecting... something mesmerizing.
This isn't a ceiling, though. This is the new floor: What you can get with more or less zero human effort and ~$100.
I expect we're going to see more and more amazing "hobby projects" that did involve both a few hundred dollars in tokens and hours of human creativity and effort in combination.
But more human effort and say a $.50 pencil would produce something more impressive for sure if the human was reasonably skilled.
I get your point, and I think it's broadly true. But I'm not building things with AI to impress others, and it could be that GP or OP are doing the same.
Unasked-for-story-time: In the early aughts, there was a PC video game called Star Trek: Elite Force. It was a FPS genre game and included a feature called "Virtual Voyager" that at the time was revolutionary because it would let you just freely explore a subset of the Starship Voyager without any linear mission requirements, the closest any Trek fans could get to gaining access to the real Paramount Studios Stage 9-and-friends sets where the TV show was filmed, which for Trekkies with no life is quite the big deal.
Wasting the GPT-6 Astra capabilities over a single month of the $200 plan, I've built out multiple decks of an Excelsior-class starship based on a deck plan of my design, dealing with creative decisions on how to adapt the limited Star Trek canon knowledge about the interior deck sizes and materials aboard an Excelsior-class vessel, and at the end of this project, all I will have will be a graphically crappy but 3d explorable virtual starship I can walk around, look at the interior of, and mentally place my daydreamed Star Trek personal fanfic that I used to pencil in a notebook, and now I direct/dictate/produce using agents to write their crappy sloppy prose, but is tuned exactly to my preferences, and now will be able to have specific camera angles and relative details because I can screenshot the 3d ship explorer I made with the crappy AI.
I've got the pencil, I've got the nerd no social life time for creative projects, and I am still writing the no-AI no-name fantasy novel that nobody will read except my family and AI-data-gorgers, but neither of the two projects above, which have provided me hours of entertainment dopamine, would have happened without AI tooling making it easy enough to be creative without having to hand render every little detail.
Art is produced for arts own sake, right? That's the phrase? Well this vibe slop story and this vibe slop 3d viewer are digital goods / digital art to _me_, because I like them for their own sake. I can't copyright any of it, its all been infected by AI stuff. I can't get a job with it, I can barely orbit in Blender and don't find it worth the time to learn to render drawing in 2d and sculpting in 3d (I've tried both over the years in pre and post AI eras, I'm from a creative family whether I wish it or not), the costs to render skills manually make it not worth my time to do.
But AI makes it worth my time because it is an enzyme that reduces the input cost of me prototyping my creative desire out without having to laboriously create every last everything.
That makes sense. The difference would be if you were to put it out there as proof of your ability rather than as proof of your imagination. Now I'd like to see what you made :)
Why would that someone want to listen something you haven't created? What have you "shared"? Claude could have made the tools that let the creator build their cities from their own imagination, but that's not what this is.
Interactive demo: https://p.migdal.pl/invisible-cities-opus-5.5/
Source code: https://github.com/stared/invisible-cities-opus-5.5
I can't imagine a worse thing happening to a better book.
One of my favorite things about the National Treasure Cinematic Universe (two movies and a TV show) is the existence of characters that I like to call "treasure grumps". Their role in the story is to warn the main characters that searching for treasure is a terrible way to spend your time, and can only lead to personal and familial ruin.
Meanwhile, here in the non-fictional world, an estimated 20-30 million dollars have been spent excavating Oak Island with absolutely zero treasure of any sort found.
Right, the treasure grumps often have a point!
To Someone who is passionate enough to write about it, do something to build about it, share about it?
unlike someone who is just whining about it?
Wtf are you to someone who has put effort to spread about the book? I never seen/heard of the book, and now I am iintersted in it thanks tot he author.
I don't think you know what the word effort means.
Don't be a smart ass be clear with your explanation.
The Opus demo was a tad garish and overwrought. The Astra one was supposed to feel lighter but it was also over-encumbered by UI. Both were hard to navigate, despite the UI trying to seem navigable.
I don't mind the AI but the human needed to put more time into curation and getting the magic right.
Understanding how the models work, is it reasonable to say the generated visualisations are copy or, at least, derived from someone else's work?
Yes, Calvino's.
Yeah, I was wondering why the two results were so similar, despite lack of detail in the prompt. Both have an array of "miniature" cities on a game-like map, all of them drawn surrounded by a circular border. Something sketchy is happening here. Either more prompt or intervention is going into it than the author claims, or it's copying some preexisting portrayal.
Does this bring any happiness? for me, No. effort need a purpose.
$74 vs $ 10 vs $25 for the same prompt, The interesting number is not actually quality, its what we can get per dollar like 6 subagents runs in parallel.
Yeah for now, this week, on that model, all of which are constantly changing. I would love to know actual cost of running those inferences on such hardware, including electricity, training, etc. i.e. what we should be paying. I can imagine that job would cost multiple hundreds if they were charging enough to actually be profitable companies.
TBH only scanned through ~20, impressive technically but disservice to Invisible Cities to render it as... uninspired, bland, miniatures. All the wrong of details to focus on to capture even notional urban verisimilitude. Looks like a 8 year old drew scribbles and their dad finished the project.
I was finishing Invisible Cities yesterday, because the book was due at my local library today. I was wondering whether the cities could be adapted into some visual art form. And obviously I thought about feeding it to Opus. 24 hours later, I see this post…
That is what happens in the city of Ternopoli.
Citizens there gather their tools at the bottom of a giant staircase and climb up countless stairs, only to find out that what they were planning to do has already been done by the time they reach the top.
Dissapointed, they descend to grab a new set of tools and make new plans for tomorrow, only for the same thing to happen again.
I copied the prompt from the article and inserted it into Claude, here is what it created:
https://claude.ai/share/6bbd421a-147d-435d-bc00-7e93b393b943
(the standalone version at the very end of the chat is the one that runs in the browser)
Well, overall, by faaaaar not that beautiful as what the author from the article received? :-D
(though, it was done with Opus 4.8)
The whole point of the OP's article is: Opus 5.5 is so much better at one-shot design! Also, your link is not public - I'm curious to see your Opus 4.8 comparison
Access granted for all people who requested.
Wondering: I clicked share & copy link - this should be correct?
Very cool, thanks for posting!
And shame on all the people shitting on it because AI made it. If you valued art for art's sake, you wouldn't be so bothered when somebody makes something art-adjacent using some technique you don't approve of. It in no way detracts from the kind of art you admire. Don't be type who just has to let others know that you find something they like to be beneath you.
It's amazing to me how miserable some people can be. Incredibly rude and dismissive to come in a comment section and shit on something creative.
I put about as much creative thought into this comment as they did their prompt
This is probably one amazing positive these models have brought forward - the ability to break out of a single modality, and utilise others to help us learn and visualise. While the visuals here are amazing, my son for example prefers to learn by listening and talking, so we convert lot of his study materials into audio and real-time voice roleplay.
Wonder how would Fable 5.1 perform. I guess to expensive for such experiment. I guess you had Max 20x. Still it taking 3 hours is significant time for agent. What thinking level you had it on?
It's pretty and cool, but it's astounding how fast I'm becoming desensitised to "I built <COOL NEW THING> with AI".
First off - I'm a senior dev with ~30 years of experience in graphics engines, game dev, web dev, WebGL stuff, blah blah. I am generally skeptical of AI, but... after reading the top of every thread in this comment section, I feel like people forgot that they could zoom in on the cities to see all the details.
The amount of _stuff_ crammed into each little city is almost unbelievable. Drawbridges open and close, glowing lines sketch patterns, camels walk across a tiny desert. Buildings fade into existence and then burst in a shower of sparks. Flags fly, smoke drifts, buckets cause ripples as they dip into underground pools. The city made of plumbing has tiny people in the bathtubs! There's a freakin' roller coaster with moving cars and a ferris wheel!
I absolutely understand the AI hate that's being commented here, but... in this instance, I just can't feel it. I keep finding more and more little details every time I click on a city, and I am astounded. I just now noticed there's a little guy in the city of strings hanging up new strings at the top of the hill. Damn. My favorite is Marozia, where dark buildings open like treasure chests and release shining towers and flocks of glowing birds.
Yes, there are also a lot of stupid AI errors - every spoked wheel has the spokes off-center for example. But holy crap, I just can't hate this. I don't care about who built it or the random bits of jank - this thing is _beautiful_ in a way that I haven't seen from AI before, and it warms my old GPU-powered heart.
Even as I was writing this I kept finding more things I didn't notice on my first few passes through the cities:
Phyllis fades from color to black and white except for two houses. The fairground has a big top with tiny bleachers and a guy inside. The happy city has people with their arms raised in a V. People in the marketplace city wear hats and some are carrying bundles on their heads. There are weird cactus plants in the hanging city. In the half-fairgroud city, there's a statue on one of the trucks. The city that copies the dead has a set of pallbearers carrying a wrapped corpse down the stairs. The people on top of Argia are laying with their ears to the ground, listening for the animated sound waves coming from below.
And on and on and on and on. I just can't even.
There's a tiny cat on one of the roofs in Raissa! The rats in Marozia hop along their paths! The center of Theodora has bookshelves with individual books!
But who cares about those details if they're just random? Why are they there? What do they mean? What was the intention? Sure it's jaded, but just cramming in elements doesn't necessarily impress, other than technically it's better at not overlapping random stuff.
both the linked camera lens example as well as the last one the author linked runs at ~5fps on vivaldi and pegs CPU 100%. i have never seen any 3d visualization website have performance this bad, i assume they're not using webgl given my GPU is 0% utilization.
On the down side, $84 spent on a slot machine might have actually yielded some benefits.
Creativity is what humans are great at—we don't need Opus for this
It reminds me of Total War. The soundtrack and way that the map is shown. Is Total War an inspiration? Anyways, I liked the visualization. But I never read the book, so I followed (partly) droidjj recommendation
While reading the book years ago, I did not visualize Isidora as a city with 9 houses, like a children’s book, or a little video game level intro, but so unlike Kublai Khan or Italo Calvino.
wow that's awesome! super cool :), fyi so inspiring that I did something similar for d&d forgotten realms, including a timeline: https://narfman0.github.io/realms-atlas/ fable drove it, opus agents did the work, took 30-45 minutes. (intend on extending to planescape, elaborating on significant historical events, and some text to speech narrating cool events)
rock on
The demo does not work on Firefox or Safari on the iPhone SE, for what it's worth. I can't scroll down to close the splash screen.
Funny note about Opus 5.5.
I once instructed it to spawn at most 20 subagents and it spawned 40, with an adversarial reviewer for each of the 20.
It doesn't follow these instructions very well.
20 sub agents, 20 doms.
Such a drag- a few carelessly lumped together Ai demos. I mean the headline says it all, there is not much more to it. ok, no thanks!
Cecilia looks like an average residential area in Russia.
Yeah. This fails to load both in Opera and in Chrome on Android. I'd tone the hyperbole down a bit.
I just have zero interest in looking at any of these "AI did a thing for me." posts.
It's one of my favorite books - and I think it deserves better; it's a conceptually very dense text and just slapping together a few 3D models misses the point
it's also not very astonishing that the current models are able to create this - it looks exactly as expected
and I actually think getting an interesting AI generated rendition of the text from a single prompt is a worthwhile endeavor
maybe even a good replacement for the pelicans: it would test if a model can engage with a text beyond the surface level
but this ain't it
The creation really is amazing. I harnet heard of the book but got a long way through it!
Loads nothing on iphone lockdown mode fwiw.
Isn’t it expected?
No, why would it be? Unless very detailed web-based microphone or camera manipulation is needed, a modern LLM can built a fallback using safer technology instead.
It’s a 3D demo and the user has disabled webgl and webgpu, for good reasons. Not much you can do.
Doesn't work right on my phone.
Vanadium (chrome)
https://imgur.com/a/nGLyJf6
Latent space is amazing isn’t it
I'm a little torn by this sort of demonstration, because while many of the scenes have obvious markers of slop (impossible intersecting geometry, bridges to nowhere, and so on), the scenes largely do work to convey the intended concept/emotion, and the low-poly aesthetic is executed decently well.
I guess I shouldn't be surprised that another visual medium is starting to fall to LLMs after the success of diffusion models for image generation. Nevertheless it's wild that a mostly text-focussed model is building little worlds like these.
this is amazing, wonder how much better it's going to get in a year from now
What is this for? It doesn't deepen the understanding of the book. The illustrations aren't attractive on their own. They're not a very good representation of the descriptions in the book.
Is creating not enough? Somebody has done something artistic and all you hear is negativity.
> Somebody has done something artistic
Nobody did it, and it’s bad.
I ran cat /dev/urandom for six hours but I didn't need to write a blog post about it
Just curious - was the result a 3D animation? If so, I'd love to read your blog post!
If you're going to spend tokens on 3D stuff, then please build me a better FreeCAD first.
https://github.com/dzervas/cadara
Thanks, I hate it. I imagine Calvino turning in his grave at such an abdication of human imagination.
"Attention whore" used to mean someone who makes garbage to get attention. Now it's a machine that uses attention to make garbage.
Well it still has the trademark "colored left-side of the boxes" ai slop, but otherwise looks (and sounds!) very impressive.
Well it still has the trademark "colored left-side of the boxes" ai slop, but otherwise looks very impressive
Apologies for the long post in advance (maybe I should turn this into a blog post).
Invisible Cities is one of my favorite books, and I've re-read it several times. It's a great "travel" read and always puts me in the mood to explore. It's also a very romantic book, obviously the cities have all named after women, so it evokes conquest in the sexual sense as well. It's also a piece of surrealist/allegorical literature in the same vein as Daumal's Mount Analogue (which also tops my list).
Back to the website. While this is an impressive tech demo (as they all are), it's clear that AI misses the forest for the trees, and here's a tangible example. I randomly clicked on Theodora. Honestly, I didn't even remember which city this was until I re-read the Calvino passage:
> Recurrent invasions racked the city of Theodora in the centuries of its history; no sooner was one enemy routed than another gained strength and threatened the survival of the inhabitants. When the sky was cleared of condors, they had to face the propagation of serpents; the spiders' extermination allowed the flies to multiply into a black swarm; the victory over the termites left the city at the mercy of the woodworms. One by one the species incompatible to the city had to succumb and were extinguished. By dint of ripping away scales and carapaces, tearing off elytra and feathers, the people gave Theodora the exclusive image of human city that still distinguishes it.
So, what do we know about this city? We can make a short sensible list:
This[1] is what AI saw. Now, this isn't bad, really. But it's so embarassingly pedestrian, it hardly counts as a visualisation. It just heard "city with animals and stuff" and made a 3D model. But it's all, in that uncanny-valley kind of feeling, wrong. Why is the city itself a grid? Why is there a library in the center?
Calvino, again, on Theodora:
> But first, for many long years, it was uncertain whether or not the final victory would not go to the last species left to fight man's posession of the city: the rats.
The vibe here is an overrun rat's nest of a city. A disgusting bloated corpse; more:
> The city, great cemetery of the animal kingdom, was closed, aseptic, over the final buried corpses with their last fleas and their last germs.
I see an overrun New York. I see a city built on a decaying graveyard. AI sees animals and humans living in a kumbaya harmony, but that is not what Theodora is: it's an eternal conflict between man and nature (and nature won a long time ago). For the record, this isn't even that particularly deep. Many authors and poets write about this eternal human struggle: the taming of the wilds. But Claude just doesn't get it. I hope readers of the book do.
Anyway, I leave you with this. I found an artist, Rian Hotton[2], who put brush to canvas to show us what he thought Theodora looks like[3]. And this... this is more like it.
[1] https://imgur.com/a/hXrKVg3
[2] https://www.facebook.com/rianhottonartist
[3] https://imgur.com/a/tR0qTYx
What a beautiful website, wow. I've been online 10,000s of hours, have seen pretty much everything and this one goes straight in my Top 10.
I'm so glad I'm able to witness this Cambrian explosion of software :).
You should avoid add model name (or at least use Open weight model names) in your title, so it would be 90 (amazing) : 10 (lost the book value) comments. /s