What AI is doing is unbelievable. Each of these examples would have cost hundreds of thousands of dollars and months of dev time to achieve, and they still wouldn't have been as polished as these are.
Before AI, in order to build something like this you had to be both creative, and you had to know how to do it. The latter part would have taken years to accumulate the experience required. Now, you just have to be creative. The days of techies gatekeeping these experiences are over.
Invisible Cities is my favorite piece of literature hands down. The audiobook narrated by Richard Higgins is also done so beautifully and can't recommend it more.
As neat as this project is, this does no justice to whatsoever to the book. On the surface Invisible Cities seems to be a book about random cities' descriptions. In reality though, this book is about semiotics, meaning, language, and it's limits.
If you haven’t read the book yet, I’m tempted to warn you against opening the visualization. Part of the joy of reading this book is letting your mind’s eye run wild, and I fear if I read it for the first time after toying with this site, I’d be remembering the cities rather than conjuring them.
That is the exact problem with the city of Dinipro.
Citizens of this city high in the sky imagine it so fully in their minds that no one there needs to open their eyes to walk the streets and live their normal lives.
But if one makes the mistake of opening their eyes and tries to see the city directly, they would immediately fall through the clouds to the ground deep below.
I heartily recommend the audiobook, narrated by John Lee.
The book is a trip and the dreamy sort of narration matched it perfectly. My wife and I would spend evenings listening to one city at a time, on a walk around our little park, one headphone in each ear. I don't know if it was just the environment or the time in our lives, but we both really look back fondly on this book.
A couple of years ago, after reading the book, I sketched some of them on Procreate with an isometric grid, but not all 55. I only arrived at number 4, since each took multiple hour...
This is probably one amazing positive these models have brought forward - the ability to break out of a single modality, and utilise others to help us learn and visualise. While the visuals here are amazing, my son for example prefers to learn by listening and talking, so we convert lot of his study materials into audio and real-time voice roleplay.
The first city I clicked on at random had the opening phrase:
> Arriving, you rejoice at its bridges, each different
The model city had like five bridges total, three of which were didn't bridge anything at all but instead sat in the middle of the river and went along its flow, the "bridges" for some reason pretending to be small islands.
both the linked camera lens example as well as the last one the author linked runs at ~5fps on vivaldi and pegs CPU 100%. i have never seen any 3d visualization website have performance this bad, i assume they're not using webgl given my GPU is 0% utilization.
I'm a little torn by this sort of demonstration, because while many of the scenes have obvious markers of slop (impossible intersecting geometry, bridges to nowhere, and so on), the scenes largely do work to convey the intended concept/emotion, and the low-poly aesthetic is executed decently well.
I guess I shouldn't be surprised that another visual medium is starting to fall to LLMs after the success of diffusion models for image generation. Nevertheless it's wild that a mostly text-focussed model is building little worlds like these.
It reminds me of Total War. The soundtrack and way that the map is shown. Is Total War an inspiration? Anyways, I liked the visualization. But I never read the book, so I followed (partly) droidjj recommendation
Six hours of continuous runtime is the underrated part here. Most one-shot demos fall apart well before that, so this doubles as an endurance benchmark.
The supervisor agent will keep the session in a loop until 6 hours have passed and eventually the agent will decide to use up the remaining time rather than fighting with it
I wonder how much co2 is being thrown into the atmosphere everyday through the steady stream of “look what I made this LLM do” and endless “benchmarking”?
One useful hint here is the API token cost - in this case "about $74 in API tokens" for Opus 5.5.
We know Anthropic run inference at a margin, so that $74 means that the cost of the electricity involved is substantially less than $74. I'd love to know the actual cost there.
Intuitively I'd say that just-for-fun AI projects emit orders of magnitude less CO2 than what just-for-fun road trips or air travel emit.
So unless you can show that it approaches all the other frivolous consumption that humans love to waste resources on, maybe we can get back to experimenting with cool technology and talking about it, on this website called "Hacker News"?
I wonder how much co2 you use to get through the day, and how much you use to remark on other people's creations in a way that suggests you disapprove of their utilization of their available resources. How much co2 do your projects emit, since you seem to be interested in those metrics?
I strongly suspect one HN post (or two now) is not even in the same ballpark. I could ask Claude to figure it out but that would be a complete waste of time and energy. ;)
I was finishing Invisible Cities yesterday, because the book was due at my local library today. I was wondering whether the cities could be adapted into some visual art form. And obviously I thought about feeding it to Opus. 24 hours later, I see this post…
Citizens there gather their tools at the bottom of a giant staircase and climb up countless stairs, only to find out that what they were planning to do has already been done by the time they reach the top.
Dissapointed, they descend to grab a new set of tools and make new plans for tomorrow, only for the same thing to happen again.
$74 vs $ 10 vs $25 for the same prompt,
The interesting number is not actually quality, its what we can get per dollar like 6 subagents runs in parallel.
Before AI, in order to build something like this you had to be both creative, and you had to know how to do it. The latter part would have taken years to accumulate the experience required. Now, you just have to be creative. The days of techies gatekeeping these experiences are over.
As neat as this project is, this does no justice to whatsoever to the book. On the surface Invisible Cities seems to be a book about random cities' descriptions. In reality though, this book is about semiotics, meaning, language, and it's limits.
Citizens of this city high in the sky imagine it so fully in their minds that no one there needs to open their eyes to walk the streets and live their normal lives. But if one makes the mistake of opening their eyes and tries to see the city directly, they would immediately fall through the clouds to the ground deep below.
The book is a trip and the dreamy sort of narration matched it perfectly. My wife and I would spend evenings listening to one city at a time, on a walk around our little park, one headphone in each ear. I don't know if it was just the environment or the time in our lives, but we both really look back fondly on this book.
The website gives you a good hint before opening up the visuals. So, I got an idea.
I don't really know about Marco Polo. My only knowledge comes from the Netflix series, which doesn't talk about these invisible cities.
Invisible Cities is a fantastical postmodern novella.
https://en.wikipedia.org/wiki/Invisible_Cities (Spoiler warning, of course.)
https://camillovisini.com/drawing/fc4jc6-le-citta-invisibili...
https://camillovisini.com/drawing/p2yrdj-le-citta-invisibili...
> Arriving, you rejoice at its bridges, each different
The model city had like five bridges total, three of which were didn't bridge anything at all but instead sat in the middle of the river and went along its flow, the "bridges" for some reason pretending to be small islands.
Opus Eutropia has a few isekai circle cities with one of them... ON the river, giving no space for ship traffic.
Slop has never been this beautiful before!
I once instructed it to spawn at most 20 subagents and it spawned 40, with an adversarial reviewer for each of the 20.
It doesn't follow these instructions very well.
Vanadium (chrome)
https://imgur.com/a/nGLyJf6
I guess I shouldn't be surprised that another visual medium is starting to fall to LLMs after the success of diffusion models for image generation. Nevertheless it's wild that a mostly text-focussed model is building little worlds like these.
The supervisor agent will keep the session in a loop until 6 hours have passed and eventually the agent will decide to use up the remaining time rather than fighting with it
Sigh
We know Anthropic run inference at a margin, so that $74 means that the cost of the electricity involved is substantially less than $74. I'd love to know the actual cost there.
So unless you can show that it approaches all the other frivolous consumption that humans love to waste resources on, maybe we can get back to experimenting with cool technology and talking about it, on this website called "Hacker News"?
What a terrible thing to say to human! Being (presumably) human, it is their inalienable, natural-born right to do so.
Shame on you.
Source code: https://github.com/stared/invisible-cities-opus-5.5
Citizens there gather their tools at the bottom of a giant staircase and climb up countless stairs, only to find out that what they were planning to do has already been done by the time they reach the top.
Dissapointed, they descend to grab a new set of tools and make new plans for tomorrow, only for the same thing to happen again.