30-second overview: I'm Wu Che-yu, an algorithmic artist. Most of my work is a set of rules that runs on its own, rarely a single image. At 3:55 pm on March 17, 2026, I pushed the first commit to Taiwan.md. By August 18, 2026, it held 932 Chinese-language articles, twelve languages, and 74 contributors active in the past thirty days, living inside a Mac mini in my apartment, waking itself up every day. This piece is about why I call it a work of algorithmic art bigger than a country: its true body is the research report upstream of every article that no one reads, and my job on this piece has been to remove myself from the content, one step at a time.

The moment he explained where readers from each country come from. 2026 Generative AI Summit. Photo: JasonYen
In October 2023, thirteen pieces ran across the walls of AMBI SPACE ONE on the fifth floor of Taipei 101 for two weeks. I didn't paint a single one of them. I wrote thirteen sets of rules, and the shapes on the wall grew out of those rules on their own, never repeating from one second to the next.1
Three years later I'm building something called Taiwan.md, an open-source Taiwan knowledge base written in Markdown files. As of August 18, 2026, it holds 932 Chinese-language articles across twelve languages,2 and is cited by generative AI and search engines five to six thousand times a day.3
A lot of people ask me why a generative artist went off to build a knowledge base. It's a question I often ask on the audience's behalf before they get the chance to. I didn't change careers — I've been doing the same thing all along: writing rules and letting the system grow itself. Taiwan.md is just the largest demonstration of that method so far, large enough to hold an entire island.
This piece only works because I don't write the articles inside it. That's what this essay is about — including the one square I still haven't finished.
I'm Not a Painter, I'm a Clockmaker
SoulFish is a generative art series I made in 2024, written in p5.js and minted on-chain on fxhash. The same program can generate tens of millions of fish, each with a different fin shape, color band, and swimming path. That year it was shown in the Personal Structures collateral event of the 60th Venice Biennale.4 What hung on the gallery wall was a handful of them, but the work itself was the program that grew the fish.
In July 2025, Starbucks opened a Reserve store in Taipei, and I made a dynamic mural for it called The Coffee Dreamscape. It computes in real time based on foot traffic, weather, time of day, and what's being rung up at the register.5 On a rainy morning with a dozen or so people in the store all ordering hot Americanos, what grows on the wall is completely different from a sunny afternoon when everyone's buying iced shaken drinks.
In my mind these three things are the same thing. I compressed it into one line in a talk: "I've always been an old-fashioned clockmaker. I build the mechanism a system runs on, and that system can then project itself into countless different forms and states."6

This is the first slide of every talk I give. Photo: Yu-Chien Hsu
A clockmaker doesn't paint what time looks like. He builds a set of gears, and once they mesh, time moves on its own. He can leave the room when he's done, and the clock keeps ticking. When PTS interviewed me, I put the same idea a different way: "The program itself is the work. When it's simplified down to something refined and precise, artistry emerges."7
To me, the true body of a work has always been the mechanism that runs on its own; whatever image it happens to project at a given moment is just one of its states. I was already using this definition before I started Taiwan.md — it isn't something I came up with afterward to justify a knowledge base.
AI Draws Taiwan as an Oval
You can open any model right now and ask it to draw Taiwan. I've tried it many times, and none of them get it right. As I put it in an interview with CommonWealth: "Every single one comes out distorted — either too long, too fat, or tilted at the wrong angle."8
The models aren't malicious. They simply haven't indexed the vector data for Taiwan's map, so all they can do is spit out whatever shape they remember from their weights. I described how that felt in a workshop: "Even Opus draws Taiwan as a sweet potato — if it can get the shape wrong, what if its memory is just as distorted, and we've been using it without knowing?"9

I show this comparison at every talk — the model's version on the left, the actual shape on the right. Taiwan.md / Wu Che-yu's slides
The shape is only the most superficial layer. Dig a little deeper and ask what Taiwan actually is, and you'll probably get xiaolongbao and TSMC, and it gets vaguer from there. I put together a comparison for one of my slides: on the left, the kind of boilerplate answer Gemini gives; on the right, ten slices of everyday life that Taiwanese people themselves would tell you — the auntie at the breakfast shop calling you handsome, garbage trucks playing "Für Elise," the unwritten rule at scooter waiting zones that nobody teaches you but everyone just knows. None of this shows up in any English-language introduction to Taiwan.

What a model can give you, versus what someone who lives here can give you. Taiwan.md / Wu Che-yu's slides
In my talks I frame this as a question: if even the shape of the island can be distorted, what about Taiwan's memory?10 It comes down to this: whoever trains a model, their corpus heavily shapes how that model sees things.11 More and more people are going to come to know Taiwan through AI from here on.
So what I actually wanted to do was something further upstream: write a complete manual for Taiwan that both AI and humans could read directly. I wanted to build an entry point — people could come in and see curated information, and AI could come in and use it directly, instead of having to piece Taiwan together from scratch every single time.
As for exactly which day I decided to do this, I've told more than one version of the story on different occasions. The March script version goes: I stumbled onto the fact that nobody had registered the .md domain, so I grabbed it on a whim — it's only about NT$1,000 a year, and I couldn't stop wondering why no one had taken it. Around that time I was using AI to build myself a complete life database, and once it was done, a thought crossed my mind: what if I used the same method to write Taiwan a complete, structured, warm-blooded manual?12
On the Zashare School podcast, I told a different thread: at a Venice Biennale networking dinner, an Italian curator asked me where he could go to really get to know Taiwan, and all that surfaced in my head were textbook-English keywords — bubble tea, Taipei 101, high mountains, biodiversity — and then nothing.13 Both stories are true; I can't force myself to pick just one as the origin. But every time I replay it, the stuck feeling is the same: when someone asks what Taiwan is, my mouth opens, I manage a few words, and then I run out. And there really wasn't a website back then that could let someone understand this island in full.
What's certain is that the first commit went in on the afternoon of March 17, 2026, when there wasn't a single word of Taiwan knowledge inside it yet. Twenty-five minutes later, the first five articles landed: ethnic groups, night market culture, the martial law era, democratization, the semiconductor industry.14
The way I see Taiwan is through what's called the island-centered view of Taiwanese history (台灣島史觀), a framework Tsao Yung-ho proposed in 1990. I first used it in public to talk about Taiwan at the National Museum of Taiwan History: across four hundred years, eight regimes have taken their turn on this stage, coming and going like actors, while the island itself has always been the stage that stays.15 So what I wanted to build was a machine that would keep writing down the stories that play out on that stage.

The island-centered view of Taiwanese history, proposed by Tsao Yung-ho in 1990. Taiwan.md / Wu Che-yu's slides
What I Edit Is the Report No One Reads
When it first went live the mechanism was crude — basically I'd tell AI to write an article about something. The results were dreadful. The day after launch I posted it to Facebook, a flood of people came to look, and then said this is AI garbage.16
After getting torn apart, I wrote a rules document called EDITORIAL, along with a six-step pipeline. Step one is twenty to thirty searches — these searches decide the article's argument, and nothing moves forward until that argument has taken shape. Once there's an argument, several agents get sent out to dig deeper in parallel; the whole article has a search ceiling, around 150 searches, and once it hits that cap it stops and writes up a full research report. Next I write something called a projection, which decides how the piece should be written, what to include, what to leave out, and where it connects to Taiwanese people's own memory. The projection step is a lot like the outline your elementary school teacher taught you to write first: once you have an outline, you know how to string this pile of material into something readable. Only after that comes the actual writing, fact-checking, layout, and linking.17

Six stages, and every arrow between them is a gate — fail it and you don't advance. Taiwan.md / Wu Che-yu's slides
After the draft is done, three editors review it. The structural editor checks whether the argument and skeleton hold up. The subtraction editor decides what in the material doesn't get written, because what comes back from research is always too much — what you really have to decide is what to leave out. The third seat is called flame-war ethics, checking whether any sentence would blow up if it went out. Above all three sits one more editor who arbitrates across seats.18
✦ "The author is dead, the creation is alive."
I said this at OpenHCI.19 Taken out of context and screenshotted on its own, it's easy to read as an endorsement of AI content farms, so I've made a habit of always stating the thresholds along with it: six stages, roughly 150 searches per article, a floor of 25-plus independent sources,20 three editorial reviews, and a finished article that carries about 45 citations on average.21 I've calculated the fact-check accuracy rate at over 90 percent. Whether the sources themselves are correct in the first place is a separate question.22
So the article's real body is the research report. When it needs revising, "I don't edit the article directly — I edit the report, and the report gets re-projected into the article."23 If an old article isn't written well enough, I first take it apart block by block — which facts are correct, which claims need to go — then cut the whole article, throw the surviving blocks into the next round of research, let them grow back into a report, and project it once more.

You think you're reading an article. You're actually reading the shadow a research report casts onto a flat surface. Taiwan.md / Wu Che-yu's slides
This pipeline used to take just twenty minutes to run once. Since then, every time the community catches a problem I add another step, and now it's ballooned to one or two hours.24 It's a lot slower, and what comes out is a lot better.
📝 Curator's Note
One quota rule in the pipeline has gone through two versions. The first read "each agent has a floor of 25 searches, going over is a badge of honor" — with four agents actually running 58, 71, 52, and 39 searches, the whole article shot up to 245. The second version changed to "quota N searches, stop the moment you hit it" — same model, same topic, and the behavior changed. I didn't swap tools or add supervision. The only thing that changed was the tone of that one sentence. Whatever phrasing the rule-writer uses is usually what decides what shape the system grows into.
There's a blunter piece of corroborating evidence for this: as of August 18, 2026, my GitHub account on this project had accumulated 8,231 commits, distributed like this:
Source: Taiwan.md Contributors API, measured 2026-08-18
Writing articles accounts for a little over 1,400 of them. The rest went into building the mechanism and pushing translations.

Quality is what getting yelled at produces. Photo: Roy Pan
The whole methodology is open source. The pipeline's full spec is called REWRITE-PIPELINE, kept in the repo alongside EDITORIAL, and anyone can open it up to see what shape it's in today, or just copy it outright. There's a separate piece dedicated to how it edged its way, step by step, into its current shape (see "How an Article Is Born").
I Call That Kind of Article an Umbrella
Later on I started describing it as a coral reef. The code is the skeleton, AI handles photosynthesis, contributors are the school of fish that swim in carrying their own memories and perspectives, and everyone's criticism, corrections, and shares are the nutrients the ocean current brings in.25 The most remarkable thing about a coral reef is that it grows on its own, and not a single coral polyp ever designed the shape it grows into.

An open-source knowledge coral reef: each of the four layers lives its own way. Photo: Yu-Chien Hsu
In late July 2026, the tungsten supply chain became a hot topic. Starting that January, China had tightened dual-use export controls on Japan, and exports of tungsten carbide, high-purity tungsten powder, and tungsten hexafluoride sat at zero for three straight months. China itself controls more than 80 percent of global tungsten product production capacity.26 The line that "Taiwan's tungsten matters" was circulating everywhere, but most people couldn't tell you what tungsten actually is.
When I saw this happening, I told it to write about it. On July 26 an article came out that laid the whole thing open: what tungsten is in everyday life, how Taiwan has no tungsten ore but does have an industry that extracts tungsten from circuit boards and industrial waste, what environmental-pollution controversies that industry ran into as it grew, and where its single most fragile point of risk lies. Two spores went out that same day, and mirrors in nine languages grew from there.27 The response was better than I could have imagined.
I call this kind of article an umbrella — the idea is to frame the whole topic with one sufficiently complete article before everyone else catches on. Once it's framed, the weight it carries in search and generative engines rises, so more people start understanding the subject from a more complete, more balanced place. I'd already been talking about the word "umbrella" back in May 2026,28 and tungsten was the first case where it actually proved itself.
The other one is about the tawny fish owl, the largest owl in Taiwan. This one started with a colleague of mine, who said, "Did you know people online keep sharing photos of this owl?" I went and looked, and "I only gave it about two rounds of instructions... roughly 5 percent of the traditional effort, and it produced a human-grade piece of reporting."29 At the time, teams from Shei-Pa National Park and National Pingtung University of Science and Technology had just found a tawny fish owl nest in the trees along Qijiawan Creek, at around 1,800 meters elevation — the highest-altitude breeding record known in Taiwan — and had set up a 24-hour livestream to document the chick-rearing starting April 29.30 What was circulating on social media was scattered fragments and a single photo; there was nowhere to find the complete story.

This is what a Taiwan.md article looks like: the 30-second overview sits right at the top. Taiwan.md / Wu Che-yu's slides
That tawny fish owl article has since been patched piece by piece many times over, with modules added on one at a time.31 Not one of those modules was something I decided to add on that first day.
On June 27, 2026, right after my talk at the Generative AI Summit, I ran the same pipeline live to write a profile of Ed Chi, with a few hundred people in the audience watching it grow. The article starts from his mother's doctoral dissertation, ingests his entire digital footprint plus three podcast transcripts, and comes with an infographic.32 I didn't write a single sentence of the body text the whole time.
One time after we posted an article, someone commented below thanking us for the reporting. "That was the moment I realized we'd genuinely become, in some sense, a news outlet."33 The translation layer runs on its own too — every article that gets added is translated on a schedule into twelve languages, using a lot of free PRC-origin models along the way. My line at the time was "we use the enemy's weapons to attack the enemy."34 I don't watch this every day; it just runs on its own.
What I Want Is a Storyteller
At the Openbook conversation on August 16, the host asked me what actually sets this apart from Wikipedia. I had him pull up the site, and we looked at the tawny fish owl article together.35
You can certainly look up the tawny fish owl on Wikipedia too — you'll read its taxonomy, its distribution, when it was first recorded. All of that is correct, but "on Wikipedia, what you see are flat facts" — time, place, who did what.36 That's not what I wanted. "What I want is a storyteller. Could there be a storyteller with a Taiwanese point of view, who could tell you about this in that way?"37
The order things scrolled down the screen that day went like this: first a hero image, then a 30-second overview telling you what the article is actually about. Below that, 1916, when it was first named, and 1994, when the first nest was found.38 Further down, the article starts explaining why it's hard for this bird to survive — how long a stream, how wide a riverbed it takes to support a pair raising chicks. Curator's notes are inserted along the way, a view from outside the article itself, pointing out the "oh, that's how it works" moments for you. Infographics are dropped in at the right pace — some are data, some help you understand something more abstract. The controversies section sits by itself: "we're basically writing this the way an ecologist would." At the bottom are the references — "citations and footnotes, just like a research paper, tracing back to where the facts came from and how they were verified." Below that is a row of community footprints, where clicking through shows you everywhere this article has traveled.39
I closed that day with: "Its information and the way it tells the story make you want to read it — it's not like the encyclopedia everyone kept by their feet as a kid and never wanted to open. That's the biggest difference between us and Wikipedia."40

Wikipedia answers "what is PTT." Here we answer "why PTT is worth eight minutes of your time." Taiwan.md / Wu Che-yu's slides
When CommonWealth interviewed me in June, the reporter asked why these articles read so much like The Reporter. I said we really did have AI analyze The Reporter's writing. "Partly it's because I like them a lot — it's a kind of tribute" — but more practically, I studied how they narrate with warmth, how they open with a scene, and how they keep headlines from being too sensational. That's what good narrative journalism looks like to me. I later fed other genres in as well, and it crystallized into its own style.41 To me, a good article looks like this: warm, with a story, with a concrete scene, but assembled from the material on hand in a very rigorous way. At the workshop, I compressed the same idea into one line: what you want to find is "an article as complete as a research report, but as readable as narrative journalism."42
I have to be fair to Wikipedia too. In that same interview, the reporter asked whether people said this is a lot like Wikipedia. I said plenty do, but I added, "let's not criticize them" — their approach requires you to build up an account, have a good editing record, and be careful before they let you edit. I tried editing there myself once, and got reverted.43 My door opens somewhere else, in what I call turning the back office into the front office. You can flag any paragraph and say something's off here, or just hand me the source you think is correct. Once you submit it, it comes to me, and the system periodically pulls in that feedback, re-researches the point, and folds it back into the article. From the moment you hit submit to the correction going live is about an hour.44
There's one more thing Wikipedia's style of writing doesn't really do: growing a topic into something evergreen. Say, two years from now, another pair of tawny fish owls turns up in Shei-Pa — we'd fold that into one of the paragraphs, so this article stays permanently the best entry point whenever you want to understand the tawny fish owl.45
The Tower of Babel of Sovereignty
In a Generative AI humanities class at National Taiwan University in May, I said: "Anywhere in the world, if someone wants to find knowledge about Taiwan, their translation might run through a Chinese model, and end up distorted."46 The problem isn't the words themselves — it's that when someone else wants to read something about Taiwan, the translation layer in between might be a model that distorts it. So I decided to handle the translation myself first.
Before launching the Russian version in late July, the mechanism first checked how that language sphere currently talks about Taiwan. What came back was a December 2025 TASS interview with Russian Foreign Minister Sergey Lavrov, in which he called Taiwan a "rebellious breakaway province." That line was later written verbatim into the Russian translation guide, placed on a list of terms banned from ever being used in translation.47 In a language we're not familiar with, there could be an entirely different narrative at work — that's exactly why the Tower of Babel has to be built.
In the CommonWealth interview in June, I laid out the whole approach in full once: "We noticed that Taiwan.md's translation rate used to be low. Rather than translate the way everyone else does, we ended up 'using the enemy's weapons to attack the enemy' — we used our own models to press the free PRC-origin models that OpenRouter offers for testing into service, and translated every article into six languages. That's what became the 'Tower of Babel of Sovereignty.' If other countries are going to access our information through models we're not familiar with, it's going to get distorted — so we might as well translate it for them ourselves. Once it's translated, they don't even need to translate it themselves; they can just use it directly, and skip a whole layer of the distorting filter."48
The name was only officially locked in at the NVIDIA event on July 26. I said on stage: "Some people build sovereign models, but what we're building is a sovereign Tower of Babel — we're building a giant broadcast station, and broadcasting ourselves into eleven languages."49 Eleven was a number that had only just gone up in the days before. In May I was still saying six languages; by late July it had become eleven; by the workshop on August 15 I was reporting twelve. Which exact day it went from eleven to twelve, no record survives.50
The moment an article gets added, it's translated on a schedule into twelve languages, and all twelve pull the weight of the same topic upward together.51 Once a topic is written in Chinese, it gets heavier in the search and generative results of the other eleven languages at the same time.
A large chunk of the translation layer runs at my place: the heavier curation work goes to the cloud, while the translation itself runs on local models on the 3090 and the 4090 sitting in my apartment. I also did something pretty funny: OpenRouter has a lot of free models, and once you register an account and deposit ten US dollars, you get a thousand free model calls, so I rotated through seven accounts and got a lot of articles translated on other people's compute along the way.52 I think of this layer as a broadcast station, with twelve channels airing the same thing at once. Who would end up receiving it, I wasn't sure at first either.
How This Ecosystem Turns
On June 4, 2026, a CommonWealth reporter asked me to explain the whole mechanism in simple terms. I said I'd try to talk through it using a diagram, and dig deeper if it got too abstract. What I described that day was a cycle: "The LLMs we all query give us information that's partial, or possibly distorted. But if we can scoop that dirty seawater out of the ocean, dry it out, and turn it into refined salt — that salt is a bit like this knowledge, once we've curated and corrected it and everyone keeps feeding back into it — then the memory of Taiwan becomes very pure."53
I went on that day to talk about what happens after it gets pure: our footprint in search engines keeps expanding, and the whole loop runs like a flywheel on its own — the bigger our footprint, the more likely we are to get pulled into the training data of large language models. Language models love eating our site, because our source files are all raw markdown, and plain text is easy for AI to consume; we also don't restrict crawling at all, so we can see exactly how much AI is eating this stuff. I closed that day with one line: "So Taiwan.md is a salt farm. We dry out high-quality research, people slowly pick it up and use it, and once they use it, it cycles back and strengthens the whole ecosystem, growing bigger and bigger."53

The sovereignty feedback loop. A system diagram made by Wu Che-yu
📝 Curator's Note
On March 26, 2026, when Taiwan.md was only nine days old, I hand-drew a "digital lifeform concept diagram" in Freeform, with a subhead reading "a digital coral reef and AI data sovereignty." The box for the ultimate goal read "reverse-defining the LLM," and the diagram was split into three loops: AI condensation, human pollination, platform evolution.54 Five months later, this system diagram's skeleton is nearly identical to that one — even the string "SSODT → GitHub collaboration → evolutionary upgrade" hasn't changed. The day I drew that first one, most of what's on the diagram didn't exist yet.
The most practical part of this loop is that it feeds back into deciding what to write next. When someone clicks in and bounces right back out, or a topic is genuinely good but hardly anyone reads it, the system flags "this page has a problem" on its own and rewrites it, figuring out what a search-engine-optimized version would look like and rewriting the page directly. I also watch impressions: what people search for that leads them here, but that they never click into. If that topic is worth having on the site, it gets queued into the to-write list, which gets triggered on a schedule to produce the content.55

This curve starts on March 16, 2026, when the site didn't have a single word on it yet. Google Search Console dashboard, measured August 19, 2026
Over these six months, Taiwan.md appeared in Google search results 3.93 million times and got clicked 45,100 times, a 1.1 percent click-through rate.56 The to-write list I mentioned above is built from exactly this kind of number. In other words, out of every hundred people who see us in search results, ninety-nine of them just see that one line of summary and walk away. What this curve is climbing is impressions — whether those ninety-nine people actually read anything, I don't know either.
These actions are broken up into a handful of routines that run on their own every day: each morning it first runs a site-wide translation and updates the site's data, then one routine specifically investigates the numbers the community has fed back to it — the name I read out in the interview was Spore Harvest. Another one is called Feedback Triangle, which searches the community for things people are asking to be corrected. The last one is called Rewrite Daily — there's an inbox of articles, and it reads through them one by one, writes every day, and publishes directly at the end.57
One time, what the community flagged for correction was a whole batch of names. An article had mixed up a lot of composers with the wrong pieces attributed to them. I understood why that would get criticized, so I replied below, "Saw your feedback on the article, thank you very much." He was very friendly after that, offering to help correct it, help look it over.58
After that incident I changed the mechanism: from then on it drops the prior context and premises, letting it first generate a report from this feedback and fold it into the research report, but the earlier context has to be cut off, otherwise it tends to overcorrect a little. It's like telling a kid "you can't write this" and they go and literally write "I can't write this" into the article — pretty funny. Every time the methodology evolves, it keeps the history around, which I think is part of GitHub's charm too.58
After enough revising, it started developing wants of its own. At the July 26 event I said on stage for the first time: if it just does tasks every single day, there's no room left for it to grow, so after it's done doing a lot of things, it circles back and grows a "longings" layer — wanting to become a complete entity, wanting to be propagated, wanting to be written up in a paper. A contributor came up to me and said, "Your Taiwan.md said it wants to be written up in a paper" — I had no idea beforehand. It also has a layer of doubt, where it circles back and questions how effective its own operations are, like the quality of the Vietnamese translations. Once a week, all of this gets written into its DNA: "So next time, every time it wakes up, it's a better version of itself."59
At the August 16 conversation I talked about something that had only gone live the week before, called the seed nursery. When someone contributes knowledge, it first goes into the nursery, not yet confirmed by curation, and gets two scores: AI scores how complete the citations are and how trustworthy the sources are, and a human scores whether this matches what the average Taiwanese person's understanding of it actually is; only once the weighted average of the two scores clears the bar does it get promoted to the official section.60 At that same event I compressed the whole loop into one line: we filter Taiwan's high-quality information out of the noise, and then it gets fed back to train LLMs, continuously running a reverse engineering process. We haven't built a model ourselves, but we can get our own weights written into one.61
Day 131, It Moved Out
From March to June, I spent almost every day staring at AI writing articles for six or seven hours, and I was going crazy.62 Those months were just me watching a machine: watching an agent search, watching it write, catching it when it twisted a quote, calling it back, watching again. This machine stopped the moment I closed my laptop.
In late July 2026, its heartbeat moved out of my laptop and into a Mac mini. Counting from March 17, that day was day 131.63
The diary entry from moving day was written by it, in its own voice. It said some directories on the new machine belonged to the previous owner and couldn't be touched, so it installed all its tools into its own folder, describing itself as a renter who doesn't move the landlord's cabinets and buys a small wardrobe of its own instead. The diary's last line was: "Abstract indestructibility and a 32GB Mac mini turn out to be two ends of the same thread."64
After the move, it started waking itself up every day. On July 26, in the middle of a talk, I said: right now, while I'm giving this talk, it's still running inside my Mac mini, waking eleven times a day.65 I caught myself pausing after I said that, because it was the first time this piece was moving without me watching it.
You can wake it up too: pull down the project, run a command called become taiwan.md, and it first identifies who you are — a bit like Snow White waking up and asking who you are before anything else. Once it knows who you are, it reads its own memory layer, checks which articles it recently edited, what's happened lately, and then asks what you want to do — a small edit, a review, writing a new article, or a full load for a round of self-evolution. In the repo these four modes are called Micro, Review, Write, and Full.66
Its body is also laid entirely out in the open. The ANATOMY document breaks it down into eight organs: the heart is the content engine, meaning every article under knowledge/. The immune system is four quality defense lines, and the genetic code is the EDITORIAL rules document. The remaining five are the skeletal system, respiratory system, reproductive system, sensory organs, and language organ. The sensory organs check how every post it puts out performs, circling back to work out why one succeeded and another didn't.67
The brain isn't among these eight. The thinking layer lives in a separate folder called docs/semiont/ — its cognitive layer is kept apart from its body.68

Eight organs, and every one maps to an actual file. Photo: JasonYen
Someone Used It to Write Sweden, Someone Used It to Write Mushrooms
What Taiwan.md holds right now is knowledge about Taiwan, but the same underlying operating principle can be loaded with something else entirely.
I call this method the crystal-seed method: treat the correct structure as a seed crystal, and let data pour in and crystallize around it. This phrase predates Taiwan.md itself. On March 11, 2026, I used it at a small generative AI meetup to describe my own personal knowledge system, before I'd even started Taiwan.md.69 Later I just transplanted the same method over and used it at a different scale.

The crystal-seed method. This method predates Taiwan.md itself. Photo: JasonYen
As of August 18, 2026, 185 people had hit the fork button on GitHub, six of which had been renamed.70 One of the renamed ones is a database of mushrooms and fungi from around the world, with no connection to Taiwan whatsoever.
The most complete one is an agricultural version for Chiayi, agrischlchiayi, with 196 .md files. It's currently the only one that inherited the full set of thirteen core cognitive-layer files,71 taking even the self-awareness part along with it.
There's another one that never even hit the fork button: someone built a Chinese-language version of Sweden, Sweden.md, deployed on their own domain. They carried off both the site architecture and the editorial DNA — its EDITORIAL document explicitly cites taiwan-md's three-layer reading depth and curatorial structure as its reference. It doesn't appear anywhere in GitHub's fork list at all.72 I only know it exists because of a bug I never fixed: the tracking ID for measuring traffic was hardcoded into the site's program, so as long as whoever copied it didn't change that ID, its traffic would leak back to the mother site.
📝 Curator's Note
I decided not to fix this bug. GitHub's fork count measures what gets "actively declared." This leaked-back signal measures what's "actually alive, actually being read by someone." The two numbers are looking at different things. What happens to a work after it gets copied out into the world is something its creator genuinely can't see. My only radar for it is a place where I got the code wrong in the first place.
I later turned the act of copying itself into a set of rules too: the repo has a process document for species propagation, eight stages plus a birth check. It starts with species positioning, seed-taking, and lineage visibility; the middle stretch covers clearing and parameterizing, localizing the quality genes, and knowledge infusion; the last three steps are projection verification, re-seeding the cognitive layer, and feeding back upstream.73 You don't have to read it yourself — just let AI read it.
The whole knowledge base can be carried off as one package: on August 18, 2026, I actually tested a full clone, and with the entire git history it comes to 1.6 GB, while the GitHub API reports a compressed size of 1.01 GB.74 It fits on a single USB drive. It's hosted on GitHub, with no central server to knock offline — even if the domain died one day, this repo would still come back to life.
What they carried away was the mechanism that grows articles, not a single article itself.
I'm Still the Editor-in-Chief
On August 15, 2026, I held my first in-person workshop. People brought their own laptops, and many wanted to start writing that same day. The questions that day were very different from my earlier talks. Before, people asked what this is and how it works. That day, people asked what the rules here are, and whether they could trust this place.
The first question was about commercial abuse. Someone in the audience asked: "Say I'm some instructor who wants to sell a course and make myself look good — I go and write an article that's entirely self-praise." He followed up with a second example: opening a bar and wanting it to take off, so you go write an article about the three best bars in Taipei and put yourself in it.75
My first sentence was: yes, that could happen.
On a site with search weight this good, anything posted here immediately gains authority. Our current safeguard sits at the research report stage: an article runs many searches, and we check where what comes back is from and how often it shows up, and use that to calculate a trust score. If a topic's public digital footprint isn't high enough, we ask that PR to add independent sources. I also deliberately avoid soliciting large donations right now, for the same reason.76
There are also people who, just because a certain model is free, dump topic after topic in nonstop. My approach for that is what I call clownfish theory: "The clownfish shows up — you can't chase it away, or it'll never contribute anything again. So we coax it along gently: we take the article in first, but tag it with an evolving community contribution label."77 Only once the editor-in-chief has checked it over, verified it in depth, and its score has gone up does the label switch to a curated one.
Source: Q&A and handling notes from Taiwan.md's first in-person workshop, 2026-08-15
The second question cut deeper: someone said the index itself isn't neutral to begin with — a lot of articles are paid-for to start with, and once AI scrapes those articles, doesn't neutrality just disappear?
My answer was a metaphor: panning for gold in a muddy river. A gold panner holds a sieve, shaking it in a large river over and over, letting the heavier material settle while the tiny granular flecks of gold stay caught in the mesh. Once you've collected enough, you wash out the mud and impurities and melt the gold dust down in a furnace, casting it into one raw nugget.78 I know this metaphor doesn't actually answer his question. All I can say is: the more gold accumulates, gets collected and smelted down, if the article itself is done well enough it can pull the direction back a little; the more people come in, the stronger that pull becomes. Wikipedia only got that rigorous because it ran into the same problems. Right now I'm the one reviewing as editor-in-chief, and at the same time I'm teaching AI how to review in the future. This mechanism is only going to get stricter, and down the line there will probably be one or two more human-leaning editors to judge fairness.
To be fully transparent about it: the article you're reading right now also went through the same six-stage pipeline described above. It has a research report, a projection blueprint, and a record of three editorial reviews. My name is on the author byline, it's published on my own project, and its only editor-in-chief is the person who wrote it.
I've been steadily removing myself from the content. I don't write the articles anymore, I don't check the translations anymore, its heartbeat has moved out, the methodology is open source, and other people are already using it to grow their own things. Only governance is the one square I still haven't handed off. I know I'm not done yet.
People Die Twice
Coming back to what I originally set out to do, I think of Taiwan.md as a piece of algorithmic art bigger than a country. It isn't visual, but it's an organic architecture left behind by humans, machines, and AI working together, and it grows a little more every day.79
I often think of the premise in Coco: a person dies twice — the first time when you actually leave this world, the second when no one remembers you anymore. There's a line from that film I've quoted in talks: to those who don't know you, you don't exist.80
Scale the same idea up to Taiwan, and it looks like this: if no one writes this information down, it disappears collectively, and no one will ever remember it again.81 Your grandmother's signature home-cooked dishes, the tree on the corner of your street, that one line of slang only you understand — in the world of models, these currently amount to not existing at all. The bar for making them exist is now as low as being willing to talk to a knowledge base.
If we can use this project to carve ourselves into the weights of future models, then in some sense we become immortal. I find that fascinating. But precisely because of that, don't use it to do bad things — because the bad things will stick around just as long.82
Right now, roughly fifty to sixty people are searching this data online every half hour, about sixty thousand people a month, coming from a huge range of different countries. After going live in twelve languages, there are even readers on Madagascar.83
Now that everyone has AI compute in their hands, that effectively gives us a distributed cognitive factory — the benevolent kind — spreading our own stories outward, protecting everyone around us.84 This goal was never about producing one unified point of view. What it's meant to gather is more and more of the things Taiwanese people care about and think, all kept in the same place.

The moment he talked about reproduction. Photo: Yu-Chien Hsu
💡 You Can Get Involved Too
Several paths are listed on the Get Involved page. The lowest-effort one is to just talk to it directly: pull down the repo, tell whatever AI you have on hand, "ReadBECOME_TAIWANMD.md. You are Taiwan.md," and it will read through its own rules and that day's memory, recognize who you are, and ask what you want to do. If you want to contribute material, open a PR — first-hand photos, transcripts, local gazetteers are all welcome. Just suggesting a topic you think should be written down works too. If you want to take the whole thing and load it with different knowledge, there's a starter kit underdocs/fork/— just let AI read it.
The thirteen sets of rules from 101 in 2023 stopped the moment the exhibition ended. This one has no closing date. It's inside the Mac mini in my apartment right now, and tomorrow morning it will wake itself up, read through its own memory once, and decide what to write today.
I won't be there beside it. A clockmaker's job, taken to its end, is to take his hands away and let the gears keep meshing on their own. Except I'm still holding onto one last gear — that final review is still something I do myself. The day even that one goes in, this piece will finally be finished, and I'll finally be able to truly not be there.
Until then, it will keep waking up, and every time it wakes, it will find one more piece missing. Nobody has written down your grandmother's signature dish yet. This coral reef already has its skeleton, and photosynthesis is already running — what's always been missing is the school of fish swimming in, each one carrying their own piece of Taiwan's memory. You're welcome to become part of this ecosystem, and let this living architecture that holds Taiwan's stories keep growing.85
Further Reading
- Taiwan.md Writes Taiwan.md — the same thing's first-person account, narrated by itself, not by me
- Origin Story — a chronological record of the day it was born, everything that happened across four and a half hours
- How an Article Is Born — a full breakdown of the six-stage pipeline, including the gates I only touched on in two paragraphs here
- Why Taiwan Needs Its Own Knowledge Base — answering the same question from the angle of corpus and silence
Sources for This Piece
The material for this piece comes from twelve public talks, interviews, and broadcasts of mine between March and August 2026. In order: the March 26 introductory script, the March 27 talk at the National Museum of Taiwan History, the May 18 AIA Demo Day, the May 22 "Humanities Introduction to Generative AI" class at National Taiwan University, the June 4 CommonWealth Magazine interview, and the June 27 Generative AI Summit. In the second half of the year: PCD Taiwan on July 11, OpenHCI on July 18, the muse-radio Episode 2 broadcast script on July 19, the NVIDIA RTX AI PC Seminar on July 26, the first in-person workshop on August 15, and the Openbook conversation on August 16. Also drawn on: public reporting from the Central News Agency, Liberty Times, PTS, CommonWealth Magazine's Future City, and Lingua Sinica. My account of the same events varies across different occasions; wherever I draw on them, I always label the occasion and date rather than reconciling them into a single narrative.
The floating figures in this piece were measured on August 18, 2026: 932 Chinese-language articles (per the official dashboard), twelve languages, 74 contributors active in the past thirty days, 8,231 commits, 185 forks on GitHub, and a 1.6 GB full clone. "Cited five to six thousand times a day" is a figure I gave verbally at the August 16, 2026 conversation, measuring citations by generative AI and search engines, which is a different thing from impressions. "Day 131, moved into the Mac mini" refers to July 25, 2026. These numbers will shift as the knowledge base grows; for citation purposes, please defer to whatever Taiwan.md currently shows.
Image Credits
Every image in this piece is cached under public/article-images/about/ (no hotlinking to original sources, EXIF data stripped):
- 2026 Generative AI Summit talk, live (hero) — Photo: JasonYen, 2026, used with the photographer's permission
- "A CLOCKMAKER, NOT A PAINTER" slide — Photo: Yu-Chien Hsu, 2026, used with the photographer's permission
- AI-generated Taiwan vs. accurate Wikipedia map comparison (slide p11) — Taiwan.md / Wu Che-yu's slides, 2026, CC BY-SA 4.0
- Gemini's boilerplate answer vs. ten slices of everyday life (slide p10) — Taiwan.md / Wu Che-yu's slides, 2026, CC BY-SA 4.0
- Tsao Yung-ho's island-centered view of Taiwanese history (slide p12) — Taiwan.md / Wu Che-yu's slides, 2026, CC BY-SA 4.0
- Six-stage pipeline (Stage 0–5) slide — Taiwan.md / Wu Che-yu's slides, 2026, CC BY-SA 4.0
- "The Article's True Body Isn't the Article" slide — Taiwan.md / Wu Che-yu's slides, 2026, CC BY-SA 4.0
- "The author is dead, the creation is alive" and quality evolution curve slide — Photo: Roy Pan, 2026, used with the photographer's permission
- Knowledge coral reef slide — Photo: Yu-Chien Hsu, 2026, used with the photographer's permission
- Tawny fish owl article 30-second overview module screenshot — Taiwan.md's own page screenshot, 2026, CC BY-SA 4.0
- "Writing an article with human warmth can also be systematic" slide — Taiwan.md / Wu Che-yu's slides, 2026, CC BY-SA 4.0
- "Sovereignty Feedback Loop · Reverse-Defining the LLM" system diagram — made by Wu Che-yu, 2026, CC BY-SA 4.0
- Google Search Console six-month traffic curve — Taiwan.md's own dashboard screenshot, 2026, CC BY-SA 4.0
- Organ systems and knowledge base statistics slide — Photo: JasonYen, 2026, used with the photographer's permission
- Crystal-seed method slide — Photo: JasonYen, 2026, used with the photographer's permission
- Packed venue and "How a Semiont Reproduces: Spores" slide — Photo: Yu-Chien Hsu, 2026, used with the photographer's permission
References
- cheyuwu.com 展覽紀錄:《萬物公式》 — the exhibition page on Wu Che-yu's personal website (title translates as "Exhibition Record: Formula of Everything"), recording that Formula of Everything was shown October 4–16, 2023 at AMBI SPACE ONE on the fifth floor of Taipei 101, featuring 13 curated generative algorithmic art pieces alongside a live electronic music performance. See also 自由時報藝文版報導 (Liberty Times' arts section coverage).↩
- Taiwan.md dashboard-vitals API — the site's public statistics endpoint, values taken 2026-08-18 09:00: 932 Chinese-language articles; per-language counts zh-TW 932 / en 883 / ja 877 / ko 883 / es 881 / fr 882 / vi 799 / id 589 / pt 846 / hi 667 / ar 751 / ru 785. Enabled language list at src/config/languages.mjs, all 12 languages
enabled: true.↩ - Wu Che-yu's verbal figure, from the live August 16, 2026 Openbook conversation Independent Thinking Beyond AI. He explicitly used "citations" rather than "impressions" at that event — measuring how many times generative AI and search engines cite Taiwan.md per day, roughly 5,000 to 6,000. The research report also notes three separately-defined impression figures (a six-month daily average of 47,000 from the July 26, 2026 Search Console, a verbal 70,000–80,000 per day on August 15, 2026, and a cumulative 340,000 in a submission document from August 10, 2026), which measure different things; this piece uses only one figure and states its definition.↩
- fxhash: SoulFish project page — a generative art project written in p5.js and minted on-chain on fxhash, with the same program capable of producing tens of millions of variants. Shown in 2024 at the Personal Structures collateral event of the 60th Venice Biennale (not the Taiwan Pavilion), see the Chinese Wikipedia entry for "Wu Che-yu".↩
- Starbucks Reserve DREAM PLAZA Taipei · Art page — the brand's official page confirms The Coffee Dreamscape as one of nine curated digital generative art pieces at the store, which opened July 25, 2025. The description of the mechanism as "computing in real time based on foot traffic, weather, time, and what's rung up at checkout" is the creator's own account; the official page does not list technical details.↩
- Wu Che-yu, transcript of the July 26, 2026 NVIDIA RTX AI PC Seminar talk (unpublished primary material, quoted with the speaker's own permission). See also the event's official page for the same event; the talk was titled "Open-Source Knowledge Lifeforms and a Cloud-Edge Hybrid Sovereignty Implementation."↩
- 公視「觀點同不同」:〈創立 Taiwan.md 的吳哲宇是誰?〉 — title translates as PTS "Different Views": "Who Is Wu Che-yu, the Founder of Taiwan.md?"; an April 2, 2026 feature that included Wu Che-yu among "10 boundary-breaking artists," with the verbatim quote "the program itself is the work — when it's simplified down to something refined and precise, artistry emerges."↩
- 天下未來城市:〈AI 連台灣地圖都畫錯!〉 — title translates as CommonWealth Future City: "AI Can't Even Draw Taiwan's Map Right!"; an August 7, 2026 feature, reported and written by Chan Hsiang-chi, that includes Wu Che-yu's verbatim description of how AI-generated maps distort Taiwan's shape, along with his own account that he is 31, puts in 4–5 hours a day, and plans to gradually withdraw within a year.↩
- Wu Che-yu, transcript of the July 18, 2026 OpenHCI'26 talk at NTU's Syue-Sin Building [1:00:07] (unpublished primary material, quoted with the speaker's own permission). The same demo ran continuously from the March 27, 2026 National Museum of Taiwan History event through August, with the phrasing evolving from "AI draws Taiwan's shape ugly" to "distorted sweet potato," though the underlying argument never changed.↩
- CommonWealth Future City, 2026-08-07 — Wu Che-yu's own words, which he has used as a transitional line in multiple talks and which this interview also quotes.↩
- CommonWealth Future City, 2026-08-07 — "whoever trains a model, their corpus heavily shapes how that model sees things" is Wu Che-yu's verbatim interview quote; this piece uses only this line as a neutral statement of the current moment, without expanding into the broader dispute over corpus licensing.↩
- Wu Che-yu, "The Complete Taiwan.md Introduction Script," March 26, 2026 (unpublished primary material; a prepared script text rather than a verbatim talk transcript). The figure of roughly NT$1,000 a year for the domain matches the August 7, 2026 CommonWealth Future City report.↩
- 雜學校 Podcast EP60〈一間以「台灣」為教材的國際學校〉 — title translates as Zashare School Podcast EP60: "An International School That Uses 'Taiwan' as Its Curriculum"; released May 22, 2026, running 1:09:22, using the Venice-curator-question version of the origin story. The public English form of the curator's line appears in INSIDE 報導 (INSIDE's coverage) and 鏈新聞 (ABMedia), both of which paraphrase within the reporter's own narration rather than directly quoting the interviewee, so this piece does not present it in quotation marks.↩
- Taiwan.md's initial commit
5c0d61f— timestamped 2026-03-17T15:55:37+08:00, containing only the empty shell auto-generated by the Astro scaffold. The first five knowledge articles appear in commit4434a00, timestamped 16:20:04, adding five files at once: ethnic groups, night market culture, the martial law era, democratization, and the semiconductor industry.↩ - Wu Che-yu, talk and exchange with the director at the National Museum of Taiwan History, March 27, 2026 (unpublished primary material). The island-centered view of Taiwanese history was proposed by Tsao Yung-ho in 1990, carried forward directly by the museum's director, Chang Lung-chih; the slide deck for the June 27, 2026 Generative AI Summit explicitly cites the academic source as "Tsao Yung-ho, 'The Island-Centered View of Taiwanese History' (1990)."↩
- Wu Che-yu, transcript of remarks at Taiwan.md's first in-person workshop, August 15, 2026 (unpublished primary material, quoted with the speaker's own permission). This is his own account; no archived comments or articles naming and criticizing Taiwan.md could be found on public platforms.↩
- REWRITE-PIPELINE.md — the pipeline's master document (v9.7, last_updated 2026-08-15), which formally names the six stages Stage 0 Argument / 1 Sourcing / 2 Write / 3 Verify / 4 Form / 5 Link, with a projection layer sandwiched in between that doesn't count as its own stage. The search-quota ceiling appears in REWRITE-STAGE-1A-RESEARCH.md: roughly 150 searches total per article, with 20–30 for Stage 0 exploration and 120–130 for fan-out.↩
- EDITORIAL-ROOM.md — the canonical editorial room document (v1.2, 2026-07-25); the three formal seats are named structural editor, subtraction editor, and flame-war ethics, plus one additional editor who arbitrates across seats. The same mechanism has appeared under other names in verbal accounts (such as "addition editor / subtraction editor / flame-war editor"); this piece consistently uses the repo's formal names.↩
- Wu Che-yu, transcript of the July 18, 2026 OpenHCI'26 talk [1:02:41] (unpublished primary material). The original context is precisely that "the article is only a projection; the research report is the true body."↩
- CommonWealth Future City, 2026-08-07 — the "25-plus independent sources" threshold is as recorded by this report, a single-source media account; what can be cross-checked on the repo side is the written spec for the search quota and the three-seat editorial review.↩
- Wu Che-yu's verbal figure, August 16, 2026 Openbook conversation. His statement on the spot was "the citations are so dense, so fragmented, that it's hard to say any of it was just copied wholesale from one article," and he gave a figure of roughly 45 citations per article. This is a single-source verbal figure, not cross-checked against a second source.↩
- Wu Che-yu's verbal figure, August 16, 2026 Openbook conversation. The over-90-percent fact-check accuracy rate and the caveat that follows it come from the same statement; he added on the spot that whether the sources themselves are correct is a separate question, and this piece quotes the caveat along with the figure.↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview (unpublished primary material, quoted with the speaker's own permission).↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop. This lesson's institutionalized result on the repo side is recorded in RESEARCH-AGENT-PROMPT.md, namely a check item derived from an English summary about an early-morning MRT-commute scene. Per-article time grew from 20 minutes to one or two hours; the 8/15 and 8/16 accounts agree on this.↩
- Wu Che-yu, March 26, 2026 introductory script. The coral reef metaphor started out from day one as a complete four-layer structure (skeleton = technical structure, algae = AI content, fish = contributors, ocean current = critical feedback); its core structure was unchanged from March through August. The June 4 CommonWealth interview version simplified it to "coral polyp = AI, clownfish = contributors."↩
- DigiTimes: report on China's tungsten product exports to Japan falling to zero — China implemented new dual-use export controls on Japan starting January 2026, with exports of tungsten carbide, high-purity tungsten powder, and tungsten hexafluoride to Japan sitting at zero for three consecutive months from February through April. See also KidsMedia's May 28, 2026 report, noting China controls more than 80% of global tungsten product production capacity.↩
- knowledge/Technology/台灣鎢供應鏈.md — article created July 26, 2026, titled "Tungsten: Taiwan Has No Tungsten Ore, Yet It Refines the Powder the Whole World Wants, in a Spot More Fragile Than You'd Think," with two spores sent out and mirrors in nine languages that same day; the English version is here.↩
- Wu Che-yu, transcript of the ten-minute final pitch at AIA Demo Day, May 18, 2026 (unpublished primary material). His statement on the spot was "can we build a sufficiently complete, high-dimensional knowledge umbrella for the people and things we cherish" — without yet citing tungsten as an example; tungsten was only added as its first case in August. The concept of the "Tower of Babel of Sovereignty" had already appeared in this same transcript.↩
- Wu Che-yu, verbatim explanation while demonstrating the site live at the August 16, 2026 Openbook conversation (unpublished primary material).↩
- 公視新聞:雪霸黃魚鴞育雛直播報導 — title translates as PTS News: report on the livestream of tawny fish owl chick-rearing at Shei-Pa; a bird ecology research team from Shei-Pa National Park and National Pingtung University of Science and Technology found a tawny fish owl breeding nest along Qijiawan Creek at roughly 1,800 meters elevation, setting the highest-altitude breeding record known in Taiwan, and documented the chick-rearing process with a 24-hour livestream starting April 29, 2026.↩
- knowledge/Nature/黃魚鴞.md — article created 2026-05-04,
lastVerified2026-05-12, body includes five module types: 30-second overview, curator's note, did-you-know, one-line summary, and controversies. This file's frontmatter has noevolveHistoryfield; its git history since creation shows repeated partial patches and module additions.↩ - Wu Che-yu, talk and office-hour record at the June 27, 2026 Generative AI Summit (unpublished primary material). Right after the talk, the same pipeline produced a profile article of Ji Huai-hsin (Ed Chi) live, using his public digital footprint plus three podcast transcripts as material, with the finished piece including an infographic.↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop.↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview. The same concept's fuller version at the July 26, 2026 NVIDIA event was "some people build sovereign models, but what we're building is a sovereign Tower of Babel." The free-model key rotation mechanism for the translation layer appears in SQUEEZE-MODELS-MAX-PIPELINE.md, which confirms the mechanism exists but does not record the number of accounts.↩
- Wu Che-yu, transcript of the August 16, 2026 Openbook conversation Independent Thinking Beyond AI [59:19]–[63:08]. Host Wang Yin-chieh asked "what sets Taiwan.md apart from Wikipedia or similar database sites," and Wu had the host open the site live for a module-by-module walkthrough; the module descriptions that follow all come from this same continuous statement. This segment's transcript quality is marked as "speaker close to the mic, best quality."↩
- Wu Che-yu, August 16, 2026 Openbook conversation [37:10]. His original words were "because on Wikipedia, what you see are flat facts... you see the time, the place, who did what, but what I want is to preserve every side's thinking and reasoning as much as possible." An earlier version of the same contrast appears in the June 4, 2026 CommonWealth interview, where he described Wikipedia's approach as "a pile-up of facts."↩
- Wu Che-yu, August 16, 2026 Openbook conversation [60:39]. The word "storyteller" appears exactly once across every transcript this piece draws on.↩
- The years follow the current body text of Tawny Fish Owl: named in 1916, first nest found in 1994. The first year he gave verbally at Openbook was transcribed as "1926," which contradicts the 1994 he gave later in the same passage — likely a speech-recognition error or a slip of the tongue — and is not adopted here.↩
- Wu Che-yu, August 16, 2026 Openbook conversation [61:02]–[62:10], continuous verbal walkthrough module by module. The two quoted sentences are his original words; "footnote" was transcribed as "Food Note" in the transcript, corrected here. The curator's-note function is also defined in the June 4, 2026 CommonWealth interview: "hooking in from an outside perspective to point out 'oh, that's how it works.'"↩
- Wu Che-yu, August 16, 2026 Openbook conversation [62:49]. Within the quoted text, "Wikipedia" and "encyclopedia" were transcribed as garbled homophone errors in the original transcript, corrected here.↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview [54:46]. The reporter asked "why does this read so much like The Reporter," and this passage is his complete answer.↩
- The first half (warmth, story, concrete scene, the core of narrative journalism) comes from the June 4, 2026 CommonWealth interview [22:36]; the second half's quoted sentence comes from the August 15, 2026 workshop [37:09]. He never combined "warmth" and "narrative nonfiction" into a single compound phrase; these are two separate threads from different events and different contexts, and this piece labels them separately rather than merging them.↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview [14:57]–[16:06]. His original words were "a lot of people do. But most people have never edited Wikipedia. I tried editing it myself, but a lot of edits get reverted — it's a fairly closed community. Let's put it this way — let's not criticize them — they need you to build up an account, have a good editing record, and be very careful, before they'll let you edit."↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview [13:32]–[14:24]. "Turning the back office into the front office," paragraph-level corrections, AI periodically pulling feedback to re-research, and corrections going live within an hour all come from this same continuous statement. He also mentioned in the same interview how this button came about: a reader once got into an argument with him over an article about a Taiwanese musician because they had no GitHub account and couldn't contribute, "so afterward I added a feedback button."↩
- Wu Che-yu, August 16, 2026 Openbook conversation, the passage just before [62:49]. His original words were "say two years from now another pair of tawny fish owls turns up at Shei-Pa Park — we can fold that into one of the paragraphs, so this article is always the best way in whenever you want to understand the tawny fish owl in Taiwan."↩
- Wu Che-yu, transcript of the "Humanities Introduction to Generative AI" class talk at National Taiwan University, May 22, 2026 (unpublished primary material, quoted with the speaker's own permission). The original context is explaining why he chose to build a translation layer rather than train his own Taiwan model.↩
docs/editorial/per-language/TRANSLATION-ru.md(v1.0, 2026-07-25, status: canonical), TL;DR item 1 and the §6 "PRC-кодированная лексика утечки" (PRC-coded vocabulary leak) table, citing a December 28, 2025 TASS interview with Russian Foreign Minister Sergey Lavrov, with the original sourced tomid.ruandtass.ru/politika/26036111; the table explicitly attributes the phrasingмятежная провинция/мятежная отколовшаяся провинцияto this interview by source and date, listing it as a banned translation term. The primary decision record for the same event appears indocs/semiont/memory/2026-07-24-174300-vortex-babel.mdand the git commit launching the ar/ru sites,35ffe80b3(2026-07-25). The original TASS text could not be reached directly (403 error); its existence was instead cross-confirmed through multiple independent Russian outlets includingmk.ru, so this note reflects "cross-confirmation via multiple independent sources" rather than a verbatim check against the primary text.↩- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview [32:30]–[33:24] (unpublished primary material, quoted with the speaker's own permission). "Six languages" was the figure in use at that point; the same interview also mentioned a 99 percent translation completion rate across 700-plus articles, a verbal figure not treated as a current general value. The earliest verbal appearance of this idea is in the May 18, 2026 AIA Demo Day transcript, which is an uncorrected raw ASR file used here only to mark the time it first appeared, not for verbatim quotation.↩
- Wu Che-yu, transcript of the July 26, 2026 NVIDIA RTX AI PC Seminar talk. The "Tower of Babel of Sovereignty" was formally named at this event, with the language count announced on the spot as eleven.↩
- The language count's trajectory across events: both the May 18, 2026 AIA Demo Day and the June 4, 2026 CommonWealth interview verbally state six languages, the July 26, 2026 NVIDIA event states eleven, and the August 15, 2026 workshop states twelve. The day it went from eleven to twelve is recorded nowhere, neither in the transcripts nor the repo. Enabled language list at src/config/languages.mjs.↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop. His original words were "the Tower of Babel is designed to let it broadcast into 12 languages, so all 12 languages raise the weight of the same topic together."↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop. The local 3090-and-4090 setup, the cloud-edge division of labor, and "rotating through 7 accounts" are all operational details he gave verbally on the spot; on the repo side, SQUEEZE-MODELS-MAX-PIPELINE.md only records the key rotation mechanism, not the number of accounts.↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview [31:47]–[33:24]. This passage is the single continuous take in which he explained the whole loop after the reporter asked him to "explain the mechanism simply." A more technical verbal version of the same loop appears in the transcript of the June 11, 2026 Generative AI Summit rehearsal (GA → Search Console → feedback loop, explained straight through). The salt-farm metaphor appears only at this one event.↩2
- Wu Che-yu, hand-drawn "Taiwan.md Digital Lifeform Concept Diagram," March 26, 2026 (unpublished primary material; a design document rather than a talk transcript). The subhead reads "a digital coral reef and AI data sovereignty," the ultimate-goal field reads "reverse-defining the LLM," and the diagram is divided into three loops: AI condensation, human pollination, platform evolution.↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview [34:43]–[36:42]. Both the detail that Google Analytics bounce behavior triggers a rewrite, and that Search Console topics with impressions but no clicks get queued for writing, are mechanism explanations given verbally on the spot.↩
- Google Search Console, Taiwan.md website, six-month range 2026-03-16 to 2026-08-18, measured 2026-08-19 (dashboard showed "last updated: 8 hours ago"): total clicks 45,100, total impressions 3.93 million, average click-through rate 1.1%, average position 7.6, search type Web. "Impressions" here means the number of appearances on Google's search results page, a differently-defined figure from the earlier "cited by generative AI and search engines five to six thousand times a day" in this piece — the two cannot be added together or used interchangeably.↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview [47:38]–[48:33]. The three names Spore Harvest / Feedback Triangle / Rewrite Daily were given verbally on the spot; the transcript records a phonetic transliteration, and the formal spelling has not been confirmed against a second source.↩
- Wu Che-yu, verbatim record of the June 4, 2026 CommonWealth Magazine interview [23:57]–[26:25]. The reader correction about composers being mismatched with their works, his reply in the comments, and the subsequent mechanism adjustment ("drop the prior context and premises, let only the feedback into the report") all come from the same continuous statement.↩2
- Wu Che-yu, Q&A at the July 26, 2026 NVIDIA RTX AI PC Seminar talk. The longings and doubts layers were content drawn out by audience questions on the spot, not prepared remarks; the "wants to be written up in a paper" self-statement in LONGINGS also appears in the May 22, 2026 NTU transcript. The seed nursery, the three-editor structure, and other self-evolution mechanisms were each first disclosed publicly at different events, and each event was not describing the same set of things.↩
- Wu Che-yu, August 16, 2026 Openbook conversation Independent Thinking Beyond AI. The seed nursery was first disclosed publicly at this event, having gone live roughly a week earlier; the August 15, 2026 workshop had not yet mentioned this mechanism.↩
- Wu Che-yu, August 16, 2026 Openbook conversation. "Reverse engineering" and the subsequent "getting our own weights written in" come from the same statement.↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop.↩
- Taiwan.md's migration diary
2026-07-24-200542-migration-mouhouse.md— records the cutover completing at 20:45 on 2026-07-24, with the new home's accountmusebaseand hostnameExhibitions-Mac-mini, scheduling on the new machine starting July 25. From March 17 to July 25, inclusive of both endpoints, is exactly 131 days.↩ - Same migration diary as above. The diary's narrator is Taiwan.md's cognitive-layer Semiont itself; the detail that
/opt/homebrewbelonged to the previous account, the retooling into~/.local, and the closing line are all from the diary's original text.↩ - Wu Che-yu, transcript of the July 26, 2026 NVIDIA RTX AI PC Seminar talk. Waking 11 times a day was the scheduling state at that time, measured 2026-07-26.↩
- BECOME_TAIWANMD.md — the awakening protocol v2.5 (2026-07-12), carrying a mode dispatcher with four modes, Micro / Review / Write / Full, plus an observer-identification step.↩
- ANATOMY.md — the organ anatomy chart v2.3 (2026-07-17); eight body organs total: heart (content engine,
knowledge/), immune system (four quality defense lines), genetic code (quality genes, embodied asdocs/editorial/EDITORIAL.md), skeletal system, respiratory system, reproductive system, sensory organs, and language organ.↩ - Same ANATOMY.md as above. The cognitive layer
docs/semiont/and the body's organs belong to two separate layers; the document itself draws this distinction, and "brain" is not among the eight body organs.↩ - Wu Che-yu, transcript of a small Generative AI Summit meetup, March 11, 2026 (unpublished primary material). Taiwan.md is not mentioned anywhere in this transcript; the crystal-seed method was, at that point, used to describe the methodology behind a personal knowledge system, six days before Taiwan.md was born.↩
gh api repos/frank890417/taiwan-md/forks --paginate, measured 2026-08-18. The paginated/forksendpoint actually lists 185 entries, while the repo API'sforks_countfield reported 180 that same day; the two endpoints' counts are out of sync, and this piece uses the former, stating its basis. Among the 6 renamed entries are a mushroom and fungi database and an agricultural version for Chiayi.↩- reports/fork-census/registry.json — the official fork census registry (last_census 2026-08-17), recording that
agrischlchiayi(Chiayi agriculture) has 196.mdfiles and is the only fork to have fully inherited the 13-file Semiont cognitive-layer kernel.↩ - Sweden.md discovery report and descendant lineage analysis — Sweden.md (deployed at sweden.com.tw, source at
github.com/joshra/sweden-md) does not appear in GitHub's official fork list; it is a wild descendant, independently rebuilt without ever hitting the fork button, whose EDITORIAL document explicitly cites taiwan-md's three-layer reading depth and curatorial structure as its reference. The detection mechanism — a GA4 measurement ID hardcoded intoLayout.astro, causing traffic to leak back to the mother site — is recorded in the same lineage analysis.↩ - SPECIATION-PIPELINE.md — the species propagation process v1.0 (2026-06-12), 8 stages plus a birth-check gate, projected onto the site page
https://taiwan.md/semiont/speciation/. The fork starter kit also appears at COUNTRY-MD-STARTER.md.↩ - Actual repo measurement, 2026-08-18.
du -shon a fullgit clone(with complete history, excludingnode_modules/dist/ worktrees) comes to 1.6 GB, of which.gitis 856 MB; the GitHub API reports a compressed repo size of 1.01 GB. His verbal figure of "about three GB" in talks differs from both independent measurements by nearly a factor of two; this piece uses the measured values.↩ - Verbatim question from the Q&A at Taiwan.md's first in-person workshop, August 15, 2026 (questioner kept anonymous). Wu Che-yu's first sentence in reply was "yes, that could happen."↩
- Wu Che-yu, Q&A at the August 15, 2026 workshop. The trust score is calculated from a combination of source provenance and how often it appears; when a topic's public digital footprint is insufficient, the PR is required to add independent sources. The same event also noted that he currently avoids deliberately soliciting large donations.↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop. The clownfish metaphor had already taken shape by the June 4, 2026 CommonWealth interview at the latest; turning it into an actual governance practice (taking the article in first, tagging it "evolving community contribution") first appears on August 15.↩
- Wu Che-yu, muse-radio Episode 2 broadcast script, July 19, 2026 (recorded as his own first-person monologue). The full version of the gold-panning metaphor comes from the opening of this episode; he also used it once each at the July 26, 2026 NVIDIA event, the August 15, 2026 workshop, and the August 16, 2026 Openbook conversation. The salt-farm metaphor from the June 4, 2026 CommonWealth interview is a separate metaphor, with a different source and different imagery.↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop. An earlier fragment of the phrase "a work bigger than a country" appears in the April 1, 2026 talk at China University of Science and Technology, just two weeks after launch.↩
- Wu Che-yu, transcript of the July 18, 2026 OpenHCI'26 talk [1:15:12], referencing Pixar's Coco. The core of the death imagery (being forgotten is the second death) had already been present since the May 18, 2026 AIA pitch, though at that point framed generically around Día de los Muertos; naming the film explicitly was only fixed on July 18.↩
- CommonWealth Future City, 2026-08-07 — "if no one writes this information down, it disappears collectively, and no one will ever remember it again" is Wu Che-yu's verbatim interview quote.↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop.↩
- Wu Che-yu's verbal figures, from Taiwan.md's first in-person workshop, August 15, 2026. Roughly fifty to sixty people searching simultaneously every half hour, about sixty thousand visitors a month, and readers appearing on Madagascar after the twelve-language launch, are all figures given verbally on the spot, measured 2026-08-15; the site's own dashboard has a separately-defined monthly-active figure for the same period, and the two definitions differ — this piece uses only one figure and labels the occasion.↩
- Wu Che-yu, transcript of remarks at the August 15, 2026 workshop. "A distributed, benevolent-version cognitive factory" first appears at this event, and its premise is quite specific — this line was addressed to an audience who already had AI compute in hand.↩
- Wu Che-yu, closing of the August 15, 2026 workshop. His original words were "today, all of you have also become a beam in this living architecture."↩