Public AI

Mozilla AI at Internet Archive Europe: Owning Your AI Stack

On 25 June, Internet Archive Europe (IAE) hosted Davide Eynard and Thomas (toto) Bille from Mozilla AI at our Amsterdam space. The afternoon covered a question that sits close to the heart of what we do: when you depend on infrastructure you don’t control, what do you actually own?

Ada, Zangemann, and the case for tinkering

Davide opened with a children’s book: Ada & Zangemann. Ada is a girl who lives in a dumpster, salvages broken hardware, and builds things entirely her own. Zangemann builds beautiful, polished technology that no one else can modify or adapt. The clash between them drives the story, but Davide used it as a frame for something more immediate: most people’s relationship with AI today looks a lot more like Zangemann’s world than Ada’s. You use what you’re given, on the terms it’s offered, for as long as the provider decides to keep it available.

This has nothing to do with technology: it’s a power relationship. And that logic applies to memory institutions every bit as much as to individual developers.

The trade-offs are concrete

To make the point, Davide rebuilt a small tool he’d originally written by hand more than twenty years ago: a script to extract train timetables from a website too clunky to use directly. He produced the rebuilt version with an AI coding assistant, and it worked. But the result lived on someone else’s platform, not his own machine. And unlike the original, which taught him Perl and regular expressions he used for years afterward, this one taught him nothing. Convenience and ownership turned out not to be the same thing.

Search as activity, not action

One of the sharpest distinctions in Davide’s talk was between search as an action and search as an activity. When you type a question into a box and accept the answer, that’s an action. When you use an agent to follow a thread, evaluate what it finds, redirect it when it goes wrong, and build toward a conclusion over time, that’s an activity. The difference matters because the second approach keeps you in the loop. You catch mistakes. You steer.

The clearest example: Davide used an agent to track down the original source of a widely cited Bill Gates quote. The agent searched, hit dead links, found partial copies, and eventually installed a subtitle extraction tool autonomously, downloaded a YouTube video, and identified the exact moment Gates said the thing. Along the way, it returned to the Wayback Machine around a dozen times, working through broken URLs until it found a usable copy.

It was a quiet illustration of something IAE says often: preserved web history is not a nostalgia project. It is working infrastructure. AI systems now depend on it to check what is actually true.

A second demo used a locally run open source model to search the Rijksmuseum’s digital collection for images connected to alchemy and the pursuit of knowledge, producing usable results entirely on Davide’s own hardware. No cloud service, no rented compute, no data leaving the machine.

Otari: making ownership practical

Thomas (toto) Bille followed with a look at the infrastructure side of the problem. Mozilla AI has built Otari as an open-source LLM gateway: a single control plane for all your interactions with language models, whether you run a local model on your own server or route through a commercial API.

The features that generated the most discussion were practical ones. Budget controls granular enough to cap spending per user, per model, or per team. Guardrails that strip personal or sensitive data before it reaches any external provider. A federated router in development that will recommend which model to use based on real usage patterns across the community, with options to prioritise cost, quality, or energy use. Otari is fully self-hostable: if you don’t want your data passing through Mozilla AI’s servers, you don’t have to. The code is on GitHub.

The honest question

Someone in the room asked how Mozilla AI intends to stay financially viable if the tools are free and open source. Toto’s answer was direct: revenue from the hosted version, income from enterprise integration work, and a long-term bet on the community. It’s the same tension that runs through almost every public-interest digital project, including IAE’s own. There’s no clean resolution, but naming it honestly matters.

Why this conversation belongs here

The afternoon drew developers, researchers, and people from across the cultural heritage sector. The discussion after the talks ran for close to an hour.

What connected the room wasn’t a shared technical interest in language models. It was a shared unease about dependency. Memory institutions know what it looks like when access to knowledge sits on infrastructure you don’t own and can’t influence. IAE has spent years arguing that the rights archives have always held offline must be protected online too. The same holds for one layer up, to the AI systems now sitting on top of those archives.

Explore Mozilla AI’s open-source tools, including AnyAgent, AnyLLM, LlamaFile, and Otari, at mozilla.ai and github.com/mozilla-ai.

You can watch a replay of the presentation and conversation on Archive.org and see the slides online.

Mozilla AI at Internet Archive Europe: Owning Your AI Stack Read Post »

Most of the Renaissance Has Never Been Translated. Source Library Is Opening It.

Ninety percent of Renaissance Latin has never been translated into a modern language. More Latin was written after 1500 than survives from all of ancient Rome, and almost none of it has been read outside a specialist library. At the current pace of human scholarship, completing that translation work would take approximately 12,000 years.

On 4 June, Internet Archive Europe attended the Source Library BETA Launch at the Embassy of the Free Mind in Amsterdam. It was a milestone worth marking.

What Source Library Is

Source Library is the world’s largest freely available collection of translated historical primary sources from the Renaissance. At launch, it holds more than 15,000 books across 55 languages, including 6,000 first-ever English translations and roughly seven billion words of original text and translation, comparable in scale to the entire English Wikipedia. Works previously readable only by Latin scholars or locked behind expensive academic editions are now open to anyone.

The project is hosted at the Bibliotheca Philosophica Hermetica, the UNESCO Memory of the World-recognised collection at the Embassy of the Free Mind: more than 25,000 volumes on alchemy, Hermetica, Kabbalah, Rosicrucianism, and the roots of modern science. Many of these books were banned at various points in history. Now they are open.

The launch carries particular meaning in light of what followed. Joost Ritman, the Amsterdam businessman who founded the Bibliotheca Philosophica Hermetica and built it into one of the world’s great collections of philosophical, religious, and esoteric knowledge, died on 5 June 2026, the day after Source Library launched in the institution he founded . He had spent sixty years guided by a conviction he traced to the Florentine Medici: that those in a position of privilege carry an obligation to culture. In 2017, he donated the library, its research institute, and the House with the Heads Monument to a cultural non-profit foundation, publicly known as the Embassy of the Free mind. This was a gift to Amsterdam and the world: he made permanent what he had spent his life assembling, and he made it public. 

AI as Accessibility, Not Replacement

Source Library places AI-powered translations directly alongside images of the original source pages, so anyone can consult the original at any point. The goal, as project creator Dr. Derek Lomas of Delft University of Technology made clear at the launch, is not to replace scholarship but to make a vast body of untranslated material discoverable for the first time. Dr. Lomas is a cognitive scientist and human-computer interaction researcher, currently a professor of Human Centred Design at Delft University of Technology, who arrived at Renaissance philosophy through a long personal engagement with the Neoplatonic tradition. That combination gives Source Library a design sensibility that most digital archive projects lack: the design starts from how people actually discover and engage with material, not from how institutions prefer to organise it. 

Dr. Lomas and the team also maintain careful transparency about data sources throughout: content drawn from the library’s own catalogue is clearly distinguished from AI-generated material, and all translations record the model, date, and prompt used to produce them. That distinction matters. The AI output is treated as useful but revisable. The primary sources are  treated as the foundation it is, complimented by academic curatorial work.

The project carries an AGPL-3 licence, the same open source licence used by the Internet Archive. It draws on open digital image standards that allow libraries and archives to share their collections freely, and it acknowledges the institutions whose digitised holdings made the work possible.

Why This Matters

The question Source Library poses is one we encounter constantly: who gets to access knowledge, and on what terms?

For centuries, the thought documented during the Renaissance has been available only to those who read Latin, have access to specialist collections, or can afford expensive critical editions. Source Library removes those barriers. It doesn’t replace careful scholarship. It makes a vast body of human thought discoverable for the first time.

This is open access made concrete: 6,000 first translations, 55 languages, no paywalls. It also matters for AI. The training data available to language models shapes what they know and how they reason. A Renaissance that remains untranslated is a Renaissance that AI cannot draw on. Source Library is building the corpus that public-interest AI will need.

Explore Source Library at sourcelibrary.org.

Most of the Renaissance Has Never Been Translated. Source Library Is Opening It. Read Post »

Maps Are Infrastructure Too

Not all infrastructure looks like a server room. Some of it looks like a map.

On 19 May, Internet Archive Europe hosted a workshop at our Amsterdam space with Bart Louwers, one of the core maintainers of MapLibre, and Tommi Marmo, OpenStreetMap contributor. The conversation moved between the technical and the political in a way that felt familiar: who builds the tools, who controls them, and what happens to the communities that depend on them when that control is concentrated in too few hands.

Two projects, one principle

OpenStreetMap is sometimes referred to as “Wikipedia for maps”. Anyone can contribute to it, anyone can use it, and no single company owns it. It does not just show you where things are. It is a database that anyone can build on, and it maps things that proprietary providers choose not to, from water fountains to informal settlements. 

MapLibre began in 2020 as a fork of a proprietary mapping tool. By 2022, it had its own sponsors, board, and charter. Today it is the rendering engine behind navigation apps, drone software, Wikimedia’s Android app, and products from companies including AWS, Microsoft, and Meta. The code is open source. The community governs the roadmap.

What makes MapLibre technically distinctive is its vector-based approach. Rather than downloading pre-rendered image tiles and stitching them together, the client receives descriptions of polygons, lines, and points and draws the map itself, in real time. The result is continuous zoom, styles that can change on the fly, and a format that is far more efficient. The entire planet fits in 136GB. You can host your own archive of it.

Why this matters beyond mapping

The question Tommi and Bart kept returning to is one we ask every day in a different context: who decides what is visible, what is preserved, and what disappears? Platforms rise and are acquired. Services shut down without notice. The communities built on them, and the data they generated, often go with them.

Open infrastructure is not a technical preference; it is the condition under which genuine public access becomes possible at all.

OpenStreetMap runs on community input. For example, StreetComplete, a map editing Android app, turns local surveying into something close to a game, sending contributors out to fill in what commercial map providers overlook. Another illustration is the Humanitarian OpenStreetMap Team project, which mobilises contributors after disasters, updating maps in real time so that aid reaches the right places. Corporations depend on OpenStreetMap as a crowdsourced map dataset, with TomTom even acknowledging this fact with their slogan “It takes the world to map the world”.

I joined this session wearing a second hat, beside my usual Internet Archive Europe one. The European Public Domain Day Working Group has been building a map on pdday.org, which tracks where public domain celebrations are taking place across Europe. That proprietary map has recently been updated and replaced with an open source version, and the community is actively looking for contributors to help keep it current. If you know of celebrations happening in your region, we would welcome your input.

Start exploring

For anyone curious about these tools: maplibre.org is the starting point. Maputnik at maputnik.github.io is a visual style editor for MapLibre maps, whose styles are in JSON rather than CSS. There are many hosted basemaps available, including non-commercial ones such as OpenFreeMap which has a quick-start guide for publishing your own maps. PMTiles lets you create and serve a permanent, self-contained archive of map data.

MapLibre’s Technical Steering Committee meets every second Wednesday of the month. Details and a community invite link are at maplibre.org/community.

The tools that remember where things are matter. So does who builds them.

We are grateful for the time and expertise that Bart and Tommi both shared with us, and the intimate group of participants who came to learn, explore, and build together.

Maps Are Infrastructure Too Read Post »

ClimateGPT is a Webby Nominee. Vote Before 16 April.

Most AI tools are black boxes. You do not know what they were trained on, who decided what counts as reliable, or whether the answers they produce can be checked against anything. ClimateGPT 3+ is built on a different premise: that climate intelligence must be open, auditable, and grounded in solid data. The Webby Awards have taken notice.

ClimateGPT 3+, a project developed by Erasmus.AI and supported by Internet Archive Europe, has been nominated in the AI: Energy and Sustainability category of the 30th Annual Webby Awards. This year more than 13,000 projects entered; ClimateGPT 3+ placed in the top 11%. A People’s Voice Award, voted on by the public, is now within reach. Voting is open until 16 April 2026.

The People’s Voice Award

The Webby People’s Voice Award is voted on by anyone, anywhere. Last year nearly 3.6 million votes were cast from more than 230 countries. The award is a signal to the sector about what kind of AI the public actually values.

ClimateGPT earned this nomination by doing something most AI platforms do not. Voting for it is a vote for the principle that AI serving the climate transition should be open, accountable, and grounded in the best available science. It should not be proprietary, locked behind paywalls, or optimised for engagement over accuracy.

Vote at vote.webbyawards.com before 16 April 2026, and visit climategpt.ai to explore the tool directly.

What ClimateGPT Is, and Why It Is Different

ClimateGPT is an open-source ensemble of large language AI models built to augment human decisions on climate change. It was trained on a corpus of over 10 billion web pages and millions of open-access academic articles, synthesising interdisciplinary research across the natural, social, and economic sciences. The model is available in more than 20 languages and is free to use for researchers.

That is not a minor technical detail. The decision to make the model open source, to publish the training data lineage, and to make it available at no cost means that a researcher in Nairobi can access the same climate intelligence as a policymaker in Brussels. Users range from individual practitioners to institutions like NASA.

The model benchmarks show ten times the efficiency on climate-specific tasks compared to general-purpose models, and a cascading machine translation approach that recovers nearly 94% of fluency performance relative to native multilingual models. Crucially, it was trained and is hosted on renewable energy.

Why Internet Archive Europe Supports ClimateGPT

Internet Archive Europe supports ClimateGPT because the initiative directly aligns with the mission of universal access to knowledge. ClimateGPT demonstrates that combining planetary-scale datasets with open, decentralised technology empowers citizens and governments to make better decisions. It is AI built for transparency and adaptation, not just automation.

This matters for governance as much as for science. Climate disinformation is not an abstract problem. It shapes legislation, investment decisions, and public understanding of risk. A model that is auditable, grounded in peer-reviewed sources, and built to counter disinformation rather than amplify it represents a different category of AI development from what currently dominates the market. The question of who builds AI, on what data, and for whose benefit is a political question as much as a technical one. ClimateGPT answers it in the public interest.That is what this nomination recognises. Vote to say it matters.

ClimateGPT is a Webby Nominee. Vote Before 16 April. Read Post »

Celebrating “Humans of AI”: A Journey Into Public-Interest Technology

Internet Archive Europe is proud to support the release of AI Lab Perspectives: Humans of AI, the seventh report by information labs, based on a series of expert video capsules.

This timely and thoughtful publication brings together voices from across Europe’s cultural heritage ecosystem to explore a question that could not be more important: what does it mean to build AI that truly serves the public?

As the report’s introduction makes clear, Humans of AI aims to spark informed conversation and critical reflection on how artificial intelligence is being applied within the cultural heritage sector across Europe and beyond. Through in-depth podcast conversations with libraries, museums, archives, artists, researchers, and digital platforms, the series highlights real-world projects, ethical challenges, and practical lessons that demonstrate how AI can support access, memory, and public engagement.

At Internet Archive Europe, this mission resonates deeply.

AI as Public-Interest Infrastructure

Across all ten case studies, one shared insight stands out: AI is neither a miracle nor a threat; it is a tool. Its impact depends on how it is governed, who shapes it, and whether it strengthens public knowledge and democratic access.

The projects featured in the report illustrate this principle in powerful ways:

  • Preserving History: From Transkribus turning historical manuscripts into searchable text to the National Library of Norway building language models grounded in local culture.
  • Innovating Access: Projects like Litte_bot bring literary characters to life, while the Museum Goggles initiative uses AI to help us understand how visitors truly experience art.
  • Memory & Ethics: The Synthetic Memories project uses AI to reconstruct lost personal histories, and Europeana uses it to identify and contextualise contested colonial terms in metadata.
  • The Web of the Past: We are especially thrilled to see Kai Jauslin’s work on the Websites van Nederland project, part of the Webarchiving Display of the Internet Archive Europe. This initiative transforms massive web archives into interactive “fields of stars,” making born-digital heritage tangible and explorable for the public.

For us, the themes of openness, shared ownership, multilingual access, ethical governance, and human oversight are not abstract ideals. They are foundational principles. We believe that digital heritage, including the web itself, belongs to everyone. AI, when responsibly developed, can help ensure that these collections remain explorable, meaningful, and usable for generations to come.

Bringing Collections to Life — Responsibly

One of the most inspiring aspects of Humans of AI is its refusal to fall into hype or fear. The report consistently underscores that the real challenges around AI are social and institutional—adoption, governance, trust, and long-term sustainability—rather than purely technical.

From reconstructing lost memories through guided conversations to using AI-powered eye-tracking to better understand museum engagement to enriching metadata across millions of records, the report shows that AI’s greatest value lies in supporting human interpretation, not replacing it.

This aligns closely with Internet Archive Europe’s commitment to:

  • Preserving digital memory at scale
  • Supporting open data and open-source innovation
  • Making archives explorable, not hidden
  • Ensuring that public knowledge is not mediated exclusively through private platforms

AI can help bring collections to life, but only if it remains grounded in public interest, transparency, and shared stewardship.

Looking Ahead: The Second Series of Case Studies

We are especially proud to see the breadth and diversity of contributors in this first edition, from national libraries and global open-source communities to artists and experimental designers. The range of voices reflects the richness of Europe’s cultural and digital heritage ecosystem.

Internet Archive Europe looks forward with great anticipation to the second series of Humans of AI case studies. Continued documentation of practical, ethical, and public-interest applications of AI will be essential in shaping a European approach that is confident, values-driven, and collaborative.

By amplifying real-world examples rather than abstract speculation, this series provides policymakers, cultural institutions, and technologists with something invaluable: grounded insight.

Celebrating “Humans of AI”: A Journey Into Public-Interest Technology Read Post »

Davos Event Spotlight: Launching ClimateGPT 3 and the Future of Public Good AI

On 22 January, at Goals House Davos, ClimateGPT 3 will be officially launched during a roundtable discussion on Planetary Boundaries – New Modes of Action in a World Beyond 1.5°C. Internet Archive Europe is proud to support this initiative, which explores how open knowledge, data, and technology can help societies understand and respond to accelerating climate risks.

A New Lens on Climate Action

As the world moves beyond the 1.5°C threshold, governments, businesses, and constituencies are showing declining engagement. Emission reduction targets and transition plans – already deemed insufficient by scientists – are being scaled back.

This raises a fundamental question: should public and private resources focus on adaptation rather than mitigation? Is it time to embrace the unimaginable—an adaptation agenda at scale—and learn to live with climate change?

In this context, new modes of change are emerging: less globally aligned, more bottom-up, technologically empowered, and often citizen-led.

Daniel Erasmus, Founder of ClimateGPT and Head of AI of Internet Archive Europe, will unveil the third iteration of Public Good AI ClimateGPT at Goals House.

ClimateGPT empowers people on the ground by combining vast decentralised datasets ranging from satellites to citizen science input: emissions, earth observation, country, company, sector, and city data to reveal their implications for human systems. The tool maps cascading risks—for example, how a storm in Indonesia can trigger socio-economic and political consequences in rice-consuming countries—helping us understand and act in a world of interconnected vulnerabilities.

Why We Support This

This initiative perfectly aligns with the Internet Archive Europe’s mission of Universal Access to all knowledge. ClimateGPT demonstrates that when we combine planetary-scale datasets with open, decentralized technology, we empower citizens and governments to make better decisions. It is AI built for transparency and adaptation, not just automation.

Join Us in Davos

The launch will feature a C-level roundtable discussion on navigating a world of interconnected vulnerabilities.

  • 🌍 Event: ClimateGPT 3: Planetary Boundaries – New Modes of Action in a World Beyond 1.5°C
  • 🎙️ Host: Daniel Erasmus – Founder, ClimateGPT, Head of AI of the Internet Archive Europe and Full Member, Club of Rome
  • 🧠 Featuring:
    • Johan Rockström – Director, Potsdam Institute for Climate Impact Research
    • Will Marshall – CEO, Planet Labs
    • Sandrine Dixson-Decleve, Executive Chair, Earth4All
    • Simon Zadek – Co Founder Morphosis
  • 🗓️ When: Thursday, 22 January | 15:30 – 17:00 CET
  • 📍 Where: Goals House, Mattastrasse 25, Davos, Switzerland

We look forward to seeing how open data can help us navigate the challenges ahead.

Davos Event Spotlight: Launching ClimateGPT 3 and the Future of Public Good AI Read Post »

Scroll to Top