Large Language Writer

This project was exhibited at Falstaff Living Design Awards Newcomer, Ars Electronica, won the S+T+ARTS Prize, and Japan AI Creative Future award.

notion image

The interfaces between humans and machines shape our perceptions of technology. Current interfaces often present technology as if it operates like magic to ensure a smooth user experience. However, this approach can breed skepticism and distrust, categorizing these interfaces as "Dishonest Interfaces." This project explores what it means to bend open the loops of AI and design interfaces that engage people rather than obscure through oversimplification.

 

Dishonest Interface

We've noticed a concerning trend in design, especially in interface design, that emerges in "fully developed" products like baking ovens, stoves, and car interiors. Take the baking oven: its practical features are optimized, but companies still need to drive sales. Instead of relying on planned obsolescence, which could tarnish their reputations, they create demand by adding features promoted through advertising. This explains the rise of WiFi in home appliances. Traditional controls are replaced by terms like "4D-Heat" or "Fish Mode." While corporate greed is a simple explanation, the nature of design itself may offer deeper insights.

Contemporary technologies of networked computational things and artificial intelligence, as well as the data capitalism they have made possible, differ from the logic of industrial production. Not only that, they fundamentally challenge the conceptual space designers have created to cope with complexity. For instance, with runtime assembly of networked services, constant atomic updates, and agile development processes, the boundary between produc- tion and consumption is almost fully dismantled.
Elisa Giaccardi, Johan Redström, Technology and More-Than-Human Design, DesignIssues: Volume 36, Number 4 Autumn 2020
notion image
notion image

Solution? → Deus ex machina

The "Deus ex machina," a figure in Greek theatre lowered by a visible crane to resolve impossible dilemmas, offers a compelling analogy. The crane's visibility allowed the audience to trust the magic without feeling deceived. In design, how much should we "expose the crane" by making the mechanisms behind technology visible, fostering trust and transparency, and enabling users to engage without feeling manipulated?

notion image

Honest Design

With all these insights, ideas, and challenges, the question remains: What does it take to design honestly? It seems that honest design, achieved through seamful[2] interactions, could be a cornerstone for fostering a more positive and confident perspective on technological development.

Good design is honest.
design needs to be anticipatory, able to craft desirable relations between people and emerging technologies, and thus proactive in the associated processes of research and development.

We set out to design a prototype of an "honest interface," along with hardware components that adhere to the same philosophy. The objective was to develop methodologies for creating such an interface, using the specific example of interfacing with Large Language Models. Without further ado, let's delve into the five principles of honest design.

notion image
 

01 Principle: Choose an AI Problem

AI is currently a buzzword, and as its tools become more widespread, significant questions about their impact arise. While generative AI in image creation remains somewhat niche, systems like ChatGPT raise broader concerns. As students use AI for homework and others rely on it for drafting documents, the value of writing as a creative practice and a human method of documentation is increasingly being questioned.

notion image
notion image

The future of the Interface

While Big Tech favors chat-based interfaces, we should remain curious about alternative methods. Beyond the outdated command-and-control approach lies the Centaur Approach, theorized by Garri Kasparow, where humans and machines collaborate, each contributing their strengths. In 2021, my collaborator Leo Mühlfeld explored this by having children and generative AI co-design toys, demonstrating how the Centaur Approach can integrate humans into the generative process.

notion image
 

02 Principle: Break down the tech

When considering "honest" tech, we encounter a paradox: while interfaces like ChatGPT obscure their workings, the most transparent interaction might be Ishan Anand's "Spreadsheets is all you need", a fully functional GPT-2 in Excel. Though transparent, it lacks usability. As a designer, I wondered if there's a middle ground, a sweet spot between transparency and usability. To explore this, I examined three essential ideas…

notion image

Tokens and Relations

In a Generative Pretrained Transformer (GPT), tokens are the smallest text units, like words or subwords. The model learns relationships between tokens by analyzing large text datasets. The transformer's attention mechanism weights each token's importance in context, capturing patterns, syntax, and semantics. This allows the model to generate coherent and contextually relevant sequences by predicting the next token based on learned relationships.

Dataset

The dataset used to train an AI model shapes the quality, scope, and biases of its generated text. A diverse and extensive dataset enables more accurate and creative outputs, but it also embeds any inherent biases. After training, the model cannot learn from new data, meaning these limitations are fixed. However, a model can be fine-tuned on additional datasets to adapt or update its knowledge, allowing for some post-training adjustments.

Probability

In AI text generation, probability drives word prediction, with the model selecting the most likely next word based on context. This often results in text that feels "in the middle," as it favors safe, generic choices to maintain coherence. Interestingly, even improbable ideas like time travel or aliens can emerge when the model calculates them as the most probable among unlikely options, giving the illusion of creativity. However, the overall focus on high-probability words often makes AI text sound neutral and predictable.

For deep-divers, we recommend LLM Visualization by Brendan Bycroft

 

03 Principle: Open up loops

In the conventional approach, the process begins with a prompt and ends with a result, which can be refined through iteration and prompt engineering. However, this method largely removes the human element from the generative part of the process. This project introduces an alternative approach that advocates for a collaborative writing mode inspired by Generative Pre-trained Transformers (GPTs). By emphasizing mutual understanding and fostering "honest" interaction with AI, this approach actively involves humans in the generative process, placing them directly at its core. During the writing process, which occurs within a dynamic feedback loop, three key mechanisms are integrated: "Dataset," "Emphasis," and "Probability."

notion image

The Main Loop

Simply put, LLW introduces a new interactive feedback loop. The process starts with a prompt, then the AI extracts a theme and generates an initial sentence. The loop then continuously offers continuation sentences based on the user’s settings. We will delve into the interaction modes (Emphasis, Dataset, Probability) shortly. Once a sentence is selected, it becomes the new "last sentence," and the loop repeats until the user concludes the writing process. This approach immerses the writer directly in the generative process.

notion image
notion image

Emphasis → steer the feedback-loop

By incorporating AI's core principles into the user interface, LLW fosters an intrinsic understanding, leading to a trustworthy and honest relationship between users and their tools. The Emphasis strategy, based on token relationships, exemplifies this approach. In Emphasis mode, users can select words or strings by holding Shift, then assign a weight (1 to 3) before exiting the mode. The generated continuation sentences will reflect these user defined weights.

notion image

Dataset → Inform the loop

We have to recognize the fundamental differences between human and more-than human intelligences and embrace them. The non-obfuscation of this reality is an important step in designing honest for AI. Human traits like inspiring moments or things, a cute present they recently got or their current surroundings can be fed into the dataset to inspire the generative process in unconventional ways. LLW is equipped with a camera-module that allow writers to capture whatever they imagine to flow into the writing.

notion image
notion image

Probability → Break the loop

An essential problem identified by implementing feedback-loop based interactions into user-interfaces is the nature of the loop itself. To brake it is one of the main challenges. Up until know we leveraged the idea of „probable continuation“ in order to break up loops and let them continue somewhere else. Instead of predicting the most probable next word or sentence as AI does naturally, on the LLW keyboard students can use the probability button to ask the AI to give the least probable continuation of a story. This helps student to understand the nature of the technology and collaborate with it to playfully expand their personal perspective.

notion image
notion image

04 Principle: Design Honest Hardware

AI is a fast changing landscape with new developments everyday. To honestly represent this ever-changing technology and give it a materiality we have to acknowledge that a design poured into the clear boundaries of a cast unibody, no longer proves sustainable. We therefore propose exploration into modular hardware like voxel-based modular design that is able to gracefully react to its obsolescence.

notion image

Modular throughout

The first prototype of the Large Language Writer consists of a 3D-printed, modular, voxel-based body. This system uses three color codings: red for volume, yellow for connection, and monochrome for function. Using this system, three main assemblies were created: 1. Display Module: This module houses a 2K e-Ink display encased in custom-cut, powder-coated aluminum sheet-metal parts, along with the computer of the LLW. 2. Keyboard Module: This module includes a custom-made circuit board with a hardware design derived from the UI/UX principles declared above. 3. Camera Module: A separate module for the camera.

notion image
notion image

Keyboard layout

We decided to design and manufacture a circuit board that adheres to the sizes of the underlying grid. This circuit board includes three function keys, a Shift key and three Mode keys: "Emphasis," "Dataset," and "Probability," which toggles the modes described above. There is also a "Write" key, comparable to a Return key. The cursor is controlled using a rotary encoder. A Pro-Micro, which sits on the underside, runs QMK. The keyboard can be connected to any computer via USB-C.

notion image
notion image

05 Principle: Involve the Real World

After developing the prototype and running the first semi-stable version of the software, five individuals from diverse backgrounds—each with a strong connection to writing, whether out of necessity or creativity—were invited to participate. They received a brief tutorial on operating the machine and were allowed to choose their own writing prompts. The participants wrote for 45 minutes to an hour and were subsequently interviewed. This initial testing phase yielded valuable insights, informing potential directions for the continuation of the project.

Philipp

Philipp, a 16-year-old upper secondary school student, contributed to the test by writing a letter to the editor. While the content of his letter wasn't focused on AI, it provides insight into how students in his age group are engaging with AI in educational contexts.

You are reading the newspaper report 'Nur keine Spompanadeln' by Michael Omasta from the weekly newspaper Falter dated June 22, 2016, and respond with a letter to the editor.

Flora

Flora, who studied law and is currently working in legislation, offers an important perspective to the tests. Questions of responsibility, accuracy, and contextual awareness, particularly in relation to AI, can be explored. Flora’s insights help to underscore the need for careful consideration of these factors when developing laws and guidelines for emerging technologies. Her contribution adds insights regarding legal and ethical implications.

A lawyer advises her client based on AI-generated legal research. Describe possible issues in her work using a case example.

Helmut

Helmut, a 56-year-old author, contributes a seasoned perspective to the test. His experience as a writer brings a unique viewpoint on how language, storytelling, and perhaps even AI intersect. While the content of his contribution is not focused on AI, his background as an author enriches these tests by offering insights into how creative professionals engage with evolving technologies.

Newspaper article about low-threshold free cultural offerings in public spaces in Vienna.

Flora