AI · MBA
The bot rejected everything important without a deadline
The triage queue I started on 13 August holds eighty-four analysed items. Sixteen came back as "relevant now". Thirty-two were dropped as noise. The remaining thirty-six went onto a "come back to this" list. The rule that sorted them read: does this content change what I do within fourteen days, measured against my project priorities. The model applied it correctly every time. The rule was the bad part, because it welded two independent things into one verdict, weight and deadline.
My claim is that a tool built around your own attention does not increase the supply of it, it moves the cost of judgement outside you, and only then can you see how much of that cost there was and what shape it had. A decision hangs on this, easy to state and uncomfortable to answer: keep the prosthesis running, or switch it off. One measurement settles it, and almost nobody collects it.
Attention is a budget, and Simon supplied the unit
Herbert Simon set this out in 1971 in his lecture "Designing Organizations for an Information-Rich World". The sentence about a wealth of information creating a poverty of attention gets quoted to death. What Simon did immediately afterwards gets quoted far less. He said an allocation problem has to be posed properly, which needs a measure of the scarce resource that cannot be stretched at will, so he rejected Shannon's bit: bit capacity depends on how a message is encoded and is therefore not an invariant. What was left was the simplest unit available. Simon proposed the time a recipient spends on a message (Simon 1971, pp. 40–41).
This is where the line runs between a metaphor and a model. If attention is a budget, a link dropped into a queue is not a saved expense but a deferred one. You drop it because right now you have nothing to pay the judgement with. The bill waits.
Where the debt comes from when you defer judgement
Sophie Leroy showed in 2009 that moving from one task to another leaves attention residue: part of your cognitive resources stays with the previous task and does not work on the current one. The residue is largest when the previous task was cut off halfway (Leroy 2009). An item dropped into a queue is exactly such a half-cut task: you watched enough to call the thing potentially valuable, and too little to do anything with it.
More important for this argument is what Leroy published with Theresa Glomb in 2018 in Organization Science. Across four studies the residue shrank when the interrupted person got a moment to write down a plan for returning: where they were and what they would do on coming back. Not to finish the task. To record a decision about it (Leroy and Glomb 2018).
That is where the mechanism of a prosthesis comes from, and it is not doing your reading for you. A prosthesis turns an open loop into a recorded decision, and a recorded decision pays off a debt that finishing the reading would not have paid. The working criterion follows in one sentence: the system works when it makes a decision you do not have to make a second time.
A rule written into code shows what you cannot choose
Separating importance from urgency is not a discovery. Covey described it in 1989 as the second quadrant of the matrix: things that are important and not urgent, the ones that lose every single day and win the decade. The maxim underneath is older; Eisenhower delivered it in Evanston on 19 August 1954 and credited an unnamed "former college president", so treat today's named attributions with care. The point is that I knew the distinction and still wrote a rule that papered over it. I wrote it out of my own reflex, not out of the book.
The same reflex came out a second time, in another tool and in a harder form. I run a weekly radar that picks three things to do out of everything that arrived in a given week. Priority was impact times confidence, divided by effort, and effort entered as a linear divisor rising from 0.5 for half a day's work to 5 for work longer than three days. With that divisor the best possible large move scores 5 times 1 over 5, which is 1.0. Average small stuff scored 2.8 to 3.6. A large correction could never win, whatever it was worth. This was not a model error or a data error. The ceiling sat in arithmetic I had written myself.
Both rules removed the same class: things important, slow and without a deadline. The categories I invented turned out to be a map of my own gaps. I saw them only once I had to write them in a form a machine executes, because a machine makes no exception for a thing whose author privately knew it mattered.
One design decision from that period held up completely. A rejecting verdict still writes one line to the wiki with a link and a reason. The rejection is therefore reversible and countable, and I can check after the fact what the system threw out and why. A bad system is one that deletes quietly.
You can call this procrastination and sometimes you would be right
The strongest objection to this thesis is not "tools do not help". It runs: building a system for reading is a form of avoiding reading, only with a better justification. The objection lands and it narrows the conclusion, so I will not leave it in a footnote.
A prosthesis defends itself under two conditions at once. First: the stream is larger than the attention budget. Eighty-four items in eighteen days is not quite five a day, and for material that has to be watched and judged that is more than I have to pay with. Second condition: the cost of keeping the prosthesis is falling. Eighteen days of triage cost 3.14 dollars in model charges, roughly 17 cents a day, and it fits inside a hard cap I set for myself. A radar run cost 17 cents, the same as one day of triage.
Honestly: those figures speak to money, which was small from the start, rather than to the cost of my attention, which is the thing actually in dispute. I do not yet measure upkeep in hours, and after eighteen days I have no right to claim the curve is falling. If the cost of upkeep is rising then it is a hobby, and a hobby is allowed as long as nobody calls it infrastructure.
I split the axes and fell into the same hole one floor up
In the radar the axes are no longer welded. A data point gets a class, impact from 1 to 5, effort, confidence from 0 to 1 and a gate flag for product ideas, each on its own. No deadline is written into any of those axes. That looks like a repair and at the level of classification it is one.
What settles the outcome is not the split but the formula that welds the axes back together. It contains a rule of mine: a product idea whose only source is content from the internet gets confidence of at most 0.4, and the rule is a sensible one, because a single creator's video does not justify changing a product. In the first real run the effect came out like this: two product ideas with impact 4 got a priority of 0.36. They landed at the bottom of the backlog, below small stuff with impact 2. The three proposals were two small things and one medium, the last of them only because I had added a separate rule guaranteeing one slot to something bigger.
The class the first version rejected outright, the second version no longer rejects. It pushes it to the end of the list with a multiplier I judged reasonable myself. The scope of the thesis therefore has to narrow: separating the axes is a necessary condition and not a sufficient one. What survives is decided by the recombination rule and by who is obliged to look at the result.
The obligation is in worse shape than the arithmetic. The radar has had one run so far, on 24 August, and it produced fourteen proposals. Six days later not one of them carries a recorded decision, and the trust ledger meant to hold my approvals and rejections is empty. Thirty-six items on the "come back to this" list plus fourteen proposals without a decision makes fifty deferred judgements. I started with one queue of unjudged things. I now have two, neatly sorted.
I wrote the condition for my own failure into the spec before I saw these numbers: if I accept fewer than 30 percent of proposals after six runs, the kill criteria apply. I have one run and zero acceptances, so the counter has not started. That is the one measurement nobody collects, and the only one that will separate a prosthesis from an expensive hobby: what share of deferred judgements ever received a recorded decision.
References
- primaryHerbert A. Simon, "Designing Organizations for an Information-Rich World", in M. Greenberger (ed.), Computers, Communications, and the Public Interest (The Johns Hopkins Press, Baltimore, 1971) — the quoted passage and the argument that an allocation problem needs a measure of the scarce resource, with the bit rejected and time proposed instead, pp. 40–41. Read in the original scan.
- primarySophie Leroy, "Why is it so hard to do my work? The challenge of attention residue when switching between work tasks", Organizational Behavior and Human Decision Processes 109(2), 2009, 168–181 — attention residue and its dependence on how the previous task was left.
- primarySophie Leroy and Theresa M. Glomb, "Tasks Interrupted: How Anticipating Time Pressure on Resumption of an Interrupted Task Causes Attention Residue and Low Performance on Interrupting Tasks and How a 'Ready-to-Resume' Plan Mitigates the Effects", Organization Science 29(3), 2018, 380–397 — four studies; a written plan for returning reduces residue without the interrupted task being finished.
- primaryStephen R. Covey, The 7 Habits of Highly Effective People (1989) — the important–urgent matrix and its second quadrant.
- secondaryQuote Investigator, "What Is Important Is Seldom Urgent and What Is Urgent Is Seldom Important" (2014) — dates the maxim to Eisenhower's address in Evanston on 19 August 1954 and records that he credited an unnamed former college president; the widespread named attributions do not hold.
Figures on the triage queue, its costs and the radar run come from the database and repository of my own tools, as of 30 August 2026.
Bot odrzucał wszystko, co ważne bez terminu
W kolejce triage'u, którą uruchomiłem 13 sierpnia, leżą osiemdziesiąt cztery przeanalizowane pozycje. Szesnaście dostało werdykt „istotne teraz". Trzydzieści dwie odpadły jako szum. Pozostałe trzydzieści sześć trafiło na listę „wrócę do tego". Reguła rozstrzygająca brzmiała tak: czy ta treść zmienia moje działanie w ciągu czternastu dni względem priorytetów projektów. Model wykonywał ją poprawnie za każdym razem. Zła była reguła, bo skleiła w jeden werdykt dwie niezależne rzeczy, wagę i termin.
Twierdzę, że narzędzie zbudowane wokół własnej uwagi nie zwiększa jej ilości, tylko wynosi koszt oceny na zewnątrz, i dopiero wtedy widać, ile tego kosztu było i jaki miał kształt. Zależy od tego decyzja łatwa do zadania i niewygodna w odpowiedzi: czy taką protezę utrzymywać dalej, czy wyłączyć. Rozstrzyga ją jedna miara, której prawie nikt nie zbiera.
Uwaga jest budżetem, a Simon podał jednostkę
Herbert Simon opisał to w 1971 roku w wykładzie „Designing Organizations for an Information-Rich World". Zdanie o tym, że bogactwo informacji tworzy ubóstwo uwagi, cytuje się do znudzenia. Rzadziej cytuje się to, co Simon zrobił zaraz potem. Powiedział, że problem alokacji trzeba postawić poprawnie, a do tego potrzeba miary rzadkiego zasobu, której nie da się dowolnie rozciągnąć, więc odrzucił bit Shannona: pojemność bitowa zależy od kodowania komunikatu i nie jest niezmiennikiem. Została jednostka najprostsza z możliwych. Simon zaproponował czas, jaki odbiorca spędza na komunikacie (Simon 1971, s. 40–41).
Tu przebiega granica między metaforą a modelem. Jeżeli uwaga jest budżetem, to link wrzucony do kolejki nie jest wydatkiem zaoszczędzonym, tylko odroczonym. Wrzucasz go, bo w tej chwili nie masz z czego zapłacić za ocenę. Rachunek czeka dalej.
Skąd bierze się dług, kiedy odkładasz ocenę
Sophie Leroy pokazała w 2009 roku, że przejście od jednego zadania do drugiego zostawia osad uwagi: część zasobów poznawczych zostaje przy zadaniu poprzednim i nie pracuje na bieżącym. Osad jest największy wtedy, gdy poprzednie zadanie zostało przerwane w połowie (Leroy 2009). Wrzut do kolejki jest właśnie takim zadaniem przerwanym w połowie: obejrzałeś dość, żeby uznać rzecz za potencjalnie wartościową, i za mało, żeby cokolwiek z nią zrobić.
Dla tego wywodu ważniejsze jest jednak to, co Leroy opublikowała z Theresą Glomb w 2018 roku w „Organization Science". W czterech badaniach osad malał, kiedy przerywana osoba dostawała chwilę na spisanie planu powrotu: gdzie jest i co zrobi, gdy wróci. Nie na dokończenie zadania. Na zapisanie rozstrzygnięcia o nim (Leroy i Glomb 2018).
Stąd bierze się mechanizm protezy, i nie jest nim wyręczanie cię w czytaniu. Proteza zamienia otwartą pętlę w zapisane rozstrzygnięcie, a zapisane rozstrzygnięcie spłaca dług, którego samo dokończenie lektury by nie spłaciło. Kryterium działania da się z tego wyprowadzić w jednym zdaniu: system pracuje wtedy, gdy podejmuje decyzję, której ty nie musisz podejmować po raz drugi.
Reguła zapisana w kodzie pokazuje, czego nie umiesz wybrać
Rozdzielenie ważności i pilności nie jest odkryciem. Covey opisał je w 1989 roku jako drugą ćwiartkę macierzy: sprawy ważne i niepilne, czyli te, które przegrywają każdy dzień z osobna i wygrywają dekadę. Maksyma, na której to stoi, jest starsza; Eisenhower wygłosił ją w Evanston 19 sierpnia 1954 roku i przypisał nieustalonemu „byłemu rektorowi", więc krążące dziś atrybucje imienne traktuj ostrożnie. Rzecz w tym, że znałem ten podział, a i tak napisałem regułę, która go zaklejała. Pisałem ją z własnego odruchu, nie z książki.
Ten sam odruch wyszedł drugi raz, w innym narzędziu i w twardszej postaci. Buduję tygodniowy radar, który wybiera trzy rzeczy do zrobienia z wszystkiego, co w danym tygodniu wpadło. Priorytet liczył się jako wpływ razy pewność, podzielone przez pracochłonność, a pracochłonność wchodziła dzielnikiem liniowym, który rósł od 0,5 przy robocie na pół dnia do 5 przy robocie dłuższej niż trzy dni. Przy takim dzielniku najlepszy możliwy duży ruch dostaje 5 razy 1 przez 5, czyli 1,0. Przeciętna drobnica liczyła 2,8 do 3,6. Duża korekta nie mogła wygrać nigdy, niezależnie od tego, ile była warta. Nie był to błąd modelu ani błąd danych. Sufit siedział w arytmetyce, którą sam wpisałem.
Obie reguły usuwały tę samą klasę: rzeczy ważne, powolne i bez terminu. Kategorie, które wymyśliłem, okazały się mapą moich luk. Zobaczyłem je dopiero wtedy, gdy musiałem zapisać je w postaci, którą wykonuje maszyna, bo maszyna nie robi wyjątku dla rzeczy, o której autor reguły w duchu wiedział, że jest ważna.
Jedna decyzja projektowa z tego okresu obroniła się w całości. Werdykt odrzucający nadal zapisuje do wiki jedną linię z odnośnikiem i powodem. Odrzucenie jest więc odwracalne i policzalne, a ja mogę po fakcie sprawdzić, co system wyrzucił i dlaczego. Zły system to ten, który kasuje po cichu.
Można to nazwać prokrastynacją i czasem będzie to trafne
Najsilniejszy zarzut wobec tej tezy nie brzmi „narzędzia nie pomagają". Brzmi tak: budowanie systemu do czytania jest formą unikania czytania, tyle że z lepszym uzasadnieniem. Zarzut jest trafny i zawęża wniosek, więc nie zostawię go w przypisie.
Proteza broni się przy dwóch warunkach naraz. Pierwszy: strumień jest większy od budżetu uwagi. Osiemdziesiąt cztery pozycje w osiemnaście dni to niecałe pięć dziennie, a dla materiału, który wymaga obejrzenia i oceny, jest to więcej, niż mam z czego zapłacić. Drugi warunek: koszt utrzymania protezy spada. Osiemnaście dni pracy triage'u kosztowało 3,14 dolara w rachunku za model, czyli około 17 centów dziennie, i mieści się w twardym limicie, który sam sobie nałożyłem. Przebieg radaru kosztował 17 centów, czyli tyle, co jeden dzień triage'u.
Uczciwie: te liczby mówią o koszcie pieniężnym, który był mały od początku, a nie o koszcie mojej uwagi, który jest właściwym przedmiotem sporu. Utrzymania nie mierzę jeszcze w godzinach i po osiemnastu dniach nie mam prawa twierdzić, że krzywa opada. Jeśli koszt utrzymania rośnie, to jest to hobby, a hobby wolno mieć, o ile nie nazywa się go infrastrukturą.
Rozdzieliłem osie i wpadłem w ten sam dół piętro wyżej
W radarze osie już nie są sklejone. Punkt dostaje osobno klasę, wpływ od 1 do 5, pracochłonność, pewność od 0 do 1 oraz znacznik bramki dla pomysłów produktowych. Termin nie jest wpisany w żadną z tych osi. Wygląda to na naprawę i na poziomie klasyfikacji nią jest.
Rozstrzyga jednak nie rozdzielenie, tylko wzór, który osie z powrotem skleja. Mam w nim regułę: pomysł produktowy, którego jedynym źródłem jest treść z internetu, dostaje pewność najwyżej 0,4, i jest to reguła sensowna, bo rolka jednego twórcy nie uzasadnia zmiany produktu. Skutek w pierwszym prawdziwym przebiegu wyszedł taki: dwa pomysły produktowe z wpływem 4 dostały priorytet 0,36. Wylądowały na dnie backlogu, pod drobnicą z wpływem 2. Do trójki propozycji weszły dwie rzeczy małe i jedna średnia, ta ostatnia wyłącznie dlatego, że dopisałem osobną regułę gwarantującą jeden slot czemuś większemu.
Klasy, którą pierwsza wersja odrzucała wprost, druga wersja już nie odrzuca. Spycha ją na koniec listy mnożnikiem, który sam uznałem za rozsądny. Zakres tezy trzeba więc zawęzić: rozdzielenie osi jest warunkiem koniecznym i nie jest wystarczające. O tym, co przeżyje, decyduje reguła składania i to, kto ma obowiązek spojrzeć na wynik.
Z obowiązkiem jest gorzej niż z arytmetyką. Radar przeszedł jak dotąd jeden przebieg, 24 sierpnia, i wyprodukował czternaście propozycji. Sześć dni później żadna z nich nie ma zapisanej decyzji, a rejestr zaufania, do którego miały trafiać moje akceptacje i odrzucenia, jest pusty. Trzydzieści sześć pozycji na liście „wrócę do tego" plus czternaście propozycji bez decyzji daje pięćdziesiąt odroczonych ocen. Zaczynałem od jednej kolejki nieocenionych rzeczy. Mam teraz dwie, za to obie ładnie posortowane.
Warunek własnej porażki zapisałem w specyfikacji, zanim zobaczyłem te liczby: jeśli po sześciu przebiegach akceptuję mniej niż 30 procent propozycji, wchodzą kryteria wyłączenia. Przebieg mam jeden, akceptacji zero, więc licznik jeszcze nie ruszył. To jest ta jedna miara, o której nikt nie zbiera danych, i jedyna, która odróżni protezę od kosztownego hobby: jaki odsetek odroczonych ocen doczekał się kiedykolwiek zapisanej decyzji.
Bibliografia
- primaryHerbert A. Simon, „Designing Organizations for an Information-Rich World", w: M. Greenberger (red.), Computers, Communications, and the Public Interest (The Johns Hopkins Press, Baltimore, 1971) — cytowany fragment oraz argument, że problem alokacji potrzebuje miary rzadkiego zasobu, z odrzuceniem bitu i propozycją czasu, s. 40–41. Czytane ze skanu pierwodruku. Przekład autora eseju.
- primarySophie Leroy, „Why is it so hard to do my work? The challenge of attention residue when switching between work tasks", Organizational Behavior and Human Decision Processes 109(2), 2009, 168–181 — osad uwagi i jego zależność od tego, w jakim stanie zostało poprzednie zadanie.
- primarySophie Leroy i Theresa M. Glomb, „Tasks Interrupted: How Anticipating Time Pressure on Resumption of an Interrupted Task Causes Attention Residue and Low Performance on Interrupting Tasks and How a »Ready-to-Resume« Plan Mitigates the Effects", Organization Science 29(3), 2018, 380–397 — cztery badania; spisany plan powrotu zmniejsza osad bez dokończenia przerwanego zadania.
- primaryStephen R. Covey, The 7 Habits of Highly Effective People (1989), wyd. pol. 7 nawyków skutecznego działania — macierz ważne–pilne i jej druga ćwiartka.
- secondaryQuote Investigator, „What Is Important Is Seldom Urgent and What Is Urgent Is Seldom Important" (2014) — datowanie maksymy na wystąpienie Eisenhowera w Evanston 19 sierpnia 1954 roku i ustalenie, że mówca przypisał ją nieustalonemu byłemu rektorowi; rozpowszechnione atrybucje imienne nie mają pokrycia.
Dane liczbowe o kolejce triage'u, kosztach i przebiegu radaru pochodzą z bazy i repozytorium moich własnych narzędzi (stan na 30 sierpnia 2026).