<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
<title>Tyler Mayberry: Notes</title>
<link>https://tylermayberry.dev/notes</link>
<atom:link href="https://tylermayberry.dev/notes/feed.xml" rel="self" type="application/rss+xml"/>
<description>Longer posts and threads I&#x27;ve written on X, kept here in one place.</description>
<language>en</language>
<lastBuildDate>Tue, 29 Sep 2026 12:00:00 +0000</lastBuildDate>
<item><title>ibara: a computer you own for your agents</title><link>https://tylermayberry.dev/notes/ibara-for-omarchy</link><guid isPermaLink="true">https://tylermayberry.dev/notes/ibara-for-omarchy</guid><pubDate>Tue, 29 Sep 2026 12:00:00 +0000</pubDate><description>&lt;p&gt;OpenAI, Meta and xAI all just gave their agents a computer.&lt;/p&gt;&lt;p&gt;But it&amp;#x27;s only available in their cloud, for a VM that&amp;#x27;s most likely less powerful than the potato in your closet.&lt;/p&gt;&lt;p&gt;ibara for Omarchy gives your agent a PC you own.&lt;/p&gt;&lt;p&gt;Free and Open Source&lt;/p&gt;&lt;p&gt;ibara gives your agents computer use across the Omarchy machines you own.&lt;/p&gt;&lt;p&gt;The agent gets a spare&amp;#x27;s desktop, browser and apps, and you watch, take control and hand back.&lt;/p&gt;&lt;p&gt;Works with any agent. Get connected with a copy/paste prompt.&lt;/p&gt;&lt;p&gt;Checked, not just claimed.&lt;/p&gt;&lt;p&gt;Your agent checks the work it made as a real user on a real machine.&lt;/p&gt;&lt;p&gt;ibara tracks every task leaving a complete log of everything the agent does for review.&lt;/p&gt;&lt;p&gt;Your hardware. Your connection.&lt;/p&gt;&lt;p&gt;Keep all your logins and access, or restrict it in any way that you want.&lt;/p&gt;&lt;p&gt;Computer use with a real home ip and no expensive cloud bill.&lt;/p&gt;&lt;p&gt;Basic hardware is plenty, ibara idles at 17.8 MiB.&lt;/p&gt;&lt;p&gt;You choose who&amp;#x27;s driving.&lt;/p&gt;&lt;p&gt;Watch the spare from the console, take control when you want, hand it back when you&amp;#x27;re done.&lt;/p&gt;&lt;p&gt;One driver at a time. By default ibara asks before an agent sends, spends or deletes, and always asks before changing access.&lt;/p&gt;&lt;p&gt;(you can also &amp;#x27;dangerously&amp;#x27; allow complete access)&lt;/p&gt;&lt;p&gt;132 of 157 runs passed the benchmark&amp;#x27;s checks.&lt;/p&gt;&lt;p&gt;Claude Code with Opus 5.5 and Codex with GPT-6-Astra each passed 22 of 23.&lt;/p&gt;&lt;p&gt;Even @opencode on the free plan passed 14 of 22 using big-pickle.&lt;/p&gt;&lt;p&gt;Big thanks to @trycua here.&lt;/p&gt;&lt;p&gt;Grok Bot, Dots and Muse give an agent its own computer in the vendor&amp;#x27;s cloud.&lt;/p&gt;&lt;p&gt;OpenClaw and Hermes run on your hardware but are the agent.&lt;/p&gt;&lt;p&gt;ibara is a separate computer you own that any agent can use.&lt;/p&gt;&lt;p&gt;Two Omarchy 4 computers, Tailscale on both, and one agent (even free @opencode works).&lt;/p&gt;&lt;p&gt;Install on each, add the spare from the console, copy the connect prompt to your agent and give it a first task.&lt;/p&gt;&lt;p&gt;I made it as plug-and-play as possible.&lt;/p&gt;&lt;p&gt;Start with one spare. Connect the rest.&lt;/p&gt;&lt;p&gt;Lightweight, built for basic hardware, free and open source.&lt;/p&gt;&lt;p&gt;Install ibara, or take over the demo in your browser first.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://ibara.app&quot;&gt;https://ibara.app&lt;/a&gt;&lt;/p&gt;</description></item>
<item><title>A theory about OpenAI, Anthropic and compute</title><link>https://tylermayberry.dev/notes/openai-anthropic-compute</link><guid isPermaLink="true">https://tylermayberry.dev/notes/openai-anthropic-compute</guid><pubDate>Sat, 26 Sep 2026 12:00:00 +0000</pubDate><description>&lt;p&gt;I have a theory about what&amp;#x27;s happening with openai and anthropic.&lt;/p&gt;&lt;p&gt;Anthropic was compute constrained for so long because everyone was using their models for work and coding.&lt;/p&gt;&lt;p&gt;This gave openai lots of compute headroom to develop, research, and hand out that compute to their users.&lt;/p&gt;&lt;p&gt;Once openai grabbed such a major market share with better models and usage, they become the ones who were compute constrained.&lt;/p&gt;&lt;p&gt;This gave anthropic the wiggle room they needed to research and develop Opus 5.5 with crazy limits.&lt;/p&gt;&lt;p&gt;Now I&amp;#x27;m thinking we might see the pendulum swing again in a month or so.&lt;/p&gt;&lt;p&gt;Or I could be blowing hot air, who knows?&lt;/p&gt;&lt;p&gt;Might be coming back to this tweet in a month.&lt;/p&gt;</description></item>
<item><title>Giving old computers to agents</title><link>https://tylermayberry.dev/notes/old-computers-for-agents</link><guid isPermaLink="true">https://tylermayberry.dev/notes/old-computers-for-agents</guid><pubDate>Fri, 25 Sep 2026 12:00:00 +0000</pubDate><description>&lt;p&gt;Give your agents instant computer use access to any Omarchy machine you own.&lt;/p&gt;&lt;p&gt;I&amp;#x27;m nearing the first release of ibara: an Omarchy-native computer use layer that turns your hardware into agent infrastructure.&lt;/p&gt;&lt;p&gt;Today I&amp;#x27;m integrating the awesome @trycua driver for computer use and polishing the UI.&lt;/p&gt;&lt;p&gt;Soon you can give every potato pc you have to your agents so they have real BARE-METAL computers to work on.&lt;/p&gt;&lt;p&gt;Remember that old Dell Optiplex tower you shoved in the attic and forgot about?&lt;/p&gt;&lt;p&gt;Install Omarchy with its insane installation speeds -&amp;gt; install the ibara plugin -&amp;gt; connect to your Omarchy fleet.&lt;/p&gt;&lt;p&gt;Boom. Your forgotten antique has a new reason to live again in less than 5 minutes. (okay maybe 10-15)&lt;/p&gt;</description></item>
<item><title>Building a personal app that agents can actually use</title><link>https://tylermayberry.dev/notes/personal-app-agents-can-use</link><guid isPermaLink="true">https://tylermayberry.dev/notes/personal-app-agents-can-use</guid><pubDate>Sun, 13 Sep 2026 12:00:00 +0000</pubDate><description>&lt;p&gt;Super App started as a health app with a place to write notes. About a week ago, I made it for my phone. It had Fitbit data and sections for food, movement, body measurements, and other health records. Pretty straightforward. Building it with AI agents, I added capabilities one at a time, and each addition changed what the underlying system needed to do.&lt;/p&gt;&lt;p&gt;A completely custom app, just for me, to track what I do and keep my data together. When I have something for my agent to record, I open it and leave a note. There are more capabilities now (Bluetooth scale readings, a personal wiki, agent updates, offline data and an export), but most of my use still starts there. A note from the app on my home screen.&lt;/p&gt;&lt;p&gt;Editing a note keeps the original, so I can clarify something later without losing what I first wrote. I want that freedom when I capture things. A meal, a workout, an observation, an idea for the app (all in the same note if that&amp;#x27;s how they come out); the agent can organize them afterwards.&lt;/p&gt;&lt;p&gt;Having to choose between 4 forms first would get in the way. As the app grew beyond health, I gave it its own project and added camera access, image attachments and editing. More ways to get something down while it&amp;#x27;s still in my head.&lt;/p&gt;&lt;p&gt;As more features arrived, I kept asking for compact controls and fewer labels. More room for the content. Some of the notes themselves became requests for improvements, because I was already using the app to capture what needed doing. I asked for pictures and reported that new notes weren&amp;#x27;t scrolling into view. When I asked an agent to implement those changes, it had the original requests.&lt;/p&gt;&lt;p&gt;The screens got a consistent look too: square edges, the dark Tokyo Night palette and shared imagery. Search and filters open when I need them. I wanted the app to do more while staying easy to open and use.&lt;/p&gt;&lt;p&gt;Body history and interactive charts, with measured weight and impedance kept separate from estimated body-composition values. The distinction matters when I look at a reading. After we connected the scale through Bluetooth, I stepped on it and the reading arrived in the app. We checked that it saved. No manufacturer&amp;#x27;s app, no extra account, no computer needed for a normal weigh-in.&lt;/p&gt;&lt;p&gt;Food records gained researched calorie and nutrient estimates (with their sources and assumptions attached), alongside Fitbit records. Notes give those numbers context; what I ate, how a workout felt and what was happening that day can help explain them. It already makes a day easier to review. I want to use that context across a longer stretch of time too.&lt;/p&gt;&lt;p&gt;When a sleep score was missing, the Fitbit code had been substituting deep-sleep minutes. Different measures, same field. We fixed the processing, and corrections to notes needed care too. A summary using an older version can show that it needs another look.&lt;/p&gt;&lt;p&gt;An agent could have explained that wrong number convincingly (that&amp;#x27;s what bothers me about the bug). I want to trace an answer back to its source, particularly once we&amp;#x27;re looking across months of records. An explanation can sound reasonable while starting from something that was never a sleep score.&lt;/p&gt;&lt;p&gt;Keep missing food logs missing, keep the assumptions attached to estimates, and avoid carrying mistakes forward. Collecting more data is only useful if we can trust what the records mean.&lt;/p&gt;&lt;p&gt;Saved progress gives another agent somewhere to resume when a session stops. We added records of completed changes too, so it can see what actually finished. Before an agent saves a review, it checks whether the notes changed while it was working. Then it reads the saved result back.&lt;/p&gt;&lt;p&gt;I have a phone interface, and the agents have direct tools for the records. Their work stays in the app after the conversation ends; procedures and results can be reused while the information behind them is still valid. When that information changes, the conclusions need checking.&lt;/p&gt;&lt;p&gt;I started treating agent use as part of the app&amp;#x27;s design. That&amp;#x27;s what I mean by agent-first. A notes review starts with finding new entries, reading the relevant material and preparing an update. The tools give the agent a way to carry that work through to a saved result.&lt;/p&gt;&lt;p&gt;To capture notes, inspect records and see more of the agents&amp;#x27; work in one place, I brought their knowledge and messages into the app. I want to see what an agent says it finished; I also want to check the update is actually there.&lt;/p&gt;&lt;p&gt;The GBrain browser gives me access to the shared knowledge system my agents use, and the Grokbot inbox holds their messages. Connected, with different jobs. A message saying the work is done doesn&amp;#x27;t establish that the result was saved.&lt;/p&gt;&lt;p&gt;One note, several things it describes. One day, several notes. A finding with sources from different places. I wanted a wiki so I could follow those connections, and the agents could find the relevant pieces too.&lt;/p&gt;&lt;p&gt;It grew to more than 4,000 generated pages covering notes, days, datasets and findings. Views of records already there. I hadn&amp;#x27;t written thousands of articles.&lt;/p&gt;&lt;p&gt;For browsing, we separated opening saved content from preparing an export, which had been slowing the wiki down. Now the app opens the page first and checks for changes in the background. From a day to its notes and measurements. From a finding back to the information supporting it.&lt;/p&gt;&lt;p&gt;To find correlations I&amp;#x27;d otherwise miss, I want a few months of connected history to work with. That&amp;#x27;s a big part of what I&amp;#x27;m building this for.&lt;/p&gt;&lt;p&gt;An unusually difficult workout is one example to investigate. What did the sleep leading up to it look like? Does that combination keep appearing? We&amp;#x27;d need wearable records, workout notes and a clear view of which days had both; a note entered late needs to link to the day it describes, while retaining when I actually wrote it. Otherwise we could be comparing the wrong days.&lt;/p&gt;&lt;p&gt;Bring the relevant evidence together through the wiki, then give me an answer I can inspect. I want to understand why the agent thinks a relationship deserves a closer look, with the records behind that judgment.&lt;/p&gt;&lt;p&gt;For any comparison, repeatable calculations and a clear account of the records included or missing. Those are things I want to see. A correlation gives us something to investigate. Working out what caused what needs more evidence.&lt;/p&gt;&lt;p&gt;Something really interesting, hopefully, over the next few months. Even a promising idea that doesn&amp;#x27;t hold up would be useful. The record needs time to grow before we can see what the comparisons are worth.&lt;/p&gt;&lt;p&gt;I want the exceptions too: the unusual week that explains most of a result, or another part of my routine changing at the same time. Show me those details. They can change what a pattern means.&lt;/p&gt;&lt;p&gt;Export readable pages, structured records and original evidence. That&amp;#x27;s available now, with GBrain and Grokbot kept outside the personal archive.&lt;/p&gt;&lt;p&gt;Without a connection, I still want access to my data. The phone keeps a copy of Fitbit data published by Halla, my server, and retains the last complete copy if Halla goes away.&lt;/p&gt;&lt;p&gt;To finish the archive, we still need to fill gaps in supporting evidence and do more work on larger archives and recovery. The app reports those gaps. I want to know what I actually have. Downloading a file shouldn&amp;#x27;t make it look as though every part of the record is there.&lt;/p&gt;&lt;p&gt;Capture in Home, review records in Health, find knowledge and agent updates in Brain. The wiki, offline copy and export in Data. Those are the 4 main areas now.&lt;/p&gt;&lt;p&gt;The original notes are still there, and everyday use still starts with opening the app, leaving a note and having the agent help organize it. Broader unattended analysis is still to build.&lt;/p&gt;&lt;p&gt;I made a custom app to track what I do. It now connects those records and gives agents a way to work with them. Over the next few months, I want to see which patterns are actually worth paying attention to. A thought left on my phone can stay useful long after I&amp;#x27;ve forgotten writing it.&lt;/p&gt;</description></item>
<item><title>You can&#x27;t just build from the top down</title><link>https://tylermayberry.dev/notes/not-just-top-down</link><guid isPermaLink="true">https://tylermayberry.dev/notes/not-just-top-down</guid><pubDate>Sun, 09 Aug 2026 12:00:00 +0000</pubDate><description>&lt;p&gt;I&amp;#x27;m starting to see a split happening with the software engineers, AI builders, founders, whatever you want to call them.&lt;/p&gt;&lt;p&gt;There&amp;#x27;s a big group that&amp;#x27;s just trying to automate everything from the top down.&lt;/p&gt;&lt;p&gt;And then there&amp;#x27;s people who are trying to get down to lower levels with their agent and actually direct it and tell it what to do instead of handing it off to another bigger orchestrator agent.&lt;/p&gt;&lt;p&gt;It&amp;#x27;s really cool and sexy that the orchestrator can actually orchestrate subagents now and build a usable product. And it&amp;#x27;s a lot of fun to use.&lt;/p&gt;&lt;p&gt;But every time I&amp;#x27;ve done that, I get a much worse result compared to if I actually use a much faster, less intelligent model and give it clear instructions and iterate over and over again over time and have much more human input and human in the loop interactions.&lt;/p&gt;&lt;p&gt;And frankly, I don&amp;#x27;t know if you&amp;#x27;ll ever actually be able to get away from this if you want to stand out.&lt;/p&gt;&lt;p&gt;While sure, the orchestrator method will always work and will always create products that function and are useful. I still use it regularly to build internal tools for myself.&lt;/p&gt;&lt;p&gt;But if you actually want that polished look and feel with real human intent, you&amp;#x27;re going to have to get into the lower levels.&lt;/p&gt;&lt;p&gt;You can&amp;#x27;t just build from the top down.&lt;/p&gt;</description></item>
<item><title>Milkbench</title><link>https://tylermayberry.dev/notes/milkbench</link><guid isPermaLink="true">https://tylermayberry.dev/notes/milkbench</guid><pubDate>Tue, 21 Jul 2026 12:00:00 +0000</pubDate><description>&lt;p&gt;All GPT-5.6 models and reasoning levels, tested for taste.&lt;/p&gt;&lt;p&gt;I created a prompt to one-shot a website for milk as if it was some sort of new revolutionary tech product. But it’s just milk.&lt;/p&gt;&lt;p&gt;We&amp;#x27;re tasting the milk.&lt;/p&gt;&lt;p&gt;Then I took that one prompt and, from a clean session on a clean slate, gave it to Codex across all different models and reasoning levels:&lt;/p&gt;&lt;p&gt;Luna: Light, Medium, High, Extra High, Max.&lt;/p&gt;&lt;p&gt;Terra: Light, Medium, High, Extra High, Max, Ultra.&lt;/p&gt;&lt;p&gt;Sol: Light, Medium, High, Extra High, Max, Ultra.&lt;/p&gt;&lt;p&gt;17 sites. Same prompt. Same product.&lt;/p&gt;&lt;p&gt;The results are pretty expected, but there were some interesting findings.&lt;/p&gt;&lt;p&gt;Sol Ultra is 32% of the total cost. One out of 17 sites was a third of the cost. Sol as a family is 73% of the total API cost, even though it only has 6 of the 17 sites.&lt;/p&gt;&lt;p&gt;Meanwhile, all five Luna runs together cost $3.97. Luna High and Extra High cost less than a dollar and take about 15 minutes. The value there is pretty outstanding.&lt;/p&gt;&lt;p&gt;Some things didn’t correlate at all. More thinking doesn’t mean a better site, turning up the effort doesn’t always mean it creates more work, and the size of the sites really didn’t have any relevance at all.&lt;/p&gt;&lt;p&gt;The heaviest site is 45x larger than the lightest: Terra Ultra is 10.9 MB, while Luna Max is only 0.24 MB. Same prompt, same product, but widely different shipping weight. But Sol consistently shipped lightweight sites with a premium feel, and Terra always shipped heavy sites.&lt;/p&gt;&lt;p&gt;Sol Ultra also made up about 24% of all token spend, using 1.33 million tokens and taking 70 minutes. Across all the runs, the lowest was Terra Medium at 59k tokens, while the highest was Sol Ultra at 1.33 million.&lt;/p&gt;&lt;p&gt;Terra Medium is a real outlier. It’s a higher effort than Terra Low, but it took less time, used fewer tokens, and was significantly cheaper. It also made a pretty terrible site.&lt;/p&gt;&lt;p&gt;Sol’s effort ladder isn’t a clean scale either. Medium is slower and has more token spend than High, while High is slower than Extra High.&lt;/p&gt;&lt;p&gt;The carton is the real skill check, and pretty much all the models fail. Sol Ultra does the best here, but it also spent 22 times more tokens than Terra Medium.&lt;/p&gt;&lt;p&gt;The hero section is another regular problem. The carton on top of the headline shows up across multiple models and effort levels. Sol Extra High is the only one that actually put the carton behind the headline, but then it put white text on a white carton.&lt;/p&gt;&lt;p&gt;Luna Light really outperformed its price tag, while Sol Light already feels very premium at a fraction of the cost. Sol Ultra is probably the best one-shot, but it’s also the worst value.&lt;/p&gt;&lt;p&gt;The family breakdown is pretty interesting.&lt;/p&gt;&lt;p&gt;Luna has the best effort-to-taste until you hit Max. Terra is the chaotic middle child. Sol is the premium one, but it comes with 73% of the total API cost.&lt;/p&gt;&lt;p&gt;Turning the reasoning up is not a quality slider. There’s a lot of variability.&lt;/p&gt;&lt;p&gt;The higher efforts fairly consistently reduce catastrophic failures and unlock animations and interactivity, but the peak taste is often around the mid-high levels, not at the absolute top of the dial.&lt;/p&gt;&lt;p&gt;Overall, Luna High and Extra High are the best value. Sol Ultra is the best website overall. And the best premium taste is Sol, especially once you break into Sol High.&lt;/p&gt;&lt;p&gt;The Milkbench overview has all the benchmark details and sites available to browse.&lt;/p&gt;&lt;p&gt;&lt;a href=&quot;https://milk.animasai.co/&quot;&gt;https://milk.animasai.co/&lt;/a&gt;&lt;/p&gt;</description></item>
<item><title>Starbucks and places to work</title><link>https://tylermayberry.dev/notes/starbucks-and-places-to-work</link><guid isPermaLink="true">https://tylermayberry.dev/notes/starbucks-and-places-to-work</guid><pubDate>Tue, 07 Jul 2026 12:00:00 +0000</pubDate><description>&lt;p&gt;Why is Starbucks getting worse at supporting people working in it?&lt;/p&gt;&lt;p&gt;These are the only tables inside this brand new Starbucks that have an outlet for laptops.&lt;/p&gt;&lt;p&gt;And you have to sit on a bench hunched over your laptop to use the only outlet spot in the whole building.&lt;/p&gt;&lt;p&gt;It was JUST BUILT. Everything is being designed to sell to you and send you on your way now.&lt;/p&gt;&lt;p&gt;Starbucks used to be my reliable place to get a few hours of work done when I&amp;#x27;m on the move.&lt;/p&gt;&lt;p&gt;Might be rethinking that now.&lt;/p&gt;&lt;p&gt;Good news is that I think culture is beginning to swing this in the other direction here in the states.&lt;/p&gt;&lt;p&gt;Not that it&amp;#x27;s going to support more electronic use, (although I don&amp;#x27;t think that&amp;#x27;s going anywhere), but people are craving human interactions again in a big way.&lt;/p&gt;&lt;p&gt;Which means people will want to spend a significant portion of their day in places like Starbucks, where they can be around other people and be productive.&lt;/p&gt;&lt;p&gt;We can only guess how it will play out from here, though.&lt;/p&gt;&lt;p&gt;I&amp;#x27;d place bets on community-centered, affordable coworking spots doing well in the future.&lt;/p&gt;</description></item>
<item><title>Launching Pip as a web app</title><link>https://tylermayberry.dev/notes/launching-pip</link><guid isPermaLink="true">https://tylermayberry.dev/notes/launching-pip</guid><pubDate>Mon, 06 Jul 2026 12:00:00 +0000</pubDate><description>&lt;p&gt;I&amp;#x27;m launching Pip today on Product Hunt.&lt;/p&gt;&lt;p&gt;This is something I started some time ago and I was gonna make it an actual app on the App Store and Play Store, but instead of jumping through all those hurdles, I just decided to release it as a web app and move on to other projects.&lt;/p&gt;&lt;p&gt;This isn&amp;#x27;t my current focus at the moment, but I do think it has real potential and I do think it&amp;#x27;s genuinely useful. I use it in my everyday life at the moment.&lt;/p&gt;&lt;p&gt;The whole idea is to help people who hate budgeting, hate spreadsheets, and wouldn&amp;#x27;t do it no matter how much it would help them.&lt;/p&gt;&lt;p&gt;Pip does it all automatically for you by accessing your transactions and bank through Plaid, taking into account how much you wanna save and your monthly bills through the cash engine, and then gives you one daily number that you can spend that day.&lt;/p&gt;&lt;p&gt;You can also ask Pip about anything about your finances and your accounts. And he&amp;#x27;ll talk to you about it.&lt;/p&gt;&lt;p&gt;I&amp;#x27;ll drop the Product Hunt link and Pip demo below.&lt;/p&gt;&lt;p&gt;Product Hunt:&lt;br&gt;&lt;a href=&quot;https://www.producthunt.com/products/pip-5&quot;&gt;https://www.producthunt.com/products/pip-5&lt;/a&gt;&lt;/p&gt;&lt;p&gt;Pip site with demo:&lt;br&gt;&lt;a href=&quot;https://spendwithpip.com&quot;&gt;https://spendwithpip.com&lt;/a&gt;&lt;/p&gt;&lt;p&gt;Also, it&amp;#x27;s open source MIT license.&lt;br&gt;&lt;a href=&quot;https://github.com/MayberryDT/Pip&quot;&gt;https://github.com/MayberryDT/Pip&lt;/a&gt;&lt;/p&gt;</description></item>
</channel>
</rss>
