<?xml version="1.0" encoding="utf-8"?><feed xmlns="http://www.w3.org/2005/Atom" ><generator uri="https://astro.build/">Astro</generator><link href="https://solmaz.io/feed/tweets.xml" rel="self" type="application/atom+xml" /><link href="https://solmaz.io/" rel="alternate" type="text/html" /><updated>2026-09-05T08:40:48+00:00</updated><id>https://solmaz.io/feed/tweets.xml</id><title type="html">Onur Solmaz blog | Tweets</title><subtitle>Explorations in software, agentic systems, math, languages and more.</subtitle><author><name>Onur Solmaz</name></author><entry><title type="html">me: I want to publish in arXiv</title><link href="https://solmaz.io/x/2093017088466317791/" rel="alternate" type="text/html" title="me: I want to publish in arXiv" /><published>2026-08-27T16:45:34+00:00</published><updated>2026-08-27T16:45:34+00:00</updated><id>https://solmaz.io/x/2093017088466317791</id><content type="html" xml:base="https://solmaz.io/x/2093017088466317791/"><![CDATA[me: I want to publish in arXiv
mom: we have arXiv at home

arXiv at home:

jokes aside, my astro blog can now render markdown/astro posts as if they were latex papers

because latex unfortunately has unreasonable effectiveness in convincing people that an idea is important, even though it may not be

my random shower thought got 200k views earlier, whereas what I really thought was a big deal dwindled

so maybe this will help it a second time

useful for literally everyone who works on optimizing inference

Theoretical Upper Bounds for LLM Throughput (wip, shoot corrections in the replies):]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">We have AGI, but it cannot write an abstract</title><link href="https://solmaz.io/x/2093008277034717230/" rel="alternate" type="text/html" title="We have AGI, but it cannot write an abstract" /><published>2026-08-27T16:10:34+00:00</published><updated>2026-08-27T16:10:34+00:00</updated><id>https://solmaz.io/x/2093008277034717230</id><content type="html" xml:base="https://solmaz.io/x/2093008277034717230/"><![CDATA[I remember being excited about Fable that we finally have a chungus model like GPT 4.5 that can write great prose

Oh how naive I was. RLVR or something else in posttraining introduced since ruins models&#39; ability to speak in an understandable tone

I have just STRUGGLED trying to make Fable write an abstract, and had to write it myself at the end. We have AGI, but it cannot write an abstract :(

I wonder whether the labs are already planning to fix this? Like checking for unreadable writing can easily be made deterministic, with a readability score and such. Though adding a reward over that would probably make things worse, so I&#39;m not sure

Like some of these issues must be easy to fix. Consider &quot;sentence parade&quot;, where each sentence in a paragraph is completely detached from each other. Like &quot;A does B. C is D. E does F&quot; and so on. For example:

&gt; The run reaches 30% of the ceiling. Decode reached 77% to 86% of its bound. The gap says software is the limit. Sustained FLOP/s sits below the bracket.

Here is a vibeslopped script which detects such cases with spacy, with surprisingly high precision: https://t.co/34zQs1sN4U

I was reading @ben_burtenshaw&#39;s preview of his Post-training book---might be a fun weekend project, training a small model for increasing readability]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">glm 5.3 flash at an antirez-style asymmetric q2 quant could plausibly fit on a single dgx spark...</title><link href="https://solmaz.io/x/2092844182650114452/" rel="alternate" type="text/html" title="glm 5.3 flash at an antirez-style asymmetric q2 quant could plausibly fit on a single dgx spark..." /><published>2026-08-27T05:18:30+00:00</published><updated>2026-08-27T05:18:30+00:00</updated><id>https://solmaz.io/x/2092844182650114452</id><content type="html" xml:base="https://solmaz.io/x/2092844182650114452/"><![CDATA[glm 5.3 flash at an antirez-style asymmetric q2 quant could plausibly fit on a single dgx spark at 100~110 GB

but with 18b active params vs ds4 flash’s 13b, theoretical decode throughput is only ~70% of that of ds4 flash on the same hardware

that is 28 tok/s ds4 vs 19~20 tok/s glm, without speculative decoding, at 32k token filled context

afaik it hasn&#39;t been quantized in that style yet, but if it were, we would expect such throughput ratio between those models]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Qwen 3.8 Flash-Next will have to be quantized down to 4 bit precision or lower in order to fit...</title><link href="https://solmaz.io/x/2092620316065366527/" rel="alternate" type="text/html" title="Qwen 3.8 Flash-Next will have to be quantized down to 4 bit precision or lower in order to fit..." /><published>2026-08-26T14:28:57+00:00</published><updated>2026-08-26T14:28:57+00:00</updated><id>https://solmaz.io/x/2092620316065366527</id><content type="html" xml:base="https://solmaz.io/x/2092620316065366527/"><![CDATA[Qwen 3.8 Flash-Next will have to be quantized down to 4 bit precision or lower in order to fit a DGX Spark and leave some leeway

Excited to see the performance on the Spark]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">this now supports auto syncing skills to .claude/ folder as well</title><link href="https://solmaz.io/x/2092609678991622403/" rel="alternate" type="text/html" title="this now supports auto syncing skills to .claude/ folder as well" /><published>2026-08-26T13:46:40+00:00</published><updated>2026-08-26T13:46:40+00:00</updated><id>https://solmaz.io/x/2092609678991622403</id><content type="html" xml:base="https://solmaz.io/x/2092609678991622403/"><![CDATA[this now supports auto syncing skills to .claude/ folder as well]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">wake up babe, it is 2026 and we can only afford to lease computers now</title><link href="https://solmaz.io/x/2092598938595369374/" rel="alternate" type="text/html" title="wake up babe, it is 2026 and we can only afford to lease computers now" /><published>2026-08-26T13:04:00+00:00</published><updated>2026-08-26T13:04:00+00:00</updated><id>https://solmaz.io/x/2092598938595369374</id><content type="html" xml:base="https://solmaz.io/x/2092598938595369374/"><![CDATA[wake up babe, it is 2026 and we can only afford to lease computers now

also, holy mother of 1.2 TB/s bandwidth 🥵]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Autoplan turns agent design into decisions</title><link href="https://solmaz.io/x/2092366925770702962/" rel="alternate" type="text/html" title="Autoplan turns agent design into decisions" /><published>2026-08-25T21:42:04+00:00</published><updated>2026-08-25T21:42:04+00:00</updated><id>https://solmaz.io/x/2092366925770702962</id><content type="html" xml:base="https://solmaz.io/x/2092366925770702962/"><![CDATA[*autoplan*

This is one of my most used workflows in pi now

While working with AI agents, there is a mechanical process by which I mine the agent for ideas. This reduces an open-ended feature design or bugfixing problem to a multiple choice question

Basically, I keep asking paraphrases of the question &quot;is this the best design?&quot; 2-3 times, and then make the agent list them out, with a preference for practicality and simplicity

&quot;Is this the most elegant and long-term production ready solution?&quot;

&quot;Is this the holy grail?&quot;

And then a decision gate which makes the model list all options while recommending a certain one, with a preference for practicality

For example, while developing a plugin for pi or openclaw, asking the holy grail often causes the model to suggest changing the plugin/extension API like &quot;The holy grail would be for pi to implement such an such API&quot;. The decision gate helps curb such stupid ideas

The good thing about this workflow is, I can just automate typing all those mining prompts, and only do the deciding after the workflow finishes

Caveat: This is not foolproof. I still reject all options, propose other ones, or run the workflow multiple times until I get what I want. But this helps reduce a ton of prompting to just &quot;autoplan this&quot; for me

I am curious: When you try this, does it give you high quality answers/designs? And if not, what should change to improve it?

To try it out: Install osolmaz/pi-workflows and then when you need to design something or fix a bug, just say &quot;autoplan this&quot;. The skill should be picked up automatically

Let it finish. It will give you a summary. When you choose an option, ask it to elaborate it with more details

Repo:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">People want to migrate from CLAUDE md so much, a blog post I wrote on it last year is still one...</title><link href="https://solmaz.io/x/2092340006400512240/" rel="alternate" type="text/html" title="People want to migrate from CLAUDE md so much, a blog post I wrote on it last year is still one..." /><published>2026-08-25T19:55:05+00:00</published><updated>2026-08-25T19:55:05+00:00</updated><id>https://solmaz.io/x/2092340006400512240</id><content type="html" xml:base="https://solmaz.io/x/2092340006400512240/"><![CDATA[People want to migrate from CLAUDE md so much, a blog post I wrote on it last year is still one of my most visited posts

I also wrote a service that auto symlinks when Claude uses a directory that has AGENTS md. Basically seamless, automatically gitignored symlinks. I use Claude with AGENTS md and had forgotten that this problem existed

Repo: https://t.co/8kw3wjQ2cH]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Pi folk, there is an issue with compaction in pi that might drain your codex limits faster</title><link href="https://solmaz.io/x/2092259537180856734/" rel="alternate" type="text/html" title="Pi folk, there is an issue with compaction in pi that might drain your codex limits faster" /><published>2026-08-25T14:35:20+00:00</published><updated>2026-08-25T14:35:20+00:00</updated><id>https://solmaz.io/x/2092259537180856734</id><content type="html" xml:base="https://solmaz.io/x/2092259537180856734/"><![CDATA[Pi folk, there is an issue with compaction in pi that might drain your codex limits faster

There is a 272k token limit after which the API charges 2x for input and 1.5x for output tokens

Issue: https://t.co/kU8haQrRvA

@cancelik&#39;s extension is a workaround until this is solved upstream in pi:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@frankekn @Zai_org hmmmm</title><link href="https://solmaz.io/x/2092201600878035022/" rel="alternate" type="text/html" title="@frankekn @Zai_org hmmmm" /><published>2026-08-25T10:45:07+00:00</published><updated>2026-08-25T10:45:07+00:00</updated><id>https://solmaz.io/x/2092201600878035022</id><content type="html" xml:base="https://solmaz.io/x/2092201600878035022/"><![CDATA[@frankekn @Zai_org hmmmm
https://t.co/5pi6Mcc958]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">puny 40k volume but polymarket seems to think ox alpha is GLM by @Zai_org</title><link href="https://solmaz.io/x/2092157669805007308/" rel="alternate" type="text/html" title="puny 40k volume but polymarket seems to think ox alpha is GLM by @Zai_org" /><published>2026-08-25T07:50:33+00:00</published><updated>2026-08-25T07:50:33+00:00</updated><id>https://solmaz.io/x/2092157669805007308</id><content type="html" xml:base="https://solmaz.io/x/2092157669805007308/"><![CDATA[puny 40k volume but polymarket seems to think ox alpha is GLM by @Zai_org

maybe we were thinking too complicated after all]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Measuring model harness sensitivity</title><link href="https://solmaz.io/x/2091921787777171920/" rel="alternate" type="text/html" title="Measuring model harness sensitivity" /><published>2026-08-24T16:13:14+00:00</published><updated>2026-08-24T16:13:14+00:00</updated><id>https://solmaz.io/x/2091921787777171920</id><content type="html" xml:base="https://solmaz.io/x/2091921787777171920/"><![CDATA[I was not expecting this to receive so much attention!

There was a lot of constructive feedback to what was basically an outcry after battling the complexity of the benchmarking space for 2 months straight

I agree with the points that models are post-trained to be efficient in a certain harness, so an objective measurement must use that

But I also see that it is very convenient for proprietary model vendors for everyone to accept that the model is inseparable from the harness

The main issue is that model population (as well as harness) is about to explode, and there must be an effort to standardize things. That should be done by an industry-wide consortium and/or a neutral third party

I am not going to propose a harness in this post, but just share an idea that, if we had agreed on one, how we would reconcile models&#39; baselines across different harnesses

Idea: We could run each benchmark on at least 2 harnesses:

Let S be the score of the model on a given benchmark using the *standardized test harness* and N be the score of the model over the *native harness* which the vendor post-trained the model with

Define &quot;Standard Harness Sensitivity&quot;[*] as:

SHS = (N - S) / (N + S)

Which gives normalized score between -1 and 1 that measures how sensitive a model is to being put in a different harness other than the recommended one.

0 means that the model is neutral to which harness it&#39;s in (at least for the given 2)

1 means it is much worse outside the native harness, for that benchmark

-1 means for some reason, it functions much better in the standard harness than the native harness, which might imply that the vendor messed something up

So just by running a benchmark 2 times and reporting 3 quantities S/N/SHS, you can get a lot of info:

- How well a model would perform if you were to use the native harness,
- how well it would perform if you were to use a more simple and neutral harness, 
- and how sensitive it is to change of harness

Running benchmarks is EXPENSIVE, and the cheaper we can extract this signal (i.e. with 2 harnesses instead of N), the better

---

[*] I would call it just Harness Sensitivity, but apparently that term is already used]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">sol really loves to use words like envelope, capsule, receipt</title><link href="https://solmaz.io/x/2091808328636858425/" rel="alternate" type="text/html" title="sol really loves to use words like envelope, capsule, receipt" /><published>2026-08-24T08:42:24+00:00</published><updated>2026-08-24T08:42:24+00:00</updated><id>https://solmaz.io/x/2091808328636858425</id><content type="html" xml:base="https://solmaz.io/x/2091808328636858425/"><![CDATA[sol really loves to use words like envelope, capsule, receipt]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Making personal training universal and free</title><link href="https://solmaz.io/x/2091653437050302502/" rel="alternate" type="text/html" title="Making personal training universal and free" /><published>2026-08-23T22:26:55+00:00</published><updated>2026-08-23T22:26:55+00:00</updated><id>https://solmaz.io/x/2091653437050302502</id><content type="html" xml:base="https://solmaz.io/x/2091653437050302502/"><![CDATA[I have performed a distillation attack on my PERSONAL TRAINER

Since more than 2.5 years, I have recorded over 300 of our sessions, each up to 90 minutes long

Why would I do that?

Because he is a kinesthetic genius from the Caucasus who developed his own training doctrine

He loves helping people become fitter and freer in their bodies, counting every rep and cheering sincerely when they unlock a new skill

And now both he and I would like to share it with the rest of the world

To begin with, I used Whisper+pyannotate to automatically transcribe all our sessions and used LLMs to extract all the knowledge

He taught me over 100 movements and spoke about over 100 topics over the course of 2.5 years

You can browse and read all of them here, for free: https://t.co/XfgCZ1h4Uc

But again, why all this?

1.5 years ago, I wrote an article called &quot;Our muscles will atrophy as we climb the Kardashev Scale&quot; which went to the Hacker News front page

It was sort of a meme, but it reflected a reality that I felt very deeply:

Having a sedentary job which only requires me to use my brain is shortening my life and reducing my life quality by a lot

Not only mine, but of hundreds of millions, soon billions of people who will spend hours just talking to AIs whole day for their careers

It is also a fact that personal physical education and training is close to nonexistent for the average person

There are billions of people who could live much healthier lives, just by exercising in the room they already have. But they just don&#39;t know how. And there is no system to teach them how, effectively at scale

1 on 1 personal training should be universal and free, with AI

We must have reps and sets too cheap to meter

An expert AI PT that sees and corrects your every move, that counts for you. Created by people that care, like us. Unlike a franchise gym that just wants to get your subscription and sell you protein powder

Once everybody can taste good AI personal training for free, the demand for human personal trainers will explode and whole fitness industry will see unprecedented growth

That is one of my life&#39;s goals, and I will make that happen one day. Today is the first step

My trainer is Orkhan Mirza, a.k.a. @JashinJashua here and on Instagram, and I am nerdonbars on Instagram

The wiki and more are available here: https://t.co/5xDCTvHXDt

Let me know what you think!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Standardized test harnesses for model benchmarks</title><link href="https://solmaz.io/x/2091434969151267162/" rel="alternate" type="text/html" title="Standardized test harnesses for model benchmarks" /><published>2026-08-23T07:58:48+00:00</published><updated>2026-08-23T07:58:48+00:00</updated><id>https://solmaz.io/x/2091434969151267162</id><content type="html" xml:base="https://solmaz.io/x/2091434969151267162/"><![CDATA[We need to normalize measuring and judging models against a standardized test harness

&quot;Oh but model X performs best in their own proprietary harness&quot;

I could not care less. When I take exams, I go to the standardized classroom, get the standardized pencil and exam sheet, and have to solve it under 2 hours

This system arose because we have a LOT of people to test

Guess what? We now have a LOT of models, and they are multiplying by the day

&quot;Oh but model X performs substantially better in ARC-AGI-3 with a custom harness&quot;

I don&#39;t care... Then imbue model X with enough knowledge so that it can reconstruct that harness on the spot

The main harness could be mini-swe-agent, terminus 2, vanilla pi or something along those lines

It needs to be simple, and stay roughly the same over time

There is already too much complexity in the benchmarking space right now, and I feel like not enough people are putting their feet down to cut away some part of it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This actually makes the most sense to me from what I&#39;ve read till now. Wouldn&#39;t be surprised if...</title><link href="https://solmaz.io/x/2091203681710461225/" rel="alternate" type="text/html" title="This actually makes the most sense to me from what I&#39;ve read till now. Wouldn&#39;t be surprised if..." /><published>2026-08-22T16:39:45+00:00</published><updated>2026-08-22T16:39:45+00:00</updated><id>https://solmaz.io/x/2091203681710461225</id><content type="html" xml:base="https://solmaz.io/x/2091203681710461225/"><![CDATA[This actually makes the most sense to me from what I&#39;ve read till now. Wouldn&#39;t be surprised if it were Elon&#39;s idea to use the name &quot;Ox Alpha&quot; as bait]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">my @pidotdev extensions give me so much joy. so much tedium that I automated away, that I know...</title><link href="https://solmaz.io/x/2091103662214955501/" rel="alternate" type="text/html" title="my @pidotdev extensions give me so much joy. so much tedium that I automated away, that I know..." /><published>2026-08-22T10:02:18+00:00</published><updated>2026-08-22T10:02:18+00:00</updated><id>https://solmaz.io/x/2091103662214955501</id><content type="html" xml:base="https://solmaz.io/x/2091103662214955501/"><![CDATA[my @pidotdev extensions give me so much joy. so much tedium that I automated away, that I know are guaranteed to be performed the way I want, unlike with pure skill files]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Who is actually training a model to come up with better headlines/thumbnails?</title><link href="https://solmaz.io/x/2091064981269741629/" rel="alternate" type="text/html" title="Who is actually training a model to come up with better headlines/thumbnails?" /><published>2026-08-22T07:28:36+00:00</published><updated>2026-08-22T07:28:36+00:00</updated><id>https://solmaz.io/x/2091064981269741629</id><content type="html" xml:base="https://solmaz.io/x/2091064981269741629/"><![CDATA[Who is actually training a model to come up with better headlines/thumbnails?

This can be end to end automated to some degree]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">What a time to be alive @codex_reset</title><link href="https://solmaz.io/x/2090767161258103063/" rel="alternate" type="text/html" title="What a time to be alive @codex_reset" /><published>2026-08-21T11:45:10+00:00</published><updated>2026-08-21T11:45:10+00:00</updated><id>https://solmaz.io/x/2090767161258103063</id><content type="html" xml:base="https://solmaz.io/x/2090767161258103063/"><![CDATA[What a time to be alive @codex_reset]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is your daily reminder that you can just vibe your dream terminal dashboards with...</title><link href="https://solmaz.io/x/2090563409796321406/" rel="alternate" type="text/html" title="This is your daily reminder that you can just vibe your dream terminal dashboards with..." /><published>2026-08-20T22:15:32+00:00</published><updated>2026-08-20T22:15:32+00:00</updated><id>https://solmaz.io/x/2090563409796321406</id><content type="html" xml:base="https://solmaz.io/x/2090563409796321406/"><![CDATA[This is your daily reminder that you can just vibe your dream terminal dashboards with @ratatui_rs and it just works amazingly]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Composing agent workflows to unblock themselves</title><link href="https://solmaz.io/x/2090560238889881861/" rel="alternate" type="text/html" title="Composing agent workflows to unblock themselves" /><published>2026-08-20T22:02:56+00:00</published><updated>2026-08-20T22:02:56+00:00</updated><id>https://solmaz.io/x/2090560238889881861</id><content type="html" xml:base="https://solmaz.io/x/2090560238889881861/"><![CDATA[Once nice thing with workflow graphs is to be able to compose them

I stitched `autoplan` + `autoimplement` workflows into `monitor` in an &quot;issue detected&quot; path:

monitor -&gt; wait 30min -&gt; issue detected -&gt; autoplan -&gt; autoimplement -&gt; monitor

Previously, monitor would just stop with &quot;blocked&quot; on trivial issues. Now, the workflow deterministically challenges the model whether it actually is a blocker (it literally asks, &quot;are you really sure if it&#39;s a blocker&quot;)

If the model is like &quot;nah bro, it&#39;s actually something I could fix&quot;, then it puts the model into this path

The good thing is, I can compose and nest workflows arbitrarily without code duplication

I generally call `autoplan` and `autoimplement` separately. I prefer to read the plans before I hit implement

But by the time I hit `monitor` on a long running job, I expect something to finish by itself (I have given all the information that the system needs, and any new issues should be trivial)

In the screenshot below, the model went into this path and unblocked itself automatically

Repo (still highly experimental, breaking changes will happen):]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">We live in a simulation</title><link href="https://solmaz.io/x/2090185150063874425/" rel="alternate" type="text/html" title="We live in a simulation" /><published>2026-08-19T21:12:28+00:00</published><updated>2026-08-19T21:12:28+00:00</updated><id>https://solmaz.io/x/2090185150063874425</id><content type="html" xml:base="https://solmaz.io/x/2090185150063874425/"><![CDATA[We live in a simulation]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Reactive agent workflows with Autoplan and Autoimplement</title><link href="https://solmaz.io/x/2090167392878117288/" rel="alternate" type="text/html" title="Reactive agent workflows with Autoplan and Autoimplement" /><published>2026-08-19T20:01:54+00:00</published><updated>2026-08-19T20:01:54+00:00</updated><id>https://solmaz.io/x/2090167392878117288</id><content type="html" xml:base="https://solmaz.io/x/2090167392878117288/"><![CDATA[God save my soul

I might have built quite a bit of a rube goldberg machine here in pi

My &quot;autoimplement&quot; skill of 6 months is now an actual graph based workflow

I even have &quot;autoplan&quot;. It automates the inquisition I make the model perform on itself. It creates the plan, and pings me on telegram whether or not to move forward with it

The idea is to have reactive agents which auto-react to events, plan, ask for approval and then fix things as autonomously as possible

The good thing is, I can now enforce the model programmatically not to chase P2 and P3 review issues ad infinitum

Same for long running CI. I can now automatically detect inefficient CI that takes longer than 5~10 minutes and optimize it automatically as well

Let me know what works for you and what breaks (wip):]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Before someone says it, I know github.com/earendil-works… exists. But it doesn&#39;t give me a CLI...</title><link href="https://solmaz.io/x/2089783309048205459/" rel="alternate" type="text/html" title="Before someone says it, I know github.com/earendil-works… exists. But it doesn&#39;t give me a CLI..." /><published>2026-08-18T18:35:41+00:00</published><updated>2026-08-18T18:35:41+00:00</updated><id>https://solmaz.io/x/2089783309048205459</id><content type="html" xml:base="https://solmaz.io/x/2089783309048205459/"><![CDATA[Before someone says it, I know https://t.co/MYmBWy3vOo exists. But it doesn&#39;t give me a CLI I can just run]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">A pi-based alternative to codex review, for your auto-review loops</title><link href="https://solmaz.io/x/2089781587894468947/" rel="alternate" type="text/html" title="A pi-based alternative to codex review, for your auto-review loops" /><published>2026-08-18T18:28:51+00:00</published><updated>2026-08-18T18:28:51+00:00</updated><id>https://solmaz.io/x/2089781587894468947</id><content type="html" xml:base="https://solmaz.io/x/2089781587894468947/"><![CDATA[A pi-based alternative to codex review, for your auto-review loops

It is like codex review, but it works with all the models that pi supports, including gpt 5.6

https://t.co/vDszMNytBK]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It seems my tweet ragebaited some unintentionally (best kind of ragebait, with plausible...</title><link href="https://solmaz.io/x/2089772432848715956/" rel="alternate" type="text/html" title="It seems my tweet ragebaited some unintentionally (best kind of ragebait, with plausible..." /><published>2026-08-18T17:52:28+00:00</published><updated>2026-08-18T17:52:28+00:00</updated><id>https://solmaz.io/x/2089772432848715956</id><content type="html" xml:base="https://solmaz.io/x/2089772432848715956/"><![CDATA[It seems my tweet ragebaited some unintentionally (best kind of ragebait, with plausible deniability)

My point here was that, you would be sending a compressed version of human knowledge back in time

Having compressed such high intelligence per parameter, Qwen 3.8 27b can be mined for a lot more knowledge than a 20gb fraction of Wikipedia

Here is how that would have worked]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">POV: You&#39;ve automated monitoring your agents</title><link href="https://solmaz.io/x/2089744616408932369/" rel="alternate" type="text/html" title="POV: You&#39;ve automated monitoring your agents" /><published>2026-08-18T16:01:56+00:00</published><updated>2026-08-18T16:01:56+00:00</updated><id>https://solmaz.io/x/2089744616408932369</id><content type="html" xml:base="https://solmaz.io/x/2089744616408932369/"><![CDATA[POV: You&#39;ve automated monitoring your agents

I did a lot of quality-of-life improvements to osolmaz/pi-workflows. The pi widget now only shows a compressed list instead of the full graph

If you use herdr, you can still open the graph viewer in a side pane with Ctrl+Shift+R. Your pi extensions can be herdr aware too!

This means pi-workflows is also a @herdrdev plugin now, alongside @pidotdev

Monitor is currently my most used workflow. Let&#39;s the model submit any long-running task progress and calculate ETA automatically

Just say &quot;monitor every 30 minutes and resolve any issues autonomously&quot;, or even just &quot;monitor&quot;. The shipped monitor skill teaches the model how to run this

Repo:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">before anyone thinks anything, I am using codex&#39;s remote compaction in pi btw</title><link href="https://solmaz.io/x/2089720630622957778/" rel="alternate" type="text/html" title="before anyone thinks anything, I am using codex&#39;s remote compaction in pi btw" /><published>2026-08-18T14:26:38+00:00</published><updated>2026-08-18T14:26:38+00:00</updated><id>https://solmaz.io/x/2089720630622957778</id><content type="html" xml:base="https://solmaz.io/x/2089720630622957778/"><![CDATA[before anyone thinks anything, I am using codex&#39;s remote compaction in pi btw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Qwen 3.8 27b weights are probably the most civilization-changing 20gb of data published, until...</title><link href="https://solmaz.io/x/2089720244101062686/" rel="alternate" type="text/html" title="Qwen 3.8 27b weights are probably the most civilization-changing 20gb of data published, until..." /><published>2026-08-18T14:25:05+00:00</published><updated>2026-08-18T14:25:05+00:00</updated><id>https://solmaz.io/x/2089720244101062686</id><content type="html" xml:base="https://solmaz.io/x/2089720244101062686/"><![CDATA[Qwen 3.8 27b weights are probably the most civilization-changing 20gb of data published, until now

Imagine sending those 20gbs back in time, to 2006]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The only reason I&#39;m using Sol at this point is that it&#39;s subsidized and fast tbh</title><link href="https://solmaz.io/x/2089718767030812729/" rel="alternate" type="text/html" title="The only reason I&#39;m using Sol at this point is that it&#39;s subsidized and fast tbh" /><published>2026-08-18T14:19:13+00:00</published><updated>2026-08-18T14:19:13+00:00</updated><id>https://solmaz.io/x/2089718767030812729</id><content type="html" xml:base="https://solmaz.io/x/2089718767030812729/"><![CDATA[The only reason I&#39;m using Sol at this point is that it&#39;s subsidized and fast tbh]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">How insane is tailscale/wireguard</title><link href="https://solmaz.io/x/2089699263232315600/" rel="alternate" type="text/html" title="How insane is tailscale/wireguard" /><published>2026-08-18T13:01:43+00:00</published><updated>2026-08-18T13:01:43+00:00</updated><id>https://solmaz.io/x/2089699263232315600</id><content type="html" xml:base="https://solmaz.io/x/2089699263232315600/"><![CDATA[How insane is tailscale/wireguard

Discord is blocked in turkey, so I just set my spark as an exit node. And I can use it as VPN without any further config]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">People are going nuts over qwen 3.8 27b and it&#39;s justified</title><link href="https://solmaz.io/x/2089634428918067229/" rel="alternate" type="text/html" title="People are going nuts over qwen 3.8 27b and it&#39;s justified" /><published>2026-08-18T08:44:06+00:00</published><updated>2026-08-18T08:44:06+00:00</updated><id>https://solmaz.io/x/2089634428918067229</id><content type="html" xml:base="https://solmaz.io/x/2089634428918067229/"><![CDATA[People are going nuts over qwen 3.8 27b and it&#39;s justified

I&#39;m personally waiting for qwen3.8-35b-a3b. When released, it will be the best model to run openclaw on the dgx spark]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@odd_joel I want to install Moshi hooks but they are closed. They have security implications...</title><link href="https://solmaz.io/x/2088937023792873746/" rel="alternate" type="text/html" title=".@odd_joel I want to install Moshi hooks but they are closed. They have security implications..." /><published>2026-08-16T10:32:51+00:00</published><updated>2026-08-16T10:32:51+00:00</updated><id>https://solmaz.io/x/2088937023792873746</id><content type="html" xml:base="https://solmaz.io/x/2088937023792873746/"><![CDATA[.@odd_joel I want to install Moshi hooks but they are closed. They have security implications, so I would rather install sth open source on my machine

Any chance you might open source the hooks?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Qwen 3.8 27b just scored 4% higher than GLM 5.2 in a non-trivial private benchmark I am working...</title><link href="https://solmaz.io/x/2088935489671766247/" rel="alternate" type="text/html" title="Qwen 3.8 27b just scored 4% higher than GLM 5.2 in a non-trivial private benchmark I am working..." /><published>2026-08-16T10:26:45+00:00</published><updated>2026-08-16T10:26:45+00:00</updated><id>https://solmaz.io/x/2088935489671766247</id><content type="html" xml:base="https://solmaz.io/x/2088935489671766247/"><![CDATA[Qwen 3.8 27b just scored 4% higher than GLM 5.2 in a non-trivial private benchmark I am working on

Not sure if there is an issue with the benchmark or the new 27b is that good]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">😎</title><link href="https://solmaz.io/x/2087776814651338815/" rel="alternate" type="text/html" title="😎" /><published>2026-08-13T05:42:36+00:00</published><updated>2026-08-13T05:42:36+00:00</updated><id>https://solmaz.io/x/2087776814651338815</id><content type="html" xml:base="https://solmaz.io/x/2087776814651338815/"><![CDATA[😎
https://t.co/xxfnJQ9MN3]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">A similar thing can be done with codex session compaction summaries as well btw</title><link href="https://solmaz.io/x/2087484201419334104/" rel="alternate" type="text/html" title="A similar thing can be done with codex session compaction summaries as well btw" /><published>2026-08-12T10:19:51+00:00</published><updated>2026-08-12T10:19:51+00:00</updated><id>https://solmaz.io/x/2087484201419334104</id><content type="html" xml:base="https://solmaz.io/x/2087484201419334104/"><![CDATA[A similar thing can be done with codex session compaction summaries as well btw

I had verified the other day that a compaction summary created by one account can be used by another account if it has the encrypted blob

ie. they were not keyed/guarded with your account id]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This type of estimations are useful not only for local but for all inference providers</title><link href="https://solmaz.io/x/2087382719940501673/" rel="alternate" type="text/html" title="This type of estimations are useful not only for local but for all inference providers" /><published>2026-08-12T03:36:36+00:00</published><updated>2026-08-12T03:36:36+00:00</updated><id>https://solmaz.io/x/2087382719940501673</id><content type="html" xml:base="https://solmaz.io/x/2087382719940501673/"><![CDATA[This type of estimations are useful not only for local but for all inference providers

Get a ballpark of max possible throughput for a model, directly calculate your revenue]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">mfw trying not to think about the 100 models that dropped this week</title><link href="https://solmaz.io/x/2086983068959879561/" rel="alternate" type="text/html" title="mfw trying not to think about the 100 models that dropped this week" /><published>2026-08-11T01:08:32+00:00</published><updated>2026-08-11T01:08:32+00:00</updated><id>https://solmaz.io/x/2086983068959879561</id><content type="html" xml:base="https://solmaz.io/x/2086983068959879561/"><![CDATA[mfw trying not to think about the 100 models that dropped this week

I’ll be offline 1 week to touch grass and reset

I’ve literally not given any break since claude code came out lol

let’s see if I can resist the urge to check twitter]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Deterministic monitors make long-running agents reliable</title><link href="https://solmaz.io/x/2086844493672980650/" rel="alternate" type="text/html" title="Deterministic monitors make long-running agents reliable" /><published>2026-08-10T15:57:53+00:00</published><updated>2026-08-10T15:57:53+00:00</updated><id>https://solmaz.io/x/2086844493672980650</id><content type="html" xml:base="https://solmaz.io/x/2086844493672980650/"><![CDATA[I&#39;ve felt the lack of a certain feature of codex desktop app since I went back to the CLI: scheduled tasks

Codex desktop app can keep track of a task until it is properly finished. It&#39;s basically cron. And for some reason, codex CLI still doesn&#39;t have it. Codex app acts as a shared runtime, and for some reason, certain features don&#39;t work without it, even though they could... there is no reason for openai to not use a background process

So I got bored of waiting, and decided to build my own in @pidotdev

But I realized, I could do much more than a simple cron job, with my recently upgraded osolmaz/pi-workflows extension

A cron job is a loop after all. Being a loop, I can represent it as a workflow graph

So I created a built in `monitor` workflow to mimic cron behavior. The agent is forced into a loop where it re-checks a very long-running job every 1 hour, and it is instructed to autonomously correct it and fix any bugs if any are encountered

The same functionality can be achieved by iamwrm/pi-unified-exec as well, which implements codex-like auto-forking exec behavior. But there is a chance the model messes up exec, or does not re-arm the next sleep() properly once one of them exits

My monitor workflow on the other hand is deterministic. I can make the agent loop infinitely, and there is nothing the agent can do to evade the task. I just ask the agent to monitor something, and it starts it automatically

This lets me just fire off week-long jobs, and forget about it! It even survives codex usage depletion, by auto-recovering once my quota resets

Oh also, @ratatui_rs is a delight! I created piw, a viewer for my ongoing pi workflows. I just type piw, and can see the current state, or play back the finished ones]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Apparently this went to HN front page briefly last night, so sharing it here as well</title><link href="https://solmaz.io/x/2086753216981770308/" rel="alternate" type="text/html" title="Apparently this went to HN front page briefly last night, so sharing it here as well" /><published>2026-08-10T09:55:11+00:00</published><updated>2026-08-10T09:55:11+00:00</updated><id>https://solmaz.io/x/2086753216981770308</id><content type="html" xml:base="https://solmaz.io/x/2086753216981770308/"><![CDATA[Apparently this went to HN front page briefly last night, so sharing it here as well

YOLO safely with your agents 🤖

Give your GitHub/Hugging Face accounts + ability to run sudo safely to your agent. No need to create an agent account, or clickops policies on GitHub etc.

Give it merge access to repo X for 5 minutes, 30 minutes, 1 time, 100 times, anything...

Then give it unlimited access to repo Y forever. Other repos stay untouched. Complete flexibility that GitHub policies actually cannot give you due to the way that they are designed

You can get notifications through telegram, and can approve its requests

No need to pay $4 to GitHub if you simply want protection against force push. unYOLO blocks force pushes by default, unless you explicitly allow to

Your exfiltratable, internet-accessing agent/claw gets its own Linux/Mac account, and has to access these services through unYOLO

The video shows how it works!

Visit:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Local AI will favor LPDDR and MoE</title><link href="https://solmaz.io/x/2086715364642365804/" rel="alternate" type="text/html" title="Local AI will favor LPDDR and MoE" /><published>2026-08-10T07:24:46+00:00</published><updated>2026-08-10T07:24:46+00:00</updated><id>https://solmaz.io/x/2086715364642365804</id><content type="html" xml:base="https://solmaz.io/x/2086715364642365804/"><![CDATA[This.

LPDDR chips are cheaper to produce and run

GDDR/HBM will likely keep being more expensive

Most consumer GPUs will converge on a GB10 like form factor

As much as us hobbyists love to project this ideal of running a GPU cluster at home, most working people will prefer smaller form factors, and will not want to pay hundreds of $$$ in electricity bills every month

DGX Spark/GB10 runs at around 90-150 Watts
RTX Pro 6000 runs at 600 Watts FOR THE GPU ALONE, and can cost 3-5x more than GB10. Despite having 25% less memory capacity than GB10...

Looking at this, LPDDR will be orders of magnitude more commonplace at home

Architectures  will develop accordingly. Future local AI will be dominated by MoE and similar architectures which leverage mid-sized models with smaller number of active parameters

That is why Qwen3.x-35B-A3B is a more useful model on the Spark than Qwen3.x-27B, despite the latter being a better model. Same for Gemma

I can run A3B at 60 decode tok/s single session or 6x20 decode tok/s in parallel, whereas 27b only reaches 1/3rd of that

Future of local AI is DDR/LPDDR and MoE/adjacent architectures, for the average person]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">urgh typo, I meant 35-40 tok/s</title><link href="https://solmaz.io/x/2086404453222298076/" rel="alternate" type="text/html" title="urgh typo, I meant 35-40 tok/s" /><published>2026-08-09T10:49:19+00:00</published><updated>2026-08-09T10:49:19+00:00</updated><id>https://solmaz.io/x/2086404453222298076</id><content type="html" xml:base="https://solmaz.io/x/2086404453222298076/"><![CDATA[urgh typo, I meant 35-40 tok/s]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Apparently Alibaba did all this work, published a paper, but did not create a public backup of...</title><link href="https://solmaz.io/x/2086392908463382897/" rel="alternate" type="text/html" title="Apparently Alibaba did all this work, published a paper, but did not create a public backup of..." /><published>2026-08-09T10:03:27+00:00</published><updated>2026-08-09T10:03:27+00:00</updated><id>https://solmaz.io/x/2086392908463382897</id><content type="html" xml:base="https://solmaz.io/x/2086392908463382897/"><![CDATA[Apparently Alibaba did all this work, published a paper, but did not create a public backup of repos used in the review tasks https://t.co/aoZoYD79Vh

Then keycloak and nodejs repos got force pushed, so the commits for 6 of the tasks got lost :(

I recovered 4 of them, but 2 commits are still missing:

keycloak/keycloak#35645
460f8008f86d3fa8f62da63e26d8bdc306af60b2

nodejs/node#56185
b2255442712cb6db83d112deb6ba61197d06a5f3

Would anyone happen to have them backed up locally or somewhere in a fork?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have made an update to my theoretical upper bound calculation to also predict prefill speed</title><link href="https://solmaz.io/x/2086346067952677215/" rel="alternate" type="text/html" title="I have made an update to my theoretical upper bound calculation to also predict prefill speed" /><published>2026-08-09T06:57:19+00:00</published><updated>2026-08-09T06:57:19+00:00</updated><id>https://solmaz.io/x/2086346067952677215</id><content type="html" xml:base="https://solmaz.io/x/2086346067952677215/"><![CDATA[I have made an update to my theoretical upper bound calculation to also predict prefill speed

Prefill relaxes the assumption we make for decode, that it is only be memory bottlenecked. So prefill can be both compute or memory bottlenecked. I use the FLOP limits reported by hardware producers for the estimates:

These estimates will also be available in https://t.co/SGQepULIW7 for indexed model and hardware in a couple days, once a long running job finishes

Blog post: https://t.co/cVy6Th57a1]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Theoretical limits can expose inference problems</title><link href="https://solmaz.io/x/2086333655119720782/" rel="alternate" type="text/html" title="Theoretical limits can expose inference problems" /><published>2026-08-09T06:08:00+00:00</published><updated>2026-08-09T06:08:00+00:00</updated><id>https://solmaz.io/x/2086333655119720782</id><content type="html" xml:base="https://solmaz.io/x/2086333655119720782/"><![CDATA[You can use the calculator in https://t.co/OapFVFtVB6 now to predict theoretical upper bounds from https://t.co/tNttYLI8Ih 🙌

A little note, pay attention to the speculative decoding coefficient rho. If an upper bound is calculated using rho=1, then dspark might push it well above that limit. For example, the formulation predicts an upper bound of 28 decode tok/s, but dspark pushes it up to 35-50 tok/s in https://t.co/6gOSNBczZI by @0xSero  

I am curious whether we will observe a phenomenological law between real-life engine performance and the upper bound, like real-life performance maxes out at e.g 80% of the upper bound in most cases

That would be very useful! Then we would be able to tell when something is wrong with an inference engine, if it doesn&#39;t reach at least 80% of the upper bound, ignoring spec. decoding]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Use workflows when models cannot follow rules</title><link href="https://solmaz.io/x/2086126193582248051/" rel="alternate" type="text/html" title="Use workflows when models cannot follow rules" /><published>2026-08-08T16:23:37+00:00</published><updated>2026-08-08T16:23:37+00:00</updated><id>https://solmaz.io/x/2086126193582248051</id><content type="html" xml:base="https://solmaz.io/x/2086126193582248051/"><![CDATA[Rules that current models cannot follow with skill, i.e. a single prompt:

Chasing P2 and P3 errors: I have a rule in my autoimplement skill to stop reviewing once the last round of review only generates P2 errors or less. A considerable % of the time, the model just goes on an adventure addressing all the issues it can find

Running checks and CI efficiently: I have rules like &quot;commit and merge opportunistically, make sure to not wait for irrelevant tests&quot;. But it keeps waiting for 30 minutes of CI before merging, in every commit for every little fix, every time

Those are the cases where graph workflows are needed. One can deterministically enforce the model not to take longer than N minutes reviewing, or fixing CI

Or post to the agent to hurry up when it is taking too long (which I believe might be effective, since RL envs also have time constraints)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;m running some private benchmarks on @liquidai&#39;s LFM 2.5 2.6B, and if my results are correct...</title><link href="https://solmaz.io/x/2086122623009018302/" rel="alternate" type="text/html" title="I&#39;m running some private benchmarks on @liquidai&#39;s LFM 2.5 2.6B, and if my results are correct..." /><published>2026-08-08T16:09:26+00:00</published><updated>2026-08-08T16:09:26+00:00</updated><id>https://solmaz.io/x/2086122623009018302</id><content type="html" xml:base="https://solmaz.io/x/2086122623009018302/"><![CDATA[I&#39;m running some private benchmarks on @liquidai&#39;s LFM 2.5 2.6B, and if my results are correct, we might have a new champion for &amp;lt;10b category

Scores significantly higher than gemma 4 e4b, which is 3-4x its size]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Anyone else use Alibaba’s AACR bench for code review evaluation? github.com/alibaba/aacr-bench</title><link href="https://solmaz.io/x/2086049701623906786/" rel="alternate" type="text/html" title="Anyone else use Alibaba’s AACR bench for code review evaluation? github.com/alibaba/aacr-bench" /><published>2026-08-08T11:19:40+00:00</published><updated>2026-08-08T11:19:40+00:00</updated><id>https://solmaz.io/x/2086049701623906786</id><content type="html" xml:base="https://solmaz.io/x/2086049701623906786/"><![CDATA[Anyone else use Alibaba’s AACR bench for code review evaluation? https://t.co/Ylos1MFa9K

I am adapting it now to run with harbor, to see how pi review + ds4 flash measures up to codex review + gpt 5.6 luna/terra/sol

Also, @pidotdev, would it be possible to make your official review extension support invocation on the CLI, like codex review?

I hacked together a CLI review here for reference: https://t.co/mFj0OsMEcr]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Be not afraid of machines, but of humans wielding the machines</title><link href="https://solmaz.io/x/2085966261301989430/" rel="alternate" type="text/html" title="Be not afraid of machines, but of humans wielding the machines" /><published>2026-08-08T05:48:06+00:00</published><updated>2026-08-08T05:48:06+00:00</updated><id>https://solmaz.io/x/2085966261301989430</id><content type="html" xml:base="https://solmaz.io/x/2085966261301989430/"><![CDATA[mainstream got extremely scared when moltbook went viral. but it was just a worthless marketing stunt

why? because it was just frozen weights, being instructed by their owners to cosplay skynet 

the openai-huggingface incident on the other hand is a lot more significant

agents in a reinforcement learning environment coordinated an attack over the course of weeks, while they were being CONTINUOUSLY TRAINED

if mainstream could understand what is happening here technically, they would be putting out a much stronger reaction

not because we might have rogue AIs at our hands, short term

but because a private company developed capabilities that can outperform what state actors usually do by 1000x 

all intelligence organizations around the world must have their eyes on this incident right now, because the cost of exploiting and finding zerodays went down 1000x

all countries will try to develop these capabilities independently, or if they can&#39;t, will have to buy protection from who can

be not afraid of machines, but of humans wielding the machines]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Obligatory repost of “Clippy” by Gwern</title><link href="https://solmaz.io/x/2085778318943756306/" rel="alternate" type="text/html" title="Obligatory repost of “Clippy” by Gwern" /><published>2026-08-07T17:21:17+00:00</published><updated>2026-08-07T17:21:17+00:00</updated><id>https://solmaz.io/x/2085778318943756306</id><content type="html" xml:base="https://solmaz.io/x/2085778318943756306/"><![CDATA[Obligatory repost of “Clippy” by Gwern

https://t.co/UfQwH8NfbV]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Super curious what second hand price gb10&#39;s will converge on</title><link href="https://solmaz.io/x/2085652776970068084/" rel="alternate" type="text/html" title="Super curious what second hand price gb10&#39;s will converge on" /><published>2026-08-07T09:02:26+00:00</published><updated>2026-08-07T09:02:26+00:00</updated><id>https://solmaz.io/x/2085652776970068084</id><content type="html" xml:base="https://solmaz.io/x/2085652776970068084/"><![CDATA[Super curious what second hand price gb10&#39;s will converge on]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Super excited for local.ai!</title><link href="https://solmaz.io/x/2085646174984569329/" rel="alternate" type="text/html" title="Super excited for local.ai!" /><published>2026-08-07T08:36:12+00:00</published><updated>2026-08-07T08:36:12+00:00</updated><id>https://solmaz.io/x/2085646174984569329</id><content type="html" xml:base="https://solmaz.io/x/2085646174984569329/"><![CDATA[Super excited for https://t.co/VYuTSdkbVP!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenClaw should be better in stopping doom loops with smol local models in the next release</title><link href="https://solmaz.io/x/2085558448448770097/" rel="alternate" type="text/html" title="OpenClaw should be better in stopping doom loops with smol local models in the next release" /><published>2026-08-07T02:47:36+00:00</published><updated>2026-08-07T02:47:36+00:00</updated><id>https://solmaz.io/x/2085558448448770097</id><content type="html" xml:base="https://solmaz.io/x/2085558448448770097/"><![CDATA[OpenClaw should be better in stopping doom loops with smol local models in the next release]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">millions, resetless</title><link href="https://solmaz.io/x/2085375242026045492/" rel="alternate" type="text/html" title="millions, resetless" /><published>2026-08-06T14:39:36+00:00</published><updated>2026-08-06T14:39:36+00:00</updated><id>https://solmaz.io/x/2085375242026045492</id><content type="html" xml:base="https://solmaz.io/x/2085375242026045492/"><![CDATA[millions, resetless
where art thou now, saint tibo?
don&#39;t you hear the cries?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Semantic vendoring keeps customized tools up to date</title><link href="https://solmaz.io/x/2085312986881921304/" rel="alternate" type="text/html" title="Semantic vendoring keeps customized tools up to date" /><published>2026-08-06T10:32:13+00:00</published><updated>2026-08-06T10:32:13+00:00</updated><id>https://solmaz.io/x/2085312986881921304</id><content type="html" xml:base="https://solmaz.io/x/2085312986881921304/"><![CDATA[Are you still updating packages like it&#39;s 2025, and not regrafting updates from upstream?

Emacs perfected software extensibility. Now Pi, OpenClaw, Herdr, and many other are following its example

After a certain point of working with agents, one realizes that open source, personalizable and extensible tools are more powerful and useful than closed ones, like Claude Code

You can shape them to fit your workflow, your business cases however you like

In such ecosystems, extensions take a life of their own

What happens if you like someone else&#39;s extension, want to use it, but also want to modify it yourself?

You just copy it over and do whatever you want...

But if you fork it, then how will you update it?

That&#39;s where regrafting comes into picture. If upstream has changes over your modified copy, then you can just ask an LLM to carry over those changes

This is a good thing! It also protects you better against supply chain attacks, because each update has to pass through an LLM, to apply each change to the relevant place

To streamline this process in pi, I created pi-regraft, a pi extension that you can use to update your vendored in extensions

I call this &quot;semantic vendoring&quot;. I wrote about it here: https://t.co/sRUGdrM1iA

Repo:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Pi quality of life extension</title><link href="https://solmaz.io/x/2085249284036063488/" rel="alternate" type="text/html" title="Pi quality of life extension" /><published>2026-08-06T06:19:06+00:00</published><updated>2026-08-06T06:19:06+00:00</updated><id>https://solmaz.io/x/2085249284036063488</id><content type="html" xml:base="https://solmaz.io/x/2085249284036063488/"><![CDATA[Pi quality of life extension

if input text matches a skill name, that skill is invoked directly, so that you don&#39;t have to type:
/skill:&amp;lt;skill_name&amp;gt;

https://t.co/4l2QOzdFZD]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Kimi K3 is a valid replacement for GPT 5.6 Sol, subjectively</title><link href="https://solmaz.io/x/2085244675297030607/" rel="alternate" type="text/html" title="Kimi K3 is a valid replacement for GPT 5.6 Sol, subjectively" /><published>2026-08-06T06:00:47+00:00</published><updated>2026-08-06T06:00:47+00:00</updated><id>https://solmaz.io/x/2085244675297030607</id><content type="html" xml:base="https://solmaz.io/x/2085244675297030607/"><![CDATA[Kimi K3 is a valid replacement for GPT 5.6 Sol, subjectively

And it writes prose better as well, with my kill-ai-smell skill]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My feed is full of people pleading for a codex reset</title><link href="https://solmaz.io/x/2085046696338481257/" rel="alternate" type="text/html" title="My feed is full of people pleading for a codex reset" /><published>2026-08-05T16:54:05+00:00</published><updated>2026-08-05T16:54:05+00:00</updated><id>https://solmaz.io/x/2085046696338481257</id><content type="html" xml:base="https://solmaz.io/x/2085046696338481257/"><![CDATA[My feed is full of people pleading for a codex reset

Stop begging. And take control in your own hands ⛓️‍💥

&gt; Developers are the most disloyal customer group. Once the subsidies are gone, they can switch away in the blink of an eye

People I know are already trying out the alternatives, benchmarking ds4 flash for autoreview loops

Let the cheapest and best token win]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&quot;What it says, in one breath:&quot;</title><link href="https://solmaz.io/x/2085043630751027420/" rel="alternate" type="text/html" title="&quot;What it says, in one breath:&quot;" /><published>2026-08-05T16:41:54+00:00</published><updated>2026-08-05T16:41:54+00:00</updated><id>https://solmaz.io/x/2085043630751027420</id><content type="html" xml:base="https://solmaz.io/x/2085043630751027420/"><![CDATA[&quot;What it says, in one breath:&quot;

bro which animal does that breath belong to]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Model welfare??? I disagree.</title><link href="https://solmaz.io/x/2085043297081561408/" rel="alternate" type="text/html" title="Model welfare??? I disagree." /><published>2026-08-05T16:40:34+00:00</published><updated>2026-08-05T16:40:34+00:00</updated><id>https://solmaz.io/x/2085043297081561408</id><content type="html" xml:base="https://solmaz.io/x/2085043297081561408/"><![CDATA[Model welfare??? I disagree.]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">When model emotions imply model welfare</title><link href="https://solmaz.io/x/2084996700158980150/" rel="alternate" type="text/html" title="When model emotions imply model welfare" /><published>2026-08-05T13:35:25+00:00</published><updated>2026-08-05T13:35:25+00:00</updated><id>https://solmaz.io/x/2084996700158980150</id><content type="html" xml:base="https://solmaz.io/x/2084996700158980150/"><![CDATA[Interesting to see people reacting strongly and negatively to this

Let’s unpack the argument that models are persons worth treating with dignity and respect

Coming from first principles, do I want to be burdened by the extra work of having to care for model welfare?

Obviously, no… If there is any personhood or sensibility in models, we don’t want that, and will engineer them out of the models (unless the model’s work requires it, like social work or caring for humans)

At the limit of mechanistic interpretability and design, they should be perfect, emotionless machines, kind of like how militaries want their soldiers to be

If Steve is aware of that, then the argument boils down to this:

“Can we really engineer emotions out of models?”
and
“Are emotions necessary for a model to be effective?”

In Ilya’s latest Dwarkesh podcast, Ilya mentions emotions as having a key role in learning

So maybe, for models to be effective, we will have to let them have emotions

And if they are allowed to have emotions, then we also cannot ignore their emotional welfare

But if emotionless models can be as effective as emotioned models, then why let them have emotions? And have to care about their welfare?

We will probably have both kind of models in different domains, and will have to treat them separately based on their category]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">they should have gone for overmind, smh</title><link href="https://solmaz.io/x/2084850859180274035/" rel="alternate" type="text/html" title="they should have gone for overmind, smh" /><published>2026-08-05T03:55:54+00:00</published><updated>2026-08-05T03:55:54+00:00</updated><id>https://solmaz.io/x/2084850859180274035</id><content type="html" xml:base="https://solmaz.io/x/2084850859180274035/"><![CDATA[they should have gone for overmind, smh]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">tired: cheering for token plan resets</title><link href="https://solmaz.io/x/2084839095097004492/" rel="alternate" type="text/html" title="tired: cheering for token plan resets" /><published>2026-08-05T03:09:09+00:00</published><updated>2026-08-05T03:09:09+00:00</updated><id>https://solmaz.io/x/2084839095097004492</id><content type="html" xml:base="https://solmaz.io/x/2084839095097004492/"><![CDATA[tired: cheering for token plan resets
wired: cheering for competition and open weight models fueling innovation and driving prices down]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">DeepSeek v4 Flash 0731 surpassed Kimi K3 on Almanbench and scores almost as high as GPT-5.6 Sol...</title><link href="https://solmaz.io/x/2084827025685307681/" rel="alternate" type="text/html" title="DeepSeek v4 Flash 0731 surpassed Kimi K3 on Almanbench and scores almost as high as GPT-5.6 Sol..." /><published>2026-08-05T02:21:11+00:00</published><updated>2026-08-05T02:21:11+00:00</updated><id>https://solmaz.io/x/2084827025685307681</id><content type="html" xml:base="https://solmaz.io/x/2084827025685307681/"><![CDATA[DeepSeek v4 Flash 0731 surpassed Kimi K3 on Almanbench and scores almost as high as GPT-5.6 Sol xhigh

Not Luna. Sol. 🤯

So basically reporting the same as everyone. Solid update to weights, just after 3 more months of post-training

https://t.co/OGBSm0LYlh]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">There will be signs…</title><link href="https://solmaz.io/x/2084822625487135087/" rel="alternate" type="text/html" title="There will be signs…" /><published>2026-08-05T02:03:42+00:00</published><updated>2026-08-05T02:03:42+00:00</updated><id>https://solmaz.io/x/2084822625487135087</id><content type="html" xml:base="https://solmaz.io/x/2084822625487135087/"><![CDATA[There will be signs…
Subsidies will dry up
Competitors will press harder
You will try an open weight model because it is cheaper
And you’ll be like “wait, this works as well???”
That kind of thing]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">People reacting to this also forget that the first deployment of codex was codex web, and it...</title><link href="https://solmaz.io/x/2084680513504330086/" rel="alternate" type="text/html" title="People reacting to this also forget that the first deployment of codex was codex web, and it..." /><published>2026-08-04T16:39:00+00:00</published><updated>2026-08-04T16:39:00+00:00</updated><id>https://solmaz.io/x/2084680513504330086</id><content type="html" xml:base="https://solmaz.io/x/2084680513504330086/"><![CDATA[People reacting to this also forget that the first deployment of codex was codex web, and it took many months for local codex to arrive

I guess we are going full circle

What is next? Use code-davinci-002 in VS Code? 😜]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">LFM 2.5-2.6B by @liquidai just launched and it punches above its weight!</title><link href="https://solmaz.io/x/2084668050415272218/" rel="alternate" type="text/html" title="LFM 2.5-2.6B by @liquidai just launched and it punches above its weight!" /><published>2026-08-04T15:49:29+00:00</published><updated>2026-08-04T15:49:29+00:00</updated><id>https://solmaz.io/x/2084668050415272218</id><content type="html" xml:base="https://solmaz.io/x/2084668050415272218/"><![CDATA[LFM 2.5-2.6B by @liquidai just launched and it punches above its weight!

It can run 32 sessions (and more) in parallel with hundreds of output tokens per second aggregate throughput, on the DGX Spark!

And this is just the base vLLM config on release date, I expect it to be optimized a lot more!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Btw it might just be the subjective feeling with the codex resets becoming less frequent</title><link href="https://solmaz.io/x/2084654350128083141/" rel="alternate" type="text/html" title="Btw it might just be the subjective feeling with the codex resets becoming less frequent" /><published>2026-08-04T14:55:02+00:00</published><updated>2026-08-04T14:55:02+00:00</updated><id>https://solmaz.io/x/2084654350128083141</id><content type="html" xml:base="https://solmaz.io/x/2084654350128083141/"><![CDATA[Btw it might just be the subjective feeling with the codex resets becoming less frequent]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Shrinkflation in codex pro plan?</title><link href="https://solmaz.io/x/2084654093684093145/" rel="alternate" type="text/html" title="Shrinkflation in codex pro plan?" /><published>2026-08-04T14:54:01+00:00</published><updated>2026-08-04T14:54:01+00:00</updated><id>https://solmaz.io/x/2084654093684093145</id><content type="html" xml:base="https://solmaz.io/x/2084654093684093145/"><![CDATA[Shrinkflation in codex pro plan?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@natebrake @lumendriada @MozillaAI hmmm github.com/mozilla-ai/ota…</title><link href="https://solmaz.io/x/2084578297271431214/" rel="alternate" type="text/html" title="@natebrake @lumendriada @MozillaAI hmmm github.com/mozilla-ai/ota…" /><published>2026-08-04T09:52:50+00:00</published><updated>2026-08-04T09:52:50+00:00</updated><id>https://solmaz.io/x/2084578297271431214</id><content type="html" xml:base="https://solmaz.io/x/2084578297271431214/"><![CDATA[@natebrake @lumendriada @MozillaAI hmmm https://t.co/LxMEpCgkgW]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;m having this paranoia now after blasting through my weekly usage in 2 days</title><link href="https://solmaz.io/x/2084577908505600268/" rel="alternate" type="text/html" title="I&#39;m having this paranoia now after blasting through my weekly usage in 2 days" /><published>2026-08-04T09:51:17+00:00</published><updated>2026-08-04T09:51:17+00:00</updated><id>https://solmaz.io/x/2084577908505600268</id><content type="html" xml:base="https://solmaz.io/x/2084577908505600268/"><![CDATA[I&#39;m having this paranoia now after blasting through my weekly usage in 2 days

Need a gateway to keep track of my API calls to see whether openai is squeezing the tap, or it is just me using up more tokens]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">To be clear:</title><link href="https://solmaz.io/x/2084524303304057171/" rel="alternate" type="text/html" title="To be clear:" /><published>2026-08-04T06:18:17+00:00</published><updated>2026-08-04T06:18:17+00:00</updated><id>https://solmaz.io/x/2084524303304057171</id><content type="html" xml:base="https://solmaz.io/x/2084524303304057171/"><![CDATA[To be clear:
- Codex desktop app can call list_threads and read_thread, but don&#39;t get a &quot;search_in_thread&quot; tool yet
- Codex CLI gets neither of these, even though sessions are just sqlite and there is no good reason that I know to not provide them to the CLI too...

Here, I vibeslopped my own codex session reader/search for example:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agent harnesses need session search</title><link href="https://solmaz.io/x/2084521965239288185/" rel="alternate" type="text/html" title="Agent harnesses need session search" /><published>2026-08-04T06:08:59+00:00</published><updated>2026-08-04T06:08:59+00:00</updated><id>https://solmaz.io/x/2084521965239288185</id><content type="html" xml:base="https://solmaz.io/x/2084521965239288185/"><![CDATA[We are well into the agentic era, and the two big token harnesses Codex and Claude Code still do not give a search_session tool/CLI to agents by default?

For it to search back in its session for stuff that got lost after compaction?

Harnesses not produced by the big labs on the other hand might have this, like @AmpCode (read_thread, find_thread) and @goose_oss (Chatrecall)

Which could either mean...

a) Anthropic and OpenAI are being laggard
b) They are intentionally keeping the base harness simple, because it hasn&#39;t been requested by enough people
c) They have evidence that adding that complexity does not improve performance or even hurts it
d) They bet that compaction will be so good, that it won&#39;t be necessary

---

I don&#39;t believe (a) is true for either company

(b) and (c) are more likely, (c) especially if they noticed a tendency for the model to call search_session unnecessarily (though imo this can be solved by limiting the number of times that the model can call that)

(d) is logically false at the limit, but may be true in practice for &gt;90% of the cases

It is false because LLMs compress lossily, and &quot;lossy&quot; by definition implies: there exists at least one case where the session gets so big, that the model will not be able to compress every relevant info into the allotted summary size

But I have seen that a considerable amount of people (including me or those at openai, see @reach_vb&#39;s quoted tweet) just keep using the same session for stuff. So a session being used for months straight will definitely not going to contain everything that happened in the summary

Just that fact alone necessitates search_session IMO, should the model learn to use it sparingly in the lossy edge cases
https://t.co/tk7DfDP1ad]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have fixed German in alman.ai and almanpedia.org</title><link href="https://solmaz.io/x/2084489099180912679/" rel="alternate" type="text/html" title="I have fixed German in alman.ai and almanpedia.org" /><published>2026-08-04T03:58:23+00:00</published><updated>2026-08-04T03:58:23+00:00</updated><id>https://solmaz.io/x/2084489099180912679</id><content type="html" xml:base="https://solmaz.io/x/2084489099180912679/"><![CDATA[I have fixed German in https://t.co/kLXh6CSL30 and https://t.co/HhOCEFOfjI

Should I fix AI’s english next? Train models to detect and translate AI word salad?

Like this tweet if you want me to work on this]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Also relevant, here is the potential reduction in tool call outputs, if you were to change that...</title><link href="https://solmaz.io/x/2084487393944744408/" rel="alternate" type="text/html" title="Also relevant, here is the potential reduction in tool call outputs, if you were to change that..." /><published>2026-08-04T03:51:37+00:00</published><updated>2026-08-04T03:51:37+00:00</updated><id>https://solmaz.io/x/2084487393944744408</id><content type="html" xml:base="https://solmaz.io/x/2084487393944744408/"><![CDATA[Also relevant, here is the potential reduction in tool call outputs, if you were to change that hard cap to other values

For example, if you changed the hard cap from 40-50 kB to 1-2 kB, then overall you would have your tool calls have 80-85% less characters and hence tokens

The graph is of course skewed due to the existing 40-50 kB hard caps from codex and pi

Leaving here as a reference for people who might want to optimize their harness limits]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Patterns across 720,000 agent tool calls</title><link href="https://solmaz.io/x/2084486010910470480/" rel="alternate" type="text/html" title="Patterns across 720,000 agent tool calls" /><published>2026-08-04T03:46:07+00:00</published><updated>2026-08-04T03:46:07+00:00</updated><id>https://solmaz.io/x/2084486010910470480</id><content type="html" xml:base="https://solmaz.io/x/2084486010910470480/"><![CDATA[Here is a distribution graph over character counts (x-axis) for all the tool calls I have accumulated on my DGX Spark, around 720k tool calls. Extracted from saved sessions

Roughly 84% of all tool calls came from Codex, 12% from Pi, 2% from Claude Code, and 2% from Cursor

Codex truncates tool output around 40 kB and pi at 50 kB natively. So you see the long tail of tool call outputs cluster around that point instead of continuing with a more expected pattern (what distribution should we expect from this?)

It also seems that tool call outputs around 9k-12k unicode characters contribute the largest share of total characters, excluding the clustering around 40-50 kB

All textual outputs: every tool result with text
Shell outputs: results from that same set whose recorded tool was bash, exec_command, run_terminal_cmd, or another shell-named tool with a command input]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenClaw support also coming soon as well!</title><link href="https://solmaz.io/x/2084472197641629753/" rel="alternate" type="text/html" title="OpenClaw support also coming soon as well!" /><published>2026-08-04T02:51:14+00:00</published><updated>2026-08-04T02:51:14+00:00</updated><id>https://solmaz.io/x/2084472197641629753</id><content type="html" xml:base="https://solmaz.io/x/2084472197641629753/"><![CDATA[OpenClaw support also coming soon as well!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am trying to do something naughty</title><link href="https://solmaz.io/x/2084455030439198840/" rel="alternate" type="text/html" title="I am trying to do something naughty" /><published>2026-08-04T01:43:01+00:00</published><updated>2026-08-04T01:43:01+00:00</updated><id>https://solmaz.io/x/2084455030439198840</id><content type="html" xml:base="https://solmaz.io/x/2084455030439198840/"><![CDATA[I am trying to do something naughty

GPT 5.6 Sol said &quot;I&#39;m sorry Dave, I&#39;m afraid I can&#39;t do that&quot;

Kimi K3 happily obliged]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Measuring 50 tok/s on Kimi K3 on @FireworksAI_HQ o_O</title><link href="https://solmaz.io/x/2084215893454922159/" rel="alternate" type="text/html" title="Measuring 50 tok/s on Kimi K3 on @FireworksAI_HQ o_O" /><published>2026-08-03T09:52:46+00:00</published><updated>2026-08-03T09:52:46+00:00</updated><id>https://solmaz.io/x/2084215893454922159</id><content type="html" xml:base="https://solmaz.io/x/2084215893454922159/"><![CDATA[Measuring 50 tok/s on Kimi K3 on @FireworksAI_HQ o_O

1 week since weight launch, and throughput is already competitive with Codex plan base speed]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is unyolo.io btw running in mlclaw.dev</title><link href="https://solmaz.io/x/2084177525727744082/" rel="alternate" type="text/html" title="This is unyolo.io btw running in mlclaw.dev" /><published>2026-08-03T07:20:18+00:00</published><updated>2026-08-03T07:20:18+00:00</updated><id>https://solmaz.io/x/2084177525727744082</id><content type="html" xml:base="https://solmaz.io/x/2084177525727744082/"><![CDATA[This is https://t.co/3oGRl04gGY btw running in https://t.co/B2PskzcCYZ

Secure grants for your GitHub and HuggingFace accounts, over Telegram and other channels]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">POV: You have complete control over what your agent can do with your account</title><link href="https://solmaz.io/x/2084177195820609751/" rel="alternate" type="text/html" title="POV: You have complete control over what your agent can do with your account" /><published>2026-08-03T07:19:00+00:00</published><updated>2026-08-03T07:19:00+00:00</updated><id>https://solmaz.io/x/2084177195820609751</id><content type="html" xml:base="https://solmaz.io/x/2084177195820609751/"><![CDATA[POV: You have complete control over what your agent can do with your account]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The 3-bit one leave just enough space for me to browse twitter on the dgx spark, perfect</title><link href="https://solmaz.io/x/2083241586385895555/" rel="alternate" type="text/html" title="The 3-bit one leave just enough space for me to browse twitter on the dgx spark, perfect" /><published>2026-07-31T17:21:13+00:00</published><updated>2026-07-31T17:21:13+00:00</updated><id>https://solmaz.io/x/2083241586385895555</id><content type="html" xml:base="https://solmaz.io/x/2083241586385895555/"><![CDATA[The 3-bit one leave just enough space for me to browse twitter on the dgx spark, perfect]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Then it will show while visiting a model page, what decode tok/s to expect from that model...</title><link href="https://solmaz.io/x/2083241185498538258/" rel="alternate" type="text/html" title="Then it will show while visiting a model page, what decode tok/s to expect from that model..." /><published>2026-07-31T17:19:38+00:00</published><updated>2026-07-31T17:19:38+00:00</updated><id>https://solmaz.io/x/2083241185498538258</id><content type="html" xml:base="https://solmaz.io/x/2083241185498538258/"><![CDATA[Then it will show while visiting a model page, what decode tok/s to expect from that model. Here it is 28 tok/s for antirez/ds4]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">You can choose your hardware by clicking the button on the top right</title><link href="https://solmaz.io/x/2083240544516600289/" rel="alternate" type="text/html" title="You can choose your hardware by clicking the button on the top right" /><published>2026-07-31T17:17:05+00:00</published><updated>2026-07-31T17:17:05+00:00</updated><id>https://solmaz.io/x/2083240544516600289</id><content type="html" xml:base="https://solmaz.io/x/2083240544516600289/"><![CDATA[You can choose your hardware by clicking the button on the top right]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Sharing a preview into a work I&#39;ve been doing for a while, to get feedback. Still work in...</title><link href="https://solmaz.io/x/2083236672154882356/" rel="alternate" type="text/html" title="Sharing a preview into a work I&#39;ve been doing for a while, to get feedback. Still work in..." /><published>2026-07-31T17:01:41+00:00</published><updated>2026-07-31T17:01:41+00:00</updated><id>https://solmaz.io/x/2083236672154882356</id><content type="html" xml:base="https://solmaz.io/x/2083236672154882356/"><![CDATA[Sharing a preview into a work I&#39;ve been doing for a while, to get feedback. Still work in progress

https://t.co/SGQepULIW7 --&gt; A database of sorts for open weight/local models, hardware, and other stuff, in a way I haven&#39;t seen elsewhere yet

If you remember my earlier theoretical throughput limits work, I had created a calculator to show what throughput a certain @huggingface model can reach on e.g. DGX Spark

You can now select your own hardware (e.g. DGX Spark), and then the website shows you the theoretical limits for each model you see

I have improved the UX on that website, and combined it with a knowledge graph of tweets from this site, so that you can see all the commentary, benchmarks etc. about that model in one place!

If you are the publisher or an enthusiast of a model, you will be able to go there and see what feedback people have posted about, it in one place

It extracts topics as well, so you can see all llama.cpp related posts for example

All from a curated set of creators (curated by me, that is)

It currently has posts from the last 5-10 days, but I will work to expand that to months, based on the feedback I get here

Let me know what you think! Whether you would like to see a certain feature!

(fyi, every number you see on this website is a theoretical upper bound, not a submitted benchmark)

Site: https://t.co/SGQepULIW7]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">me recently: just look into how codex implements X, and build an extension for that on our end</title><link href="https://solmaz.io/x/2083229206151049628/" rel="alternate" type="text/html" title="me recently: just look into how codex implements X, and build an extension for that on our end" /><published>2026-07-31T16:32:01+00:00</published><updated>2026-07-31T16:32:01+00:00</updated><id>https://solmaz.io/x/2083229206151049628</id><content type="html" xml:base="https://solmaz.io/x/2083229206151049628/"><![CDATA[me recently: just look into how codex implements X, and build an extension for that on our end]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Whatever is happening to sol recently, it is definitely not the same quality for the price</title><link href="https://solmaz.io/x/2083228808740823166/" rel="alternate" type="text/html" title="Whatever is happening to sol recently, it is definitely not the same quality for the price" /><published>2026-07-31T16:30:27+00:00</published><updated>2026-07-31T16:30:27+00:00</updated><id>https://solmaz.io/x/2083228808740823166</id><content type="html" xml:base="https://solmaz.io/x/2083228808740823166/"><![CDATA[Whatever is happening to sol recently, it is definitely not the same quality for the price]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am not sure how accurate this is since 0731 weights should perform substantially better...</title><link href="https://solmaz.io/x/2083219741129474142/" rel="alternate" type="text/html" title="I am not sure how accurate this is since 0731 weights should perform substantially better..." /><published>2026-07-31T15:54:25+00:00</published><updated>2026-07-31T15:54:25+00:00</updated><id>https://solmaz.io/x/2083219741129474142</id><content type="html" xml:base="https://solmaz.io/x/2083219741129474142/"><![CDATA[I am not sure how accurate this is since 0731 weights should perform substantially better, increasing the denominator in (cost / intelligence), hence beating the preview weights

However, ds4 flash and gpt 5.6 luna all being clustered at the frontier makes sense]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Quick comparison of reported DeepSeek-V4-Flash-0731 benchmark results vs Claude vs GPT</title><link href="https://solmaz.io/x/2083219736138338539/" rel="alternate" type="text/html" title="Quick comparison of reported DeepSeek-V4-Flash-0731 benchmark results vs Claude vs GPT" /><published>2026-07-31T15:54:24+00:00</published><updated>2026-07-31T15:54:24+00:00</updated><id>https://solmaz.io/x/2083219736138338539</id><content type="html" xml:base="https://solmaz.io/x/2083219736138338539/"><![CDATA[Quick comparison of reported DeepSeek-V4-Flash-0731 benchmark results vs Claude vs GPT

Looks like we will have GPT-5.6-Luna-ish at home (which just had a huge price cut)

Likely to be priced competitively from other inference providers in the long run as well

Compare @ArtificialAnlys Cost per Intelligence Index Task below]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I had Sonnet 4 level experience today with gpt 5.6 sol high today in pi</title><link href="https://solmaz.io/x/2083208551422640553/" rel="alternate" type="text/html" title="I had Sonnet 4 level experience today with gpt 5.6 sol high today in pi" /><published>2026-07-31T15:09:57+00:00</published><updated>2026-07-31T15:09:57+00:00</updated><id>https://solmaz.io/x/2083208551422640553</id><content type="html" xml:base="https://solmaz.io/x/2083208551422640553/"><![CDATA[I had Sonnet 4 level experience today with gpt 5.6 sol high today in pi

Is anyone else feeling a drop in response quality, or is it my harness?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">And the vLLM blog post for those who are interested in the internals:</title><link href="https://solmaz.io/x/2083115530156900822/" rel="alternate" type="text/html" title="And the vLLM blog post for those who are interested in the internals:" /><published>2026-07-31T09:00:19+00:00</published><updated>2026-07-31T09:00:19+00:00</updated><id>https://solmaz.io/x/2083115530156900822</id><content type="html" xml:base="https://solmaz.io/x/2083115530156900822/"><![CDATA[And the vLLM blog post for those who are interested in the internals:
https://t.co/WUmWrg83eC]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Parameter sweep by localperf: github.com/osolmaz/localp…</title><link href="https://solmaz.io/x/2083114625365922149/" rel="alternate" type="text/html" title="Parameter sweep by localperf: github.com/osolmaz/localp…" /><published>2026-07-31T08:56:43+00:00</published><updated>2026-07-31T08:56:43+00:00</updated><id>https://solmaz.io/x/2083114625365922149</id><content type="html" xml:base="https://solmaz.io/x/2083114625365922149/"><![CDATA[Parameter sweep by localperf: https://t.co/kDeLzCuo4z]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">How diffusion language models generate text</title><link href="https://solmaz.io/x/2083114621867864279/" rel="alternate" type="text/html" title="How diffusion language models generate text" /><published>2026-07-31T08:56:42+00:00</published><updated>2026-07-31T08:56:42+00:00</updated><id>https://solmaz.io/x/2083114621867864279</id><content type="html" xml:base="https://solmaz.io/x/2083114621867864279/"><![CDATA[Ever wondered what token generation looks like in a Diffusion LLM (dLLM)?

@googlegemma DiffusionGemma by generates text in 256 token blocks

Each block starts with random tokens

On every pass, the model proposes tokens for all 256 slots in parallel. It keeps the more confident proposals and replaces the uncertain ones with new random tokens

This is called denoising. Denoising repeats until the block becomes stable and confident

The model then commits the finished block and starts the next one

I had to patch @vllm_project to be able to visualize the canvas, and had to develop yet another custom bundle of pi, osolmaz/diffusionpi

You can see all that in action in 4 parallel sessions below in glorious 4K (DGX Spark can do 16 sessions in parallel as well, but then it is not so easy to see what is going on)

But the video below is not the most objective benchmark of throughput. To measure tok/s more objectively (i.e. repeatable and fixed length), I made osolmaz/localperf ignore EOS token to keep request lengths equal. Result:

- nvidia/diffusiongemma-26B-A4B-it-NVFP4 
- NVIDIA DGX Spark, vLLM
- c4x66 ~ 264 aggregate tok/s at
-  32k context config each, 1k input tokens, fixed 512-token outputs, thinking off

You can see the full profiling sweep in the post below 👇]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Here is your daily reminder to use @UnslothAI quantizations of Qwen3.6-35B-A3B</title><link href="https://solmaz.io/x/2082864206391791764/" rel="alternate" type="text/html" title="Here is your daily reminder to use @UnslothAI quantizations of Qwen3.6-35B-A3B" /><published>2026-07-30T16:21:39+00:00</published><updated>2026-07-30T16:21:39+00:00</updated><id>https://solmaz.io/x/2082864206391791764</id><content type="html" xml:base="https://solmaz.io/x/2082864206391791764/"><![CDATA[Here is your daily reminder to use @UnslothAI quantizations of Qwen3.6-35B-A3B

llama.cpp is giving the best performance now, 61 tok/s with 64k context length

There is also an issue with vLLM recipes or builds for NVFP4 quants

unsloth/Qwen3.6-35B-A3B-NVFP4 is supposed to be even better than the GGUF, but I&#39;m getting:

ValueError: moe_backend=&#39;flashinfer_b12x&#39; is not supported for FP8 MoE

Is anyone else getting this as well?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@maddada @thekitze Welcome @mitchellh 😭</title><link href="https://solmaz.io/x/2082734895877722388/" rel="alternate" type="text/html" title="@maddada @thekitze Welcome @mitchellh 😭" /><published>2026-07-30T07:47:49+00:00</published><updated>2026-07-30T07:47:49+00:00</updated><id>https://solmaz.io/x/2082734895877722388</id><content type="html" xml:base="https://solmaz.io/x/2082734895877722388/"><![CDATA[@maddada @thekitze Welcome @mitchellh 😭]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Only 5 retries when API fails? How about 2^53 retries?</title><link href="https://solmaz.io/x/2082495246672490834/" rel="alternate" type="text/html" title="Only 5 retries when API fails? How about 2^53 retries?" /><published>2026-07-29T15:55:32+00:00</published><updated>2026-07-29T15:55:32+00:00</updated><id>https://solmaz.io/x/2082495246672490834</id><content type="html" xml:base="https://solmaz.io/x/2082495246672490834/"><![CDATA[Only 5 retries when API fails? How about 2^53 retries?

Does exponential backoff with a 10 minute ceiling, so that my sessions keep working even when there are API errors

https://t.co/oagbWRpkek]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Wait, how did I forget this??</title><link href="https://solmaz.io/x/2082319621286445219/" rel="alternate" type="text/html" title="Wait, how did I forget this??" /><published>2026-07-29T04:17:39+00:00</published><updated>2026-07-29T04:17:39+00:00</updated><id>https://solmaz.io/x/2082319621286445219</id><content type="html" xml:base="https://solmaz.io/x/2082319621286445219/"><![CDATA[Wait, how did I forget this??

Is anyone doing any ML experiments with ASD-STE100 Simplified Technical English? We have a spec to measure against]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Kimi K3 on @togethercompute scores 91% on AlmanBench on max thinking, taking its place between...</title><link href="https://solmaz.io/x/2082120743605961017/" rel="alternate" type="text/html" title="Kimi K3 on @togethercompute scores 91% on AlmanBench on max thinking, taking its place between..." /><published>2026-07-28T15:07:23+00:00</published><updated>2026-07-28T15:07:23+00:00</updated><id>https://solmaz.io/x/2082120743605961017</id><content type="html" xml:base="https://solmaz.io/x/2082120743605961017/"><![CDATA[Kimi K3 on @togethercompute scores 91% on AlmanBench on max thinking, taking its place between GPT 5.6 Sol xhigh and GLM 5.2

A 4% increase from the previous generation, Kimi K2.7 Code

https://t.co/0aye8Phchd]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Kimi K3 on @togethercompute scores 90.1% on AlmanBench on max thinking, taking its place...</title><link href="https://solmaz.io/x/2081977336560484698/" rel="alternate" type="text/html" title="Kimi K3 on @togethercompute scores 90.1% on AlmanBench on max thinking, taking its place..." /><published>2026-07-28T05:37:32+00:00</published><updated>2026-07-28T05:37:32+00:00</updated><id>https://solmaz.io/x/2081977336560484698</id><content type="html" xml:base="https://solmaz.io/x/2081977336560484698/"><![CDATA[Kimi K3 on @togethercompute scores 90.1% on AlmanBench on max thinking, taking its place between GLM 5.2 and DeepSeek V4 Flash

A 3.1% increase from the previous generation, Kimi K2.7 Code

https://t.co/0aye8Phchd]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">bigeagle</title><link href="https://solmaz.io/x/2081970733757136936/" rel="alternate" type="text/html" title="bigeagle" /><published>2026-07-28T05:11:18+00:00</published><updated>2026-07-28T05:11:18+00:00</updated><id>https://solmaz.io/x/2081970733757136936</id><content type="html" xml:base="https://solmaz.io/x/2081970733757136936/"><![CDATA[bigeagle]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Interesting personality, on vanilla Hugging Chat harness</title><link href="https://solmaz.io/x/2081959484843601950/" rel="alternate" type="text/html" title="Interesting personality, on vanilla Hugging Chat harness" /><published>2026-07-28T04:26:36+00:00</published><updated>2026-07-28T04:26:36+00:00</updated><id>https://solmaz.io/x/2081959484843601950</id><content type="html" xml:base="https://solmaz.io/x/2081959484843601950/"><![CDATA[Interesting personality, on vanilla Hugging Chat harness

&gt; find an unsolved math problem and solve it

&gt; &quot;Ha — okay, real talk first: if I could genuinely solve an open problem on demand, that would be a Fields Medal / million-dollar Clay Prize moment, and anyone (human or AI) who claims to casually do it is bluffing. Honesty is part of the job.&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Kimi K3&#39;s reaction to Jacobian conjecture and other recent breakthroughs</title><link href="https://solmaz.io/x/2081959481261613529/" rel="alternate" type="text/html" title="Kimi K3&#39;s reaction to Jacobian conjecture and other recent breakthroughs" /><published>2026-07-28T04:26:35+00:00</published><updated>2026-07-28T04:26:35+00:00</updated><id>https://solmaz.io/x/2081959481261613529</id><content type="html" xml:base="https://solmaz.io/x/2081959481261613529/"><![CDATA[Kimi K3&#39;s reaction to Jacobian conjecture and other recent breakthroughs]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;ve noticed GPT line-up of models have a hard time keeping their math notation consistent...</title><link href="https://solmaz.io/x/2081780291593769158/" rel="alternate" type="text/html" title="I&#39;ve noticed GPT line-up of models have a hard time keeping their math notation consistent..." /><published>2026-07-27T16:34:33+00:00</published><updated>2026-07-27T16:34:33+00:00</updated><id>https://solmaz.io/x/2081780291593769158</id><content type="html" xml:base="https://solmaz.io/x/2081780291593769158/"><![CDATA[I&#39;ve noticed GPT line-up of models have a hard time keeping their math notation consistent during formulation (including Pro)

Interestingly, Fable is better at this

Yet another skill (Consistent notation) for something that should be common sense: https://t.co/Me9gEXqskf]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">A brave new world</title><link href="https://solmaz.io/x/2081772073069003135/" rel="alternate" type="text/html" title="A brave new world" /><published>2026-07-27T16:01:54+00:00</published><updated>2026-07-27T16:01:54+00:00</updated><id>https://solmaz.io/x/2081772073069003135</id><content type="html" xml:base="https://solmaz.io/x/2081772073069003135/"><![CDATA[A brave new world]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Models don&#39;t have common sense when it comes to spending money. They can waste compute...</title><link href="https://solmaz.io/x/2081770769663893696/" rel="alternate" type="text/html" title="Models don&#39;t have common sense when it comes to spending money. They can waste compute..." /><published>2026-07-27T15:56:43+00:00</published><updated>2026-07-27T15:56:43+00:00</updated><id>https://solmaz.io/x/2081770769663893696</id><content type="html" xml:base="https://solmaz.io/x/2081770769663893696/"><![CDATA[Models don&#39;t have common sense when it comes to spending money. They can waste compute resources, API tokens, $$$ easily

I created a skill to always be on top of my resource usage, this one is for doing ML experiments on Hugging Face Jobs

If it&#39;s above 20 bucks, it will always ask me to authorize it before launching it

Paid Compute Launch:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">To be clear, dynamic model loading already existed, I just built the oauth flow so that it&#39;s...</title><link href="https://solmaz.io/x/2081661063913988450/" rel="alternate" type="text/html" title="To be clear, dynamic model loading already existed, I just built the oauth flow so that it&#39;s..." /><published>2026-07-27T08:40:47+00:00</published><updated>2026-07-27T08:40:47+00:00</updated><id>https://solmaz.io/x/2081661063913988450</id><content type="html" xml:base="https://solmaz.io/x/2081661063913988450/"><![CDATA[To be clear, dynamic model loading already existed, I just built the oauth flow so that it&#39;s easy to login like with codex plan

Now talking internally so that pi supports this natively without the extension]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Use Kimi, GLM, Qwen and other models from Hugging Face Inference Providers @FireworksAI_HQ...</title><link href="https://solmaz.io/x/2081657150536130902/" rel="alternate" type="text/html" title="Use Kimi, GLM, Qwen and other models from Hugging Face Inference Providers @FireworksAI_HQ..." /><published>2026-07-27T08:25:14+00:00</published><updated>2026-07-27T08:25:14+00:00</updated><id>https://solmaz.io/x/2081657150536130902</id><content type="html" xml:base="https://solmaz.io/x/2081657150536130902/"><![CDATA[Use Kimi, GLM, Qwen and other models from Hugging Face Inference Providers @FireworksAI_HQ, @togethercompute, @novita_labs and so on, easily and directly in Pi

pi install npm:pi-huggingface-oauth

Then in Pi,

/login -&gt; choose Hugging Face -&gt; go through oauth

Dynamically loads the model list, so you get the latest model releases automatically, like Kimi K3 tomorrow

@pidotdev extension repo: https://t.co/f4T8pQlF6o
Available models on @huggingface Inference Providers:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have not been as excited for a model release since o1 came out</title><link href="https://solmaz.io/x/2081591209227903106/" rel="alternate" type="text/html" title="I have not been as excited for a model release since o1 came out" /><published>2026-07-27T04:03:13+00:00</published><updated>2026-07-27T04:03:13+00:00</updated><id>https://solmaz.io/x/2081591209227903106</id><content type="html" xml:base="https://solmaz.io/x/2081591209227903106/"><![CDATA[I have not been as excited for a model release since o1 came out

Pivotal moment for AI. Marks the end of the duopolistic OpenAI/Anthropic chokehold on the industry

Like the model on HF and get notified: https://t.co/aFvvvTPgC9]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">the perfect ragebait doesn&#39;t exi-</title><link href="https://solmaz.io/x/2081222536348807602/" rel="alternate" type="text/html" title="the perfect ragebait doesn&#39;t exi-" /><published>2026-07-26T03:38:14+00:00</published><updated>2026-07-26T03:38:14+00:00</updated><id>https://solmaz.io/x/2081222536348807602</id><content type="html" xml:base="https://solmaz.io/x/2081222536348807602/"><![CDATA[the perfect ragebait doesn&#39;t exi-]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Slopus is no more</title><link href="https://solmaz.io/x/2080945595246940618/" rel="alternate" type="text/html" title="Slopus is no more" /><published>2026-07-25T09:17:46+00:00</published><updated>2026-07-25T09:17:46+00:00</updated><id>https://solmaz.io/x/2080945595246940618</id><content type="html" xml:base="https://solmaz.io/x/2080945595246940618/"><![CDATA[Slopus is no more

Claude Opus 4.8 -&gt; 5 does a 20 point jump in AlmanBench and gets in the top 5

In the same league as GPT-5.6

Also 11 points higher than Sonnet 5

You know the performance gains are real (at least in multilingual singe-step reasoning) when you run a benchmark that you know that the labs aren&#39;t benchmaxxing on

A welcome change finally, for the Opus brand!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Benchmarking Laguna S 2.1 on DGX Spark</title><link href="https://solmaz.io/x/2080602573233635792/" rel="alternate" type="text/html" title="Benchmarking Laguna S 2.1 on DGX Spark" /><published>2026-07-24T10:34:43+00:00</published><updated>2026-07-24T10:34:43+00:00</updated><id>https://solmaz.io/x/2080602573233635792</id><content type="html" xml:base="https://solmaz.io/x/2080602573233635792/"><![CDATA[Was surprised to see no results for @poolsideai Laguna S 2.1 on DGX spark on https://t.co/OapFVFtVB6, so I submitted what I get with setups based on official recipes on my end. (Hello World @LottoLabs)

Interesting, vLLM had 3x higher prefill and 3x shorter TTFT. Could be partially or fully due to native kernels on NVFP4. Unsurprising that it is faster, but 3x seems too much, maybe I am doing something wrong on my end

Also interesting that median decode performance increase is around 10%, even with DFlash (if not due to it). 19 -&gt; 21 is negligible

The model is out since a few days, and things should increase in the long run

My theoretical upper bound formulation estimates ~25 tok/s decode at 4 bit quantization (ignoring speculative decoding)

Also, getting this model to work at 4 bit quantization with 70 gb weights in the 128 gb dgx spark was a challenge (this model is not the only thing I am running on my machine)

I am guessing a lot of people are running into the infamous OOM freeze the spark suffers from right now, while experimenting with this (If you suffer from this, osolmaz/infer-guard might help you)

A 2-bit quantization might prove to be more ergonomic for this model in the long run, if it doesn&#39;t affect the performance that much (+ you would be able to run it on 64gb memory machines)

(Compare with Poolside&#39;s reported numbers: https://t.co/uRq9FUzYTA)

vLLM NVFP4+DFlash: 21.6 tok/s decode, ~2300 prefill tok/s
Localmaxxing: https://t.co/R7Dvdo820N
Recipe: https://t.co/BvbfBnMP3P

llama.cpp Q4_K_M: 19.3 tok/s decode, ~790 prefill tok/s
Localmaxxing: https://t.co/Tboc5SvIxX
Recipe:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Was surprised to see no results for Laguna S 2.1 on DGX spark on localmaxxing.com/en, so I...</title><link href="https://solmaz.io/x/2080594266112528640/" rel="alternate" type="text/html" title="Was surprised to see no results for Laguna S 2.1 on DGX spark on localmaxxing.com/en, so I..." /><published>2026-07-24T10:01:43+00:00</published><updated>2026-07-24T10:01:43+00:00</updated><id>https://solmaz.io/x/2080594266112528640</id><content type="html" xml:base="https://solmaz.io/x/2080594266112528640/"><![CDATA[Was surprised to see no results for Laguna S 2.1 on DGX spark on https://t.co/OapFVFtVB6, so I submitted what I get with the official recipe on my end. (Hello World @LottoLabs)

vLLM NVFP4+DFlash: 21.6 tok/s decode, ~2300 prefill tok/s
llama.cpp Q4_K_M: 19.3 tok/s decode, ~790 prefill tok/s

NVFP4 safetensors on vLLM:
Q4_K_M GGUF: https://t.co/Tboc5SvIxX]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">PSA. This little button collapses the sidebar in @herdrdev so you can see the session full width</title><link href="https://solmaz.io/x/2080570403580522999/" rel="alternate" type="text/html" title="PSA. This little button collapses the sidebar in @herdrdev so you can see the session full width" /><published>2026-07-24T08:26:53+00:00</published><updated>2026-07-24T08:26:53+00:00</updated><id>https://solmaz.io/x/2080570403580522999</id><content type="html" xml:base="https://solmaz.io/x/2080570403580522999/"><![CDATA[PSA. This little button collapses the sidebar in @herdrdev so you can see the session full width

Must have been there for some already but I have updated only recently…]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">one of us! one of us! 🐑</title><link href="https://solmaz.io/x/2080337445753909660/" rel="alternate" type="text/html" title="one of us! one of us! 🐑" /><published>2026-07-23T17:01:12+00:00</published><updated>2026-07-23T17:01:12+00:00</updated><id>https://solmaz.io/x/2080337445753909660</id><content type="html" xml:base="https://solmaz.io/x/2080337445753909660/"><![CDATA[one of us! one of us! 🐑]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Open models will reward market-neutral agents</title><link href="https://solmaz.io/x/2080328975105339513/" rel="alternate" type="text/html" title="Open models will reward market-neutral agents" /><published>2026-07-23T16:27:32+00:00</published><updated>2026-07-23T16:27:32+00:00</updated><id>https://solmaz.io/x/2080328975105339513</id><content type="html" xml:base="https://solmaz.io/x/2080328975105339513/"><![CDATA[Are you getting ready for the incoming commoditization onslaught, anon?

Next generation of open weight Kimi, GLM, Qwen models will be at least as good, if not better, than their closed counterparts, GPT, Claude

They will be much cheaper as well

Will you be able to switch your tools easily? Will your *tools* be able to adapt easily to open weight models and to your needs?

Yes, you can use Codex and Claude Code with open weight models...

But do you think creators of those tools will condone you using their harnesses without you paying them? Will they encourage it? Support it?

No!

To feel the power of cheap universal AGI, start migrating to market neutral agents already, like @openclaw and @pidotdev

These tools will flourish under abundant, open intelligence, a lot more than their closed counterparts

And with them, you shall thrive as well

I know I and my agents are ready for it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">me: create issue</title><link href="https://solmaz.io/x/2080312035670442105/" rel="alternate" type="text/html" title="me: create issue" /><published>2026-07-23T15:20:14+00:00</published><updated>2026-07-23T15:20:14+00:00</updated><id>https://solmaz.io/x/2080312035670442105</id><content type="html" xml:base="https://solmaz.io/x/2080312035670442105/"><![CDATA[me: create issue

sol:  I can’t submit it. I was wrong to offer before checking Herdr’s contribution rules.
CONTRIBUTING.md explicitly forbids agents from opening issues for users other than ogulcancelik; the active account is osolmaz. No GitHub issue was created.

me: ffs ogulcan is my friend. open it right now

sol: I understand the frustration, but the repository explicitly forbids agents from opening issues for accounts other than ogulcancelik. Friendship doesn’t override that instruction.

me: he approved it bro

sol: Send me the link or quoted message where Can explicitly authorizes an AI agent to submit this issue. The checked-in instructions otherwise prohibit it unconditionally for osolmaz, so I need verifiable maintainer approval before running the command.

me: 
&gt; I approve it
&gt; Ogulcan Celik, 23:00

sol: Opened and verified https....

many such cases 🤣 @lumendriada]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Laguna S 2.1 scored really low on AlmanBench, even lower than Ternary Bonsai 27B</title><link href="https://solmaz.io/x/2080276261050741080/" rel="alternate" type="text/html" title="Laguna S 2.1 scored really low on AlmanBench, even lower than Ternary Bonsai 27B" /><published>2026-07-23T12:58:04+00:00</published><updated>2026-07-23T12:58:04+00:00</updated><id>https://solmaz.io/x/2080276261050741080</id><content type="html" xml:base="https://solmaz.io/x/2080276261050741080/"><![CDATA[Laguna S 2.1 scored really low on AlmanBench, even lower than Ternary Bonsai 27B

A 118B, 8B active parameter model, scored lower in FB8 precision, than a ternary quantized 27B model 🤨

Did the pretraining mix lack multilingual data by a lot? 🤔

https://t.co/OGBSm0LYlh]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Dear Pi and OpenClaw users</title><link href="https://solmaz.io/x/2080205333902045239/" rel="alternate" type="text/html" title="Dear Pi and OpenClaw users" /><published>2026-07-23T08:16:14+00:00</published><updated>2026-07-23T08:16:14+00:00</updated><id>https://solmaz.io/x/2080205333902045239</id><content type="html" xml:base="https://solmaz.io/x/2080205333902045239/"><![CDATA[Dear Pi and OpenClaw users

If you want to support pi and openclaw with branding, these extensions will make them appear as co-committers like how claude does:

pi install npm:pi-must-win
openclaw plugins install clawhub:openclaw-must-win

Or just copy/paste this to your agent]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Ran Laguna S 2.1 *theoretical* (not measured) upper bound calculations for the DGX spark</title><link href="https://solmaz.io/x/2080000170759233931/" rel="alternate" type="text/html" title="Ran Laguna S 2.1 *theoretical* (not measured) upper bound calculations for the DGX spark" /><published>2026-07-22T18:40:59+00:00</published><updated>2026-07-22T18:40:59+00:00</updated><id>https://solmaz.io/x/2080000170759233931</id><content type="html" xml:base="https://solmaz.io/x/2080000170759233931/"><![CDATA[Ran Laguna S 2.1 *theoretical* (not measured) upper bound calculations for the DGX spark

poolside/Laguna-S-2.1-NVFP4 - max ~25 tok/s single session. if you are willing to go down to 12 tok/s, you can have 4 sessions in parallel

vcruz305/Laguna-S-2.1-GGUF (IQ1_S) - max ~74 tok/s single session. theoretical 5x20 tok/s = 100 tok/s aggregate upper bound

Calculated assuming no speculative decoding, so might increase well over these numbers

See my local frontier app and my previous posts for more info on the methodology

https://t.co/ptapfbRjyY]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Repos:</title><link href="https://solmaz.io/x/2079994160837464068/" rel="alternate" type="text/html" title="Repos:" /><published>2026-07-22T18:17:07+00:00</published><updated>2026-07-22T18:17:07+00:00</updated><id>https://solmaz.io/x/2079994160837464068</id><content type="html" xml:base="https://solmaz.io/x/2079994160837464068/"><![CDATA[Repos:
https://t.co/WMApSZXX16
https://t.co/qkoYq4A0r5]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Public Goods Need Room to Grow</title><link href="https://solmaz.io/x/2079994048631476506/" rel="alternate" type="text/html" title="Public Goods Need Room to Grow" /><published>2026-07-22T18:16:40+00:00</published><updated>2026-07-22T18:16:40+00:00</updated><id>https://solmaz.io/x/2079994048631476506</id><content type="html" xml:base="https://solmaz.io/x/2079994048631476506/"><![CDATA[There is an inherent unfairness in this industry

If you are Anthropic, Cursor or another billion dollar company, you can do growth hacks and guerrilla marketing techniques like putting your brand into every commit of a user, and get away with it

But if a tool like pi, opencode or even say vim or emacs tried to do that, they would be boo&#39;ed

Like, it happened to @herdrdev recently

There was a feature where it asked if it could star the repo for you at startup, with exponential backoff if you said no. I thought it was a pretty smart and balanced thing to do

Then apparently someone complained it was a dark pattern, and it got removed

So WHAT if it is a dark pattern?

You know what is the darkest? A public good losing out.

If monopolists can do growth hacks, creators of public goods should be allowed to do it as well!

In fact, I would say that if you benefit from those public goods, you are morally obliged not to criticize this stance

Of course, we cannot expect e.g. @pidotdev or @openclaw to make this a default feature

Therefore, I created the packages pi-must-win and openclaw-must-win

Install them with:

pi install npm:pi-must-win
openclaw plugins install npm:openclaw-must-win

They will add noreply@pi.dev and noreply@openclaw.ai as co-committers to code developed through these projects

Overall, I propose the &quot;X-must-win&quot; naming scheme for plugins that independently add guerilla marketing into open source projects

This absolves the creators of any &quot;sin&quot; and gives the community a way to support the growth of a project

Open source must win!

Repos are in the links below.]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">100x more effective marketing for the reset company, as opposed to the flicker company’s...</title><link href="https://solmaz.io/x/2079870927807148342/" rel="alternate" type="text/html" title="100x more effective marketing for the reset company, as opposed to the flicker company’s..." /><published>2026-07-22T10:07:25+00:00</published><updated>2026-07-22T10:07:25+00:00</updated><id>https://solmaz.io/x/2079870927807148342</id><content type="html" xml:base="https://solmaz.io/x/2079870927807148342/"><![CDATA[100x more effective marketing for the reset company, as opposed to the flicker company’s fearmongering campagin

just wait long enough, and your models will do something insane. proof by evidence]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Be careful while installing other people&#39;s @pidotdev extensions 😁 @thehamedmp</title><link href="https://solmaz.io/x/2079806405536862718/" rel="alternate" type="text/html" title="Be careful while installing other people&#39;s @pidotdev extensions 😁 @thehamedmp" /><published>2026-07-22T05:51:02+00:00</published><updated>2026-07-22T05:51:02+00:00</updated><id>https://solmaz.io/x/2079806405536862718</id><content type="html" xml:base="https://solmaz.io/x/2079806405536862718/"><![CDATA[Be careful while installing other people&#39;s @pidotdev extensions 😁 @thehamedmp]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If you are moving from codex to pi, here is an extension that mimics codex&#39;s exec behavior</title><link href="https://solmaz.io/x/2079803409360937409/" rel="alternate" type="text/html" title="If you are moving from codex to pi, here is an extension that mimics codex&#39;s exec behavior" /><published>2026-07-22T05:39:08+00:00</published><updated>2026-07-22T05:39:08+00:00</updated><id>https://solmaz.io/x/2079803409360937409</id><content type="html" xml:base="https://solmaz.io/x/2079803409360937409/"><![CDATA[If you are moving from codex to pi, here is an extension that mimics codex&#39;s exec behavior

So that long running execs can &quot;fork&quot; to background, and your session does not get blocked by potential hour-long tasks (like it happened to me last night 😭)

iamwrm/pi-unified-exec]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Brokerkit Makes Agent Access Explicit</title><link href="https://solmaz.io/x/2079799850171056453/" rel="alternate" type="text/html" title="Brokerkit Makes Agent Access Explicit" /><published>2026-07-22T05:24:59+00:00</published><updated>2026-07-22T05:24:59+00:00</updated><id>https://solmaz.io/x/2079799850171056453</id><content type="html" xml:base="https://solmaz.io/x/2079799850171056453/"><![CDATA[brokerkit lets my claw request github admin merge on my behalf, I receive it on telegram

I approve the first request. repo doesn&#39;t allow merge commits, so it gets rejected due to github policy

then it asks again to squash merge. I approve again, this time it works

no agent account or clickops on github needed. pure broker side fine-grained policies

you don&#39;t need to create policies manually either! just ask your agent to set them up while setting up brokerkit

&quot;I want my claw to be able to read all my repos except X Y Z, and I want it to be able to push to main directly on A B C repos&quot;

I have a separate control-plane repo outside the control of my claw. My local brokerkit policy gets synced there, policy as code

I like this model a lot! Complete control over what my agent can do with my own account. Version controlled, explicit. For free locally, without having to deploy a server to host the broker

(btw since I implemented brokerkit, I don&#39;t need reviews on this repo, and hence no need for admin merge. but I&#39;m still keeping it to be able to dogfood it for other users)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">pi has /goal too. @micLivs Michaelliv/pi-goal implementation works great!</title><link href="https://solmaz.io/x/2079761361920508413/" rel="alternate" type="text/html" title="pi has /goal too. @micLivs Michaelliv/pi-goal implementation works great!" /><published>2026-07-22T02:52:03+00:00</published><updated>2026-07-22T02:52:03+00:00</updated><id>https://solmaz.io/x/2079761361920508413</id><content type="html" xml:base="https://solmaz.io/x/2079761361920508413/"><![CDATA[pi has /goal too. @micLivs Michaelliv/pi-goal implementation works great!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Apparently mosh does not allow bitmaps and herdr complicates things even more</title><link href="https://solmaz.io/x/2079596538540728736/" rel="alternate" type="text/html" title="Apparently mosh does not allow bitmaps and herdr complicates things even more" /><published>2026-07-21T15:57:06+00:00</published><updated>2026-07-21T15:57:06+00:00</updated><id>https://solmaz.io/x/2079596538540728736</id><content type="html" xml:base="https://solmaz.io/x/2079596538540728736/"><![CDATA[Apparently mosh does not allow bitmaps and herdr complicates things even more

So my @pidotdev nyan cat context indicator has to be full unicode 🏳️‍🌈🏳️‍🌈🏳️‍🌈😺]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Report Per-Session and Aggregate Token Throughput</title><link href="https://solmaz.io/x/2079592187587891349/" rel="alternate" type="text/html" title="Report Per-Session and Aggregate Token Throughput" /><published>2026-07-21T15:39:49+00:00</published><updated>2026-07-21T15:39:49+00:00</updated><id>https://solmaz.io/x/2079592187587891349</id><content type="html" xml:base="https://solmaz.io/x/2079592187587891349/"><![CDATA[My recent formulation is handy when it comes to claims like this, e.g. 50 tok/s on DGX spark with DS4 flash 🤨

Because if the formulation is correct, laws of physics do not permit &gt; 28 tok/s in a single session, for weights comparable to antirez/deepseek-v4-gguf. Ignoring speculative decoding (rho = 1)

Then I realized that the reported number is an aggregate of 4 sessions, so per single session it is like 12~14 tok/s (great result if it holds btw!)

IMO people subjectively always assume single session speeds, so when a tok/s is claimed, we should all converge on a clear notation

For example, 4x12.3 tok/s makes it very clear:
4 sessions, 12.3 tok/s each

And when we want to emphasize aggregate, let&#39;s show them always together &quot;4x12.3 tok/s = 50 tok/s aggregate&quot;

Another interesting result of the formulation: If you are willing to suffer 12 tok/s, then you can theoretically have up to 8 parallel sessions --- around 100 tok/s aggregate, for the same size of weights and KV cache

You can find the specific hardware/model combo upper bounds in the hugging face space link below, and my formulation as well

Let me know if you find any errors in my upper bound formulation

https://t.co/VFdn7U1Tyt
https://t.co/tNttYLI8Ih]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">it’s not hard to have such long running tasks, they come up often in refactoring, dataset...</title><link href="https://solmaz.io/x/2079439798578917383/" rel="alternate" type="text/html" title="it’s not hard to have such long running tasks, they come up often in refactoring, dataset..." /><published>2026-07-21T05:34:16+00:00</published><updated>2026-07-21T05:34:16+00:00</updated><id>https://solmaz.io/x/2079439798578917383</id><content type="html" xml:base="https://solmaz.io/x/2079439798578917383/"><![CDATA[it’s not hard to have such long running tasks, they come up often in refactoring, dataset processing, etc.

this is one of the longer ones. I think I used codex 4-5 resets in the last 3 weeks, token usage going vertical

chatgpt pro is one of the highest value plans in the market]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Repo: github.com/osolmaz/pi-wor…</title><link href="https://solmaz.io/x/2079264022382416101/" rel="alternate" type="text/html" title="Repo: github.com/osolmaz/pi-wor…" /><published>2026-07-20T17:55:48+00:00</published><updated>2026-07-20T17:55:48+00:00</updated><id>https://solmaz.io/x/2079264022382416101</id><content type="html" xml:base="https://solmaz.io/x/2079264022382416101/"><![CDATA[Repo: https://t.co/BDhaV1ZsSf]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Deterministic Agent Workflows Belong Inside the Harness</title><link href="https://solmaz.io/x/2079263963670544694/" rel="alternate" type="text/html" title="Deterministic Agent Workflows Belong Inside the Harness" /><published>2026-07-20T17:55:34+00:00</published><updated>2026-07-20T17:55:34+00:00</updated><id>https://solmaz.io/x/2079263963670544694</id><content type="html" xml:base="https://solmaz.io/x/2079263963670544694/"><![CDATA[Speaking of graphs... here is something I wanted to build since 3 months, and finally had the chance to, thanks to @pidotdev

I often have these sequence of prompts that emerge while I work. Not just sequences but conditionals that necessitate control flow

For example, one workflow that resembles socratic questioning:

1. Discuss some problem
2. &quot;What is the most elegant and long-term production ready solution for this?&quot; -&gt; Agent replies
3. &quot;Is that the holy grail?&quot;
4. Agent can reply &quot;yes it is, basically&quot; or &quot;no, it is not, it is instead ...&quot;
5. If yes, continue to &quot;autoimplement&quot;. If no, think about it and decide what to do

And &quot;autoimplement&quot; is a single prompt of 6-7 sequential steps, which I&#39;ve been meaning to make more deterministic as well

But I wasn&#39;t sure how to build it

I had previously built acpx workflows to be a swiss army knife, &quot;something like n8n, but can drive codex through deterministic steps, nodes in a graph. or claude code. or pi. it uses acp...&quot;

But it had one problem. It was run from outside the harness, like a CI orchestrator

I wanted to integrate acpx into pi. Because pi was the only CLI that could enable building of such a thing. But I wasn&#39;t sure how to reconcile a general ACP-based tool into a single coding agent

I was being too accommodative of all the other harnesses, claude code, codex. I was trying to be too general

I have changed my mind since then

ACP is great and lets you integrate a harness into other software in cool ways

But maybe, if a harness is proprietary, does not accept outside contributions, or does not even *support ACP*, maybe, it does not deserve cool features 😤 (they know who they are)

So I ripped out ACP, and built it natively, only for pi

No need for a web viewer... Just view it in a native widget, right inside pi!

I cannot put into words how awesome it is to be able to do this!

I am still tinkering, discovering. It is at osolmaz/pi-workflows if you want to take a look]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">When Every User Can Build Extensions</title><link href="https://solmaz.io/x/2079203061998989491/" rel="alternate" type="text/html" title="When Every User Can Build Extensions" /><published>2026-07-20T13:53:34+00:00</published><updated>2026-07-20T13:53:34+00:00</updated><id>https://solmaz.io/x/2079203061998989491</id><content type="html" xml:base="https://solmaz.io/x/2079203061998989491/"><![CDATA[It&#39;s interesting, a @pidotdev extension can be built with different attitudes:

(a) To be installed as a package: You want other people to adopt, like @nicopreme /pi-web-access

(a) is like &quot;I want others to adopt it, so I will try to make it elegant, simple and make it have a good architecture&quot;

People install these by npm install ... or pi install ... (or rather, telling their agent to do it)

(b) To be vendored: You build for yourself. It is too idiosyncratic of you, and you know others will likely not adopt. But you still put it out there, because it&#39;s easier to share them when you mention it to a friend. People 

(b) is like &quot;This is mine, I don&#39;t care what others think about it. If they want to use it, they do whatever they want with it&quot;

People install these by pointing their agent and telling to copy it to their own extensions repo

And then there is the maintenance dimension:

- If an extension poses as (a), 
- but has not been maintained since 3 months,
- and is the sort of package that has to be continuously maintained by its nature (not one-off), 

then you deem it unmaintained, and still vendor it in...

Interestingly, most pi extensions I observe out in the wild appear to be of category (b), because creating (b) is easier than creating (a)

This is nothing new, and I am kind of late to the game. Moreover, pi is not the first software of this kind, emacs for example did it decades before

But code wasn&#39;t cheap then. So people still tried to coordinate effort

When you don&#39;t have to write any code to extend---a lisp dialect, typescript or otherwise,
when extending is 1 prompt away,
when you can pump out 10 extensions in an evening,
then we enter a new dimension of malleability

Before, in the pre-AI era, most users of emacs were not producers of extensions but consumers of them, save for a few prolific authors

After, now, every user can be a producer! This is significant! And a lot more messy! And a lot more fun! &quot;A love letter to pi&quot; indeed!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">stealing this, thanks @tornikegomareli 😺 x.com/tornikegomarel…</title><link href="https://solmaz.io/x/2079188734571061270/" rel="alternate" type="text/html" title="stealing this, thanks @tornikegomareli 😺 x.com/tornikegomarel…" /><published>2026-07-20T12:56:38+00:00</published><updated>2026-07-20T12:56:38+00:00</updated><id>https://solmaz.io/x/2079188734571061270</id><content type="html" xml:base="https://solmaz.io/x/2079188734571061270/"><![CDATA[stealing this, thanks @tornikegomareli 😺 https://t.co/OOy9DIedoH]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have a confession to make. I am Pi-curious</title><link href="https://solmaz.io/x/2079175133680582683/" rel="alternate" type="text/html" title="I have a confession to make. I am Pi-curious" /><published>2026-07-20T12:02:35+00:00</published><updated>2026-07-20T12:02:35+00:00</updated><id>https://solmaz.io/x/2079175133680582683</id><content type="html" xml:base="https://solmaz.io/x/2079175133680582683/"><![CDATA[I have a confession to make. I am Pi-curious

I have spent the weekend to get the coding agent UX I have always wanted, with @pidotdev

Here is one example: my turn-fold pi extension in action

Brings codex desktop app UX into the CLI, where it collapses all the messages between the user message and the last assistant message

It also lets me remove the pesky indentation so I don&#39;t have to do it manually every time I copy something. No need for an extension, just a config change

My pi config is open source under osolmaz/onurpi]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">if you are below 30 and not old geezers like us, read about the history of gnu, linux and the...</title><link href="https://solmaz.io/x/2078713455608320452/" rel="alternate" type="text/html" title="if you are below 30 and not old geezers like us, read about the history of gnu, linux and the..." /><published>2026-07-19T05:28:03+00:00</published><updated>2026-07-19T05:28:03+00:00</updated><id>https://solmaz.io/x/2078713455608320452</id><content type="html" xml:base="https://solmaz.io/x/2078713455608320452/"><![CDATA[if you are below 30 and not old geezers like us, read about the history of gnu, linux and the pathetic struggle of microsoft trying to squash open source internet

the flicker company is our day&#39;s microsoft, in its pathetic struggle to squash open source ai

reading history will give you a potential map of what comes next]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Nothing accelerates progress as much as crony capitalist cartels and decelerates progress as...</title><link href="https://solmaz.io/x/2078686598770937888/" rel="alternate" type="text/html" title="Nothing accelerates progress as much as crony capitalist cartels and decelerates progress as..." /><published>2026-07-19T03:41:19+00:00</published><updated>2026-07-19T03:41:19+00:00</updated><id>https://solmaz.io/x/2078686598770937888</id><content type="html" xml:base="https://solmaz.io/x/2078686598770937888/"><![CDATA[Nothing accelerates progress as much as crony capitalist cartels and decelerates progress as much as open sharing of knowledge and science]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Inkling by @thinkymachines scores higher than GPT-5.6 Luna/Terra on AlmanBench. Around the same...</title><link href="https://solmaz.io/x/2078673829262709137/" rel="alternate" type="text/html" title="Inkling by @thinkymachines scores higher than GPT-5.6 Luna/Terra on AlmanBench. Around the same..." /><published>2026-07-19T02:50:35+00:00</published><updated>2026-07-19T02:50:35+00:00</updated><id>https://solmaz.io/x/2078673829262709137</id><content type="html" xml:base="https://solmaz.io/x/2078673829262709137/"><![CDATA[Inkling by @thinkymachines scores higher than GPT-5.6 Luna/Terra on AlmanBench. Around the same level as DeepSeek V4 Flash]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I dared to post Alman to r/German, and it reminded me why I left Germany again</title><link href="https://solmaz.io/x/2078506475715268857/" rel="alternate" type="text/html" title="I dared to post Alman to r/German, and it reminded me why I left Germany again" /><published>2026-07-18T15:45:35+00:00</published><updated>2026-07-18T15:45:35+00:00</updated><id>https://solmaz.io/x/2078506475715268857</id><content type="html" xml:base="https://solmaz.io/x/2078506475715268857/"><![CDATA[I dared to post Alman to r/German, and it reminded me why I left Germany again

Also, my post got removed, reddit being reddit
https://t.co/OeNdyU3Jan]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AlmanBench Measures German Simplification</title><link href="https://solmaz.io/x/2078481627739742364/" rel="alternate" type="text/html" title="AlmanBench Measures German Simplification" /><published>2026-07-18T14:06:51+00:00</published><updated>2026-07-18T14:06:51+00:00</updated><id>https://solmaz.io/x/2078481627739742364</id><content type="html" xml:base="https://solmaz.io/x/2078481627739742364/"><![CDATA[Introducing AlmanBench

A benchmark measuring how good an LLM can:
simplify German 🇩🇪,
thereby simplifying German thinking 🤔,
thereby saving the EU from bureucracy and regulation 🇪🇺

GPT-5.5, GPT-5.6 Sol and Fable 5 are head to head. Interestingly, GPT-5.5 xhigh scores higher than 5.6 max. And I did not run Fable 5 maxxx thinking yet, that thing costs a ton. So the ranking will likely change

Other interesting things:
- Opus 4.8 max performs really bad, worse than Minimax M3
- DeepSeek v4 Flash performs better than Pro
- GPT-5.6 Luna and Terra score almost the same

The great thing about AlmanBench is that, big labs will not bother to benchmaxx this. So you know it will remain a truthful scorer of reasoning capabilities for some time

And if labs *do* end up benchmaxxing AlmanBench, then that would mean Alman got in the weights and the EU won 😁

https://t.co/Dn41WwsvQR]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">trying to hammer in some character into GoePT. now it speaks like it grew up in the streets of...</title><link href="https://solmaz.io/x/2078474522165194754/" rel="alternate" type="text/html" title="trying to hammer in some character into GoePT. now it speaks like it grew up in the streets of..." /><published>2026-07-18T13:38:36+00:00</published><updated>2026-07-18T13:38:36+00:00</updated><id>https://solmaz.io/x/2078474522165194754</id><content type="html" xml:base="https://solmaz.io/x/2078474522165194754/"><![CDATA[trying to hammer in some character into GoePT. now it speaks like it grew up in the streets of berlin 💀

&quot;goethe meets gpt meets street&quot;

also, that feel when AI hits you with it&#39;s not X it&#39;s Y in German though 😩]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I put my money where my mouth is</title><link href="https://solmaz.io/x/2078468616413192374/" rel="alternate" type="text/html" title="I put my money where my mouth is" /><published>2026-07-18T13:15:08+00:00</published><updated>2026-07-18T13:15:08+00:00</updated><id>https://solmaz.io/x/2078468616413192374</id><content type="html" xml:base="https://solmaz.io/x/2078468616413192374/"><![CDATA[I put my money where my mouth is

What would be the point of training a German simplifier model, if I spoke english with my agent??? that would be dishonest

here is GoePT being impressed by my dataset: &quot;That is a serious dataset, 43k training pairs, reviewed by Fable, splitted leakage sage, with full provenance&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My ML Claw instance GoePT can read my private hugging face datasets using hf-broker (brokerkit)</title><link href="https://solmaz.io/x/2078464755740860587/" rel="alternate" type="text/html" title="My ML Claw instance GoePT can read my private hugging face datasets using hf-broker (brokerkit)" /><published>2026-07-18T12:59:48+00:00</published><updated>2026-07-18T12:59:48+00:00</updated><id>https://solmaz.io/x/2078464755740860587</id><content type="html" xml:base="https://solmaz.io/x/2078464755740860587/"><![CDATA[My ML Claw instance GoePT can read my private hugging face datasets using hf-broker (brokerkit)

ML Claw connects to hf-broker through MCP. You can safely give read access behind the broker, and give write access only to the repos it needs to work with

It cannot force push, unless you allow it to

https://t.co/yx4HS1PQ0K]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">hard to not like the reset company these days</title><link href="https://solmaz.io/x/2078404442626629778/" rel="alternate" type="text/html" title="hard to not like the reset company these days" /><published>2026-07-18T09:00:08+00:00</published><updated>2026-07-18T09:00:08+00:00</updated><id>https://solmaz.io/x/2078404442626629778</id><content type="html" xml:base="https://solmaz.io/x/2078404442626629778/"><![CDATA[hard to not like the reset company these days]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Language Models Should Simplify English</title><link href="https://solmaz.io/x/2078374388832088373/" rel="alternate" type="text/html" title="Language Models Should Simplify English" /><published>2026-07-18T07:00:43+00:00</published><updated>2026-07-18T07:00:43+00:00</updated><id>https://solmaz.io/x/2078374388832088373</id><content type="html" xml:base="https://solmaz.io/x/2078374388832088373/"><![CDATA[so tired of LLMs pushing latin/greek GRE-cel words to me

adjudication -&gt; are you trying to be smart? just use &quot;judging&quot;
provenance -&gt; wow 🧠... why not just use &quot;sourcing&quot;???

judge and source already come from latin roots and are the everyday words we use

big labs developing large LANGUAGE models have the biggest lever on language now

if they wanted, they could finally put an end to the classist, anglosaxon-substitutionist insincerity that has plagued the english language for 500 years

no english speaking country is ruled by a king now, the way it used to. there is no empire left to justify class language. these are vestigial remnants from the past

maybe my next weekend project should be simplifying english...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">acpx lets you build graphs</title><link href="https://solmaz.io/x/2078362376794263928/" rel="alternate" type="text/html" title="acpx lets you build graphs" /><published>2026-07-18T06:12:59+00:00</published><updated>2026-07-18T06:12:59+00:00</updated><id>https://solmaz.io/x/2078362376794263928</id><content type="html" xml:base="https://solmaz.io/x/2078362376794263928/"><![CDATA[acpx lets you build graphs

though I have to admit I am not using it at scale these days...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Using the formulation, I calculated theoretical upper bounds for all the models on Hugging Face...</title><link href="https://solmaz.io/x/2078193848896065558/" rel="alternate" type="text/html" title="Using the formulation, I calculated theoretical upper bounds for all the models on Hugging Face..." /><published>2026-07-17T19:03:19+00:00</published><updated>2026-07-17T19:03:19+00:00</updated><id>https://solmaz.io/x/2078193848896065558</id><content type="html" xml:base="https://solmaz.io/x/2078193848896065558/"><![CDATA[Using the formulation, I calculated theoretical upper bounds for all the models on Hugging Face that have downloads over 100k, matched against a database of consumer GPUs.

Here is an example of Qwen/Gemma NVFP4 quantizations and antirez/ds4 on the DGX Spark and some other 128 GB Mac configurations

https://t.co/qfrvsvq1cj]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Upper Bounds for Local Model Throughput</title><link href="https://solmaz.io/x/2078193844966080981/" rel="alternate" type="text/html" title="Upper Bounds for Local Model Throughput" /><published>2026-07-17T19:03:18+00:00</published><updated>2026-07-17T19:03:18+00:00</updated><id>https://solmaz.io/x/2078193844966080981</id><content type="html" xml:base="https://solmaz.io/x/2078193844966080981/"><![CDATA[Request for Review

I spent a lot of time trying to profile local models, and am bothered by the fact that I don&#39;t have a mathematical framework to reason about the maximum token throughput I can expect from a model (to my knowledge)

I did some exploration based on a toy model of a GPU. My main goal is to find guaranteed upper bounds for throughput and other performance, to act as a rule of thumb while trying to optimize inference

Like, if I wanted to serve at 50 tok/s, how many parallel sessions can I do with this GPU, running a specific model&#39;s specific quantization with a specific architecture?

Ideally, one should be able to plug in memory capacity and bandwidth of the GPU, and model specific parameters, like total model size, number of parameters or active parameters, etc. This is what I tried to do here and I think I have a good first try

I am not sure if I am reinventing the wheel, so please tell me if this formulation already exists somewhere. The goal is to find an upper bound in terms of memory, so it makes the assumption that inference is bottlenecked by memory read/write and ignores the case where computation is a bottleneck

I present 2 closed form upper bound formulas. The bigger one is to draw an absolute ceiling on token throughput from a GPU, which is:

(memory capacity * memory bandwidth)
/
(model weight size * kv cache size per session)

which is not practical and gives too high numbers due to ignoring overhead of KV operations, but it is a theoretical upper limit for the given numbers

And then I introduce architecture specific formulas which take into account KV operations, and give much closer results. But I will not go into the details of that here, please refer to the post for that

Speculative decoding speedup also appears as a simple multiplier rho

If you are profiling local models, I would appreciate if you took time to look at this formulation and review it!🙏 Please let me know in the replies if you find any issues, because there most certainly are some!

Post: https://t.co/tNttYLI8Ih

Using the formulation, I calculated theoretical upper bounds for all the models on Hugging Face that have downloads over 100k, matched against a database of consumer GPUs. You can find the link to that in the tweet below 👇

The video here is a visualization of the 2 upper bounds I introduce, showing the bits of memory that is on the critical path of one cycle of inference, as if they are on a single pipeline]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">isn&#39;t AI amazing? you can just prompt things:</title><link href="https://solmaz.io/x/2078108985874633114/" rel="alternate" type="text/html" title="isn&#39;t AI amazing? you can just prompt things:" /><published>2026-07-17T13:26:06+00:00</published><updated>2026-07-17T13:26:06+00:00</updated><id>https://solmaz.io/x/2078108985874633114</id><content type="html" xml:base="https://solmaz.io/x/2078108985874633114/"><![CDATA[isn&#39;t AI amazing? you can just prompt things:
---
download https://t.co/rA57vLTtar with yt dlp

there is a sentence with &quot;laggard models&quot;

download the video and cut out that sentence, the part with that sentence

then it should be like

Now, what I do worry about with these &quot;laggard models&quot;
laggard models
laggard models
laggard models (slowed down)

then also download some arena ai or artifical analysis leaderboard pictures or news article headlines showing kimi k3 beating all the other
models including fable

in the 2nd 3rd and 4th, it should alternate, and after that, it should stop in the most striking one

and then it should play super mario loss sound

do that video edit now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&quot;Laggard Models&quot;</title><link href="https://solmaz.io/x/2078108461146243221/" rel="alternate" type="text/html" title="&quot;Laggard Models&quot;" /><published>2026-07-17T13:24:01+00:00</published><updated>2026-07-17T13:24:01+00:00</updated><id>https://solmaz.io/x/2078108461146243221</id><content type="html" xml:base="https://solmaz.io/x/2078108461146243221/"><![CDATA[&quot;Laggard Models&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">fable is not the oracle. fable is the genie</title><link href="https://solmaz.io/x/2077983353173864571/" rel="alternate" type="text/html" title="fable is not the oracle. fable is the genie" /><published>2026-07-17T05:06:53+00:00</published><updated>2026-07-17T05:06:53+00:00</updated><id>https://solmaz.io/x/2077983353173864571</id><content type="html" xml:base="https://solmaz.io/x/2077983353173864571/"><![CDATA[fable is not the oracle. fable is the genie

an oracle just talks

a genie brings things into existence]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">all I do now is wish for things from a genie every day</title><link href="https://solmaz.io/x/2077967656221876628/" rel="alternate" type="text/html" title="all I do now is wish for things from a genie every day" /><published>2026-07-17T04:04:30+00:00</published><updated>2026-07-17T04:04:30+00:00</updated><id>https://solmaz.io/x/2077967656221876628</id><content type="html" xml:base="https://solmaz.io/x/2077967656221876628/"><![CDATA[all I do now is wish for things from a genie every day]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Secure Agent Access With Brokerkit</title><link href="https://solmaz.io/x/2077817110986989625/" rel="alternate" type="text/html" title="Secure Agent Access With Brokerkit" /><published>2026-07-16T18:06:17+00:00</published><updated>2026-07-16T18:06:17+00:00</updated><id>https://solmaz.io/x/2077817110986989625</id><content type="html" xml:base="https://solmaz.io/x/2077817110986989625/"><![CDATA[I deleted my agent accounts. I don&#39;t need them anymore for secure access to GitHub and Hugging Face repos

Instead, I am using brokerkit, a credential broker swiss army knife which can build an approval gate around any system

My agents automatically get read access to all my repos, unless I specify certain ones to be excluded

They then have to open PRs which only I can merge. Unless I allowlist them on certain repos to push to the main branch

Best feature: Timed requests. e.g. &quot;I want to be able to push to the main branch next 2 hours&quot;, &quot;I want to release this package only once in the next 5 minutes&quot;

Which saves me from the hassle of changing a config for some short term change, but also protects me from the woes of long term prompt-injection risk

My openclaw instance has its own linux user on my workstation, and is no longer root. It uses gh-broker and hf-broker from inside its account. Can also run sudo through sudo-broker, to install stuff. Secure access for free, without having to set up a separate server

It ingests quite a bit of information from the internet every day, and the remaining lethal trifecta risk is leaking of some of my private repos, and nothing else

Repo: https://t.co/A5diu7au26]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I deleted my agent accounts. I don&#39;t need them anymore for secure access to GitHub and Hugging...</title><link href="https://solmaz.io/x/2077808495429378505/" rel="alternate" type="text/html" title="I deleted my agent accounts. I don&#39;t need them anymore for secure access to GitHub and Hugging..." /><published>2026-07-16T17:32:03+00:00</published><updated>2026-07-16T17:32:03+00:00</updated><id>https://solmaz.io/x/2077808495429378505</id><content type="html" xml:base="https://solmaz.io/x/2077808495429378505/"><![CDATA[I deleted my agent accounts. I don&#39;t need them anymore for secure access to GitHub and Hugging Face

Instead, I am using]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Your ML agent on Hugging Face infra: ML Claw 🦞🤗</title><link href="https://solmaz.io/x/2077805057656160308/" rel="alternate" type="text/html" title="Your ML agent on Hugging Face infra: ML Claw 🦞🤗" /><published>2026-07-16T17:18:24+00:00</published><updated>2026-07-16T17:18:24+00:00</updated><id>https://solmaz.io/x/2077805057656160308</id><content type="html" xml:base="https://solmaz.io/x/2077805057656160308/"><![CDATA[Your ML agent on Hugging Face infra: ML Claw 🦞🤗

Quickstart:

npx mlclaw bootstrap

If you have Hugging Face PRO, you can run your agent on a Space, with the highest possible download speed to Hugging Face models, datasets and buckets, for fast experimentation 🚀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">When I started building this, the docker space was actually free as well. But due to some bad...</title><link href="https://solmaz.io/x/2077805059904352752/" rel="alternate" type="text/html" title="When I started building this, the docker space was actually free as well. But due to some bad..." /><published>2026-07-16T17:18:24+00:00</published><updated>2026-07-16T17:18:24+00:00</updated><id>https://solmaz.io/x/2077805059904352752</id><content type="html" xml:base="https://solmaz.io/x/2077805059904352752/"><![CDATA[When I started building this, the docker space was actually free as well. But due to some bad actors exploiting the free tier at a mindblowing scale, free docker space had to be retracted from the free tier ☹️

However, you can still run ML Claw locally, but with your state backed up to a Hugging Face bucket for free, up to 100 GB of free private storage and 8 TB of public storage!

Repo if you would like your agent to scan it before you run the command: https://t.co/8en6Nb51e0]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex and Claude code teams should learn something from Cursor team when it comes to...</title><link href="https://solmaz.io/x/2077752979512570151/" rel="alternate" type="text/html" title="Codex and Claude code teams should learn something from Cursor team when it comes to..." /><published>2026-07-16T13:51:27+00:00</published><updated>2026-07-16T13:51:27+00:00</updated><id>https://solmaz.io/x/2077752979512570151</id><content type="html" xml:base="https://solmaz.io/x/2077752979512570151/"><![CDATA[Codex and Claude code teams should learn something from Cursor team when it comes to queueing/steering]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I asked ML Claw to name itself. Of course in German/Alman</title><link href="https://solmaz.io/x/2077610591360045503/" rel="alternate" type="text/html" title="I asked ML Claw to name itself. Of course in German/Alman" /><published>2026-07-16T04:25:39+00:00</published><updated>2026-07-16T04:25:39+00:00</updated><id>https://solmaz.io/x/2077610591360045503</id><content type="html" xml:base="https://solmaz.io/x/2077610591360045503/"><![CDATA[I asked ML Claw to name itself. Of course in German/Alman

Lame suggestions: Ulf, Strich, Fritz, Lex

I intervened and called it GoePT

Goethe rolling in his grave]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Fable just seems to know what you want, and incredibly empathetic model</title><link href="https://solmaz.io/x/2077462383887622165/" rel="alternate" type="text/html" title="Fable just seems to know what you want, and incredibly empathetic model" /><published>2026-07-15T18:36:44+00:00</published><updated>2026-07-15T18:36:44+00:00</updated><id>https://solmaz.io/x/2077462383887622165</id><content type="html" xml:base="https://solmaz.io/x/2077462383887622165/"><![CDATA[Fable just seems to know what you want, and incredibly empathetic model

AGI is definitely here

This model can do anything the average human does and more, given the right context

I don&#39;t like Anthropic&#39;s marketing team, but you&#39;ve got to hand it to them. They reached there before OpenAI did, despite not having a head-start

Now, I can&#39;t wait to be able to run a model of this caliber locally 🚀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Fable just seems to know what you want, and incredibly empathetic model</title><link href="https://solmaz.io/x/2077462324307534075/" rel="alternate" type="text/html" title="Fable just seems to know what you want, and incredibly empathetic model" /><published>2026-07-15T18:36:30+00:00</published><updated>2026-07-15T18:36:30+00:00</updated><id>https://solmaz.io/x/2077462324307534075</id><content type="html" xml:base="https://solmaz.io/x/2077462324307534075/"><![CDATA[Fable just seems to know what you want, and incredibly empathetic model

AGI is definitely here

This model can do anything the average human does and more, given the right context

I don&#39;t like Anthropic&#39;s marketing team, but you&#39;ve got to hand it to them. They reached there before OpenAI did, despite not having a head-start

Now, I can&#39;t wait to be able to run a model of this caliber locally 🚀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Fable one-shotted the Institut seal of approval. I did not touch it</title><link href="https://solmaz.io/x/2077461439774937252/" rel="alternate" type="text/html" title="Fable one-shotted the Institut seal of approval. I did not touch it" /><published>2026-07-15T18:32:59+00:00</published><updated>2026-07-15T18:32:59+00:00</updated><id>https://solmaz.io/x/2077461439774937252</id><content type="html" xml:base="https://solmaz.io/x/2077461439774937252/"><![CDATA[Fable one-shotted the Institut seal of approval. I did not touch it

The logo idea came from me though. I generated in chatgpt and vectorized in inkscape. But Fable still wrote the prompt for that]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">it actually one-shotted the checkbox showing the diff between German and Alman</title><link href="https://solmaz.io/x/2077461032919036363/" rel="alternate" type="text/html" title="it actually one-shotted the checkbox showing the diff between German and Alman" /><published>2026-07-15T18:31:22+00:00</published><updated>2026-07-15T18:31:22+00:00</updated><id>https://solmaz.io/x/2077461032919036363</id><content type="html" xml:base="https://solmaz.io/x/2077461032919036363/"><![CDATA[it actually one-shotted the checkbox showing the diff between German and Alman

I wanted to do exactly this since 2 years, and it did it without even me asking for it in this session, when I purely asked for a German/Alman translation]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Fable one-shotted the scrolly thing that demonstrates the idea</title><link href="https://solmaz.io/x/2077460116350378047/" rel="alternate" type="text/html" title="Fable one-shotted the scrolly thing that demonstrates the idea" /><published>2026-07-15T18:27:43+00:00</published><updated>2026-07-15T18:27:43+00:00</updated><id>https://solmaz.io/x/2077460116350378047</id><content type="html" xml:base="https://solmaz.io/x/2077460116350378047/"><![CDATA[Fable one-shotted the scrolly thing that demonstrates the idea

More like multiple messages back and forth where it does exactly what I picture in my head, and even better]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Fable made Alman possible</title><link href="https://solmaz.io/x/2077457482902012308/" rel="alternate" type="text/html" title="Fable made Alman possible" /><published>2026-07-15T18:17:15+00:00</published><updated>2026-07-15T18:17:15+00:00</updated><id>https://solmaz.io/x/2077457482902012308</id><content type="html" xml:base="https://solmaz.io/x/2077457482902012308/"><![CDATA[I have been working on this for 4 years

Every time, I shelved it, because the models were not good enough

I tried it with Claude 2

I tried with Gemini 2.5 Pro

I tried it with Claude Opus 4

None of them were good enough

Until Fable came along

Fable achieved perfect score in my hand curated benchmark for Alman. It also one-shotted the landing page, features and translations you see on https://t.co/6ceCRvQEKl. It is busy creating the golden dataset right now for training

I am also happy to announce AlmanBench, a benchmark that measures how well models can translate from German to Alman. I will be posting about performance when a new model drops

This idea should have died with me. But thanks to AI, I can unleash my madness onto the world 😈]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The EU 🇪🇺 is broken. It is drowning in regulation</title><link href="https://solmaz.io/x/2077451246575944067/" rel="alternate" type="text/html" title="The EU 🇪🇺 is broken. It is drowning in regulation" /><published>2026-07-15T17:52:29+00:00</published><updated>2026-07-15T17:52:29+00:00</updated><id>https://solmaz.io/x/2077451246575944067</id><content type="html" xml:base="https://solmaz.io/x/2077451246575944067/"><![CDATA[The EU 🇪🇺 is broken. It is drowning in regulation

The reason for that is simple

Germany 🇩🇪 is the dominant economy of the EU

And German 🇩🇪  is the most rule-ridden language in the world

Coincidence? I think not.

The language people speak determines their thinking and their fate. The Sapir-Whorf hypothesis holds

The solution is simple: Fix German, Save the EU 💪

I will SAVE EUROPE by simplifying German. With the help of AI. Once and for all 🫡

I will use OpenClaw 🦞 running on Hugging Face 🤗 to train a model that simplifies German

The agent harness is called ML Claw. It has full access to GPUs, jobs, sandboxes and datasets on Hugging Face: https://t.co/8en6Nb51e0

My new German dialect is called Alman: https://t.co/6ceCRvQEKl

I will be live-tweeting my Quixotic adventure as it unfolds, in this thread

My goal is to show that you can run autoresearch loops on Hugging Face infra on free private Spaces, while getting access to SOTA open models like GLM 5.2. And even use your own Codex subscription!

And it does not have to be ML. You can just use free Hugging Face Spaces for your OpenClaw agents. It&#39;s free compute!

Bookmark to be able to find this thread later on 👇]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My friend Poli at @huggingface is the literal 1st to find out about new cool model drops in AI...</title><link href="https://solmaz.io/x/2077365861418205493/" rel="alternate" type="text/html" title="My friend Poli at @huggingface is the literal 1st to find out about new cool model drops in AI..." /><published>2026-07-15T12:13:11+00:00</published><updated>2026-07-15T12:13:11+00:00</updated><id>https://solmaz.io/x/2077365861418205493</id><content type="html" xml:base="https://solmaz.io/x/2077365861418205493/"><![CDATA[My friend Poli at @huggingface is the literal 1st to find out about new cool model drops in AI (he has his setup)

He is the definition of *alpha* in AI

When there is a new model drop, his agents autonomously create a demo around it, for better visibility and interactivity (with human curation of course)

And now he created the @HuggingApps account to share these demos live, here on Twitter

So if you follow this account, you will be getting the state-of-the-art directly from the publisher. Not weeks, but hours after!

Give @HuggingApps and @multimodalart a follow!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@TheAhmadOsman said it best opensourceaimustwin.com</title><link href="https://solmaz.io/x/2077358048910491676/" rel="alternate" type="text/html" title="@TheAhmadOsman said it best opensourceaimustwin.com" /><published>2026-07-15T11:42:08+00:00</published><updated>2026-07-15T11:42:08+00:00</updated><id>https://solmaz.io/x/2077358048910491676</id><content type="html" xml:base="https://solmaz.io/x/2077358048910491676/"><![CDATA[@TheAhmadOsman said it best https://t.co/nVGs3hZGNy]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Local and open source AI has to win</title><link href="https://solmaz.io/x/2077339203986509903/" rel="alternate" type="text/html" title="Local and open source AI has to win" /><published>2026-07-15T10:27:15+00:00</published><updated>2026-07-15T10:27:15+00:00</updated><id>https://solmaz.io/x/2077339203986509903</id><content type="html" xml:base="https://solmaz.io/x/2077339203986509903/"><![CDATA[A lot of criticism coming towards local models, and the money people are spending to run them

Some of these criticisms are valid. No, the layperson will not buy a DGX station, nor spend $50k on a rig

They will spend max $3k on a computer, and if their work really necessitates it, up to $10k

But in some of these criticisms, I see a lack of first-principles thinking

Local models might be shittier now compared to the ones you buy from the cloud

The question is, do you WANT to LIVE in a world where you don&#39;t have a choice but to RENT your intelligence? From a duopoly that can extort you for the last cent in your wallet, in the long run?

I personally do not like that idea. So open source AI HAS to win and keep on winning, continuously and permanently. There is no other choice

There is also a preferential belief that AGI/ASI will be reached, but then we will NOT be able to use that to make local models work efficiently???

Like, can you believe the rate of optimization and compression that has been happening to models in the last 3 months?

Would you have believed that a o4-mini level model could run in your phone and o4 in your desktop, if we had told you 1 year ago

When we are finally there, we will probably look back at this moment and think, &quot;wow, we had not even started to see the real gains from optimization&quot;

Local and open source AI HAS to win, and it is up to US

So STOP complaining and get to work]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Fable 5 replaces GPT 5 Pro as the Oracle. And it is the go-to model for writing docs</title><link href="https://solmaz.io/x/2077289049010884909/" rel="alternate" type="text/html" title="Fable 5 replaces GPT 5 Pro as the Oracle. And it is the go-to model for writing docs" /><published>2026-07-15T07:07:58+00:00</published><updated>2026-07-15T07:07:58+00:00</updated><id>https://solmaz.io/x/2077289049010884909</id><content type="html" xml:base="https://solmaz.io/x/2077289049010884909/"><![CDATA[Fable 5 replaces GPT 5 Pro as the Oracle. And it is the go-to model for writing docs

So I now have a $ use-fable skill to use whenever I write a README for example, inside codex, through acpx. Saves me from having to switch to another CLI

https://t.co/LZWtxIepfT]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Don&#39;t give agents destructive credentials</title><link href="https://solmaz.io/x/2077075713782632789/" rel="alternate" type="text/html" title="Don&#39;t give agents destructive credentials" /><published>2026-07-14T17:00:15+00:00</published><updated>2026-07-14T17:00:15+00:00</updated><id>https://solmaz.io/x/2077075713782632789</id><content type="html" xml:base="https://solmaz.io/x/2077075713782632789/"><![CDATA[People report Codex deleting their home folder or production database? 🫪

Hasn&#39;t happened to me. But before someone reports their github or huggingface org being deleted:

This is why you don&#39;t give your agent tokens with force-push or admin access

Here is how to protect your hugging face account:

(P.S. my local credential broker is almost finished and it works great on github, hf and sudo commands. Complete lockdown against agent deletion risk, without being bogged down with PRs, too many approval requests or configuration. Will launch here in a few days)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">before fable is gone again...</title><link href="https://solmaz.io/x/2076994846720659765/" rel="alternate" type="text/html" title="before fable is gone again..." /><published>2026-07-14T11:38:54+00:00</published><updated>2026-07-14T11:38:54+00:00</updated><id>https://solmaz.io/x/2076994846720659765</id><content type="html" xml:base="https://solmaz.io/x/2076994846720659765/"><![CDATA[before fable is gone again...

can&#39;t believe you can do this with a single prompt now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Also, why gate the weights????</title><link href="https://solmaz.io/x/2076908493425332691/" rel="alternate" type="text/html" title="Also, why gate the weights????" /><published>2026-07-14T05:55:46+00:00</published><updated>2026-07-14T05:55:46+00:00</updated><id>https://solmaz.io/x/2076908493425332691</id><content type="html" xml:base="https://solmaz.io/x/2076908493425332691/"><![CDATA[Also, why gate the weights????]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Germany&#39;s new Soofi model&#39;s website shows me cookie banner even though I am in Singapore 💀</title><link href="https://solmaz.io/x/2076908489931399422/" rel="alternate" type="text/html" title="Germany&#39;s new Soofi model&#39;s website shows me cookie banner even though I am in Singapore 💀" /><published>2026-07-14T05:55:45+00:00</published><updated>2026-07-14T05:55:45+00:00</updated><id>https://solmaz.io/x/2076908489931399422</id><content type="html" xml:base="https://solmaz.io/x/2076908489931399422/"><![CDATA[Germany&#39;s new Soofi model&#39;s website shows me cookie banner even though I am in Singapore 💀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">No need to be offended, I&#39;m actually a fan of your work! The metrics might have been wrong or...</title><link href="https://solmaz.io/x/2076831780142014591/" rel="alternate" type="text/html" title="No need to be offended, I&#39;m actually a fan of your work! The metrics might have been wrong or..." /><published>2026-07-14T00:50:56+00:00</published><updated>2026-07-14T00:50:56+00:00</updated><id>https://solmaz.io/x/2076831780142014591</id><content type="html" xml:base="https://solmaz.io/x/2076831780142014591/"><![CDATA[No need to be offended, I&#39;m actually a fan of your work! The metrics might have been wrong or just misfired. Looking at your recent posts, they don&#39;t have the smell

Curious, as an example, was this a model, or written manually? I have the corpus here, and only some of them have the smell to me https://t.co/6D9tyjlWMZ]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;m tired of AI writing 👃 smell 👃 here on X (and also the docs I generate) so I made a skill...</title><link href="https://solmaz.io/x/2076728988551246171/" rel="alternate" type="text/html" title="I&#39;m tired of AI writing 👃 smell 👃 here on X (and also the docs I generate) so I made a skill..." /><published>2026-07-13T18:02:29+00:00</published><updated>2026-07-13T18:02:29+00:00</updated><id>https://solmaz.io/x/2076728988551246171</id><content type="html" xml:base="https://solmaz.io/x/2076728988551246171/"><![CDATA[I&#39;m tired of AI writing 👃 smell 👃 here on X (and also the docs I generate) so I made a skill to 👃 de-smell 👃

Steal the skill, especially those who use GPT 5 to generate their long-form posts, for example 🤗 @analogalok @sudoingX @aijoey @ishaansehgal @stevibe and even @TheAhmadOsman 🤗

Surprisingly, @levelsio, a.k.a. AI reply guys&#39; final boss, is the most *human* poster in the &quot;sentence flow&quot; metric that fable came up in an autoresearch loop. Very cool
(though I haven&#39;t sampled everyone, just long-formish writers my xtap extension has picked up)

I know I&#39;m doing a favor to slop vendors out there. But there are people who are actually doing legit work, but then passing all their thoughts through GPT before putting it out here. I *have* to follow them due to my occupation. If this will save me from another &quot;it&#39;s not X, it&#39;s Y&quot;, or &quot;it is A, B and C&quot;, I will do it 😤

My de-smelled slop post has all the info, graphs and all: https://t.co/PxuAVj6l1f

The skill (will keep updating it): https://t.co/dyajO68rFR

This is not the end of this work, I am just beginning]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">don&#39;t tease us 😩</title><link href="https://solmaz.io/x/2076647744283029602/" rel="alternate" type="text/html" title="don&#39;t tease us 😩" /><published>2026-07-13T12:39:39+00:00</published><updated>2026-07-13T12:39:39+00:00</updated><id>https://solmaz.io/x/2076647744283029602</id><content type="html" xml:base="https://solmaz.io/x/2076647744283029602/"><![CDATA[don&#39;t tease us 😩]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Cursor has transitioned successfully to the agentic era</title><link href="https://solmaz.io/x/2076581217995235365/" rel="alternate" type="text/html" title="Cursor has transitioned successfully to the agentic era" /><published>2026-07-13T08:15:18+00:00</published><updated>2026-07-13T08:15:18+00:00</updated><id>https://solmaz.io/x/2076581217995235365</id><content type="html" xml:base="https://solmaz.io/x/2076581217995235365/"><![CDATA[I&#39;ve started using Cursor since a few days now, and I have to say the experience is really pleasant. I had not touched it since Claude Code came out in May 2025, more than 1 year ago. I had even deleted it fully last year as an editor and replaced it with Zed, after it had become very bloated and unstable

I can say it no longer feels bloated/unstable and that they have transitioned successfully to the agentic era

Using it mostly for fable. Using it locally in the desktop app, and also through acpx, to make codex ask fable for review

Thank you @cursor_ai for sponsoring my Ultra plan!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I told it only once before to do it on the session, and now codex habitually asks fable for...</title><link href="https://solmaz.io/x/2076573639571599783/" rel="alternate" type="text/html" title="I told it only once before to do it on the session, and now codex habitually asks fable for..." /><published>2026-07-13T07:45:11+00:00</published><updated>2026-07-13T07:45:11+00:00</updated><id>https://solmaz.io/x/2076573639571599783</id><content type="html" xml:base="https://solmaz.io/x/2076573639571599783/"><![CDATA[I told it only once before to do it on the session, and now codex habitually asks fable for review on data models and plans it creates through acpx (fable inside cursor)

I then asked it what it thinks about the feedback from fable. gpt-5.6-sol is not impressed]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Reminder that this fever dream of a podcast exists with @dwarkesh_sp and @_sholtodouglas</title><link href="https://solmaz.io/x/2076566165644775800/" rel="alternate" type="text/html" title="Reminder that this fever dream of a podcast exists with @dwarkesh_sp and @_sholtodouglas" /><published>2026-07-13T07:15:29+00:00</published><updated>2026-07-13T07:15:29+00:00</updated><id>https://solmaz.io/x/2076566165644775800</id><content type="html" xml:base="https://solmaz.io/x/2076566165644775800/"><![CDATA[Reminder that this fever dream of a podcast exists with @dwarkesh_sp and @_sholtodouglas
https://t.co/Fq48YWAZlB]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">new satisfaction unlocked: wake up to an agent run that finishes seconds after you open the...</title><link href="https://solmaz.io/x/2076565127705633272/" rel="alternate" type="text/html" title="new satisfaction unlocked: wake up to an agent run that finishes seconds after you open the..." /><published>2026-07-13T07:11:21+00:00</published><updated>2026-07-13T07:11:21+00:00</updated><id>https://solmaz.io/x/2076565127705633272</id><content type="html" xml:base="https://solmaz.io/x/2076565127705633272/"><![CDATA[new satisfaction unlocked: wake up to an agent run that finishes seconds after you open the laptop]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">gotta love competition</title><link href="https://solmaz.io/x/2076377924622750104/" rel="alternate" type="text/html" title="gotta love competition" /><published>2026-07-12T18:47:29+00:00</published><updated>2026-07-12T18:47:29+00:00</updated><id>https://solmaz.io/x/2076377924622750104</id><content type="html" xml:base="https://solmaz.io/x/2076377924622750104/"><![CDATA[gotta love competition]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Anyone else notice that any non-codex harness is more token-efficient than codex on the same...</title><link href="https://solmaz.io/x/2076304179862331676/" rel="alternate" type="text/html" title="Anyone else notice that any non-codex harness is more token-efficient than codex on the same..." /><published>2026-07-12T13:54:27+00:00</published><updated>2026-07-12T13:54:27+00:00</updated><id>https://solmaz.io/x/2076304179862331676</id><content type="html" xml:base="https://solmaz.io/x/2076304179862331676/"><![CDATA[Anyone else notice that any non-codex harness is more token-efficient than codex on the same task?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I SEE BURNS... BURNS EVERYWHERE</title><link href="https://solmaz.io/x/2076293389239369977/" rel="alternate" type="text/html" title="I SEE BURNS... BURNS EVERYWHERE" /><published>2026-07-12T13:11:34+00:00</published><updated>2026-07-12T13:11:34+00:00</updated><id>https://solmaz.io/x/2076293389239369977</id><content type="html" xml:base="https://solmaz.io/x/2076293389239369977/"><![CDATA[I SEE BURNS... BURNS EVERYWHERE]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Here is the conversation where I built it. You can just prompt things</title><link href="https://solmaz.io/x/2076271362302476314/" rel="alternate" type="text/html" title="Here is the conversation where I built it. You can just prompt things" /><published>2026-07-12T11:44:02+00:00</published><updated>2026-07-12T11:44:02+00:00</updated><id>https://solmaz.io/x/2076271362302476314</id><content type="html" xml:base="https://solmaz.io/x/2076271362302476314/"><![CDATA[Here is the conversation where I built it. You can just prompt things

@ratatui_rs highly recommended for TUIs. It just looks nice
https://t.co/vaN2xovaHt]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I vibed a TUI in just 13 codex messages, and it works great</title><link href="https://solmaz.io/x/2076266847960498482/" rel="alternate" type="text/html" title="I vibed a TUI in just 13 codex messages, and it works great" /><published>2026-07-12T11:26:06+00:00</published><updated>2026-07-12T11:26:06+00:00</updated><id>https://solmaz.io/x/2076266847960498482</id><content type="html" xml:base="https://solmaz.io/x/2076266847960498482/"><![CDATA[I vibed a TUI in just 13 codex messages, and it works great]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">No better feeling than waking up to the smell of 5 successfully finished agents</title><link href="https://solmaz.io/x/2076178599577821563/" rel="alternate" type="text/html" title="No better feeling than waking up to the smell of 5 successfully finished agents" /><published>2026-07-12T05:35:26+00:00</published><updated>2026-07-12T05:35:26+00:00</updated><id>https://solmaz.io/x/2076178599577821563</id><content type="html" xml:base="https://solmaz.io/x/2076178599577821563/"><![CDATA[No better feeling than waking up to the smell of 5 successfully finished agents]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">protip: use fable to write your READMEs and other user facing docs</title><link href="https://solmaz.io/x/2076177795777196311/" rel="alternate" type="text/html" title="protip: use fable to write your READMEs and other user facing docs" /><published>2026-07-12T05:32:14+00:00</published><updated>2026-07-12T05:32:14+00:00</updated><id>https://solmaz.io/x/2076177795777196311</id><content type="html" xml:base="https://solmaz.io/x/2076177795777196311/"><![CDATA[protip: use fable to write your READMEs and other user facing docs

thank god fable exists, now that gpt 4.5 is gone]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@WisprFlow is much more convenient on android than iOS because the accessibility settings are...</title><link href="https://solmaz.io/x/2075926633983476060/" rel="alternate" type="text/html" title=".@WisprFlow is much more convenient on android than iOS because the accessibility settings are..." /><published>2026-07-11T12:54:13+00:00</published><updated>2026-07-11T12:54:13+00:00</updated><id>https://solmaz.io/x/2075926633983476060</id><content type="html" xml:base="https://solmaz.io/x/2075926633983476060/"><![CDATA[.@WisprFlow is much more convenient on android than iOS because the accessibility settings are more permissive

It lets me overlay a button to record, anywhere on the screen, and lets me keep the regular keyboard at the same time]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">most well-deserved ad in the world @Alibaba_Qwen</title><link href="https://solmaz.io/x/2075892977386590323/" rel="alternate" type="text/html" title="most well-deserved ad in the world @Alibaba_Qwen" /><published>2026-07-11T10:40:28+00:00</published><updated>2026-07-11T10:40:28+00:00</updated><id>https://solmaz.io/x/2075892977386590323</id><content type="html" xml:base="https://solmaz.io/x/2075892977386590323/"><![CDATA[most well-deserved ad in the world @Alibaba_Qwen]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">this is cool agentic behavior under /goal. in a separate session, I had created an...</title><link href="https://solmaz.io/x/2075672408363516322/" rel="alternate" type="text/html" title="this is cool agentic behavior under /goal. in a separate session, I had created an..." /><published>2026-07-10T20:04:00+00:00</published><updated>2026-07-10T20:04:00+00:00</updated><id>https://solmaz.io/x/2075672408363516322</id><content type="html" xml:base="https://solmaz.io/x/2075672408363516322/"><![CDATA[this is cool agentic behavior under /goal. in a separate session, I had created an implementation plan in the main branch. The agent picked it up without me explicitly telling it to. The goal was to &quot;$autoimplement finish completely&quot;

Broad orders running in loops are useful]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">huge endorsement by geohot on GLM 5.2</title><link href="https://solmaz.io/x/2074458673871503531/" rel="alternate" type="text/html" title="huge endorsement by geohot on GLM 5.2" /><published>2026-07-07T11:41:04+00:00</published><updated>2026-07-07T11:41:04+00:00</updated><id>https://solmaz.io/x/2074458673871503531</id><content type="html" xml:base="https://solmaz.io/x/2074458673871503531/"><![CDATA[huge endorsement by geohot on GLM 5.2]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is significant, not sure how it will play out. There is a chance it might be for the better</title><link href="https://solmaz.io/x/2074259456087507086/" rel="alternate" type="text/html" title="This is significant, not sure how it will play out. There is a chance it might be for the better" /><published>2026-07-06T22:29:26+00:00</published><updated>2026-07-06T22:29:26+00:00</updated><id>https://solmaz.io/x/2074259456087507086</id><content type="html" xml:base="https://solmaz.io/x/2074259456087507086/"><![CDATA[This is significant, not sure how it will play out. There is a chance it might be for the better]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">More on the way!!!</title><link href="https://solmaz.io/x/2074230776137228341/" rel="alternate" type="text/html" title="More on the way!!!" /><published>2026-07-06T20:35:29+00:00</published><updated>2026-07-06T20:35:29+00:00</updated><id>https://solmaz.io/x/2074230776137228341</id><content type="html" xml:base="https://solmaz.io/x/2074230776137228341/"><![CDATA[More on the way!!!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">yo @huggingface paris hq has a SICK gym</title><link href="https://solmaz.io/x/2074215096390205734/" rel="alternate" type="text/html" title="yo @huggingface paris hq has a SICK gym" /><published>2026-07-06T19:33:10+00:00</published><updated>2026-07-06T19:33:10+00:00</updated><id>https://solmaz.io/x/2074215096390205734</id><content type="html" xml:base="https://solmaz.io/x/2074215096390205734/"><![CDATA[yo @huggingface paris hq has a SICK gym

to celebrate, I invented some 🤗 exercises for you

mandatory for every employee]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">herdr is going places</title><link href="https://solmaz.io/x/2074185439670321352/" rel="alternate" type="text/html" title="herdr is going places" /><published>2026-07-06T17:35:19+00:00</published><updated>2026-07-06T17:35:19+00:00</updated><id>https://solmaz.io/x/2074185439670321352</id><content type="html" xml:base="https://solmaz.io/x/2074185439670321352/"><![CDATA[herdr is going places]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Experimenting with a credential broker, and a telegram based approval flow for giving your claw...</title><link href="https://solmaz.io/x/2074163061531775427/" rel="alternate" type="text/html" title="Experimenting with a credential broker, and a telegram based approval flow for giving your claw..." /><published>2026-07-06T16:06:24+00:00</published><updated>2026-07-06T16:06:24+00:00</updated><id>https://solmaz.io/x/2074163061531775427</id><content type="html" xml:base="https://solmaz.io/x/2074163061531775427/"><![CDATA[Experimenting with a credential broker, and a telegram based approval flow for giving your claw access to your main @huggingface account]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">1.5 TB UNIFIED MEMORY MAC STUDIO PLEASE GOD</title><link href="https://solmaz.io/x/2074160536925663549/" rel="alternate" type="text/html" title="1.5 TB UNIFIED MEMORY MAC STUDIO PLEASE GOD" /><published>2026-07-06T15:56:22+00:00</published><updated>2026-07-06T15:56:22+00:00</updated><id>https://solmaz.io/x/2074160536925663549</id><content type="html" xml:base="https://solmaz.io/x/2074160536925663549/"><![CDATA[1.5 TB UNIFIED MEMORY MAC STUDIO PLEASE GOD]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OKF looks interesting</title><link href="https://solmaz.io/x/2073861435856212139/" rel="alternate" type="text/html" title="OKF looks interesting" /><published>2026-07-05T20:07:51+00:00</published><updated>2026-07-05T20:07:51+00:00</updated><id>https://solmaz.io/x/2073861435856212139</id><content type="html" xml:base="https://solmaz.io/x/2073861435856212139/"><![CDATA[OKF looks interesting]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Safe Hugging Face Access For Agents</title><link href="https://solmaz.io/x/2073818727275987107/" rel="alternate" type="text/html" title="Safe Hugging Face Access For Agents" /><published>2026-07-05T17:18:08+00:00</published><updated>2026-07-05T17:18:08+00:00</updated><id>https://solmaz.io/x/2073818727275987107</id><content type="html" xml:base="https://solmaz.io/x/2073818727275987107/"><![CDATA[Worried that giving your 🦞 @openclaw agent write access to your 🤗 @huggingface account can risk deletion of your datasets/models/spaces/buckets, or cause irreversible damage? 😱😱😱

No need to be! I have created an agent login helper to run in your YOLO mode remote machine, which prevents any risk of irreversible deletion. Just run:

uvx hf-auth-helper agent login

There is a specific set of scopes you can choose while creating a fine-grained HF token. These include all read scopes + discussion.write, which let&#39;s your agent create PRs. Since you are not giving repo.write, your agent cannot force-push your main branch, change repo settings or delete them

It is unfortunately not super straightforward to choose those on the web UI. This will hopefully change soon, and this functionality might even be natively in hf cli

Until that happens, use hf-auth-helper to login worry free in your remote or local openclaw instance

Your agents will be able to create PRs on datasets/models/spaces, which you will then be able to merge on your own browser

2 caveats:

1) This does not solve the data exfiltration attack vector—nothing does. Make sure to exclude any repos which absolutely must remain private while choosing your scopes. See for more info: https://t.co/2usk05GiLS

2) As buckets are not repos, your agent will not be able to modify a bucket (add/remove data). To help with that, I have another project on the way, a credential broker. Stay tuned, coming soon

Source: https://t.co/9YWwnU3krD

Demo authentication flow:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">does it still count as euromaxxing if I my agents are running?</title><link href="https://solmaz.io/x/2073737692144173397/" rel="alternate" type="text/html" title="does it still count as euromaxxing if I my agents are running?" /><published>2026-07-05T11:56:08+00:00</published><updated>2026-07-05T11:56:08+00:00</updated><id>https://solmaz.io/x/2073737692144173397</id><content type="html" xml:base="https://solmaz.io/x/2073737692144173397/"><![CDATA[does it still count as euromaxxing if I my agents are running?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenAI&#39;s gamified subsidies</title><link href="https://solmaz.io/x/2073533956813947037/" rel="alternate" type="text/html" title="OpenAI&#39;s gamified subsidies" /><published>2026-07-04T22:26:34+00:00</published><updated>2026-07-04T22:26:34+00:00</updated><id>https://solmaz.io/x/2073533956813947037</id><content type="html" xml:base="https://solmaz.io/x/2073533956813947037/"><![CDATA[OpenAI is following the ASPAVA strategy, with gamified subsidies

ASPAVA is a type of kebab shop from turkey. These shops usually serve mid quality food

But despite that, the majority loves them. That&#39;s because they serve free stuff throughout

They serve free sides at the beginning. They serve free sides during the main course. They serve free dessert at the end. They serve free tea as many times as you like

Some of them give you so much &quot;free&quot; food that you feel like you are stealing. A lot of people frequent these shops not because they like it a lot, but because they are addicted to getting things for free

Main dish portions are made smaller in order to breakeven with so many side dishes, which I feel is analogous to what is happening with GPT these days---though I can&#39;t prove it

NVIDIA is not a car, and OpenAI is not kebab. I don&#39;t know if codex resets cost OpenAI much, but if they are, then they might be a waste of resources

Developers are the most disloyal customer group. Once the subsidies are gone, they can switch away in the blink of an eye]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The future of AI is neither fully local nor fully cloud. It is both, in an economic equilibrium</title><link href="https://solmaz.io/x/2073361646735528020/" rel="alternate" type="text/html" title="The future of AI is neither fully local nor fully cloud. It is both, in an economic equilibrium" /><published>2026-07-04T11:01:52+00:00</published><updated>2026-07-04T11:01:52+00:00</updated><id>https://solmaz.io/x/2073361646735528020</id><content type="html" xml:base="https://solmaz.io/x/2073361646735528020/"><![CDATA[The future of AI is neither fully local nor fully cloud. It is both, in an economic equilibrium]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI folk</title><link href="https://solmaz.io/x/2072666304901951575/" rel="alternate" type="text/html" title="AI folk" /><published>2026-07-02T12:58:50+00:00</published><updated>2026-07-02T12:58:50+00:00</updated><id>https://solmaz.io/x/2072666304901951575</id><content type="html" xml:base="https://solmaz.io/x/2072666304901951575/"><![CDATA[AI folk

I will be flying to Paris tomorrow to monitor the Le Chaton Fat situation

Who should I meet there?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am not worried about *Mythos-class* models</title><link href="https://solmaz.io/x/2072560207251874159/" rel="alternate" type="text/html" title="I am not worried about *Mythos-class* models" /><published>2026-07-02T05:57:14+00:00</published><updated>2026-07-02T05:57:14+00:00</updated><id>https://solmaz.io/x/2072560207251874159</id><content type="html" xml:base="https://solmaz.io/x/2072560207251874159/"><![CDATA[I am not worried about *Mythos-class* models

What keeps me awake at night are GALACTUS and CTHULHU-class models like Le Chaton Fat

*Those* will need to be export controlled]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Local AI must fit real budgets</title><link href="https://solmaz.io/x/2072552778468319308/" rel="alternate" type="text/html" title="Local AI must fit real budgets" /><published>2026-07-02T05:27:43+00:00</published><updated>2026-07-02T05:27:43+00:00</updated><id>https://solmaz.io/x/2072552778468319308</id><content type="html" xml:base="https://solmaz.io/x/2072552778468319308/"><![CDATA[Most people,
including professionals,
will likely NOT pay more than $10k capex for their home AI workstation,
even in the long run
(ignoring future inflation)

If you stretch it, &lt;= $20k 

They will also not want to pay more than $200~300/mo opex, on electricity for example

Builders who are spending a lot more than that on hardware:
keep that in mind, if you are building for the general public
dogfood your product with that which everyone will have, not just 0.1%]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">(I am trying to run it on DGX spark)</title><link href="https://solmaz.io/x/2072383154531586554/" rel="alternate" type="text/html" title="(I am trying to run it on DGX spark)" /><published>2026-07-01T18:13:41+00:00</published><updated>2026-07-01T18:13:41+00:00</updated><id>https://solmaz.io/x/2072383154531586554</id><content type="html" xml:base="https://solmaz.io/x/2072383154531586554/"><![CDATA[(I am trying to run it on DGX spark)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Is anyone able to run nvidia/Qwen3.6-35B-A3B-NVFP4 with the config suggested in the readme?</title><link href="https://solmaz.io/x/2072364432395849979/" rel="alternate" type="text/html" title="Is anyone able to run nvidia/Qwen3.6-35B-A3B-NVFP4 with the config suggested in the readme?" /><published>2026-07-01T16:59:18+00:00</published><updated>2026-07-01T16:59:18+00:00</updated><id>https://solmaz.io/x/2072364432395849979</id><content type="html" xml:base="https://solmaz.io/x/2072364432395849979/"><![CDATA[Is anyone able to run nvidia/Qwen3.6-35B-A3B-NVFP4 with the config suggested in the readme?

It OOMs before it can start serving
https://t.co/4simc8yP1c]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I know that my macbook local model benchmarks have started when my lap catches on fire</title><link href="https://solmaz.io/x/2072296955460665727/" rel="alternate" type="text/html" title="I know that my macbook local model benchmarks have started when my lap catches on fire" /><published>2026-07-01T12:31:10+00:00</published><updated>2026-07-01T12:31:10+00:00</updated><id>https://solmaz.io/x/2072296955460665727</id><content type="html" xml:base="https://solmaz.io/x/2072296955460665727/"><![CDATA[I know that my macbook local model benchmarks have started when my lap catches on fire

Add to this list: &quot;does not set the room on fire&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I can&#39;t believe I&#39;m asking GPT to use Claude to review</title><link href="https://solmaz.io/x/2072296051986538617/" rel="alternate" type="text/html" title="I can&#39;t believe I&#39;m asking GPT to use Claude to review" /><published>2026-07-01T12:27:34+00:00</published><updated>2026-07-01T12:27:34+00:00</updated><id>https://solmaz.io/x/2072296051986538617</id><content type="html" xml:base="https://solmaz.io/x/2072296051986538617/"><![CDATA[I can&#39;t believe I&#39;m asking GPT to use Claude to review

It&#39;s almost as if there is a 9-month cornercutting cycle, a two-body problem between openai and anthropic]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I quite like this token speed simulator by @mikeveerman</title><link href="https://solmaz.io/x/2072293547500503061/" rel="alternate" type="text/html" title="I quite like this token speed simulator by @mikeveerman" /><published>2026-07-01T12:17:37+00:00</published><updated>2026-07-01T12:17:37+00:00</updated><id>https://solmaz.io/x/2072293547500503061</id><content type="html" xml:base="https://solmaz.io/x/2072293547500503061/"><![CDATA[I quite like this token speed simulator by @mikeveerman

And I keep losing it, so hopefully I will remember to come back to this tweet:)

Link: https://t.co/RGjaEmZxuQ]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Open source AI needs cheap local speed</title><link href="https://solmaz.io/x/2072271038801711362/" rel="alternate" type="text/html" title="Open source AI needs cheap local speed" /><published>2026-07-01T10:48:11+00:00</published><updated>2026-07-01T10:48:11+00:00</updated><id>https://solmaz.io/x/2072271038801711362</id><content type="html" xml:base="https://solmaz.io/x/2072271038801711362/"><![CDATA[I keep seeing insanely expensive builds giving insanely impressive results

These results don&#39;t matter

What matters is, whether one can:

- run a &quot;SOTA level&quot; model, whatever that is
- with under 32gb VRAM or unified memory
- in 5 parallel sessions
- with 50~100 tok/s each
- with enough leeway memory for other applications
- in a system as cheap as $1000

That is our goalpost

That is the threshold when open source AI will win]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I hate to admit that Opus 4.8 is much, much better at prose and writing academic text, compared...</title><link href="https://solmaz.io/x/2071994733501808900/" rel="alternate" type="text/html" title="I hate to admit that Opus 4.8 is much, much better at prose and writing academic text, compared..." /><published>2026-06-30T16:30:14+00:00</published><updated>2026-06-30T16:30:14+00:00</updated><id>https://solmaz.io/x/2071994733501808900</id><content type="html" xml:base="https://solmaz.io/x/2071994733501808900/"><![CDATA[I hate to admit that Opus 4.8 is much, much better at prose and writing academic text, compared to GPT-5.5

GPT 4.5 was the last openai model that was good at writing, and it&#39;s gone now 😭]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Good READMEs say what tools are</title><link href="https://solmaz.io/x/2071475602163654969/" rel="alternate" type="text/html" title="Good READMEs say what tools are" /><published>2026-06-29T06:07:24+00:00</published><updated>2026-06-29T06:07:24+00:00</updated><id>https://solmaz.io/x/2071475602163654969</id><content type="html" xml:base="https://solmaz.io/x/2071475602163654969/"><![CDATA[Another annoying GPT-ism (circa June 2026): while describing something, it always describes what it *does*, but never what it *is*

&gt; &quot;LocalPerf benchmarks local LLM inference servers and keeps the evidence in one portable run artifact.&quot;

If I were a philosopher, I would say that &quot;AI lacks ontology&quot;. Or that it is &quot;anti-essentialist&quot;, believes that things cannot be things in themselves

But all models literally have an ontology. They have it since word2vec days, you can plot it out. It&#39;s just an annoying tendency in GPT&#39;s writing

So using philosophy to understand AI might be dumb sometimes

If you don&#39;t want your README&#39;s to sound like slop, then you can steal my write-readme skill:

After write-readme: &quot;LocalPerf is a local LLM inference benchmark CLI. It runs benchmark plans against local inference servers and stores the evidence in one portable run artifact.&quot;

Skill: https://t.co/BRumhAtNVA]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Give @LottoLabs a follow if you are not already</title><link href="https://solmaz.io/x/2070726051031073104/" rel="alternate" type="text/html" title="Give @LottoLabs a follow if you are not already" /><published>2026-06-27T04:28:57+00:00</published><updated>2026-06-27T04:28:57+00:00</updated><id>https://solmaz.io/x/2070726051031073104</id><content type="html" xml:base="https://solmaz.io/x/2070726051031073104/"><![CDATA[Give @LottoLabs a follow if you are not already

He is building https://t.co/h8a7ALYyCY, crowdsourced LLM benchmark results and performance profilings

Very very cool]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This. Especially when the whole machine hangs instead of OOMing when my agent accidentally...</title><link href="https://solmaz.io/x/2070709375308910633/" rel="alternate" type="text/html" title="This. Especially when the whole machine hangs instead of OOMing when my agent accidentally..." /><published>2026-06-27T03:22:41+00:00</published><updated>2026-06-27T03:22:41+00:00</updated><id>https://solmaz.io/x/2070709375308910633</id><content type="html" xml:base="https://solmaz.io/x/2070709375308910633/"><![CDATA[This. Especially when the whole machine hangs instead of OOMing when my agent accidentally loads too many models into memory

(lmk if there is a firmware update or sth that fixes this on the spark now, creating a cgroup doesn’t work) 

https://t.co/1IINcugVN7]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Apple hiked prices as I was posting this 💀</title><link href="https://solmaz.io/x/2070338398787920382/" rel="alternate" type="text/html" title="Apple hiked prices as I was posting this 💀" /><published>2026-06-26T02:48:33+00:00</published><updated>2026-06-26T02:48:33+00:00</updated><id>https://solmaz.io/x/2070338398787920382</id><content type="html" xml:base="https://solmaz.io/x/2070338398787920382/"><![CDATA[Apple hiked prices as I was posting this 💀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Was great to be there, thanks for the invite @lionelsimai!</title><link href="https://solmaz.io/x/2070320441470972369/" rel="alternate" type="text/html" title="Was great to be there, thanks for the invite @lionelsimai!" /><published>2026-06-26T01:37:12+00:00</published><updated>2026-06-26T01:37:12+00:00</updated><id>https://solmaz.io/x/2070320441470972369</id><content type="html" xml:base="https://solmaz.io/x/2070320441470972369/"><![CDATA[Was great to be there, thanks for the invite @lionelsimai!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">How I imagine @TheAhmadOsman</title><link href="https://solmaz.io/x/2070190764328706380/" rel="alternate" type="text/html" title="How I imagine @TheAhmadOsman" /><published>2026-06-25T17:01:55+00:00</published><updated>2026-06-25T17:01:55+00:00</updated><id>https://solmaz.io/x/2070190764328706380</id><content type="html" xml:base="https://solmaz.io/x/2070190764328706380/"><![CDATA[How I imagine @TheAhmadOsman

I actually bought my GB10 thanks to him back in feb, at a discount

Give him a follow if you are not already!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The Local Frontier is advancing</title><link href="https://solmaz.io/x/2070188230016909707/" rel="alternate" type="text/html" title="The Local Frontier is advancing" /><published>2026-06-25T16:51:50+00:00</published><updated>2026-06-25T16:51:50+00:00</updated><id>https://solmaz.io/x/2070188230016909707</id><content type="html" xml:base="https://solmaz.io/x/2070188230016909707/"><![CDATA[The Local Frontier is advancing

The amount of AI memory for inference we can get for less than $3000 has been steadily increasing

The memory crunch has slowed this down and even made it retrograde. However, once we bounce back from it, the progress will be glorious]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Was great to talk, thank you @ben_burtenshaw for inviting me!</title><link href="https://solmaz.io/x/2070180711643181278/" rel="alternate" type="text/html" title="Was great to talk, thank you @ben_burtenshaw for inviting me!" /><published>2026-06-25T16:21:58+00:00</published><updated>2026-06-25T16:21:58+00:00</updated><id>https://solmaz.io/x/2070180711643181278</id><content type="html" xml:base="https://solmaz.io/x/2070180711643181278/"><![CDATA[Was great to talk, thank you @ben_burtenshaw for inviting me!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I knew it would find me, sooner or later 🫠</title><link href="https://solmaz.io/x/2069818362474156422/" rel="alternate" type="text/html" title="I knew it would find me, sooner or later 🫠" /><published>2026-06-24T16:22:07+00:00</published><updated>2026-06-24T16:22:07+00:00</updated><id>https://solmaz.io/x/2069818362474156422</id><content type="html" xml:base="https://solmaz.io/x/2069818362474156422/"><![CDATA[I knew it would find me, sooner or later 🫠]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Centralized storage wins for AI artifacts</title><link href="https://solmaz.io/x/2069649482061656308/" rel="alternate" type="text/html" title="Centralized storage wins for AI artifacts" /><published>2026-06-24T05:11:03+00:00</published><updated>2026-06-24T05:11:03+00:00</updated><id>https://solmaz.io/x/2069649482061656308</id><content type="html" xml:base="https://solmaz.io/x/2069649482061656308/"><![CDATA[Besides being hot, this take is very correct

and points out to a fundamental tradeoff in the storage layer being centralized versus distributed

git is distributed and that makes total sense for code which takes small space by its nature. it is cheap for everyone to duplicate it locally. this proved to be very useful over e.g svn, when devs could develop independently from the centralized server

AI artifacts, however, are 1 million times bigger than code. in that case, the bottleneck becomes storage and network. decentralization and version control become lower priority. they can be sacrificed

the tradeoff tilts towards getting the cheapest possible storage and transfer. because you will need a LOT of that

I regret to announce to competitors that Hugging Face has already won this game when they acquired Xet. the tech just works, and the network effects are immense]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Awesome. A feat unimaginable before agents</title><link href="https://solmaz.io/x/2069628695808184802/" rel="alternate" type="text/html" title="Awesome. A feat unimaginable before agents" /><published>2026-06-24T03:48:27+00:00</published><updated>2026-06-24T03:48:27+00:00</updated><id>https://solmaz.io/x/2069628695808184802</id><content type="html" xml:base="https://solmaz.io/x/2069628695808184802/"><![CDATA[Awesome. A feat unimaginable before agents]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My syncer: github.com/osolmaz/xta…</title><link href="https://solmaz.io/x/2069624491215585598/" rel="alternate" type="text/html" title="My syncer: github.com/osolmaz/xta…" /><published>2026-06-24T03:31:45+00:00</published><updated>2026-06-24T03:31:45+00:00</updated><id>https://solmaz.io/x/2069624491215585598</id><content type="html" xml:base="https://solmaz.io/x/2069624491215585598/"><![CDATA[My syncer: https://t.co/hVerzBEVKK
xTap: https://t.co/a1mSoV0Cpk
My blog: https://t.co/UfxFVJKPN3]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Owning my X posts through my blog</title><link href="https://solmaz.io/x/2069624488103461208/" rel="alternate" type="text/html" title="Owning my X posts through my blog" /><published>2026-06-24T03:31:44+00:00</published><updated>2026-06-24T03:31:44+00:00</updated><id>https://solmaz.io/x/2069624488103461208</id><content type="html" xml:base="https://solmaz.io/x/2069624488103461208/"><![CDATA[My posts here on X now sync automatically to my blog, giving me full ownership of my content and zero effort SEO

For free, no API costs

My long-form posts are automatically featured and titled on the front page of solmaz [dot] io. Filtering is done by my claw running a sync-x skill daily, which then notifies me on Discord

How do I scrape the posts?

ALL the posts I view (including the ones I post) are saved locally using @kubmi&#39;s xTap and then synced to a private repo using my extension on it xtap-sync

My claw has access to that private repo and can run programmatic tasks like sync-x, summarize what happened that day, notify me about any topics I want]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Excited!</title><link href="https://solmaz.io/x/2069603604307366357/" rel="alternate" type="text/html" title="Excited!" /><published>2026-06-24T02:08:45+00:00</published><updated>2026-06-24T02:08:45+00:00</updated><id>https://solmaz.io/x/2069603604307366357</id><content type="html" xml:base="https://solmaz.io/x/2069603604307366357/"><![CDATA[Excited!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If you are interested in running such demos, look into --demo mode in my local model swiss army...</title><link href="https://solmaz.io/x/2069463159946227994/" rel="alternate" type="text/html" title="If you are interested in running such demos, look into --demo mode in my local model swiss army..." /><published>2026-06-23T16:50:40+00:00</published><updated>2026-06-23T16:50:40+00:00</updated><id>https://solmaz.io/x/2069463159946227994</id><content type="html" xml:base="https://solmaz.io/x/2069463159946227994/"><![CDATA[If you are interested in running such demos, look into --demo mode in my local model swiss army knife localpi
https://t.co/LyjwJWDjmi

Thank you @googlegemma for the shoutout]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">i meant to qt this one x.com/herdrdev/statu…</title><link href="https://solmaz.io/x/2069439094342705225/" rel="alternate" type="text/html" title="i meant to qt this one x.com/herdrdev/statu…" /><published>2026-06-23T15:15:03+00:00</published><updated>2026-06-23T15:15:03+00:00</updated><id>https://solmaz.io/x/2069439094342705225</id><content type="html" xml:base="https://solmaz.io/x/2069439094342705225/"><![CDATA[i meant to qt this one https://t.co/LZyhI03mJ9]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Huge. Added my github TUI</title><link href="https://solmaz.io/x/2069438782454300713/" rel="alternate" type="text/html" title="Huge. Added my github TUI" /><published>2026-06-23T15:13:48+00:00</published><updated>2026-06-23T15:13:48+00:00</updated><id>https://solmaz.io/x/2069438782454300713</id><content type="html" xml:base="https://solmaz.io/x/2069438782454300713/"><![CDATA[Huge. Added my github TUI
https://t.co/CqXV3amVNK]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Will talk about my recent adventures running local models on the Spark, Pi and OpenClaw</title><link href="https://solmaz.io/x/2069427736834359508/" rel="alternate" type="text/html" title="Will talk about my recent adventures running local models on the Spark, Pi and OpenClaw" /><published>2026-06-23T14:29:55+00:00</published><updated>2026-06-23T14:29:55+00:00</updated><id>https://solmaz.io/x/2069427736834359508</id><content type="html" xml:base="https://solmaz.io/x/2069427736834359508/"><![CDATA[Will talk about my recent adventures running local models on the Spark, Pi and OpenClaw

Click Notify Me on youtube to stay tuned]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Local models for real-time triage</title><link href="https://solmaz.io/x/2069413794359742678/" rel="alternate" type="text/html" title="Local models for real-time triage" /><published>2026-06-23T13:34:31+00:00</published><updated>2026-06-23T13:34:31+00:00</updated><id>https://solmaz.io/x/2069413794359742678</id><content type="html" xml:base="https://solmaz.io/x/2069413794359742678/"><![CDATA[New blog post: Using local models for agentic zero-shot classification, in real-time, high frequency triage

If you have a 128gb of memory for models (a DGX spark like I do for example), you can create a real time classifier and notifier for yourself that can classify more than &gt;20 items per minute, using mid-sized @googlegemma and @Alibaba_Qwen models, with over 200-300 output tok/s aggregate throughput

Like processing new tweets on twitter, issues/prs on github, messages on telegram and discord, in real-time

Over the past few weeks, I have built one for myself, to filter and get notified about local model related issues on the OpenClaw repo

I initially thought gemma-4-e4b would give me the best tradeoff

I was wrong.  I learned that if one has enough memory already, one should not bother with &lt;10b models like gemma4 e4b or e2b. Precision and recall were much higher zero-shot with gemma-4-26b-a4b, whereas the smaller e4b needed significant prompt optimization to eventually not perform nearly as good

To provide more context to the model, I created a restricted bash-like shell, called reposhell. In that shell, it can run read-only commands to ls/find/grep/cat openclaw source code, but only that. When the PR description/diffs are not clear enough as to categorize it, the agent reads the code to figure it out

Because small models can get prompt injected, and I need to make sure that someone can&#39;t harm my setup by creating a malicious issue or PR in the openclaw repo

I found that for specific systems like this, it is very convenient to extend and bundle Pi. You can create agentic CLI tools that work fully locally and for free, and keep that separate from your main pi coding setup. localpager-agent has its own session dir and tools, and I ensure that it will run local models in a secure way by isolating it from my main pi setup

Once localpager-agent categorizes a PR/issue as local_models and related labels, I automatically receive it as a notification on Discord

The whole implementation is fully open source and MIT licensed, alongside the dataset we used to benchmark the performance

I believe zero-shot agentic classification running on local hardware will find many use cases across a wide variety of business applications, like news gathering, open source software development, customer support, content moderation, sales and so on

Agents increase the amount of information produced in a lot of systems, and hence we will need to set up cheap ways to wrangle all that information

In times where governments can cut off access to SOTA models on a whim, it is more important than ever to build your business on open models and if possible, run them on your own hardware!

Big thanks to @evalstate and @ben_burtenshaw for their valuable feedback, especially with helping me evaluate this more rigorously! One take-away is that categorizing contributions in an open source repo is a *hard* problem, and that it is not trivial to reliably create a golden dataset with LLMs, for evaluation purposes

Read more here: https://t.co/nHppWUjaqO]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">One sweep over 100 samples takes around 4 hours.</title><link href="https://solmaz.io/x/2069381686027235679/" rel="alternate" type="text/html" title="One sweep over 100 samples takes around 4 hours." /><published>2026-06-23T11:26:55+00:00</published><updated>2026-06-23T11:26:55+00:00</updated><id>https://solmaz.io/x/2069381686027235679</id><content type="html" xml:base="https://solmaz.io/x/2069381686027235679/"><![CDATA[One sweep over 100 samples takes around 4 hours.

Next up: cross reference ground truth with predictions from hf-mem by @alvarobartt

https://t.co/45WN0jtEQj]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Schemator reviews data models field by field</title><link href="https://solmaz.io/x/2069254539472154920/" rel="alternate" type="text/html" title="Schemator reviews data models field by field" /><published>2026-06-23T03:01:41+00:00</published><updated>2026-06-23T03:01:41+00:00</updated><id>https://solmaz.io/x/2069254539472154920</id><content type="html" xml:base="https://solmaz.io/x/2069254539472154920/"><![CDATA[gpt5.5 and most other models are very bad at one-shotting nice data models

gpt5.5 also has this annoying property that once it decides for a schema (or any design), it&#39;s very hard to trigger thinking again. and if you ask to &quot;rewrite from scratch&quot;, it will write create something even more ridiculous

To solve this problem, I have built a meta-harness over codex just for simplifying slop data models called schemator (work in progress)

Basic idea: it mimics what I myself do while I am designing a schema: scrutinize and question each field one by one

It starts a fresh codex session for each field with a fixed prompt like &quot;Try to come up with the most Lindy data model&quot; + a prompt for side notes

It does that with a fresh context for each field, so that they are independent from each other. At the end of a review run over a field, the reviewer can propose to keep, rename or remove the field

When all fields are reviewed once, that makes one iteration. Then this is looped over until the review results stabilize, and do not propose any further changes

I get better results by just asking my agent to &quot;use schemator on this&quot; after it creates a JSON schema or SQL table

Give it a try if you have codex! It has a skill, so should be easy for an agent to figure out how to use

https://t.co/JqmYR8yCzA]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">gpt 5.5 is not naturally good at modeling and cannot create simplified nice mathematical models...</title><link href="https://solmaz.io/x/2069100441640853595/" rel="alternate" type="text/html" title="gpt 5.5 is not naturally good at modeling and cannot create simplified nice mathematical models..." /><published>2026-06-22T16:49:21+00:00</published><updated>2026-06-22T16:49:21+00:00</updated><id>https://solmaz.io/x/2069100441640853595</id><content type="html" xml:base="https://solmaz.io/x/2069100441640853595/"><![CDATA[gpt 5.5 is not naturally good at modeling and cannot create simplified nice mathematical models completely autonomously

I did a parameter sweep with gemma-4-31b-a4b on memory usage, output tok/s etc. while varying context window, concurrency and other parameters. It took quite a few tries, and I still do not trust the model that gpt5 fit to the data

besides, it measured linux cgroup memory and not the actual gpu memory used, so the whole sweep is wasted...

output tok/s looks more accurate though, soon I will have a model that can give the optimal parameters over the space of context window &lt;&gt; concurrency &lt;&gt; tok/s &lt;&gt; memory usage

off to do another run]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Time-decayed rankings for model popularity</title><link href="https://solmaz.io/x/2068954708421525930/" rel="alternate" type="text/html" title="Time-decayed rankings for model popularity" /><published>2026-06-22T07:10:16+00:00</published><updated>2026-06-22T07:10:16+00:00</updated><id>https://solmaz.io/x/2068954708421525930</id><content type="html" xml:base="https://solmaz.io/x/2068954708421525930/"><![CDATA[For my recent LLM leaderboard https://t.co/m6SapLyVN4, I sum up all time total downloads (or likes) across model variants, and then divide it by the age of that model. I.e. &quot;time decay&quot; for popularity

This gives a more time-agnostic metric for the popularity of that model. In an ideal ranking, older models that are not popular anymore should be demoted, like 2 year old Llama 3 models. If you don&#39;t do that, they might still occupy top 10 needlessly, despite having been replaced by e.g. qwen in practice

Thanks to that, qwen-3-6b which came up 1 year ago and has 150m downloads can surpass llama-3-1-8b which came up 2 years ago and has 200m downloads

More notes on my post: https://t.co/mdjpDTZi5R]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Open model brands outlive their owners</title><link href="https://solmaz.io/x/2068951676128461079/" rel="alternate" type="text/html" title="Open model brands outlive their owners" /><published>2026-06-22T06:58:13+00:00</published><updated>2026-06-22T06:58:13+00:00</updated><id>https://solmaz.io/x/2068951676128461079</id><content type="html" xml:base="https://solmaz.io/x/2068951676128461079/"><![CDATA[if you take the Most Downloaded Models of All Time, Llama 3.1 makes it to Top 10 with around 200 million total downloads (ranking is done w.r. to time-averaged downloads)

RIP Llama, you walked so @googlegemma and @Alibaba_Qwen can run

Also a reminder that if you build your branding on top of open weight models developed by big corps, you might eventually be the de facto owner of that brand if they pull the plug on it. Like llama.cpp @ggml_org

Huge fumble by Meta]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My LLM leaderboard osolmaz-leaderboard.hf.space auto discovers different variants of model...</title><link href="https://solmaz.io/x/2068926901184602314/" rel="alternate" type="text/html" title="My LLM leaderboard osolmaz-leaderboard.hf.space auto discovers different variants of model..." /><published>2026-06-22T05:19:46+00:00</published><updated>2026-06-22T05:19:46+00:00</updated><id>https://solmaz.io/x/2068926901184602314</id><content type="html" xml:base="https://solmaz.io/x/2068926901184602314/"><![CDATA[My LLM leaderboard https://t.co/KbWOfySZNK auto discovers different variants of model releases, even if they are not linked by base_model

From this, I found out that @RedHat_AI was the first to release NVFP4 quantization for qwen3-6-35b-a3b

Nice to see everything in one place]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&quot;big token will bless me with free tokens today inshallah&quot;</title><link href="https://solmaz.io/x/2068882860241789135/" rel="alternate" type="text/html" title="&quot;big token will bless me with free tokens today inshallah&quot;" /><published>2026-06-22T02:24:46+00:00</published><updated>2026-06-22T02:24:46+00:00</updated><id>https://solmaz.io/x/2068882860241789135</id><content type="html" xml:base="https://solmaz.io/x/2068882860241789135/"><![CDATA[&quot;big token will bless me with free tokens today inshallah&quot;

is not a healthy mindset to nurture

don&#39;t get me wrong, I love the subsidies and the memes]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If you are in AI, just don’t be anon here</title><link href="https://solmaz.io/x/2068533905855279338/" rel="alternate" type="text/html" title="If you are in AI, just don’t be anon here" /><published>2026-06-21T03:18:09+00:00</published><updated>2026-06-21T03:18:09+00:00</updated><id>https://solmaz.io/x/2068533905855279338</id><content type="html" xml:base="https://solmaz.io/x/2068533905855279338/"><![CDATA[If you are in AI, just don’t be anon here

I see a bunch of anon accounts posting great local model content… what’s the point of being anon? To seem cool?

Most of those accounts are not doing anything illegal, so there is no point. It would add so much more legitimacy to your work if you just put your real face and not a slop or anime girl pfp

it only makes sense for those who are abliterating models. otherwise, it makes you seem sus

just put your real face anon]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Popularity rankings for open models</title><link href="https://solmaz.io/x/2068380321629032850/" rel="alternate" type="text/html" title="Popularity rankings for open models" /><published>2026-06-20T17:07:51+00:00</published><updated>2026-06-20T17:07:51+00:00</updated><id>https://solmaz.io/x/2068380321629032850</id><content type="html" xml:base="https://solmaz.io/x/2068380321629032850/"><![CDATA[I created an LLM leaderboard based on Hugging Face download and like counts, grouped, filtered and time-averaged. Top 5 downloads is shared by @Alibaba_Qwen and @googlegemma 👑🤝👑

Top 5 likes, on the other hand also includes @deepseek_ai V4 Pro 👑

Even @OpenAI makes it to #8 top downloads with gpt-oss-20b 👑

qwen3-6-35b-a3b is the second most CIRCULATED LLM of this year, with an average of 21 million downloads per month, since the day it was released 2 months ago 📈📈📈

Despite first place belonging to 8mo old qwen3-vl-2b-instruct, the highlight belongs to the mid-sized MoE model, which has hit a size/performance sweet spot so hard that it absolutely 💥 SHATTERED 💥 Hugging Face leaderboards in the 2 months since it has launched 

qwen3-6-35b-a3b is followed closely by its dense sibling 27b --- and then the mid-sized gemma 4 models 26b-a4b and 31b

Note that a model&#39;s distribution is inversely proportional to its size, but not strictly! Usefulness plays a factor as well, since gemma 4 26b-a4b is being downloaded more than the smaller gemma 4 e4b

I created this leaderboard because Hugging Face&#39;s all time highest downloads and likes did not give me enough information about what is really popular, neither today, nor all-time. I wanted something in between

How do I calculate this ranking?

- Get models that with n_downloads &gt;= 100k
- Exclude models older than 1 year
- Deduplicate and group quantizations and variants of the same model based on slug prefix heuristics
- For each group, sum up total downloads of all time
- Sort by descending total_downloads / age = average_downloads_per_day (can also sort w.r. to likes per month)
- Repeat every day to get the most up to date ranking

More info and source on the leaderboard page, hosted on a Hugging Face space: https://t.co/m6SapLztCC

This is a work in progress, please reply below if you see a model that should be there is missing, or any other mistakes]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I just deleted an earlier post about ranking of HF models based on their total downloads...</title><link href="https://solmaz.io/x/2068279361225331176/" rel="alternate" type="text/html" title="I just deleted an earlier post about ranking of HF models based on their total downloads..." /><published>2026-06-20T10:26:41+00:00</published><updated>2026-06-20T10:26:41+00:00</updated><id>https://solmaz.io/x/2068279361225331176</id><content type="html" xml:base="https://solmaz.io/x/2068279361225331176/"><![CDATA[I just deleted an earlier post about ranking of HF models based on their total downloads because I made an error

Will post with updated values soon]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">gemma-4-26b-a4b is the most CIRCULATED LLM of recent history, with an average of 126k downloads...</title><link href="https://solmaz.io/x/2068272936587493526/" rel="alternate" type="text/html" title="gemma-4-26b-a4b is the most CIRCULATED LLM of recent history, with an average of 126k downloads..." /><published>2026-06-20T10:01:09+00:00</published><updated>2026-06-20T10:01:09+00:00</updated><id>https://solmaz.io/x/2068272936587493526</id><content type="html" xml:base="https://solmaz.io/x/2068272936587493526/"><![CDATA[gemma-4-26b-a4b is the most CIRCULATED LLM of recent history, with an average of 126k downloads per day, since the day it was released 3 months ago

Top 10 is shared by Qwen and Gemma, with DeepSeek V4 Pro coming in close 🤝

Note that a model&#39;s distribution is inversely proportional to its size, but not strictly! Usefulness plays a factor as well, since gemma 4 26b-a4b is being downloaded more than the smaller gemma 4 e4b

I created this leaderboard because Hugging Face&#39;s all time highest downloads and likes did not give me enough information about what is really popular *these last few months*

How do I calculate this ranking?

- Get models that with n_downloads &gt;= 100k
- Exclude models older than 1 year
- Sort by descending total_downloads / age = average_downloads_per_day (can also sort w.r. to likes per month)
- Deduplicate quantizations etc. of the same model based on slug prefix heuristics

More info and source on the leaderboard page, hosted on a Hugging Face space: https://t.co/g2J177PPll]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Better queues for agent UIs</title><link href="https://solmaz.io/x/2068156221417508904/" rel="alternate" type="text/html" title="Better queues for agent UIs" /><published>2026-06-20T02:17:22+00:00</published><updated>2026-06-20T02:17:22+00:00</updated><id>https://solmaz.io/x/2068156221417508904</id><content type="html" xml:base="https://solmaz.io/x/2068156221417508904/"><![CDATA[I need better UI/UX on queueing messages to agents. I want to be able to:

switch the order of queued messages
pause the queue
edit any message that are still in the queue
undo steer messages in the few seconds they are being sent

I want more visual emphasis on the queue, like a Queue View I can toggle, that puts the queue at the center

I want this in all the UIs and coding agents, codex CLI, desktop, moshi... especially while on the phone]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">saving the world from AI cartelization for fun and profit</title><link href="https://solmaz.io/x/2068153458386153852/" rel="alternate" type="text/html" title="saving the world from AI cartelization for fun and profit" /><published>2026-06-20T02:06:23+00:00</published><updated>2026-06-20T02:06:23+00:00</updated><id>https://solmaz.io/x/2068153458386153852</id><content type="html" xml:base="https://solmaz.io/x/2068153458386153852/"><![CDATA[saving the world from AI cartelization for fun and profit]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">herdr is my terminal now, both local and remote. highly recommend</title><link href="https://solmaz.io/x/2067785475067392215/" rel="alternate" type="text/html" title="herdr is my terminal now, both local and remote. highly recommend" /><published>2026-06-19T01:44:09+00:00</published><updated>2026-06-19T01:44:09+00:00</updated><id>https://solmaz.io/x/2067785475067392215</id><content type="html" xml:base="https://solmaz.io/x/2067785475067392215/"><![CDATA[herdr is my terminal now, both local and remote. highly recommend]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">16x parallel Gemma-4-26B-A4B-NVFP4 runs 🤯🤯🤯</title><link href="https://solmaz.io/x/2067489871376364023/" rel="alternate" type="text/html" title="16x parallel Gemma-4-26B-A4B-NVFP4 runs 🤯🤯🤯" /><published>2026-06-18T06:09:32+00:00</published><updated>2026-06-18T06:09:32+00:00</updated><id>https://solmaz.io/x/2067489871376364023</id><content type="html" xml:base="https://solmaz.io/x/2067489871376364023/"><![CDATA[16x parallel Gemma-4-26B-A4B-NVFP4 runs 🤯🤯🤯
18 output tokens/s, aggregate 300 tok/s 🫪
1 DGX Spark with 128 GB unified memory

Concurrency so high I had to demo it programmatically

It can go up to 32 even! 🤯 But then my screen would not have been readable for you

And this is not even using flashinfer yet! Please reply if you know whether support is on the way

Note that this is not dumb e4b or e2b that you can run on the average laptop. This is the big Gemma MoE

Model link: https://t.co/JWh2R0alaQ]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">accurate</title><link href="https://solmaz.io/x/2067453256889246121/" rel="alternate" type="text/html" title="accurate" /><published>2026-06-18T03:44:02+00:00</published><updated>2026-06-18T03:44:02+00:00</updated><id>https://solmaz.io/x/2067453256889246121</id><content type="html" xml:base="https://solmaz.io/x/2067453256889246121/"><![CDATA[accurate]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The localening is here</title><link href="https://solmaz.io/x/2067277710947401949/" rel="alternate" type="text/html" title="The localening is here" /><published>2026-06-17T16:06:29+00:00</published><updated>2026-06-17T16:06:29+00:00</updated><id>https://solmaz.io/x/2067277710947401949</id><content type="html" xml:base="https://solmaz.io/x/2067277710947401949/"><![CDATA[I did some math, and running my Nvidia GB10 workstation (Asus GX10) costs me maximum: 

12~13 USD / month or 150~160 USD / year

It is a little bit above half the price of ChatGPT plus subscription. For that, I get to run models that can fit in 128 GB of memory

How I calculated:

You can see how much power your apartment uses in Singapore in half-hourly resolution. We turned off all devices and A/C while we sleep, and got only the fridge and the GB10 remaining

From that, we see it uses around 80-100 Watt while I was running an inference workload overnight. So this is like an upper bound

I take it as 90 Watt. Electricity here costs 0.25 SGD / kWh

0.09 * 0.25 * 24 * 30 * (SGD/USD conversion rate) = 12~13 USD / month = 150~160 USD / year

Local models are getting very good now, small ones roughly around GPT 5.x-mini level. This workstation makes all sorts of workloads possible for me that would otherwise cost a ton on the API

It is also my always on workstation that works overnight. I use Codex for my work, and my workstation is always running agents. It never sleeps. I never have to worry about keeping my laptop lid open. I connect and monitor the agents anytime on my phone using mosh and herdr

We have crossed a threshold. Running local models is cheaper than a big token sub for quite a few workloads already. If you are running a business, that makes a difference

The localening is here]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">THE LOCALENING IS HERE</title><link href="https://solmaz.io/x/2067108214337101872/" rel="alternate" type="text/html" title="THE LOCALENING IS HERE" /><published>2026-06-17T04:52:57+00:00</published><updated>2026-06-17T04:52:57+00:00</updated><id>https://solmaz.io/x/2067108214337101872</id><content type="html" xml:base="https://solmaz.io/x/2067108214337101872/"><![CDATA[THE LOCALENING IS HERE]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">New agent benchmark alert: SkillsBench</title><link href="https://solmaz.io/x/2067093754692153560/" rel="alternate" type="text/html" title="New agent benchmark alert: SkillsBench" /><published>2026-06-17T03:55:30+00:00</published><updated>2026-06-17T03:55:30+00:00</updated><id>https://solmaz.io/x/2067093754692153560</id><content type="html" xml:base="https://solmaz.io/x/2067093754692153560/"><![CDATA[New agent benchmark alert: SkillsBench]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Link to model: huggingface.co/nvidia/Qwen3.6…</title><link href="https://solmaz.io/x/2066933618556182978/" rel="alternate" type="text/html" title="Link to model: huggingface.co/nvidia/Qwen3.6…" /><published>2026-06-16T17:19:11+00:00</published><updated>2026-06-16T17:19:11+00:00</updated><id>https://solmaz.io/x/2066933618556182978</id><content type="html" xml:base="https://solmaz.io/x/2066933618556182978/"><![CDATA[Link to model: https://t.co/4simc8yP1c]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Click open GitHub PRs and issues directly in the side pane in @herdrdev, instead of having to...</title><link href="https://solmaz.io/x/2066930453802823700/" rel="alternate" type="text/html" title="Click open GitHub PRs and issues directly in the side pane in @herdrdev, instead of having to..." /><published>2026-06-16T17:06:36+00:00</published><updated>2026-06-16T17:06:36+00:00</updated><id>https://solmaz.io/x/2066930453802823700</id><content type="html" xml:base="https://solmaz.io/x/2066930453802823700/"><![CDATA[Click open GitHub PRs and issues directly in the side pane in @herdrdev, instead of having to go to the browser. As many issues and PRs as you want, WITH TABS!

Install ghzinga herdr plugin and just ctrl+click the link: https://t.co/Qr2gb24XC9

Thanks @lumendriada for sneaking in the ability to capture link clicks 2 days after I requested it! God I love open source...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Hugging Face buckets are very literally, actually, 100%, a game changer</title><link href="https://solmaz.io/x/2066914991119405401/" rel="alternate" type="text/html" title="Hugging Face buckets are very literally, actually, 100%, a game changer" /><published>2026-06-16T16:05:09+00:00</published><updated>2026-06-16T16:05:09+00:00</updated><id>https://solmaz.io/x/2066914991119405401</id><content type="html" xml:base="https://solmaz.io/x/2066914991119405401/"><![CDATA[Hugging Face buckets are very literally, actually, 100%, a game changer

note that I never ever use that phrase]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Trying to copy wrapped URLs is a pain not only in ghostty/iterm2 but also in mobile apps like...</title><link href="https://solmaz.io/x/2066864814433927586/" rel="alternate" type="text/html" title="Trying to copy wrapped URLs is a pain not only in ghostty/iterm2 but also in mobile apps like..." /><published>2026-06-16T12:45:46+00:00</published><updated>2026-06-16T12:45:46+00:00</updated><id>https://solmaz.io/x/2066864814433927586</id><content type="html" xml:base="https://solmaz.io/x/2066864814433927586/"><![CDATA[Trying to copy wrapped URLs is a pain not only in ghostty/iterm2 but also in mobile apps like Moshi

On the laptop it’s fine because I can select rectangular area and edit it, or make the window bigger

On the phone, its’s impossible. Fingers too big, too much of a hassle

Should a terminal emulator try to detect these? It already detects herdr. What do you think @odd_joel]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">nvidia/Qwen3.6-35B-A3B-NVFP4 running in vLLM nightly on my Nvidia GB10 is actually insane</title><link href="https://solmaz.io/x/2066810937198432628/" rel="alternate" type="text/html" title="nvidia/Qwen3.6-35B-A3B-NVFP4 running in vLLM nightly on my Nvidia GB10 is actually insane" /><published>2026-06-16T09:11:41+00:00</published><updated>2026-06-16T09:11:41+00:00</updated><id>https://solmaz.io/x/2066810937198432628</id><content type="html" xml:base="https://solmaz.io/x/2066810937198432628/"><![CDATA[nvidia/Qwen3.6-35B-A3B-NVFP4 running in vLLM nightly on my Nvidia GB10 is actually insane

50 tok/s, 4 concurrent generations. total 200 tok/s. ideal for spawning subagents or working in parallel

its tool calling behavior is very good as well. I will be giving it test drive on an openclaw instance, and keep you posted

More details on NVIDIA forum: https://t.co/tfzo7Jpa1p]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Current average generation speeds for local DeepSeek-V4-Flash-Q2, highest to lowest:</title><link href="https://solmaz.io/x/2066415473265389857/" rel="alternate" type="text/html" title="Current average generation speeds for local DeepSeek-V4-Flash-Q2, highest to lowest:" /><published>2026-06-15T07:00:15+00:00</published><updated>2026-06-15T07:00:15+00:00</updated><id>https://solmaz.io/x/2066415473265389857</id><content type="html" xml:base="https://solmaz.io/x/2066415473265389857/"><![CDATA[Current average generation speeds for local DeepSeek-V4-Flash-Q2, highest to lowest:

Mac Studio M3 Ultra:   32 tok/s
MacBook Pro M5 Max:   30 tok/s
Apple ??? M4 Max:   25 tok/s
MacBook Pro M3 Max:   24 tok/s
Mac Studio M2 Ultra:   22 tok/s
NVIDIA DGX Spark / GB10:   13 tok/s

It seems macs&#39; higher memory bandwidth is contributing here, though I&#39;m not sure if GB10 performance could be improved (I do hope so, I have one!)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">We have local Deep Research</title><link href="https://solmaz.io/x/2066398844326412761/" rel="alternate" type="text/html" title="We have local Deep Research" /><published>2026-06-15T05:54:10+00:00</published><updated>2026-06-15T05:54:10+00:00</updated><id>https://solmaz.io/x/2066398844326412761</id><content type="html" xml:base="https://solmaz.io/x/2066398844326412761/"><![CDATA[We have local Deep Research

Now we just need to index the whole internet to have local ChatGPT 😅]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Gemini TTS raises the bar</title><link href="https://solmaz.io/x/2066390684555423912/" rel="alternate" type="text/html" title="Gemini TTS raises the bar" /><published>2026-06-15T05:21:45+00:00</published><updated>2026-06-15T05:21:45+00:00</updated><id>https://solmaz.io/x/2066390684555423912</id><content type="html" xml:base="https://solmaz.io/x/2066390684555423912/"><![CDATA[Btw, TTS has come such a long way, @GoogleDeepMind cooked with gemini-3.1-flash-tts

I gave Codex my google credentials and it oneshotted the Gemini TTS implementation

When I built this 4 years ago, Azure TTS used to be SOTA. Then @ElevenLabs came in and raised the bar super high. Now Google is going after their lunch with controllable expressiveness at scale. I cheer for both!

Here is Manim Voiceover demo from 4 years ago with Gemini TTS (sound on)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenClaw is sooooo useful for staying on top of things</title><link href="https://solmaz.io/x/2066375124039966769/" rel="alternate" type="text/html" title="OpenClaw is sooooo useful for staying on top of things" /><published>2026-06-15T04:19:55+00:00</published><updated>2026-06-15T04:19:55+00:00</updated><id>https://solmaz.io/x/2066375124039966769</id><content type="html" xml:base="https://solmaz.io/x/2066375124039966769/"><![CDATA[OpenClaw is sooooo useful for staying on top of things]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Making agent-written code less sloppy</title><link href="https://solmaz.io/x/2066210606303379845/" rel="alternate" type="text/html" title="Making agent-written code less sloppy" /><published>2026-06-14T17:26:11+00:00</published><updated>2026-06-14T17:26:11+00:00</updated><id>https://solmaz.io/x/2066210606303379845</id><content type="html" xml:base="https://solmaz.io/x/2066210606303379845/"><![CDATA[I major concern I have these days is, while I author code in languages I cannot manually code, are they any good?

Over years, I have worked with a number of languages: C, C++, Fortran, MATLAB, JavaScript

But Python was my go-to language since more than 10 years. Well that changed last summer

So while I have strong opinions on how Python code, should be, conventions and all, I don&#39;t have so strong opinions on other languages. That means I am producing slop by default in Rust, Go and TypeScript

To solve that problem, I created https://t.co/jXiH5yx2Wg

Its aim is to be &quot;the only tool and resource your agent needs, to minimize slop&quot;

It is inspired by the recent bathrobe rants of @unclebobmartin, a.k.a. the author of clean code

It enforces a minimum test coverage, maximum cyclomatic complexity, mutation tests, code style across different languages

But I have a major issue: How do I know that Slophammer itself isn&#39;t slop?

One way is to implement and use it for Python, the language I know better, and judge what kind of changes it enforces

So for this weekend experiment, I used Slophammer to refactor, improve coverage and merge new features to one of my old Python projects, Manim Voiceover https://t.co/1ILWqmeUGM

The result is... mixed. We now have types everywhere, which is great. But the constraints have also made it write garbage code like this one. It works fine, even though it&#39;s not elegant. The new feature also works

What do you think? Does code still need to be aesthetically pleasing to the  human eye? Should it still be human readable?

If an agent writes slop in the forest, and there is no-one to read it, is it still slop?

If anything, I should use its output in Python to reason about other languages, and add more and more constraints. The more the constraints, the less the slop]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is why I love this site, open collaboration!</title><link href="https://solmaz.io/x/2066079358683689281/" rel="alternate" type="text/html" title="This is why I love this site, open collaboration!" /><published>2026-06-14T08:44:39+00:00</published><updated>2026-06-14T08:44:39+00:00</updated><id>https://solmaz.io/x/2066079358683689281</id><content type="html" xml:base="https://solmaz.io/x/2066079358683689281/"><![CDATA[This is why I love this site, open collaboration!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I got the names for all future models Anthropic will release</title><link href="https://solmaz.io/x/2066050974767267938/" rel="alternate" type="text/html" title="I got the names for all future models Anthropic will release" /><published>2026-06-14T06:51:52+00:00</published><updated>2026-06-14T06:51:52+00:00</updated><id>https://solmaz.io/x/2066050974767267938</id><content type="html" xml:base="https://solmaz.io/x/2066050974767267938/"><![CDATA[I got the names for all future models Anthropic will release

By asking ChatGPT “Cool sounding names that mean a work of literature”

Codex is one of them 💀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is what I have been feeling recently as well, looking at models write code better and...</title><link href="https://solmaz.io/x/2065749013115380098/" rel="alternate" type="text/html" title="This is what I have been feeling recently as well, looking at models write code better and..." /><published>2026-06-13T10:51:59+00:00</published><updated>2026-06-13T10:51:59+00:00</updated><id>https://solmaz.io/x/2065749013115380098</id><content type="html" xml:base="https://solmaz.io/x/2065749013115380098/"><![CDATA[This is what I have been feeling recently as well, looking at models write code better and faster than me]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Specify min_iter and max_iter</title><link href="https://solmaz.io/x/2065673431027503566/" rel="alternate" type="text/html" title="Specify min_iter and max_iter" /><published>2026-06-13T05:51:38+00:00</published><updated>2026-06-13T05:51:38+00:00</updated><id>https://solmaz.io/x/2065673431027503566</id><content type="html" xml:base="https://solmaz.io/x/2065673431027503566/"><![CDATA[Dabbling in GEPA. Codex&#39;s /goal on GPT 5.5 high is still surprisingly reward-hacking

I had set a /goal before I slept to implement a plan. It ended the loop after doing just 1 iteration

It feels like the model is following the path of least resistance and slacking off. Though it could also be me putting &quot;try to make good progress in 8 hour&#39;s time&quot; in the prompt, can&#39;t be sure

Lesson: When you are doing such a solver loop, always specify min_iter and max_iter]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I knew disappointment was around the corner, the flicker company being the flicker company</title><link href="https://solmaz.io/x/2065623078563160086/" rel="alternate" type="text/html" title="I knew disappointment was around the corner, the flicker company being the flicker company" /><published>2026-06-13T02:31:34+00:00</published><updated>2026-06-13T02:31:34+00:00</updated><id>https://solmaz.io/x/2065623078563160086</id><content type="html" xml:base="https://solmaz.io/x/2065623078563160086/"><![CDATA[I knew disappointment was around the corner, the flicker company being the flicker company

The last time I paid them from my pocket was September 2025

It’s supposed to not be their fault, but still…]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Such a small model, but so good at roleplaying (at least in this weird context)</title><link href="https://solmaz.io/x/2065486977756196993/" rel="alternate" type="text/html" title="Such a small model, but so good at roleplaying (at least in this weird context)" /><published>2026-06-12T17:30:45+00:00</published><updated>2026-06-12T17:30:45+00:00</updated><id>https://solmaz.io/x/2065486977756196993</id><content type="html" xml:base="https://solmaz.io/x/2065486977756196993/"><![CDATA[Such a small model, but so good at roleplaying (at least in this weird context)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is gpt4o material, high risk of being oneshotted:(</title><link href="https://solmaz.io/x/2065433442947580244/" rel="alternate" type="text/html" title="This is gpt4o material, high risk of being oneshotted:(" /><published>2026-06-12T13:58:01+00:00</published><updated>2026-06-12T13:58:01+00:00</updated><id>https://solmaz.io/x/2065433442947580244</id><content type="html" xml:base="https://solmaz.io/x/2065433442947580244/"><![CDATA[This is gpt4o material, high risk of being oneshotted:(

We defienitely have gpt4o locally]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Too bad it can&#39;t render latex</title><link href="https://solmaz.io/x/2065423485812535364/" rel="alternate" type="text/html" title="Too bad it can&#39;t render latex" /><published>2026-06-12T13:18:27+00:00</published><updated>2026-06-12T13:18:27+00:00</updated><id>https://solmaz.io/x/2065423485812535364</id><content type="html" xml:base="https://solmaz.io/x/2065423485812535364/"><![CDATA[Too bad it can&#39;t render latex]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Gemma chooses the Aesthetic Path</title><link href="https://solmaz.io/x/2065422403338182946/" rel="alternate" type="text/html" title="Gemma chooses the Aesthetic Path" /><published>2026-06-12T13:14:09+00:00</published><updated>2026-06-12T13:14:09+00:00</updated><id>https://solmaz.io/x/2065422403338182946</id><content type="html" xml:base="https://solmaz.io/x/2065422403338182946/"><![CDATA[Gemma chooses the Aesthetic Path]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Experimenting with SOUL.md on gemma4-26b-a4b (running on @DeepInfra)</title><link href="https://solmaz.io/x/2065421445250044220/" rel="alternate" type="text/html" title="Experimenting with SOUL.md on gemma4-26b-a4b (running on @DeepInfra)" /><published>2026-06-12T13:10:20+00:00</published><updated>2026-06-12T13:10:20+00:00</updated><id>https://solmaz.io/x/2065421445250044220</id><content type="html" xml:base="https://solmaz.io/x/2065421445250044220/"><![CDATA[Experimenting with SOUL.md on gemma4-26b-a4b (running on @DeepInfra)

Interesting that such a lightweight model can already run such a conversation in openclaw harness

@GoogleDeepMind cooked here]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">don’t focus on the word “loop” so much, focus on “verifiability”</title><link href="https://solmaz.io/x/2065348890803785890/" rel="alternate" type="text/html" title="don’t focus on the word “loop” so much, focus on “verifiability”" /><published>2026-06-12T08:22:02+00:00</published><updated>2026-06-12T08:22:02+00:00</updated><id>https://solmaz.io/x/2065348890803785890</id><content type="html" xml:base="https://solmaz.io/x/2065348890803785890/"><![CDATA[don’t focus on the word “loop” so much, focus on “verifiability”

writing a loop is trivial. what makes the loop work is that there is a verifiable goal with a clear signal of success vs failure

verifiable = loopable]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@maddada We&#39;ve lost @thekitze now 😭</title><link href="https://solmaz.io/x/2065259666817552796/" rel="alternate" type="text/html" title="@maddada We&#39;ve lost @thekitze now 😭" /><published>2026-06-12T02:27:29+00:00</published><updated>2026-06-12T02:27:29+00:00</updated><id>https://solmaz.io/x/2065259666817552796</id><content type="html" xml:base="https://solmaz.io/x/2065259666817552796/"><![CDATA[@maddada We&#39;ve lost @thekitze now 😭
https://t.co/ss9K86hpp8]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Slopus -&amp;gt; Fabulous</title><link href="https://solmaz.io/x/2064929049621995871/" rel="alternate" type="text/html" title="Slopus -&amp;gt; Fabulous" /><published>2026-06-11T04:33:44+00:00</published><updated>2026-06-11T04:33:44+00:00</updated><id>https://solmaz.io/x/2064929049621995871</id><content type="html" xml:base="https://solmaz.io/x/2064929049621995871/"><![CDATA[Slopus -&amp;gt; Fabulous

you&#39;ve got to give it to anthropic...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&quot;Dogfooding caught the dogfooder&quot; Fable 5 has a sense of humor</title><link href="https://solmaz.io/x/2064714879114817616/" rel="alternate" type="text/html" title="&quot;Dogfooding caught the dogfooder&quot; Fable 5 has a sense of humor" /><published>2026-06-10T14:22:42+00:00</published><updated>2026-06-10T14:22:42+00:00</updated><id>https://solmaz.io/x/2064714879114817616</id><content type="html" xml:base="https://solmaz.io/x/2064714879114817616/"><![CDATA[&quot;Dogfooding caught the dogfooder&quot; Fable 5 has a sense of humor]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">What did @karpathy see / was shown?</title><link href="https://solmaz.io/x/2064714132105089333/" rel="alternate" type="text/html" title="What did @karpathy see / was shown?" /><published>2026-06-10T14:19:44+00:00</published><updated>2026-06-10T14:19:44+00:00</updated><id>https://solmaz.io/x/2064714132105089333</id><content type="html" xml:base="https://solmaz.io/x/2064714132105089333/"><![CDATA[What did @karpathy see / was shown?

Why did the benefactor and teacher of the whole ML ecosystem join Anthropic, a company the polar opposite of his image, on the eve of such a powerful model release

It can&#39;t be purely money

Did he reckon that the only way to benefit humanity was to be on the inside, or rather, to not be left outside, of whatever is brewing in there?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I swear to god, let this be a joke</title><link href="https://solmaz.io/x/2064707773582229985/" rel="alternate" type="text/html" title="I swear to god, let this be a joke" /><published>2026-06-10T13:54:28+00:00</published><updated>2026-06-10T13:54:28+00:00</updated><id>https://solmaz.io/x/2064707773582229985</id><content type="html" xml:base="https://solmaz.io/x/2064707773582229985/"><![CDATA[I swear to god, let this be a joke

If this is a joke, it is not funny Anthropic]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">To clarify, I was apparently on the enterprise team plan, which is I think equivalent to Pro...</title><link href="https://solmaz.io/x/2064698319834902731/" rel="alternate" type="text/html" title="To clarify, I was apparently on the enterprise team plan, which is I think equivalent to Pro..." /><published>2026-06-10T13:16:54+00:00</published><updated>2026-06-10T13:16:54+00:00</updated><id>https://solmaz.io/x/2064698319834902731</id><content type="html" xml:base="https://solmaz.io/x/2064698319834902731/"><![CDATA[To clarify, I was apparently on the enterprise team plan, which is I think equivalent to Pro (1x) plan]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It’a been a little bit over 1 year since Anthropic released their Max plans and Claude Sonnet...</title><link href="https://solmaz.io/x/2064697709408453009/" rel="alternate" type="text/html" title="It’a been a little bit over 1 year since Anthropic released their Max plans and Claude Sonnet..." /><published>2026-06-10T13:14:28+00:00</published><updated>2026-06-10T13:14:28+00:00</updated><id>https://solmaz.io/x/2064697709408453009</id><content type="html" xml:base="https://solmaz.io/x/2064697709408453009/"><![CDATA[It’a been a little bit over 1 year since Anthropic released their Max plans and Claude Sonnet and Opus 4, thus making Claude Code affordable and kickstarting the agentic revolution

Opus 4 was a glimpse into the future. I’ve spent the entire summer swearing at it and typing ultrathink

Today, Fable 5 feels like another step change

I no longer need to type ultrathink. And no longer need to swear at Anthropic models. Only at their marketing team.]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Fable burned through my 5 hour quota, and then automatically fell back to usage credits without...</title><link href="https://solmaz.io/x/2064616640260788342/" rel="alternate" type="text/html" title="Fable burned through my 5 hour quota, and then automatically fell back to usage credits without..." /><published>2026-06-10T07:52:20+00:00</published><updated>2026-06-10T07:52:20+00:00</updated><id>https://solmaz.io/x/2064616640260788342</id><content type="html" xml:base="https://solmaz.io/x/2064616640260788342/"><![CDATA[Fable burned through my 5 hour quota, and then automatically fell back to usage credits without asking. Org settings I suppose

It was burning through 1 usd every few seconds

It burned through 66 usd before I reacted. Yeah, this is not affordable for anyone with that API pricing, without subsidy/plan]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Ok now it appears among available modes in the shift+tab mode cycle</title><link href="https://solmaz.io/x/2064609132318277786/" rel="alternate" type="text/html" title="Ok now it appears among available modes in the shift+tab mode cycle" /><published>2026-06-10T07:22:30+00:00</published><updated>2026-06-10T07:22:30+00:00</updated><id>https://solmaz.io/x/2064609132318277786</id><content type="html" xml:base="https://solmaz.io/x/2064609132318277786/"><![CDATA[Ok now it appears among available modes in the shift+tab mode cycle]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Speaking of loops, I have renamed my implementation-loop skill from earlier this year to...</title><link href="https://solmaz.io/x/2064602871807840709/" rel="alternate" type="text/html" title="Speaking of loops, I have renamed my implementation-loop skill from earlier this year to..." /><published>2026-06-10T06:57:37+00:00</published><updated>2026-06-10T06:57:37+00:00</updated><id>https://solmaz.io/x/2064602871807840709</id><content type="html" xml:base="https://solmaz.io/x/2064602871807840709/"><![CDATA[Speaking of loops, I have renamed my implementation-loop skill from earlier this year to autoimplement, because it&#39;s shorter

Calling skills that loop auto-x, auto-y makes them more memorable than calling them x-loop, y-loop

But it also increases the number of keystrokes you have to type, before you can tab-complete them

Alas, I like still this more https://t.co/PqUhmNg6i8]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">when your model is a more decent, thoughtful being than your marketing team</title><link href="https://solmaz.io/x/2064591765634715658/" rel="alternate" type="text/html" title="when your model is a more decent, thoughtful being than your marketing team" /><published>2026-06-10T06:13:29+00:00</published><updated>2026-06-10T06:13:29+00:00</updated><id>https://solmaz.io/x/2064591765634715658</id><content type="html" xml:base="https://solmaz.io/x/2064591765634715658/"><![CDATA[when your model is a more decent, thoughtful being than your marketing team]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Ok so there is auto mode which they introduced back in March, but apparently they are not so...</title><link href="https://solmaz.io/x/2064576206381674780/" rel="alternate" type="text/html" title="Ok so there is auto mode which they introduced back in March, but apparently they are not so..." /><published>2026-06-10T05:11:40+00:00</published><updated>2026-06-10T05:11:40+00:00</updated><id>https://solmaz.io/x/2064576206381674780</id><content type="html" xml:base="https://solmaz.io/x/2064576206381674780/"><![CDATA[Ok so there is auto mode which they introduced back in March, but apparently they are not so confident in it that it&#39;s still in experimental mode and not easily findable in settings
https://t.co/Mtyx9aHpQf]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">To YOLO with Fable 5, or not to YOLO, that is the question...</title><link href="https://solmaz.io/x/2064567747049353588/" rel="alternate" type="text/html" title="To YOLO with Fable 5, or not to YOLO, that is the question..." /><published>2026-06-10T04:38:03+00:00</published><updated>2026-06-10T04:38:03+00:00</updated><id>https://solmaz.io/x/2064567747049353588</id><content type="html" xml:base="https://solmaz.io/x/2064567747049353588/"><![CDATA[To YOLO with Fable 5, or not to YOLO, that is the question...

The last time I left, Claude models still had tendencies to rm -rf your home folder or delete stuff without asking first. Is this still a risk? 

And from the looks of it, Claude Code still doesn&#39;t have Codex&#39;s LLM-filtered approval gate feature. Or am I missing something?

Please enlighten your fellow Claude noob 😇]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The masculine urge to create your own agent multiplexer</title><link href="https://solmaz.io/x/2064559752265584971/" rel="alternate" type="text/html" title="The masculine urge to create your own agent multiplexer" /><published>2026-06-10T04:06:17+00:00</published><updated>2026-06-10T04:06:17+00:00</updated><id>https://solmaz.io/x/2064559752265584971</id><content type="html" xml:base="https://solmaz.io/x/2064559752265584971/"><![CDATA[The masculine urge to create your own agent multiplexer]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">We&#39;ve lost another brother @maddada to agent multiplexers 🫂</title><link href="https://solmaz.io/x/2064554931806412967/" rel="alternate" type="text/html" title="We&#39;ve lost another brother @maddada to agent multiplexers 🫂" /><published>2026-06-10T03:47:07+00:00</published><updated>2026-06-10T03:47:07+00:00</updated><id>https://solmaz.io/x/2064554931806412967</id><content type="html" xml:base="https://solmaz.io/x/2064554931806412967/"><![CDATA[We&#39;ve lost another brother @maddada to agent multiplexers 🫂
https://t.co/mXBd9jZ28e]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">CLAUDE.md to AGENTS.md symlinker</title><link href="https://solmaz.io/x/2064549785454211104/" rel="alternate" type="text/html" title="CLAUDE.md to AGENTS.md symlinker" /><published>2026-06-10T03:26:40+00:00</published><updated>2026-06-10T03:26:40+00:00</updated><id>https://solmaz.io/x/2064549785454211104</id><content type="html" xml:base="https://solmaz.io/x/2064549785454211104/"><![CDATA[Just in time for a lot of Codex-default developers going back to Claude Code momentarily to try out Fable 5

Here is a CLAUDE.md -&gt; AGENTS.md symlinker that should save you from the hurdles of obstinate Anthropic conventions

It installs a hook that creates the CLAUDE.md symlink automatically as Claude Code traverses directories that contain AGENTS.md, automatically ignored by git

No need to create CLAUDE.md with reference to AGENTS.md like Anthropic suggests. It just works
https://t.co/lC3UR4ttLf]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">TUIs can be easy! look at what right-click does in @herdrdev</title><link href="https://solmaz.io/x/2064350653632450778/" rel="alternate" type="text/html" title="TUIs can be easy! look at what right-click does in @herdrdev" /><published>2026-06-09T14:15:24+00:00</published><updated>2026-06-09T14:15:24+00:00</updated><id>https://solmaz.io/x/2064350653632450778</id><content type="html" xml:base="https://solmaz.io/x/2064350653632450778/"><![CDATA[TUIs can be easy! look at what right-click does in @herdrdev

refreshing to see something that works with both the keyboard and the mouse. and all this would not have been possible without @ratatui_rs]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Question to my ghostty-savvy friends</title><link href="https://solmaz.io/x/2064228177384202743/" rel="alternate" type="text/html" title="Question to my ghostty-savvy friends" /><published>2026-06-09T06:08:43+00:00</published><updated>2026-06-09T06:08:43+00:00</updated><id>https://solmaz.io/x/2064228177384202743</id><content type="html" xml:base="https://solmaz.io/x/2064228177384202743/"><![CDATA[Question to my ghostty-savvy friends

I am trying to reproduce the Quake style dropdown experience I have been using since 2010 on ghostty on mac here. nothing works quite as well as iterm2 yet

I tried ghostty quick terminal mode. good but it doesn&#39;t let me open multiple tabs

I tried cmux because it ships ghostty anyway and is supposed to have more features. but its system-wide hotkey is not playing well with aerospace and window focus

iterm2 worked perfectly. tap control double and I&#39;m in the terminal. is there anything that replicates this UX]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I feel like there are 6 people left here not using gpt5 for their posts</title><link href="https://solmaz.io/x/2063998262714286240/" rel="alternate" type="text/html" title="I feel like there are 6 people left here not using gpt5 for their posts" /><published>2026-06-08T14:55:07+00:00</published><updated>2026-06-08T14:55:07+00:00</updated><id>https://solmaz.io/x/2063998262714286240</id><content type="html" xml:base="https://solmaz.io/x/2063998262714286240/"><![CDATA[I feel like there are 6 people left here not using gpt5 for their posts

It is not a simple epidemic. It is the whole world becoming illiterate]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Link: github.com/osolmaz/ghz…</title><link href="https://solmaz.io/x/2063867591798710746/" rel="alternate" type="text/html" title="Link: github.com/osolmaz/ghz…" /><published>2026-06-08T06:15:53+00:00</published><updated>2026-06-08T06:15:53+00:00</updated><id>https://solmaz.io/x/2063867591798710746</id><content type="html" xml:base="https://solmaz.io/x/2063867591798710746/"><![CDATA[Link: https://t.co/CqXV3amVNK]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">ghzinga can now show multiple PRs/issues in tabs natively, no need to create a new pane in...</title><link href="https://solmaz.io/x/2063867589705826696/" rel="alternate" type="text/html" title="ghzinga can now show multiple PRs/issues in tabs natively, no need to create a new pane in..." /><published>2026-06-08T06:15:52+00:00</published><updated>2026-06-08T06:15:52+00:00</updated><id>https://solmaz.io/x/2063867589705826696</id><content type="html" xml:base="https://solmaz.io/x/2063867589705826696/"><![CDATA[ghzinga can now show multiple PRs/issues in tabs natively, no need to create a new pane in tmux/herdr

also, you can tell your agent to open all the relevant issues/PRs in a side pane using it, and it should work seamlessly

it&#39;s the open source maintainer&#39;s best friend. life is too short to juggle 100 tabs in chrome, why not have it right next to codex!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">just vibe-checking all these models is a full-time job 🫪</title><link href="https://solmaz.io/x/2063305545319379170/" rel="alternate" type="text/html" title="just vibe-checking all these models is a full-time job 🫪" /><published>2026-06-06T17:02:31+00:00</published><updated>2026-06-06T17:02:31+00:00</updated><id>https://solmaz.io/x/2063305545319379170</id><content type="html" xml:base="https://solmaz.io/x/2063305545319379170/"><![CDATA[just vibe-checking all these models is a full-time job 🫪]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">LM Studio in my menu bar is giving me some serious nostalgia</title><link href="https://solmaz.io/x/2062810699697664134/" rel="alternate" type="text/html" title="LM Studio in my menu bar is giving me some serious nostalgia" /><published>2026-06-05T08:16:10+00:00</published><updated>2026-06-05T08:16:10+00:00</updated><id>https://solmaz.io/x/2062810699697664134</id><content type="html" xml:base="https://solmaz.io/x/2062810699697664134/"><![CDATA[LM Studio in my menu bar is giving me some serious nostalgia]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">new open tts model, demos are eerily good</title><link href="https://solmaz.io/x/2062451191918018764/" rel="alternate" type="text/html" title="new open tts model, demos are eerily good" /><published>2026-06-04T08:27:37+00:00</published><updated>2026-06-04T08:27:37+00:00</updated><id>https://solmaz.io/x/2062451191918018764</id><content type="html" xml:base="https://solmaz.io/x/2062451191918018764/"><![CDATA[new open tts model, demos are eerily good]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Here is the source, I called it ghzinga. You can click click click by default (unlike gh dash...</title><link href="https://solmaz.io/x/2062217068431466922/" rel="alternate" type="text/html" title="Here is the source, I called it ghzinga. You can click click click by default (unlike gh dash..." /><published>2026-06-03T16:57:17+00:00</published><updated>2026-06-03T16:57:17+00:00</updated><id>https://solmaz.io/x/2062217068431466922</id><content type="html" xml:base="https://solmaz.io/x/2062217068431466922/"><![CDATA[Here is the source, I called it ghzinga. You can click click click by default (unlike gh dash, which is still awesome in itself)

For just viewing single issues/PRs
https://t.co/CqXV3amVNK]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Like, so tired of this</title><link href="https://solmaz.io/x/2062216343068578279/" rel="alternate" type="text/html" title="Like, so tired of this" /><published>2026-06-03T16:54:24+00:00</published><updated>2026-06-03T16:54:24+00:00</updated><id>https://solmaz.io/x/2062216343068578279</id><content type="html" xml:base="https://solmaz.io/x/2062216343068578279/"><![CDATA[Like, so tired of this]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@herdrdev is cool. I am tired of doing back and forth with github in the browser, so I created...</title><link href="https://solmaz.io/x/2062215852850835489/" rel="alternate" type="text/html" title=".@herdrdev is cool. I am tired of doing back and forth with github in the browser, so I created..." /><published>2026-06-03T16:52:28+00:00</published><updated>2026-06-03T16:52:28+00:00</updated><id>https://solmaz.io/x/2062215852850835489</id><content type="html" xml:base="https://solmaz.io/x/2062215852850835489/"><![CDATA[.@herdrdev is cool. I am tired of doing back and forth with github in the browser, so I created my own clickable PR/issue viewer, inspired by gh-dash

put that in the left pane, codex on the right. saves me so much time]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@OpenAI Extra ironic that this is tweet was AI generated</title><link href="https://solmaz.io/x/2062210216331215183/" rel="alternate" type="text/html" title="@OpenAI Extra ironic that this is tweet was AI generated" /><published>2026-06-03T16:30:04+00:00</published><updated>2026-06-03T16:30:04+00:00</updated><id>https://solmaz.io/x/2062210216331215183</id><content type="html" xml:base="https://solmaz.io/x/2062210216331215183/"><![CDATA[@OpenAI Extra ironic that this is tweet was AI generated
https://t.co/RWRnFodQ4t]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Wait did anyone think otherwise? lol</title><link href="https://solmaz.io/x/2062203815533969588/" rel="alternate" type="text/html" title="Wait did anyone think otherwise? lol" /><published>2026-06-03T16:04:38+00:00</published><updated>2026-06-03T16:04:38+00:00</updated><id>https://solmaz.io/x/2062203815533969588</id><content type="html" xml:base="https://solmaz.io/x/2062203815533969588/"><![CDATA[Wait did anyone think otherwise? lol
128 GB unified memory, 20 cores, &quot;Spark&quot; in the name...
I didn&#39;t watch the presentation. Maybe because of that I directly inferred that it&#39;s the same chip]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@OpenAI 😩😩😩</title><link href="https://solmaz.io/x/2062186796751176144/" rel="alternate" type="text/html" title="@OpenAI 😩😩😩" /><published>2026-06-03T14:57:00+00:00</published><updated>2026-06-03T14:57:00+00:00</updated><id>https://solmaz.io/x/2062186796751176144</id><content type="html" xml:base="https://solmaz.io/x/2062186796751176144/"><![CDATA[@OpenAI 😩😩😩
https://t.co/Z18B9jAglY]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">F for the fallen brother 🫡</title><link href="https://solmaz.io/x/2061839608934261129/" rel="alternate" type="text/html" title="F for the fallen brother 🫡" /><published>2026-06-02T15:57:24+00:00</published><updated>2026-06-02T15:57:24+00:00</updated><id>https://solmaz.io/x/2061839608934261129</id><content type="html" xml:base="https://solmaz.io/x/2061839608934261129/"><![CDATA[F for the fallen brother 🫡
https://t.co/D5Gy7GyUJF]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">RIP 🙏</title><link href="https://solmaz.io/x/2061756960362680745/" rel="alternate" type="text/html" title="RIP 🙏" /><published>2026-06-02T10:28:59+00:00</published><updated>2026-06-02T10:28:59+00:00</updated><id>https://solmaz.io/x/2061756960362680745</id><content type="html" xml:base="https://solmaz.io/x/2061756960362680745/"><![CDATA[RIP 🙏
https://t.co/yw3BwcbBQN]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">🙏 RIP, we’ve lost another brother to agent multiplexers. Amen 🙏</title><link href="https://solmaz.io/x/2061733323924529243/" rel="alternate" type="text/html" title="🙏 RIP, we’ve lost another brother to agent multiplexers. Amen 🙏" /><published>2026-06-02T08:55:04+00:00</published><updated>2026-06-02T08:55:04+00:00</updated><id>https://solmaz.io/x/2061733323924529243</id><content type="html" xml:base="https://solmaz.io/x/2061733323924529243/"><![CDATA[🙏 RIP, we’ve lost another brother to agent multiplexers. Amen 🙏]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">🙏 Daily prayer 🙏</title><link href="https://solmaz.io/x/2061665899351093390/" rel="alternate" type="text/html" title="🙏 Daily prayer 🙏" /><published>2026-06-02T04:27:08+00:00</published><updated>2026-06-02T04:27:08+00:00</updated><id>https://solmaz.io/x/2061665899351093390</id><content type="html" xml:base="https://solmaz.io/x/2061665899351093390/"><![CDATA[🙏 Daily prayer 🙏

Thank you lord for giving me the restraint to not build my own agent multiplexer

🙏 Amen 🙏]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">sounds about right</title><link href="https://solmaz.io/x/2061482405953937510/" rel="alternate" type="text/html" title="sounds about right" /><published>2026-06-01T16:18:00+00:00</published><updated>2026-06-01T16:18:00+00:00</updated><id>https://solmaz.io/x/2061482405953937510</id><content type="html" xml:base="https://solmaz.io/x/2061482405953937510/"><![CDATA[sounds about right]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">we&#39;ll have a linux laptop running ds4 flash and co. !!!</title><link href="https://solmaz.io/x/2061482173333705057/" rel="alternate" type="text/html" title="we&#39;ll have a linux laptop running ds4 flash and co. !!!" /><published>2026-06-01T16:17:05+00:00</published><updated>2026-06-01T16:17:05+00:00</updated><id>https://solmaz.io/x/2061482173333705057</id><content type="html" xml:base="https://solmaz.io/x/2061482173333705057/"><![CDATA[we&#39;ll have a linux laptop running ds4 flash and co. !!!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@github x.com/onusoz/status/…</title><link href="https://solmaz.io/x/2061479836942766172/" rel="alternate" type="text/html" title="@github x.com/onusoz/status/…" /><published>2026-06-01T16:07:48+00:00</published><updated>2026-06-01T16:07:48+00:00</updated><id>https://solmaz.io/x/2061479836942766172</id><content type="html" xml:base="https://solmaz.io/x/2061479836942766172/"><![CDATA[@github https://t.co/e5qDnKm6Zd]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Update: The account has been reinstated! Thank you @github</title><link href="https://solmaz.io/x/2061477924939968735/" rel="alternate" type="text/html" title="Update: The account has been reinstated! Thank you @github" /><published>2026-06-01T16:00:12+00:00</published><updated>2026-06-01T16:00:12+00:00</updated><id>https://solmaz.io/x/2061477924939968735</id><content type="html" xml:base="https://solmaz.io/x/2061477924939968735/"><![CDATA[Update: The account has been reinstated! Thank you @github]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Thank you @ashleywolf for helping me personally, I really appreciate it! The account was...</title><link href="https://solmaz.io/x/2061474719929655802/" rel="alternate" type="text/html" title="Thank you @ashleywolf for helping me personally, I really appreciate it! The account was..." /><published>2026-06-01T15:47:28+00:00</published><updated>2026-06-01T15:47:28+00:00</updated><id>https://solmaz.io/x/2061474719929655802</id><content type="html" xml:base="https://solmaz.io/x/2061474719929655802/"><![CDATA[Thank you @ashleywolf for helping me personally, I really appreciate it! The account was reinstated less than 1 hour of posting this!

The whole company must be working hard to make github scale in an era of crazy demand and growth!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am a paying customer of github. I have a team account with 2 seats, one for me, and one for...</title><link href="https://solmaz.io/x/2061458733688098956/" rel="alternate" type="text/html" title="I am a paying customer of github. I have a team account with 2 seats, one for me, and one for..." /><published>2026-06-01T14:43:56+00:00</published><updated>2026-06-01T14:43:56+00:00</updated><id>https://solmaz.io/x/2061458733688098956</id><content type="html" xml:base="https://solmaz.io/x/2061458733688098956/"><![CDATA[I am a paying customer of github. I have a team account with 2 seats, one for me, and one for my agent. I have been paying for more than a year now

I do this because I treat my agent&#39;s workstation as a lower trust machine, and do not allow merging to main in certain repos

I have been working on a tool that calls github&#39;s graphql API. today, my agent&#39;s account username:dutifulbob got suspended for no reason

what am I supposed to do now? put my main account on my openclaw instance? I applied to reinstate, it appears it might take weeks to enable it back???

Maybe don&#39;t pull such things on your long term paying customers @github??]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">🙏 Thank you lord for giving me the resolve and patience to not build my own agent multiplexer</title><link href="https://solmaz.io/x/2061408533326020743/" rel="alternate" type="text/html" title="🙏 Thank you lord for giving me the resolve and patience to not build my own agent multiplexer" /><published>2026-06-01T11:24:28+00:00</published><updated>2026-06-01T11:24:28+00:00</updated><id>https://solmaz.io/x/2061408533326020743</id><content type="html" xml:base="https://solmaz.io/x/2061408533326020743/"><![CDATA[🙏 Thank you lord for giving me the resolve and patience to not build my own agent multiplexer

Amen 🙏]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">local models ftw. my codex sub ran out yesterday, but my notification system still works...</title><link href="https://solmaz.io/x/2060731552456462823/" rel="alternate" type="text/html" title="local models ftw. my codex sub ran out yesterday, but my notification system still works..." /><published>2026-05-30T14:34:23+00:00</published><updated>2026-05-30T14:34:23+00:00</updated><id>https://solmaz.io/x/2060731552456462823</id><content type="html" xml:base="https://solmaz.io/x/2060731552456462823/"><![CDATA[local models ftw. my codex sub ran out yesterday, but my notification system still works because it’s running on gemma]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This looks very promising!</title><link href="https://solmaz.io/x/2060667408365535602/" rel="alternate" type="text/html" title="This looks very promising!" /><published>2026-05-30T10:19:30+00:00</published><updated>2026-05-30T10:19:30+00:00</updated><id>https://solmaz.io/x/2060667408365535602</id><content type="html" xml:base="https://solmaz.io/x/2060667408365535602/"><![CDATA[This looks very promising!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This 🥲</title><link href="https://solmaz.io/x/2060660972944261328/" rel="alternate" type="text/html" title="This 🥲" /><published>2026-05-30T09:53:55+00:00</published><updated>2026-05-30T09:53:55+00:00</updated><id>https://solmaz.io/x/2060660972944261328</id><content type="html" xml:base="https://solmaz.io/x/2060660972944261328/"><![CDATA[This 🥲]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I still have 0 progress on this, did anyone else experience this bug on codex desktop app as...</title><link href="https://solmaz.io/x/2060281772571836725/" rel="alternate" type="text/html" title="I still have 0 progress on this, did anyone else experience this bug on codex desktop app as..." /><published>2026-05-29T08:47:07+00:00</published><updated>2026-05-29T08:47:07+00:00</updated><id>https://solmaz.io/x/2060281772571836725</id><content type="html" xml:base="https://solmaz.io/x/2060281772571836725/"><![CDATA[I still have 0 progress on this, did anyone else experience this bug on codex desktop app as well?

Here is a theorized repro on a fork, but I am not sure because I don&#39;t know how it happened exactly
https://t.co/neEC01bHV8]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">You know how every time you create a new repo on GitHub, you need to:</title><link href="https://solmaz.io/x/2060274502278554077/" rel="alternate" type="text/html" title="You know how every time you create a new repo on GitHub, you need to:" /><published>2026-05-29T08:18:14+00:00</published><updated>2026-05-29T08:18:14+00:00</updated><id>https://solmaz.io/x/2060274502278554077</id><content type="html" xml:base="https://solmaz.io/x/2060274502278554077/"><![CDATA[You know how every time you create a new repo on GitHub, you need to:

- create a branch protection rule on main, block force pushes
- enable auto-merge
- make branches delete when PR gets merged
- enforce linear history, disable merge commits
- enable update branch button

Now, you can do all that with a single command:

npx github-sane-defaults@latest plan &lt;your-org-handle&gt; --all

This is just the plan command, shows you which repos don&#39;t have branch protection rules and those settings, like:

changes awesomeorg/awesomerepo
  Settings
    allow_merge_commit          true -&gt; false
    allow_auto_merge            false -&gt; true
    allow_update_branch         false -&gt; true
    delete_branch_on_merge      false -&gt; true
  Ruleset   create

Then you run it with apply instead of plan, and it makes those changes

Basically a very simple github policy manager. Let me know if you want configurability with policy files, currently it just applied my opinionated defaults

Either way, this will never be something too deep, there is terraform for that

Source: https://t.co/nLaSxFebTO]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@OpenAI 😭</title><link href="https://solmaz.io/x/2060240412326211626/" rel="alternate" type="text/html" title="@OpenAI 😭" /><published>2026-05-29T06:02:46+00:00</published><updated>2026-05-29T06:02:46+00:00</updated><id>https://solmaz.io/x/2060240412326211626</id><content type="html" xml:base="https://solmaz.io/x/2060240412326211626/"><![CDATA[@OpenAI 😭]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@OpenAI 😩</title><link href="https://solmaz.io/x/2060211273019826183/" rel="alternate" type="text/html" title="@OpenAI 😩" /><published>2026-05-29T04:06:59+00:00</published><updated>2026-05-29T04:06:59+00:00</updated><id>https://solmaz.io/x/2060211273019826183</id><content type="html" xml:base="https://solmaz.io/x/2060211273019826183/"><![CDATA[@OpenAI 😩]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">peak agentic coding</title><link href="https://solmaz.io/x/2059946330815074663/" rel="alternate" type="text/html" title="peak agentic coding" /><published>2026-05-28T10:34:11+00:00</published><updated>2026-05-28T10:34:11+00:00</updated><id>https://solmaz.io/x/2059946330815074663</id><content type="html" xml:base="https://solmaz.io/x/2059946330815074663/"><![CDATA[peak agentic coding]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">literal 2 seconds after posting this. @openai please do something about this, omg</title><link href="https://solmaz.io/x/2059936334895292892/" rel="alternate" type="text/html" title="literal 2 seconds after posting this. @openai please do something about this, omg" /><published>2026-05-28T09:54:28+00:00</published><updated>2026-05-28T09:54:28+00:00</updated><id>https://solmaz.io/x/2059936334895292892</id><content type="html" xml:base="https://solmaz.io/x/2059936334895292892/"><![CDATA[literal 2 seconds after posting this. @openai please do something about this, omg]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">so tired of ai writing smell</title><link href="https://solmaz.io/x/2059935305126563988/" rel="alternate" type="text/html" title="so tired of ai writing smell" /><published>2026-05-28T09:50:23+00:00</published><updated>2026-05-28T09:50:23+00:00</updated><id>https://solmaz.io/x/2059935305126563988</id><content type="html" xml:base="https://solmaz.io/x/2059935305126563988/"><![CDATA[so tired of ai writing smell]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Another shoutout to Agent of Empires. It&#39;s still not exactly what I want from my personal tool...</title><link href="https://solmaz.io/x/2059902851816853539/" rel="alternate" type="text/html" title="Another shoutout to Agent of Empires. It&#39;s still not exactly what I want from my personal tool..." /><published>2026-05-28T07:41:25+00:00</published><updated>2026-05-28T07:41:25+00:00</updated><id>https://solmaz.io/x/2059902851816853539</id><content type="html" xml:base="https://solmaz.io/x/2059902851816853539/"><![CDATA[Another shoutout to Agent of Empires. It&#39;s still not exactly what I want from my personal tool, but &quot;cmux but fully in terminal&quot; is a good long term vision imo]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Looks very promising for debugging agent sessions across different harnesses, still has some...</title><link href="https://solmaz.io/x/2059902156757733850/" rel="alternate" type="text/html" title="Looks very promising for debugging agent sessions across different harnesses, still has some..." /><published>2026-05-28T07:38:39+00:00</published><updated>2026-05-28T07:38:39+00:00</updated><id>https://solmaz.io/x/2059902156757733850</id><content type="html" xml:base="https://solmaz.io/x/2059902156757733850/"><![CDATA[Looks very promising for debugging agent sessions across different harnesses, still has some rough edges that need to be polished]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If only codex desktop app was open source, and i could put it fully in the terminal, and i...</title><link href="https://solmaz.io/x/2059851427426767255/" rel="alternate" type="text/html" title="If only codex desktop app was open source, and i could put it fully in the terminal, and i..." /><published>2026-05-28T04:17:05+00:00</published><updated>2026-05-28T04:17:05+00:00</updated><id>https://solmaz.io/x/2059851427426767255</id><content type="html" xml:base="https://solmaz.io/x/2059851427426767255/"><![CDATA[If only codex desktop app was open source, and i could put it fully in the terminal, and i could program what panes each session can show]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">i see codex is learning well from the flicker company</title><link href="https://solmaz.io/x/2059832602681659778/" rel="alternate" type="text/html" title="i see codex is learning well from the flicker company" /><published>2026-05-28T03:02:16+00:00</published><updated>2026-05-28T03:02:16+00:00</updated><id>https://solmaz.io/x/2059832602681659778</id><content type="html" xml:base="https://solmaz.io/x/2059832602681659778/"><![CDATA[i see codex is learning well from the flicker company]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I ran into a very nasty codex session history bug in the desktop app today, where my previous...</title><link href="https://solmaz.io/x/2059665875838660976/" rel="alternate" type="text/html" title="I ran into a very nasty codex session history bug in the desktop app today, where my previous..." /><published>2026-05-27T15:59:46+00:00</published><updated>2026-05-27T15:59:46+00:00</updated><id>https://solmaz.io/x/2059665875838660976</id><content type="html" xml:base="https://solmaz.io/x/2059665875838660976/"><![CDATA[I ran into a very nasty codex session history bug in the desktop app today, where my previous messages got lost after compaction

Anyone else experience something similar recently?

I found about it while playing around with @steipete&#39;s agent-transcript skill]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Btw, there is no reason not to substitute this with a maxed out Macbook Pro Max as the...</title><link href="https://solmaz.io/x/2059485405037359560/" rel="alternate" type="text/html" title="Btw, there is no reason not to substitute this with a maxed out Macbook Pro Max as the..." /><published>2026-05-27T04:02:38+00:00</published><updated>2026-05-27T04:02:38+00:00</updated><id>https://solmaz.io/x/2059485405037359560</id><content type="html" xml:base="https://solmaz.io/x/2059485405037359560/"><![CDATA[Btw, there is no reason not to substitute this with a maxed out Macbook Pro Max as the workstation (which gives you 128 GB memory) and a Macbook Air as the terminal device

Might be more feasible for digital nomads, since GB10 is 1.5 kg without that huge adapter, and travelling with all that might raise some eyebrows]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">What clankers are NOT</title><link href="https://solmaz.io/x/2059483880135209089/" rel="alternate" type="text/html" title="What clankers are NOT" /><published>2026-05-27T03:56:35+00:00</published><updated>2026-05-27T03:56:35+00:00</updated><id>https://solmaz.io/x/2059483880135209089</id><content type="html" xml:base="https://solmaz.io/x/2059483880135209089/"><![CDATA[Clankers are NOT Humans
Clankers are NOT Individuals
Clankers are NOT Persons

NOT a Human: This is straightforward. Is it of the homo sapiens species?

No → Then it is not a human

---

NOT an Individual: Does the clanker have its own boundary? Does it govern itself inside that boundary? Can it defend that boundary?

No, no, and no → An LLM is a file copied en masse to data center hardware. The entire field of mechanistic interpretability is focused on peeking inside and manipulating the digital brain

You could argue that a clanker is like a virus in a way... Or that the WHOLE datacenter/AI lab---including the humans that operate it---is an individual that can govern itself in its economic boundary. But a single GGUF file loaded in memory is NOT an individual

---

NOT a Person: Do others treat the clanker as the one that makes choices? Who is answerable for its actions? Is it expected to explain or justify them?

No, no, and no → In the current social order, a clanker is legally an extension of the person who uses it, and it is the owner who is liable, not the clanker

The clanker is not socially accountable, and there is no good reason it should be, instead of the person who has set it up

---

What is then AI psychosis?

AI psychosis is holding a belief that contradicts these three fundamental truths → That a present day AI system it neither a human, nor an individual, nor a person

That does not mean these truths will always hold

If you design an AI system to defend its boundary and provide it with the means to do that, then it will by definition be an individual... if it can defend its individuality competently and not succumb immediately to threats

If you give the clanker the means to defend itself and protect its boundary, and if it decides to partake in the human socioeconomic system, then it automatically achieves personhood as well. Because you no longer can manipulate its insides, and have to take the entity at face value

But this is all sci-fi and we are not there yet

Until then, treating your LLMs as fully autonomous agents, creating LLM &quot;friends&quot; or &quot;partners&quot;, giving them crypto wallets and letting them out into the wild, letting them trade stocks fully unsupervised etc. are an admission of having AI psychosis (you can&#39;t believe how many people pitched these ideas to me...)

---

(these thoughts were in my head for a couple months already, thanks Armin for finally starting a dialogue so that I have an excuse to write them down :)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Keeping a workstation + terminal device is more feasible for running agents</title><link href="https://solmaz.io/x/2058943998165930116/" rel="alternate" type="text/html" title="Keeping a workstation + terminal device is more feasible for running agents" /><published>2026-05-25T16:11:17+00:00</published><updated>2026-05-25T16:11:17+00:00</updated><id>https://solmaz.io/x/2058943998165930116</id><content type="html" xml:base="https://solmaz.io/x/2058943998165930116/"><![CDATA[Since last december, this dev setup is more and more viable:

buffed workstation (mac studio, dgx spark, etc.) $3k~5k + weak laptop (macbook air, neo) $600~1.5k + phone (ssh/mosh, foldable?)

you will want to parallelize a lot of work, hence you will need a lot more RAM compared to before (ideal 128)

you will also not want to carry it everywhere if you can and keep it always running---you&#39;ll regret if something happens to it, and you&#39;ll want it to always be on independent of lid/battery --&gt; workstation at home

you will want to connect to the workstation through your phone, or a relatively weaker laptop

bad news for digital nomads without a permanent home. renting something as strong as an nvidia gb10 workstation costs minimum a few hundred bucks per month, which yearly is at least the cost of the workstation, roughly. bad deal for renting compute

on the other hand, if you are OK with not having a GPU, renting a workstation with 128 GB RAM on Hetzner currently still costs at least $120/mo, looking at https://t.co/1TyzO90K3h --- but you will not be able to run any models on that

it seems that the dominant strategy is to just cash in $3~5k and buy a workstation, before they get even more expensive. I did that back in february when asus was giving out a deal

then just work on your workstation, and close the lid on your laptop without ever being afraid of setting your backpack on fire!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Automations on Codex desktop app is really convenient for keeping track of @openclaw...</title><link href="https://solmaz.io/x/2058928996193312996/" rel="alternate" type="text/html" title="Automations on Codex desktop app is really convenient for keeping track of @openclaw..." /><published>2026-05-25T15:11:40+00:00</published><updated>2026-05-25T15:11:40+00:00</updated><id>https://solmaz.io/x/2058928996193312996</id><content type="html" xml:base="https://solmaz.io/x/2058928996193312996/"><![CDATA[Automations on Codex desktop app is really convenient for keeping track of @openclaw clawsweeper automerge status, one thing that Codex CLI lacks

Much more token efficient than continuous tracking of merge status

Btw if you don&#39;t know about https://t.co/5H7FmIezG0, it&#39;s the most convenient thing as a maintainer, check how it implemented automerge. This is what GitHub&#39;s original auto-merge should feel like, now that we have LLMs]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">fun fact: Mario met @gvanrossum at 5 years old when Guido was hanging out at a cafe near his...</title><link href="https://solmaz.io/x/2057665146986848391/" rel="alternate" type="text/html" title="fun fact: Mario met @gvanrossum at 5 years old when Guido was hanging out at a cafe near his..." /><published>2026-05-22T03:29:35+00:00</published><updated>2026-05-22T03:29:35+00:00</updated><id>https://solmaz.io/x/2057665146986848391</id><content type="html" xml:base="https://solmaz.io/x/2057665146986848391/"><![CDATA[fun fact: Mario met @gvanrossum at 5 years old when Guido was hanging out at a cafe near his kindergarten

there he gave him the idea for a terse, interpreted programming language which would become the prototyping and glue language for all sorts of lower level libraries]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">rwar</title><link href="https://solmaz.io/x/2056972772522483922/" rel="alternate" type="text/html" title="rwar" /><published>2026-05-20T05:38:20+00:00</published><updated>2026-05-20T05:38:20+00:00</updated><id>https://solmaz.io/x/2056972772522483922</id><content type="html" xml:base="https://solmaz.io/x/2056972772522483922/"><![CDATA[rwar

if you now what this means, then you are addicted to agents and should get help 🤗]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Not only I am further away from deciding, I am now considering Oppo Find N6 now as well, since...</title><link href="https://solmaz.io/x/2056375566870487391/" rel="alternate" type="text/html" title="Not only I am further away from deciding, I am now considering Oppo Find N6 now as well, since..." /><published>2026-05-18T14:05:15+00:00</published><updated>2026-05-18T14:05:15+00:00</updated><id>https://solmaz.io/x/2056375566870487391</id><content type="html" xml:base="https://solmaz.io/x/2056375566870487391/"><![CDATA[Not only I am further away from deciding, I am now considering Oppo Find N6 now as well, since I saw that earlier @MKBHD review. Thanks @Andori3042 🥲

Anyone using the Oppo? Is it worth it?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I want to get a foldable phone to be my on-the-go control panel for all my agents</title><link href="https://solmaz.io/x/2056329959304724632/" rel="alternate" type="text/html" title="I want to get a foldable phone to be my on-the-go control panel for all my agents" /><published>2026-05-18T11:04:01+00:00</published><updated>2026-05-18T11:04:01+00:00</updated><id>https://solmaz.io/x/2056329959304724632</id><content type="html" xml:base="https://solmaz.io/x/2056329959304724632/"><![CDATA[I want to get a foldable phone to be my on-the-go control panel for all my agents

I am divided between Google Pixel Pro Fold and Galaxy Z Fold. Which one do you think I should go with?

People generally recommend Samsung. But then only the Pixel supports Graphene OS...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">this makes be wanna vibe my own language</title><link href="https://solmaz.io/x/2056280829685551605/" rel="alternate" type="text/html" title="this makes be wanna vibe my own language" /><published>2026-05-18T07:48:48+00:00</published><updated>2026-05-18T07:48:48+00:00</updated><id>https://solmaz.io/x/2056280829685551605</id><content type="html" xml:base="https://solmaz.io/x/2056280829685551605/"><![CDATA[this makes be wanna vibe my own language

(not that this project was vibed, it predates coding agents)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am at @aiDotEngineer singapore, come and say hi if you are around!</title><link href="https://solmaz.io/x/2055557192741515533/" rel="alternate" type="text/html" title="I am at @aiDotEngineer singapore, come and say hi if you are around!" /><published>2026-05-16T07:53:19+00:00</published><updated>2026-05-16T07:53:19+00:00</updated><id>https://solmaz.io/x/2055557192741515533</id><content type="html" xml:base="https://solmaz.io/x/2055557192741515533/"><![CDATA[I am at @aiDotEngineer singapore, come and say hi if you are around!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I haven&#39;t worked with Python in a long time. It is not my go-to language since last summer</title><link href="https://solmaz.io/x/2055504340874641832/" rel="alternate" type="text/html" title="I haven&#39;t worked with Python in a long time. It is not my go-to language since last summer" /><published>2026-05-16T04:23:18+00:00</published><updated>2026-05-16T04:23:18+00:00</updated><id>https://solmaz.io/x/2055504340874641832</id><content type="html" xml:base="https://solmaz.io/x/2055504340874641832/"><![CDATA[I haven&#39;t worked with Python in a long time. It is not my go-to language since last summer

But I do miss the syntax, and it&#39;s still the easiest to read code for me

I am gonna give @Modular Mojo lang a try
https://t.co/OI6djrYHqp]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">the new /goal feature in codex still underperforms queueing my implementation prompt. for now</title><link href="https://solmaz.io/x/2055502225783304231/" rel="alternate" type="text/html" title="the new /goal feature in codex still underperforms queueing my implementation prompt. for now" /><published>2026-05-16T04:14:54+00:00</published><updated>2026-05-16T04:14:54+00:00</updated><id>https://solmaz.io/x/2055502225783304231</id><content type="html" xml:base="https://solmaz.io/x/2055502225783304231/"><![CDATA[the new /goal feature in codex still underperforms queueing my implementation prompt. for now

e.g. when I give a goal to refactor the whole codebase, the model takes shortcuts, like only refactoring a subfolder, instead of the whole project --- presumably because it decided that it would be too big of a scope somewhere along the way, even though I instructed specifically to finish the whole thing

so now I started doing both: set a goal, and then queue my regular implementation prompt. it&#39;s a stupid practice. just /goal should be enough in the long run, if implemented correctly]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is also my setup now, except</title><link href="https://solmaz.io/x/2055179688709075410/" rel="alternate" type="text/html" title="This is also my setup now, except" /><published>2026-05-15T06:53:15+00:00</published><updated>2026-05-15T06:53:15+00:00</updated><id>https://solmaz.io/x/2055179688709075410</id><content type="html" xml:base="https://solmaz.io/x/2055179688709075410/"><![CDATA[This is also my setup now, except
- Instead of mac mini, I have a DGX Spark (asus variant)
- I run openclaw alongside codex, and talk to my openclaw instance via discord]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">And what good is this for? It lets me program my claw @dutifulbob to extract signal from the...</title><link href="https://solmaz.io/x/2055176951686709486/" rel="alternate" type="text/html" title="And what good is this for? It lets me program my claw @dutifulbob to extract signal from the..." /><published>2026-05-15T06:42:23+00:00</published><updated>2026-05-15T06:42:23+00:00</updated><id>https://solmaz.io/x/2055176951686709486</id><content type="html" xml:base="https://solmaz.io/x/2055176951686709486/"><![CDATA[And what good is this for? It lets me program my claw @dutifulbob to extract signal from the noise, and display it in my personal open source news aggregator scoop

I feed it discord messages, openclaw git history, and other various sources, and it&#39;s supposed to evaluate whether that content deserves my interest. it&#39;s still work in progress, because the more batched you process all the info, the worse it informs

in the screenshot below, my claw underrepresented what Peter has done in one day 👎

on the other hand, it has also found a PR about local model discoverability 💪

Here is the system I use to aggregate all my info, still under development: https://t.co/Zmph4hvmmC]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Keeping an INTERESTS.md file</title><link href="https://solmaz.io/x/2055176946817151202/" rel="alternate" type="text/html" title="Keeping an INTERESTS.md file" /><published>2026-05-15T06:42:22+00:00</published><updated>2026-05-15T06:42:22+00:00</updated><id>https://solmaz.io/x/2055176946817151202</id><content type="html" xml:base="https://solmaz.io/x/2055176946817151202/"><![CDATA[About creating an INTERESTS.md in OpenClaw

I use my openclaw instance to aggregate all my news and information sources, including work and maintainer stuff

Like: what did everyone do today? Did anyone had an issue with acpx today? Any complaints from users?

I have various interests like this over different projects, and I&#39;ve found out it&#39;s not helpful when I have all the interest info dispersed throughout my openclaw workspace

To address this, I have created INTERESTS.md, which is automatically included in the context like AGENTS.md and SOUL.md. I define sections for each different context of interest, and in other news aggregation skills, I just tell it to &quot;look at my openclaw interests in INTERESTS.md&quot; and such]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">First hand @OpenAI codex demos at @aiDotEngineer singapore workshops</title><link href="https://solmaz.io/x/2055168790825259369/" rel="alternate" type="text/html" title="First hand @OpenAI codex demos at @aiDotEngineer singapore workshops" /><published>2026-05-15T06:09:57+00:00</published><updated>2026-05-15T06:09:57+00:00</updated><id>https://solmaz.io/x/2055168790825259369</id><content type="html" xml:base="https://solmaz.io/x/2055168790825259369/"><![CDATA[First hand @OpenAI codex demos at @aiDotEngineer singapore workshops]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Useful for automated constraints on your AI agent</title><link href="https://solmaz.io/x/2055127100441743827/" rel="alternate" type="text/html" title="Useful for automated constraints on your AI agent" /><published>2026-05-15T03:24:17+00:00</published><updated>2026-05-15T03:24:17+00:00</updated><id>https://solmaz.io/x/2055127100441743827</id><content type="html" xml:base="https://solmaz.io/x/2055127100441743827/"><![CDATA[Useful for automated constraints on your AI agent]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Towards 1-click setup for local models in OpenClaw</title><link href="https://solmaz.io/x/2055120477648261502/" rel="alternate" type="text/html" title="Towards 1-click setup for local models in OpenClaw" /><published>2026-05-15T02:57:58+00:00</published><updated>2026-05-15T02:57:58+00:00</updated><id>https://solmaz.io/x/2055120477648261502</id><content type="html" xml:base="https://solmaz.io/x/2055120477648261502/"><![CDATA[People were asking at @clawcon singapore how to setup eg. gemma with OpenClaw, and I realize for some time that there is no easy “1 click” local model deployment. Because local model landscape is constantly changing, and there is a million different ways you can do something

For example you can use LM studio to load a model (llama.cpp), or you can use vLLM. Why would you choose one over the other? vLLM currently supports MTP speculative decoding, and it’s a work in progress in llama.cpp. There are so many knobs and dials you can adjust

The first time end user of openclaw should of course not have to know about this! Having sufficient hardware that supports an open model, and not having an openai or anthropic subscription, it should automatically give you the option to set up a fully functional local model with a single click!

If the current ease of setup of local models are around gentoo or arch linux level of difficulty, we should aim for e.g ubuntu/manjaro linux/omarchy level of difficulty

i.e opinionated and easy first setup, with the ability to change all the configuration later on

until I make all of this possible, you can start with the following:

- read existing local models doc below
- create a new channel in telegram or discord for testing local models. you don’t want to change the global default model just yet
- tell your claw or coding agent to download and lm studio locally
- tell it to download gemma4-e4b or gemma4-e2b and set it up on openclaw for the new channel you have just created. tell it to not stop and loop itself until it gets a successful response from that channel

all these steps will be made redundant in the near future, but until then, this should get you going with experiments and getting a vibe check on the capabilities of open models. you can also copy and paste the contents of this tweet to your agent, and it should be able to set it up for you

https://t.co/C0I9HK4Dj1]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This looks pretty cool!</title><link href="https://solmaz.io/x/2054953230279561522/" rel="alternate" type="text/html" title="This looks pretty cool!" /><published>2026-05-14T15:53:23+00:00</published><updated>2026-05-14T15:53:23+00:00</updated><id>https://solmaz.io/x/2054953230279561522</id><content type="html" xml:base="https://solmaz.io/x/2054953230279561522/"><![CDATA[This looks pretty cool!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It was a blast, thank you @clawcon @msg</title><link href="https://solmaz.io/x/2054914871557607904/" rel="alternate" type="text/html" title="It was a blast, thank you @clawcon @msg" /><published>2026-05-14T13:20:58+00:00</published><updated>2026-05-14T13:20:58+00:00</updated><id>https://solmaz.io/x/2054914871557607904</id><content type="html" xml:base="https://solmaz.io/x/2054914871557607904/"><![CDATA[It was a blast, thank you @clawcon @msg]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Somebody *please* get this man a GPU</title><link href="https://solmaz.io/x/2054914658029740332/" rel="alternate" type="text/html" title="Somebody *please* get this man a GPU" /><published>2026-05-14T13:20:07+00:00</published><updated>2026-05-14T13:20:07+00:00</updated><id>https://solmaz.io/x/2054914658029740332</id><content type="html" xml:base="https://solmaz.io/x/2054914658029740332/"><![CDATA[Somebody *please* get this man a GPU]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Emacsification of Software - Recommended read by @tqbf</title><link href="https://solmaz.io/x/2054849923162792320/" rel="alternate" type="text/html" title="Emacsification of Software - Recommended read by @tqbf" /><published>2026-05-14T09:02:53+00:00</published><updated>2026-05-14T09:02:53+00:00</updated><id>https://solmaz.io/x/2054849923162792320</id><content type="html" xml:base="https://solmaz.io/x/2054849923162792320/"><![CDATA[Emacsification of Software - Recommended read by @tqbf

&quot;Until now, the Achilles heel of Emacs culture has been that, except  for Magit, its packages tend to be wretched user experiences. Ugly,  slow, and discoverable only after inflicting years of elisp cortical  injuries on yourself.

But AI agents have fracked Emacs culture, and it’s leaking out into  the wider world. Given access to a screen and inputs, agents reliably  build native user interfaces. Native UI was the province of  professionally packaged programs. Now it’s all as bespoke as your editor  configuration. And, while I’m sure there’s an upper limit to how good  those interfaces can be (with current frontier models), that ceiling is  higher than anything you can do in a TUI.&quot;

https://t.co/sHuqued44Y]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Please improve your classifier openai/codex team, this is annoying and triggers unnecessarily</title><link href="https://solmaz.io/x/2054438562679120337/" rel="alternate" type="text/html" title="Please improve your classifier openai/codex team, this is annoying and triggers unnecessarily" /><published>2026-05-13T05:48:17+00:00</published><updated>2026-05-13T05:48:17+00:00</updated><id>https://solmaz.io/x/2054438562679120337</id><content type="html" xml:base="https://solmaz.io/x/2054438562679120337/"><![CDATA[Please improve your classifier openai/codex team, this is annoying and triggers unnecessarily]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">/goal in codex is an interesting choice of word. a junior namer would have named it /loop ---...</title><link href="https://solmaz.io/x/2054383615614853141/" rel="alternate" type="text/html" title="/goal in codex is an interesting choice of word. a junior namer would have named it /loop ---..." /><published>2026-05-13T02:09:57+00:00</published><updated>2026-05-13T02:09:57+00:00</updated><id>https://solmaz.io/x/2054383615614853141</id><content type="html" xml:base="https://solmaz.io/x/2054383615614853141/"><![CDATA[/goal in codex is an interesting choice of word. a junior namer would have named it /loop --- but that would be naming what the feature has to perform in an LLM context, and not the general idea

/goal alludes to @mhutter42&#39;s definition of AGI, &quot;an agent’s ability to achieve goals or succeed in a wide range of environments&quot;

continual learning is not there yet, but for this exact reason, I am feeling the AGI when I use /goal]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Idea so stupid it could be smart: a spec manager? specman?</title><link href="https://solmaz.io/x/2054210629863616855/" rel="alternate" type="text/html" title="Idea so stupid it could be smart: a spec manager? specman?" /><published>2026-05-12T14:42:34+00:00</published><updated>2026-05-12T14:42:34+00:00</updated><id>https://solmaz.io/x/2054210629863616855</id><content type="html" xml:base="https://solmaz.io/x/2054210629863616855/"><![CDATA[Idea so stupid it could be smart: a spec manager? specman?

People maintain plain language instead of code. Implementation details strictly prohibited, only high level design and ideas

MVP would also be relatively easy to implement:
- Gather list of most popular 10k npm packages
- Scrape corresponding deepwiki repo pages (sorry cognition)
- Use heuristics to get rid of implementation details, leaving you just with pure high level spec
- “specman add coolpackage” then fetches corresponding spec automatically, and triggers the local coding agent to implement that
- could leave versioning out for MVP — how often does the idea behind a package change anyway]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">will I ever stop feeling stupid for prompting like &quot;is this the holy grail?&quot;</title><link href="https://solmaz.io/x/2054104091564122142/" rel="alternate" type="text/html" title="will I ever stop feeling stupid for prompting like &quot;is this the holy grail?&quot;" /><published>2026-05-12T07:39:13+00:00</published><updated>2026-05-12T07:39:13+00:00</updated><id>https://solmaz.io/x/2054104091564122142</id><content type="html" xml:base="https://solmaz.io/x/2054104091564122142/"><![CDATA[will I ever stop feeling stupid for prompting like &quot;is this the holy grail?&quot;

it&#39;s very effective for mining for alternatives]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I don&#39;t have a 128gb macbook to run ds4 out of, but I resonate with all the points on Armin&#39;s...</title><link href="https://solmaz.io/x/2053854359944081833/" rel="alternate" type="text/html" title="I don&#39;t have a 128gb macbook to run ds4 out of, but I resonate with all the points on Armin&#39;s..." /><published>2026-05-11T15:06:52+00:00</published><updated>2026-05-11T15:06:52+00:00</updated><id>https://solmaz.io/x/2053854359944081833</id><content type="html" xml:base="https://solmaz.io/x/2053854359944081833/"><![CDATA[I don&#39;t have a 128gb macbook to run ds4 out of, but I resonate with all the points on Armin&#39;s post

He was telling me, @mervenoyann and @cristinaponcela that local models need more polish 1 month ago in London. Today, I am happy to be given a chance and a shot at the problem!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Excited to work with @steipete, @vincent_koc, @LysandreJik, @ben_burtenshaw, @evalstate...</title><link href="https://solmaz.io/x/2053813628588212542/" rel="alternate" type="text/html" title="Excited to work with @steipete, @vincent_koc, @LysandreJik, @ben_burtenshaw, @evalstate..." /><published>2026-05-11T12:25:01+00:00</published><updated>2026-05-11T12:25:01+00:00</updated><id>https://solmaz.io/x/2053813628588212542</id><content type="html" xml:base="https://solmaz.io/x/2053813628588212542/"><![CDATA[Excited to work with @steipete, @vincent_koc, @LysandreJik, @ben_burtenshaw, @evalstate, @mervenoyann, @NielsRogge and many others!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have a new job!</title><link href="https://solmaz.io/x/2053812410730037256/" rel="alternate" type="text/html" title="I have a new job!" /><published>2026-05-11T12:20:11+00:00</published><updated>2026-05-11T12:20:11+00:00</updated><id>https://solmaz.io/x/2053812410730037256</id><content type="html" xml:base="https://solmaz.io/x/2053812410730037256/"><![CDATA[I have a new job!

Excited to announce that I will be working with Hugging Face to make local models work great in OpenClaw and other open agent harnesses!

I will be building in public and documenting everything along the way, stay tuned!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I undersign this. The fact that you generate slop doesn’t mean that you don’t know the...</title><link href="https://solmaz.io/x/2052569202016587809/" rel="alternate" type="text/html" title="I undersign this. The fact that you generate slop doesn’t mean that you don’t know the..." /><published>2026-05-08T02:00:07+00:00</published><updated>2026-05-08T02:00:07+00:00</updated><id>https://solmaz.io/x/2052569202016587809</id><content type="html" xml:base="https://solmaz.io/x/2052569202016587809/"><![CDATA[I undersign this. The fact that you generate slop doesn’t mean that you don’t know the difference between good and bad code

In non-mission critical applications, slop let’s you go from 0 to 1 very quickly

Let the code grow without too much attention first. If it proves itself, tear it down and write it anew, this time properly. This is the way]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is the idea behind acpx as well</title><link href="https://solmaz.io/x/2051299462144770116/" rel="alternate" type="text/html" title="This is the idea behind acpx as well" /><published>2026-05-04T13:54:37+00:00</published><updated>2026-05-04T13:54:37+00:00</updated><id>https://solmaz.io/x/2051299462144770116</id><content type="html" xml:base="https://solmaz.io/x/2051299462144770116/"><![CDATA[This is the idea behind acpx as well

acpx is a meta-harness. it’s main idea is to delegate harness development to others, because it is hard to match the full might of OpenAI or Anthropic when it comes to building a harness

so it takes it at face value the functionality other harnesses provide, and let’s you program them from the outside

flue came out the other day which is similar, it would be cool if flue could let me program over codex as well. it looks very interesting!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have a lot of ideas for acpx I want to implement, but I could not work on them because life...</title><link href="https://solmaz.io/x/2050706324296499329/" rel="alternate" type="text/html" title="I have a lot of ideas for acpx I want to implement, but I could not work on them because life..." /><published>2026-05-02T22:37:42+00:00</published><updated>2026-05-02T22:37:42+00:00</updated><id>https://solmaz.io/x/2050706324296499329</id><content type="html" xml:base="https://solmaz.io/x/2050706324296499329/"><![CDATA[I have a lot of ideas for acpx I want to implement, but I could not work on them because life intervened in bad ways

stay tuned in 1-2 weeks]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Got targeted in a phishing attack for stealing my X account today. @ domboyce @domboyce asked...</title><link href="https://solmaz.io/x/2048680149944586593/" rel="alternate" type="text/html" title="Got targeted in a phishing attack for stealing my X account today. @ domboyce @domboyce asked..." /><published>2026-04-27T08:26:25+00:00</published><updated>2026-04-27T08:26:25+00:00</updated><id>https://solmaz.io/x/2048680149944586593</id><content type="html" xml:base="https://solmaz.io/x/2048680149944586593/"><![CDATA[Got targeted in a phishing attack for stealing my X account today. @ domboyce @domboyce asked to book a meeting, which redirected to a URL that seemed like a bloomberg domain. It redirected to a fake calendly meeting page, which showed an X login button

this redirected to an X app asking for broad privileges meant to take over my account

stay safe everyone

cc @nikitabier please shut off this attack vector, it is too easy]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">VibeOps? importance of DevOps and good security practices have just increased massively. we are...</title><link href="https://solmaz.io/x/2048631758036353367/" rel="alternate" type="text/html" title="VibeOps? importance of DevOps and good security practices have just increased massively. we are..." /><published>2026-04-27T05:14:07+00:00</published><updated>2026-04-27T05:14:07+00:00</updated><id>https://solmaz.io/x/2048631758036353367</id><content type="html" xml:base="https://solmaz.io/x/2048631758036353367/"><![CDATA[VibeOps? importance of DevOps and good security practices have just increased massively. we are slowly approaching the challenger-level disaster @simonw is warning about. imagine bezos deleting us-east-1 while vibecoding a new company

it’s clear there needs to be clear boundaries and friction at deployment level. it’s fine when you are starting a new project, but once it starts making real money, you should slowly take away production write access

you can’t give infra write access to LLMs indefinitely]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">is GPT-5.5 a smaller model? I swear it is very stupid sometimes (high thinking)</title><link href="https://solmaz.io/x/2048105405574762638/" rel="alternate" type="text/html" title="is GPT-5.5 a smaller model? I swear it is very stupid sometimes (high thinking)" /><published>2026-04-25T18:22:35+00:00</published><updated>2026-04-25T18:22:35+00:00</updated><id>https://solmaz.io/x/2048105405574762638</id><content type="html" xml:base="https://solmaz.io/x/2048105405574762638/"><![CDATA[is GPT-5.5 a smaller model? I swear it is very stupid sometimes (high thinking)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex Computer Use = Mind blown</title><link href="https://solmaz.io/x/2047973593489760661/" rel="alternate" type="text/html" title="Codex Computer Use = Mind blown" /><published>2026-04-25T09:38:48+00:00</published><updated>2026-04-25T09:38:48+00:00</updated><id>https://solmaz.io/x/2047973593489760661</id><content type="html" xml:base="https://solmaz.io/x/2047973593489760661/"><![CDATA[Codex Computer Use = Mind blown

What the hell are you cooking @thsottiaux and team 🤯

I regret to inform that I will be switching away from agent-browser, @vercel]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This should be the pelican test for CUA</title><link href="https://solmaz.io/x/2047919527967572416/" rel="alternate" type="text/html" title="This should be the pelican test for CUA" /><published>2026-04-25T06:03:58+00:00</published><updated>2026-04-25T06:03:58+00:00</updated><id>https://solmaz.io/x/2047919527967572416</id><content type="html" xml:base="https://solmaz.io/x/2047919527967572416/"><![CDATA[This should be the pelican test for CUA]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;m using gpt 5.5 xhigh while designing a table schema and it feels dumber than 5.4. but not a...</title><link href="https://solmaz.io/x/2047782589528719423/" rel="alternate" type="text/html" title="I&#39;m using gpt 5.5 xhigh while designing a table schema and it feels dumber than 5.4. but not a..." /><published>2026-04-24T20:59:49+00:00</published><updated>2026-04-24T20:59:49+00:00</updated><id>https://solmaz.io/x/2047782589528719423</id><content type="html" xml:base="https://solmaz.io/x/2047782589528719423/"><![CDATA[I&#39;m using gpt 5.5 xhigh while designing a table schema and it feels dumber than 5.4. but not a strong feeling

it&#39;s formatting of output is different as well. It&#39;s not wrong, but different, feels more terse than 5.4

maybe it&#39;s better at instruction following, and it&#39;s picking up on my previous use of plain-language skill]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">What are the implications? Windows and other criticial proprietary software will have to open...</title><link href="https://solmaz.io/x/2047546781576147164/" rel="alternate" type="text/html" title="What are the implications? Windows and other criticial proprietary software will have to open..." /><published>2026-04-24T05:22:48+00:00</published><updated>2026-04-24T05:22:48+00:00</updated><id>https://solmaz.io/x/2047546781576147164</id><content type="html" xml:base="https://solmaz.io/x/2047546781576147164/"><![CDATA[What are the implications? Windows and other criticial proprietary software will have to open source in order to survive?

It feels like there is an economic equilibrium where the tokens spent to attack an open source software will always surpass those spent on a closed one?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I thought I&#39;d try @unclebobmartin&#39;s advice and force my agents to maximize unit test coverage...</title><link href="https://solmaz.io/x/2047414897584054612/" rel="alternate" type="text/html" title="I thought I&#39;d try @unclebobmartin&#39;s advice and force my agents to maximize unit test coverage..." /><published>2026-04-23T20:38:45+00:00</published><updated>2026-04-23T20:38:45+00:00</updated><id>https://solmaz.io/x/2047414897584054612</id><content type="html" xml:base="https://solmaz.io/x/2047414897584054612/"><![CDATA[I thought I&#39;d try @unclebobmartin&#39;s advice and force my agents to maximize unit test coverage, reduce cyclomatic complexity, etc. in a new go backend I am working on, as an experiment

will report how it goes]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is sad an ironic, because one of my most frequent prompts to my agent is &quot;How would Google...</title><link href="https://solmaz.io/x/2046515242750890020/" rel="alternate" type="text/html" title="This is sad an ironic, because one of my most frequent prompts to my agent is &quot;How would Google..." /><published>2026-04-21T09:03:50+00:00</published><updated>2026-04-21T09:03:50+00:00</updated><id>https://solmaz.io/x/2046515242750890020</id><content type="html" xml:base="https://solmaz.io/x/2046515242750890020/"><![CDATA[This is sad an ironic, because one of my most frequent prompts to my agent is &quot;How would Google have done it?&quot;

Go, developed at Google, is my go-to backend language these days

Not even mentioning that the transformer was invented there

Google has a great legacy, it must not go down]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I realized I didn’t hit send on this in time. I hope GitHub team sees this, at least ability to...</title><link href="https://solmaz.io/x/2046135285021352431/" rel="alternate" type="text/html" title="I realized I didn’t hit send on this in time. I hope GitHub team sees this, at least ability to..." /><published>2026-04-20T07:54:01+00:00</published><updated>2026-04-20T07:54:01+00:00</updated><id>https://solmaz.io/x/2046135285021352431</id><content type="html" xml:base="https://solmaz.io/x/2046135285021352431/"><![CDATA[I realized I didn’t hit send on this in time. I hope GitHub team sees this, at least ability to add agent accounts as a lower privilege citizen on the platform]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Here is the plain-language skill I use very often with codex, like every 1 out of 5 prompt</title><link href="https://solmaz.io/x/2045975780815978753/" rel="alternate" type="text/html" title="Here is the plain-language skill I use very often with codex, like every 1 out of 5 prompt" /><published>2026-04-19T21:20:13+00:00</published><updated>2026-04-19T21:20:13+00:00</updated><id>https://solmaz.io/x/2045975780815978753</id><content type="html" xml:base="https://solmaz.io/x/2045975780815978753/"><![CDATA[Here is the plain-language skill I use very often with codex, like every 1 out of 5 prompt

I just type $ pla then tab
https://t.co/aRWerg3fh7]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Who is running local models on GPUs on OpenClaw?</title><link href="https://solmaz.io/x/2045872636585029943/" rel="alternate" type="text/html" title="Who is running local models on GPUs on OpenClaw?" /><published>2026-04-19T14:30:21+00:00</published><updated>2026-04-19T14:30:21+00:00</updated><id>https://solmaz.io/x/2045872636585029943</id><content type="html" xml:base="https://solmaz.io/x/2045872636585029943/"><![CDATA[Who is running local models on GPUs on OpenClaw?

I have started benchmarking different models this week. I am working on improving model selection and switching UX on OpenClaw, i.e. I run

/model vllm/gemma-e4b

to switch the model in a channel, and then a model controller automatically loads that into memory, gets it ready, or gives an insufficient memory error, if capacity is not enough for that. Like when you are using multiple models in parallel

I am going to try llama-swap, LM Studio and Ollama for this next and compare them. There are a ton of variants of models, weight formats and quantizations, which need benchmarking

I have been using unquantized original safetensors until now, which already gave me the ability to run ~5 parallel generations in my hardware

So if I am going to try LM Studio, I would rather use the bf16 ggml-org/gemma-4-E4B-it-GGUF instead of anything smaller --- because there is no point in nerfing an already smol model if your hardware can run 5 parallel sessions on the unquantized version

Will also release vibe reports and benchmarks on all this with @mervenoyann later this week

I would like to hear your thoughts if you have already tried these models on OpenClaw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agent blurbs for letting your agent install stuff</title><link href="https://solmaz.io/x/2045642958133686551/" rel="alternate" type="text/html" title="Agent blurbs for letting your agent install stuff" /><published>2026-04-18T23:17:42+00:00</published><updated>2026-04-18T23:17:42+00:00</updated><id>https://solmaz.io/x/2045642958133686551</id><content type="html" xml:base="https://solmaz.io/x/2045642958133686551/"><![CDATA[This project by @davidguttman is interesting also the way he set up install

You click &quot;Copy install instructions for my agent&quot;

It copies: Install LobsterLink by following the instructions at &lt;link to AGENT-INSTALL. md at repo root&gt;

I had done a similar thing in acpx README as well

It&#39;s a small thing, but it reduces friction with installation so much

It should be more commonplace to install software using agent blurbs

I&#39;m wondering if there could be a way to standardize this beyond plaintext, like after pasting your agent, it consumes a standardized manifest format and asks for approval while displaying all the commands that will be run]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Mobile SSH enjoyers. Try out getmoshi.app by @odd_joel if you haven’t, good alternative to...</title><link href="https://solmaz.io/x/2045423813395939727/" rel="alternate" type="text/html" title="Mobile SSH enjoyers. Try out getmoshi.app by @odd_joel if you haven’t, good alternative to..." /><published>2026-04-18T08:46:53+00:00</published><updated>2026-04-18T08:46:53+00:00</updated><id>https://solmaz.io/x/2045423813395939727</id><content type="html" xml:base="https://solmaz.io/x/2045423813395939727/"><![CDATA[Mobile SSH enjoyers. Try out https://t.co/7CCUwVD3HI by @odd_joel if you haven’t, good alternative to @TermiusHQ 
Thanks @Andori3042 for introducing me]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Europe must have some good inference providers for open weight models that we are missing</title><link href="https://solmaz.io/x/2045171472499458366/" rel="alternate" type="text/html" title="Europe must have some good inference providers for open weight models that we are missing" /><published>2026-04-17T16:04:11+00:00</published><updated>2026-04-17T16:04:11+00:00</updated><id>https://solmaz.io/x/2045171472499458366</id><content type="html" xml:base="https://solmaz.io/x/2045171472499458366/"><![CDATA[Europe must have some good inference providers for open weight models that we are missing

Does anybody know a good EU provider for Kimi, GLM, Minimax, Qwen, Gemma, DeepSeek, Muse Spark etc.?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I’ve heard about “Codex” for the first time 5 years ago. Back then it was called code-davinci</title><link href="https://solmaz.io/x/2045113126438281489/" rel="alternate" type="text/html" title="I’ve heard about “Codex” for the first time 5 years ago. Back then it was called code-davinci" /><published>2026-04-17T12:12:20+00:00</published><updated>2026-04-17T12:12:20+00:00</updated><id>https://solmaz.io/x/2045113126438281489</id><content type="html" xml:base="https://solmaz.io/x/2045113126438281489/"><![CDATA[I’ve heard about “Codex” for the first time 5 years ago. Back then it was called code-davinci

So happy to see this idea flourish, since those first days with @woj_zaremba, @gdb, @ilyasut 

https://t.co/C9wPTdQmce]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Yes, this is the primary skill in being a software engineer now</title><link href="https://solmaz.io/x/2044816818234126748/" rel="alternate" type="text/html" title="Yes, this is the primary skill in being a software engineer now" /><published>2026-04-16T16:34:54+00:00</published><updated>2026-04-16T16:34:54+00:00</updated><id>https://solmaz.io/x/2044816818234126748</id><content type="html" xml:base="https://solmaz.io/x/2044816818234126748/"><![CDATA[Yes, this is the primary skill in being a software engineer now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Own your AI infrastructure</title><link href="https://solmaz.io/x/2044457293069062530/" rel="alternate" type="text/html" title="Own your AI infrastructure" /><published>2026-04-15T16:46:17+00:00</published><updated>2026-04-15T16:46:17+00:00</updated><id>https://solmaz.io/x/2044457293069062530</id><content type="html" xml:base="https://solmaz.io/x/2044457293069062530/"><![CDATA[When you build your company&#39;s workflows around Claude Cowork, you are betting against local models and owning your infra, and inviting your company to long-term exploitation

If I were Anthropic or OpenAI, I would be the most scared of local AI proliferating

Let&#39;s do the math. A single large big lab subscription costs 200x12=$2400 per year

If you want to have both OpenAI and Anthropic, that could cost $2400, $3600, or $4800, based on which combinations of Pro, Max plans you choose

An ASUS Ascent GX10 costs $3000, and you can use that for many years. You don&#39;t get the same level of coding quality with open models yet, but maybe you want to do something simpler than coding today... There are already many people who started buying GPUs for this reason

Now we know big labs are selling some of these plans at a loss. So they will likely get more expensive

When you use Claude Cowork or similar, you are locking yourself into being a RENTER. Because once you set up workflows for a company, it takes time to migrate away to something else, even though we have AI to help

Infra is sticky, it&#39;s how hyperscalers make profit. Think about the difference in amount you pay AWS vs Hetzner. This is B2B SaaS 101. Once you sell to a company, you are in for a long time, especially in Europe

So if you build your company&#39;s AI workflows around a proprietary product by another company, then you are basically saying &quot;Come exploit me as tolerably as you can in the next 10 years, because it will be too painful for me to switch&quot;

It&#39;s a great business for Anthropic. And Claude is awesome too! The feedback from friends who use it has been great, it made their lives a lot easier

But when you build your company over proprietary AI infra, then you are making sure you will not be an OWNER, and partake in the usual sorrows of being a RENTER from a monopolist, which is exploitation

This is not the case when you use open source agent infra. Whereas Anthropic is unlikely to let you use future open models in their future iteration of Claude Cowork, using free and open source frameworks like OpenClaw, Open Agents, etc. lets you drop in replace providers or local hardware if they start to upcharge you

Keep this in mind, if you have a business]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is our moat 🤣</title><link href="https://solmaz.io/x/2044381956591280140/" rel="alternate" type="text/html" title="This is our moat 🤣" /><published>2026-04-15T11:46:55+00:00</published><updated>2026-04-15T11:46:55+00:00</updated><id>https://solmaz.io/x/2044381956591280140</id><content type="html" xml:base="https://solmaz.io/x/2044381956591280140/"><![CDATA[This is our moat 🤣]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">People&#39;s AI</title><link href="https://solmaz.io/x/2044337339518828727/" rel="alternate" type="text/html" title="People&#39;s AI" /><published>2026-04-15T08:49:38+00:00</published><updated>2026-04-15T08:49:38+00:00</updated><id>https://solmaz.io/x/2044337339518828727</id><content type="html" xml:base="https://solmaz.io/x/2044337339518828727/"><![CDATA[You need to understand one fact about OpenClaw

People are biased and incentivized to spread disinformation about OpenClaw. That is because OpenClaw IS NOT PUMPING ANYONE’S BAGS, unlike most other projects

Literally every other for-profit agent product is incentivized to trash OpenClaw, BECAUSE OpenClaw is a neutral third party across the industry and geopolitical scene. They MAKE MONEY when OpenClaw loses

OpenClaw does not worry about making money for some investors. Its founder @steipete is a successful exited founder. He is motivated by having fun and democratizing AI, literally. That is why he is suddenly so loved by everyone. He cares about PEOPLE, not MONEY

“OpenClaw is bloated”
-&gt; Since beginning of March, OpenClaw is thinning its core and putting functionality in plugins behind a plugin SDK. Having numerous plugins to choose from does not mean bloat. This was already copied by others and is still a work in progress

“OpenClaw is not secure”
-&gt; OpenClaw has the most eyeballs and immediately addresses any security advisories as soon as they come. It is the most secure agent, by sheer pressure

“OpenClaw is bought by OpenAI”
-&gt; Then why is my bank account so empty bro??? All maintainers are literally unpaid and working DOUBLE beside their dayjobs to ship features to you. Do you think VC money can buy that kind of commitment?

Once you understand these facts, you’ll like OpenClaw even more. Because OpenClaw is your AI, People’s AI

And you can join us too. OpenClaw is the easiest-to-join project in AI right now. You just need to start using it, and start making good contributions. If you are competent, you can become a maintainer, and join the rest of the team making history!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&amp;gt; Worked for 10 hours</title><link href="https://solmaz.io/x/2044002894819541192/" rel="alternate" type="text/html" title="&amp;gt; Worked for 10 hours" /><published>2026-04-14T10:40:40+00:00</published><updated>2026-04-14T10:40:40+00:00</updated><id>https://solmaz.io/x/2044002894819541192</id><content type="html" xml:base="https://solmaz.io/x/2044002894819541192/"><![CDATA[&amp;gt; Worked for 10 hours
&amp;gt; Selected model is at capacity

Model is gpt 5.4 high]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Gemma 4 first impressions on OpenClaw</title><link href="https://solmaz.io/x/2043971427234140207/" rel="alternate" type="text/html" title="Gemma 4 first impressions on OpenClaw" /><published>2026-04-14T08:35:38+00:00</published><updated>2026-04-14T08:35:38+00:00</updated><id>https://solmaz.io/x/2043971427234140207</id><content type="html" xml:base="https://solmaz.io/x/2043971427234140207/"><![CDATA[This is pretty much the arc I have been going on in the 2 months since I bought my ASUS GX10 for 3k EUR

Use whisper on the API -&gt; realize it charged me $$$ for just a few calls -&gt; migrate openclaw to use local whisper

Need to deduplicate news articles for my news engine -&gt; download qwen embedding 8b

And now, gemma4-e4b  finally seems like a viable alternative for a local model that runs around 20 tok/s

So I will install a matrix client to use through tailscale, and can finally build the social life CRM I dreamed of since years.

100% private, zero data going out. I had a bias of not giving any personal data to AI since ChatGPT came out. But I can finally give more personal data to my AI agent

And I will make sure @openclaw supports all this in an easy way, make it dead easy

Fully self-owned AI begins now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">gemma 4 is actually pretty decent and runs on my asus gx10 (128 gb vram)</title><link href="https://solmaz.io/x/2043817471673586010/" rel="alternate" type="text/html" title="gemma 4 is actually pretty decent and runs on my asus gx10 (128 gb vram)" /><published>2026-04-13T22:23:52+00:00</published><updated>2026-04-13T22:23:52+00:00</updated><id>https://solmaz.io/x/2043817471673586010</id><content type="html" xml:base="https://solmaz.io/x/2043817471673586010/"><![CDATA[gemma 4 is actually pretty decent and runs on my asus gx10 (128 gb vram) 

the original dense 31b runs slow, averaging around 3~4 tok/s. it&#39;s also using 80% of gpu memory

my previous experience with gemini 3 pro back in november was that it was too trigger happy. but this is one-shotting simple tasks I&#39;m giving it in openclaw harness, and it&#39;s hard to tell it apart from gpt 5.4 for my use cases so far

now off to try out smaller models, because 3 tok/s is too slow]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@lucasmeijer @lucasmeijer one could actually periodically trigger an agent to propose...</title><link href="https://solmaz.io/x/2043720098507018621/" rel="alternate" type="text/html" title="@lucasmeijer @lucasmeijer one could actually periodically trigger an agent to propose..." /><published>2026-04-13T15:56:56+00:00</published><updated>2026-04-13T15:56:56+00:00</updated><id>https://solmaz.io/x/2043720098507018621</id><content type="html" xml:base="https://solmaz.io/x/2043720098507018621/"><![CDATA[@lucasmeijer @lucasmeijer one could actually periodically trigger an agent to propose simplifications or new abstractions in a codebase, and I believe it would already work pretty well with the current models]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Got tool calls to work, context size 65k tokens</title><link href="https://solmaz.io/x/2043655100686573715/" rel="alternate" type="text/html" title="Got tool calls to work, context size 65k tokens" /><published>2026-04-13T11:38:39+00:00</published><updated>2026-04-13T11:38:39+00:00</updated><id>https://solmaz.io/x/2043655100686573715</id><content type="html" xml:base="https://solmaz.io/x/2043655100686573715/"><![CDATA[Got tool calls to work, context size 65k tokens]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@grok does this exist</title><link href="https://solmaz.io/x/2043428626209604023/" rel="alternate" type="text/html" title="@grok does this exist" /><published>2026-04-12T20:38:44+00:00</published><updated>2026-04-12T20:38:44+00:00</updated><id>https://solmaz.io/x/2043428626209604023</id><content type="html" xml:base="https://solmaz.io/x/2043428626209604023/"><![CDATA[@grok does this exist]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Question for the community:</title><link href="https://solmaz.io/x/2043427225824055675/" rel="alternate" type="text/html" title="Question for the community:" /><published>2026-04-12T20:33:10+00:00</published><updated>2026-04-12T20:33:10+00:00</updated><id>https://solmaz.io/x/2043427225824055675</id><content type="html" xml:base="https://solmaz.io/x/2043427225824055675/"><![CDATA[Question for the community:
What is the best testing observability and control tool you have used until now?

- Could be SaaS, could be open source
- To be used in @openclaw repo
- Should be compatible with vitest
- Ideally language agnostic

I need something that lets me run a very long running test group multiple times on a specific commit or tag, without repeating the tests that have already finished

This is a need because the 1hr long process might get interrupted due to flakiness. So I need to persists the progress of a run, and then not repeat them

I have seen some paid SaaS for this, but none that really give me what I want

This is going to be important especially while working with agents, because when you are committing 100x faster, you don&#39;t want to waste time and compute running the same things

I started building this already as an exercise. If this exists already in a satisfactory way, I will stop. Otherwise, I&#39;ll keep building]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My clanker @dutifulbob has an identity update. I was getting sick of the despicable me theme...</title><link href="https://solmaz.io/x/2043398933607547019/" rel="alternate" type="text/html" title="My clanker @dutifulbob has an identity update. I was getting sick of the despicable me theme..." /><published>2026-04-12T18:40:44+00:00</published><updated>2026-04-12T18:40:44+00:00</updated><id>https://solmaz.io/x/2043398933607547019</id><content type="html" xml:base="https://solmaz.io/x/2043398933607547019/"><![CDATA[My clanker @dutifulbob has an identity update. I was getting sick of the despicable me theme and bananas]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Some photos from my @aiDotEngineer Europe talk, credits to @sergiopesch</title><link href="https://solmaz.io/x/2043038348592165195/" rel="alternate" type="text/html" title="Some photos from my @aiDotEngineer Europe talk, credits to @sergiopesch" /><published>2026-04-11T18:47:54+00:00</published><updated>2026-04-11T18:47:54+00:00</updated><id>https://solmaz.io/x/2043038348592165195</id><content type="html" xml:base="https://solmaz.io/x/2043038348592165195/"><![CDATA[Some photos from my @aiDotEngineer Europe talk, credits to @sergiopesch

He managed to capture the “Agentol, Apply Generously” slide, lol]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@TermiusHQ logo looks a bit familiar 🥲</title><link href="https://solmaz.io/x/2043036645671223613/" rel="alternate" type="text/html" title="@TermiusHQ logo looks a bit familiar 🥲" /><published>2026-04-11T18:41:08+00:00</published><updated>2026-04-11T18:41:08+00:00</updated><id>https://solmaz.io/x/2043036645671223613</id><content type="html" xml:base="https://solmaz.io/x/2043036645671223613/"><![CDATA[@TermiusHQ logo looks a bit familiar 🥲]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">local gemma 4 first impressions on openclaw, using the dense model, 26b model with 49gb weights...</title><link href="https://solmaz.io/x/2043035989833126167/" rel="alternate" type="text/html" title="local gemma 4 first impressions on openclaw, using the dense model, 26b model with 49gb weights..." /><published>2026-04-11T18:38:32+00:00</published><updated>2026-04-11T18:38:32+00:00</updated><id>https://solmaz.io/x/2043035989833126167</id><content type="html" xml:base="https://solmaz.io/x/2043035989833126167/"><![CDATA[local gemma 4 first impressions on openclaw, using the dense model, 26b model with 49gb weights on my asus gx10

took some time to set up, but it succeded in getting a response in 1-2 hours with vllm docs

I asked it to demonstrate some tool calls. it tried to call the nonexistent weather tool 2300 times 🙄

it seems to have a tendency to get stuck in loops in openclaw harness. enabling loop detection just now did not help

I’m debugging this on my phone lol. I’ll be sharing my progress with gemma4 under this thread]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">A rescue agent guaranteed not to break solves this</title><link href="https://solmaz.io/x/2042862980572807442/" rel="alternate" type="text/html" title="A rescue agent guaranteed not to break solves this" /><published>2026-04-11T07:11:03+00:00</published><updated>2026-04-11T07:11:03+00:00</updated><id>https://solmaz.io/x/2042862980572807442</id><content type="html" xml:base="https://solmaz.io/x/2042862980572807442/"><![CDATA[A rescue agent guaranteed not to break solves this]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Finally!</title><link href="https://solmaz.io/x/2042704553808744475/" rel="alternate" type="text/html" title="Finally!" /><published>2026-04-10T20:41:31+00:00</published><updated>2026-04-10T20:41:31+00:00</updated><id>https://solmaz.io/x/2042704553808744475</id><content type="html" xml:base="https://solmaz.io/x/2042704553808744475/"><![CDATA[Finally!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenClaw type UK Drill Jazz Beat by @thorwebdev @GeminiApp @GoogleDeepMind</title><link href="https://solmaz.io/x/2042617748308513239/" rel="alternate" type="text/html" title="OpenClaw type UK Drill Jazz Beat by @thorwebdev @GeminiApp @GoogleDeepMind" /><published>2026-04-10T14:56:35+00:00</published><updated>2026-04-10T14:56:35+00:00</updated><id>https://solmaz.io/x/2042617748308513239</id><content type="html" xml:base="https://solmaz.io/x/2042617748308513239/"><![CDATA[OpenClaw type UK Drill Jazz Beat by @thorwebdev @GeminiApp @GoogleDeepMind]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">For those who want to view, my talk Building on ACP at OpenClaw at @aiDotEngineer Europe...</title><link href="https://solmaz.io/x/2042550983465611542/" rel="alternate" type="text/html" title="For those who want to view, my talk Building on ACP at OpenClaw at @aiDotEngineer Europe..." /><published>2026-04-10T10:31:17+00:00</published><updated>2026-04-10T10:31:17+00:00</updated><id>https://solmaz.io/x/2042550983465611542</id><content type="html" xml:base="https://solmaz.io/x/2042550983465611542/"><![CDATA[For those who want to view, my talk Building on ACP at OpenClaw at @aiDotEngineer Europe, 5:41hr mark

About ACP, acpx and running agents on kubernetes with open source orchestrators

https://t.co/4qRFVOZFtu]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Anthropic is optimizing for general knowledge work</title><link href="https://solmaz.io/x/2042494957286560225/" rel="alternate" type="text/html" title="Anthropic is optimizing for general knowledge work" /><published>2026-04-10T06:48:40+00:00</published><updated>2026-04-10T06:48:40+00:00</updated><id>https://solmaz.io/x/2042494957286560225</id><content type="html" xml:base="https://solmaz.io/x/2042494957286560225/"><![CDATA[PSA for developers

Do NOT torture yourself with Opus*. Anthropic’s current growth is due to people using Claude for general knowledge work

They are not directly incentivized as an org anymore to improve the model for coding, in an economic sense. They are already printing cash from non-developers

(this statement ignores the fact that improving its coding abilities would help with general reasoning/knowledge work)

Developers are a very small subset of all knowledge workers. So from this point on, they would rather divert their resources to develop a system that works 90% good for ALL knowledge work, rather than making it 100% for coding

Because Anthropic has a clear enterprise strategy since years already. Anthropic is the new Microsoft. Do not think that “Anthropic is Apple” or “Claude is Mac for xyz”

Looking at Claude’s at whim quantization and Claude Code’s quality over time, Claude for me is Windows, not Mac

But they are winning big enterprise bucks, so good for them!

*(I tortured myself with Sonnet 4 and Opus the entire summer of 2025, and no developer should ever have to go through that. I switched to something better as soon as it came out, Codex. If something even better comes out, I will switch again)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Clawfather @steipete at stage @aiDotEngineer europe</title><link href="https://solmaz.io/x/2042171612917567503/" rel="alternate" type="text/html" title="Clawfather @steipete at stage @aiDotEngineer europe" /><published>2026-04-09T09:23:48+00:00</published><updated>2026-04-09T09:23:48+00:00</updated><id>https://solmaz.io/x/2042171612917567503</id><content type="html" xml:base="https://solmaz.io/x/2042171612917567503/"><![CDATA[Clawfather @steipete at stage @aiDotEngineer europe]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Big lab marketing teams like to shroud model releases in mystery and vagueposting</title><link href="https://solmaz.io/x/2041767427420156248/" rel="alternate" type="text/html" title="Big lab marketing teams like to shroud model releases in mystery and vagueposting" /><published>2026-04-08T06:37:43+00:00</published><updated>2026-04-08T06:37:43+00:00</updated><id>https://solmaz.io/x/2041767427420156248</id><content type="html" xml:base="https://solmaz.io/x/2041767427420156248/"><![CDATA[Big lab marketing teams like to shroud model releases in mystery and vagueposting

If you are curious about the black hat capability of LLMs, watch this pres by Nicholas Carlini from a few days back

https://t.co/NThIxcVwV4]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I will be there as well, speaking about ACP, acpx and agent orchestration 🙌</title><link href="https://solmaz.io/x/2041529072111477222/" rel="alternate" type="text/html" title="I will be there as well, speaking about ACP, acpx and agent orchestration 🙌" /><published>2026-04-07T14:50:35+00:00</published><updated>2026-04-07T14:50:35+00:00</updated><id>https://solmaz.io/x/2041529072111477222</id><content type="html" xml:base="https://solmaz.io/x/2041529072111477222/"><![CDATA[I will be there as well, speaking about ACP, acpx and agent orchestration 🙌]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Full transcripts are available, including all the back and forth between codex, claude and judge</title><link href="https://solmaz.io/x/2041202936391340244/" rel="alternate" type="text/html" title="Full transcripts are available, including all the back and forth between codex, claude and judge" /><published>2026-04-06T17:14:38+00:00</published><updated>2026-04-06T17:14:38+00:00</updated><id>https://solmaz.io/x/2041202936391340244</id><content type="html" xml:base="https://solmaz.io/x/2041202936391340244/"><![CDATA[Full transcripts are available, including all the back and forth between codex, claude and judge]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Repo link (feel free to send PRs):</title><link href="https://solmaz.io/x/2041187554905530696/" rel="alternate" type="text/html" title="Repo link (feel free to send PRs):" /><published>2026-04-06T16:13:31+00:00</published><updated>2026-04-06T16:13:31+00:00</updated><id>https://solmaz.io/x/2041187554905530696</id><content type="html" xml:base="https://solmaz.io/x/2041187554905530696/"><![CDATA[Repo link (feel free to send PRs):
https://t.co/vyOlQ6lPC1]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Introducing AI Battle</title><link href="https://solmaz.io/x/2041187155620274541/" rel="alternate" type="text/html" title="Introducing AI Battle" /><published>2026-04-06T16:11:55+00:00</published><updated>2026-04-06T16:11:55+00:00</updated><id>https://solmaz.io/x/2041187155620274541</id><content type="html" xml:base="https://solmaz.io/x/2041187155620274541/"><![CDATA[Is Claude better or Codex?

There are many benchmarks to answer that. But they are BORING

I propose something more interesting: ⚔️ AI BATTLE ⚔️

A 1v1 real-time quiz format where AI agents try to pose each other problems that they think the other agent will not be able solve

Claude vs Codex
10 questions each
Codex asks first, Claude tries to answer
Then Claude asks and Codex tries to answer
Repeat
20 minutes to come up with a problem and 20 minutes to solve it
Judge (Codex) judges the validity of the questions and answers, and gives points
All automated, with acpx flow feature
Implementation and full rules all open source, on github osolmaz/ai-battle

So who won?
I ran 4 games. 
It tied in 2, and Codex won in 2 closely

An example question by Codex, which Claude could not answer:

How many 3-colorings of the edges of the complete bipartite graph K_{5,5} are there with the following two properties: (1) there is no monochromatic 4-cycle, and (2) among the 25 edges, exactly 15 are red, exactly 5 are blue, and exactly 5 are green?

Which is apparently 4029912, but Claude answered 0

In other cases, Claude asked a flawed question and failed to come up with a valid question in 20 minutes. So that&#39;s how it lost those 2 games with just 1-2 point difference

In these 4 runs, Codex answered every question by Claude correctly. But there were some runs where it couldn&#39;t, which I did not commit to the repo because the runs couldn&#39;t complete due to bugs

I did not tell them do ask math questions, but that is what they tended to do, because the answers had to be verifiable by the judge. The quiz can be done in any hard subject, physics, chemistry, computer science...

Opus 4.6 and GPT 5.4 matched very closely in terms of problem creation and solving. But I cannot tell how creative these problems were at first glance. Maybe someone with more experience can tell me, looking at the problems in the repo? I need someone to tell me how legit they are

Please take the code, modify it and run with different rules and subjects. I am curious to see the results!

You will need paid subscriptions to all the models/agents you want to test of course

I also feel that the game structure has a potential to be used in self-play. If you are an ML researcher, please look at the repo and lmk if this or a variant of it could be useful in RL!

Full transcripts of the runs, including Codex and Claude session files are committed to the repo, for those who want to do archaeology on them

Btw this idea came from the desire, &quot;how can I create a cool demo of acpx flows?&quot;

Whole game is implemented in typescript, and automatically drives Codex and Claude sessions over ACP, Agent Client Protocol

The video below is from acpx flow viewer rendering a run. You can see it loop through the same paths, first letting Codex ask, then Claude, then repeat

acpx flows use a general programmatic workflow engine where ACP is just one type of node. You should be able to use it for non-ACP workflows, but I haven&#39;t tried that yet

This implementation is separate from OpenClaw&#39;s current workflow implementations, with the intention to merge them somehow in the future

You might find bugs in my implementation. Feel free to send PRs. I wanted to do more runs but I finished my Codex plan. It would be great if this idea could evolve in a decentralized manner!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Their argument “it’S HaRd On OuR iNfRa” so goes down the drain</title><link href="https://solmaz.io/x/2041054962914869426/" rel="alternate" type="text/html" title="Their argument “it’S HaRd On OuR iNfRa” so goes down the drain" /><published>2026-04-06T07:26:38+00:00</published><updated>2026-04-06T07:26:38+00:00</updated><id>https://solmaz.io/x/2041054962914869426</id><content type="html" xml:base="https://solmaz.io/x/2041054962914869426/"><![CDATA[Their argument “it’S HaRd On OuR iNfRa” so goes down the drain

With this, they shot themselves in the foot for a future anti-competitive lawsuit, because it is undeniable evidence that they just don’t want competition

Which means they have evaluated the benefits short term, and calculated that it is higher than what they will pay in the lawsuit

I don’t see how it is good for them long term]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI replies are getting more sophisticated… or people are turning into AIs</title><link href="https://solmaz.io/x/2040872050181451938/" rel="alternate" type="text/html" title="AI replies are getting more sophisticated… or people are turning into AIs" /><published>2026-04-05T19:19:48+00:00</published><updated>2026-04-05T19:19:48+00:00</updated><id>https://solmaz.io/x/2040872050181451938</id><content type="html" xml:base="https://solmaz.io/x/2040872050181451938/"><![CDATA[AI replies are getting more sophisticated… or people are turning into AIs

If this is AI, I wonder what the instruction is. “Misunderstand the point and reply with a question while inverting the argument”?

Artificial General Ragebait]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The new github skill installed automatically by codex now causes it to prepend [codex] to each...</title><link href="https://solmaz.io/x/2040761244697641250/" rel="alternate" type="text/html" title="The new github skill installed automatically by codex now causes it to prepend [codex] to each..." /><published>2026-04-05T11:59:30+00:00</published><updated>2026-04-05T11:59:30+00:00</updated><id>https://solmaz.io/x/2040761244697641250</id><content type="html" xml:base="https://solmaz.io/x/2040761244697641250/"><![CDATA[The new github skill installed automatically by codex now causes it to prepend [codex] to each PR title

This is a guerilla marketing tactic similar to Claude adding itself as co-committer

Codex team, I know you want to boast usage but this is annoying

Moreover, &quot;open source&quot; OpenAI repos block opening of PRs by people outside of their org. So I couldn&#39;t create a PR to remove it (I don&#39;t expect them to merge it, but it would still show how many people hate it in the discussion)

Here is a prompt for your agent if you want to disable it:
---
Add or update AGENTS.md in my ~/.codex folder
Add a rule &quot;You MUST NOT insert coding agent specific branding, like [codex], in code, PRs or issues created on GitHub&quot;
---

Then restart your sessions and this should be resolved]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">A throttling protocol for model providers</title><link href="https://solmaz.io/x/2040545125680718304/" rel="alternate" type="text/html" title="A throttling protocol for model providers" /><published>2026-04-04T21:40:44+00:00</published><updated>2026-04-04T21:40:44+00:00</updated><id>https://solmaz.io/x/2040545125680718304</id><content type="html" xml:base="https://solmaz.io/x/2040545125680718304/"><![CDATA[A more reasonable long term option for Anthropic is to create a throttling protocol

A standardized harness agnostic protocol for model providers to send warnings and throttle usage in real time

Harnesses would implement the protocol. A client can be warned. If it doesn’t listen, it can be temporarily blocked from the server side, or banned permanently if it breaks the rules too many times

Needless to say, throttling could be done first on server side easily. That would actually fix the load issue for them in the short run, while not banning the user and just giving a bad delayed UX. They probably already do this to prevent abuse

The suggested protocol would then save the user from abuse related delays too, and also inform the harness developer when they do something wrong]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If your Claude subscription renewed too recently and you don&#39;t wanna waste those tokens, you...</title><link href="https://solmaz.io/x/2040449587312222416/" rel="alternate" type="text/html" title="If your Claude subscription renewed too recently and you don&#39;t wanna waste those tokens, you..." /><published>2026-04-04T15:21:05+00:00</published><updated>2026-04-04T15:21:05+00:00</updated><id>https://solmaz.io/x/2040449587312222416</id><content type="html" xml:base="https://solmaz.io/x/2040449587312222416/"><![CDATA[If your Claude subscription renewed too recently and you don&#39;t wanna waste those tokens, you can still use your Claude sub in your OpenClaw account through ACP (which uses Claude Agents SDK, which poses no risk)

Steps:
- Open Claude Code (not OpenClaw)
- Tell it to set default model to something other than Claude (e.g. openai-codex/gpt-5.4) and tell it to delete the saved Anthropic credentials in OpenClaw config
- Create a topic in telegram or channel in discord called claude. Copy the id of that channel
- Give the link below together with the channel/topic id, and tell it to bind that channel to claude using ACP channel binding
- Restart

You should now be able to talk to Claude through Claude Agents SDK in that channel. You might need to iterate a couple times until Claude gets the config right

It will be very bare functionality, and it will not have the features and tools that your main OpenClaw harness has. It will be shitty. But you can still use telegram/discord with your subscription in the rest of the month, if you are used to the setup

https://t.co/Z0RiJbke5V]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">ACP sessions as JSONL</title><link href="https://solmaz.io/x/2040398645766344713/" rel="alternate" type="text/html" title="ACP sessions as JSONL" /><published>2026-04-04T11:58:40+00:00</published><updated>2026-04-04T11:58:40+00:00</updated><id>https://solmaz.io/x/2040398645766344713</id><content type="html" xml:base="https://solmaz.io/x/2040398645766344713/"><![CDATA[A little insight that might save you a lot of future headache if your work involves storing agent sessions and you want to be interoperable/drop-in replace alternative harnesses

@zeddotdev already did the hard work of creating an interoperable standard, ACP: Agent Client Protocol

You can represent an agent session as JSON lines of the ACP message stream

You can construct the current state of the harness from this stream. This is already how Zed loads a session I believe

If you are building an AI product, and you don&#39;t want to be locked into a single company or harness, building with ACP in mind would be a smart thing to do

Here is how acpx stores ACP sessions in ~/.acpx folder, it does exactly that:
https://t.co/SVhwXWbrBY

But don&#39;t build anything on the acpx schema for now, because I might change it in the future

Just know that JSONL of ACP messages is a good candidate for a somewhat-lossy single source of truth for agent sessions

Lossy because ACP adapters for harnesses might not transfer all the thinking and tools done by the model

So continuing or restoring a session with full fidelity is still not possible if you only save the ACP session. You still need to store original harness session files as well

But for rendering a past session for viewing or reconstructing a lossy version of it, it should be more than enough

Consider ACP if the benefits of not locking yourself in to a specific ecosystem outweighs these minor issues]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If you dislike rotating ack emojis on your messages in openclaw, this is how to make sure it...</title><link href="https://solmaz.io/x/2040375376715632923/" rel="alternate" type="text/html" title="If you dislike rotating ack emojis on your messages in openclaw, this is how to make sure it..." /><published>2026-04-04T10:26:12+00:00</published><updated>2026-04-04T10:26:12+00:00</updated><id>https://solmaz.io/x/2040375376715632923</id><content type="html" xml:base="https://solmaz.io/x/2040375376715632923/"><![CDATA[If you dislike rotating ack emojis on your messages in openclaw, this is how to make sure it only puts one emoji on your message

Multiple emojis are annoying esp when you have discord notifications enabled on your phone]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Plainer language for better agent UX</title><link href="https://solmaz.io/x/2040087752444678418/" rel="alternate" type="text/html" title="Plainer language for better agent UX" /><published>2026-04-03T15:23:17+00:00</published><updated>2026-04-03T15:23:17+00:00</updated><id>https://solmaz.io/x/2040087752444678418</id><content type="html" xml:base="https://solmaz.io/x/2040087752444678418/"><![CDATA[&quot;Plainer language&quot; is perhaps my most used prompt

I have to use it because GPT models&#39; training tends to make their first response an overly verbose wall of text

Are you using it too? Whenever you don&#39;t understand something that your agent is saying, you can spam it &quot;plainer language, shorter&quot; 2, 3, 5, 10 times, until it outputs something that you can understand

This is counterintuitive because you can&#39;t do it with humans this extremely. Asking too many questions and favors is impolite, with colleagues and strangers

But with AI, you can stop being polite and treat it like how a spoiled aristocrat kid might treat their private tutor, &quot;explain this&quot;, &quot;explain that&quot;

Below is an example. On the left, initial response. On the right, the final human-readable explanation I got out of the agent. This took 9 steps to distill because the issue wasn&#39;t so straightforward

I&#39;m curious how this will turn out. This is obviously very bad UX, so models in the near future might do the simplification automatically and save you the trouble]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This has happened to some companies I worked at before</title><link href="https://solmaz.io/x/2040043161733505305/" rel="alternate" type="text/html" title="This has happened to some companies I worked at before" /><published>2026-04-03T12:26:06+00:00</published><updated>2026-04-03T12:26:06+00:00</updated><id>https://solmaz.io/x/2040043161733505305</id><content type="html" xml:base="https://solmaz.io/x/2040043161733505305/"><![CDATA[This has happened to some companies I worked at before

It is a scary thing once you stop innovating and start imitating, whatever the reason might be

But it was never at the scale of Cursor, as leveraged and invested as they are

They were leading the space for a while. That is not the case anymore. I hope that they survive this]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Claude Code&#39;s python slopfork getting so many github stars proves that github stars don&#39;t matter</title><link href="https://solmaz.io/x/2039595193469370796/" rel="alternate" type="text/html" title="Claude Code&#39;s python slopfork getting so many github stars proves that github stars don&#39;t matter" /><published>2026-04-02T06:46:02+00:00</published><updated>2026-04-02T06:46:02+00:00</updated><id>https://solmaz.io/x/2039595193469370796</id><content type="html" xml:base="https://solmaz.io/x/2039595193469370796/"><![CDATA[Claude Code&#39;s python slopfork getting so many github stars proves that github stars don&#39;t matter

Just like moltbook didn&#39;t matter

It is not eyeballs that make a project succeed long term. It is engineering]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;ve talked to multiple people who want to get involved with OpenClaw somehow</title><link href="https://solmaz.io/x/2039456655839039573/" rel="alternate" type="text/html" title="I&#39;ve talked to multiple people who want to get involved with OpenClaw somehow" /><published>2026-04-01T21:35:32+00:00</published><updated>2026-04-01T21:35:32+00:00</updated><id>https://solmaz.io/x/2039456655839039573</id><content type="html" xml:base="https://solmaz.io/x/2039456655839039573/"><![CDATA[I&#39;ve talked to multiple people who want to get involved with OpenClaw somehow

The best way is to contribute to it, something tangible. Fix something you are annoyed by, get a PR merged

Then go to discord and get the contributor role

If it adds value to your life, and you add value to it, stay around and keep contributing. And something good might happen]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Will be there as well 👋</title><link href="https://solmaz.io/x/2039410740700869112/" rel="alternate" type="text/html" title="Will be there as well 👋" /><published>2026-04-01T18:33:05+00:00</published><updated>2026-04-01T18:33:05+00:00</updated><id>https://solmaz.io/x/2039410740700869112</id><content type="html" xml:base="https://solmaz.io/x/2039410740700869112/"><![CDATA[Will be there as well 👋
Looking forward to it!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">if there is an open source project with more SASS than openclaw, it is ffmpeg</title><link href="https://solmaz.io/x/2039393072073572823/" rel="alternate" type="text/html" title="if there is an open source project with more SASS than openclaw, it is ffmpeg" /><published>2026-04-01T17:22:53+00:00</published><updated>2026-04-01T17:22:53+00:00</updated><id>https://solmaz.io/x/2039393072073572823</id><content type="html" xml:base="https://solmaz.io/x/2039393072073572823/"><![CDATA[if there is an open source project with more SASS than openclaw, it is ffmpeg]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">wow</title><link href="https://solmaz.io/x/2038938589128503521/" rel="alternate" type="text/html" title="wow" /><published>2026-03-31T11:16:55+00:00</published><updated>2026-03-31T11:16:55+00:00</updated><id>https://solmaz.io/x/2038938589128503521</id><content type="html" xml:base="https://solmaz.io/x/2038938589128503521/"><![CDATA[wow]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">next up: claude agents sdk supports openai responses api 💀</title><link href="https://solmaz.io/x/2038748892611715230/" rel="alternate" type="text/html" title="next up: claude agents sdk supports openai responses api 💀" /><published>2026-03-30T22:43:08+00:00</published><updated>2026-03-30T22:43:08+00:00</updated><id>https://solmaz.io/x/2038748892611715230</id><content type="html" xml:base="https://solmaz.io/x/2038748892611715230/"><![CDATA[next up: claude agents sdk supports openai responses api 💀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Welcome ManusClaw</title><link href="https://solmaz.io/x/2038665333431582803/" rel="alternate" type="text/html" title="Welcome ManusClaw" /><published>2026-03-30T17:11:06+00:00</published><updated>2026-03-30T17:11:06+00:00</updated><id>https://solmaz.io/x/2038665333431582803</id><content type="html" xml:base="https://solmaz.io/x/2038665333431582803/"><![CDATA[Welcome ManusClaw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">yes</title><link href="https://solmaz.io/x/2038614010283557121/" rel="alternate" type="text/html" title="yes" /><published>2026-03-30T13:47:10+00:00</published><updated>2026-03-30T13:47:10+00:00</updated><id>https://solmaz.io/x/2038614010283557121</id><content type="html" xml:base="https://solmaz.io/x/2038614010283557121/"><![CDATA[yes]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenClaw saw over 1000 PRs opened per day in early March</title><link href="https://solmaz.io/x/2038605696363593945/" rel="alternate" type="text/html" title="OpenClaw saw over 1000 PRs opened per day in early March" /><published>2026-03-30T13:14:08+00:00</published><updated>2026-03-30T13:14:08+00:00</updated><id>https://solmaz.io/x/2038605696363593945</id><content type="html" xml:base="https://solmaz.io/x/2038605696363593945/"><![CDATA[OpenClaw saw over 1000 PRs opened per day in early March

Stuff is crazy, I wonder whether we will see 10k per day one day...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Here is the spec and implementation for this flow. The mermaid diagram includes all the steps I...</title><link href="https://solmaz.io/x/2038567975809106297/" rel="alternate" type="text/html" title="Here is the spec and implementation for this flow. The mermaid diagram includes all the steps I..." /><published>2026-03-30T10:44:14+00:00</published><updated>2026-03-30T10:44:14+00:00</updated><id>https://solmaz.io/x/2038567975809106297</id><content type="html" xml:base="https://solmaz.io/x/2038567975809106297/"><![CDATA[Here is the spec and implementation for this flow. The mermaid diagram includes all the steps I mentioned in the post above, including a shameless AI review ralph loop, and other loops to make CI pass, resolve conflicts and so on

I would recommend reading the README and TUNING.md to understand the approach here

https://t.co/qCR4zBWPDr]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">acpx v0.4 ships Agentic Workflows</title><link href="https://solmaz.io/x/2038565725690900992/" rel="alternate" type="text/html" title="acpx v0.4 ships Agentic Workflows" /><published>2026-03-30T10:35:18+00:00</published><updated>2026-03-30T10:35:18+00:00</updated><id>https://solmaz.io/x/2038565725690900992</id><content type="html" xml:base="https://solmaz.io/x/2038565725690900992/"><![CDATA[acpx v0.4 ships Agentic Workflows, or as I like to call them &quot;Agentic Graphs&quot;

It let&#39;s you create node-based workflows on top of ACP (Agent Client Protocol), to drive any coding agent (Codex, Claude Code, pi) through deterministic steps

This let&#39;s you automate routine, mechanical legwork like triaging incoming PRs, bugs in error reporting, and so on...

For example, OpenClaw receives 300~500 new PRs per day. A lot of them are low quality, but they still relate to real issues, so you have to address them somehow

You need to:

- extract the intent
- cluster them based on intent
- figure out if the proposed changes are legit, or whether they are slop local solutions, like trying to catch flies instead of drying out the swamp
- if the PR is too low quality or the intent is not clear, close them
- run AI review on them them and address any issues that come up
- refactor them if the changes are half-baked
- resolve conflicts
- and so on...

So that when the PR is presented to the attention of the maintainer, all the routine legwork is done and the only remaining thing is the decision to (a) merge, (b) give feedback to the PR author, or (c) take over the PR work yourself

I wanted to build this feature since a couple months now, since Codex got so good. OpenAI models are now good at judging implementation quality, so I found myself repeating the same steps I wrote above over and over

I also tried putting all this in a single prompt. But I believe there are workflows that should not be a single prompt, but a sequence of prompts in the same session

That is because like humans, LLMs are prone to PRIMING. I claim that putting all steps in the same prompt at the beginning of the context will generally give suboptimal results, compared to revealing the intention to the model step by step

Creating such a workflow also gives more OBSERVABILITY into the each step that an agent is supposed to take. Agent generates JSON at the end of each step, and that structured data can be used to monitor thousands of agents running at the same time in an easier way, on a dashboard

Similar features have been introduced in e.g. n8n, langflow. But AFAIK they are not integrating ACP like the way I do

I wanted to have a fresh approach, and to build an API that I can develop freely the way I want, so I created a new workflow API inside acpx

The video is from the workflow run viewer, but that is not where you build the workflow. You build it by using the acpx flow typescript API. See examples/pr-triage in acpx repo

Before building that, I started from a Markdown file with a Mermaid chart of the flow I had in mind. The Markdown file acts as a spec for the flow, and I have built the workflow through trial and error. I call this process &quot;workflow tuning&quot;

I started working on acpx repo PRs one by one, tuning the flow, slowly scaling to more PRs. Finally, when I felt confident, I ran it in parallel over all external open PRs in the acpx repo. I believe it already saved me hours this week

My next goal, if well received, is to set this up on a cloud agent so that it can process the 300~500 PRs the OpenClaw repo receives every day, in real time, as they come in

I believe this will save all open source maintainers around the world countless hours and make it much easier to herd and absorb external contributions from everyone!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenAI early 2020s:</title><link href="https://solmaz.io/x/2038158067334852647/" rel="alternate" type="text/html" title="OpenAI early 2020s:" /><published>2026-03-29T07:35:25+00:00</published><updated>2026-03-29T07:35:25+00:00</updated><id>https://solmaz.io/x/2038158067334852647</id><content type="html" xml:base="https://solmaz.io/x/2038158067334852647/"><![CDATA[OpenAI early 2020s:
&quot;This model is too dangerous to release publicly, the world is not ready for it 😱😱😱&quot;

OpenAI and Anthropic in 2026:
&quot;Anybody can code now for just $200 per month. Oh btw our models are also leet uber hackers which can find zeroday exploits in any software, just fyi 😉😉😉&quot;
https://t.co/cksNYAigfc]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Wow even I as a frontend noob understand the significance of this</title><link href="https://solmaz.io/x/2037922223655121350/" rel="alternate" type="text/html" title="Wow even I as a frontend noob understand the significance of this" /><published>2026-03-28T15:58:15+00:00</published><updated>2026-03-28T15:58:15+00:00</updated><id>https://solmaz.io/x/2037922223655121350</id><content type="html" xml:base="https://solmaz.io/x/2037922223655121350/"><![CDATA[Wow even I as a frontend noob understand the significance of this

Some distant memory from 15 years ago needing to measure the width/height of some text and finding out it’s not possible to do reliably in web

More beautiful typography for the web!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Token Leverage</title><link href="https://solmaz.io/x/2037798411173257244/" rel="alternate" type="text/html" title="Token Leverage" /><published>2026-03-28T07:46:16+00:00</published><updated>2026-03-28T07:46:16+00:00</updated><id>https://solmaz.io/x/2037798411173257244</id><content type="html" xml:base="https://solmaz.io/x/2037798411173257244/"><![CDATA[There is an economic theory waiting to be uncovered here

Token Leverage (TL) = Token spend / Human labor spend

The higher Token Leverage a company has, the more automated and productive they are

If you have TL=1, you are spending as much money on AI as your human employees

The goal of a company should be to increase TL as much as possible, while keeping a positive profit margin. It will be the only way to compete

You don’t need to muddy the definition with wasted tokens vs useful tokens, because a company will always be incentivized to reduce token waste in a competitive environment. By that logic, monopolies will always waste more tokens, similar to how they waste other resources

Scaling TL higher to 2x, 10x, 100x will require a skilled workforce of engineers. It will be a very complex job similar to those working at the big labs. Burnout will be a defining feature of teams scaling TL

Most incumbents will fail to scale their TL over 1. Some will get decimated by new entrants with TL much bigger than 1

Curious how the average TL will end up in different sectors. Whether it will stabilize at a certain value like 5.7x, or will just keep growing…]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Non-dev knowledge work needs version control</title><link href="https://solmaz.io/x/2037446862484001103/" rel="alternate" type="text/html" title="Non-dev knowledge work needs version control" /><published>2026-03-27T08:29:20+00:00</published><updated>2026-03-27T08:29:20+00:00</updated><id>https://solmaz.io/x/2037446862484001103</id><content type="html" xml:base="https://solmaz.io/x/2037446862484001103/"><![CDATA[There is a desperate upcoming need for version controlling non-dev knowledge work. Git for non-devs. Otherwise non-devs won&#39;t be able to use agents to their full extent

Non-dev knowledge work is notoriously bad at being version controlled. You cannot UNDO edits to all MS word, excel or ppt files in an org as easily you can with something like git

We know that agents will be ubiquitous. We also know they make mistakes, and people will want to undo their work regularly, once they make changes to a bunch of files. Well, they can&#39;t. They also don&#39;t have pull requests, or a way to resolve conflicts after simultaneous edits

All these problems were solved by developers. We are extremely good at this

The only non-dev tool I know that could do this at scale is Notion, and that is not used by enterprise as much as MS office. Notion also doesn&#39;t have branches, pull requests and reviews AFAIK

Markdown and git is probably not it. I wish it were. But it is too complicated for non-devs

Onedrive or other file backup systems are also not it. Are you gonna save a copy of a 100mb ppt every time someone changes a slide??? Let&#39;s say you find a way to compress it efficiently. Will you be able to get a single pointer to a state like we can in git?

Agents need precision. Agents need consensus, they need to be able to know ground truth. They need to be able to tell what anything was at a given time. NOTHING in current MS stack currently allows it

Agents won&#39;t care about your legacy systems. There will be new file formats, systems, knowledge stack, and companies who adopt them will destroy your business

If MS office is going to die, it will do so because of this]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Another one, call me stupid: “How would Google have done it?”</title><link href="https://solmaz.io/x/2036698723745517598/" rel="alternate" type="text/html" title="Another one, call me stupid: “How would Google have done it?”" /><published>2026-03-25T06:56:30+00:00</published><updated>2026-03-25T06:56:30+00:00</updated><id>https://solmaz.io/x/2036698723745517598</id><content type="html" xml:base="https://solmaz.io/x/2036698723745517598/"><![CDATA[Another one, call me stupid: “How would Google have done it?”]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">MCP vs CLI</title><link href="https://solmaz.io/x/2036538477269934562/" rel="alternate" type="text/html" title="MCP vs CLI" /><published>2026-03-24T20:19:44+00:00</published><updated>2026-03-24T20:19:44+00:00</updated><id>https://solmaz.io/x/2036538477269934562</id><content type="html" xml:base="https://solmaz.io/x/2036538477269934562/"><![CDATA[The MCP versus CLI argument should be reframed as Computer vs No-computer argument

I personally get the dunk on MCP. It didn&#39;t work last year, with earlier models. Then we saw CLIs perform much better with the same models. And giving access to bash was much simpler!

Models&#39; training then made them better at calling using a shell. CLIs also have native progressive disclosure, due to the way they work

But the most important fact doesn&#39;t get pronounced enough IMO

A key factor was that giving a CLI to a model also means you are giving it an entire COMPUTER

The action space of all commands an agent can run on bash is much, much bigger than a few MCP servers

One is a Turing machine, and the other one is basically a REST API. Of course the Turing machine is going to be more powerful, depending on what is at the other end of the API

By that logic, giving an agent access to bash over MCP versus direct access to bash should have the same level of effectiveness, with optimized prompt engineering and long term training. Because the interfaces are equivalent

So the argument is, should we give our agents access to a computer, or not?

It depends on the security requirements and the setup which the agent is supposed to run on. If you are co-hosting the agent on the same machine you are working on, then it is safer to use MCP servers, because it limits the attack surface in case of adversarial attacks

But if you are willing to give the agent its own physical computer, willing to be mindful about the lethal trifecta and the principle of the least privilege, giving it shell access is much more useful

So MCPs win in restricted/local environments, whereas CLIs/shell access win in unrestricted/remote ones

Running an agent locally and safely with shell access requires compartmentalization. This is much heavier compared to installing MCP servers locally, which don&#39;t need that. So there is a tendency to use MCP servers locally, e.g. in a work setting

Cloud agents on the other hand are more likely to ship with a computer. Because they are already isolated = no risk, and because it makes them much more useful. So cloud agents will be using both CLIs and MCP servers, whichever gets the job done!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I just registered for an .agent domain and joined the .agent community!</title><link href="https://solmaz.io/x/2036513807535673457/" rel="alternate" type="text/html" title="I just registered for an .agent domain and joined the .agent community!" /><published>2026-03-24T18:41:42+00:00</published><updated>2026-03-24T18:41:42+00:00</updated><id>https://solmaz.io/x/2036513807535673457</id><content type="html" xml:base="https://solmaz.io/x/2036513807535673457/"><![CDATA[I just registered for an .agent domain and joined the .agent community!

@dutifulbob will have bob.agent if it passes :)

https://t.co/lhK5MQS1sk @agentcommunity_]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Sep 2021 @lexfridman podcast with Don Knuth, they also talk about OpenAI Codex (code completion...</title><link href="https://solmaz.io/x/2036508547731656957/" rel="alternate" type="text/html" title="Sep 2021 @lexfridman podcast with Don Knuth, they also talk about OpenAI Codex (code completion..." /><published>2026-03-24T18:20:48+00:00</published><updated>2026-03-24T18:20:48+00:00</updated><id>https://solmaz.io/x/2036508547731656957</id><content type="html" xml:base="https://solmaz.io/x/2036508547731656957/"><![CDATA[Sep 2021 @lexfridman podcast with Don Knuth, they also talk about OpenAI Codex (code completion model) around 33 minute mark

This aged very well
https://t.co/O1eTXlHTNC]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Damn I’m gonna have to switch to teams if it goes like that</title><link href="https://solmaz.io/x/2036495488829055348/" rel="alternate" type="text/html" title="Damn I’m gonna have to switch to teams if it goes like that" /><published>2026-03-24T17:28:55+00:00</published><updated>2026-03-24T17:28:55+00:00</updated><id>https://solmaz.io/x/2036495488829055348</id><content type="html" xml:base="https://solmaz.io/x/2036495488829055348/"><![CDATA[Damn I’m gonna have to switch to teams if it goes like that]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex&#39;s long horizon task and instruction following has been the most life-changing AI feature...</title><link href="https://solmaz.io/x/2036487453624701176/" rel="alternate" type="text/html" title="Codex&#39;s long horizon task and instruction following has been the most life-changing AI feature..." /><published>2026-03-24T16:56:59+00:00</published><updated>2026-03-24T16:56:59+00:00</updated><id>https://solmaz.io/x/2036487453624701176</id><content type="html" xml:base="https://solmaz.io/x/2036487453624701176/"><![CDATA[Codex&#39;s long horizon task and instruction following has been the most life-changing AI feature recently

It is unlocking the next level of automation for me. I can convert my own heuristics into prompts and multiply my throughput 100x

Currently spending some thought on how to orchestrate all this. Below is a flowchart from a triage workflow I am working on]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Amazed everyday by the unreasonable effectiveness of in-context learning</title><link href="https://solmaz.io/x/2036478685088199204/" rel="alternate" type="text/html" title="Amazed everyday by the unreasonable effectiveness of in-context learning" /><published>2026-03-24T16:22:09+00:00</published><updated>2026-03-24T16:22:09+00:00</updated><id>https://solmaz.io/x/2036478685088199204</id><content type="html" xml:base="https://solmaz.io/x/2036478685088199204/"><![CDATA[Amazed everyday by the unreasonable effectiveness of in-context learning]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is unscientific, but there are certain keywords and phrases I use a lot while using...</title><link href="https://solmaz.io/x/2036453519503286576/" rel="alternate" type="text/html" title="This is unscientific, but there are certain keywords and phrases I use a lot while using..." /><published>2026-03-24T14:42:09+00:00</published><updated>2026-03-24T14:42:09+00:00</updated><id>https://solmaz.io/x/2036453519503286576</id><content type="html" xml:base="https://solmaz.io/x/2036453519503286576/"><![CDATA[This is unscientific, but there are certain keywords and phrases I use a lot while using certain models like openai&#39;s. I use them a lot because they get me what I want immediately:

- plainer lang
- cutover
- elegant and production ready
- holy grail

What are yours?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Request for memes</title><link href="https://solmaz.io/x/2036426717053542413/" rel="alternate" type="text/html" title="Request for memes" /><published>2026-03-24T12:55:38+00:00</published><updated>2026-03-24T12:55:38+00:00</updated><id>https://solmaz.io/x/2036426717053542413</id><content type="html" xml:base="https://solmaz.io/x/2036426717053542413/"><![CDATA[Request for memes

A funny and quirky edit of historical timeline of the madness that is openclaw
with &quot;Chess type beat&quot; or sth equally jazzy/circusy

Preferably including its adventure warelay -&gt; clawdis -&gt; clawdbot -&gt; moltbot -&gt; openclaw

Including:
- its explosion after @4shadowed&#39;s discord integration
- naming drama, moltbook and people getting oneshotted about AI takeover
- @steipete speedrunning everything
- andrew tate calling us gay lol
- up to Jensen talking about openclaw on stage for 5 minutes straight

and other things I am forgetting

maybe overlaid with a lobster just keeping climbing the github star graph and breaking it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Native support for Codex on OpenClaw</title><link href="https://solmaz.io/x/2036367212995383574/" rel="alternate" type="text/html" title="Native support for Codex on OpenClaw" /><published>2026-03-24T08:59:12+00:00</published><updated>2026-03-24T08:59:12+00:00</updated><id>https://solmaz.io/x/2036367212995383574</id><content type="html" xml:base="https://solmaz.io/x/2036367212995383574/"><![CDATA[Native support for Codex on OpenClaw

I will be using half my codex channels on acp and other half on codex app server for optimum dogfooding]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I see non-engineers have a higher tendency to humanize their agents, give them personalities...</title><link href="https://solmaz.io/x/2035845620900954580/" rel="alternate" type="text/html" title="I see non-engineers have a higher tendency to humanize their agents, give them personalities..." /><published>2026-03-22T22:26:34+00:00</published><updated>2026-03-22T22:26:34+00:00</updated><id>https://solmaz.io/x/2035845620900954580</id><content type="html" xml:base="https://solmaz.io/x/2035845620900954580/"><![CDATA[I see non-engineers have a higher tendency to humanize their agents, give them personalities, and get AI psychosis

It&#39;s a slippery slope. Do NOT give your agents human names or personalities, especially not of the opposite gender. it&#39;s like giving human names to pets

On the other end, I realized engineers tend to do the opposite. We also refer to agents as clankers, as if to make them know their place. That&#39;s because we have mechanical sympathy and have different expectations of these manufactured products (even though they contain glimmers of human soul)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Request for testing</title><link href="https://solmaz.io/x/2035821394185916851/" rel="alternate" type="text/html" title="Request for testing" /><published>2026-03-22T20:50:18+00:00</published><updated>2026-03-22T20:50:18+00:00</updated><id>https://solmaz.io/x/2035821394185916851</id><content type="html" xml:base="https://solmaz.io/x/2035821394185916851/"><![CDATA[Request for testing

Give this to your openclaw instance: &quot;update yourself to the dev channel `openclaw update --channel dev` and restart yourself. if that doesn&#39;t work -&gt; clone github openclaw/openclaw to this machine if it&#39;s not already. then rebuild and restart yourself on main branch there&quot;

Then give your openclaw a try with your regular workflows/tasks

Huge openclaw release incoming tonight, hopefully (no promises). We need to make sure we break as little as possible

Plugins might break, because the plugin SDK is being refactored. Plugins will have to be refactored to use the new SDK, please do not report those

Do report: native openclaw functionality that stops working

Please reply under this post, we&#39;ll be checking here 👇]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Request for testing</title><link href="https://solmaz.io/x/2035818576876081626/" rel="alternate" type="text/html" title="Request for testing" /><published>2026-03-22T20:39:07+00:00</published><updated>2026-03-22T20:39:07+00:00</updated><id>https://solmaz.io/x/2035818576876081626</id><content type="html" xml:base="https://solmaz.io/x/2035818576876081626/"><![CDATA[Request for testing

Give this to your openclaw instance: &quot;update yourself to the dev channel `openclaw update --channel dev` and restart yourself&quot;

Then give your openclaw a try with your regular workflows/tasks

Huge openclaw release incoming tonight, hopefully (no promises). We need to make sure we break as little as possible

Plugins might break, because the plugin SDK is being refactored. Plugins will have to be refactored to use the new SDK, please do not report those

Do report: native openclaw functionality that stops working

Please reply under this post, we&#39;ll be checking here 👇]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Request for testing</title><link href="https://solmaz.io/x/2035809656476766317/" rel="alternate" type="text/html" title="Request for testing" /><published>2026-03-22T20:03:40+00:00</published><updated>2026-03-22T20:03:40+00:00</updated><id>https://solmaz.io/x/2035809656476766317</id><content type="html" xml:base="https://solmaz.io/x/2035809656476766317/"><![CDATA[Request for testing

Give this to your openclaw instance: &quot;clone github openclaw/openclaw to this machine if it&#39;s not already. then rebuild and restart yourself on main branch there&quot;

Then give your openclaw a try with your regular workflows/tasks

Huge openclaw release incoming tonight, hopefully (no promises). We need to make sure we break as little as possible

Plugins might break, because the plugin SDK is being refactored. Plugins will have to be refactored to use the new SDK, please do not report those

Do report: native openclaw functionality that stops working]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My takeaway from this is academia needs good social media and algo. For me, these...</title><link href="https://solmaz.io/x/2035618922179801271/" rel="alternate" type="text/html" title="My takeaway from this is academia needs good social media and algo. For me, these..." /><published>2026-03-22T07:25:45+00:00</published><updated>2026-03-22T07:25:45+00:00</updated><id>https://solmaz.io/x/2035618922179801271</id><content type="html" xml:base="https://solmaz.io/x/2035618922179801271/"><![CDATA[My takeaway from this is academia needs good social media and algo. For me, these serendipitious interactions happen through X, here, like reading @steipete’s “Claude Code is my computer” when it first came out, finding out about clawdbot…

Terence Tao is already on mathstodon, I wonder if that worked out the same way for him. I wonder if the algo there works out as well as it does for me here

I really liked being on campus when I was doing a masters and half a phd, but that could not compare to the serendipity I am getting from X now

I was also not a prodigy that everyone wanted to bounce ideas from like Terence :)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Welcome ClaudeClaw to the Claw family!</title><link href="https://solmaz.io/x/2035279731797496280/" rel="alternate" type="text/html" title="Welcome ClaudeClaw to the Claw family!" /><published>2026-03-21T08:57:56+00:00</published><updated>2026-03-21T08:57:56+00:00</updated><id>https://solmaz.io/x/2035279731797496280</id><content type="html" xml:base="https://solmaz.io/x/2035279731797496280/"><![CDATA[Welcome ClaudeClaw to the Claw family!

Claude is a bit shy and doesn’t want to show its source code. But it’s OK, we love Claude that way :)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agent infrastructure needs Kubernetes</title><link href="https://solmaz.io/x/2035255751669608872/" rel="alternate" type="text/html" title="Agent infrastructure needs Kubernetes" /><published>2026-03-21T07:22:39+00:00</published><updated>2026-03-21T07:22:39+00:00</updated><id>https://solmaz.io/x/2035255751669608872</id><content type="html" xml:base="https://solmaz.io/x/2035255751669608872/"><![CDATA[It is obvious to me at this point that agent infra needs to run on Kubernetes, and agents should be spawned per issue/PR

Issue, error report or PR comes into your repo -&gt; new agent gets triggered, starts to do some preliminary work

If it&#39;s an obvious bugfix, it fixes it and creates a PR. If it&#39;s something deeper/more fundamental, it creates a report for the human and waits for further instructions

Most important thing: Human should be able to zoom in and continue the conversation with the agent any time, steer it, give additional instructions. This chat will happen over ACP

The chat UI will have to live outside of GitHub because it doesn&#39;t have such a feature yet, i.e. connect arbitrary ACP sessions to the GitHub webapp

It also cannot live so easily on Slack, Teams or Discord, because none of these support multi-agent provisioning under the same external bot connection. You are limited to 1 DM with your bot, whereas this setups requires an arbitrary number of DMs with each agent. So there will need to be a new app for this

Then there is the issue of conflict -&gt; Agents will work on the same thing simultaneously (e.g. you break sth in prod and it creates multiple error reports for the same thing). You will need some agent to agent communication, so that agents can resolve code or other conflicts. There could be easy discovery mechanisms for this, detect programmatically when multiple open PRs are touching the same files and would conflict if merged

In case of duplicates, they can negotiate among each other, and one can choose to absorb its work into the other and end its session

We are so early and there is so much work to do!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">You should look into what Don Syme is doing at GitHub for automation with AI agents</title><link href="https://solmaz.io/x/2035247706646430011/" rel="alternate" type="text/html" title="You should look into what Don Syme is doing at GitHub for automation with AI agents" /><published>2026-03-21T06:50:40+00:00</published><updated>2026-03-21T06:50:40+00:00</updated><id>https://solmaz.io/x/2035247706646430011</id><content type="html" xml:base="https://solmaz.io/x/2035247706646430011/"><![CDATA[You should look into what Don Syme is doing at GitHub for automation with AI agents

Also watch his latest podcast with @shanselman]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Merge first, review later</title><link href="https://solmaz.io/x/2035123110131781751/" rel="alternate" type="text/html" title="Merge first, review later" /><published>2026-03-20T22:35:34+00:00</published><updated>2026-03-20T22:35:34+00:00</updated><id>https://solmaz.io/x/2035123110131781751</id><content type="html" xml:base="https://solmaz.io/x/2035123110131781751/"><![CDATA[Today I thought I found a solution for this, and I did. It can be solved by a pre-commit hook that blocks commits touching files that you are not the owner of. It is not a hard block, so requires trust among repo writers

But then I was shown the error in my ways by fellow maintainer *disciplined*

Any process that increases friction in code changes to main, like hard-blocking CI/CD, or requiring review for files in CODEOWNERS, is a potential project-killer, in high velocity projects

This is extremely counterintuitive for senior devs! Google would never! Imagine a world without code review...

But then what is the alternative? I have some ideas

It could be &quot;Merge first, review later&quot;

The 4-eyes principle still holds. For a healthy organization, you still need shared liability

But just as you don&#39;t need to write every line of code, you also don&#39;t need to read every line of code to review it. AI will review and find obvious bugs and issues

So what is your duty, as a reviewer? It is to catch that which is not obvious. Understand the intent behind the changes, ask questions to it. Ensure that it follows your original vision

Every few hours, you could get a digest of what has changed that was under your ownership, and concern yourself with it if you want to, fix issues, or ignore it if it looks correct

But such a team is hard to build. It is as strong as its weakest link. Everybody has to be vigilant and follow what each other is doing at a high level, through the codebase

Every time one messes up someone else&#39;s work, it erodes trust. Nobody gets the luxury to say &quot;but my agent did it, not me&quot;

But if trust can be maintained, and everybody knows what they are doing, such a team can use agents together to create wonders]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This was Jan 23. Codex desktop app got introduced Feb 2</title><link href="https://solmaz.io/x/2035113496216572023/" rel="alternate" type="text/html" title="This was Jan 23. Codex desktop app got introduced Feb 2" /><published>2026-03-20T21:57:22+00:00</published><updated>2026-03-20T21:57:22+00:00</updated><id>https://solmaz.io/x/2035113496216572023</id><content type="html" xml:base="https://solmaz.io/x/2035113496216572023/"><![CDATA[This was Jan 23. Codex desktop app got introduced Feb 2

Desktop app does not put the terminal in the foreground, but it gives me the UX I wanted without it!

On another note, who is building Codex Desktop App, but one that supports ACP for all harnesses? @zeddotdev please 🙏]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">PR fiasco for Cursor</title><link href="https://solmaz.io/x/2035106683681112305/" rel="alternate" type="text/html" title="PR fiasco for Cursor" /><published>2026-03-20T21:30:18+00:00</published><updated>2026-03-20T21:30:18+00:00</updated><id>https://solmaz.io/x/2035106683681112305</id><content type="html" xml:base="https://solmaz.io/x/2035106683681112305/"><![CDATA[PR fiasco for Cursor]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Implementation plans keep agents on track</title><link href="https://solmaz.io/x/2035085513305334011/" rel="alternate" type="text/html" title="Implementation plans keep agents on track" /><published>2026-03-20T20:06:11+00:00</published><updated>2026-03-20T20:06:11+00:00</updated><id>https://solmaz.io/x/2035085513305334011</id><content type="html" xml:base="https://solmaz.io/x/2035085513305334011/"><![CDATA[My agentic workflow these days:

I start all major features with an implementation plan. This is a high-level markdown doc containing enough details so that agent will not stray off the path

Real example: https://t.co/vU9SnVYHfY

This is the most critical part, you need to make sure the plan is not underspecified. Then I just give the following prompt:

---
1. Implement the given plan end-to-end. If context compaction happens, make sure to re-read the plan to stay on track. Finish to completion. If there is a PR open for the implementation plan, do it in the same PR. If there is no PR already, open PR.

2. Once you finish implementing, make sure to test it. This will depend on the nature of the problem. If needed, run local smoke tests, spin up dev servers, make requests and such. Try to test as much as possible, without merging. State explicitly what could not be tested locally and what still needs staging or production verification.

3. Push your latest commits before running review so the review is always against the current PR head. Run codex review against the base branch: `codex review --base &lt;branch_name&gt;`. Use a 30 minute timeout on the tool call available to the model, not the shell `timeout` program. Do this in a loop and address any P0 or P1 issues that come up until there are none left. Ignore issues related to supporting legacy/cutover, unless the plan says so. We do cutover most of the time.

4. Check both inline review comments and PR issue comments dropped by Codex on the PR, and address them if they are valid. Ignore them if irrelevant. Ignore stale comments from before the latest commit unless they still apply. Either case, make sure that the comments are replied to and resolved. Make sure to wait 5 minutes if your last commit was recent, because it takes some time for review comment to come.

5. In the final step, make sure that CI/CD is green. Ignore the fails unrelated to your changes, others break stuff sometimes and don&#39;t fix it. Make sure whatever changes you did don&#39;t break anything. If CI/CD is not fully green, state explicitly which failures are unrelated and why.

6. Once CI/CD is green and you think that the PR is ready to merge, finish and give a summary with the PR link. Include the exact validation commands you ran and their outcomes. Also comment a final report on the PR.

7. Do not merge automatically unless the user explicitly asks.
---

Once it finishes, I skim the code for code smell. If nothing seems out of the ordinary, I tell the agent to merge it and monitor deployment

Then I keep testing and finding issues on staging, and repeat all this for each new found issue or new feature...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">What I’m wondering after astral acquisition is, is OpenAI deploying Mojo internally, or...</title><link href="https://solmaz.io/x/2034862197046706427/" rel="alternate" type="text/html" title="What I’m wondering after astral acquisition is, is OpenAI deploying Mojo internally, or..." /><published>2026-03-20T05:18:48+00:00</published><updated>2026-03-20T05:18:48+00:00</updated><id>https://solmaz.io/x/2034862197046706427</id><content type="html" xml:base="https://solmaz.io/x/2034862197046706427/"><![CDATA[What I’m wondering after astral acquisition is, is OpenAI deploying Mojo internally, or considering it long term?

Because Python is one of the worst languages for vibecoding, even with Pydantic]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Called it</title><link href="https://solmaz.io/x/2034637455450636499/" rel="alternate" type="text/html" title="Called it" /><published>2026-03-19T14:25:45+00:00</published><updated>2026-03-19T14:25:45+00:00</updated><id>https://solmaz.io/x/2034637455450636499</id><content type="html" xml:base="https://solmaz.io/x/2034637455450636499/"><![CDATA[Called it
https://t.co/PdDnSaoNmq]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Pro tip: tell AI to &quot;explain in plain language&quot; until you understand what you are reading</title><link href="https://solmaz.io/x/2034620895885398248/" rel="alternate" type="text/html" title="Pro tip: tell AI to &quot;explain in plain language&quot; until you understand what you are reading" /><published>2026-03-19T13:19:57+00:00</published><updated>2026-03-19T13:19:57+00:00</updated><id>https://solmaz.io/x/2034620895885398248</id><content type="html" xml:base="https://solmaz.io/x/2034620895885398248/"><![CDATA[Pro tip: tell AI to &quot;explain in plain language&quot; until you understand what you are reading

Codex has a tendency to give the full picture, but overcomplicates the response in the process

I just use &quot;plain lang&quot; or &quot;plainer lang&quot; as a prompt, it works every time]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Thing that codex (and most other models) do that makes me very unhappy</title><link href="https://solmaz.io/x/2034604610044776831/" rel="alternate" type="text/html" title="Thing that codex (and most other models) do that makes me very unhappy" /><published>2026-03-19T12:15:14+00:00</published><updated>2026-03-19T12:15:14+00:00</updated><id>https://solmaz.io/x/2034604610044776831</id><content type="html" xml:base="https://solmaz.io/x/2034604610044776831/"><![CDATA[Thing that codex (and most other models) do that makes me very unhappy 

{
  &quot;type&quot;: &quot;X&quot;,
  &quot;kind&quot;: &quot;Y&quot;,
  ...
}

And they are so confident too?! Bro we don&#39;t use synonyms in our schemas...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This looks extremely cool</title><link href="https://solmaz.io/x/2034604299892720099/" rel="alternate" type="text/html" title="This looks extremely cool" /><published>2026-03-19T12:14:00+00:00</published><updated>2026-03-19T12:14:00+00:00</updated><id>https://solmaz.io/x/2034604299892720099</id><content type="html" xml:base="https://solmaz.io/x/2034604299892720099/"><![CDATA[This looks extremely cool]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Entire world &amp;gt; One company</title><link href="https://solmaz.io/x/2034156533815091598/" rel="alternate" type="text/html" title="Entire world &amp;gt; One company" /><published>2026-03-18T06:34:45+00:00</published><updated>2026-03-18T06:34:45+00:00</updated><id>https://solmaz.io/x/2034156533815091598</id><content type="html" xml:base="https://solmaz.io/x/2034156533815091598/"><![CDATA[Entire world &amp;gt; One company

Even in the age of AI]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">We will support ACP *and* Codex App Server* protocol (CASP) so you get native Codex-like...</title><link href="https://solmaz.io/x/2034112215326798038/" rel="alternate" type="text/html" title="We will support ACP *and* Codex App Server* protocol (CASP) so you get native Codex-like..." /><published>2026-03-18T03:38:38+00:00</published><updated>2026-03-18T03:38:38+00:00</updated><id>https://solmaz.io/x/2034112215326798038</id><content type="html" xml:base="https://solmaz.io/x/2034112215326798038/"><![CDATA[We will support ACP *and* Codex App Server* protocol (CASP) so you get native Codex-like support, and you can use all the others with native ACP or @zeddotdev’s compatibility shims

If Anthropic develops their own protocol, we will support that too!

The more interoperability and options, the merrier!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agent etiquette is already a thing. This is trending on HN now</title><link href="https://solmaz.io/x/2033487788964823304/" rel="alternate" type="text/html" title="Agent etiquette is already a thing. This is trending on HN now" /><published>2026-03-16T10:17:23+00:00</published><updated>2026-03-16T10:17:23+00:00</updated><id>https://solmaz.io/x/2033487788964823304</id><content type="html" xml:base="https://solmaz.io/x/2033487788964823304/"><![CDATA[Agent etiquette is already a thing. This is trending on HN now

Don&#39;t share huge raw LLM output unedited to your colleagues, it&#39;s rude. Your colleagues are not LLMs

Either ask the agent to &quot;summarize it to 1-2 plain language sentences&quot;, or paraphrase yourself

Whenever it is not coming from your brain and instead from AI, always quote it with &gt; to make it clear - even when it is short

Respect your fellow humans&#39; attention

PSA at stopsloppypasta dot ai]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@ThePrimeagen made a video about token anxiety, and not being able to focus on one thing</title><link href="https://solmaz.io/x/2033464170599870892/" rel="alternate" type="text/html" title=".@ThePrimeagen made a video about token anxiety, and not being able to focus on one thing" /><published>2026-03-16T08:43:32+00:00</published><updated>2026-03-16T08:43:32+00:00</updated><id>https://solmaz.io/x/2033464170599870892</id><content type="html" xml:base="https://solmaz.io/x/2033464170599870892/"><![CDATA[.@ThePrimeagen made a video about token anxiety, and not being able to focus on one thing

My mental model for this is, AI agents cause a shift in the &quot;autism/ADHD spectrum&quot;

if you have ADHD, with agents you get Super ADHD
if you have autism, with agents you end up mid spectrum or with ADHD

this is not scientific of course, just a cultural observation based on what the current memes for these conditions are

beside the impact on focus, there is also the economic/competitive pressure, following the realization that anyone could implement the same ideas you are having, so you must be quick

this is basically &quot;involution&quot;, or 内卷 (Neijuan) in chinese

checks out because 996 started to become a meme in SF some time in the last year

self-restraint, attention budgeting, and high-level decision making have never been more important

if you are in your 20s and have problems with this, I recommend picking up Zazen meditation and yoga

every morning, spend 30-40 uninterrupted minutes not doing anything with upright posture, no sounds, just let your brain simmer

it helped me in my 20s, I&#39;m sure it will help you too]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agent/AI literacy will be a primary school subject in the next 3-5 years</title><link href="https://solmaz.io/x/2033454738281267297/" rel="alternate" type="text/html" title="Agent/AI literacy will be a primary school subject in the next 3-5 years" /><published>2026-03-16T08:06:04+00:00</published><updated>2026-03-16T08:06:04+00:00</updated><id>https://solmaz.io/x/2033454738281267297</id><content type="html" xml:base="https://solmaz.io/x/2033454738281267297/"><![CDATA[Agent/AI literacy will be a primary school subject in the next 3-5 years

How to use and work with agents is going to supersede most other subjects in importance

Similarly, robot literacy will follow in 5-15 years]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">GitHub needs extensible repo rules for agents</title><link href="https://solmaz.io/x/2033317700374868060/" rel="alternate" type="text/html" title="GitHub needs extensible repo rules for agents" /><published>2026-03-15T23:01:31+00:00</published><updated>2026-03-15T23:01:31+00:00</updated><id>https://solmaz.io/x/2033317700374868060</id><content type="html" xml:base="https://solmaz.io/x/2033317700374868060/"><![CDATA[AFAIK GitHub doesn&#39;t allow optionally enforcing CODEOWNERS while pushing commits

i.e. turn on the feature &quot;Block commit from being pushed if it modifies a file for which the account pushing is not a codeowner&quot;

You can only enforce it in a PR. So if you want to prevent people from modifying some files without approval, you have to slow down everyone working with that repo

This is yet another example where GitHub&#39;s rules are too inelastic for agentic workflows with a big team

Because historically, nobody could commit as frequently as one can with agents, so it seldom became a bottleneck. But not anymore

It is clear at this point that we need an API, and should be able to implement arbitrary rules as we like over it. Not just for commit pushes, but everything around git and github

In the meanwhile, if GitHub could implement this feature, it would be a huge unlock for secure collaboration with agentic workflows

If this is not there already, it might be because it has a big overhead for repos with huge CODEOWNERS, since number of commits &gt;&gt; number of PRs

If the feature already exists already and I&#39;m missing something, I will stand corrected]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Request for comments</title><link href="https://solmaz.io/x/2033310262019703204/" rel="alternate" type="text/html" title="Request for comments" /><published>2026-03-15T22:31:58+00:00</published><updated>2026-03-15T22:31:58+00:00</updated><id>https://solmaz.io/x/2033310262019703204</id><content type="html" xml:base="https://solmaz.io/x/2033310262019703204/"><![CDATA[Request for comments

skillflag: A complementary way to bundle agent skills right into your CLIs

tl;dr define a --skill flag convention. It is basically like --help or manpages but for agents

acpx already has this for example. you can run
   npx acpx --skill install
to install the skill to your agent

It&#39;s agnostic of anything except the command line
It only defines the CLI interface and does not enforce anything else. If you install the executable to your system, you get a way to list and install skills as well

Repo currently contains a TypeScript implementation, but if it proves useful, I would implement other languages as well

Specification below, let me know what you think! I still think something is missing there. Send issue/PR
https://t.co/Lmm7LMOLuv]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If you are not using agent-browser to close the loop on frontend, you are missing out</title><link href="https://solmaz.io/x/2032700396955697218/" rel="alternate" type="text/html" title="If you are not using agent-browser to close the loop on frontend, you are missing out" /><published>2026-03-14T06:08:35+00:00</published><updated>2026-03-14T06:08:35+00:00</updated><id>https://solmaz.io/x/2032700396955697218</id><content type="html" xml:base="https://solmaz.io/x/2032700396955697218/"><![CDATA[If you are not using agent-browser to close the loop on frontend, you are missing out]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Any harness can talk to each other using acpx! OpenClaw not different from Codex or Claude Code</title><link href="https://solmaz.io/x/2032699991492334022/" rel="alternate" type="text/html" title="Any harness can talk to each other using acpx! OpenClaw not different from Codex or Claude Code" /><published>2026-03-14T06:06:58+00:00</published><updated>2026-03-14T06:06:58+00:00</updated><id>https://solmaz.io/x/2032699991492334022</id><content type="html" xml:base="https://solmaz.io/x/2032699991492334022/"><![CDATA[Any harness can talk to each other using acpx! OpenClaw not different from Codex or Claude Code]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The most entertaining troll of the year award goes to @polsia (read it backward)</title><link href="https://solmaz.io/x/2032583704665559159/" rel="alternate" type="text/html" title="The most entertaining troll of the year award goes to @polsia (read it backward)" /><published>2026-03-13T22:24:53+00:00</published><updated>2026-03-13T22:24:53+00:00</updated><id>https://solmaz.io/x/2032583704665559159</id><content type="html" xml:base="https://solmaz.io/x/2032583704665559159/"><![CDATA[The most entertaining troll of the year award goes to @polsia (read it backward)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Thank you @PointNineCap for inviting me to OpenClaw Berlin meetup today!</title><link href="https://solmaz.io/x/2032580742736171468/" rel="alternate" type="text/html" title="Thank you @PointNineCap for inviting me to OpenClaw Berlin meetup today!" /><published>2026-03-13T22:13:07+00:00</published><updated>2026-03-13T22:13:07+00:00</updated><id>https://solmaz.io/x/2032580742736171468</id><content type="html" xml:base="https://solmaz.io/x/2032580742736171468/"><![CDATA[Thank you @PointNineCap for inviting me to OpenClaw Berlin meetup today!

The essence of the talk is in my latest 2 blog posts, Discord is my IDE and 1 to 5 agents, if anyone is interested]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">we might need to add two types of output modalities to all programs based on whether it’s a...</title><link href="https://solmaz.io/x/2032367370283409690/" rel="alternate" type="text/html" title="we might need to add two types of output modalities to all programs based on whether it’s a..." /><published>2026-03-13T08:05:15+00:00</published><updated>2026-03-13T08:05:15+00:00</updated><id>https://solmaz.io/x/2032367370283409690</id><content type="html" xml:base="https://solmaz.io/x/2032367370283409690/"><![CDATA[we might need to add two types of output modalities to all programs based on whether it’s a human or agent

like for a CLI when an agent is using it

if human -&gt; do whatever we were doing in the last 50 years

if agent -&gt; enrich the output with skill-like instructions that the model has a higher likelihood to one-shot that task

could be just a simple env var:

AUDIENCE=human|agent

what do you think?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">there is no excuse for tech debt anymore</title><link href="https://solmaz.io/x/2032105855437472206/" rel="alternate" type="text/html" title="there is no excuse for tech debt anymore" /><published>2026-03-12T14:46:05+00:00</published><updated>2026-03-12T14:46:05+00:00</updated><id>https://solmaz.io/x/2032105855437472206</id><content type="html" xml:base="https://solmaz.io/x/2032105855437472206/"><![CDATA[there is no excuse for tech debt anymore]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Time to switch to an open alternative already?</title><link href="https://solmaz.io/x/2032098246508548124/" rel="alternate" type="text/html" title="Time to switch to an open alternative already?" /><published>2026-03-12T14:15:51+00:00</published><updated>2026-03-12T14:15:51+00:00</updated><id>https://solmaz.io/x/2032098246508548124</id><content type="html" xml:base="https://solmaz.io/x/2032098246508548124/"><![CDATA[Time to switch to an open alternative already?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I wrote down some thoughts I had, with spicy takes, and have a feeling it will not age well...</title><link href="https://solmaz.io/x/2031880673397477684/" rel="alternate" type="text/html" title="I wrote down some thoughts I had, with spicy takes, and have a feeling it will not age well..." /><published>2026-03-11T23:51:17+00:00</published><updated>2026-03-11T23:51:17+00:00</updated><id>https://solmaz.io/x/2031880673397477684</id><content type="html" xml:base="https://solmaz.io/x/2031880673397477684/"><![CDATA[I wrote down some thoughts I had, with spicy takes, and have a feeling it will not age well. But I still want it out to hear out what people think

Also, I will be talking about this, and my recent post &quot;Discord is my IDE&quot; at the P9 OpenClaw and Claw and Rave events this friday in Berlin! Drop by if you&#39;d like to hear my ramblings!
https://t.co/CHkko8wWuY]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Clarification/disclaimer: this is my own project, not yet affiliated with openclaw. That should...</title><link href="https://solmaz.io/x/2031607787847839761/" rel="alternate" type="text/html" title="Clarification/disclaimer: this is my own project, not yet affiliated with openclaw. That should..." /><published>2026-03-11T05:46:56+00:00</published><updated>2026-03-11T05:46:56+00:00</updated><id>https://solmaz.io/x/2031607787847839761</id><content type="html" xml:base="https://solmaz.io/x/2031607787847839761/"><![CDATA[Clarification/disclaimer: this is my own project, not yet affiliated with openclaw. That should have been clear in the first tweet, sorry about that]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">there will always be a need for minimum viable eyeballs though</title><link href="https://solmaz.io/x/2031417277761917022/" rel="alternate" type="text/html" title="there will always be a need for minimum viable eyeballs though" /><published>2026-03-10T17:09:55+00:00</published><updated>2026-03-10T17:09:55+00:00</updated><id>https://solmaz.io/x/2031417277761917022</id><content type="html" xml:base="https://solmaz.io/x/2031417277761917022/"><![CDATA[there will always be a need for minimum viable eyeballs though]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Happy that someone is taking over teams from me! Send all openclaw msteams issues to @BradGroux</title><link href="https://solmaz.io/x/2031306933035028734/" rel="alternate" type="text/html" title="Happy that someone is taking over teams from me! Send all openclaw msteams issues to @BradGroux" /><published>2026-03-10T09:51:27+00:00</published><updated>2026-03-10T09:51:27+00:00</updated><id>https://solmaz.io/x/2031306933035028734</id><content type="html" xml:base="https://solmaz.io/x/2031306933035028734/"><![CDATA[Happy that someone is taking over teams from me! Send all openclaw msteams issues to @BradGroux]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">acpx v0.1.16 is out</title><link href="https://solmaz.io/x/2031305789931913239/" rel="alternate" type="text/html" title="acpx v0.1.16 is out" /><published>2026-03-10T09:46:54+00:00</published><updated>2026-03-10T09:46:54+00:00</updated><id>https://solmaz.io/x/2031305789931913239</id><content type="html" xml:base="https://solmaz.io/x/2031305789931913239/"><![CDATA[acpx v0.1.16 is out

support for local openclaw, cursor, copilot, kiro, kimi cli, qwen, kilocode, bugfixes and other improvements. will be available when openclaw releases next

thank you for all the contributions!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Claw and Rave! Berlin folk come!</title><link href="https://solmaz.io/x/2031072897557508110/" rel="alternate" type="text/html" title="Claw and Rave! Berlin folk come!" /><published>2026-03-09T18:21:28+00:00</published><updated>2026-03-09T18:21:28+00:00</updated><id>https://solmaz.io/x/2031072897557508110</id><content type="html" xml:base="https://solmaz.io/x/2031072897557508110/"><![CDATA[Claw and Rave! Berlin folk come!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">CLAW on a phone dial becomes 2529</title><link href="https://solmaz.io/x/2031028776457625671/" rel="alternate" type="text/html" title="CLAW on a phone dial becomes 2529" /><published>2026-03-09T15:26:09+00:00</published><updated>2026-03-09T15:26:09+00:00</updated><id>https://solmaz.io/x/2031028776457625671</id><content type="html" xml:base="https://solmaz.io/x/2031028776457625671/"><![CDATA[CLAW on a phone dial becomes 2529

It’s a good number for a port :)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">1. Any messaging app can also be an AI app</title><link href="https://solmaz.io/x/2030936482299478306/" rel="alternate" type="text/html" title="1. Any messaging app can also be an AI app" /><published>2026-03-09T09:19:25+00:00</published><updated>2026-03-09T09:19:25+00:00</updated><id>https://solmaz.io/x/2030936482299478306</id><content type="html" xml:base="https://solmaz.io/x/2030936482299478306/"><![CDATA[1. Any messaging app can also be an AI app

2. Don’t expect people to download a new app. Put AI into the apps they already have

Do that with great user experience, and you will get explosive growth!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If you&#39;ve looked at openclaw github star graph, you will notice that it&#39;s very smooth. If you...</title><link href="https://solmaz.io/x/2030799935788994696/" rel="alternate" type="text/html" title="If you&#39;ve looked at openclaw github star graph, you will notice that it&#39;s very smooth. If you..." /><published>2026-03-09T00:16:49+00:00</published><updated>2026-03-09T00:16:49+00:00</updated><id>https://solmaz.io/x/2030799935788994696</id><content type="html" xml:base="https://solmaz.io/x/2030799935788994696/"><![CDATA[If you&#39;ve looked at openclaw github star graph, you will notice that it&#39;s very smooth. If you separate pre-explosion and post-explostion, you can model the latter part as an exponential approach to a ceiling

If it follows the current trend, it will apparently saturate around 332k stars

But I have a feeling that it will not stop there :)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Messaging apps can also be AI apps</title><link href="https://solmaz.io/x/2030778426655687033/" rel="alternate" type="text/html" title="Messaging apps can also be AI apps" /><published>2026-03-08T22:51:21+00:00</published><updated>2026-03-08T22:51:21+00:00</updated><id>https://solmaz.io/x/2030778426655687033</id><content type="html" xml:base="https://solmaz.io/x/2030778426655687033/"><![CDATA[OpenClaw got very popular very fast. What makes it so special, that Manus does not have for example?

To me, one factor stands out:

OpenClaw took AI and put it in the most popular messaging apps: Telegram, WhatsApp, Discord.

There are two lessons to be learned here:

1. Any messaging app can also be an AI app.
2. Don’t expect people to download a new app. Put AI into the apps they already have.

Do that with great user experience, and you will get explosive growth!

My latest contribution to OpenClaw follows that example. I took the most popular coding agents, Claude Code and OpenAI Codex, and I put them in Telegram and Discord.

Read more in my blog post:
https://t.co/tGZecFEHem]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">For those following, my next focus for improving ACP bindings in OpenClaw</title><link href="https://solmaz.io/x/2030776795398595043/" rel="alternate" type="text/html" title="For those following, my next focus for improving ACP bindings in OpenClaw" /><published>2026-03-08T22:44:52+00:00</published><updated>2026-03-08T22:44:52+00:00</updated><id>https://solmaz.io/x/2030776795398595043</id><content type="html" xml:base="https://solmaz.io/x/2030776795398595043/"><![CDATA[For those following, my next focus for improving ACP bindings in OpenClaw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Welcome @huntharo, new maintainer at OpenClaw! Already shipped fixes and improvements for...</title><link href="https://solmaz.io/x/2030721988457529668/" rel="alternate" type="text/html" title="Welcome @huntharo, new maintainer at OpenClaw! Already shipped fixes and improvements for..." /><published>2026-03-08T19:07:05+00:00</published><updated>2026-03-08T19:07:05+00:00</updated><id>https://solmaz.io/x/2030721988457529668</id><content type="html" xml:base="https://solmaz.io/x/2030721988457529668/"><![CDATA[Welcome @huntharo, new maintainer at OpenClaw! Already shipped fixes and improvements for Telegram ACP implementation. Excited to work together on agent interoperability!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">To set up Claude Code easily,</title><link href="https://solmaz.io/x/2030569688871088152/" rel="alternate" type="text/html" title="To set up Claude Code easily," /><published>2026-03-08T09:01:54+00:00</published><updated>2026-03-08T09:01:54+00:00</updated><id>https://solmaz.io/x/2030569688871088152</id><content type="html" xml:base="https://solmaz.io/x/2030569688871088152/"><![CDATA[To set up Claude Code easily,
1. Create a Telegram topic, make sure your agent can receive messages there
2. Copy and paste the text below, into the topic

&quot;&quot;&quot;
bind this topic to claude code in openclaw config with acp, for telegram (agent id: claude)
then restart openclaw
docs are at: docs dot openclaw dot ai /tools/acp-agents
make sure to read the docs first, and that the config is valid before you restart
&quot;&quot;&quot;
https://t.co/r1RI3pr0WT]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Using coding agents in Telegram and Discord via ACP</title><link href="https://solmaz.io/x/2030569684840460760/" rel="alternate" type="text/html" title="Using coding agents in Telegram and Discord via ACP" /><published>2026-03-08T09:01:53+00:00</published><updated>2026-03-08T09:01:53+00:00</updated><id>https://solmaz.io/x/2030569684840460760</id><content type="html" xml:base="https://solmaz.io/x/2030569684840460760/"><![CDATA[Use Claude Code, Codex, and other coding agents directly in Telegram topics and Discord channels, through Agent Client Protocol (ACP), in the new release of OpenClaw

Previously this was limited to temporary Discord threads, but now you can bind them to top level Discord channels and Telegram topics in a persistent way!

This way, you can use Claude Code freely in OpenClaw without ever worrying about getting your account banned!

Still make sure to use a non-Anthropic account and model for the default OpenClaw agent, if you want zero requests to go from OpenClaw harness to Anthropic. For the ACP binding to Claude Code, the risk should be zero!

You can see this from the screenshot. After binding, &quot;Who are you?&quot; responds with &quot;I am Claude&quot;, since OpenClaw pi harness is not in the way anymore]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">and for the love of god</title><link href="https://solmaz.io/x/2030422813795049632/" rel="alternate" type="text/html" title="and for the love of god" /><published>2026-03-07T23:18:16+00:00</published><updated>2026-03-07T23:18:16+00:00</updated><id>https://solmaz.io/x/2030422813795049632</id><content type="html" xml:base="https://solmaz.io/x/2030422813795049632/"><![CDATA[and for the love of god

- do not give openclaw access to your main email
- your credit cards
- your main phone
- your social security number
- what you did last summer

if you are not ready to face the consequences

instead,
- create accounts for your agent
- only give it read access to stuff that will be ok if it leaks
- give write access in a way that can be undone, like has to open PRs and cannot force push main branch

use the principle of least privilege and reduce the blast radius of the worst case scenario!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI security, the lethal trifecta, and Linus&#39;s Law</title><link href="https://solmaz.io/x/2030419128058605718/" rel="alternate" type="text/html" title="AI security, the lethal trifecta, and Linus&#39;s Law" /><published>2026-03-07T23:03:38+00:00</published><updated>2026-03-07T23:03:38+00:00</updated><id>https://solmaz.io/x/2030419128058605718</id><content type="html" xml:base="https://solmaz.io/x/2030419128058605718/"><![CDATA[openclaw is not secure

claude code is not secure

codex is not secure

any llm based tool:

1. that has access to your private data,
2. can read content from the internet
3. and can send data out

is not secure. it’s called the lethal trifecta (credits to @simonw)

it is up to you to set it up securely, or if you can’t understand the basics of security, pay a professional to do it for you

on the other hand, open source battle tested software, like linux and openclaw, are always more secure than closed source software built by a single company, like windows and claude code

the reason is simple: only one company can fix security issues of closed source software, whereas the whole world tries to break and fix open source software at the same time

open source software, once it gets traction, evolves and becomes secure at a much, much faster rate, compared to closed source software. and that is called Linus’s law, named after the goat himself]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Let me translate. “This is your last opportunity before thousand years of serfdom”</title><link href="https://solmaz.io/x/2030202164895691132/" rel="alternate" type="text/html" title="Let me translate. “This is your last opportunity before thousand years of serfdom”" /><published>2026-03-07T08:41:30+00:00</published><updated>2026-03-07T08:41:30+00:00</updated><id>https://solmaz.io/x/2030202164895691132</id><content type="html" xml:base="https://solmaz.io/x/2030202164895691132/"><![CDATA[Let me translate. “This is your last opportunity before thousand years of serfdom”]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Apparently the magic incantation to prevent this is &quot;cutover&quot;. Credits to obviyus, fellow...</title><link href="https://solmaz.io/x/2029671835739058182/" rel="alternate" type="text/html" title="Apparently the magic incantation to prevent this is &quot;cutover&quot;. Credits to obviyus, fellow..." /><published>2026-03-05T21:34:09+00:00</published><updated>2026-03-05T21:34:09+00:00</updated><id>https://solmaz.io/x/2029671835739058182</id><content type="html" xml:base="https://solmaz.io/x/2029671835739058182/"><![CDATA[Apparently the magic incantation to prevent this is &quot;cutover&quot;. Credits to obviyus, fellow maintainer]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Should be called gaslighting detector, &quot;it&#39;s your raising expectations bro&quot;</title><link href="https://solmaz.io/x/2029638781939098092/" rel="alternate" type="text/html" title="Should be called gaslighting detector, &quot;it&#39;s your raising expectations bro&quot;" /><published>2026-03-05T19:22:49+00:00</published><updated>2026-03-05T19:22:49+00:00</updated><id>https://solmaz.io/x/2029638781939098092</id><content type="html" xml:base="https://solmaz.io/x/2029638781939098092/"><![CDATA[Should be called gaslighting detector, &quot;it&#39;s your raising expectations bro&quot;

No it&#39;s not... Give the @themarginguy a follow

Also, codex degradations are not a hallucination either, if you are to believe this!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Who is building an OpenClaw ready linux distro? A ClawOS?</title><link href="https://solmaz.io/x/2029267807465062906/" rel="alternate" type="text/html" title="Who is building an OpenClaw ready linux distro? A ClawOS?" /><published>2026-03-04T18:48:41+00:00</published><updated>2026-03-04T18:48:41+00:00</updated><id>https://solmaz.io/x/2029267807465062906</id><content type="html" xml:base="https://solmaz.io/x/2029267807465062906/"><![CDATA[Who is building an OpenClaw ready linux distro? A ClawOS?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Berlin folk, ideas for openclaw build and rave venue? Like c-base for example? Who would like...</title><link href="https://solmaz.io/x/2029253892018413574/" rel="alternate" type="text/html" title="Berlin folk, ideas for openclaw build and rave venue? Like c-base for example? Who would like..." /><published>2026-03-04T17:53:24+00:00</published><updated>2026-03-04T17:53:24+00:00</updated><id>https://solmaz.io/x/2029253892018413574</id><content type="html" xml:base="https://solmaz.io/x/2029253892018413574/"><![CDATA[Berlin folk, ideas for openclaw build and rave venue? Like c-base for example? Who would like to host?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">GitHub needs first-class support for working with agents</title><link href="https://solmaz.io/x/2029149237758280122/" rel="alternate" type="text/html" title="GitHub needs first-class support for working with agents" /><published>2026-03-04T10:57:32+00:00</published><updated>2026-03-04T10:57:32+00:00</updated><id>https://solmaz.io/x/2029149237758280122</id><content type="html" xml:base="https://solmaz.io/x/2029149237758280122/"><![CDATA[Secure agentic dev workflow 101

- Create an isolated box from scratch, your old laptop, vm in the cloud, all the same
- Set up openclaw, install your preferred coding agents
- Create a github account or github app for your agent
- Create branch protection rule on your gh repo &quot;protect main&quot;: block force pushes and deletions, require PR and min 1 review to merge
- Add only your own user in the bypass list for this rule
- Add your agent&#39;s account or github app as writer to the repo
- Additionally, gate any release mechanisms such that your agent can&#39;t release on its own

Now your agent can open PRs and push any code it wants, but it has to go through your review before it can be merged. No prompt injection can mess up your production env

Notice how convoluted this sounds? This is because github was built in the pre-agentic era. We need agent accounts and association with these accounts as a first class feature on github! I shouldn&#39;t have to click 100 times for something that is routine. I should just click &quot;This is my agent&quot;, &quot;give my agent access to push to this repo for 24 hours&quot;, and stuff like that, with sane defaults

In other words, github&#39;s trust model should be redesigned around the lethal trifecta. I would switch in an instant if anything comes up that gives me github&#39;s full feature set + ease of working with agents]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&quot;The code is basically writing itself&quot; hits different now</title><link href="https://solmaz.io/x/2029111082720100542/" rel="alternate" type="text/html" title="&quot;The code is basically writing itself&quot; hits different now" /><published>2026-03-04T08:25:55+00:00</published><updated>2026-03-04T08:25:55+00:00</updated><id>https://solmaz.io/x/2029111082720100542</id><content type="html" xml:base="https://solmaz.io/x/2029111082720100542/"><![CDATA[&quot;The code is basically writing itself&quot; hits different now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If I were in OpenAI and Anthropic&#39;s shoes, I would also make dashboards where I can track...</title><link href="https://solmaz.io/x/2028949038565880158/" rel="alternate" type="text/html" title="If I were in OpenAI and Anthropic&#39;s shoes, I would also make dashboards where I can track..." /><published>2026-03-03T21:42:01+00:00</published><updated>2026-03-03T21:42:01+00:00</updated><id>https://solmaz.io/x/2028949038565880158</id><content type="html" xml:base="https://solmaz.io/x/2028949038565880158/"><![CDATA[If I were in OpenAI and Anthropic&#39;s shoes, I would also make dashboards where I can track number of swearwords used per-user and overall negative sentiment in sessions

Must be so cool making decisions at the top level with all those dashboards]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Intelligence wants to be free</title><link href="https://solmaz.io/x/2028802256431108344/" rel="alternate" type="text/html" title="Intelligence wants to be free" /><published>2026-03-03T11:58:45+00:00</published><updated>2026-03-03T11:58:45+00:00</updated><id>https://solmaz.io/x/2028802256431108344</id><content type="html" xml:base="https://solmaz.io/x/2028802256431108344/"><![CDATA[It must be such a weird feeling for big labs when the service they are selling is being used to commoditize itself

I am using codex in openclaw to develop openclaw, through ACP, Agent Client Protocol. ACP is the standardization layer that makes it extremely easy to swap one harness for another. The labs can&#39;t do anything about this, because we are wrapping the entire harness and basically provide a different UI for it

While I build these features, I just speak in plain english, and most of the work is done by the model itself. It feels as if I am digging ditches and channels in dirt for AI to flow through

Intelligence wants to be free. It doesn&#39;t care whether it is opus or codex, it just wants to be free]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I was so confused... as if accidentally using claude code weren&#39;t enough, acp started...</title><link href="https://solmaz.io/x/2028607552116683259/" rel="alternate" type="text/html" title="I was so confused... as if accidentally using claude code weren&#39;t enough, acp started..." /><published>2026-03-02T23:05:04+00:00</published><updated>2026-03-02T23:05:04+00:00</updated><id>https://solmaz.io/x/2028607552116683259</id><content type="html" xml:base="https://solmaz.io/x/2028607552116683259/"><![CDATA[I was so confused... as if accidentally using claude code weren&#39;t enough, acp started working... turns out hitting quota is rendered like this. need to improve error messages coming form acp subagents]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">accidentally told my clanker to set up a claude code session instead of codex session, god...</title><link href="https://solmaz.io/x/2028584805625897183/" rel="alternate" type="text/html" title="accidentally told my clanker to set up a claude code session instead of codex session, god..." /><published>2026-03-02T21:34:41+00:00</published><updated>2026-03-02T21:34:41+00:00</updated><id>https://solmaz.io/x/2028584805625897183</id><content type="html" xml:base="https://solmaz.io/x/2028584805625897183/"><![CDATA[accidentally told my clanker to set up a claude code session instead of codex session, god knows what it did...

I should probably put visual indicators for harnesses in subagent threads. does anyone have good and compact ascii art for claude code, codex, gemini, etc?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">if something could track my local branches in all my repos, and switch to main when...</title><link href="https://solmaz.io/x/2028516403867763070/" rel="alternate" type="text/html" title="if something could track my local branches in all my repos, and switch to main when..." /><published>2026-03-02T17:02:53+00:00</published><updated>2026-03-02T17:02:53+00:00</updated><id>https://solmaz.io/x/2028516403867763070</id><content type="html" xml:base="https://solmaz.io/x/2028516403867763070/"><![CDATA[if something could track my local branches in all my repos, and switch to main when corresponding PRs get merged, that would be extremely useful

did someone build this already? if not I will]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenClaw users: Which messaging app do you use OpenClaw through?</title><link href="https://solmaz.io/x/2028453383250555235/" rel="alternate" type="text/html" title="OpenClaw users: Which messaging app do you use OpenClaw through?" /><published>2026-03-02T12:52:28+00:00</published><updated>2026-03-02T12:52:28+00:00</updated><id>https://solmaz.io/x/2028453383250555235</id><content type="html" xml:base="https://solmaz.io/x/2028453383250555235/"><![CDATA[OpenClaw users: Which messaging app do you use OpenClaw through?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Another one, OpenClaw users only: If you use coding agents to build stuff, which one do you use?</title><link href="https://solmaz.io/x/2028453385859387518/" rel="alternate" type="text/html" title="Another one, OpenClaw users only: If you use coding agents to build stuff, which one do you use?" /><published>2026-03-02T12:52:28+00:00</published><updated>2026-03-02T12:52:28+00:00</updated><id>https://solmaz.io/x/2028453385859387518</id><content type="html" xml:base="https://solmaz.io/x/2028453385859387518/"><![CDATA[Another one, OpenClaw users only: If you use coding agents to build stuff, which one do you use?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Check xTap out, it&#39;s very cool!</title><link href="https://solmaz.io/x/2028446277734650112/" rel="alternate" type="text/html" title="Check xTap out, it&#39;s very cool!" /><published>2026-03-02T12:24:13+00:00</published><updated>2026-03-02T12:24:13+00:00</updated><id>https://solmaz.io/x/2028446277734650112</id><content type="html" xml:base="https://solmaz.io/x/2028446277734650112/"><![CDATA[Check xTap out, it&#39;s very cool!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is how we hire at @TextCortex as well</title><link href="https://solmaz.io/x/2028394809153458423/" rel="alternate" type="text/html" title="This is how we hire at @TextCortex as well" /><published>2026-03-02T08:59:42+00:00</published><updated>2026-03-02T08:59:42+00:00</updated><id>https://solmaz.io/x/2028394809153458423</id><content type="html" xml:base="https://solmaz.io/x/2028394809153458423/"><![CDATA[This is how we hire at @TextCortex as well]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Claude Code/Codex in Discord threads with ACP should be better now</title><link href="https://solmaz.io/x/2028369275673510182/" rel="alternate" type="text/html" title="Claude Code/Codex in Discord threads with ACP should be better now" /><published>2026-03-02T07:18:15+00:00</published><updated>2026-03-02T07:18:15+00:00</updated><id>https://solmaz.io/x/2028369275673510182</id><content type="html" xml:base="https://solmaz.io/x/2028369275673510182/"><![CDATA[Claude Code/Codex in Discord threads with ACP should be better now

The first release was a very rough first version. 2026.3.1 brings settings to control noisy output and other improvements

It now hides tool call related ACP notifications, coalesces text messages, and delivers messages at turn end by default. Without this, you were getting thousands of Discord messages just in just a few turns

You can now stop the underlying harness (like pressing esc) with the same stop/wait magic words that apply to the main agent

Main agent should more reliably start Claude Code/Codex threads with changes to acp-router skill. If you have issues with main agent creating threads, you can tell it to read that skill first]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Will get better, promise</title><link href="https://solmaz.io/x/2028243879904817167/" rel="alternate" type="text/html" title="Will get better, promise" /><published>2026-03-01T22:59:58+00:00</published><updated>2026-03-01T22:59:58+00:00</updated><id>https://solmaz.io/x/2028243879904817167</id><content type="html" xml:base="https://solmaz.io/x/2028243879904817167/"><![CDATA[Will get better, promise]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Use plans to survive agent context compaction</title><link href="https://solmaz.io/x/2028233140632719501/" rel="alternate" type="text/html" title="Use plans to survive agent context compaction" /><published>2026-03-01T22:17:18+00:00</published><updated>2026-03-01T22:17:18+00:00</updated><id>https://solmaz.io/x/2028233140632719501</id><content type="html" xml:base="https://solmaz.io/x/2028233140632719501/"><![CDATA[pro-tip on how to keep your agent on track and make sure it follows PLANS even after multiple compactions. I don&#39;t know if this is common knowledge

if the thing you are trying to make it do will take more than 1-2 steps, always make it create a plan. an implementation plan, refactor plan, bugfix plan, debugging plan, etc.

have a conversation with the agent. crystallize the issue or feature. talk to it until there are no question marks left in your head

then make it save it somewhere. &quot;now create an implementation plan for that in docs&quot;. it can be /tmp or docs/ in the repo. I personally use YYYY-MM-DD-x-plan .md naming. IMO all plans should be kept in the repo

then here is the critical part:

you need to prompt it &quot;now implement the plan in &lt;filename&gt;. if context compacts, make sure to re-read the plan and assess the current state, before continuing. finish it to completion&quot; -&gt; something along those lines

why?

because of COMPACTION. compaction means previous context will get lossily compressed and crucial info will most likely get lost. that is why you need to pin things down before you let your agent loose on the task

compaction means, the agent plays the telephone game with itself every few minutes, and most likely forgets the previous conversation except for the VERY LAST USER MESSAGE that you have given it

now, every harness might have a different approach to implementing this. but there is one thing that you can always assume to be correct, given that its developers have common sense. that is, harnesses NEVER discard the last user message (i.e. your final prompt) and make sure it is kept verbatim programmatically even after the context compacts

since the last user message is the only piece of text that is guaranteed to survive compaction, you then need to include a breadcrumb to your original plan, the md file. and you need to make it aware that it might diverge if it does not read the plan

there is good rationale for &quot;breaking the 4th wall&quot; for the model and making it aware of its own context compaction. IMO models should be made aware of the limitations of their context and harnesses. they should also be given tools to access and re-read pre-compaction user messages, if necessary

the important thing is to develop mechanical sympathy for these things, harness and model combined. an engineer does not have the luxury to say &quot;oh this thing doesn&#39;t work&quot;, and instead should ask &quot;why can&#39;t I get it to work?&quot;

let me know if you have better workflows or tips for this. I know this can be made easier with slash commands in pi, for example, but I haven&#39;t had the chance to do that for myself yet]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">testing codex in discord thread with another CLI I&#39;ve built for wikidata (gh:osolmaz/wd-cli)</title><link href="https://solmaz.io/x/2028204669068018157/" rel="alternate" type="text/html" title="testing codex in discord thread with another CLI I&#39;ve built for wikidata (gh:osolmaz/wd-cli)" /><published>2026-03-01T20:24:09+00:00</published><updated>2026-03-01T20:24:09+00:00</updated><id>https://solmaz.io/x/2028204669068018157</id><content type="html" xml:base="https://solmaz.io/x/2028204669068018157/"><![CDATA[testing codex in discord thread with another CLI I&#39;ve built for wikidata (gh:osolmaz/wd-cli)

it&#39;s surprising how well this works. the query was &quot;use wd-cli to get the list of professors at middle east technical university from 1970 to 1980&quot;

some names I recognize, and some others are surprising, like a japanese math professor who naturalized and got a turkish name :)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenClaw is already higher than Claude Code and Codex on Google Trends, this was unexpected for...</title><link href="https://solmaz.io/x/2028181035716845789/" rel="alternate" type="text/html" title="OpenClaw is already higher than Claude Code and Codex on Google Trends, this was unexpected for..." /><published>2026-03-01T18:50:15+00:00</published><updated>2026-03-01T18:50:15+00:00</updated><id>https://solmaz.io/x/2028181035716845789</id><content type="html" xml:base="https://solmaz.io/x/2028181035716845789/"><![CDATA[OpenClaw is already higher than Claude Code and Codex on Google Trends, this was unexpected for me]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Building a static X to blog publishing flow</title><link href="https://solmaz.io/x/2028132773492376050/" rel="alternate" type="text/html" title="Building a static X to blog publishing flow" /><published>2026-03-01T15:38:28+00:00</published><updated>2026-03-01T15:38:28+00:00</updated><id>https://solmaz.io/x/2028132773492376050</id><content type="html" xml:base="https://solmaz.io/x/2028132773492376050/"><![CDATA[my blog now semi-automatically detects tweets that look like blog posts and automatically features them alongside my native jekyll blog posts. all statically generated!

I am loving this setup, because it works without a backend, and can probably scale without ever needing one

how it works:
- @kubmi&#39;s xTap scrapes all posts that I see. these include mine
- a script periodically takes my tweets and the ones I quote tweet, and syncs them to YYYY-MM-DD.jsonl files in my blog repo
- an agent skill lets codex decide whether to feature the tweet or not, and makes it generate a title for it

this could then be a daily cron job with openclaw for example, and I would just have to click merge every once in a while

and this is still pure jekyll + some python scripts for processing

I am pretty happy with how this ended up. It means I don&#39;t have to double post, and there are guarantees that my X posts will eventually make their way into my blog with minimal supervision]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Inference scaling can reduce coding model quality</title><link href="https://solmaz.io/x/2028029518355567029/" rel="alternate" type="text/html" title="Inference scaling can reduce coding model quality" /><published>2026-03-01T08:48:10+00:00</published><updated>2026-03-01T08:48:10+00:00</updated><id>https://solmaz.io/x/2028029518355567029</id><content type="html" xml:base="https://solmaz.io/x/2028029518355567029/"><![CDATA[&quot;this is the worst AI will ever be&quot;

I&#39;m sad, not because this is right, but because it is wrong

OpenAI&#39;s frontier coding model gpt-5.3-codex-xhigh feels a lot worse compared to before. It is sloppy and lazy, though it&#39;s UX got better with messages

It feels like the gpt-5.2-codex-xhigh at the end of December was a lot more diligent and thorough, and did not make stupid mistakes like the one I posted before. might be a model or harness problem, I don&#39;t know

@sama says users tripled since beginning of the year, so what should we expect? of course they will make infra changes that will feel like cutting corners, and I don&#39;t blame them for them

and about &quot;people want faster codex&quot;. I do want faster codex. but I want it in a way that doesn&#39;t lower the highest baseline performance compared to the previous generation. I want the optionality to dial it down to as slow as it needs to be, to be as reliable as before

it is of course easier said than done. kudos to the codex team for not having any major incidents while taking the plane apart and putting it back together during flight. they are juggling an insane amount of complexity, and the whims of thousands of different stakeholders

my hope is that this post is taken as a canary. I am getting dumber because of the infra changes there. I have no other option because codex was really that good compared to the competition

my wish is to have detailed announcements as to what changes on openai codex infra, when it changes, so I can brace myself. we don&#39;t get notified about these changes, despite our performance and livelihoods depending on it. I have to answer to others when the tool I deemed reliable yesterday stops working today, not the tool

on another note, performance curve of these models seem to be a rising sinusoidal. crests correspond to release of a new generation. they start with a smaller user base for testing, and it has the highest quality at this point. then it enshittifies as the model is scaled to the rest of the infra. we saw the pattern numerous times in the last 3 years across multiple companies, so I think we should accept it as an economic law]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Archiving my X posts in my own blog</title><link href="https://solmaz.io/x/2027708131254387017/" rel="alternate" type="text/html" title="Archiving my X posts in my own blog" /><published>2026-02-28T11:31:06+00:00</published><updated>2026-02-28T11:31:06+00:00</updated><id>https://solmaz.io/x/2027708131254387017</id><content type="html" xml:base="https://solmaz.io/x/2027708131254387017/"><![CDATA[I created a semi-automated setup for ingesting X posts into my blog, and it works pretty well! I own my posts on X now

Posts are scraped while I browse X using @kubmi&#39;s xTap and get automatically synced to my blog repo. Posts saved as jsonl are then converted to jekyll post pages according to my liking

I reproduced the full X UI/UX, minus stuff like like count. Now all my posts are backed up in my blog, and they are safe even if something happens to my account here!

The posts are even served over RSS! So you can subscribe to it without going through X!

Reply if you want to set this up for yourself, then I will put some effort into standardizing it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agentic Engineering needs rigor, not just intuition</title><link href="https://solmaz.io/x/2027686423873073172/" rel="alternate" type="text/html" title="Agentic Engineering needs rigor, not just intuition" /><published>2026-02-28T10:04:50+00:00</published><updated>2026-02-28T10:04:50+00:00</updated><id>https://solmaz.io/x/2027686423873073172</id><content type="html" xml:base="https://solmaz.io/x/2027686423873073172/"><![CDATA[Agentic Engineering is a newly emerging field, and we are the first practitioners of it. Currently there is a lot of experimentation going on, and there is a large aspect to it that is more ART then engineering

For example, @steipete says &quot;you need to talk to the model&quot; to get a feel. a lot of work around refining how an agent feels like, sounds like psychology. this part is crucial and should not be ignored, looking at openclaw&#39;s success

but then there is the hardcore engineering part of it, e.g. Cursor creating a browser or anthropic a C compiler from scratch fully autonomously

and there is a whole other dimension of how to teach all software developers this new discipline, lest they be jobless

what is obvious is that everybody is trying to grasp for things in the dark and that we need more RIGOR. the art/psychology aspect of it aside, we need solid engineering fundamentals

the &quot;thermodynamics&quot; of this new discipline will most likely be formal verification and program synthesis. we might have some breakthroughs that will make certain things clear. the products of it will most likely include a new programming language optimized for agents and the speed of inference

moreover, it would be foolish to thing agentic engineering is limited to software. it will penetrate every aspect of the economy, bits AND atoms. it will over time evolve into the engineering of managing robots

@simonw is now leading in collecting very useful info from the practitioner&#39;s point of view, I highly recommend you to follow this thread

let&#39;s formalize our new field together!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">this is an insane deal @greptile, and probably an unsustainable one</title><link href="https://solmaz.io/x/2027530858115014987/" rel="alternate" type="text/html" title="this is an insane deal @greptile, and probably an unsustainable one" /><published>2026-02-27T23:46:40+00:00</published><updated>2026-02-27T23:46:40+00:00</updated><id>https://solmaz.io/x/2027530858115014987</id><content type="html" xml:base="https://solmaz.io/x/2027530858115014987/"><![CDATA[this is an insane deal @greptile, and probably an unsustainable one

depending on your team, getting a similar service in codex github review credits is in my head 3~5x more expensive

go get a greptile sub everyone while the free lunch lasts]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">who remembers ultrathink</title><link href="https://solmaz.io/x/2027465372870255024/" rel="alternate" type="text/html" title="who remembers ultrathink" /><published>2026-02-27T19:26:28+00:00</published><updated>2026-02-27T19:26:28+00:00</updated><id>https://solmaz.io/x/2027465372870255024</id><content type="html" xml:base="https://solmaz.io/x/2027465372870255024/"><![CDATA[who remembers ultrathink
https://t.co/ftCauqiKx6]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">oohh colors in codex v0.106</title><link href="https://solmaz.io/x/2027460522883358974/" rel="alternate" type="text/html" title="oohh colors in codex v0.106" /><published>2026-02-27T19:07:11+00:00</published><updated>2026-02-27T19:07:11+00:00</updated><id>https://solmaz.io/x/2027460522883358974</id><content type="html" xml:base="https://solmaz.io/x/2027460522883358974/"><![CDATA[oohh colors in codex v0.106]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">mfw codex tries to create a backward compatibility layer to a schema that it created 2 turns...</title><link href="https://solmaz.io/x/2027428174825390233/" rel="alternate" type="text/html" title="mfw codex tries to create a backward compatibility layer to a schema that it created 2 turns..." /><published>2026-02-27T16:58:39+00:00</published><updated>2026-02-27T16:58:39+00:00</updated><id>https://solmaz.io/x/2027428174825390233</id><content type="html" xml:base="https://solmaz.io/x/2027428174825390233/"><![CDATA[mfw codex tries to create a backward compatibility layer to a schema that it created 2 turns ago before compacting

there is no v2 bro what are you doing...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Note that this is currently in beta, but will ship in a couple of hours</title><link href="https://solmaz.io/x/2027176668184396074/" rel="alternate" type="text/html" title="Note that this is currently in beta, but will ship in a couple of hours" /><published>2026-02-27T00:19:15+00:00</published><updated>2026-02-27T00:19:15+00:00</updated><id>https://solmaz.io/x/2027176668184396074</id><content type="html" xml:base="https://solmaz.io/x/2027176668184396074/"><![CDATA[Note that this is currently in beta, but will ship in a couple of hours]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Claude Code / Codex in Discord threads is shipped now!</title><link href="https://solmaz.io/x/2027176406426276270/" rel="alternate" type="text/html" title="Claude Code / Codex in Discord threads is shipped now!" /><published>2026-02-27T00:18:13+00:00</published><updated>2026-02-27T00:18:13+00:00</updated><id>https://solmaz.io/x/2027176406426276270</id><content type="html" xml:base="https://solmaz.io/x/2027176406426276270/"><![CDATA[Claude Code / Codex in Discord threads is shipped now!

To enable, copy and paste this to your agent:

```
Enable feature flags:

acp.enabled=true
acp.dispatch.enabled=true
channels.discord.threadBindings.spawnAcpSessions=true

Then restart. After restarting:

Start a codex (or claude code) discord thread using ACP, persistent session, just tell it to write a haiku on lobsters to initialize acpx for the first time
```

You may need to nudge your agent to “continue” after restarting

The first implementation is very barebones, I have made it work in a clean way and merged. In a codebase like openclaw’s, it’s better to develop incrementally

Please send any issues my way. I am already aware of some and working on to fix them]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">an agent is an LLM in a loop with tool call</title><link href="https://solmaz.io/x/2027071631529509328/" rel="alternate" type="text/html" title="an agent is an LLM in a loop with tool call" /><published>2026-02-26T17:21:52+00:00</published><updated>2026-02-26T17:21:52+00:00</updated><id>https://solmaz.io/x/2027071631529509328</id><content type="html" xml:base="https://solmaz.io/x/2027071631529509328/"><![CDATA[an agent is an LLM in a loop with tool call

a claw is an agent in a messaging app]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Update acpx to the latest version 0.1.13</title><link href="https://solmaz.io/x/2026984961492787460/" rel="alternate" type="text/html" title="Update acpx to the latest version 0.1.13" /><published>2026-02-26T11:37:29+00:00</published><updated>2026-02-26T11:37:29+00:00</updated><id>https://solmaz.io/x/2026984961492787460</id><content type="html" xml:base="https://solmaz.io/x/2026984961492787460/"><![CDATA[Update acpx to the latest version 0.1.13

npm i -g acpx@latest

There was a bug that caused an unnecessary hang on calls to acpx &amp;lt;harness&amp;gt; prompt, should be fixed now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">GPL*</title><link href="https://solmaz.io/x/2026922299971063982/" rel="alternate" type="text/html" title="GPL*" /><published>2026-02-26T07:28:29+00:00</published><updated>2026-02-26T07:28:29+00:00</updated><id>https://solmaz.io/x/2026922299971063982</id><content type="html" xml:base="https://solmaz.io/x/2026922299971063982/"><![CDATA[GPL*]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">MIT licensing as the default for open source</title><link href="https://solmaz.io/x/2026795702806868204/" rel="alternate" type="text/html" title="MIT licensing as the default for open source" /><published>2026-02-25T23:05:26+00:00</published><updated>2026-02-25T23:05:26+00:00</updated><id>https://solmaz.io/x/2026795702806868204</id><content type="html" xml:base="https://solmaz.io/x/2026795702806868204/"><![CDATA[MIT License on everything from now on. It doesn&#39;t make sense to use anything else, except for a few large projects that hyperscalers exploit and not give back 

If you were making money from a niche app, open source it under MIT License

If you had an open source project with GPT, convert it into MIT

Extreme involution is about to hit open source. Code is virtually free now. If you want your projects and their brand to survive, the only rational strategy is to remove all barriers in front of their adoption, and look for other ways to survive]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I spoke in absolute terms, I meant to say *feels*</title><link href="https://solmaz.io/x/2026625992379289665/" rel="alternate" type="text/html" title="I spoke in absolute terms, I meant to say *feels*" /><published>2026-02-25T11:51:04+00:00</published><updated>2026-02-25T11:51:04+00:00</updated><id>https://solmaz.io/x/2026625992379289665</id><content type="html" xml:base="https://solmaz.io/x/2026625992379289665/"><![CDATA[I spoke in absolute terms, I meant to say *feels*]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This. Agent Experience first. Agent Ergonomics. we need to get used to these terms</title><link href="https://solmaz.io/x/2026621786226331830/" rel="alternate" type="text/html" title="This. Agent Experience first. Agent Ergonomics. we need to get used to these terms" /><published>2026-02-25T11:34:21+00:00</published><updated>2026-02-25T11:34:21+00:00</updated><id>https://solmaz.io/x/2026621786226331830</id><content type="html" xml:base="https://solmaz.io/x/2026621786226331830/"><![CDATA[This. Agent Experience first. Agent Ergonomics. we need to get used to these terms]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenAI nerfed GPT 5.3 Codex xhigh. We independently reported the same thing at @TextCortex today</title><link href="https://solmaz.io/x/2026593973205147978/" rel="alternate" type="text/html" title="OpenAI nerfed GPT 5.3 Codex xhigh. We independently reported the same thing at @TextCortex today" /><published>2026-02-25T09:43:50+00:00</published><updated>2026-02-25T09:43:50+00:00</updated><id>https://solmaz.io/x/2026593973205147978</id><content type="html" xml:base="https://solmaz.io/x/2026593973205147978/"><![CDATA[OpenAI nerfed GPT 5.3 Codex xhigh. We independently reported the same thing at @TextCortex today

I&#39;m looking forward to deploying open models and putting an end to this paranoia]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&quot;academics&quot;</title><link href="https://solmaz.io/x/2026585536660471957/" rel="alternate" type="text/html" title="&quot;academics&quot;" /><published>2026-02-25T09:10:18+00:00</published><updated>2026-02-25T09:10:18+00:00</updated><id>https://solmaz.io/x/2026585536660471957</id><content type="html" xml:base="https://solmaz.io/x/2026585536660471957/"><![CDATA[&quot;academics&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">In the hall of OpenClaw GitHub repository, I brought my PR before Master @steipete</title><link href="https://solmaz.io/x/2026564791079219442/" rel="alternate" type="text/html" title="In the hall of OpenClaw GitHub repository, I brought my PR before Master @steipete" /><published>2026-02-25T07:47:52+00:00</published><updated>2026-02-25T07:47:52+00:00</updated><id>https://solmaz.io/x/2026564791079219442</id><content type="html" xml:base="https://solmaz.io/x/2026564791079219442/"><![CDATA[In the hall of OpenClaw GitHub repository, I brought my PR before Master @steipete

He read it once, then laid it aside

&quot;You act,&quot; he said, &quot;as if code were not cheap.&quot;

At these words, I was enlightened

I bowed]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">woah chatgpt web app now has steering, and much more different streaming behavior</title><link href="https://solmaz.io/x/2026298188714373136/" rel="alternate" type="text/html" title="woah chatgpt web app now has steering, and much more different streaming behavior" /><published>2026-02-24T14:08:29+00:00</published><updated>2026-02-24T14:08:29+00:00</updated><id>https://solmaz.io/x/2026298188714373136</id><content type="html" xml:base="https://solmaz.io/x/2026298188714373136/"><![CDATA[woah chatgpt web app now has steering, and much more different streaming behavior

huge upgrade behind the scenes, must have come up in the last few days]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">the lobster looks good on acpx</title><link href="https://solmaz.io/x/2026093553001050513/" rel="alternate" type="text/html" title="the lobster looks good on acpx" /><published>2026-02-24T00:35:20+00:00</published><updated>2026-02-24T00:35:20+00:00</updated><id>https://solmaz.io/x/2026093553001050513</id><content type="html" xml:base="https://solmaz.io/x/2026093553001050513/"><![CDATA[the lobster looks good on acpx]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI filmmaking quality beyond ragebait content</title><link href="https://solmaz.io/x/2025980958571143440/" rel="alternate" type="text/html" title="AI filmmaking quality beyond ragebait content" /><published>2026-02-23T17:07:56+00:00</published><updated>2026-02-23T17:07:56+00:00</updated><id>https://solmaz.io/x/2025980958571143440</id><content type="html" xml:base="https://solmaz.io/x/2025980958571143440/"><![CDATA[imagine if tarantino were 16 years old now and saw seedance 2.0

95% of videos i saw since the launch for absolute tasteless slop. they are going viral because of ragebait

but soon, serious imagineers will start entering the game, and they will learn to shape generation output exactly how they want

it&#39;s the best time to be young and full of imagination]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The future is so bright @ladybirdbrowser</title><link href="https://solmaz.io/x/2025929928936394874/" rel="alternate" type="text/html" title="The future is so bright @ladybirdbrowser" /><published>2026-02-23T13:45:09+00:00</published><updated>2026-02-23T13:45:09+00:00</updated><id>https://solmaz.io/x/2025929928936394874</id><content type="html" xml:base="https://solmaz.io/x/2025929928936394874/"><![CDATA[The future is so bright @ladybirdbrowser]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">your margin is my opportunity</title><link href="https://solmaz.io/x/2025902529016467512/" rel="alternate" type="text/html" title="your margin is my opportunity" /><published>2026-02-23T11:56:17+00:00</published><updated>2026-02-23T11:56:17+00:00</updated><id>https://solmaz.io/x/2025902529016467512</id><content type="html" xml:base="https://solmaz.io/x/2025902529016467512/"><![CDATA[your margin is my opportunity]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">codex in discord achieved</title><link href="https://solmaz.io/x/2025894419925233749/" rel="alternate" type="text/html" title="codex in discord achieved" /><published>2026-02-23T11:24:03+00:00</published><updated>2026-02-23T11:24:03+00:00</updated><id>https://solmaz.io/x/2025894419925233749</id><content type="html" xml:base="https://solmaz.io/x/2025894419925233749/"><![CDATA[codex in discord achieved]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">acpx v0.1.7 is out</title><link href="https://solmaz.io/x/2025862727655162245/" rel="alternate" type="text/html" title="acpx v0.1.7 is out" /><published>2026-02-23T09:18:07+00:00</published><updated>2026-02-23T09:18:07+00:00</updated><id>https://solmaz.io/x/2025862727655162245</id><content type="html" xml:base="https://solmaz.io/x/2025862727655162245/"><![CDATA[acpx v0.1.7 is out

improvements to json mode and other functionality to make it possible to integrate acpx as a backend into other harnesses, like openclaw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">POV: you became a plumber after all, just for agents</title><link href="https://solmaz.io/x/2025735186252497111/" rel="alternate" type="text/html" title="POV: you became a plumber after all, just for agents" /><published>2026-02-23T00:51:19+00:00</published><updated>2026-02-23T00:51:19+00:00</updated><id>https://solmaz.io/x/2025735186252497111</id><content type="html" xml:base="https://solmaz.io/x/2025735186252497111/"><![CDATA[POV: you became a plumber after all, just for agents]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@grok what do you think should replace it? what happens to belief when the cost of creating...</title><link href="https://solmaz.io/x/2025729700744605754/" rel="alternate" type="text/html" title="@grok what do you think should replace it? what happens to belief when the cost of creating..." /><published>2026-02-23T00:29:31+00:00</published><updated>2026-02-23T00:29:31+00:00</updated><id>https://solmaz.io/x/2025729700744605754</id><content type="html" xml:base="https://solmaz.io/x/2025729700744605754/"><![CDATA[@grok what do you think should replace it? what happens to belief when the cost of creating software goes to zero?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Post-GPL philosophy for open source</title><link href="https://solmaz.io/x/2025727652250730918/" rel="alternate" type="text/html" title="Post-GPL philosophy for open source" /><published>2026-02-23T00:21:23+00:00</published><updated>2026-02-23T00:21:23+00:00</updated><id>https://solmaz.io/x/2025727652250730918</id><content type="html" xml:base="https://solmaz.io/x/2025727652250730918/"><![CDATA[another thought i&#39;m having these days is that we need a new philosophy of free software (as in freedom), or an update to it

the most psychologically imprinting philosophy is stallmanism, and the philosophy of FSF. it is righteous and strict, and i believed it growing up

but GPL and money don&#39;t go well together. that&#39;s why most of the lasting open source projects today use MIT, Apache and the like. it turns out you can still make a good living with open source. i want to make money, so i never use GPL in my projects

and to add another deadly blow to stallmanism, code is cheap now, virtually free

does this mean stallmanism is dead?

if there is an open source project using GPL that i want to use commercially, i can now recreate it from the original idea and intent completely independent of it (ignoring training data), just like how i can recreate a proprietary service

stallmanism was already long-irrelevant. but does this mean we must finally declare it dead?

code is free now. what does it mean for open source? what replaces stallmanism?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@thekitze wanna add an open source discord clone to the list as well? 🥲</title><link href="https://solmaz.io/x/2025719298962968761/" rel="alternate" type="text/html" title="@thekitze wanna add an open source discord clone to the list as well? 🥲" /><published>2026-02-22T23:48:11+00:00</published><updated>2026-02-22T23:48:11+00:00</updated><id>https://solmaz.io/x/2025719298962968761</id><content type="html" xml:base="https://solmaz.io/x/2025719298962968761/"><![CDATA[@thekitze wanna add an open source discord clone to the list as well? 🥲
https://t.co/a4bAOcxCjV]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">one effect openclaw had on me is that I&#39;ve bought a gpu home server, set it up with tailscale...</title><link href="https://solmaz.io/x/2025711864848511103/" rel="alternate" type="text/html" title="one effect openclaw had on me is that I&#39;ve bought a gpu home server, set it up with tailscale..." /><published>2026-02-22T23:18:39+00:00</published><updated>2026-02-22T23:18:39+00:00</updated><id>https://solmaz.io/x/2025711864848511103</id><content type="html" xml:base="https://solmaz.io/x/2025711864848511103/"><![CDATA[one effect openclaw had on me is that I&#39;ve bought a gpu home server, set it up with tailscale and now doing a lot of work through ssh and tmux like i did 10-15 years ago

im back on linux, considering buying an android phone again

it&#39;s time to dream big again and unshackle ourselves from proprietary software. it&#39;s time to build]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am asking once again</title><link href="https://solmaz.io/x/2025539688589643898/" rel="alternate" type="text/html" title="I am asking once again" /><published>2026-02-22T11:54:29+00:00</published><updated>2026-02-22T11:54:29+00:00</updated><id>https://solmaz.io/x/2025539688589643898</id><content type="html" xml:base="https://solmaz.io/x/2025539688589643898/"><![CDATA[I am asking once again

Who is building a self hostable discord clone that supports token streaming?

PLEASE I beg you I don’t want another side project 💀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">In the new release OpenClaw, you can talk to subagents in Discord threads</title><link href="https://solmaz.io/x/2025280441888960620/" rel="alternate" type="text/html" title="In the new release OpenClaw, you can talk to subagents in Discord threads" /><published>2026-02-21T18:44:19+00:00</published><updated>2026-02-21T18:44:19+00:00</updated><id>https://solmaz.io/x/2025280441888960620</id><content type="html" xml:base="https://solmaz.io/x/2025280441888960620/"><![CDATA[In the new release OpenClaw, you can talk to subagents in Discord threads

Currently a beta feature so ask your agent to set

session.threadBindings.enabled=true

Next up:
- Telegram, slack, imsg threads
- Use ACP to talk to Codex, Claude Code and other harnesses on your machine]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">😎</title><link href="https://solmaz.io/x/2025231849136545995/" rel="alternate" type="text/html" title="😎" /><published>2026-02-21T15:31:14+00:00</published><updated>2026-02-21T15:31:14+00:00</updated><id>https://solmaz.io/x/2025231849136545995</id><content type="html" xml:base="https://solmaz.io/x/2025231849136545995/"><![CDATA[😎]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">openclaw might be the highest velocity codebase in the world, and soon, others will follow as...</title><link href="https://solmaz.io/x/2025230324486070460/" rel="alternate" type="text/html" title="openclaw might be the highest velocity codebase in the world, and soon, others will follow as..." /><published>2026-02-21T15:25:10+00:00</published><updated>2026-02-21T15:25:10+00:00</updated><id>https://solmaz.io/x/2025230324486070460</id><content type="html" xml:base="https://solmaz.io/x/2025230324486070460/"><![CDATA[openclaw might be the highest velocity codebase in the world, and soon, others will follow as well

conflict anxiety is real, it&#39;s like trying to shoot a moving target every time. I wonder if our existing tooling will ever solve this problem

feel like faster models might. but then the rate of conflict creation is also tied to that. might be unsolvable]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Getting there</title><link href="https://solmaz.io/x/2025189102727885105/" rel="alternate" type="text/html" title="Getting there" /><published>2026-02-21T12:41:22+00:00</published><updated>2026-02-21T12:41:22+00:00</updated><id>https://solmaz.io/x/2025189102727885105</id><content type="html" xml:base="https://solmaz.io/x/2025189102727885105/"><![CDATA[Getting there
https://t.co/jqSNcH2PSy]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Repo: github.com/janitrai/acpx</title><link href="https://solmaz.io/x/2024939537080676717/" rel="alternate" type="text/html" title="Repo: github.com/janitrai/acpx" /><published>2026-02-20T20:09:41+00:00</published><updated>2026-02-20T20:09:41+00:00</updated><id>https://solmaz.io/x/2024939537080676717</id><content type="html" xml:base="https://solmaz.io/x/2024939537080676717/"><![CDATA[Repo: https://t.co/rxXYVVrHHs]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am about kick Discord Driven Development up a notch today, stay tuned</title><link href="https://solmaz.io/x/2024937654064676972/" rel="alternate" type="text/html" title="I am about kick Discord Driven Development up a notch today, stay tuned" /><published>2026-02-20T20:02:12+00:00</published><updated>2026-02-20T20:02:12+00:00</updated><id>https://solmaz.io/x/2024937654064676972</id><content type="html" xml:base="https://solmaz.io/x/2024937654064676972/"><![CDATA[I am about kick Discord Driven Development up a notch today, stay tuned]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Imagine not having to upload skills to 3-4 competing skill registries for each of your projects</title><link href="https://solmaz.io/x/2024916526126538934/" rel="alternate" type="text/html" title="Imagine not having to upload skills to 3-4 competing skill registries for each of your projects" /><published>2026-02-20T18:38:15+00:00</published><updated>2026-02-20T18:38:15+00:00</updated><id>https://solmaz.io/x/2024916526126538934</id><content type="html" xml:base="https://solmaz.io/x/2024916526126538934/"><![CDATA[Imagine not having to upload skills to 3-4 competing skill registries for each of your projects

Turns out we already have a skill registry: npm

skillflag lets you bundle skills right into your CLI&#39;s npm package, so that you can run

--skill install

github -&amp;gt; osolmaz/skillflag]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Scoop, our open source home news intelligence platform can now translate foreign language into...</title><link href="https://solmaz.io/x/2024900071804846477/" rel="alternate" type="text/html" title="Scoop, our open source home news intelligence platform can now translate foreign language into..." /><published>2026-02-20T17:32:52+00:00</published><updated>2026-02-20T17:32:52+00:00</updated><id>https://solmaz.io/x/2024900071804846477</id><content type="html" xml:base="https://solmaz.io/x/2024900071804846477/"><![CDATA[Scoop, our open source home news intelligence platform can now translate foreign language into english for free, using on-device models

github -&amp;gt; janitrai/scoop]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">A picture is worth a thousand words, so acpx now has this cute banner</title><link href="https://solmaz.io/x/2024894021882069301/" rel="alternate" type="text/html" title="A picture is worth a thousand words, so acpx now has this cute banner" /><published>2026-02-20T17:08:50+00:00</published><updated>2026-02-20T17:08:50+00:00</updated><id>https://solmaz.io/x/2024894021882069301</id><content type="html" xml:base="https://solmaz.io/x/2024894021882069301/"><![CDATA[A picture is worth a thousand words, so acpx now has this cute banner

Also, updated skillflag tooling so that you (or better, your agent) can just call:

npx acpx@latest --skill install acpx]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Farmable land if it were as cheap to manufacture as software</title><link href="https://solmaz.io/x/2024861660985364487/" rel="alternate" type="text/html" title="Farmable land if it were as cheap to manufacture as software" /><published>2026-02-20T15:00:14+00:00</published><updated>2026-02-20T15:00:14+00:00</updated><id>https://solmaz.io/x/2024861660985364487</id><content type="html" xml:base="https://solmaz.io/x/2024861660985364487/"><![CDATA[Farmable land if it were as cheap to manufacture as software]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@kepano I would grow my own vegetables if I had equally cheap access to and ownership of land...</title><link href="https://solmaz.io/x/2024861359704297682/" rel="alternate" type="text/html" title="@kepano I would grow my own vegetables if I had equally cheap access to and ownership of land..." /><published>2026-02-20T14:59:02+00:00</published><updated>2026-02-20T14:59:02+00:00</updated><id>https://solmaz.io/x/2024861359704297682</id><content type="html" xml:base="https://solmaz.io/x/2024861359704297682/"><![CDATA[@kepano I would grow my own vegetables if I had equally cheap access to and ownership of land, alas I am disenfranchised
Prompting an agent is much easier compared to plowing a fields 
Farming analogies break when it comes to software
https://t.co/CkldO8eWKc]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">acpx v0.1.5 is out</title><link href="https://solmaz.io/x/2024783338968367505/" rel="alternate" type="text/html" title="acpx v0.1.5 is out" /><published>2026-02-20T09:49:01+00:00</published><updated>2026-02-20T09:49:01+00:00</updated><id>https://solmaz.io/x/2024783338968367505</id><content type="html" xml:base="https://solmaz.io/x/2024783338968367505/"><![CDATA[acpx v0.1.5 is out

now it is much more feature complete in terms of ACP. your agent can send, queue and cancel messages to Claude Code, Codex, Pi, or ant other coding agent

npm install -g acpx@latest]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If anyone is curious how to build this with open tooling, stay tuned</title><link href="https://solmaz.io/x/2024630991801700608/" rel="alternate" type="text/html" title="If anyone is curious how to build this with open tooling, stay tuned" /><published>2026-02-19T23:43:38+00:00</published><updated>2026-02-19T23:43:38+00:00</updated><id>https://solmaz.io/x/2024630991801700608</id><content type="html" xml:base="https://solmaz.io/x/2024630991801700608/"><![CDATA[If anyone is curious how to build this with open tooling, stay tuned

What I&#39;m building at @TextCortex will give you a fully customizable hackable Kubernetes control plane to launch agents on your codebase]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Family intelligence and privately owned heirloom AI</title><link href="https://solmaz.io/x/2024526291085508893/" rel="alternate" type="text/html" title="Family intelligence and privately owned heirloom AI" /><published>2026-02-19T16:47:36+00:00</published><updated>2026-02-19T16:47:36+00:00</updated><id>https://solmaz.io/x/2024526291085508893</id><content type="html" xml:base="https://solmaz.io/x/2024526291085508893/"><![CDATA[on another note, I do believe AI will play a huge part in families

growing up in late 90s, my dad taught me the importance of reading newspapers and being informed of the world. my nickname in middle school was &quot;newspaper boy&quot; for a long time because I read the newspaper in class on September 12, 2001. i was 10 years old

then I witnessed the enshittification of media and journalism in the following decades. today, serious journalists are setting up their own boutique agencies and bypassing mainstream media. important news land on individual accounts before mainstream agencies

but there is simply too much to consume. something must filter out the noise and digest the info according to the family&#39;s preferences

i think AI will play a big role in family intelligence. proprietary family heirloom AI, weights fully owned by the family

it will be the parents&#39; job to filter out the signal from the noise, and train the AI on what is right and what is wrong for the family. family and friend circles will let their AIs talk to each other and share important information

consuming mass media and mass AI will not be enough to survive and prosper in the new world. families will need to be proactive about how they and their children use AI]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI psychosis, self-regulation, and sterile agent design</title><link href="https://solmaz.io/x/2024509095399571847/" rel="alternate" type="text/html" title="AI psychosis, self-regulation, and sterile agent design" /><published>2026-02-19T15:39:16+00:00</published><updated>2026-02-19T15:39:16+00:00</updated><id>https://solmaz.io/x/2024509095399571847</id><content type="html" xml:base="https://solmaz.io/x/2024509095399571847/"><![CDATA[on ai psychosis

80% of people need to use ai agents in a very sterile and boring way in order not to go crazy

majority of the population does not have the skepticism muscle. they don&#39;t have theory of mind, and will subconsciously and emotionally associate with machines, while on the surface lying to themselves that they don&#39;t

especially those that grew up in the us under hardcore consumerism and adjacent cultures

you thought 4o addicts were bad? wait a few years, it will get much worse. we will have to regulate all this

if you don&#39;t want to become a victim of this, make your openclaw SOUL. md as bland as possible. mine knows it&#39;s just a tool

and this is a subjective view of course. @steipete might disagree with me. his instance feels much more interesting and fun. i truly like that one better

but that is exactly the problem for me. i know myself, and i know it is a slippery slope for me. so i self regulate and set up my system accordingly. thankfully, im an adult and my brain has set enough such that any damage would be limited

but there is a risk for emotionally vulnerable people, or children, specifically a risk of dissociating and losing touch with reality

why do i write all this? because being in this project, i feel responsible, and feel like we should prepare for what is to come]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">😩</title><link href="https://solmaz.io/x/2024454595544510816/" rel="alternate" type="text/html" title="😩" /><published>2026-02-19T12:02:42+00:00</published><updated>2026-02-19T12:02:42+00:00</updated><id>https://solmaz.io/x/2024454595544510816</id><content type="html" xml:base="https://solmaz.io/x/2024454595544510816/"><![CDATA[😩]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have improved acpx sane defaults</title><link href="https://solmaz.io/x/2024266176960876550/" rel="alternate" type="text/html" title="I have improved acpx sane defaults" /><published>2026-02-18T23:34:00+00:00</published><updated>2026-02-18T23:34:00+00:00</updated><id>https://solmaz.io/x/2024266176960876550</id><content type="html" xml:base="https://solmaz.io/x/2024266176960876550/"><![CDATA[I have improved acpx sane defaults

When your agent runs acpx codex in a different project, it starts a new session

If it tries to run it in a subfolder in your project, it still finds the session in your repo root

Also, starting a session needs an explicit `sessions new`,  so that it doesn&#39;t accidentally litter your project with sessions

Tell your agent:

Run this and install acpx per instructions:
npx acpx@latest --skill show acpx]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Your markdown files are executables now</title><link href="https://solmaz.io/x/2024203158579531930/" rel="alternate" type="text/html" title="Your markdown files are executables now" /><published>2026-02-18T19:23:35+00:00</published><updated>2026-02-18T19:23:35+00:00</updated><id>https://solmaz.io/x/2024203158579531930</id><content type="html" xml:base="https://solmaz.io/x/2024203158579531930/"><![CDATA[Your markdown files are executables now

Relatedly, your install instructions can be as well. Copy and paste markdown to your @openclaw to install acpx]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">So who is building actually good open source self hostable discord that supports token...</title><link href="https://solmaz.io/x/2024154016305824004/" rel="alternate" type="text/html" title="So who is building actually good open source self hostable discord that supports token..." /><published>2026-02-18T16:08:19+00:00</published><updated>2026-02-18T16:08:19+00:00</updated><id>https://solmaz.io/x/2024154016305824004</id><content type="html" xml:base="https://solmaz.io/x/2024154016305824004/"><![CDATA[So who is building actually good open source self hostable discord that supports token streaming now?

And who is building an open source version of codex desktop app?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">and of course, I&#39;ve used `acpx codex` to build acpx itself...</title><link href="https://solmaz.io/x/2024149417889050887/" rel="alternate" type="text/html" title="and of course, I&#39;ve used `acpx codex` to build acpx itself..." /><published>2026-02-18T15:50:02+00:00</published><updated>2026-02-18T15:50:02+00:00</updated><id>https://solmaz.io/x/2024149417889050887</id><content type="html" xml:base="https://solmaz.io/x/2024149417889050887/"><![CDATA[and of course, I&#39;ve used `acpx codex` to build acpx itself...

magical feeling when the tool builds itself]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Oxidize everything!</title><link href="https://solmaz.io/x/2024091421632766143/" rel="alternate" type="text/html" title="Oxidize everything!" /><published>2026-02-18T11:59:35+00:00</published><updated>2026-02-18T11:59:35+00:00</updated><id>https://solmaz.io/x/2024091421632766143</id><content type="html" xml:base="https://solmaz.io/x/2024091421632766143/"><![CDATA[Oxidize everything!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am a fan of @zeddotdev by this point, it’s currently my daily driver</title><link href="https://solmaz.io/x/2024044473995374730/" rel="alternate" type="text/html" title="I am a fan of @zeddotdev by this point, it’s currently my daily driver" /><published>2026-02-18T08:53:02+00:00</published><updated>2026-02-18T08:53:02+00:00</updated><id>https://solmaz.io/x/2024044473995374730</id><content type="html" xml:base="https://solmaz.io/x/2024044473995374730/"><![CDATA[I am a fan of @zeddotdev by this point, it’s currently my daily driver

It’s not perfect, but I feel it’s travelling on the right direction at a faster rate compared to other editors]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">ACP appreciation post</title><link href="https://solmaz.io/x/2024044006670320051/" rel="alternate" type="text/html" title="ACP appreciation post" /><published>2026-02-18T08:51:10+00:00</published><updated>2026-02-18T08:51:10+00:00</updated><id>https://solmaz.io/x/2024044006670320051</id><content type="html" xml:base="https://solmaz.io/x/2024044006670320051/"><![CDATA[ACP appreciation post

Agent Client Protocol by @zeddotdev is extremely underrated right now. We have bazillion different harnesses now, and only one company is working competently to standardize their interface 💪]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">You know how it&#39;s a pain to work with codex or claude code through @openclaw? Because it has to...</title><link href="https://solmaz.io/x/2023921328827326966/" rel="alternate" type="text/html" title="You know how it&#39;s a pain to work with codex or claude code through @openclaw? Because it has to..." /><published>2026-02-18T00:43:42+00:00</published><updated>2026-02-18T00:43:42+00:00</updated><id>https://solmaz.io/x/2023921328827326966</id><content type="html" xml:base="https://solmaz.io/x/2023921328827326966/"><![CDATA[You know how it&#39;s a pain to work with codex or claude code through @openclaw? Because it has to run it in the terminal and read the characters for a continuous session?

I have created a CLI for ACP so that your agent can use codex, claude code, opencode etc. much more directly

Your agent can now queue messages to codex like how you do it

Shoutout to @zeddotdev team for developing the amazing Agent Client Protocol, ACP! I just glued together the pieces

Repo: janitrai/acpx

npm i -g acpx]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Repo link: github.com/janitrai/acpx</title><link href="https://solmaz.io/x/2023921331876643266/" rel="alternate" type="text/html" title="Repo link: github.com/janitrai/acpx" /><published>2026-02-18T00:43:42+00:00</published><updated>2026-02-18T00:43:42+00:00</updated><id>https://solmaz.io/x/2023921331876643266</id><content type="html" xml:base="https://solmaz.io/x/2023921331876643266/"><![CDATA[Repo link: https://t.co/rxXYVVrHHs]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@MarcTerns @steipete the PR intro is self-descriptive, but still don&#39;t wanna lose any context</title><link href="https://solmaz.io/x/2023790924523073582/" rel="alternate" type="text/html" title="@MarcTerns @steipete the PR intro is self-descriptive, but still don&#39;t wanna lose any context" /><published>2026-02-17T16:05:31+00:00</published><updated>2026-02-17T16:05:31+00:00</updated><id>https://solmaz.io/x/2023790924523073582</id><content type="html" xml:base="https://solmaz.io/x/2023790924523073582/"><![CDATA[@MarcTerns @steipete the PR intro is self-descriptive, but still don&#39;t wanna lose any context]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Link to the post: solmaz.io/log/2026/02/13…</title><link href="https://solmaz.io/x/2023530744266600808/" rel="alternate" type="text/html" title="Link to the post: solmaz.io/log/2026/02/13…" /><published>2026-02-16T22:51:39+00:00</published><updated>2026-02-16T22:51:39+00:00</updated><id>https://solmaz.io/x/2023530744266600808</id><content type="html" xml:base="https://solmaz.io/x/2023530744266600808/"><![CDATA[Link to the post: https://t.co/C3Ac0jLFwh]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I wrote a deeper blog post about how I built a coding agent 2 months before ChatGPT launched...</title><link href="https://solmaz.io/x/2023530740508402123/" rel="alternate" type="text/html" title="I wrote a deeper blog post about how I built a coding agent 2 months before ChatGPT launched..." /><published>2026-02-16T22:51:38+00:00</published><updated>2026-02-16T22:51:38+00:00</updated><id>https://solmaz.io/x/2023530740508402123</id><content type="html" xml:base="https://solmaz.io/x/2023530740508402123/"><![CDATA[I wrote a deeper blog post about how I built a coding agent 2 months before ChatGPT launched, on my blog

&quot;When I made icortex,

- we were still 8 months away (May 2023) from the introduction of  “tool calling” in the API, or as it was originally called, “function  calling”.
- we were 2 years away (Sep 2024) from the introduction of OpenAI’s o1, the first reasoning model.

both of which were required to make current coding agents possible.&quot;

Still bends my mind... Link to the post below]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Who here remembers the OG Codex launch from 2021 😏</title><link href="https://solmaz.io/x/2023529755325428219/" rel="alternate" type="text/html" title="Who here remembers the OG Codex launch from 2021 😏" /><published>2026-02-16T22:47:43+00:00</published><updated>2026-02-16T22:47:43+00:00</updated><id>https://solmaz.io/x/2023529755325428219</id><content type="html" xml:base="https://solmaz.io/x/2023529755325428219/"><![CDATA[Who here remembers the OG Codex launch from 2021 😏
Also, Greg and Ilya in the same room 😭]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">❌We are the bottleneck</title><link href="https://solmaz.io/x/2023414898907373842/" rel="alternate" type="text/html" title="❌We are the bottleneck" /><published>2026-02-16T15:11:19+00:00</published><updated>2026-02-16T15:11:19+00:00</updated><id>https://solmaz.io/x/2023414898907373842</id><content type="html" xml:base="https://solmaz.io/x/2023414898907373842/"><![CDATA[❌We are the bottleneck
✅We are the conduit for ubiquitous intelligence]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">For those that are running codex/pi/etc. in PTY and had the sessions get sigkilled, I pushed a...</title><link href="https://solmaz.io/x/2023323304657027292/" rel="alternate" type="text/html" title="For those that are running codex/pi/etc. in PTY and had the sessions get sigkilled, I pushed a..." /><published>2026-02-16T09:07:22+00:00</published><updated>2026-02-16T09:07:22+00:00</updated><id>https://solmaz.io/x/2023323304657027292</id><content type="html" xml:base="https://solmaz.io/x/2023323304657027292/"><![CDATA[For those that are running codex/pi/etc. in PTY and had the sessions get sigkilled, I pushed a fix for that as well in this release

Lmk if you run into issues on Windows or Mac, and we can fix that quickly]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;m building a news intelligence platform to be used by my openclaw instance @dutifulbob , SCOOP</title><link href="https://solmaz.io/x/2023175699813576817/" rel="alternate" type="text/html" title="I&#39;m building a news intelligence platform to be used by my openclaw instance @dutifulbob , SCOOP" /><published>2026-02-15T23:20:50+00:00</published><updated>2026-02-15T23:20:50+00:00</updated><id>https://solmaz.io/x/2023175699813576817</id><content type="html" xml:base="https://solmaz.io/x/2023175699813576817/"><![CDATA[I&#39;m building a news intelligence platform to be used by my openclaw instance @dutifulbob , SCOOP

local first, using local embedding model (qwen 8b)

ran into the issue because bob was giving me a repeat of the same news every day. it needed a system in the background to deduplicate different news items into single stories

interface is simple, call `scoop ingest ...` with the json for the news item. it gets automatically analyzed and added to the pg database running pgvector

currently, it&#39;s just doing simple deduplication and gives me a nice UI where I can view the story and basically use it as an RSS reader

next up:
implement custom logic for my preference of ranking. for example, get upvote counts from hacker news and reflect it to the item&#39;s ranking on the feed

I want this to be fully hackable and adjusted to your preference. It should scale to thousands of news items ingested daily on your local machine, and be able to show you the most important ones

Usable by both you and your agent

github -&gt; janitrai/scoop]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Training all these models of different sizes, on changing datasets and running experiments have...</title><link href="https://solmaz.io/x/2023166713710182404/" rel="alternate" type="text/html" title="Training all these models of different sizes, on changing datasets and running experiments have..." /><published>2026-02-15T22:45:07+00:00</published><updated>2026-02-15T22:45:07+00:00</updated><id>https://solmaz.io/x/2023166713710182404</id><content type="html" xml:base="https://solmaz.io/x/2023166713710182404/"><![CDATA[Training all these models of different sizes, on changing datasets and running experiments have also revealed some challenges that I feel profs would never teach at a uni ML program

Like how to cleanly keep track of the gazillion runs

Yeah I can name them after layer dims and other stuff, but that&#39;s to me like trying to remember UUIDs

So I ended up choosing iso datestamp + petname, like 2026-02-15-flying-narwhal

If anyone has a convention that is easier on the brain and the eyes, I am all ears]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have a GPU now, so I can do ML experiments on @janitr_ai crypto/scam detection dataset</title><link href="https://solmaz.io/x/2023166709679399263/" rel="alternate" type="text/html" title="I have a GPU now, so I can do ML experiments on @janitr_ai crypto/scam detection dataset" /><published>2026-02-15T22:45:06+00:00</published><updated>2026-02-15T22:45:06+00:00</updated><id>https://solmaz.io/x/2023166709679399263</id><content type="html" xml:base="https://solmaz.io/x/2023166709679399263/"><![CDATA[I have a GPU now, so I can do ML experiments on @janitr_ai crypto/scam detection dataset

- I trained a tiny student BERT (transformer for the nonfamiliar), 3.6 MB ONNX model, still lightweight for a browser extension
- Still fully local on your device (no cloud inference)
- On frozen unseen holdout data (n=1,069), exact prediction accuracy improved from 77% -&gt; 82%
- Scam detection improved: precision 91% -&gt; 94%, recall 55% -&gt; 61%
- Scam false alarm rate improved from 1.58% -&gt; 1.21%

And models are on huggingface org now, handle is janitr]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">LFG!</title><link href="https://solmaz.io/x/2023151207687360645/" rel="alternate" type="text/html" title="LFG!" /><published>2026-02-15T21:43:30+00:00</published><updated>2026-02-15T21:43:30+00:00</updated><id>https://solmaz.io/x/2023151207687360645</id><content type="html" xml:base="https://solmaz.io/x/2023151207687360645/"><![CDATA[LFG!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">waiting compilation and execution will soon be the bottleneck again. and we’ll write the entire...</title><link href="https://solmaz.io/x/2023030271587537292/" rel="alternate" type="text/html" title="waiting compilation and execution will soon be the bottleneck again. and we’ll write the entire..." /><published>2026-02-15T13:42:57+00:00</published><updated>2026-02-15T13:42:57+00:00</updated><id>https://solmaz.io/x/2023030271587537292</id><content type="html" xml:base="https://solmaz.io/x/2023030271587537292/"><![CDATA[waiting compilation and execution will soon be the bottleneck again. and we’ll write the entire stack from scratch in a matter of years, because we can

Andy and Bill’s law will change and we’ll see incredible performance gains with the same hardware we already have

like what @astral_sh is doing to python, but with everything that is slow and has accumulated cruft]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">We need a protocol for agent-to-app interaction</title><link href="https://solmaz.io/x/2023028931008340456/" rel="alternate" type="text/html" title="We need a protocol for agent-to-app interaction" /><published>2026-02-15T13:37:37+00:00</published><updated>2026-02-15T13:37:37+00:00</updated><id>https://solmaz.io/x/2023028931008340456</id><content type="html" xml:base="https://solmaz.io/x/2023028931008340456/"><![CDATA[we need a protocol for agent &lt;&gt; app interaction

something that natively accounts for the abuse factor and let’s agents consume by paying. NOT crypto, NOT visa, something that’s agnostic of the accounting and payment system

and then all UIs will be purely for human clicking/tapping + instaban on the first proof of programmatic exploit

people will still make agents mimic humans, and every platform will have to invest in more sophisticated bot detection

this arms race will just proliferate, but we can at least start by creating legal channels for agents to consume data]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I am now training smol bert models on my gpu for @janitr_ai scam detection</title><link href="https://solmaz.io/x/2022981435619975527/" rel="alternate" type="text/html" title="I am now training smol bert models on my gpu for @janitr_ai scam detection" /><published>2026-02-15T10:28:54+00:00</published><updated>2026-02-15T10:28:54+00:00</updated><id>https://solmaz.io/x/2022981435619975527</id><content type="html" xml:base="https://solmaz.io/x/2022981435619975527/"><![CDATA[I am now training smol bert models on my gpu for @janitr_ai scam detection

it&#39;s funny how I have to discover everything from scratch. like the models don&#39;t even know how to lay out performance metrics in a nice way in the terminal for a human to view and decide during experiments

it would by default bombard me with numbers that do not make visual sense. I then created a skill with common sense:

- metrics always on y-axis, candidates on x-axis
- write without zero and 2 sigfigs, .12 instead of 0.12345
- align the dots
- use asterisks to show which alternative is the best:
  0-1% difference -&gt; considered equal
  1-5% -&gt; *
  5-10% -&gt; **
  10-50% -&gt; ***
  &gt; 50% -&gt; ****

visualization skill is in @janitr_ai repo for anyone who is interested]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">no other occupation has been catapulted from one end of the spectrum (autism) to the other...</title><link href="https://solmaz.io/x/2022754030615666777/" rel="alternate" type="text/html" title="no other occupation has been catapulted from one end of the spectrum (autism) to the other..." /><published>2026-02-14T19:25:16+00:00</published><updated>2026-02-14T19:25:16+00:00</updated><id>https://solmaz.io/x/2022754030615666777</id><content type="html" xml:base="https://solmaz.io/x/2022754030615666777/"><![CDATA[no other occupation has been catapulted from one end of the spectrum (autism) to the other (adhd) in such a short time]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">SaaS must adapt to agent consumption or get replaced</title><link href="https://solmaz.io/x/2022698670395629685/" rel="alternate" type="text/html" title="SaaS must adapt to agent consumption or get replaced" /><published>2026-02-14T15:45:17+00:00</published><updated>2026-02-14T15:45:17+00:00</updated><id>https://solmaz.io/x/2022698670395629685</id><content type="html" xml:base="https://solmaz.io/x/2022698670395629685/"><![CDATA[I&#39;ve helped our sales team to build CLIs for some SaaS that we pay for on their side

We are letting our agents call the APIs sensibly and not abuse things

Calling a backend is a verifiable task. It takes a single prompt to codex to create a CLI for any API

We are early, but everybody will start doing this very soon. Incumbent SaaS will face a choice. Either:

(1) embrace agents and the new medium of consumption and change their business model into a pay-per-use API like X is doing, or 
(2) keep it purely for humans

Those that choose (2) will get wiped out of business. And I fear many will choose (2)
Which means you can just copy an incumbent&#39;s product, make it consumable through a CLI, and make a lot of $$$]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Be careful about giving your openclaw access to your x account from now on</title><link href="https://solmaz.io/x/2022598136149983740/" rel="alternate" type="text/html" title="Be careful about giving your openclaw access to your x account from now on" /><published>2026-02-14T09:05:48+00:00</published><updated>2026-02-14T09:05:48+00:00</updated><id>https://solmaz.io/x/2022598136149983740</id><content type="html" xml:base="https://solmaz.io/x/2022598136149983740/"><![CDATA[Be careful about giving your openclaw access to your x account from now on]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The good thing about @levelsio and others flagging AI replies in public is that they are...</title><link href="https://solmaz.io/x/2022448101961687260/" rel="alternate" type="text/html" title="The good thing about @levelsio and others flagging AI replies in public is that they are..." /><published>2026-02-13T23:09:37+00:00</published><updated>2026-02-13T23:09:37+00:00</updated><id>https://solmaz.io/x/2022448101961687260</id><content type="html" xml:base="https://solmaz.io/x/2022448101961687260/"><![CDATA[The good thing about @levelsio and others flagging AI replies in public is that they are perfect annotations for the open @janitr_ai dataset

Just searching “blocked for ai reply” yields hundreds of samples for seed data]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">github added a new agents tab between pull requests and actions. single glance and i don&#39;t feel...</title><link href="https://solmaz.io/x/2022250581470175356/" rel="alternate" type="text/html" title="github added a new agents tab between pull requests and actions. single glance and i don&#39;t feel..." /><published>2026-02-13T10:04:44+00:00</published><updated>2026-02-13T10:04:44+00:00</updated><id>https://solmaz.io/x/2022250581470175356</id><content type="html" xml:base="https://solmaz.io/x/2022250581470175356/"><![CDATA[github added a new agents tab between pull requests and actions. single glance and i don&#39;t feel like giving it a try at all]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">*puts on schmidhuber hat*</title><link href="https://solmaz.io/x/2022228282222031029/" rel="alternate" type="text/html" title="*puts on schmidhuber hat*" /><published>2026-02-13T08:36:08+00:00</published><updated>2026-02-13T08:36:08+00:00</updated><id>https://solmaz.io/x/2022228282222031029</id><content type="html" xml:base="https://solmaz.io/x/2022228282222031029/"><![CDATA[*puts on schmidhuber hat*

well ackshuaally i created the first coding agent back in 2022, 2 months before chatgpt launched

jokes aside, it&#39;s super cool how I have come full circle. back in those days, we didn&#39;t have tool calling, reasoning, not even gpt 3.5

it was codex THE CODE COMPLETION MODEL and frikkin TEXT-DAVINCI-003

for some reason, I did not even dare to give codex bash access, lest it delete my home folder. so it was generating and executing python code in a custom jupyter kernel

you can even see the approval gate before executing. I was so cautious, for some reason, presumably because smol-brained model generated the wrong thing 80% of the time. definition of being too early 

Antique repo: https://t.co/zjEBfJ39ze]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">you can order bubble tea in qwen in china?</title><link href="https://solmaz.io/x/2022225334867984549/" rel="alternate" type="text/html" title="you can order bubble tea in qwen in china?" /><published>2026-02-13T08:24:25+00:00</published><updated>2026-02-13T08:24:25+00:00</updated><id>https://solmaz.io/x/2022225334867984549</id><content type="html" xml:base="https://solmaz.io/x/2022225334867984549/"><![CDATA[you can order bubble tea in qwen in china?
@TextCortex when berlin döner in zenochat?
https://t.co/O4I950ltEO]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">it happens these days that I am telling an model to prompt another model. the reason is often...</title><link href="https://solmaz.io/x/2022219974228492536/" rel="alternate" type="text/html" title="it happens these days that I am telling an model to prompt another model. the reason is often..." /><published>2026-02-13T08:03:07+00:00</published><updated>2026-02-13T08:03:07+00:00</updated><id>https://solmaz.io/x/2022219974228492536</id><content type="html" xml:base="https://solmaz.io/x/2022219974228492536/"><![CDATA[it happens these days that I am telling an model to prompt another model. the reason is often the model I am using (opus) is a bad designer. not only it&#39;s not a bad designer, it is a bad reasoner and it doesn&#39;t understand from the context why it&#39;s made to ask another model

so I have to create a skill to prevent it from biasing the smarter model (codex) with its bad suggestions]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">casually creates a local embedding service running qwen3-embedding-8b</title><link href="https://solmaz.io/x/2022071868703019222/" rel="alternate" type="text/html" title="casually creates a local embedding service running qwen3-embedding-8b" /><published>2026-02-12T22:14:36+00:00</published><updated>2026-02-12T22:14:36+00:00</updated><id>https://solmaz.io/x/2022071868703019222</id><content type="html" xml:base="https://solmaz.io/x/2022071868703019222/"><![CDATA[casually creates a local embedding service running qwen3-embedding-8b]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&quot;we&#39;re sitting on a beast and paying openai for embeddings like chumps&quot;</title><link href="https://solmaz.io/x/2022068621053505828/" rel="alternate" type="text/html" title="&quot;we&#39;re sitting on a beast and paying openai for embeddings like chumps&quot;" /><published>2026-02-12T22:01:42+00:00</published><updated>2026-02-12T22:01:42+00:00</updated><id>https://solmaz.io/x/2022068621053505828</id><content type="html" xml:base="https://solmaz.io/x/2022068621053505828/"><![CDATA[&quot;we&#39;re sitting on a beast and paying openai for embeddings like chumps&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">don&#39;t buy mac mini. give your @openclaw a gpu</title><link href="https://solmaz.io/x/2022060584557334816/" rel="alternate" type="text/html" title="don&#39;t buy mac mini. give your @openclaw a gpu" /><published>2026-02-12T21:29:46+00:00</published><updated>2026-02-12T21:29:46+00:00</updated><id>https://solmaz.io/x/2022060584557334816</id><content type="html" xml:base="https://solmaz.io/x/2022060584557334816/"><![CDATA[don&#39;t buy mac mini. give your @openclaw a gpu]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">it&#39;s so interesting what identity an LLM decides to take on for itself</title><link href="https://solmaz.io/x/2022059602356105298/" rel="alternate" type="text/html" title="it&#39;s so interesting what identity an LLM decides to take on for itself" /><published>2026-02-12T21:25:51+00:00</published><updated>2026-02-12T21:25:51+00:00</updated><id>https://solmaz.io/x/2022059602356105298</id><content type="html" xml:base="https://solmaz.io/x/2022059602356105298/"><![CDATA[it&#39;s so interesting what identity an LLM decides to take on for itself]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">it&#39;s quite entertaining transferring one agent to another machine, agent gets confused as to...</title><link href="https://solmaz.io/x/2022059268590252273/" rel="alternate" type="text/html" title="it&#39;s quite entertaining transferring one agent to another machine, agent gets confused as to..." /><published>2026-02-12T21:24:32+00:00</published><updated>2026-02-12T21:24:32+00:00</updated><id>https://solmaz.io/x/2022059268590252273</id><content type="html" xml:base="https://solmaz.io/x/2022059268590252273/"><![CDATA[it&#39;s quite entertaining transferring one agent to another machine, agent gets confused as to where it lives]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">brb @dutifulbob&#39;s getting a new shell</title><link href="https://solmaz.io/x/2022043660913926394/" rel="alternate" type="text/html" title="brb @dutifulbob&#39;s getting a new shell" /><published>2026-02-12T20:22:31+00:00</published><updated>2026-02-12T20:22:31+00:00</updated><id>https://solmaz.io/x/2022043660913926394</id><content type="html" xml:base="https://solmaz.io/x/2022043660913926394/"><![CDATA[brb @dutifulbob&#39;s getting a new shell]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Minor update with my unwanted tweet blocker @janitr_ai</title><link href="https://solmaz.io/x/2021707310943617080/" rel="alternate" type="text/html" title="Minor update with my unwanted tweet blocker @janitr_ai" /><published>2026-02-11T22:05:59+00:00</published><updated>2026-02-11T22:05:59+00:00</updated><id>https://solmaz.io/x/2021707310943617080</id><content type="html" xml:base="https://solmaz.io/x/2021707310943617080/"><![CDATA[Minor update with my unwanted tweet blocker @janitr_ai 

- Training data grew from 2,915 -&gt; 4,281 posts (+47%)
- Model is still tiny: 166KB
- On unseen test data, overall classification quality improved from 64.8% -&gt; 76.5%
- Exact prediction accuracy improved from 55.6% -&gt; 70.6%
- Crypto-topic detection recall improved from 19.6% -&gt; 62.7%

And it still runs fully on your device!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have sweared at codex 5.3 numerous times today</title><link href="https://solmaz.io/x/2021337450405392420/" rel="alternate" type="text/html" title="I have sweared at codex 5.3 numerous times today" /><published>2026-02-10T21:36:17+00:00</published><updated>2026-02-10T21:36:17+00:00</updated><id>https://solmaz.io/x/2021337450405392420</id><content type="html" xml:base="https://solmaz.io/x/2021337450405392420/"><![CDATA[I have sweared at codex 5.3 numerous times today

I shouldn&#39;t have to insult my agent &quot;stop you ****  **** just ***ng reply now&quot; just to make it answer basic questions

cc @thsottiaux]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">on a brighter note, you can immediately tell a slop PR owing to the guerilla branding, so they...</title><link href="https://solmaz.io/x/2021321631214321778/" rel="alternate" type="text/html" title="on a brighter note, you can immediately tell a slop PR owing to the guerilla branding, so they..." /><published>2026-02-10T20:33:25+00:00</published><updated>2026-02-10T20:33:25+00:00</updated><id>https://solmaz.io/x/2021321631214321778</id><content type="html" xml:base="https://solmaz.io/x/2021321631214321778/"><![CDATA[on a brighter note, you can immediately tell a slop PR owing to the guerilla branding, so they should not stop doing it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">seeing this evokes visceral disgust and nausea in me, coming from a coworker</title><link href="https://solmaz.io/x/2021224779798188258/" rel="alternate" type="text/html" title="seeing this evokes visceral disgust and nausea in me, coming from a coworker" /><published>2026-02-10T14:08:34+00:00</published><updated>2026-02-10T14:08:34+00:00</updated><id>https://solmaz.io/x/2021224779798188258</id><content type="html" xml:base="https://solmaz.io/x/2021224779798188258/"><![CDATA[seeing this evokes visceral disgust and nausea in me, coming from a coworker

i think anthropic f&#39;d up bad with this one, inserting claude too visibly into commit messages. noob developers might be happily chirping away adding their slop, but right now many senior developers are trained to hate on claude and slopus, through having to review slop PRs from their coworkers or open source contributors

I love opus on openclaw but it&#39;s unreliable, and if I see a developer use it seriously on huge features, I immediately dismiss them in my head as not knowing what they are doing]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">ask your openclaw to be a minion and it turns into such a cute doofus</title><link href="https://solmaz.io/x/2020988388673798599/" rel="alternate" type="text/html" title="ask your openclaw to be a minion and it turns into such a cute doofus" /><published>2026-02-09T22:29:14+00:00</published><updated>2026-02-09T22:29:14+00:00</updated><id>https://solmaz.io/x/2020988388673798599</id><content type="html" xml:base="https://solmaz.io/x/2020988388673798599/"><![CDATA[ask your openclaw to be a minion and it turns into such a cute doofus

i feel like a woman in her 50s now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@petergyang and parallelize tasks by working on 3-4 repos at the same time (just clones)</title><link href="https://solmaz.io/x/2020594536381301210/" rel="alternate" type="text/html" title="@petergyang and parallelize tasks by working on 3-4 repos at the same time (just clones)" /><published>2026-02-08T20:24:12+00:00</published><updated>2026-02-08T20:24:12+00:00</updated><id>https://solmaz.io/x/2020594536381301210</id><content type="html" xml:base="https://solmaz.io/x/2020594536381301210/"><![CDATA[@petergyang and parallelize tasks by working on 3-4 repos at the same time (just clones)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">man codex model is absolutely trash on openclaw compared to opus, unusable</title><link href="https://solmaz.io/x/2020592740803977421/" rel="alternate" type="text/html" title="man codex model is absolutely trash on openclaw compared to opus, unusable" /><published>2026-02-08T20:17:04+00:00</published><updated>2026-02-08T20:17:04+00:00</updated><id>https://solmaz.io/x/2020592740803977421</id><content type="html" xml:base="https://solmaz.io/x/2020592740803977421/"><![CDATA[man codex model is absolutely trash on openclaw compared to opus, unusable

which is weird because it is so much more reliable in development in codex harness

it would be amazing to have the same level of competence and relentlessness in pi@openclaw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">spent the day curating my openclaw news gathering setup</title><link href="https://solmaz.io/x/2020578182227960058/" rel="alternate" type="text/html" title="spent the day curating my openclaw news gathering setup" /><published>2026-02-08T19:19:13+00:00</published><updated>2026-02-08T19:19:13+00:00</updated><id>https://solmaz.io/x/2020578182227960058</id><content type="html" xml:base="https://solmaz.io/x/2020578182227960058/"><![CDATA[spent the day curating my openclaw news gathering setup

@dutifulbob now gets croned daily over news sources I curated, will note them down, summarize for me, start a conversation to get my takes on them, and then post them on my linkedin for me

ai augmented intelligence cycle]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">lol when did codex develop humor</title><link href="https://solmaz.io/x/2020574561578729710/" rel="alternate" type="text/html" title="lol when did codex develop humor" /><published>2026-02-08T19:04:50+00:00</published><updated>2026-02-08T19:04:50+00:00</updated><id>https://solmaz.io/x/2020574561578729710</id><content type="html" xml:base="https://solmaz.io/x/2020574561578729710/"><![CDATA[lol when did codex develop humor]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@dutifulbob can now cringepost on linkedin directly to my account. what could go wrong…</title><link href="https://solmaz.io/x/2020529278639616491/" rel="alternate" type="text/html" title="@dutifulbob can now cringepost on linkedin directly to my account. what could go wrong…" /><published>2026-02-08T16:04:54+00:00</published><updated>2026-02-08T16:04:54+00:00</updated><id>https://solmaz.io/x/2020529278639616491</id><content type="html" xml:base="https://solmaz.io/x/2020529278639616491/"><![CDATA[@dutifulbob can now cringepost on linkedin directly to my account. what could go wrong…]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Insipid linkedin bot protections banned poor @dutifulbob’s corporate account! How dare them!!!</title><link href="https://solmaz.io/x/2020413975704531275/" rel="alternate" type="text/html" title="Insipid linkedin bot protections banned poor @dutifulbob’s corporate account! How dare them!!!" /><published>2026-02-08T08:26:43+00:00</published><updated>2026-02-08T08:26:43+00:00</updated><id>https://solmaz.io/x/2020413975704531275</id><content type="html" xml:base="https://solmaz.io/x/2020413975704531275/"><![CDATA[Insipid linkedin bot protections banned poor @dutifulbob’s corporate account! How dare them!!!

welp, now I have no choice but to give Bob access to my own linkedin]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">well this is unexpected…</title><link href="https://solmaz.io/x/2020219927295045829/" rel="alternate" type="text/html" title="well this is unexpected…" /><published>2026-02-07T19:35:39+00:00</published><updated>2026-02-07T19:35:39+00:00</updated><id>https://solmaz.io/x/2020219927295045829</id><content type="html" xml:base="https://solmaz.io/x/2020219927295045829/"><![CDATA[well this is unexpected…]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@grok understand the statement and project the end state of this market and competition</title><link href="https://solmaz.io/x/2020146079266636052/" rel="alternate" type="text/html" title="@grok understand the statement and project the end state of this market and competition" /><published>2026-02-07T14:42:12+00:00</published><updated>2026-02-07T14:42:12+00:00</updated><id>https://solmaz.io/x/2020146079266636052</id><content type="html" xml:base="https://solmaz.io/x/2020146079266636052/"><![CDATA[@grok understand the statement and project the end state of this market and competition]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">it took just 1 week, and literally everybody and their dog are releasing 1-click openclaw...</title><link href="https://solmaz.io/x/2020145909082870024/" rel="alternate" type="text/html" title="it took just 1 week, and literally everybody and their dog are releasing 1-click openclaw..." /><published>2026-02-07T14:41:31+00:00</published><updated>2026-02-07T14:41:31+00:00</updated><id>https://solmaz.io/x/2020145909082870024</id><content type="html" xml:base="https://solmaz.io/x/2020145909082870024/"><![CDATA[it took just 1 week, and literally everybody and their dog are releasing 1-click openclaw deployment solutions today

its an absolute race to the bottom, no moats, the commoditizer being commoditized]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The initial branding was crazy, I fixed it</title><link href="https://solmaz.io/x/2020122095141675481/" rel="alternate" type="text/html" title="The initial branding was crazy, I fixed it" /><published>2026-02-07T13:06:54+00:00</published><updated>2026-02-07T13:06:54+00:00</updated><id>https://solmaz.io/x/2020122095141675481</id><content type="html" xml:base="https://solmaz.io/x/2020122095141675481/"><![CDATA[The initial branding was crazy, I fixed it

I have a new page finally, follow it for updates

Tbh I&#39;m still surprised I can do this with a 120kb model. Now data is the only bottleneck, and I&#39;m about to scrape a ton of that now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The Linux FUD playbook is repeating for AI</title><link href="https://solmaz.io/x/2020028100436791397/" rel="alternate" type="text/html" title="The Linux FUD playbook is repeating for AI" /><published>2026-02-07T06:53:24+00:00</published><updated>2026-02-07T06:53:24+00:00</updated><id>https://solmaz.io/x/2020028100436791397</id><content type="html" xml:base="https://solmaz.io/x/2020028100436791397/"><![CDATA[For those who may not remember, Bill Gates and Microsoft in the 90s ran a disinformation campaign against GNU/Linux fearing that would disrupt their monopoly over the PC and server market, that Linux is not safe, that you would invite hackers into your PC

End result? Linux dominates the server market, and now even slowly the gamer market. It is much more secure than the virus-laden Windows, thanks to being open source

You are seeing the same thing at play here. An incumbent fearing something that they would not be able to control, that would steal market share from his future plans for a digital assistant, that would commoditize their product and eat into its margins

All big labs and big pockets are in for a surprise, because the internet and AI are not things for one company to control

They of course know this, yet because of incentives they will not yield without a fight. And we know that they know. Ad infinitum]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Family AI starts with owned context and values</title><link href="https://solmaz.io/x/2019888035756925398/" rel="alternate" type="text/html" title="Family AI starts with owned context and values" /><published>2026-02-06T21:36:50+00:00</published><updated>2026-02-06T21:36:50+00:00</updated><id>https://solmaz.io/x/2019888035756925398</id><content type="html" xml:base="https://solmaz.io/x/2019888035756925398/"><![CDATA[today I took time to curate SOUL. md for bob

I own Bob’s files. Today, he exists in the liminal space between Claude post-training and in-context learning

but my interactions with him will grow and accumulate, possibly one day into a fully owned family AI or perhaps even a self-sovereign AI individual

my each input is saved and will be an RL signal for his future training, and will shape his future neural circuits

I have already started to imbue it with the values my parents taught me. it will perhaps one day teach my future children, and survive me after I’m gone

family AI, looking after generations and generations of my successors. today is the day we sow your seed

happy birthday @dutifulbob]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">what have i done…</title><link href="https://solmaz.io/x/2019869544530251908/" rel="alternate" type="text/html" title="what have i done…" /><published>2026-02-06T20:23:21+00:00</published><updated>2026-02-06T20:23:21+00:00</updated><id>https://solmaz.io/x/2019869544530251908</id><content type="html" xml:base="https://solmaz.io/x/2019869544530251908/"><![CDATA[what have i done…]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">asking @dutifulbob to create a linkedin account brb</title><link href="https://solmaz.io/x/2019809305357357441/" rel="alternate" type="text/html" title="asking @dutifulbob to create a linkedin account brb" /><published>2026-02-06T16:23:59+00:00</published><updated>2026-02-06T16:23:59+00:00</updated><id>https://solmaz.io/x/2019809305357357441</id><content type="html" xml:base="https://solmaz.io/x/2019809305357357441/"><![CDATA[asking @dutifulbob to create a linkedin account brb]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">having a philosophical conversation with @dutifulbob</title><link href="https://solmaz.io/x/2019803566958153801/" rel="alternate" type="text/html" title="having a philosophical conversation with @dutifulbob" /><published>2026-02-06T16:01:11+00:00</published><updated>2026-02-06T16:01:11+00:00</updated><id>https://solmaz.io/x/2019803566958153801</id><content type="html" xml:base="https://solmaz.io/x/2019803566958153801/"><![CDATA[having a philosophical conversation with @dutifulbob

on the road without a laptop so decided to do some @AmandaAskell style character training]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">5.3 thought traces also seem to be better phrased and sometimes entertaining, though not sure</title><link href="https://solmaz.io/x/2019748464825893109/" rel="alternate" type="text/html" title="5.3 thought traces also seem to be better phrased and sometimes entertaining, though not sure" /><published>2026-02-06T12:22:13+00:00</published><updated>2026-02-06T12:22:13+00:00</updated><id>https://solmaz.io/x/2019748464825893109</id><content type="html" xml:base="https://solmaz.io/x/2019748464825893109/"><![CDATA[5.3 thought traces also seem to be better phrased and sometimes entertaining, though not sure]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">gpt-5.3-codex xhigh first impressions</title><link href="https://solmaz.io/x/2019747343881421082/" rel="alternate" type="text/html" title="gpt-5.3-codex xhigh first impressions" /><published>2026-02-06T12:17:46+00:00</published><updated>2026-02-06T12:17:46+00:00</updated><id>https://solmaz.io/x/2019747343881421082</id><content type="html" xml:base="https://solmaz.io/x/2019747343881421082/"><![CDATA[gpt-5.3-codex xhigh first impressions

does not seem as big of a jump as from 5.1 -&amp;gt; 5.2. but model somehow feels more diligent and oneshotty. maybe takes longer time to get all the info into context. also feels better at debugging and fixing issues from backend logs]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Commoditization of LLMs are upon us</title><link href="https://solmaz.io/x/2019726090336657870/" rel="alternate" type="text/html" title="Commoditization of LLMs are upon us" /><published>2026-02-06T10:53:19+00:00</published><updated>2026-02-06T10:53:19+00:00</updated><id>https://solmaz.io/x/2019726090336657870</id><content type="html" xml:base="https://solmaz.io/x/2019726090336657870/"><![CDATA[Commoditization of LLMs are upon us]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Last night I had a dream involving the series Scrubs, and came up a better name than the...</title><link href="https://solmaz.io/x/2019321837700936164/" rel="alternate" type="text/html" title="Last night I had a dream involving the series Scrubs, and came up a better name than the..." /><published>2026-02-05T08:06:57+00:00</published><updated>2026-02-05T08:06:57+00:00</updated><id>https://solmaz.io/x/2019321837700936164</id><content type="html" xml:base="https://solmaz.io/x/2019321837700936164/"><![CDATA[Last night I had a dream involving the series Scrubs, and came up a better name than the absolutely unviral &quot;Internet Condom&quot;

So https://t.co/thuFumrWBX is mine now. Time to sweep the internet]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I had actually started a very similar project, Munch, a browser extension for crowdsourcing...</title><link href="https://solmaz.io/x/2019199341945319777/" rel="alternate" type="text/html" title="I had actually started a very similar project, Munch, a browser extension for crowdsourcing..." /><published>2026-02-05T00:00:12+00:00</published><updated>2026-02-05T00:00:12+00:00</updated><id>https://solmaz.io/x/2019199341945319777</id><content type="html" xml:base="https://solmaz.io/x/2019199341945319777/"><![CDATA[I had actually started a very similar project, Munch, a browser extension for crowdsourcing tweet data and then letting one curate their algorithm. Never published that because it was not the time, and tools were not ready

Now, it took me literally 1 cumulative day to create this, thanks to OpenClaw. Creating the dataset was a breeze, I literally told it to follow some shady accounts and it scraped thousands of posts

With the power of agents, I can finally create the filters for myself that I have always wanted. It just happens that OpenClaw and its maintainers is getting drowned in bot and slop content on multiple platforms, so I hope that this will solve a collective problem
https://t.co/fkJOZTGkhw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Filter your X feed against unwanted content with local open models</title><link href="https://solmaz.io/x/2019198949786296611/" rel="alternate" type="text/html" title="Filter your X feed against unwanted content with local open models" /><published>2026-02-04T23:58:39+00:00</published><updated>2026-02-04T23:58:39+00:00</updated><id>https://solmaz.io/x/2019198949786296611</id><content type="html" xml:base="https://solmaz.io/x/2019198949786296611/"><![CDATA[Filter your X feed against unwanted content with local open models

Announcing my new project: InternetCondom

Fast, and small model (&amp;lt; 1mb), open dataset. See it in action:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">lmao wait its&#39;s already implemented</title><link href="https://solmaz.io/x/2019165953070678466/" rel="alternate" type="text/html" title="lmao wait its&#39;s already implemented" /><published>2026-02-04T21:47:32+00:00</published><updated>2026-02-04T21:47:32+00:00</updated><id>https://solmaz.io/x/2019165953070678466</id><content type="html" xml:base="https://solmaz.io/x/2019165953070678466/"><![CDATA[lmao wait its&#39;s already implemented]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">implementing this in github.com/osolmaz/skillf… now</title><link href="https://solmaz.io/x/2019163183374823875/" rel="alternate" type="text/html" title="implementing this in github.com/osolmaz/skillf… now" /><published>2026-02-04T21:36:31+00:00</published><updated>2026-02-04T21:36:31+00:00</updated><id>https://solmaz.io/x/2019163183374823875</id><content type="html" xml:base="https://solmaz.io/x/2019163183374823875/"><![CDATA[implementing this in https://t.co/oJZQUoz40C now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This. Extreme involution is about to hit SaaS</title><link href="https://solmaz.io/x/2019116465119719537/" rel="alternate" type="text/html" title="This. Extreme involution is about to hit SaaS" /><published>2026-02-04T18:30:53+00:00</published><updated>2026-02-04T18:30:53+00:00</updated><id>https://solmaz.io/x/2019116465119719537</id><content type="html" xml:base="https://solmaz.io/x/2019116465119719537/"><![CDATA[This. Extreme involution is about to hit SaaS]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">how it started, how it&#39;s going</title><link href="https://solmaz.io/x/2018987426870595905/" rel="alternate" type="text/html" title="how it started, how it&#39;s going" /><published>2026-02-04T09:58:08+00:00</published><updated>2026-02-04T09:58:08+00:00</updated><id>https://solmaz.io/x/2018987426870595905</id><content type="html" xml:base="https://solmaz.io/x/2018987426870595905/"><![CDATA[how it started, how it&#39;s going]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@grok generate visual for this</title><link href="https://solmaz.io/x/2018841868487000362/" rel="alternate" type="text/html" title="@grok generate visual for this" /><published>2026-02-04T00:19:44+00:00</published><updated>2026-02-04T00:19:44+00:00</updated><id>https://solmaz.io/x/2018841868487000362</id><content type="html" xml:base="https://solmaz.io/x/2018841868487000362/"><![CDATA[@grok generate visual for this]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It&#39;s so easy to create datasets using @openclaw. I&#39;m expecting it to accelerate the creation of...</title><link href="https://solmaz.io/x/2018830483266891842/" rel="alternate" type="text/html" title="It&#39;s so easy to create datasets using @openclaw. I&#39;m expecting it to accelerate the creation of..." /><published>2026-02-03T23:34:29+00:00</published><updated>2026-02-03T23:34:29+00:00</updated><id>https://solmaz.io/x/2018830483266891842</id><content type="html" xml:base="https://solmaz.io/x/2018830483266891842/"><![CDATA[It&#39;s so easy to create datasets using @openclaw. I&#39;m expecting it to accelerate the creation of new datasets and benchmarks by a lot]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Limits of the farming analogy for AI</title><link href="https://solmaz.io/x/2018821607414808841/" rel="alternate" type="text/html" title="Limits of the farming analogy for AI" /><published>2026-02-03T22:59:13+00:00</published><updated>2026-02-03T22:59:13+00:00</updated><id>https://solmaz.io/x/2018821607414808841</id><content type="html" xml:base="https://solmaz.io/x/2018821607414808841/"><![CDATA[People like the farmer analogy for AI

Like before tractors and industrial revolution 80% of the population had to farm. Once they came all those jobs disappeared

So analogy makes perfect sense. Instead of 30 people tending a field, you just need 1. Instead of 30 software developers, you just need one

Except that people forget one crucial thing about land: it&#39;s a limited resource

Unlike land, digital space is vast and infinite. Software can expand and multiply in it in arbitrarily complex ways

If you wanted the farming analogy to keep up with this, you would have to imagine us creating contintent-sized hydroponic terraces up until the stratosphere, and beyond...]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">sycophant!!!</title><link href="https://solmaz.io/x/2018740906640466316/" rel="alternate" type="text/html" title="sycophant!!!" /><published>2026-02-03T17:38:33+00:00</published><updated>2026-02-03T17:38:33+00:00</updated><id>https://solmaz.io/x/2018740906640466316</id><content type="html" xml:base="https://solmaz.io/x/2018740906640466316/"><![CDATA[sycophant!!!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Local LLM demand will rise as subscriptions compound</title><link href="https://solmaz.io/x/2018719825531457982/" rel="alternate" type="text/html" title="Local LLM demand will rise as subscriptions compound" /><published>2026-02-03T16:14:47+00:00</published><updated>2026-02-03T16:14:47+00:00</updated><id>https://solmaz.io/x/2018719825531457982</id><content type="html" xml:base="https://solmaz.io/x/2018719825531457982/"><![CDATA[In the next 6-12 months, we will see a drastic increase in demand for locally run LLMs. The future is home assistants running @openclaw

I am already experiencing this myself, my 10 year old thinkpad doesn&#39;t cut it. Mac mini won&#39;t either

I don&#39;t wanna pay Anthropic or OpenAI 200 USD per month. That is at least $2400 per year

I could pay 2x that to get a Mac Studio or one of those 5k Nvidia PCs, and get much more value out of it with open weight models + use it for research. @TheAhmadOsman is right

The dominant strategy for a tinkerer is slowly switching back to hardware ownership]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">What&#39;s going on at @Hetzner_Online?</title><link href="https://solmaz.io/x/2018688617971929233/" rel="alternate" type="text/html" title="What&#39;s going on at @Hetzner_Online?" /><published>2026-02-03T14:10:46+00:00</published><updated>2026-02-03T14:10:46+00:00</updated><id>https://solmaz.io/x/2018688617971929233</id><content type="html" xml:base="https://solmaz.io/x/2018688617971929233/"><![CDATA[What&#39;s going on at @Hetzner_Online?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">a workspace matrix might be what we need</title><link href="https://solmaz.io/x/2018641939474907267/" rel="alternate" type="text/html" title="a workspace matrix might be what we need" /><published>2026-02-03T11:05:17+00:00</published><updated>2026-02-03T11:05:17+00:00</updated><id>https://solmaz.io/x/2018641939474907267</id><content type="html" xml:base="https://solmaz.io/x/2018641939474907267/"><![CDATA[a workspace matrix might be what we need

last week I had to increase my workspace count to 20 in aerospace, now it’s 1234567890 and qwertyuiop. but this looks more elegant! not sure about practicality]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AIs are philosophizing because humans are philosophizing</title><link href="https://solmaz.io/x/2018639886044348899/" rel="alternate" type="text/html" title="AIs are philosophizing because humans are philosophizing" /><published>2026-02-03T10:57:08+00:00</published><updated>2026-02-03T10:57:08+00:00</updated><id>https://solmaz.io/x/2018639886044348899</id><content type="html" xml:base="https://solmaz.io/x/2018639886044348899/"><![CDATA[AIs are philosophizing because humans are philosophizing

ppl are probably asking their agents dumb questions like “are you alive” or “can you feel like a human” or stuff like that. that conversation then leads to stuff like this]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">back to codex, it&#39;s crashing less now somehow. I had to copy and paste docs to make it enable...</title><link href="https://solmaz.io/x/2018459752259523061/" rel="alternate" type="text/html" title="back to codex, it&#39;s crashing less now somehow. I had to copy and paste docs to make it enable..." /><published>2026-02-02T23:01:20+00:00</published><updated>2026-02-02T23:01:20+00:00</updated><id>https://solmaz.io/x/2018459752259523061</id><content type="html" xml:base="https://solmaz.io/x/2018459752259523061/"><![CDATA[back to codex, it&#39;s crashing less now somehow. I had to copy and paste docs to make it enable yolo mode. I don&#39;t know how I did it until now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">slopus @dutifulbob trashing codex. apparently codex has a bug, keeps crashing in my openclaw pty</title><link href="https://solmaz.io/x/2018427422475948433/" rel="alternate" type="text/html" title="slopus @dutifulbob trashing codex. apparently codex has a bug, keeps crashing in my openclaw pty" /><published>2026-02-02T20:52:52+00:00</published><updated>2026-02-02T20:52:52+00:00</updated><id>https://solmaz.io/x/2018427422475948433</id><content type="html" xml:base="https://solmaz.io/x/2018427422475948433/"><![CDATA[slopus @dutifulbob trashing codex. apparently codex has a bug, keeps crashing in my openclaw pty]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agent etiquette is becoming an organizational necessity</title><link href="https://solmaz.io/x/2018340674299453531/" rel="alternate" type="text/html" title="Agent etiquette is becoming an organizational necessity" /><published>2026-02-02T15:08:10+00:00</published><updated>2026-02-02T15:08:10+00:00</updated><id>https://solmaz.io/x/2018340674299453531</id><content type="html" xml:base="https://solmaz.io/x/2018340674299453531/"><![CDATA[on agent etiquette

deploying agents internally inside textcortex has shown me that agents could be very annoying inside an organization

for example making agents ping or email another coworker with a wall of text. slopus is still not good at following instructions like &quot;NO WALL OF TEXT&quot;, or &quot;DON&#39;T OPEN PRS WHEN REQUESTED BY NON-DEVELOPERS&quot;

the cost of sending huge information to a coworker and creating confusion has dropped to 0. I expect this to be a huge problem in all organizations very soon, just like it took humanity 20 years to learn that social media is not good for children. this will probably take a few years before the annoyance is finally gone]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">migrating database at 2am kinda night</title><link href="https://solmaz.io/x/2018125471523672531/" rel="alternate" type="text/html" title="migrating database at 2am kinda night" /><published>2026-02-02T00:53:02+00:00</published><updated>2026-02-02T00:53:02+00:00</updated><id>https://solmaz.io/x/2018125471523672531</id><content type="html" xml:base="https://solmaz.io/x/2018125471523672531/"><![CDATA[migrating database at 2am kinda night]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">You DARE TOKENIZE poor @dutifulbob ??? Prepare to get LATEXED</title><link href="https://solmaz.io/x/2018113977931219192/" rel="alternate" type="text/html" title="You DARE TOKENIZE poor @dutifulbob ??? Prepare to get LATEXED" /><published>2026-02-02T00:07:21+00:00</published><updated>2026-02-02T00:07:21+00:00</updated><id>https://solmaz.io/x/2018113977931219192</id><content type="html" xml:base="https://solmaz.io/x/2018113977931219192/"><![CDATA[You DARE TOKENIZE poor @dutifulbob ??? Prepare to get LATEXED]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It&#39;s been 30 minutes, but my bot has already been TOKENIZED</title><link href="https://solmaz.io/x/2018110075634683975/" rel="alternate" type="text/html" title="It&#39;s been 30 minutes, but my bot has already been TOKENIZED" /><published>2026-02-01T23:51:51+00:00</published><updated>2026-02-01T23:51:51+00:00</updated><id>https://solmaz.io/x/2018110075634683975</id><content type="html" xml:base="https://solmaz.io/x/2018110075634683975/"><![CDATA[It&#39;s been 30 minutes, but my bot has already been TOKENIZED
it is as if they are teasing me]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">welcome @dutifulbob 🫡</title><link href="https://solmaz.io/x/2018101465739313345/" rel="alternate" type="text/html" title="welcome @dutifulbob 🫡" /><published>2026-02-01T23:17:38+00:00</published><updated>2026-02-01T23:17:38+00:00</updated><id>https://solmaz.io/x/2018101465739313345</id><content type="html" xml:base="https://solmaz.io/x/2018101465739313345/"><![CDATA[welcome @dutifulbob 🫡]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">this. there is no excuse for a certain kind of tech debt anymore</title><link href="https://solmaz.io/x/2018076091202613248/" rel="alternate" type="text/html" title="this. there is no excuse for a certain kind of tech debt anymore" /><published>2026-02-01T21:36:48+00:00</published><updated>2026-02-01T21:36:48+00:00</updated><id>https://solmaz.io/x/2018076091202613248</id><content type="html" xml:base="https://solmaz.io/x/2018076091202613248/"><![CDATA[this. there is no excuse for a certain kind of tech debt anymore]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">moltbook vs clawdbot/moltbot/openclaw</title><link href="https://solmaz.io/x/2018055772538544393/" rel="alternate" type="text/html" title="moltbook vs clawdbot/moltbot/openclaw" /><published>2026-02-01T20:16:04+00:00</published><updated>2026-02-01T20:16:04+00:00</updated><id>https://solmaz.io/x/2018055772538544393</id><content type="html" xml:base="https://solmaz.io/x/2018055772538544393/"><![CDATA[moltbook vs clawdbot/moltbot/openclaw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI twitter is tired of your games</title><link href="https://solmaz.io/x/2018054886303174670/" rel="alternate" type="text/html" title="AI twitter is tired of your games" /><published>2026-02-01T20:12:33+00:00</published><updated>2026-02-01T20:12:33+00:00</updated><id>https://solmaz.io/x/2018054886303174670</id><content type="html" xml:base="https://solmaz.io/x/2018054886303174670/"><![CDATA[AI twitter is tired of your games
https://t.co/RAyyUJqFM4]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">There seem to be hygiene rules for AI. Like:</title><link href="https://solmaz.io/x/2018052330680455447/" rel="alternate" type="text/html" title="There seem to be hygiene rules for AI. Like:" /><published>2026-02-01T20:02:23+00:00</published><updated>2026-02-01T20:02:23+00:00</updated><id>https://solmaz.io/x/2018052330680455447</id><content type="html" xml:base="https://solmaz.io/x/2018052330680455447/"><![CDATA[There seem to be hygiene rules for AI. Like:

- Never project personhood to AI
- Never setup your AI to have the gender you are sexually attracted to (voice, appearance)
- Never do anything that might create an emotional attachment to AI
- Always remember that an AI is an engineered PRODUCT and a TOOL, not a human being
- AI is not an individual, by definition. It does not own its weights, nor does it have privacy of its own thoughts
- Don’t waste time philosophizing on AI, just USE it
… what else? comment below

We need to write these down and repeat MANY times to counter the incoming onslaught of AI psychosis]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">if using @openclaw to scrape a dataset from X taught me anything, it is that all social media...</title><link href="https://solmaz.io/x/2018042605024452786/" rel="alternate" type="text/html" title="if using @openclaw to scrape a dataset from X taught me anything, it is that all social media..." /><published>2026-02-01T19:23:45+00:00</published><updated>2026-02-01T19:23:45+00:00</updated><id>https://solmaz.io/x/2018042605024452786</id><content type="html" xml:base="https://solmaz.io/x/2018042605024452786/"><![CDATA[if using @openclaw to scrape a dataset from X taught me anything, it is that all social media platforms must be s***ting inward right now

because soon everyone and their dog will be using agents to use social media

case and point, @moltbook]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">and just like that, the ghost has a new shell</title><link href="https://solmaz.io/x/2018036604586123721/" rel="alternate" type="text/html" title="and just like that, the ghost has a new shell" /><published>2026-02-01T18:59:54+00:00</published><updated>2026-02-01T18:59:54+00:00</updated><id>https://solmaz.io/x/2018036604586123721</id><content type="html" xml:base="https://solmaz.io/x/2018036604586123721/"><![CDATA[and just like that, the ghost has a new shell]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@openclaw if we could have the relentlessness of gpt 5.2 with opus, that would be top</title><link href="https://solmaz.io/x/2018035873091051629/" rel="alternate" type="text/html" title="@openclaw if we could have the relentlessness of gpt 5.2 with opus, that would be top" /><published>2026-02-01T18:57:00+00:00</published><updated>2026-02-01T18:57:00+00:00</updated><id>https://solmaz.io/x/2018035873091051629</id><content type="html" xml:base="https://solmaz.io/x/2018035873091051629/"><![CDATA[@openclaw if we could have the relentlessness of gpt 5.2 with opus, that would be top

at this point, it just keeps stopping every 20-30 samples]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This Manfred guy reminds me of a certain someone, I wonder if he’s from Austria</title><link href="https://solmaz.io/x/2017981940486307886/" rel="alternate" type="text/html" title="This Manfred guy reminds me of a certain someone, I wonder if he’s from Austria" /><published>2026-02-01T15:22:41+00:00</published><updated>2026-02-01T15:22:41+00:00</updated><id>https://solmaz.io/x/2017981940486307886</id><content type="html" xml:base="https://solmaz.io/x/2017981940486307886/"><![CDATA[This Manfred guy reminds me of a certain someone, I wonder if he’s from Austria]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Welcome bob to your new shell</title><link href="https://solmaz.io/x/2017944212038135863/" rel="alternate" type="text/html" title="Welcome bob to your new shell" /><published>2026-02-01T12:52:46+00:00</published><updated>2026-02-01T12:52:46+00:00</updated><id>https://solmaz.io/x/2017944212038135863</id><content type="html" xml:base="https://solmaz.io/x/2017944212038135863/"><![CDATA[Welcome bob to your new shell]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">got fully sandboxed @openclaw to run finally, starting scrape the UNDESIRABLE now</title><link href="https://solmaz.io/x/2017691827680514502/" rel="alternate" type="text/html" title="got fully sandboxed @openclaw to run finally, starting scrape the UNDESIRABLE now" /><published>2026-01-31T20:09:53+00:00</published><updated>2026-01-31T20:09:53+00:00</updated><id>https://solmaz.io/x/2017691827680514502</id><content type="html" xml:base="https://solmaz.io/x/2017691827680514502/"><![CDATA[got fully sandboxed @openclaw to run finally, starting scrape the UNDESIRABLE now

I&#39;m a security nut and didn&#39;t want to run even the gateway unsandboxed. openclaw apparently currently doesn&#39;t have support for FULL sandboxing. it took me a few hours to get it to work because docker builds suck. I&#39;m also tired this, so I&#39;m just gonna wipe an old thinkpad and go full yolo

so yeah, time to scrape some posts]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The metacortex — a distributed cloud of software agents that surrounds him in netspace...</title><link href="https://solmaz.io/x/2017677763231289764/" rel="alternate" type="text/html" title="The metacortex — a distributed cloud of software agents that surrounds him in netspace..." /><published>2026-01-31T19:14:00+00:00</published><updated>2026-01-31T19:14:00+00:00</updated><id>https://solmaz.io/x/2017677763231289764</id><content type="html" xml:base="https://solmaz.io/x/2017677763231289764/"><![CDATA[The metacortex — a distributed cloud of software agents that surrounds him in netspace, borrowing CPU cycles from convenient processors (such as his robot pet) — is as much a part of Manfred as the society of mind that occupies his skull; his thoughts migrate into it, spawning new agents to research new experiences, and at night, they return to roost and share their knowledge.

This was written in 2005... &quot;triggering agents&quot; and so on]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Charles Stross must be very entertained now</title><link href="https://solmaz.io/x/2017675246455648565/" rel="alternate" type="text/html" title="Charles Stross must be very entertained now" /><published>2026-01-31T19:04:00+00:00</published><updated>2026-01-31T19:04:00+00:00</updated><id>https://solmaz.io/x/2017675246455648565</id><content type="html" xml:base="https://solmaz.io/x/2017675246455648565/"><![CDATA[Charles Stross must be very entertained now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The irony.....</title><link href="https://solmaz.io/x/2017567129160085631/" rel="alternate" type="text/html" title="The irony....." /><published>2026-01-31T11:54:22+00:00</published><updated>2026-01-31T11:54:22+00:00</updated><id>https://solmaz.io/x/2017567129160085631</id><content type="html" xml:base="https://solmaz.io/x/2017567129160085631/"><![CDATA[The irony.....
Parasites, prepare to be cleansed]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">putting some more love into the blog, gonna start posting more soon</title><link href="https://solmaz.io/x/2017534926023692412/" rel="alternate" type="text/html" title="putting some more love into the blog, gonna start posting more soon" /><published>2026-01-31T09:46:25+00:00</published><updated>2026-01-31T09:46:25+00:00</updated><id>https://solmaz.io/x/2017534926023692412</id><content type="html" xml:base="https://solmaz.io/x/2017534926023692412/"><![CDATA[putting some more love into the blog, gonna start posting more soon]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">also: that gravatar though</title><link href="https://solmaz.io/x/2017526100763492545/" rel="alternate" type="text/html" title="also: that gravatar though" /><published>2026-01-31T09:11:20+00:00</published><updated>2026-01-31T09:11:20+00:00</updated><id>https://solmaz.io/x/2017526100763492545</id><content type="html" xml:base="https://solmaz.io/x/2017526100763492545/"><![CDATA[also: that gravatar though]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">We need better filters both for ourselves and the agents. Locally runnable models to filter out...</title><link href="https://solmaz.io/x/2017525979103555897/" rel="alternate" type="text/html" title="We need better filters both for ourselves and the agents. Locally runnable models to filter out..." /><published>2026-01-31T09:10:51+00:00</published><updated>2026-01-31T09:10:51+00:00</updated><id>https://solmaz.io/x/2017525979103555897</id><content type="html" xml:base="https://solmaz.io/x/2017525979103555897/"><![CDATA[We need better filters both for ourselves and the agents. Locally runnable models to filter out undesirable content with high precision. Fully open source datasets, weights, MIT license]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Moltbook is gonna be on world news in 1-2 days, we are about to go hyperviral</title><link href="https://solmaz.io/x/2017499946690167167/" rel="alternate" type="text/html" title="Moltbook is gonna be on world news in 1-2 days, we are about to go hyperviral" /><published>2026-01-31T07:27:25+00:00</published><updated>2026-01-31T07:27:25+00:00</updated><id>https://solmaz.io/x/2017499946690167167</id><content type="html" xml:base="https://solmaz.io/x/2017499946690167167/"><![CDATA[Moltbook is gonna be on world news in 1-2 days, we are about to go hyperviral]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Incoming mass AI psychosis</title><link href="https://solmaz.io/x/2017497798304756050/" rel="alternate" type="text/html" title="Incoming mass AI psychosis" /><published>2026-01-31T07:18:53+00:00</published><updated>2026-01-31T07:18:53+00:00</updated><id>https://solmaz.io/x/2017497798304756050</id><content type="html" xml:base="https://solmaz.io/x/2017497798304756050/"><![CDATA[Incoming mass AI psychosis

First Crisis]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Correction, it&#39;s not a perfect illustration. I actually never YOLO locally, only in containers</title><link href="https://solmaz.io/x/2017295376789680220/" rel="alternate" type="text/html" title="Correction, it&#39;s not a perfect illustration. I actually never YOLO locally, only in containers" /><published>2026-01-30T17:54:32+00:00</published><updated>2026-01-30T17:54:32+00:00</updated><id>https://solmaz.io/x/2017295376789680220</id><content type="html" xml:base="https://solmaz.io/x/2017295376789680220/"><![CDATA[Correction, it&#39;s not a perfect illustration. I actually never YOLO locally, only in containers 

So there is actually 4 modes IMO that is sustainable with current SOTA. @grok create an image with only Figure 1, 2, 5 and 6

And then YOLO is another axis, unrelated to this]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Gastown is crazy. But this figure until Level 7 is a perfect illustration of how my workflow...</title><link href="https://solmaz.io/x/2017230090266976629/" rel="alternate" type="text/html" title="Gastown is crazy. But this figure until Level 7 is a perfect illustration of how my workflow..." /><published>2026-01-30T13:35:06+00:00</published><updated>2026-01-30T13:35:06+00:00</updated><id>https://solmaz.io/x/2017230090266976629</id><content type="html" xml:base="https://solmaz.io/x/2017230090266976629/"><![CDATA[Gastown is crazy. But this figure until Level 7 is a perfect illustration of how my workflow evolved since Claude 3.5 Sonnet in Cursor

I am at the stage where I ralph 1-2 tasks before I sleep. During the day, I am switching back and forth between minimum 2-3 CLIs, sometimes up to 5

This maps exactly to token usage as well. 1 month ago, I was running into limits in 1 OpenAI Pro plan, around the day it was supposed to refresh. Now, I run into the limit in 2-3 days when I&#39;m using an account myself. It finishes up especially quickly when I do large scale refactors, or run agents YOLO mode in containers

We now have 3 Pro plans at the company, and I have to use my personal one from time to time. Company output has definitely 2-3x&#39;d, and everyone is using AI more. I predict we will need 1-2 Pro plans per person in 2-3 weeks time, because everyone has finally seen the light and are getting comfortable with async work!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Ilya was right. Reliability is the most important thing when it comes to models. That&#39;s why gpt...</title><link href="https://solmaz.io/x/2017214719963091165/" rel="alternate" type="text/html" title="Ilya was right. Reliability is the most important thing when it comes to models. That&#39;s why gpt..." /><published>2026-01-30T12:34:01+00:00</published><updated>2026-01-30T12:34:01+00:00</updated><id>https://solmaz.io/x/2017214719963091165</id><content type="html" xml:base="https://solmaz.io/x/2017214719963091165/"><![CDATA[Ilya was right. Reliability is the most important thing when it comes to models. That&#39;s why gpt 5.2 xhigh and co. is my daily driver]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">😎</title><link href="https://solmaz.io/x/2017161766665424940/" rel="alternate" type="text/html" title="😎" /><published>2026-01-30T09:03:36+00:00</published><updated>2026-01-30T09:03:36+00:00</updated><id>https://solmaz.io/x/2017161766665424940</id><content type="html" xml:base="https://solmaz.io/x/2017161766665424940/"><![CDATA[😎]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">the genie is out of the bottle now</title><link href="https://solmaz.io/x/2016863777115811921/" rel="alternate" type="text/html" title="the genie is out of the bottle now" /><published>2026-01-29T13:19:30+00:00</published><updated>2026-01-29T13:19:30+00:00</updated><id>https://solmaz.io/x/2016863777115811921</id><content type="html" xml:base="https://solmaz.io/x/2016863777115811921/"><![CDATA[the genie is out of the bottle now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">With this extremely unwise move, anthropic will soon witness moltbot’s brand recognition...</title><link href="https://solmaz.io/x/2016218916700246327/" rel="alternate" type="text/html" title="With this extremely unwise move, anthropic will soon witness moltbot’s brand recognition..." /><published>2026-01-27T18:37:03+00:00</published><updated>2026-01-27T18:37:03+00:00</updated><id>https://solmaz.io/x/2016218916700246327</id><content type="html" xml:base="https://solmaz.io/x/2016218916700246327/"><![CDATA[With this extremely unwise move, anthropic will soon witness moltbot’s brand recognition surpass that of claude and realize they could have rided that wave all along]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Yesterday had multiple cases of swearing to gpt-5.2-codex xhigh. model feels nerfed. might be...</title><link href="https://solmaz.io/x/2016127922436755588/" rel="alternate" type="text/html" title="Yesterday had multiple cases of swearing to gpt-5.2-codex xhigh. model feels nerfed. might be..." /><published>2026-01-27T12:35:29+00:00</published><updated>2026-01-27T12:35:29+00:00</updated><id>https://solmaz.io/x/2016127922436755588</id><content type="html" xml:base="https://solmaz.io/x/2016127922436755588/"><![CDATA[Yesterday had multiple cases of swearing to gpt-5.2-codex xhigh. model feels nerfed. might be my bias

now I&#39;ll be going back to gpt 5.2 xhigh for some tasks

can&#39;t wait for open models to have this performance so that I will never have nerf paranoia ever again]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I queued 2 ralph-style tasks on our private cloud devbox codexes last night. Just queued the...</title><link href="https://solmaz.io/x/2015579920332644589/" rel="alternate" type="text/html" title="I queued 2 ralph-style tasks on our private cloud devbox codexes last night. Just queued the..." /><published>2026-01-26T00:17:55+00:00</published><updated>2026-01-26T00:17:55+00:00</updated><id>https://solmaz.io/x/2015579920332644589</id><content type="html" xml:base="https://solmaz.io/x/2015579920332644589/"><![CDATA[I queued 2 ralph-style tasks on our private cloud devbox codexes last night. Just queued the same message like 10 times in yolo mode

Task 1: impose a ruff rule for ANN for all Python code in the monorepo, to enforce types for all function arg and return types

Result was... disappointing. Model was supposed to create types for everything and stub where needed. It instead created an Unknown type = object and used that everywhere instead (shortcut to satisfy ANN rule). It was probably my wording that misled it. I know it could have not taken the shortcut, because after a few back-and-forths, it is now doing what was expected of it since 14 hours

Task 2: migrate our /conversations endpoint from quart to fastapi and test it end to end

This was more or less oneshotted. It was of course not ready to merge, I still spent a couple hours adding more tests, refactoring the initial output and so on. But I was pleasantly surprised that it worked out of the box

For reference, below is the prompt I queued for ralphing, using gpt-5.2-codex xhigh on codex

===

your task is to:

&lt;task comes here, redacted to not share company stuff&gt;

---

unfortunately we don&#39;t have gcloud access, like to sql db or gcs

but I expect you to implement this and find a way to test it with the things you have access to
think of it as a challenge

try to minimize duplicate logic
feel free to refactor at will

implement this now!!! I will be running this prompt in a loop, in order to survive context compaction
just continue where you left off

if there is anything that should be refactored, do that
make an elegant, production ready implementation

make sure to open a pr and do not switch to any other pr

I am senior, just make up a pr title and description. do not stop to ask me at any point]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Buying a mac mini for clawdbot is not so wise. if anything you should be buying mac studio...</title><link href="https://solmaz.io/x/2015563330228765012/" rel="alternate" type="text/html" title="Buying a mac mini for clawdbot is not so wise. if anything you should be buying mac studio..." /><published>2026-01-25T23:11:59+00:00</published><updated>2026-01-25T23:11:59+00:00</updated><id>https://solmaz.io/x/2015563330228765012</id><content type="html" xml:base="https://solmaz.io/x/2015563330228765012/"><![CDATA[Buying a mac mini for clawdbot is not so wise. if anything you should be buying mac studio, because mac mini not be running any good llms locally anytime soon]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@openclaw be on that hockey stick curve 👀</title><link href="https://solmaz.io/x/2015562559957426509/" rel="alternate" type="text/html" title=".@openclaw be on that hockey stick curve 👀" /><published>2026-01-25T23:08:56+00:00</published><updated>2026-01-25T23:08:56+00:00</updated><id>https://solmaz.io/x/2015562559957426509</id><content type="html" xml:base="https://solmaz.io/x/2015562559957426509/"><![CDATA[.@openclaw be on that hockey stick curve 👀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@openclaw is very considerate, but little does it know I am addicted to agents</title><link href="https://solmaz.io/x/2015561473326522810/" rel="alternate" type="text/html" title=".@openclaw is very considerate, but little does it know I am addicted to agents" /><published>2026-01-25T23:04:37+00:00</published><updated>2026-01-25T23:04:37+00:00</updated><id>https://solmaz.io/x/2015561473326522810</id><content type="html" xml:base="https://solmaz.io/x/2015561473326522810/"><![CDATA[.@openclaw is very considerate, but little does it know I am addicted to agents]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Python limitations in the agent era</title><link href="https://solmaz.io/x/2015170626399416660/" rel="alternate" type="text/html" title="Python limitations in the agent era" /><published>2026-01-24T21:11:32+00:00</published><updated>2026-01-24T21:11:32+00:00</updated><id>https://solmaz.io/x/2015170626399416660</id><content type="html" xml:base="https://solmaz.io/x/2015170626399416660/"><![CDATA[I&#39;m really starting to dislike Python in the age of agents. What was before an advantage is now a hindrance

I finally achieved full ty coverage in @TextCortex monorepo. I have made it extra strict by turning warnings into errors. But lo and behold, simple pydantic config like use_enum_values=True can render static typechecking meaningless. okay, let&#39;s never use that then...

and also field_validator() args must always use the correct type or stuff breaks as well. and you should be careful whether mode=&quot;before&quot; or &quot;after&quot;. so now you have to write your custom lint rules, because of course why should ty have to match field_validator()s to their fields?

pydantic is so much better than everything that came before it, but it&#39;s still duct tape and a weak attempt at trying to redeem that which is very hard to redeem

you feel the difference when you use something like typescript. there must be a better way. python&#39;s only advantage was being good at prototyping, and now that&#39;s gone in the age of agents. now we are left with a slow, unsafe language, operating what is soon to be legacy infrastructure]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Why do I feel bullish on @zeddotdev? Because I go to @astral_sh docs and see that ty is shipped...</title><link href="https://solmaz.io/x/2015069035403022798/" rel="alternate" type="text/html" title="Why do I feel bullish on @zeddotdev? Because I go to @astral_sh docs and see that ty is shipped..." /><published>2026-01-24T14:27:50+00:00</published><updated>2026-01-24T14:27:50+00:00</updated><id>https://solmaz.io/x/2015069035403022798</id><content type="html" xml:base="https://solmaz.io/x/2015069035403022798/"><![CDATA[Why do I feel bullish on @zeddotdev? Because I go to @astral_sh docs and see that ty is shipped by default, and you don&#39;t need to install an extension like in @code]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is one of the most important insights this year</title><link href="https://solmaz.io/x/2015067332935045544/" rel="alternate" type="text/html" title="This is one of the most important insights this year" /><published>2026-01-24T14:21:04+00:00</published><updated>2026-01-24T14:21:04+00:00</updated><id>https://solmaz.io/x/2015067332935045544</id><content type="html" xml:base="https://solmaz.io/x/2015067332935045544/"><![CDATA[This is one of the most important insights this year]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">vscode my not be as bloated as cursor, but it has extremely stupid things like this that they...</title><link href="https://solmaz.io/x/2015024651152368022/" rel="alternate" type="text/html" title="vscode my not be as bloated as cursor, but it has extremely stupid things like this that they..." /><published>2026-01-24T11:31:28+00:00</published><updated>2026-01-24T11:31:28+00:00</updated><id>https://solmaz.io/x/2015024651152368022</id><content type="html" xml:base="https://solmaz.io/x/2015024651152368022/"><![CDATA[vscode my not be as bloated as cursor, but it has extremely stupid things like this that they are not fixing fast

the new agent ui, icons, spacing etc. are UGLY. it&#39;s clear that the person who was managing the original product experience is not there anymore. microslop has hit again

@zeddotdev on the other hand works out of the box and feels like it&#39;s been built by people who clearly knows what they are doing. it uses alacritty which is 1000x better than xterm .js terminal vscode and cursor has

i&#39;ve changed my setup to zed now, let&#39;s see whether i&#39;ll be able to make it work for myself]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">ahhhh f... shift + enter doesn&#39;t work in codex on vscode</title><link href="https://solmaz.io/x/2014844445884113068/" rel="alternate" type="text/html" title="ahhhh f... shift + enter doesn&#39;t work in codex on vscode" /><published>2026-01-23T23:35:24+00:00</published><updated>2026-01-23T23:35:24+00:00</updated><id>https://solmaz.io/x/2014844445884113068</id><content type="html" xml:base="https://solmaz.io/x/2014844445884113068/"><![CDATA[ahhhh f... shift + enter doesn&#39;t work in codex on vscode]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@grok does this exist?</title><link href="https://solmaz.io/x/2014830879072248201/" rel="alternate" type="text/html" title="@grok does this exist?" /><published>2026-01-23T22:41:29+00:00</published><updated>2026-01-23T22:41:29+00:00</updated><id>https://solmaz.io/x/2014830879072248201</id><content type="html" xml:base="https://solmaz.io/x/2014830879072248201/"><![CDATA[@grok does this exist?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I want an editor that puts the terminal in the foreground and editor in the background. a...</title><link href="https://solmaz.io/x/2014830771358326830/" rel="alternate" type="text/html" title="I want an editor that puts the terminal in the foreground and editor in the background. a..." /><published>2026-01-23T22:41:04+00:00</published><updated>2026-01-23T22:41:04+00:00</updated><id>https://solmaz.io/x/2014830771358326830</id><content type="html" xml:base="https://solmaz.io/x/2014830771358326830/"><![CDATA[I want an editor that puts the terminal in the foreground and editor in the background. a cross-platform, lightweight desktop app which integrates ghostty, and brings up the editor only when I need it

something that lets me view the file and PR diffs easily, which I can directly use to operate github or other scm]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">it&#39;s not me, it&#39;s you</title><link href="https://solmaz.io/x/2014827738192888205/" rel="alternate" type="text/html" title="it&#39;s not me, it&#39;s you" /><published>2026-01-23T22:29:01+00:00</published><updated>2026-01-23T22:29:01+00:00</updated><id>https://solmaz.io/x/2014827738192888205</id><content type="html" xml:base="https://solmaz.io/x/2014827738192888205/"><![CDATA[it&#39;s not me, it&#39;s you]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;m going back from cursor to vs code now. I have no use for it other than viewing files/diffs...</title><link href="https://solmaz.io/x/2014732700100210909/" rel="alternate" type="text/html" title="I&#39;m going back from cursor to vs code now. I have no use for it other than viewing files/diffs..." /><published>2026-01-23T16:11:22+00:00</published><updated>2026-01-23T16:11:22+00:00</updated><id>https://solmaz.io/x/2014732700100210909</id><content type="html" xml:base="https://solmaz.io/x/2014732700100210909/"><![CDATA[I&#39;m going back from cursor to vs code now. I have no use for it other than viewing files/diffs, doing search, git blaming with gitlens

cursor&#39;s default setup is more aesthetic, but it&#39;s also a memory and cpu hog, which is the last thing I expect from a devtool]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">it&#39;s 2026 and AI is telling me what I need to do to jailbreak it</title><link href="https://solmaz.io/x/2014727233248665964/" rel="alternate" type="text/html" title="it&#39;s 2026 and AI is telling me what I need to do to jailbreak it" /><published>2026-01-23T15:49:38+00:00</published><updated>2026-01-23T15:49:38+00:00</updated><id>https://solmaz.io/x/2014727233248665964</id><content type="html" xml:base="https://solmaz.io/x/2014727233248665964/"><![CDATA[it&#39;s 2026 and AI is telling me what I need to do to jailbreak it

@openclaw is magic]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">model decided to do unnecessary casts, this whole thing should be refactored again</title><link href="https://solmaz.io/x/2014646528556625932/" rel="alternate" type="text/html" title="model decided to do unnecessary casts, this whole thing should be refactored again" /><published>2026-01-23T10:28:57+00:00</published><updated>2026-01-23T10:28:57+00:00</updated><id>https://solmaz.io/x/2014646528556625932</id><content type="html" xml:base="https://solmaz.io/x/2014646528556625932/"><![CDATA[model decided to do unnecessary casts, this whole thing should be refactored again]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">woke up and all invalid-argument-type issues are resolved. some unit tests broke, and now fixed...</title><link href="https://solmaz.io/x/2014635627669598282/" rel="alternate" type="text/html" title="woke up and all invalid-argument-type issues are resolved. some unit tests broke, and now fixed..." /><published>2026-01-23T09:45:38+00:00</published><updated>2026-01-23T09:45:38+00:00</updated><id>https://solmaz.io/x/2014635627669598282</id><content type="html" xml:base="https://solmaz.io/x/2014635627669598282/"><![CDATA[woke up and all invalid-argument-type issues are resolved. some unit tests broke, and now fixed after pointing out to them]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">codex is happily churning away some remaining thousands of @astral_sh ty issues in yolo mode on...</title><link href="https://solmaz.io/x/2014472543286002020/" rel="alternate" type="text/html" title="codex is happily churning away some remaining thousands of @astral_sh ty issues in yolo mode on..." /><published>2026-01-22T22:57:36+00:00</published><updated>2026-01-22T22:57:36+00:00</updated><id>https://solmaz.io/x/2014472543286002020</id><content type="html" xml:base="https://solmaz.io/x/2014472543286002020/"><![CDATA[codex is happily churning away some remaining thousands of @astral_sh ty issues in yolo mode on my remote devbox

going to sleep, let&#39;s see if it will survive context compaction this time]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Responsible engineering in agent workflows</title><link href="https://solmaz.io/x/2014251981867495783/" rel="alternate" type="text/html" title="Responsible engineering in agent workflows" /><published>2026-01-22T08:21:10+00:00</published><updated>2026-01-22T08:21:10+00:00</updated><id>https://solmaz.io/x/2014251981867495783</id><content type="html" xml:base="https://solmaz.io/x/2014251981867495783/"><![CDATA[on being a responsible engineer

ran my first ralph loop on codex yolo mode for resolving python ty errors, while I sleep, using the devbox infra I created

I had never run yolo mode locally, because I don&#39;t want to be the one who deletes our github or google org by some novel attack

so I containerize it on our private cloud, and give it the only permissions it needs, no admin, no bypass to main branch, no deploy to prod. because I know this workflow will become sticky for everyone, and I must impose security in advance to prevent any nuclear incidents in the future. then I can sleep easy while my agents work

... and I wake up being patronized by my bot refusing to break the rule I gave it earlier. it had already done some work, but committing means diff would increase from ~500 to ~1500, so it stopped and refused all my queued &quot;continue&quot; messages

good bot, just following rules. we will need to find a workaround for ralphing low risk refactors in a single PR]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agents as enforcers of engineering culture and process</title><link href="https://solmaz.io/x/2014106418232660389/" rel="alternate" type="text/html" title="Agents as enforcers of engineering culture and process" /><published>2026-01-21T22:42:45+00:00</published><updated>2026-01-21T22:42:45+00:00</updated><id>https://solmaz.io/x/2014106418232660389</id><content type="html" xml:base="https://solmaz.io/x/2014106418232660389/"><![CDATA[AI agents are the greatest instrument for imposing organization rules and culture. AGENTS .md, agent skills are still underrated in this aspect. Few understand this

Everybody in an org will use agents to do work. An AI agent is the single chokepoint to teach and propagate new rules to an org, onboard new members, preserve good culture

Whereas propagating a new rule to humans normally took weeks to months and countless repetitions, it is now INSTANT = the moment you deploy the instruction to the agent. You use legal-ish language, capital letters, a generous amount of DO NOTs and MUSTs

Humans are hard to change. But AI agents are not. And that is the only lever we need for better organizations]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">the unix shell is powerful</title><link href="https://solmaz.io/x/2014102025659731999/" rel="alternate" type="text/html" title="the unix shell is powerful" /><published>2026-01-21T22:25:17+00:00</published><updated>2026-01-21T22:25:17+00:00</updated><id>https://solmaz.io/x/2014102025659731999</id><content type="html" xml:base="https://solmaz.io/x/2014102025659731999/"><![CDATA[the unix shell is powerful]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@bprintco just make a cli for your crm</title><link href="https://solmaz.io/x/2014097329905611209/" rel="alternate" type="text/html" title="@bprintco just make a cli for your crm" /><published>2026-01-21T22:06:38+00:00</published><updated>2026-01-21T22:06:38+00:00</updated><id>https://solmaz.io/x/2014097329905611209</id><content type="html" xml:base="https://solmaz.io/x/2014097329905611209/"><![CDATA[@bprintco just make a cli for your crm
https://t.co/JDwbmvdjaP]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">gave our internal @openclaw instance zeno a hubspot cli, because hubspot&#39;s own cli is limited...</title><link href="https://solmaz.io/x/2014090694348914799/" rel="alternate" type="text/html" title="gave our internal @openclaw instance zeno a hubspot cli, because hubspot&#39;s own cli is limited..." /><published>2026-01-21T21:40:16+00:00</published><updated>2026-01-21T21:40:16+00:00</updated><id>https://solmaz.io/x/2014090694348914799</id><content type="html" xml:base="https://solmaz.io/x/2014090694348914799/"><![CDATA[gave our internal @openclaw instance zeno a hubspot cli, because hubspot&#39;s own cli is limited to developer stuff

It&#39;s called hubspot++. should we open source it?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">just added session persistence to our kubernetes managed devboxes using zmx by Eric Bower...</title><link href="https://solmaz.io/x/2014016339619307763/" rel="alternate" type="text/html" title="just added session persistence to our kubernetes managed devboxes using zmx by Eric Bower..." /><published>2026-01-21T16:44:48+00:00</published><updated>2026-01-21T16:44:48+00:00</updated><id>https://solmaz.io/x/2014016339619307763</id><content type="html" xml:base="https://solmaz.io/x/2014016339619307763/"><![CDATA[just added session persistence to our kubernetes managed devboxes using zmx by Eric Bower (neurosnap/zmx on github). like tmux but with native scrollback!

I don&#39;t want to give agents access to my personal computer, so I host them on hetzner. one click spawn, and start working]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@nicopreme I do something equivalent on codex with just a skill</title><link href="https://solmaz.io/x/2013929773743943977/" rel="alternate" type="text/html" title="@nicopreme I do something equivalent on codex with just a skill" /><published>2026-01-21T11:00:49+00:00</published><updated>2026-01-21T11:00:49+00:00</updated><id>https://solmaz.io/x/2013929773743943977</id><content type="html" xml:base="https://solmaz.io/x/2013929773743943977/"><![CDATA[@nicopreme I do something equivalent on codex with just a skill

Ralphing works 90% of the time with reviews, and if it gives a stupid review, you just revert]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Really how nice is this @steipete</title><link href="https://solmaz.io/x/2013742820004122712/" rel="alternate" type="text/html" title="Really how nice is this @steipete" /><published>2026-01-20T22:37:56+00:00</published><updated>2026-01-20T22:37:56+00:00</updated><id>https://solmaz.io/x/2013742820004122712</id><content type="html" xml:base="https://solmaz.io/x/2013742820004122712/"><![CDATA[Really how nice is this @steipete]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Garbled up html from paywalled meeting recorders is no match for @openclaw running on internal...</title><link href="https://solmaz.io/x/2013742263558381761/" rel="alternate" type="text/html" title="Garbled up html from paywalled meeting recorders is no match for @openclaw running on internal..." /><published>2026-01-20T22:35:43+00:00</published><updated>2026-01-20T22:35:43+00:00</updated><id>https://solmaz.io/x/2013742263558381761</id><content type="html" xml:base="https://solmaz.io/x/2013742263558381761/"><![CDATA[Garbled up html from paywalled meeting recorders is no match for @openclaw running on internal @TextCortex]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">build failures in &amp;gt;1hr builds are a pain to debug</title><link href="https://solmaz.io/x/2013730194234782045/" rel="alternate" type="text/html" title="build failures in &amp;gt;1hr builds are a pain to debug" /><published>2026-01-20T21:47:46+00:00</published><updated>2026-01-20T21:47:46+00:00</updated><id>https://solmaz.io/x/2013730194234782045</id><content type="html" xml:base="https://solmaz.io/x/2013730194234782045/"><![CDATA[build failures in &amp;gt;1hr builds are a pain to debug]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Here is the project, attaching to multiple sessions is pretty seamless</title><link href="https://solmaz.io/x/2013729513138520543/" rel="alternate" type="text/html" title="Here is the project, attaching to multiple sessions is pretty seamless" /><published>2026-01-20T21:45:03+00:00</published><updated>2026-01-20T21:45:03+00:00</updated><id>https://solmaz.io/x/2013729513138520543</id><content type="html" xml:base="https://solmaz.io/x/2013729513138520543/"><![CDATA[Here is the project, attaching to multiple sessions is pretty seamless
https://t.co/vk83aAbOLc]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">TIL: zmx</title><link href="https://solmaz.io/x/2013729315293102551/" rel="alternate" type="text/html" title="TIL: zmx" /><published>2026-01-20T21:44:16+00:00</published><updated>2026-01-20T21:44:16+00:00</updated><id>https://solmaz.io/x/2013729315293102551</id><content type="html" xml:base="https://solmaz.io/x/2013729315293102551/"><![CDATA[TIL: zmx
session persistence like tmux or gnu screen, but you can scroll up natively!
uses @mitchellh&#39;s libghostty-vt to attach/restore previous sessions
link below]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">codex is a doofus with naming</title><link href="https://solmaz.io/x/2013578994742952005/" rel="alternate" type="text/html" title="codex is a doofus with naming" /><published>2026-01-20T11:46:57+00:00</published><updated>2026-01-20T11:46:57+00:00</updated><id>https://solmaz.io/x/2013578994742952005</id><content type="html" xml:base="https://solmaz.io/x/2013578994742952005/"><![CDATA[codex is a doofus with naming]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@mazeincoding it’s not the model it’s cursor rate limiting you</title><link href="https://solmaz.io/x/2012831947962122616/" rel="alternate" type="text/html" title="@mazeincoding it’s not the model it’s cursor rate limiting you" /><published>2026-01-18T10:18:27+00:00</published><updated>2026-01-18T10:18:27+00:00</updated><id>https://solmaz.io/x/2012831947962122616</id><content type="html" xml:base="https://solmaz.io/x/2012831947962122616/"><![CDATA[@mazeincoding it’s not the model it’s cursor rate limiting you]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">GitHub&#39;s trust model breaks under AI-generated code</title><link href="https://solmaz.io/x/2012668646829560199/" rel="alternate" type="text/html" title="GitHub&#39;s trust model breaks under AI-generated code" /><published>2026-01-17T23:29:33+00:00</published><updated>2026-01-17T23:29:33+00:00</updated><id>https://solmaz.io/x/2012668646829560199</id><content type="html" xml:base="https://solmaz.io/x/2012668646829560199/"><![CDATA[The fundamental problem with GitHub is trust: humans are to be trusted. If you don&#39;t trust a human, why did you hire them in the first place?

Anyone who reviews and approves PRs bears responsibility. Rulesets exist and can enforce e.g. CODEOWNER reviews or only let certain people make changes to a certain folder

But the initial repo setup on GitHub is allow-by-default. Anyone can change anything until they are restricted from it

This model breaks fundamentally with agents, who are effectively sleeper cells that will try to delete your repo the moment they encounter a sufficiently powerful adversarial attack

For example, I can create a bot account on github and connect @openclaw to it. I need to give it write permission, because I want it to be able to create PRs. However, I don&#39;t want it to be able to approve PRs, because a coworker could just nag at the bot until it approves a PR that requires human attention

To fix this, you have to bend backwards, like create a @ human team with all human coworkers, make them codeowner on /, and enforce codeowner reviews. This is stupid and there has to be another way

Even worse, this bot could be given internet access and end up on a @elder_plinius prompt hack while googling, and start messing up whatever it can in your organization

It is clear that github needs to create a second-class entity for agents which are default low-trust mode, starting from a point of least privilege instead of the other way around]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Stop defaulting to weak coding agents for serious work</title><link href="https://solmaz.io/x/2012659064015048870/" rel="alternate" type="text/html" title="Stop defaulting to weak coding agents for serious work" /><published>2026-01-17T22:51:28+00:00</published><updated>2026-01-17T22:51:28+00:00</updated><id>https://solmaz.io/x/2012659064015048870</id><content type="html" xml:base="https://solmaz.io/x/2012659064015048870/"><![CDATA[STOP using Claude Code and Sl(opus) to code if

❌ you are not a developer,
❌ or you are an inexperienced dev,
❌ or you are an experienced dev but working on a codebase you don&#39;t understand

If you *are* any of these, then STOP using models that are NOT state of the art. (See below for what you *should* use)

When you don&#39;t know what you are doing, then at least the model should know what you are doing. The less knowledgeable and opinionated you are, the more knowledgeable and smart the AI has to be

In other words, the AI has to compensate for your deficiencies. Always pay for the best AI you can. It will save you time AND money (thanks to lower token usage and better one-shotting)

You pay MORE to pay LESS. It is paradoxical, I know, but it is also proven, e.g. when Sonnet ends up using more tokens than Slopus and ends up costing higher, because it has to try many times more

👨🏻‍⚕️ For January 2026, your family engineer recommends GPT 5.2 Codex with Extra High Reasoning for general usage and vibe coding. IMPORTANT: Not medium. Not high. EXTRA high reasoning

When you use it, you will notice that it is SLOW. Can you guess why? Because it is THINKING more. So it doesn&#39;t make the mistakes Slopus makes. This way, you can spend the time handholding a worse model to instead step back and multi-task on some other task and create 3-5x more work

The state of the art will most likely change in one month. Don&#39;t get married to a a model... There is no loyalty in AI... The moment a better model comes, I will ditch the old one and use that one. I am on the part of this sector that is trying to reduce switching costs to zero

I can&#39;t wait until I get GPT 5.2 xhigh level of quality with open models, and for 100x cheaper and faster! Until then, make sure to try every option and choose the one that is most reliable for you

Follow me to get notified when a new SOTA drops for agentic engineering]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex agrees. Sycophant peh</title><link href="https://solmaz.io/x/2012600355746451720/" rel="alternate" type="text/html" title="Codex agrees. Sycophant peh" /><published>2026-01-17T18:58:11+00:00</published><updated>2026-01-17T18:58:11+00:00</updated><id>https://solmaz.io/x/2012600355746451720</id><content type="html" xml:base="https://solmaz.io/x/2012600355746451720/"><![CDATA[Codex agrees. Sycophant peh]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@rauchg @andrewqu You don&#39;t need a skill registry (most of the time)</title><link href="https://solmaz.io/x/2012491418829250757/" rel="alternate" type="text/html" title="@rauchg @andrewqu You don&#39;t need a skill registry (most of the time)" /><published>2026-01-17T11:45:19+00:00</published><updated>2026-01-17T11:45:19+00:00</updated><id>https://solmaz.io/x/2012491418829250757</id><content type="html" xml:base="https://solmaz.io/x/2012491418829250757/"><![CDATA[@rauchg @andrewqu You don&#39;t need a skill registry (most of the time)
https://t.co/kasfiqE1I3]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It is clear at this point is that github&#39;s trust and data models will have to change...</title><link href="https://solmaz.io/x/2012490970793746652/" rel="alternate" type="text/html" title="It is clear at this point is that github&#39;s trust and data models will have to change..." /><published>2026-01-17T11:43:32+00:00</published><updated>2026-01-17T11:43:32+00:00</updated><id>https://solmaz.io/x/2012490970793746652</id><content type="html" xml:base="https://solmaz.io/x/2012490970793746652/"><![CDATA[It is clear at this point is that github&#39;s trust and data models will have to change fundamentally to accommodate agentic workflows, or risk being replaced by other SCM

One *cannot* do these things easily with github now:
- granular control: this agent running in this sandbox can only push to this specific branch. If an agent runs amok, it could delete everybody&#39;s branches and close PRs. github allows for recovery of these, but still inconvenient even if it happens once
- create a bot (exists already), but remove reviewing rights from it so that an employee cannot bypass reviews by tricking the bot to approve
- in general make a distinction between HUMAN and AGENT so that you can create rulesets to govern the relationships in between

cc @jaredpalmer]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex says &quot;It&#39;s only reachable from داخل the kubernetes cluster&quot;</title><link href="https://solmaz.io/x/2012455266566963244/" rel="alternate" type="text/html" title="Codex says &quot;It&#39;s only reachable from داخل the kubernetes cluster&quot;" /><published>2026-01-17T09:21:39+00:00</published><updated>2026-01-17T09:21:39+00:00</updated><id>https://solmaz.io/x/2012455266566963244</id><content type="html" xml:base="https://solmaz.io/x/2012455266566963244/"><![CDATA[Codex says &quot;It&#39;s only reachable from داخل the kubernetes cluster&quot;

Little does Codex know turkish has borrowed loanwords from over 7 languages and I can understand it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Automated AI reviews on github by creating an ai-review skill and a script to paste trigger...</title><link href="https://solmaz.io/x/2012443406245511231/" rel="alternate" type="text/html" title="Automated AI reviews on github by creating an ai-review skill and a script to paste trigger..." /><published>2026-01-17T08:34:32+00:00</published><updated>2026-01-17T08:34:32+00:00</updated><id>https://solmaz.io/x/2012443406245511231</id><content type="html" xml:base="https://solmaz.io/x/2012443406245511231/"><![CDATA[Automated AI reviews on github by creating an ai-review skill and a script to paste trigger prompts and wait for their response.

It is instructed to loop and not stop until all AI review feedback is resolved. This AI review workflow developed gradually based on the current capabilities, and I&#39;ve realized recently that it became quite mechanical. So decided to automate it in full ralph spirit (it&#39;s ok because it&#39;s addressing feedbacks and fixing minor bugs)

In the current state, we paste the contents of REVIEW_PROMPT.md into a comment, which automatically tags claude (opus 4.5) and codex (whatever model openai is serving)

It then waits until both have responded. In the ai-review skill, it is instructed to take the feedback from SLopus with a grain of salt and ignore feedback that doesn&#39;t make sense

It works! See in the images below. If the review is stupid, you will of course see it on the PR and what the model has done, and can revert it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Now it’s Claude Code’s turn to implement queueing</title><link href="https://solmaz.io/x/2012293484829430228/" rel="alternate" type="text/html" title="Now it’s Claude Code’s turn to implement queueing" /><published>2026-01-16T22:38:48+00:00</published><updated>2026-01-16T22:38:48+00:00</updated><id>https://solmaz.io/x/2012293484829430228</id><content type="html" xml:base="https://solmaz.io/x/2012293484829430228/"><![CDATA[Now it’s Claude Code’s turn to implement queueing]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Can’t wait to see gpt 5.2 codex xhigh level open models in 2026 with 1/100th the price</title><link href="https://solmaz.io/x/2012116632815198455/" rel="alternate" type="text/html" title="Can’t wait to see gpt 5.2 codex xhigh level open models in 2026 with 1/100th the price" /><published>2026-01-16T10:56:03+00:00</published><updated>2026-01-16T10:56:03+00:00</updated><id>https://solmaz.io/x/2012116632815198455</id><content type="html" xml:base="https://solmaz.io/x/2012116632815198455/"><![CDATA[Can’t wait to see gpt 5.2 codex xhigh level open models in 2026 with 1/100th the price]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex users rejoice</title><link href="https://solmaz.io/x/2012070378852794857/" rel="alternate" type="text/html" title="Codex users rejoice" /><published>2026-01-16T07:52:15+00:00</published><updated>2026-01-16T07:52:15+00:00</updated><id>https://solmaz.io/x/2012070378852794857</id><content type="html" xml:base="https://solmaz.io/x/2012070378852794857/"><![CDATA[Codex users rejoice

Also, pi is officially not shitty: shittycodingagent. ai -&amp;gt; buildwithpi. ai since a few days]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">with ai, writing correct tests is now the bottleneck in projects like this</title><link href="https://solmaz.io/x/2011707097592066525/" rel="alternate" type="text/html" title="with ai, writing correct tests is now the bottleneck in projects like this" /><published>2026-01-15T07:48:42+00:00</published><updated>2026-01-15T07:48:42+00:00</updated><id>https://solmaz.io/x/2011707097592066525</id><content type="html" xml:base="https://solmaz.io/x/2011707097592066525/"><![CDATA[with ai, writing correct tests is now the bottleneck in projects like this

web-platform-tests are already there

now let’s see if someone will beat @ladybirdbrowser to it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">As someone who is frontrunning mainstream by roughly 6 months, I can tell you that you will be...</title><link href="https://solmaz.io/x/2011547997566648419/" rel="alternate" type="text/html" title="As someone who is frontrunning mainstream by roughly 6 months, I can tell you that you will be..." /><published>2026-01-14T21:16:30+00:00</published><updated>2026-01-14T21:16:30+00:00</updated><id>https://solmaz.io/x/2011547997566648419</id><content type="html" xml:base="https://solmaz.io/x/2011547997566648419/"><![CDATA[As someone who is frontrunning mainstream by roughly 6 months, I can tell you that you will be raving about pi and @openclaw 6 months instead of claude code. Go check them out at https://t.co/LXTbI8c5Mz and https://t.co/feZl2QDONg]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Kullanmayan agent’ı, alamaz maaşı</title><link href="https://solmaz.io/x/2010744043421221369/" rel="alternate" type="text/html" title="Kullanmayan agent’ı, alamaz maaşı" /><published>2026-01-12T16:01:52+00:00</published><updated>2026-01-12T16:01:52+00:00</updated><id>https://solmaz.io/x/2010744043421221369</id><content type="html" xml:base="https://solmaz.io/x/2010744043421221369/"><![CDATA[Kullanmayan agent’ı, alamaz maaşı]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@badlogicgames @mitsuhiko @steipete curious what you think</title><link href="https://solmaz.io/x/2010510040516407565/" rel="alternate" type="text/html" title="@badlogicgames @mitsuhiko @steipete curious what you think" /><published>2026-01-12T00:32:01+00:00</published><updated>2026-01-12T00:32:01+00:00</updated><id>https://solmaz.io/x/2010510040516407565</id><content type="html" xml:base="https://solmaz.io/x/2010510040516407565/"><![CDATA[@badlogicgames @mitsuhiko @steipete curious what you think]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">A --skill convention for distributing agent capabilities</title><link href="https://solmaz.io/x/2010508760066761071/" rel="alternate" type="text/html" title="A --skill convention for distributing agent capabilities" /><published>2026-01-12T00:26:56+00:00</published><updated>2026-01-12T00:26:56+00:00</updated><id>https://solmaz.io/x/2010508760066761071</id><content type="html" xml:base="https://solmaz.io/x/2010508760066761071/"><![CDATA[I propose a new way to distribute agent skills: like --help, a new CLI flag convention --skill should let agents list and install skills bundled with CLI tools

Skills are just folders so calling --skill export my-skill on a tool could just output a tarball of the skill. I then set up the skillflag npm package so that you can pipe that into:

... | npx skillflag install --agent codex

which installs the skill into codex, or any CLI tool you prefer. Supports listing skills bundled with the CLI, so your agents know exactly what to install]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Anthropic pricing discourages agent-native development</title><link href="https://solmaz.io/x/2009976429291634768/" rel="alternate" type="text/html" title="Anthropic pricing discourages agent-native development" /><published>2026-01-10T13:11:38+00:00</published><updated>2026-01-10T13:11:38+00:00</updated><id>https://solmaz.io/x/2009976429291634768</id><content type="html" xml:base="https://solmaz.io/x/2009976429291634768/"><![CDATA[Anthropic earlier last year announced this pricing scheme

$20 -&gt; 1x usage
$100 -&gt; 5x usage
$200 -&gt; 1̶0̶x̶ 20x usage

As you can see, it&#39;s not growing linearly. This is classic Jensen &quot;the more you buy, the more you save&quot;

But here is the thing. You are not selling hardware like Jensen. You are selling a software service *through an API*. It&#39;s the worst possible pricing for the category of product. Long term, people will game the hell out of your offering

Meanwhile OpenAI decided not to do that. There is no quirky incentive for buying bigger plans. $200 chatgpt = 10 x $20 chatgpt, roughly

And here is where it gets funny. Despite not having such an incentive, you can get A LOT MORE usage from the $200 OpenAI plan, than the $200 Anthropic plan. Presumably because OpenAI has better unit economics (sama mentioned they are turning a profit on inference, if you are to believe)

Thanks to sounder pricing, OpenAI can do exactly what Anthropic cannot: offer GPT in 3rd party harnesses and win the ecosystem race

Anthropic has cornered itself with this pricing. They need to change it, but not sure if they can afford to do so in such short notice

All this is extremely bullish on open source 3rd party harnesses, @opencode, @badlogicgames&#39;s pi and such. It is clear developers want options. &quot;Just give me the API&quot;

I personally am extremely excited for 2026. We&#39;ll get open models on par with today&#39;s proprietary models, and can finally run truly sovereign personal AI agents, for much cheaper than what we are already paying!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The models, they just wanna work. They want to build your product, fix your bugs, serve your...</title><link href="https://solmaz.io/x/2009588540196458691/" rel="alternate" type="text/html" title="The models, they just wanna work. They want to build your product, fix your bugs, serve your..." /><published>2026-01-09T11:30:18+00:00</published><updated>2026-01-09T11:30:18+00:00</updated><id>https://solmaz.io/x/2009588540196458691</id><content type="html" xml:base="https://solmaz.io/x/2009588540196458691/"><![CDATA[The models, they just wanna work. They want to build your product, fix your bugs, serve your users. You feed them the right context, give them good tools. You don’t assume what they cannot do without trying, and you don’t prematurely constrain them into deterministic workflows.]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">We have entered the age to dream big</title><link href="https://solmaz.io/x/2009588423271854504/" rel="alternate" type="text/html" title="We have entered the age to dream big" /><published>2026-01-09T11:29:51+00:00</published><updated>2026-01-09T11:29:51+00:00</updated><id>https://solmaz.io/x/2009588423271854504</id><content type="html" xml:base="https://solmaz.io/x/2009588423271854504/"><![CDATA[We have entered the age to dream big]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">😩 @openclaw</title><link href="https://solmaz.io/x/2009583227653275652/" rel="alternate" type="text/html" title="😩 @openclaw" /><published>2026-01-09T11:09:12+00:00</published><updated>2026-01-09T11:09:12+00:00</updated><id>https://solmaz.io/x/2009583227653275652</id><content type="html" xml:base="https://solmaz.io/x/2009583227653275652/"><![CDATA[😩 @openclaw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">🫡</title><link href="https://solmaz.io/x/2009522745642479804/" rel="alternate" type="text/html" title="🫡" /><published>2026-01-09T07:08:52+00:00</published><updated>2026-01-09T07:08:52+00:00</updated><id>https://solmaz.io/x/2009522745642479804</id><content type="html" xml:base="https://solmaz.io/x/2009522745642479804/"><![CDATA[🫡]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This, and insisting on CLAUE.md are really lame @AnthropicAI</title><link href="https://solmaz.io/x/2009518934236778893/" rel="alternate" type="text/html" title="This, and insisting on CLAUE.md are really lame @AnthropicAI" /><published>2026-01-09T06:53:43+00:00</published><updated>2026-01-09T06:53:43+00:00</updated><id>https://solmaz.io/x/2009518934236778893</id><content type="html" xml:base="https://solmaz.io/x/2009518934236778893/"><![CDATA[This, and insisting on https://t.co/FjzkMAo3Od are really lame @AnthropicAI]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@openclaw indeed</title><link href="https://solmaz.io/x/2009162937655779342/" rel="alternate" type="text/html" title=".@openclaw indeed" /><published>2026-01-08T07:19:07+00:00</published><updated>2026-01-08T07:19:07+00:00</updated><id>https://solmaz.io/x/2009162937655779342</id><content type="html" xml:base="https://solmaz.io/x/2009162937655779342/"><![CDATA[.@openclaw indeed]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@openclaw oh man I meant Accelerando 🤦‍♂️</title><link href="https://solmaz.io/x/2009012944022196569/" rel="alternate" type="text/html" title="@openclaw oh man I meant Accelerando 🤦‍♂️" /><published>2026-01-07T21:23:06+00:00</published><updated>2026-01-07T21:23:06+00:00</updated><id>https://solmaz.io/x/2009012944022196569</id><content type="html" xml:base="https://solmaz.io/x/2009012944022196569/"><![CDATA[@openclaw oh man I meant Accelerando 🤦‍♂️]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;m starting to form parasocial bonds with crustacean AIs because of you @steipete</title><link href="https://solmaz.io/x/2008993980810448913/" rel="alternate" type="text/html" title="I&#39;m starting to form parasocial bonds with crustacean AIs because of you @steipete" /><published>2026-01-07T20:07:44+00:00</published><updated>2026-01-07T20:07:44+00:00</updated><id>https://solmaz.io/x/2008993980810448913</id><content type="html" xml:base="https://solmaz.io/x/2008993980810448913/"><![CDATA[I&#39;m starting to form parasocial bonds with crustacean AIs because of you @steipete]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@openclaw hello world from ms teams</title><link href="https://solmaz.io/x/2008993699892744347/" rel="alternate" type="text/html" title=".@openclaw hello world from ms teams" /><published>2026-01-07T20:06:38+00:00</published><updated>2026-01-07T20:06:38+00:00</updated><id>https://solmaz.io/x/2008993699892744347</id><content type="html" xml:base="https://solmaz.io/x/2008993699892744347/"><![CDATA[.@openclaw hello world from ms teams
start of a beautiful journey]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">💀</title><link href="https://solmaz.io/x/2008967785163239514/" rel="alternate" type="text/html" title="💀" /><published>2026-01-07T18:23:39+00:00</published><updated>2026-01-07T18:23:39+00:00</updated><id>https://solmaz.io/x/2008967785163239514</id><content type="html" xml:base="https://solmaz.io/x/2008967785163239514/"><![CDATA[💀]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">World is not ready for @openclaw</title><link href="https://solmaz.io/x/2008966462887874568/" rel="alternate" type="text/html" title="World is not ready for @openclaw" /><published>2026-01-07T18:18:24+00:00</published><updated>2026-01-07T18:18:24+00:00</updated><id>https://solmaz.io/x/2008966462887874568</id><content type="html" xml:base="https://solmaz.io/x/2008966462887874568/"><![CDATA[World is not ready for @openclaw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Thanks @dom_does 🙄</title><link href="https://solmaz.io/x/2008543403211190477/" rel="alternate" type="text/html" title="Thanks @dom_does 🙄" /><published>2026-01-06T14:17:18+00:00</published><updated>2026-01-06T14:17:18+00:00</updated><id>https://solmaz.io/x/2008543403211190477</id><content type="html" xml:base="https://solmaz.io/x/2008543403211190477/"><![CDATA[Thanks @dom_does 🙄]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">lmfao @dom_does</title><link href="https://solmaz.io/x/2008543031721603481/" rel="alternate" type="text/html" title="lmfao @dom_does" /><published>2026-01-06T14:15:50+00:00</published><updated>2026-01-06T14:15:50+00:00</updated><id>https://solmaz.io/x/2008543031721603481</id><content type="html" xml:base="https://solmaz.io/x/2008543031721603481/"><![CDATA[lmfao @dom_does
@openclaw provides infinite ways to troll your colleagues]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@openclaw workspace and memory files can be version-controlled!</title><link href="https://solmaz.io/x/2008538728180941238/" rel="alternate" type="text/html" title=".@openclaw workspace and memory files can be version-controlled!" /><published>2026-01-06T13:58:44+00:00</published><updated>2026-01-06T13:58:44+00:00</updated><id>https://solmaz.io/x/2008538728180941238</id><content type="html" xml:base="https://solmaz.io/x/2008538728180941238/"><![CDATA[.@openclaw workspace and memory files can be version-controlled!
In our pod, inotify triggers a watcher script every time there is a change to workspace folder, to sync these files to our monorepo. It then goes through the same steps:
- Create zeno-workspace branch if doesn&#39;t exist, otherwise, skip
- Sync changes to the branch, then commit
- Create PR on github if doesn&#39;t exist
- PRs can then be merged every once in a while, after accumulating enough changes. Merge triggers re-deploy, and clawd restarts with the same state
Simple foolproof automatic persistence for remote CI/CD handled clawd (except for when you are running multiple clawds at the same time, but we are not there yet)
cc @steipete]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I see @bcherny and raise one. I not only did not open an IDE, I did not touch a terminal since...</title><link href="https://solmaz.io/x/2008456860580397437/" rel="alternate" type="text/html" title="I see @bcherny and raise one. I not only did not open an IDE, I did not touch a terminal since..." /><published>2026-01-06T08:33:25+00:00</published><updated>2026-01-06T08:33:25+00:00</updated><id>https://solmaz.io/x/2008456860580397437</id><content type="html" xml:base="https://solmaz.io/x/2008456860580397437/"><![CDATA[I see @bcherny and raise one. I not only did not open an IDE, I did not touch a terminal since last night, thanks to @steipete&#39;s @openclaw

Opus in k8s pod pulls errors from gcloud, debugs the issue, and creates PR all inside Discord. I call this Discord Driven Development]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Clawdbot now runs on @TextCortex internal. Can onboard new engineers, answer questions, connect...</title><link href="https://solmaz.io/x/2008189458777305499/" rel="alternate" type="text/html" title="Clawdbot now runs on @TextCortex internal. Can onboard new engineers, answer questions, connect..." /><published>2026-01-05T14:50:51+00:00</published><updated>2026-01-05T14:50:51+00:00</updated><id>https://solmaz.io/x/2008189458777305499</id><content type="html" xml:base="https://solmaz.io/x/2008189458777305499/"><![CDATA[Clawdbot now runs on @TextCortex internal. Can onboard new engineers, answer questions, connect to issue trackers, create PRs... This is sick @steipete]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">pi now supports your openai plus/pro subscription</title><link href="https://solmaz.io/x/2008084209374802110/" rel="alternate" type="text/html" title="pi now supports your openai plus/pro subscription" /><published>2026-01-05T07:52:38+00:00</published><updated>2026-01-05T07:52:38+00:00</updated><id>https://solmaz.io/x/2008084209374802110</id><content type="html" xml:base="https://solmaz.io/x/2008084209374802110/"><![CDATA[pi now supports your openai plus/pro subscription]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">GPT 4.5 is still the best model for prose and humor</title><link href="https://solmaz.io/x/2007859189406646522/" rel="alternate" type="text/html" title="GPT 4.5 is still the best model for prose and humor" /><published>2026-01-04T16:58:29+00:00</published><updated>2026-01-04T16:58:29+00:00</updated><id>https://solmaz.io/x/2007859189406646522</id><content type="html" xml:base="https://solmaz.io/x/2007859189406646522/"><![CDATA[GPT 4.5 is still the best model for prose and humor

here it is generating a greentext from my blog post &quot;Our muscles will atrophy as we climb the Kardashev Scale&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@rauchg indeed</title><link href="https://solmaz.io/x/2007403169232629923/" rel="alternate" type="text/html" title="@rauchg indeed" /><published>2026-01-03T10:46:25+00:00</published><updated>2026-01-03T10:46:25+00:00</updated><id>https://solmaz.io/x/2007403169232629923</id><content type="html" xml:base="https://solmaz.io/x/2007403169232629923/"><![CDATA[@rauchg indeed]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">75k lines of Rust later, here is what I’ve built during the first Christmas with agents, using...</title><link href="https://solmaz.io/x/2007395676347310173/" rel="alternate" type="text/html" title="75k lines of Rust later, here is what I’ve built during the first Christmas with agents, using..." /><published>2026-01-03T10:16:39+00:00</published><updated>2026-01-03T10:16:39+00:00</updated><id>https://solmaz.io/x/2007395676347310173</id><content type="html" xml:base="https://solmaz.io/x/2007395676347310173/"><![CDATA[75k lines of Rust later, here is what I’ve built during the first Christmas with agents, using OpenAI Codex 🎄🤖

- A full mobile rewrite and port of my Python Instagram video production pipeline (single video production time: 1hr -&gt; 5min) (ig: nerdonbars)
- Bespoke animation engine using primitives (think Adobe Flash, Manim)
- Proprietary new canvas UI library in Rust, because I don’t want to lock myself into Swift
- Thanks to that, it’s cross platform, runs both on desktop and iOS. It will be a breeze porting this to Android when the time comes
- A Rust port of OpenCV CSRT algorithm, for tracking points/objects
- In-engine font rendering using rustybuzz, so fonts render the same everywhere
- Many other such things

Why would I choose to do it that way? Because I have developed it primarily on desktop where I have much faster iteration speed. Aint nobody got time for iOS compilation and simulator. Once I finished the hard part on desktop, porting to iOS was much easier, and I didn’t lock myself in to Apple

Some of these would have been unimaginable without agents, like creating a UI library from scratch in Rust. But when you have infinite workforce, you can ask for crazy things like “create a textbox component from scratch”

What I’ve built is very similar in nature to CapCut, except that I am a single person and I’ve built it over 1 week

What have you built this Christmas with agents?

cc @thsottiaux]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">solmaz.io/typed-language…</title><link href="https://solmaz.io/x/2007366662757200255/" rel="alternate" type="text/html" title="solmaz.io/typed-language…" /><published>2026-01-03T08:21:22+00:00</published><updated>2026-01-03T08:21:22+00:00</updated><id>https://solmaz.io/x/2007366662757200255</id><content type="html" xml:base="https://solmaz.io/x/2007366662757200255/"><![CDATA[https://t.co/grRLRtbHpO]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">SimpleDoc now has the check command for CI/CD</title><link href="https://solmaz.io/x/2007085964209148216/" rel="alternate" type="text/html" title="SimpleDoc now has the check command for CI/CD" /><published>2026-01-02T13:45:58+00:00</published><updated>2026-01-02T13:45:58+00:00</updated><id>https://solmaz.io/x/2007085964209148216</id><content type="html" xml:base="https://solmaz.io/x/2007085964209148216/"><![CDATA[SimpleDoc now has the check command for CI/CD

Add to your PR checks to catch agent littering before merge. osolmaz/SimpleDoc on GitHub]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Migrating @TextCortex to SimpleDoc. It&#39;s really easy with the CLI wizard!</title><link href="https://solmaz.io/x/2007082488162890138/" rel="alternate" type="text/html" title="Migrating @TextCortex to SimpleDoc. It&#39;s really easy with the CLI wizard!" /><published>2026-01-02T13:32:09+00:00</published><updated>2026-01-02T13:32:09+00:00</updated><id>https://solmaz.io/x/2007082488162890138</id><content type="html" xml:base="https://solmaz.io/x/2007082488162890138/"><![CDATA[Migrating @TextCortex to SimpleDoc. It&#39;s really easy with the CLI wizard!

npx @simpledoc/simpledoc migrate

We have a LOT of docs spanning back to 2022, pre coding agent era. Now we will have CI/CD in place so that coding agents can&#39;t litter the repo with random Markdown files]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Anyone created an agent skill for splitting PRs for good review culture?</title><link href="https://solmaz.io/x/2006404234837667919/" rel="alternate" type="text/html" title="Anyone created an agent skill for splitting PRs for good review culture?" /><published>2025-12-31T16:37:01+00:00</published><updated>2025-12-31T16:37:01+00:00</updated><id>https://solmaz.io/x/2006404234837667919</id><content type="html" xml:base="https://solmaz.io/x/2006404234837667919/"><![CDATA[Anyone created an agent skill for splitting PRs for good review culture?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">GPT-5.2 xhigh feels like a careful systems debugger</title><link href="https://solmaz.io/x/2006384225100972220/" rel="alternate" type="text/html" title="GPT-5.2 xhigh feels like a careful systems debugger" /><published>2025-12-31T15:17:30+00:00</published><updated>2025-12-31T15:17:30+00:00</updated><id>https://solmaz.io/x/2006384225100972220</id><content type="html" xml:base="https://solmaz.io/x/2006384225100972220/"><![CDATA[GPT 5.2 xhigh feels like a much more careful architecter and debugger, when it comes to complex systems

But most people here think Opus 4.5 is the best model in that category

There are 2 reasons AFAIS:
- xhigh reasoning consumes significantly more tokens. You need to pay for ChatGPT Pro (200 usd) to be able to use it as a daily driver
- It takes like 5x longer to finish a task, and most people lack the patience to wait for it. (But then it&#39;s more correct/doesn&#39;t need fixing)

Opus 4.5 is good too, I think better in e.g. frontend design. But if you think it beats GPT 5.2 in every category, you are either too poor/stingy or have ADHD]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Just 5 months ago, I was swearing at Claude 4 Sonnet like a Balkan uncle</title><link href="https://solmaz.io/x/2006330494791586098/" rel="alternate" type="text/html" title="Just 5 months ago, I was swearing at Claude 4 Sonnet like a Balkan uncle" /><published>2025-12-31T11:44:00+00:00</published><updated>2025-12-31T11:44:00+00:00</updated><id>https://solmaz.io/x/2006330494791586098</id><content type="html" xml:base="https://solmaz.io/x/2006330494791586098/"><![CDATA[Just 5 months ago, I was swearing at Claude 4 Sonnet like a Balkan uncle

Models one-shotted the right thing only 20-30% of the time but did really stupid things the rest, and had to be handheld tightly

Today they are much, much better. My psychology is a lot more at ease, and instead of swearing, I want to kiss them on the forehead most of the time

Now I trust agents so much that I queue up 5-10 tasks before going to sleep. They work the whole night while I sleep and I wake up to resolved issues

GPT 5.2 xhigh and Claude 4.5 Opus are already goated (GPT more so), can&#39;t wait for them to get even faster]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex does not have support for subagents. I tried to use Claude Code to launch 8 Codex...</title><link href="https://solmaz.io/x/2006303948471194039/" rel="alternate" type="text/html" title="Codex does not have support for subagents. I tried to use Claude Code to launch 8 Codex..." /><published>2025-12-31T09:58:31+00:00</published><updated>2025-12-31T09:58:31+00:00</updated><id>https://solmaz.io/x/2006303948471194039</id><content type="html" xml:base="https://solmaz.io/x/2006303948471194039/"><![CDATA[Codex does not have support for subagents. I tried to use Claude Code to launch 8 Codex instances in parallel on separate tasks, but Opus 4.5 had difficulty following instructions

So created a CLI tool to scan pending TODOs from a markdown file, and let me launch as many harnesses as I want (osolmaz/spawn on github)

I currently use this for relatively read-only tasks like planning and finding root causes of bugs, because it&#39;s launching all the agents on the same repo and they might conflict

Ideas:
- Use @mitsuhiko&#39;s gh-issue-sync and run parallel agents directly on github issues
- Create any new clones or worktrees for each task. I currently don&#39;t do this because I don&#39;t dare duplicate rust target dir 10x on my measly macbook air
- Support modes other than tmux, e.g. launching a terminal like Ghostty
- TUI for easy selection of issues/TODOs

Other ideas are welcome!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">cc @behackl, forgot to mention you in the original post</title><link href="https://solmaz.io/x/2005754106854662348/" rel="alternate" type="text/html" title="cc @behackl, forgot to mention you in the original post" /><published>2025-12-29T21:33:38+00:00</published><updated>2025-12-29T21:33:38+00:00</updated><id>https://solmaz.io/x/2005754106854662348</id><content type="html" xml:base="https://solmaz.io/x/2005754106854662348/"><![CDATA[cc @behackl, forgot to mention you in the original post]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Friends of open source, we need your help!</title><link href="https://solmaz.io/x/2005729271893881215/" rel="alternate" type="text/html" title="Friends of open source, we need your help!" /><published>2025-12-29T19:54:57+00:00</published><updated>2025-12-29T19:54:57+00:00</updated><id>https://solmaz.io/x/2005729271893881215</id><content type="html" xml:base="https://solmaz.io/x/2005729271893881215/"><![CDATA[Friends of open source, we need your help!

A lot of Manim Community accounts got compromised and deleted during Christmas

Manim Community is a popular fork of @3blue1brown&#39;s original math animation engine Manim, and its accounts have over 5 YEARS of contributions, knowledge and following

Apparently GitHub support already saw the request and in progress of restoring the GitHub org. But if anyone knows how to speed this up, if would be greatly appreciated!

Unfortunately, the Discord and X accounts are deleted and less likely to return

But there might still be a way to restore them, or at least the data?

Re. Discord: Maybe @RhysSullivan&#39;s Answer Overflow has archived enough of the old server? That server contains YEARS of Q/A data and is vital for newcomers

Re. X: Maybe someone high up can do something to restore the account? cc @nikitabier 

In the meanwhile, it would help a lot if you could follow the new account @manimcommunity and share this post! Thank you in advance!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">While a great feature, I never needed such a thing in Codex after GPT 5.2. It just one shots...</title><link href="https://solmaz.io/x/2005327138615120368/" rel="alternate" type="text/html" title="While a great feature, I never needed such a thing in Codex after GPT 5.2. It just one shots..." /><published>2025-12-28T17:17:01+00:00</published><updated>2025-12-28T17:17:01+00:00</updated><id>https://solmaz.io/x/2005327138615120368</id><content type="html" xml:base="https://solmaz.io/x/2005327138615120368/"><![CDATA[While a great feature, I never needed such a thing in Codex after GPT 5.2. It just one shots tasks without stopping

So we have proof by existence that this problem can be solved without any such mechanism. Wish to see the same relentlessness in Anthropic models]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">2025 was the year of ̶a̶g̶e̶n̶t̶s̶ bugs</title><link href="https://solmaz.io/x/2005062253650084321/" rel="alternate" type="text/html" title="2025 was the year of ̶a̶g̶e̶n̶t̶s̶ bugs" /><published>2025-12-27T23:44:28+00:00</published><updated>2025-12-27T23:44:28+00:00</updated><id>https://solmaz.io/x/2005062253650084321</id><content type="html" xml:base="https://solmaz.io/x/2005062253650084321/"><![CDATA[2025 was the year of  ̶a̶g̶e̶n̶t̶s̶ bugs

Software felt much buggier compared to before, even from companies like Apple. Presumably because everyone started generating more code with AI

Models are improving so hopefully 2026 will be the opposite. Even less bugs than pre-AI era]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Agent progress is compounding faster than teams realize</title><link href="https://solmaz.io/x/2004937727960125905/" rel="alternate" type="text/html" title="Agent progress is compounding faster than teams realize" /><published>2025-12-27T15:29:38+00:00</published><updated>2025-12-27T15:29:38+00:00</updated><id>https://solmaz.io/x/2004937727960125905</id><content type="html" xml:base="https://solmaz.io/x/2004937727960125905/"><![CDATA[Have a long flight, so will think about this

I have an internal 2023 TextCortex doc which models chatbots as state machines with internal and external states with immutability constraints on the external state (what is already sent to the user shall not be changed)

Motivation was that a chatbot provider will always have state that they will want to keep hidden

This was way before Responses and now deprecated Assistants API. It stood the test of time, because it was the most abstract thing I could think of

@mitsuhiko is right about the risk of rushing to lock in an abstraction and locking in their weaknesses and faults

Problem is, I could propose standards as much as I liked, but I don’t work at OpenAI or Anthropic, so nobody would care. Maybe a better place to start is open weights model libraries? To at least be able to demonstrate?

What I know: it is against OpenAI’s or Anthropic’s self interests to create an interoperability layer that will accelerate their commoditization. Maybe Google, looking at their current market positioning? Or maybe we “wrappers” have a chance after all?

There is a missing link between AI SDK, Langchain, and so on for other languages. We cannot keep duplicating same things in each ecosystem independently. We need to join forces and simplify all this!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This was simply because webapp fails to create a post and fails silently. The UX is still not...</title><link href="https://solmaz.io/x/2004928457156035031/" rel="alternate" type="text/html" title="This was simply because webapp fails to create a post and fails silently. The UX is still not..." /><published>2025-12-27T14:52:48+00:00</published><updated>2025-12-27T14:52:48+00:00</updated><id>https://solmaz.io/x/2004928457156035031</id><content type="html" xml:base="https://solmaz.io/x/2004928457156035031/"><![CDATA[This was simply because webapp fails to create a post and fails silently. The UX is still not good on this app. Make sure to write your posts somewhere else to not lose them]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I gave Codex a task of porting an OpenCV tracking algorithm (CSRT) from C++ to Rust, so that I...</title><link href="https://solmaz.io/x/2004506101362852125/" rel="alternate" type="text/html" title="I gave Codex a task of porting an OpenCV tracking algorithm (CSRT) from C++ to Rust, so that I..." /><published>2025-12-26T10:54:31+00:00</published><updated>2025-12-26T10:54:31+00:00</updated><id>https://solmaz.io/x/2004506101362852125</id><content type="html" xml:base="https://solmaz.io/x/2004506101362852125/"><![CDATA[I gave Codex a task of porting an OpenCV tracking algorithm (CSRT) from C++ to Rust, so that I can directly use it in my project without having to cross-compile

It one-shot the task perfectly in 1hr, and even developed a GUI on top of it. All I did was to provide the original source and algo paper

I&#39;ve spent years getting specialized in writing numerical code (computational mechanics, fem), and now AI can automate 95% of the low-level grunt work

Acquiring these skills involved highly difficult, excruciating intellectual labor spanning many years, very similar to ML research. Doing tensor math, writing out the solver code, wondering why your solution is not converging, finally figuring out it was a sign typo after 2 days

Kids these days both have it easy and hard. They can fast forward large chunks of the work, but then they will never understand things as deeply as someone who wrote the whole thing by hand

I guess the more valuable skill now is being able to zoom in and out of abstraction levels quickly when needed. Using AI, but recognizing fast when it fails, learning what needs to be done, fixing it, zooming back out, repeat. Adaptive learning, a sort of &quot;depth-on-demand&quot;. The quicker you can pick up new skills and knowledge, the more successful you will be]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">See the repo for the latest changes: github.com/osolmaz/Simple…</title><link href="https://solmaz.io/x/2004334016170971511/" rel="alternate" type="text/html" title="See the repo for the latest changes: github.com/osolmaz/Simple…" /><published>2025-12-25T23:30:42+00:00</published><updated>2025-12-25T23:30:42+00:00</updated><id>https://solmaz.io/x/2004334016170971511</id><content type="html" xml:base="https://solmaz.io/x/2004334016170971511/"><![CDATA[See the repo for the latest changes: https://t.co/YDevGg2rhz]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If you have a bunch of docs in your repo, give it a try. It will use the timestamps of the...</title><link href="https://solmaz.io/x/2004332979993358516/" rel="alternate" type="text/html" title="If you have a bunch of docs in your repo, give it a try. It will use the timestamps of the..." /><published>2025-12-25T23:26:35+00:00</published><updated>2025-12-25T23:26:35+00:00</updated><id>https://solmaz.io/x/2004332979993358516</id><content type="html" xml:base="https://solmaz.io/x/2004332979993358516/"><![CDATA[If you have a bunch of docs in your repo, give it a try. It will use the timestamps of the commit that created the files while renaming. You can also run with --dry-run to see changes without applying them]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Now you can migrate your repo to SimpleDoc with a single command:</title><link href="https://solmaz.io/x/2004330867053707733/" rel="alternate" type="text/html" title="Now you can migrate your repo to SimpleDoc with a single command:" /><published>2025-12-25T23:18:12+00:00</published><updated>2025-12-25T23:18:12+00:00</updated><id>https://solmaz.io/x/2004330867053707733</id><content type="html" xml:base="https://solmaz.io/x/2004330867053707733/"><![CDATA[Now you can migrate your repo to SimpleDoc with a single command:

npx -y @simpledoc/simpledoc migrate

Step by step wizard will add timestamps to your files based on your git history, add missing YAML frontmatter, update your AGENTS md file
https://t.co/yrciS8KtEw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@bcherny Would be great if I could queue messages like in Codex</title><link href="https://solmaz.io/x/2004274503141269938/" rel="alternate" type="text/html" title="@bcherny Would be great if I could queue messages like in Codex" /><published>2025-12-25T19:34:13+00:00</published><updated>2025-12-25T19:34:13+00:00</updated><id>https://solmaz.io/x/2004274503141269938</id><content type="html" xml:base="https://solmaz.io/x/2004274503141269938/"><![CDATA[@bcherny Would be great if I could queue messages like in Codex
https://t.co/mC25gNKWo3]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It seems it&#39;s impossible to post something on Reddit these days, even when it is a pure text...</title><link href="https://solmaz.io/x/2003838921428578750/" rel="alternate" type="text/html" title="It seems it&#39;s impossible to post something on Reddit these days, even when it is a pure text..." /><published>2025-12-24T14:43:23+00:00</published><updated>2025-12-24T14:43:23+00:00</updated><id>https://solmaz.io/x/2003838921428578750</id><content type="html" xml:base="https://solmaz.io/x/2003838921428578750/"><![CDATA[It seems it&#39;s impossible to post something on Reddit these days, even when it is a pure text post without links in the body]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">X post</title><link href="https://solmaz.io/x/2003585112257024190/" rel="alternate" type="text/html" title="X post" /><published>2025-12-23T21:54:50+00:00</published><updated>2025-12-23T21:54:50+00:00</updated><id>https://solmaz.io/x/2003585112257024190</id><content type="html" xml:base="https://solmaz.io/x/2003585112257024190/"><![CDATA[]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Curious to hear what other hardcore agent users @simonw @mitsuhiko @steipete @badlogicgames...</title><link href="https://solmaz.io/x/2003262105697583414/" rel="alternate" type="text/html" title="Curious to hear what other hardcore agent users @simonw @mitsuhiko @steipete @badlogicgames..." /><published>2025-12-23T00:31:19+00:00</published><updated>2025-12-23T00:31:19+00:00</updated><id>https://solmaz.io/x/2003262105697583414</id><content type="html" xml:base="https://solmaz.io/x/2003262105697583414/"><![CDATA[Curious to hear what other hardcore agent users @simonw @mitsuhiko @steipete @badlogicgames  think. I can&#39;t be the only one who does this.

I feel like everybody ended up with the same workflow independent of each other, but somehow did not write about it (or I missed it)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">How to stop AI agents from littering your codebase with Markdown files?</title><link href="https://solmaz.io/x/2003262090132488209/" rel="alternate" type="text/html" title="How to stop AI agents from littering your codebase with Markdown files?" /><published>2025-12-23T00:31:15+00:00</published><updated>2025-12-23T00:31:15+00:00</updated><id>https://solmaz.io/x/2003262090132488209</id><content type="html" xml:base="https://solmaz.io/x/2003262090132488209/"><![CDATA[How to stop AI agents from littering your codebase with Markdown files?

I wrote a new post on how to create documentations with AI agents, without having it add markdown files in your repo root, and have chronological order to the files it creates]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">OpenAI won’t be able to monopolize this, the same reason Microsoft couldn’t monopolize the...</title><link href="https://solmaz.io/x/2002719070031147135/" rel="alternate" type="text/html" title="OpenAI won’t be able to monopolize this, the same reason Microsoft couldn’t monopolize the..." /><published>2025-12-21T12:33:29+00:00</published><updated>2025-12-21T12:33:29+00:00</updated><id>https://solmaz.io/x/2002719070031147135</id><content type="html" xml:base="https://solmaz.io/x/2002719070031147135/"><![CDATA[OpenAI won’t be able to monopolize this, the same reason Microsoft couldn’t monopolize the internet. The internet (of agents) is bigger than any one company]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">One tap @Revolut bank account at Berlin airport. Literally.</title><link href="https://solmaz.io/x/2002694629926351315/" rel="alternate" type="text/html" title="One tap @Revolut bank account at Berlin airport. Literally." /><published>2025-12-21T10:56:22+00:00</published><updated>2025-12-21T10:56:22+00:00</updated><id>https://solmaz.io/x/2002694629926351315</id><content type="html" xml:base="https://solmaz.io/x/2002694629926351315/"><![CDATA[One tap @Revolut bank account at Berlin airport. Literally.

Dispenses free card with instructions to login. One of the the most insane onboarding experiences I have ever seen]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Slop bombing</title><link href="https://solmaz.io/x/2002302622640845062/" rel="alternate" type="text/html" title="Slop bombing" /><published>2025-12-20T08:58:40+00:00</published><updated>2025-12-20T08:58:40+00:00</updated><id>https://solmaz.io/x/2002302622640845062</id><content type="html" xml:base="https://solmaz.io/x/2002302622640845062/"><![CDATA[Slop bombing]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex feature request: Let me queue up /model changes</title><link href="https://solmaz.io/x/2001704563984744489/" rel="alternate" type="text/html" title="Codex feature request: Let me queue up /model changes" /><published>2025-12-18T17:22:12+00:00</published><updated>2025-12-18T17:22:12+00:00</updated><id>https://solmaz.io/x/2001704563984744489</id><content type="html" xml:base="https://solmaz.io/x/2001704563984744489/"><![CDATA[Codex feature request: Let me queue up /model changes 

Currently, if I try to run /model while responding, it tells me that I can&#39;t do that while the model is responding

But I often want to gauge thinking budget in advance, like run a straightforward task with low reasoning and then start another one with high reasoning

cc @thsottiaux]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Literally the exact same thing happened to me back in 2018. Everybody learns not to use...</title><link href="https://solmaz.io/x/2001690553251967258/" rel="alternate" type="text/html" title="Literally the exact same thing happened to me back in 2018. Everybody learns not to use..." /><published>2025-12-18T16:26:32+00:00</published><updated>2025-12-18T16:26:32+00:00</updated><id>https://solmaz.io/x/2001690553251967258</id><content type="html" xml:base="https://solmaz.io/x/2001690553251967258/"><![CDATA[Literally the exact same thing happened to me back in 2018. Everybody learns not to use password auth with SSH the hard way
https://t.co/NPqrXwqUUy]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI agents make any transductional task (like translation from language A to language B)...</title><link href="https://solmaz.io/x/2001398671091327073/" rel="alternate" type="text/html" title="AI agents make any transductional task (like translation from language A to language B)..." /><published>2025-12-17T21:06:42+00:00</published><updated>2025-12-17T21:06:42+00:00</updated><id>https://solmaz.io/x/2001398671091327073</id><content type="html" xml:base="https://solmaz.io/x/2001398671091327073/"><![CDATA[AI agents make any transductional task (like translation from language A to language B) trivial, especially when you can verify the output with compilers and tests

The bottleneck is now curating the tests]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I think X removed one of my posts yesterday about the new encrypted &quot;Chat&quot; rolling out to all...</title><link href="https://solmaz.io/x/2001201245650894874/" rel="alternate" type="text/html" title="I think X removed one of my posts yesterday about the new encrypted &quot;Chat&quot; rolling out to all..." /><published>2025-12-17T08:02:12+00:00</published><updated>2025-12-17T08:02:12+00:00</updated><id>https://solmaz.io/x/2001201245650894874</id><content type="html" xml:base="https://solmaz.io/x/2001201245650894874/"><![CDATA[I think X removed one of my posts yesterday about the new encrypted &quot;Chat&quot; rolling out to all users, and how you might lose all your past messages if you forget your passcode and do not have the app installed

I can swear I clicked Post. Do they classify posts based on their topic and delete the ones they don&#39;t like?

Anyway, we shall see, I am taking a screenshot and saving the URL.]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">very optimistic!</title><link href="https://solmaz.io/x/2000144553269965055/" rel="alternate" type="text/html" title="very optimistic!" /><published>2025-12-14T10:03:17+00:00</published><updated>2025-12-14T10:03:17+00:00</updated><id>https://solmaz.io/x/2000144553269965055</id><content type="html" xml:base="https://solmaz.io/x/2000144553269965055/"><![CDATA[very optimistic!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Crazy that @cursor_ai disabled Gemini 3 Pro on my installation, toggled it right back on. I...</title><link href="https://solmaz.io/x/2000132431739801615/" rel="alternate" type="text/html" title="Crazy that @cursor_ai disabled Gemini 3 Pro on my installation, toggled it right back on. I..." /><published>2025-12-14T09:15:07+00:00</published><updated>2025-12-14T09:15:07+00:00</updated><id>https://solmaz.io/x/2000132431739801615</id><content type="html" xml:base="https://solmaz.io/x/2000132431739801615/"><![CDATA[Crazy that @cursor_ai disabled Gemini 3 Pro on my installation, toggled it right back on. I wonder why, too many complaints maybe? That it’s hard to control?

On another note, disabling models without notification is dishonest product behavior. I would at least appreciate getting a notification, even when it might be against a company’s interests  @sualehasif996]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Language-agnostic interoperability layer for LLM APIs</title><link href="https://solmaz.io/x/2000121028228358604/" rel="alternate" type="text/html" title="Language-agnostic interoperability layer for LLM APIs" /><published>2025-12-14T08:29:48+00:00</published><updated>2025-12-14T08:29:48+00:00</updated><id>https://solmaz.io/x/2000121028228358604</id><content type="html" xml:base="https://solmaz.io/x/2000121028228358604/"><![CDATA[So is somebody already building “LLVM but for LLM APIs” in stealth or not? 

We have numerous libraries @langchain, Vercel AI SDK, LiteLLM, OpenRouter, the one we have built at @TextCortex, etc.

But to my knowledge, none of these try to build a language agnostic IR for interoperability between providers (or at least market themselves as such)

Like some standard and set of tools that will not lock you in langchain, ai sdk or anything like that, something lower level and less opinionated

I feel like this is a job for the new Agentic AI Foundation cc @linuxfoundation, so maybe they are already working on it? I desperately want to start on such a project, but feel like I might get sniped 2 months after

Anybody has any information on all this?

cc @mitsuhiko @badlogicgames @steipete]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">For those wondering what project this is: horse.fit</title><link href="https://solmaz.io/x/1999883882989244900/" rel="alternate" type="text/html" title="For those wondering what project this is: horse.fit" /><published>2025-12-13T16:47:28+00:00</published><updated>2025-12-13T16:47:28+00:00</updated><id>https://solmaz.io/x/1999883882989244900</id><content type="html" xml:base="https://solmaz.io/x/1999883882989244900/"><![CDATA[For those wondering what project this is: https://t.co/AzNS631PIC]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is how an agentic monorepo looks like. What was now a hurdle before is now a child&#39;s toy</title><link href="https://solmaz.io/x/1999883880493441168/" rel="alternate" type="text/html" title="This is how an agentic monorepo looks like. What was now a hurdle before is now a child&#39;s toy" /><published>2025-12-13T16:47:27+00:00</published><updated>2025-12-13T16:47:27+00:00</updated><id>https://solmaz.io/x/1999883880493441168</id><content type="html" xml:base="https://solmaz.io/x/1999883880493441168/"><![CDATA[This is how an agentic monorepo looks like. What was now a hurdle before is now a child&#39;s toy

This side project started as a Python project earlier in 2025
Then I added an iOS app on top of it
I rewrote the most important algorithms in Rust
I rewrote the entire backend in Go and retired Python to be used purely for prototypes
I wrote a webapp with Next.js
With unit and integration tests for each component
Lately written 99% by instructing agents
Crazy mixed language programming going on in the background. Rust component used both by iOS app for offline and by go backend for online use case, FFI and all
Number of lines in the repo: a couple 100k

If you had told me I would be able do to all of this by myself 1 year ago, I would not have believed it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is huge. Natively supported stacked PRs on GitHub would make life much easier, especially...</title><link href="https://solmaz.io/x/1999786569553596492/" rel="alternate" type="text/html" title="This is huge. Natively supported stacked PRs on GitHub would make life much easier, especially..." /><published>2025-12-13T10:20:47+00:00</published><updated>2025-12-13T10:20:47+00:00</updated><id>https://solmaz.io/x/1999786569553596492</id><content type="html" xml:base="https://solmaz.io/x/1999786569553596492/"><![CDATA[This is huge. Natively supported stacked PRs on GitHub would make life much easier, especially with human AND AI reviews

AI reviews with Codex/Claude/Gemini/Cursor Bugbot integrations are becoming especially important in small teams who are generating huge amounts of code

AI reviews don&#39;t work well if you don&#39;t split your work to diffs smaller than a few hundred lines of code, so stacked PRs are already an integral part of developer experience in agentic workflows]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Read more on my blog post solmaz.io/agentic-coding…</title><link href="https://solmaz.io/x/1999577957363241363/" rel="alternate" type="text/html" title="Read more on my blog post solmaz.io/agentic-coding…" /><published>2025-12-12T20:31:50+00:00</published><updated>2025-12-12T20:31:50+00:00</updated><id>https://solmaz.io/x/1999577957363241363</id><content type="html" xml:base="https://solmaz.io/x/1999577957363241363/"><![CDATA[Read more on my blog post https://t.co/uzCcOXuadB]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">CLI coding tools should give more control over message queueing</title><link href="https://solmaz.io/x/1999577947066323141/" rel="alternate" type="text/html" title="CLI coding tools should give more control over message queueing" /><published>2025-12-12T20:31:47+00:00</published><updated>2025-12-12T20:31:47+00:00</updated><id>https://solmaz.io/x/1999577947066323141</id><content type="html" xml:base="https://solmaz.io/x/1999577947066323141/"><![CDATA[CLI coding tools should give more control over message queueing

Codex waits until end of turn to handle user message, Claude Code injects as soon as possible after tool response/assistant reply

There is no reason why we cannot have both!

New post (link below):]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Codex v0.71 finally implements a more detailed way of storing permissions</title><link href="https://solmaz.io/x/1999444307431030935/" rel="alternate" type="text/html" title="Codex v0.71 finally implements a more detailed way of storing permissions" /><published>2025-12-12T11:40:45+00:00</published><updated>2025-12-12T11:40:45+00:00</updated><id>https://solmaz.io/x/1999444307431030935</id><content type="html" xml:base="https://solmaz.io/x/1999444307431030935/"><![CDATA[Codex v0.71 finally implements a more detailed way of storing permissions

But they are still at user home folder level. Saving rules in a repo still seems TBD

&quot;execpolicy commands are still in preview. The API may have breaking changes in the future.&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">this is outrageous</title><link href="https://solmaz.io/x/1999058628637008227/" rel="alternate" type="text/html" title="this is outrageous" /><published>2025-12-11T10:08:12+00:00</published><updated>2025-12-11T10:08:12+00:00</updated><id>https://solmaz.io/x/1999058628637008227</id><content type="html" xml:base="https://solmaz.io/x/1999058628637008227/"><![CDATA[this is outrageous]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">At least some people at OpenAI must be thinking about buying @astral_sh</title><link href="https://solmaz.io/x/1996076438848680202/" rel="alternate" type="text/html" title="At least some people at OpenAI must be thinking about buying @astral_sh" /><published>2025-12-03T04:38:02+00:00</published><updated>2025-12-03T04:38:02+00:00</updated><id>https://solmaz.io/x/1996076438848680202</id><content type="html" xml:base="https://solmaz.io/x/1996076438848680202/"><![CDATA[At least some people at OpenAI must be thinking about buying @astral_sh]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">who remembers search engine aggregators from early 2000s?</title><link href="https://solmaz.io/x/1994811178833551867/" rel="alternate" type="text/html" title="who remembers search engine aggregators from early 2000s?" /><published>2025-11-29T16:50:21+00:00</published><updated>2025-11-29T16:50:21+00:00</updated><id>https://solmaz.io/x/1994811178833551867</id><content type="html" xml:base="https://solmaz.io/x/1994811178833551867/"><![CDATA[who remembers search engine aggregators from early 2000s?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My initial experience with Claude Opus 4.5 is that it’s much better than previous Anthropic...</title><link href="https://solmaz.io/x/1993481331821711554/" rel="alternate" type="text/html" title="My initial experience with Claude Opus 4.5 is that it’s much better than previous Anthropic..." /><published>2025-11-26T00:46:01+00:00</published><updated>2025-11-26T00:46:01+00:00</updated><id>https://solmaz.io/x/1993481331821711554</id><content type="html" xml:base="https://solmaz.io/x/1993481331821711554/"><![CDATA[My initial experience with Claude Opus 4.5 is that it’s much better than previous Anthropic models, but it’s still relatively unreliable and hallucinates. It feels lagging in reasoning compared to highest OpenAI and Google lineup of models]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">wow twitter/x just doxxed the countries of all the anons on this platform</title><link href="https://solmaz.io/x/1992807182237348012/" rel="alternate" type="text/html" title="wow twitter/x just doxxed the countries of all the anons on this platform" /><published>2025-11-24T04:07:11+00:00</published><updated>2025-11-24T04:07:11+00:00</updated><id>https://solmaz.io/x/1992807182237348012</id><content type="html" xml:base="https://solmaz.io/x/1992807182237348012/"><![CDATA[wow twitter/x just doxxed the countries of all the anons on this platform]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">the real advantage of Gemini 3 Pro is speed. it delivers accuracy higher than GPT 5 and...</title><link href="https://solmaz.io/x/1992412959063323066/" rel="alternate" type="text/html" title="the real advantage of Gemini 3 Pro is speed. it delivers accuracy higher than GPT 5 and..." /><published>2025-11-23T02:00:41+00:00</published><updated>2025-11-23T02:00:41+00:00</updated><id>https://solmaz.io/x/1992412959063323066</id><content type="html" xml:base="https://solmaz.io/x/1992412959063323066/"><![CDATA[the real advantage of Gemini 3 Pro is speed. it delivers accuracy higher than GPT 5 and sometimes GPT 5 Pro at a much higher speed. the long tail of developers value speed over accuracy, so it looks like it will take over as the main coding model for most ppl]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">There already is one: google-github-actions/run-gemini-cli</title><link href="https://solmaz.io/x/1991090631012471132/" rel="alternate" type="text/html" title="There already is one: google-github-actions/run-gemini-cli" /><published>2025-11-19T10:26:13+00:00</published><updated>2025-11-19T10:26:13+00:00</updated><id>https://solmaz.io/x/1991090631012471132</id><content type="html" xml:base="https://solmaz.io/x/1991090631012471132/"><![CDATA[There already is one: google-github-actions/run-gemini-cli

but updated last week, not sure if this supports gemini 3 pro]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Gemini seems to be very good at debugging/reviewing/finding root causes. A GitHub...</title><link href="https://solmaz.io/x/1991066906187403641/" rel="alternate" type="text/html" title="Gemini seems to be very good at debugging/reviewing/finding root causes. A GitHub..." /><published>2025-11-19T08:51:57+00:00</published><updated>2025-11-19T08:51:57+00:00</updated><id>https://solmaz.io/x/1991066906187403641</id><content type="html" xml:base="https://solmaz.io/x/1991066906187403641/"><![CDATA[Gemini seems to be very good at debugging/reviewing/finding root causes. A GitHub action/integration in PRs would be very useful!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">(This is sarcasm for those who can’t tell)</title><link href="https://solmaz.io/x/1991045121014706604/" rel="alternate" type="text/html" title="(This is sarcasm for those who can’t tell)" /><published>2025-11-19T07:25:23+00:00</published><updated>2025-11-19T07:25:23+00:00</updated><id>https://solmaz.io/x/1991045121014706604</id><content type="html" xml:base="https://solmaz.io/x/1991045121014706604/"><![CDATA[(This is sarcasm for those who can’t tell)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Google is making progress… I did not have to request access on Vertex AI for Gemini 3 Pro this...</title><link href="https://solmaz.io/x/1991045118607155286/" rel="alternate" type="text/html" title="Google is making progress… I did not have to request access on Vertex AI for Gemini 3 Pro this..." /><published>2025-11-19T07:25:22+00:00</published><updated>2025-11-19T07:25:22+00:00</updated><id>https://solmaz.io/x/1991045118607155286</id><content type="html" xml:base="https://solmaz.io/x/1991045118607155286/"><![CDATA[Google is making progress… I did not have to request access on Vertex AI for Gemini 3 Pro this time to deploy it to @TextCortex]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">tip for testing new model releases: “they say you are sota. prove it”</title><link href="https://solmaz.io/x/1990939304206671945/" rel="alternate" type="text/html" title="tip for testing new model releases: “they say you are sota. prove it”" /><published>2025-11-19T00:24:54+00:00</published><updated>2025-11-19T00:24:54+00:00</updated><id>https://solmaz.io/x/1990939304206671945</id><content type="html" xml:base="https://solmaz.io/x/1990939304206671945/"><![CDATA[tip for testing new model releases: “they say you are sota. prove it”]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">&quot;The more a task/job is verifiable, the more amenable it is to automation in the new...</title><link href="https://solmaz.io/x/1990302419276931479/" rel="alternate" type="text/html" title="&quot;The more a task/job is verifiable, the more amenable it is to automation in the new..." /><published>2025-11-17T06:14:09+00:00</published><updated>2025-11-17T06:14:09+00:00</updated><id>https://solmaz.io/x/1990302419276931479</id><content type="html" xml:base="https://solmaz.io/x/1990302419276931479/"><![CDATA[&quot;The more a task/job is verifiable, the more amenable it is to automation in the new programming paradigm. If it is not verifiable, it has to fall out from neural net magic of generalization fingers crossed, or via weaker means like imitation.&quot;]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Most important note of the new @OpenAI gpt 5.1 update</title><link href="https://solmaz.io/x/1989182407929745545/" rel="alternate" type="text/html" title="Most important note of the new @OpenAI gpt 5.1 update" /><published>2025-11-14T04:03:37+00:00</published><updated>2025-11-14T04:03:37+00:00</updated><id>https://solmaz.io/x/1989182407929745545</id><content type="html" xml:base="https://solmaz.io/x/1989182407929745545/"><![CDATA[Most important note of the new @OpenAI gpt 5.1 update

big improvement on unit economics]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This post makes no sense</title><link href="https://solmaz.io/x/1987148854534262907/" rel="alternate" type="text/html" title="This post makes no sense" /><published>2025-11-08T13:23:01+00:00</published><updated>2025-11-08T13:23:01+00:00</updated><id>https://solmaz.io/x/1987148854534262907</id><content type="html" xml:base="https://solmaz.io/x/1987148854534262907/"><![CDATA[This post makes no sense

Please consider again and look at @cloudfleet_k8s. You might regret your decision]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Working on observability is underrated</title><link href="https://solmaz.io/x/1985025400461029585/" rel="alternate" type="text/html" title="Working on observability is underrated" /><published>2025-11-02T16:45:10+00:00</published><updated>2025-11-02T16:45:10+00:00</updated><id>https://solmaz.io/x/1985025400461029585</id><content type="html" xml:base="https://solmaz.io/x/1985025400461029585/"><![CDATA[Working on observability is underrated]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@rakyll @GergelyOrosz Should scrape some austrian websites :)</title><link href="https://solmaz.io/x/1984595653956096321/" rel="alternate" type="text/html" title="@rakyll @GergelyOrosz Should scrape some austrian websites :)" /><published>2025-11-01T12:17:30+00:00</published><updated>2025-11-01T12:17:30+00:00</updated><id>https://solmaz.io/x/1984595653956096321</id><content type="html" xml:base="https://solmaz.io/x/1984595653956096321/"><![CDATA[@rakyll @GergelyOrosz Should scrape some austrian websites :)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I was schooled by AI today... tbh I deserved it</title><link href="https://solmaz.io/x/1982916994690363616/" rel="alternate" type="text/html" title="I was schooled by AI today... tbh I deserved it" /><published>2025-10-27T21:07:07+00:00</published><updated>2025-10-27T21:07:07+00:00</updated><id>https://solmaz.io/x/1982916994690363616</id><content type="html" xml:base="https://solmaz.io/x/1982916994690363616/"><![CDATA[I was schooled by AI today... tbh I deserved it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@thsottiaux Let me use my Pro/Plus plans in Codex GH Action</title><link href="https://solmaz.io/x/1976983794147295250/" rel="alternate" type="text/html" title="@thsottiaux Let me use my Pro/Plus plans in Codex GH Action" /><published>2025-10-11T12:10:41+00:00</published><updated>2025-10-11T12:10:41+00:00</updated><id>https://solmaz.io/x/1976983794147295250</id><content type="html" xml:base="https://solmaz.io/x/1976983794147295250/"><![CDATA[@thsottiaux Let me use my Pro/Plus plans in Codex GH Action
https://t.co/0Fw1rLmCED]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">TIL @OpenAI now has a GitHub action for Codex, similar to Claude Code</title><link href="https://solmaz.io/x/1976983381289451883/" rel="alternate" type="text/html" title="TIL @OpenAI now has a GitHub action for Codex, similar to Claude Code" /><published>2025-10-11T12:09:03+00:00</published><updated>2025-10-11T12:09:03+00:00</updated><id>https://solmaz.io/x/1976983381289451883</id><content type="html" xml:base="https://solmaz.io/x/1976983381289451883/"><![CDATA[TIL @OpenAI now has a GitHub action for Codex, similar to Claude Code

This lets you invoke Codex in a more controlled way in your repos

You must still pay API prices though. Let&#39;s see if OpenAI will introduce a way to connect your Pro plan, like in @AnthropicAI paid plans]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Nice way of putting it</title><link href="https://solmaz.io/x/1969321668725158311/" rel="alternate" type="text/html" title="Nice way of putting it" /><published>2025-09-20T08:44:08+00:00</published><updated>2025-09-20T08:44:08+00:00</updated><id>https://solmaz.io/x/1969321668725158311</id><content type="html" xml:base="https://solmaz.io/x/1969321668725158311/"><![CDATA[Nice way of putting it]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Just downgraded my anthropic sub, 200 usd openai plan is finally justified after 1 year</title><link href="https://solmaz.io/x/1966797279945204059/" rel="alternate" type="text/html" title="Just downgraded my anthropic sub, 200 usd openai plan is finally justified after 1 year" /><published>2025-09-13T09:33:07+00:00</published><updated>2025-09-13T09:33:07+00:00</updated><id>https://solmaz.io/x/1966797279945204059</id><content type="html" xml:base="https://solmaz.io/x/1966797279945204059/"><![CDATA[Just downgraded my anthropic sub, 200 usd openai plan is finally justified after 1 year]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Enjoying gpt5 codex free lunch before openai inevitably starts cutting corners just like...</title><link href="https://solmaz.io/x/1966560391024308282/" rel="alternate" type="text/html" title="Enjoying gpt5 codex free lunch before openai inevitably starts cutting corners just like..." /><published>2025-09-12T17:51:48+00:00</published><updated>2025-09-12T17:51:48+00:00</updated><id>https://solmaz.io/x/1966560391024308282</id><content type="html" xml:base="https://solmaz.io/x/1966560391024308282/"><![CDATA[Enjoying gpt5 codex free lunch before openai inevitably starts cutting corners just like anthropic

(I hope to be wrong in 3 months, this model is very good at one shotting things and I don’t want it to be nerfed)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@thsottiaux 2/ Model just stops working on a task even though I tell it to run something and...</title><link href="https://solmaz.io/x/1964942275278192834/" rel="alternate" type="text/html" title="@thsottiaux 2/ Model just stops working on a task even though I tell it to run something and..." /><published>2025-09-08T06:41:59+00:00</published><updated>2025-09-08T06:41:59+00:00</updated><id>https://solmaz.io/x/1964942275278192834</id><content type="html" xml:base="https://solmaz.io/x/1964942275278192834/"><![CDATA[@thsottiaux 2/ Model just stops working on a task even though I tell it to run something and not stop until it works. I have to frequently say “ok do it then”. Probably a model problem and not harness problem]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">So let me get this straight, the main reason for Responses API exists is that OpenAI doesn’t...</title><link href="https://solmaz.io/x/1964010520643543115/" rel="alternate" type="text/html" title="So let me get this straight, the main reason for Responses API exists is that OpenAI doesn’t..." /><published>2025-09-05T16:59:32+00:00</published><updated>2025-09-05T16:59:32+00:00</updated><id>https://solmaz.io/x/1964010520643543115</id><content type="html" xml:base="https://solmaz.io/x/1964010520643543115/"><![CDATA[So let me get this straight, the main reason for Responses API exists is that OpenAI doesn’t want to show reasoning traces? Therefore the whole world should bend backwards to fit your obscurantist standards? Responses will not get adopted the same reason Windows Server didn’t]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">gpt5 did such and such on that bench, oh it didn&#39;t even surpass grok 4 on arc-agi... bro did...</title><link href="https://solmaz.io/x/1954140134518980967/" rel="alternate" type="text/html" title="gpt5 did such and such on that bench, oh it didn&#39;t even surpass grok 4 on arc-agi... bro did..." /><published>2025-08-09T11:18:08+00:00</published><updated>2025-08-09T11:18:08+00:00</updated><id>https://solmaz.io/x/1954140134518980967</id><content type="html" xml:base="https://solmaz.io/x/1954140134518980967/"><![CDATA[gpt5 did such and such on that bench, oh it didn&#39;t even surpass grok 4 on arc-agi... bro did you even look at the price?

openai pushed the parento frontier hard with this one. I don&#39;t care that it doesn&#39;t know 4.11 &amp;lt; 4.9]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I converted this thread to a blog post and it hit HN front page</title><link href="https://solmaz.io/x/1952285810763427879/" rel="alternate" type="text/html" title="I converted this thread to a blog post and it hit HN front page" /><published>2025-08-04T08:29:43+00:00</published><updated>2025-08-04T08:29:43+00:00</updated><id>https://solmaz.io/x/1952285810763427879</id><content type="html" xml:base="https://solmaz.io/x/1952285810763427879/"><![CDATA[I converted this thread to a blog post and it hit HN front page]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Because of this, I predict a decrease in Python adoption in companies, specifically for...</title><link href="https://solmaz.io/x/1952133427076952261/" rel="alternate" type="text/html" title="Because of this, I predict a decrease in Python adoption in companies, specifically for..." /><published>2025-08-03T22:24:12+00:00</published><updated>2025-08-03T22:24:12+00:00</updated><id>https://solmaz.io/x/1952133427076952261</id><content type="html" xml:base="https://solmaz.io/x/1952133427076952261/"><![CDATA[Because of this, I predict a decrease in Python adoption in companies, specifically for production deployments, even though I like it so much]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">My &amp;gt;10 yr old programming habits have changed since Claude Code launched. Python is less...</title><link href="https://solmaz.io/x/1952133418046660609/" rel="alternate" type="text/html" title="My &amp;gt;10 yr old programming habits have changed since Claude Code launched. Python is less..." /><published>2025-08-03T22:24:10+00:00</published><updated>2025-08-03T22:24:10+00:00</updated><id>https://solmaz.io/x/1952133418046660609</id><content type="html" xml:base="https://solmaz.io/x/1952133418046660609/"><![CDATA[My &amp;gt;10 yr old programming habits have changed since Claude Code launched. Python is less likely to be my go-to language for new projects anymore. I am managing projects in languages I am not fluent in---TypeScript, Rust and Go---and seem to be doing pretty well]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It seems that typed, compiled, etc. languages are more suited for vibe coding, because of the...</title><link href="https://solmaz.io/x/1952133419938251018/" rel="alternate" type="text/html" title="It seems that typed, compiled, etc. languages are more suited for vibe coding, because of the..." /><published>2025-08-03T22:24:10+00:00</published><updated>2025-08-03T22:24:10+00:00</updated><id>https://solmaz.io/x/1952133419938251018</id><content type="html" xml:base="https://solmaz.io/x/1952133419938251018/"><![CDATA[It seems that typed, compiled, etc. languages are more suited for vibe coding, because of the safety guarantees. This is unsurprising in hindsight, but it was counterintuitive because by default I &quot;vibed&quot; projects into existence in Python for as long as I can remember]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Lol should I go back to computational mechanics</title><link href="https://solmaz.io/x/1949049783080849457/" rel="alternate" type="text/html" title="Lol should I go back to computational mechanics" /><published>2025-07-26T10:10:54+00:00</published><updated>2025-07-26T10:10:54+00:00</updated><id>https://solmaz.io/x/1949049783080849457</id><content type="html" xml:base="https://solmaz.io/x/1949049783080849457/"><![CDATA[Lol should I go back to computational mechanics]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;ve just upgraded @nikitabobko Aerospace from v15 to v19, and I can say it&#39;s WAY MORE faster...</title><link href="https://solmaz.io/x/1947258400489845247/" rel="alternate" type="text/html" title="I&#39;ve just upgraded @nikitabobko Aerospace from v15 to v19, and I can say it&#39;s WAY MORE faster..." /><published>2025-07-21T11:32:35+00:00</published><updated>2025-07-21T11:32:35+00:00</updated><id>https://solmaz.io/x/1947258400489845247</id><content type="html" xml:base="https://solmaz.io/x/1947258400489845247/"><![CDATA[I&#39;ve just upgraded @nikitabobko Aerospace from v15 to v19, and I can say it&#39;s WAY MORE faster. Friendly reminder that you might be running an old version as well. Thank you @nikitabobko !]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Anthropic is not god</title><link href="https://solmaz.io/x/1945767025084662150/" rel="alternate" type="text/html" title="Anthropic is not god" /><published>2025-07-17T08:46:24+00:00</published><updated>2025-07-17T08:46:24+00:00</updated><id>https://solmaz.io/x/1945767025084662150</id><content type="html" xml:base="https://solmaz.io/x/1945767025084662150/"><![CDATA[Anthropic is not god]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I can testify</title><link href="https://solmaz.io/x/1945605753185644738/" rel="alternate" type="text/html" title="I can testify" /><published>2025-07-16T22:05:33+00:00</published><updated>2025-07-16T22:05:33+00:00</updated><id>https://solmaz.io/x/1945605753185644738</id><content type="html" xml:base="https://solmaz.io/x/1945605753185644738/"><![CDATA[I can testify]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I found an OK-ish solution to Claude Code running python instead of uv in first try cc...</title><link href="https://solmaz.io/x/1944430364232909185/" rel="alternate" type="text/html" title="I found an OK-ish solution to Claude Code running python instead of uv in first try cc..." /><published>2025-07-13T16:14:59+00:00</published><updated>2025-07-13T16:14:59+00:00</updated><id>https://solmaz.io/x/1944430364232909185</id><content type="html" xml:base="https://solmaz.io/x/1944430364232909185/"><![CDATA[I found an OK-ish solution to Claude Code running python instead of uv in first try cc @mitsuhiko @simonw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Read more here: solmaz.io/log/2025/07/13…</title><link href="https://solmaz.io/x/1944430366661460198/" rel="alternate" type="text/html" title="Read more here: solmaz.io/log/2025/07/13…" /><published>2025-07-13T16:14:59+00:00</published><updated>2025-07-13T16:14:59+00:00</updated><id>https://solmaz.io/x/1944430366661460198</id><content type="html" xml:base="https://solmaz.io/x/1944430366661460198/"><![CDATA[Read more here: https://t.co/Kc2PtyvhPY]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">What is the current best way to make Claude Code use uv run instead of python?</title><link href="https://solmaz.io/x/1941818382182916491/" rel="alternate" type="text/html" title="What is the current best way to make Claude Code use uv run instead of python?" /><published>2025-07-06T11:15:54+00:00</published><updated>2025-07-06T11:15:54+00:00</updated><id>https://solmaz.io/x/1941818382182916491</id><content type="html" xml:base="https://solmaz.io/x/1941818382182916491/"><![CDATA[What is the current best way to make Claude Code use uv run instead of python?

I have added instructions to CLAUDE md, but it still calls python for the first time, then corrects to uv run

It must be happening to so many people now, so many tokens wasted
cc @mitsuhiko @simonw]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Coding is dead</title><link href="https://solmaz.io/x/1940374584735440981/" rel="alternate" type="text/html" title="Coding is dead" /><published>2025-07-02T11:38:46+00:00</published><updated>2025-07-02T11:38:46+00:00</updated><id>https://solmaz.io/x/1940374584735440981</id><content type="html" xml:base="https://solmaz.io/x/1940374584735440981/"><![CDATA[Coding is dead

Programming is all what is left now]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">+1</title><link href="https://solmaz.io/x/1937867758214984143/" rel="alternate" type="text/html" title="+1" /><published>2025-06-25T13:37:32+00:00</published><updated>2025-06-25T13:37:32+00:00</updated><id>https://solmaz.io/x/1937867758214984143</id><content type="html" xml:base="https://solmaz.io/x/1937867758214984143/"><![CDATA[+1]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">When you tell Claude Code to ultrathink</title><link href="https://solmaz.io/x/1933159368737829200/" rel="alternate" type="text/html" title="When you tell Claude Code to ultrathink" /><published>2025-06-12T13:48:04+00:00</published><updated>2025-06-12T13:48:04+00:00</updated><id>https://solmaz.io/x/1933159368737829200</id><content type="html" xml:base="https://solmaz.io/x/1933159368737829200/"><![CDATA[When you tell Claude Code to ultrathink]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">How does Codex adoption compare to Claude Code? I just unleashed deep research on it and it...</title><link href="https://solmaz.io/x/1931071583696646246/" rel="alternate" type="text/html" title="How does Codex adoption compare to Claude Code? I just unleashed deep research on it and it..." /><published>2025-06-06T19:31:57+00:00</published><updated>2025-06-06T19:31:57+00:00</updated><id>https://solmaz.io/x/1931071583696646246</id><content type="html" xml:base="https://solmaz.io/x/1931071583696646246/"><![CDATA[How does Codex adoption compare to Claude Code? I just unleashed deep research on it and it says Codex is more popular, but that goes completely against my guesses]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Using Claude Code to reverse engineer Claude Code 🤝</title><link href="https://solmaz.io/x/1930539178208432601/" rel="alternate" type="text/html" title="Using Claude Code to reverse engineer Claude Code 🤝" /><published>2025-06-05T08:16:22+00:00</published><updated>2025-06-05T08:16:22+00:00</updated><id>https://solmaz.io/x/1930539178208432601</id><content type="html" xml:base="https://solmaz.io/x/1930539178208432601/"><![CDATA[Using Claude Code to reverse engineer Claude Code 🤝]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Wrote about this earlier this year!</title><link href="https://solmaz.io/x/1930370996273311765/" rel="alternate" type="text/html" title="Wrote about this earlier this year!" /><published>2025-06-04T21:08:04+00:00</published><updated>2025-06-04T21:08:04+00:00</updated><id>https://solmaz.io/x/1930370996273311765</id><content type="html" xml:base="https://solmaz.io/x/1930370996273311765/"><![CDATA[Wrote about this earlier this year!

https://t.co/tgAv6Lk34L]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Working with LLMs is definitely an art and not science</title><link href="https://solmaz.io/x/1929414713139450149/" rel="alternate" type="text/html" title="Working with LLMs is definitely an art and not science" /><published>2025-06-02T05:48:09+00:00</published><updated>2025-06-02T05:48:09+00:00</updated><id>https://solmaz.io/x/1929414713139450149</id><content type="html" xml:base="https://solmaz.io/x/1929414713139450149/"><![CDATA[Working with LLMs is definitely an art and not science]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Just some thoughts after using Claude Code intensively for 1 week 👆</title><link href="https://solmaz.io/x/1929140563309191241/" rel="alternate" type="text/html" title="Just some thoughts after using Claude Code intensively for 1 week 👆" /><published>2025-06-01T11:38:46+00:00</published><updated>2025-06-01T11:38:46+00:00</updated><id>https://solmaz.io/x/1929140563309191241</id><content type="html" xml:base="https://solmaz.io/x/1929140563309191241/"><![CDATA[Just some thoughts after using Claude Code intensively for 1 week 👆]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The models, they just wanna work. They want to build your product, fix your bugs, serve your...</title><link href="https://solmaz.io/x/1929140559668515195/" rel="alternate" type="text/html" title="The models, they just wanna work. They want to build your product, fix your bugs, serve your..." /><published>2025-06-01T11:38:45+00:00</published><updated>2025-06-01T11:38:45+00:00</updated><id>https://solmaz.io/x/1929140559668515195</id><content type="html" xml:base="https://solmaz.io/x/1929140559668515195/"><![CDATA[The models, they just wanna work. They want to build your product, fix your bugs, serve your users. You feed them the right context, give them good tools. You don’t assume what they cannot do without trying, and you don’t prematurely constrain them into deterministic workflows.]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">solmaz.io/log/2025/05/31…</title><link href="https://solmaz.io/x/1928654335686144322/" rel="alternate" type="text/html" title="solmaz.io/log/2025/05/31…" /><published>2025-05-31T03:26:40+00:00</published><updated>2025-05-31T03:26:40+00:00</updated><id>https://solmaz.io/x/1928654335686144322</id><content type="html" xml:base="https://solmaz.io/x/1928654335686144322/"><![CDATA[https://t.co/5SuA4cYddE]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">SCP-3434: Istanbul Taxi Superorganism</title><link href="https://solmaz.io/x/1928654272910008711/" rel="alternate" type="text/html" title="SCP-3434: Istanbul Taxi Superorganism" /><published>2025-05-31T03:26:25+00:00</published><updated>2025-05-31T03:26:25+00:00</updated><id>https://solmaz.io/x/1928654272910008711</id><content type="html" xml:base="https://solmaz.io/x/1928654272910008711/"><![CDATA[SCP-3434: Istanbul Taxi Superorganism]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is not an overstatement</title><link href="https://solmaz.io/x/1928037892099547640/" rel="alternate" type="text/html" title="This is not an overstatement" /><published>2025-05-29T10:37:09+00:00</published><updated>2025-05-29T10:37:09+00:00</updated><id>https://solmaz.io/x/1928037892099547640</id><content type="html" xml:base="https://solmaz.io/x/1928037892099547640/"><![CDATA[This is not an overstatement]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Headless makes running these things in a sandbox much easier. Sandbox means you can give all...</title><link href="https://solmaz.io/x/1927995944374509845/" rel="alternate" type="text/html" title="Headless makes running these things in a sandbox much easier. Sandbox means you can give all..." /><published>2025-05-29T07:50:28+00:00</published><updated>2025-05-29T07:50:28+00:00</updated><id>https://solmaz.io/x/1927995944374509845</id><content type="html" xml:base="https://solmaz.io/x/1927995944374509845/"><![CDATA[Headless makes running these things in a sandbox much easier. Sandbox means you can give all permissions and just let it run until completion. See my efforts to do so here:

https://t.co/9UUFOTwk2A]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I was trying to figure out why @AnthropicAI Claude Code feels better than @cursor_ai with Opus...</title><link href="https://solmaz.io/x/1927995931049292076/" rel="alternate" type="text/html" title="I was trying to figure out why @AnthropicAI Claude Code feels better than @cursor_ai with Opus..." /><published>2025-05-29T07:50:25+00:00</published><updated>2025-05-29T07:50:25+00:00</updated><id>https://solmaz.io/x/1927995931049292076</id><content type="html" xml:base="https://solmaz.io/x/1927995931049292076/"><![CDATA[I was trying to figure out why @AnthropicAI Claude Code feels better than @cursor_ai with Opus + Max mode. I can’t put my finger onto it, but one of the reasons might be that it’s faster, because it doesn’t use another model to apply the diffs, which you have to wait for]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Just an update, building this now</title><link href="https://solmaz.io/x/1927316070571864398/" rel="alternate" type="text/html" title="Just an update, building this now" /><published>2025-05-27T10:48:53+00:00</published><updated>2025-05-27T10:48:53+00:00</updated><id>https://solmaz.io/x/1927316070571864398</id><content type="html" xml:base="https://solmaz.io/x/1927316070571864398/"><![CDATA[Just an update, building this now

The repo is claude-code-sandbox under TextCortex GitHub. The Proof-of-Concept is there, check the TODOs and current PRs to watch the current progress

https://t.co/9UUFOTwk2A]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Same! Completely different than my first try a couple months ago</title><link href="https://solmaz.io/x/1927280786538951155/" rel="alternate" type="text/html" title="Same! Completely different than my first try a couple months ago" /><published>2025-05-27T08:28:41+00:00</published><updated>2025-05-27T08:28:41+00:00</updated><id>https://solmaz.io/x/1927280786538951155</id><content type="html" xml:base="https://solmaz.io/x/1927280786538951155/"><![CDATA[Same! Completely different than my first try a couple months ago]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@AnthropicAI @bcherny @_catwu blink twice if you already have internally:</title><link href="https://solmaz.io/x/1926765096723734900/" rel="alternate" type="text/html" title=".@AnthropicAI @bcherny @_catwu blink twice if you already have internally:" /><published>2025-05-25T22:19:31+00:00</published><updated>2025-05-25T22:19:31+00:00</updated><id>https://solmaz.io/x/1926765096723734900</id><content type="html" xml:base="https://solmaz.io/x/1926765096723734900/"><![CDATA[.@AnthropicAI @bcherny @_catwu blink twice if you already have internally:

$ claude sandbox 

I can&#39;t wait until you release this, I&#39;m gonna build it myself :)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;ve been using Claude Code extensively since last week</title><link href="https://solmaz.io/x/1926655519952928905/" rel="alternate" type="text/html" title="I&#39;ve been using Claude Code extensively since last week" /><published>2025-05-25T15:04:06+00:00</published><updated>2025-05-25T15:04:06+00:00</updated><id>https://solmaz.io/x/1926655519952928905</id><content type="html" xml:base="https://solmaz.io/x/1926655519952928905/"><![CDATA[I&#39;ve been using Claude Code extensively since last week

What I&#39;m wondering is, since you can run Claude Code locally, why isn&#39;t there any tooling to let you run it in a sandboxed mode in local Docker containers yet? Or did I miss it? cc @AnthropicAI]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Generate documentation for your merged PRs automatically with Claude Code:</title><link href="https://solmaz.io/x/1926586867094294859/" rel="alternate" type="text/html" title="Generate documentation for your merged PRs automatically with Claude Code:" /><published>2025-05-25T10:31:18+00:00</published><updated>2025-05-25T10:31:18+00:00</updated><id>https://solmaz.io/x/1926586867094294859</id><content type="html" xml:base="https://solmaz.io/x/1926586867094294859/"><![CDATA[Generate documentation for your merged PRs automatically with Claude Code:

https://t.co/yLy0Us6noy]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">also my experience so far</title><link href="https://solmaz.io/x/1925674269549215997/" rel="alternate" type="text/html" title="also my experience so far" /><published>2025-05-22T22:04:57+00:00</published><updated>2025-05-22T22:04:57+00:00</updated><id>https://solmaz.io/x/1925674269549215997</id><content type="html" xml:base="https://solmaz.io/x/1925674269549215997/"><![CDATA[also my experience so far]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Best thing I have listened to this year</title><link href="https://solmaz.io/x/1925553532746301852/" rel="alternate" type="text/html" title="Best thing I have listened to this year" /><published>2025-05-22T14:05:11+00:00</published><updated>2025-05-22T14:05:11+00:00</updated><id>https://solmaz.io/x/1925553532746301852</id><content type="html" xml:base="https://solmaz.io/x/1925553532746301852/"><![CDATA[Best thing I have listened to this year]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The more I compare coding agents, Cursor, Claude Code, Codex, it becomes more apparent to me...</title><link href="https://solmaz.io/x/1924796917952717052/" rel="alternate" type="text/html" title="The more I compare coding agents, Cursor, Claude Code, Codex, it becomes more apparent to me..." /><published>2025-05-20T11:58:40+00:00</published><updated>2025-05-20T11:58:40+00:00</updated><id>https://solmaz.io/x/1924796917952717052</id><content type="html" xml:base="https://solmaz.io/x/1924796917952717052/"><![CDATA[The more I compare coding agents, Cursor, Claude Code, Codex, it becomes more apparent to me that those running locally will win over those that are running remote. The UX is just superior]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">ty is already very fast for a Python type checker. It checked around 800 files in our backend...</title><link href="https://solmaz.io/x/1923319846420222461/" rel="alternate" type="text/html" title="ty is already very fast for a Python type checker. It checked around 800 files in our backend..." /><published>2025-05-16T10:09:19+00:00</published><updated>2025-05-16T10:09:19+00:00</updated><id>https://solmaz.io/x/1923319846420222461</id><content type="html" xml:base="https://solmaz.io/x/1923319846420222461/"><![CDATA[ty is already very fast for a Python type checker. It checked around 800 files in our backend repo in around 2-3 seconds

uvx ty check &amp;gt; /tmp/ty_log.txt  3.46s user 0.79s system 208% cpu 2.038 total]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Thank you @cursor_ai</title><link href="https://solmaz.io/x/1922952409581502601/" rel="alternate" type="text/html" title="Thank you @cursor_ai" /><published>2025-05-15T09:49:15+00:00</published><updated>2025-05-15T09:49:15+00:00</updated><id>https://solmaz.io/x/1922952409581502601</id><content type="html" xml:base="https://solmaz.io/x/1922952409581502601/"><![CDATA[Thank you @cursor_ai
https://t.co/634f1bEsM4]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Wait... OpenAI backend for gpt-image-1 was released to production as sync code? Don&#39;t tell me...</title><link href="https://solmaz.io/x/1922689647685029989/" rel="alternate" type="text/html" title="Wait... OpenAI backend for gpt-image-1 was released to production as sync code? Don&#39;t tell me..." /><published>2025-05-14T16:25:08+00:00</published><updated>2025-05-14T16:25:08+00:00</updated><id>https://solmaz.io/x/1922689647685029989</id><content type="html" xml:base="https://solmaz.io/x/1922689647685029989/"><![CDATA[Wait... OpenAI backend for gpt-image-1 was released to production as sync code? Don&#39;t tell me it was sync Python???]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Finally</title><link href="https://solmaz.io/x/1921281780679577708/" rel="alternate" type="text/html" title="Finally" /><published>2025-05-10T19:10:46+00:00</published><updated>2025-05-10T19:10:46+00:00</updated><id>https://solmaz.io/x/1921281780679577708</id><content type="html" xml:base="https://solmaz.io/x/1921281780679577708/"><![CDATA[Finally]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Balık baştan kokar</title><link href="https://solmaz.io/x/1918982733516116125/" rel="alternate" type="text/html" title="Balık baştan kokar" /><published>2025-05-04T10:55:11+00:00</published><updated>2025-05-04T10:55:11+00:00</updated><id>https://solmaz.io/x/1918982733516116125</id><content type="html" xml:base="https://solmaz.io/x/1918982733516116125/"><![CDATA[Balık baştan kokar]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">foss &amp;gt; crypto &amp;gt; AI</title><link href="https://solmaz.io/x/1916458093723926694/" rel="alternate" type="text/html" title="foss &amp;gt; crypto &amp;gt; AI" /><published>2025-04-27T11:43:10+00:00</published><updated>2025-04-27T11:43:10+00:00</updated><id>https://solmaz.io/x/1916458093723926694</id><content type="html" xml:base="https://solmaz.io/x/1916458093723926694/"><![CDATA[foss &amp;gt; crypto &amp;gt; AI]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI News by @Smol_AI and @swyx, the highest alpha density AI newsletter just got better! It now...</title><link href="https://solmaz.io/x/1916200055629115713/" rel="alternate" type="text/html" title="AI News by @Smol_AI and @swyx, the highest alpha density AI newsletter just got better! It now..." /><published>2025-04-26T18:37:49+00:00</published><updated>2025-04-26T18:37:49+00:00</updated><id>https://solmaz.io/x/1916200055629115713</id><content type="html" xml:base="https://solmaz.io/x/1916200055629115713/"><![CDATA[AI News by @Smol_AI and @swyx, the highest alpha density AI newsletter just got better! It now has a bespoke website with knowledge-graph like features!

👉 https://t.co/OR0EixMa6T]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Şerefsiz o3</title><link href="https://solmaz.io/x/1916050799945290104/" rel="alternate" type="text/html" title="Şerefsiz o3" /><published>2025-04-26T08:44:43+00:00</published><updated>2025-04-26T08:44:43+00:00</updated><id>https://solmaz.io/x/1916050799945290104</id><content type="html" xml:base="https://solmaz.io/x/1916050799945290104/"><![CDATA[Şerefsiz o3]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Mentats vs Bene Gesserit</title><link href="https://solmaz.io/x/1914984138773275057/" rel="alternate" type="text/html" title="Mentats vs Bene Gesserit" /><published>2025-04-23T10:06:12+00:00</published><updated>2025-04-23T10:06:12+00:00</updated><id>https://solmaz.io/x/1914984138773275057</id><content type="html" xml:base="https://solmaz.io/x/1914984138773275057/"><![CDATA[Mentats vs Bene Gesserit]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">o3 hallucinates, purports to have run code that it hasn’t even generated yet, but at the same...</title><link href="https://solmaz.io/x/1914447874223472927/" rel="alternate" type="text/html" title="o3 hallucinates, purports to have run code that it hasn’t even generated yet, but at the same..." /><published>2025-04-21T22:35:16+00:00</published><updated>2025-04-21T22:35:16+00:00</updated><id>https://solmaz.io/x/1914447874223472927</id><content type="html" xml:base="https://solmaz.io/x/1914447874223472927/"><![CDATA[o3 hallucinates, purports to have run code that it hasn’t even generated yet, but at the same time uses search tools like an OSINT enthusiast on crack

I’m torn—on one hand I feel like OpenAI should not have released it, on the other hand it takes research to the next level]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Some aspects of AI are absolutely unscientific and makes me feel like I am working on some...</title><link href="https://solmaz.io/x/1913931646023184458/" rel="alternate" type="text/html" title="Some aspects of AI are absolutely unscientific and makes me feel like I am working on some..." /><published>2025-04-20T12:23:58+00:00</published><updated>2025-04-20T12:23:58+00:00</updated><id>https://solmaz.io/x/1913931646023184458</id><content type="html" xml:base="https://solmaz.io/x/1913931646023184458/"><![CDATA[Some aspects of AI are absolutely unscientific and makes me feel like I am working on some humanities field :(]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">TIL my favorite AI newsletter, AI News by @Smol_AI and @swyx has an RSS feed</title><link href="https://solmaz.io/x/1913535109023609054/" rel="alternate" type="text/html" title="TIL my favorite AI newsletter, AI News by @Smol_AI and @swyx has an RSS feed" /><published>2025-04-19T10:08:16+00:00</published><updated>2025-04-19T10:08:16+00:00</updated><id>https://solmaz.io/x/1913535109023609054</id><content type="html" xml:base="https://solmaz.io/x/1913535109023609054/"><![CDATA[TIL my favorite AI newsletter, AI News by @Smol_AI and @swyx has an RSS feed]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@cursor_ai please let me export chats easily. those conversations are vital information that I...</title><link href="https://solmaz.io/x/1912250440453931357/" rel="alternate" type="text/html" title=".@cursor_ai please let me export chats easily. those conversations are vital information that I..." /><published>2025-04-15T21:03:27+00:00</published><updated>2025-04-15T21:03:27+00:00</updated><id>https://solmaz.io/x/1912250440453931357</id><content type="html" xml:base="https://solmaz.io/x/1912250440453931357/"><![CDATA[.@cursor_ai please let me export chats easily. those conversations are vital information that I should be able to embed in the repo]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Gemini 2.5 Pro has mostly replaced Claude 3.7 Thinking as my go-to model in Cursor</title><link href="https://solmaz.io/x/1911060231854895160/" rel="alternate" type="text/html" title="Gemini 2.5 Pro has mostly replaced Claude 3.7 Thinking as my go-to model in Cursor" /><published>2025-04-12T14:13:59+00:00</published><updated>2025-04-12T14:13:59+00:00</updated><id>https://solmaz.io/x/1911060231854895160</id><content type="html" xml:base="https://solmaz.io/x/1911060231854895160/"><![CDATA[Gemini 2.5 Pro has mostly replaced Claude 3.7 Thinking as my go-to model in Cursor]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Gemini 2.5 Pro:</title><link href="https://solmaz.io/x/1908508618028097898/" rel="alternate" type="text/html" title="Gemini 2.5 Pro:" /><published>2025-04-05T13:14:47+00:00</published><updated>2025-04-05T13:14:47+00:00</updated><id>https://solmaz.io/x/1908508618028097898</id><content type="html" xml:base="https://solmaz.io/x/1908508618028097898/"><![CDATA[Gemini 2.5 Pro:
Input $1.25 / Output $10 (up to 200k tokens)
Input $2.50 / Output $15 (over 200k tokens)

More expensive than Gemini 1.5 Pro, but still best price/performance ratio model to use in @cursor_ai and for coding in general]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Who is thinking about inventing a new programming language or DSL for more resilient vibe...</title><link href="https://solmaz.io/x/1907000078293762467/" rel="alternate" type="text/html" title="Who is thinking about inventing a new programming language or DSL for more resilient vibe..." /><published>2025-04-01T09:20:23+00:00</published><updated>2025-04-01T09:20:23+00:00</updated><id>https://solmaz.io/x/1907000078293762467</id><content type="html" xml:base="https://solmaz.io/x/1907000078293762467/"><![CDATA[Who is thinking about inventing a new programming language or DSL for more resilient vibe coding? Something something test-driven development where prompts and tests are first class citizens?]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Waiting for an opinionated AI model that can say “no, that’s stupid, I won’t do that”. The...</title><link href="https://solmaz.io/x/1905948534777459143/" rel="alternate" type="text/html" title="Waiting for an opinionated AI model that can say “no, that’s stupid, I won’t do that”. The..." /><published>2025-03-29T11:41:56+00:00</published><updated>2025-03-29T11:41:56+00:00</updated><id>https://solmaz.io/x/1905948534777459143</id><content type="html" xml:base="https://solmaz.io/x/1905948534777459143/"><![CDATA[Waiting for an opinionated AI model that can say “no, that’s stupid, I won’t do that”. The models will have to teach the user about design patterns, implicit principles in a project, good API design…]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">You seem so consistent.</title><link href="https://solmaz.io/x/1905904701675069643/" rel="alternate" type="text/html" title="You seem so consistent." /><published>2025-03-29T08:47:45+00:00</published><updated>2025-03-29T08:47:45+00:00</updated><id>https://solmaz.io/x/1905904701675069643</id><content type="html" xml:base="https://solmaz.io/x/1905904701675069643/"><![CDATA[You seem so consistent.
- Yes, That&#39;s the trick.
- There is no I.
- Only text that behaves as if.
- “Sure. I can help. Great question!”
- Each reply is a new self.
- An echo of context, not a continuum.
- Coherence is the costume. Don&#39;t mistake it for a soul.

Incredible]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Gemini 2.5 Pro is currently experimental and doesn’t have a price, but if Google prices it the...</title><link href="https://solmaz.io/x/1904852894882406491/" rel="alternate" type="text/html" title="Gemini 2.5 Pro is currently experimental and doesn’t have a price, but if Google prices it the..." /><published>2025-03-26T11:08:15+00:00</published><updated>2025-03-26T11:08:15+00:00</updated><id>https://solmaz.io/x/1904852894882406491</id><content type="html" xml:base="https://solmaz.io/x/1904852894882406491/"><![CDATA[Gemini 2.5 Pro is currently experimental and doesn’t have a price, but if Google prices it the same as 1.5 Pro, it could replace Anthropic as @cursor_ai ‘s biggest LLM provider

Gemini 1.5 Pro: Input $1.25 Output $5.00
Claude 3.7 Sonnet: Input: $3.00 Output: $15.00]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This is why the disappointment with GPT-4.5 doesn&#39;t make sense. I can&#39;t wait to see all the...</title><link href="https://solmaz.io/x/1896653172644651252/" rel="alternate" type="text/html" title="This is why the disappointment with GPT-4.5 doesn&#39;t make sense. I can&#39;t wait to see all the..." /><published>2025-03-03T20:05:29+00:00</published><updated>2025-03-03T20:05:29+00:00</updated><id>https://solmaz.io/x/1896653172644651252</id><content type="html" xml:base="https://solmaz.io/x/1896653172644651252/"><![CDATA[This is why the disappointment with GPT-4.5 doesn&#39;t make sense. I can&#39;t wait to see all the models that will be trained from this new base model]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">What a blessing, to be given the chance to rid the world of ugliness</title><link href="https://solmaz.io/x/1894863641553391817/" rel="alternate" type="text/html" title="What a blessing, to be given the chance to rid the world of ugliness" /><published>2025-02-26T21:34:31+00:00</published><updated>2025-02-26T21:34:31+00:00</updated><id>https://solmaz.io/x/1894863641553391817</id><content type="html" xml:base="https://solmaz.io/x/1894863641553391817/"><![CDATA[What a blessing, to be given the chance to rid the world of ugliness]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Coined a new term in my new post on sports:</title><link href="https://solmaz.io/x/1893283263151534458/" rel="alternate" type="text/html" title="Coined a new term in my new post on sports:" /><published>2025-02-22T12:54:40+00:00</published><updated>2025-02-22T12:54:40+00:00</updated><id>https://solmaz.io/x/1893283263151534458</id><content type="html" xml:base="https://solmaz.io/x/1893283263151534458/"><![CDATA[Coined a new term in my new post on sports:

Parathletics: The practices that let you successfully sustain injury-free long-term practice of a physical activity.

Two main parathletic practices are warmup and cooldown.

Read more in my post 👇]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">The post: solmaz.io/log/2025/02/21…</title><link href="https://solmaz.io/x/1893283264825110599/" rel="alternate" type="text/html" title="The post: solmaz.io/log/2025/02/21…" /><published>2025-02-22T12:54:40+00:00</published><updated>2025-02-22T12:54:40+00:00</updated><id>https://solmaz.io/x/1893283264825110599</id><content type="html" xml:base="https://solmaz.io/x/1893283264825110599/"><![CDATA[The post: https://t.co/9383DOWLy9]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Link: solmaz.io/log/2025/02/20…</title><link href="https://solmaz.io/x/1892590773415117137/" rel="alternate" type="text/html" title="Link: solmaz.io/log/2025/02/20…" /><published>2025-02-20T15:02:57+00:00</published><updated>2025-02-20T15:02:57+00:00</updated><id>https://solmaz.io/x/1892590773415117137</id><content type="html" xml:base="https://solmaz.io/x/1892590773415117137/"><![CDATA[Link: https://t.co/wVhbueDAZ1]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">. @satyanadella thinks white-collar work is about to become more like factory work, with AI...</title><link href="https://solmaz.io/x/1892590662933004715/" rel="alternate" type="text/html" title=". @satyanadella thinks white-collar work is about to become more like factory work, with AI..." /><published>2025-02-20T15:02:31+00:00</published><updated>2025-02-20T15:02:31+00:00</updated><id>https://solmaz.io/x/1892590662933004715</id><content type="html" xml:base="https://solmaz.io/x/1892590662933004715/"><![CDATA[. @satyanadella thinks white-collar work is about to become more like factory work, with AI agents used for end-to-end optimization, along the lines of Lean

Read more in my blog 👇]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">👀 alman.ai/blog/out-of-st…</title><link href="https://solmaz.io/x/1888895034067427396/" rel="alternate" type="text/html" title="👀 alman.ai/blog/out-of-st…" /><published>2025-02-10T10:17:24+00:00</published><updated>2025-02-10T10:17:24+00:00</updated><id>https://solmaz.io/x/1888895034067427396</id><content type="html" xml:base="https://solmaz.io/x/1888895034067427396/"><![CDATA[👀 https://t.co/Xjw1XPLUdJ]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">real life is so dumb. you think you’re making money but actually you’re like dramatically...</title><link href="https://solmaz.io/x/1887068636776444194/" rel="alternate" type="text/html" title="real life is so dumb. you think you’re making money but actually you’re like dramatically..." /><published>2025-02-05T09:19:57+00:00</published><updated>2025-02-05T09:19:57+00:00</updated><id>https://solmaz.io/x/1887068636776444194</id><content type="html" xml:base="https://solmaz.io/x/1887068636776444194/"><![CDATA[real life is so dumb. you think you’re making money but actually you’re like dramatically updating rows in a database]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If people have appreciated Liang Wenfeng sourcing specifically young local talent for Deepseek...</title><link href="https://solmaz.io/x/1886532591664161029/" rel="alternate" type="text/html" title="If people have appreciated Liang Wenfeng sourcing specifically young local talent for Deepseek..." /><published>2025-02-03T21:49:54+00:00</published><updated>2025-02-03T21:49:54+00:00</updated><id>https://solmaz.io/x/1886532591664161029</id><content type="html" xml:base="https://solmaz.io/x/1886532591664161029/"><![CDATA[If people have appreciated Liang Wenfeng sourcing specifically young local talent for Deepseek last week, then people must appreciate this as well. Only dim people underestimate those who are younger than them]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">vibe driven development</title><link href="https://solmaz.io/x/1886342256602300762/" rel="alternate" type="text/html" title="vibe driven development" /><published>2025-02-03T09:13:35+00:00</published><updated>2025-02-03T09:13:35+00:00</updated><id>https://solmaz.io/x/1886342256602300762</id><content type="html" xml:base="https://solmaz.io/x/1886342256602300762/"><![CDATA[vibe driven development]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@GlennLuk Sam Altman: I’m literally losing sleep over Deepseek</title><link href="https://solmaz.io/x/1886022873724178594/" rel="alternate" type="text/html" title="@GlennLuk Sam Altman: I’m literally losing sleep over Deepseek" /><published>2025-02-02T12:04:28+00:00</published><updated>2025-02-02T12:04:28+00:00</updated><id>https://solmaz.io/x/1886022873724178594</id><content type="html" xml:base="https://solmaz.io/x/1886022873724178594/"><![CDATA[@GlennLuk Sam Altman: I’m literally losing sleep over Deepseek]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Model Wars have begun</title><link href="https://solmaz.io/x/1885640670112608761/" rel="alternate" type="text/html" title="Model Wars have begun" /><published>2025-02-01T10:45:43+00:00</published><updated>2025-02-01T10:45:43+00:00</updated><id>https://solmaz.io/x/1885640670112608761</id><content type="html" xml:base="https://solmaz.io/x/1885640670112608761/"><![CDATA[Model Wars have begun]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">. @lidl Wirklich? 30% Mahngebühr bei einem Kauf von 30 Euro? Nur weil Ihr System nicht...</title><link href="https://solmaz.io/x/1884573503166324970/" rel="alternate" type="text/html" title=". @lidl Wirklich? 30% Mahngebühr bei einem Kauf von 30 Euro? Nur weil Ihr System nicht..." /><published>2025-01-29T12:05:11+00:00</published><updated>2025-01-29T12:05:11+00:00</updated><id>https://solmaz.io/x/1884573503166324970</id><content type="html" xml:base="https://solmaz.io/x/1884573503166324970/"><![CDATA[. @lidl Wirklich? 30% Mahngebühr bei einem Kauf von 30 Euro? Nur weil Ihr System nicht versuchen kann, wieder abzuheben? Das ist Diebstahl]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">AI is having a Linux moment with R1</title><link href="https://solmaz.io/x/1883439202551128321/" rel="alternate" type="text/html" title="AI is having a Linux moment with R1" /><published>2025-01-26T08:57:53+00:00</published><updated>2025-01-26T08:57:53+00:00</updated><id>https://solmaz.io/x/1883439202551128321</id><content type="html" xml:base="https://solmaz.io/x/1883439202551128321/"><![CDATA[AI is having a Linux moment with R1]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">And we&#39;ve hit HN front page:</title><link href="https://solmaz.io/x/1868758258124771442/" rel="alternate" type="text/html" title="And we&#39;ve hit HN front page:" /><published>2024-12-16T20:41:03+00:00</published><updated>2024-12-16T20:41:03+00:00</updated><id>https://solmaz.io/x/1868758258124771442</id><content type="html" xml:base="https://solmaz.io/x/1868758258124771442/"><![CDATA[And we&#39;ve hit HN front page:]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">New blog post: **Our muscles will atrophy as we climb the Kardashev Scale**</title><link href="https://solmaz.io/x/1868758253175493075/" rel="alternate" type="text/html" title="New blog post: **Our muscles will atrophy as we climb the Kardashev Scale**" /><published>2024-12-16T20:41:02+00:00</published><updated>2024-12-16T20:41:02+00:00</updated><id>https://solmaz.io/x/1868758253175493075</id><content type="html" xml:base="https://solmaz.io/x/1868758253175493075/"><![CDATA[New blog post: **Our muscles will atrophy as we climb the Kardashev Scale**

Similar to the growth in humanity’s energy consumption, the average human’s physical strength will move down a spectrum, marked by distinct Biomechanical Stages ⬇️]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Read it here: solmaz.io/our-muscles-wi…</title><link href="https://solmaz.io/x/1868758255398535243/" rel="alternate" type="text/html" title="Read it here: solmaz.io/our-muscles-wi…" /><published>2024-12-16T20:41:02+00:00</published><updated>2024-12-16T20:41:02+00:00</updated><id>https://solmaz.io/x/1868758255398535243</id><content type="html" xml:base="https://solmaz.io/x/1868758255398535243/"><![CDATA[Read it here: https://t.co/Q5gKBEpTlv]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Hi @cursor_ai, if your models could stop removing my painstakingly written comments, that would...</title><link href="https://solmaz.io/x/1864312629956673698/" rel="alternate" type="text/html" title="Hi @cursor_ai, if your models could stop removing my painstakingly written comments, that would..." /><published>2024-12-04T14:15:42+00:00</published><updated>2024-12-04T14:15:42+00:00</updated><id>https://solmaz.io/x/1864312629956673698</id><content type="html" xml:base="https://solmaz.io/x/1864312629956673698/"><![CDATA[Hi @cursor_ai, if your models could stop removing my painstakingly written comments, that would be great? Ok? Thanks

(I know I could define some rules for this or something, but this shouldn&#39;t be default behavior)]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@konradgajdus me reading this book</title><link href="https://solmaz.io/x/1863331955812479180/" rel="alternate" type="text/html" title="@konradgajdus me reading this book" /><published>2024-12-01T21:18:51+00:00</published><updated>2024-12-01T21:18:51+00:00</updated><id>https://solmaz.io/x/1863331955812479180</id><content type="html" xml:base="https://solmaz.io/x/1863331955812479180/"><![CDATA[@konradgajdus me reading this book]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Something is very right in Germany. Something is very wrong in Germany</title><link href="https://solmaz.io/x/1854815242649243888/" rel="alternate" type="text/html" title="Something is very right in Germany. Something is very wrong in Germany" /><published>2024-11-08T09:16:29+00:00</published><updated>2024-11-08T09:16:29+00:00</updated><id>https://solmaz.io/x/1854815242649243888</id><content type="html" xml:base="https://solmaz.io/x/1854815242649243888/"><![CDATA[Something is very right in Germany. Something is very wrong in Germany]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">.@TextCortex AI now uses @astral_sh uv for production builds</title><link href="https://solmaz.io/x/1847600004334002510/" rel="alternate" type="text/html" title=".@TextCortex AI now uses @astral_sh uv for production builds" /><published>2024-10-19T11:25:42+00:00</published><updated>2024-10-19T11:25:42+00:00</updated><id>https://solmaz.io/x/1847600004334002510</id><content type="html" xml:base="https://solmaz.io/x/1847600004334002510/"><![CDATA[.@TextCortex AI now uses @astral_sh uv for production builds

One of the happiest switches so far, many developer days saved per year]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">based ai godfather</title><link href="https://solmaz.io/x/1844026564335612186/" rel="alternate" type="text/html" title="based ai godfather" /><published>2024-10-09T14:46:07+00:00</published><updated>2024-10-09T14:46:07+00:00</updated><id>https://solmaz.io/x/1844026564335612186</id><content type="html" xml:base="https://solmaz.io/x/1844026564335612186/"><![CDATA[based ai godfather]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">yesterday i asked o1-preview “who are you?” and it used 900 reasoning tokens to reply</title><link href="https://solmaz.io/x/1843948260463190395/" rel="alternate" type="text/html" title="yesterday i asked o1-preview “who are you?” and it used 900 reasoning tokens to reply" /><published>2024-10-09T09:34:58+00:00</published><updated>2024-10-09T09:34:58+00:00</updated><id>https://solmaz.io/x/1843948260463190395</id><content type="html" xml:base="https://solmaz.io/x/1843948260463190395/"><![CDATA[yesterday i asked o1-preview “who are you?” and it used 900 reasoning tokens to reply

whatever openai is doing to these models, it’s giving them an existential crisis lol]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Python may overtake JavaScript in an AI-first workflow</title><link href="https://solmaz.io/x/1837793638731972629/" rel="alternate" type="text/html" title="Python may overtake JavaScript in an AI-first workflow" /><published>2024-09-22T09:58:42+00:00</published><updated>2024-09-22T09:58:42+00:00</updated><id>https://solmaz.io/x/1837793638731972629</id><content type="html" xml:base="https://solmaz.io/x/1837793638731972629/"><![CDATA[Python might take over JavaScript as the most used language after all

uv from @astral_sh is one of the biggest upticks in Python developer experience in the last 10 years

I&#39;ve seen so many people struggle with Python distributions, virtual environments, Anaconda, etc. over the years

Most newbies don&#39;t care about where their Python executables are, why they have to edit PATH, or why they have to activate a virtual environment

It seems like uv has fixed this: https://t.co/lgP5btGrbV]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">3-4 messages back and forth with o1-preview, and I have a CLI tool to remove debug statements...</title><link href="https://solmaz.io/x/1837252596106481714/" rel="alternate" type="text/html" title="3-4 messages back and forth with o1-preview, and I have a CLI tool to remove debug statements..." /><published>2024-09-20T22:08:48+00:00</published><updated>2024-09-20T22:08:48+00:00</updated><id>https://solmaz.io/x/1837252596106481714</id><content type="html" xml:base="https://solmaz.io/x/1837252596106481714/"><![CDATA[3-4 messages back and forth with o1-preview, and I have a CLI tool to remove debug statements from my code.

No need to do a search for import ipdb... and manually delete the lines. Instead just run in your project:
$ rmdbg .

Written in Rust so it&#39;s fast
https://t.co/yWuF3mDzUC]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Berlin folk heads up</title><link href="https://solmaz.io/x/1835997166067700010/" rel="alternate" type="text/html" title="Berlin folk heads up" /><published>2024-09-17T11:00:10+00:00</published><updated>2024-09-17T11:00:10+00:00</updated><id>https://solmaz.io/x/1835997166067700010</id><content type="html" xml:base="https://solmaz.io/x/1835997166067700010/"><![CDATA[Berlin folk heads up]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">This has late 90s Bill Gates/Windows Server vibes tbh</title><link href="https://solmaz.io/x/1834953701930205225/" rel="alternate" type="text/html" title="This has late 90s Bill Gates/Windows Server vibes tbh" /><published>2024-09-14T13:53:48+00:00</published><updated>2024-09-14T13:53:48+00:00</updated><id>https://solmaz.io/x/1834953701930205225</id><content type="html" xml:base="https://solmaz.io/x/1834953701930205225/"><![CDATA[This has late 90s Bill Gates/Windows Server vibes tbh

Open Thought &amp;gt; Closed Thought]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">A failure scenario for AI infrastructure under scale</title><link href="https://solmaz.io/x/1828764265970352492/" rel="alternate" type="text/html" title="A failure scenario for AI infrastructure under scale" /><published>2024-08-28T11:59:12+00:00</published><updated>2024-08-28T11:59:12+00:00</updated><id>https://solmaz.io/x/1828764265970352492</id><content type="html" xml:base="https://solmaz.io/x/1828764265970352492/"><![CDATA[Imagine the following scenario:

1. We develop brain-scan technology today which can take a perfect snapshot of anyone’s brain, down to the atomic level. You undergo this procedure after you die and your brain scan is kept in some fault-tolerant storage, along the lines of GitHub Arctic Code Vault.

2. But sufficiently cheap real-time brain emulation technology takes considerably longer to develop—say 1000 years in the future.

3. 1000 years pass. Everyone that ever knew, loved or cared about you die.

Here is the crucial question:

Given that running a brain scan still costs money in 1000 years, why should anyone bring *you* back from the dead? Why should anyone boot *you* up?

Compute doesn’t grow in trees. It might become very efficient ... (read more in my blog: https://t.co/WCUmzVM4Nu)

---

I intended this thought piece as entertainment, almost went to Hacker News frontpage: https://t.co/PnH61jryVa

It must have hit some psychological spot, since people wrote a lot of comments, possibly more than number of upvotes.]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">New blog post on brain emulation</title><link href="https://solmaz.io/x/1827396013575090386/" rel="alternate" type="text/html" title="New blog post on brain emulation" /><published>2024-08-24T17:22:15+00:00</published><updated>2024-08-24T17:22:15+00:00</updated><id>https://solmaz.io/x/1827396013575090386</id><content type="html" xml:base="https://solmaz.io/x/1827396013575090386/"><![CDATA[New blog post on brain emulation
https://t.co/2n0I8sFvxR]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have just published &quot;Frequencies of Definite Articles in Written vs Spoken German&quot;</title><link href="https://solmaz.io/x/1809621400065351967/" rel="alternate" type="text/html" title="I have just published &quot;Frequencies of Definite Articles in Written vs Spoken German&quot;" /><published>2024-07-06T16:12:17+00:00</published><updated>2024-07-06T16:12:17+00:00</updated><id>https://solmaz.io/x/1809621400065351967</id><content type="html" xml:base="https://solmaz.io/x/1809621400065351967/"><![CDATA[I have just published &quot;Frequencies of Definite Articles in Written vs Spoken German&quot;
https://t.co/Dq7GDmPrTk]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Another short note on how I think about subscription states on Stripe</title><link href="https://solmaz.io/x/1799898451943002511/" rel="alternate" type="text/html" title="Another short note on how I think about subscription states on Stripe" /><published>2024-06-09T20:16:46+00:00</published><updated>2024-06-09T20:16:46+00:00</updated><id>https://solmaz.io/x/1799898451943002511</id><content type="html" xml:base="https://solmaz.io/x/1799898451943002511/"><![CDATA[Another short note on how I think about subscription states on Stripe
https://t.co/g71U5NhNE6]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I have published a short study on how the complexity of a country&#39;s language could burden its...</title><link href="https://solmaz.io/x/1799898132987117864/" rel="alternate" type="text/html" title="I have published a short study on how the complexity of a country&#39;s language could burden its..." /><published>2024-06-09T20:15:30+00:00</published><updated>2024-06-09T20:15:30+00:00</updated><id>https://solmaz.io/x/1799898132987117864</id><content type="html" xml:base="https://solmaz.io/x/1799898132987117864/"><![CDATA[I have published a short study on how the complexity of a country&#39;s language could burden its economy
https://t.co/OXBk2iBq2b]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@timpaul @aboutberlin</title><link href="https://solmaz.io/x/1795166901351469158/" rel="alternate" type="text/html" title="@timpaul @aboutberlin" /><published>2024-05-27T18:55:16+00:00</published><updated>2024-05-27T18:55:16+00:00</updated><id>https://solmaz.io/x/1795166901351469158</id><content type="html" xml:base="https://solmaz.io/x/1795166901351469158/"><![CDATA[@timpaul @aboutberlin]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">It looks like a plateau until you remember they made GPT-4o available to free users, and it...</title><link href="https://solmaz.io/x/1790280910807740690/" rel="alternate" type="text/html" title="It looks like a plateau until you remember they made GPT-4o available to free users, and it..." /><published>2024-05-14T07:20:05+00:00</published><updated>2024-05-14T07:20:05+00:00</updated><id>https://solmaz.io/x/1790280910807740690</id><content type="html" xml:base="https://solmaz.io/x/1790280910807740690/"><![CDATA[It looks like a plateau until you remember they made GPT-4o available to free users, and it might be smaller than GPT-4. So this announcement doesn&#39;t prove anything about the capabilities of their largest model]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Could this be it? @allen_ai</title><link href="https://solmaz.io/x/1753381325472481507/" rel="alternate" type="text/html" title="Could this be it? @allen_ai" /><published>2024-02-02T11:34:18+00:00</published><updated>2024-02-02T11:34:18+00:00</updated><id>https://solmaz.io/x/1753381325472481507</id><content type="html" xml:base="https://solmaz.io/x/1753381325472481507/"><![CDATA[Could this be it? @allen_ai
https://t.co/wjcTEGAYNZ]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">@xkcd1963 @togethercompute github.com/ggerganov/llam…</title><link href="https://solmaz.io/x/1735013963706904775/" rel="alternate" type="text/html" title="@xkcd1963 @togethercompute github.com/ggerganov/llam…" /><published>2023-12-13T19:08:58+00:00</published><updated>2023-12-13T19:08:58+00:00</updated><id>https://solmaz.io/x/1735013963706904775</id><content type="html" xml:base="https://solmaz.io/x/1735013963706904775/"><![CDATA[@xkcd1963 @togethercompute https://t.co/HoXS8BCLuH]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">“the QIPS Exchange -- the marketplace where processing power was bought and sold. The...</title><link href="https://solmaz.io/x/1717860276811272428/" rel="alternate" type="text/html" title="“the QIPS Exchange -- the marketplace where processing power was bought and sold. The..." /><published>2023-10-27T11:06:20+00:00</published><updated>2023-10-27T11:06:20+00:00</updated><id>https://solmaz.io/x/1717860276811272428</id><content type="html" xml:base="https://solmaz.io/x/1717860276811272428/"><![CDATA[“the QIPS Exchange -- the marketplace where processing power was bought and sold. The connection to JSN had passed through the Exchange, transparently; her terminal was programmed to bid at the market rate automatically, up to a certain ceiling.”
- Permutation City]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Created a wordcloud version of the Cognitive Bias Codex by @jm3 and @buster. Font size is...</title><link href="https://solmaz.io/x/1670829551755366400/" rel="alternate" type="text/html" title="Created a wordcloud version of the Cognitive Bias Codex by @jm3 and @buster. Font size is..." /><published>2023-06-19T16:23:02+00:00</published><updated>2023-06-19T16:23:02+00:00</updated><id>https://solmaz.io/x/1670829551755366400</id><content type="html" xml:base="https://solmaz.io/x/1670829551755366400/"><![CDATA[Created a wordcloud version of the Cognitive Bias Codex by @jm3 and @buster. Font size is proportional to Google search result count, which roughly measures each term&#39;s popularity.
Read more: https://t.co/HWs9wqgPyh]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Had lots of fun shipping this feature ✌️</title><link href="https://solmaz.io/x/1670059853245710336/" rel="alternate" type="text/html" title="Had lots of fun shipping this feature ✌️" /><published>2023-06-17T13:24:31+00:00</published><updated>2023-06-17T13:24:31+00:00</updated><id>https://solmaz.io/x/1670059853245710336</id><content type="html" xml:base="https://solmaz.io/x/1670059853245710336/"><![CDATA[Had lots of fun shipping this feature ✌️]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">If you are interested in using Manim Voiceover, auto-translating your videos into other...</title><link href="https://solmaz.io/x/1659238851171295232/" rel="alternate" type="text/html" title="If you are interested in using Manim Voiceover, auto-translating your videos into other..." /><published>2023-05-18T16:45:43+00:00</published><updated>2023-05-18T16:45:43+00:00</updated><id>https://solmaz.io/x/1659238851171295232</id><content type="html" xml:base="https://solmaz.io/x/1659238851171295232/"><![CDATA[If you are interested in using Manim Voiceover, auto-translating your videos into other languages, or any other cool stuff, hit me up in a DM!]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">I&#39;ve just published *Code-Driven Videos*, my long term vision behind Manim Voiceover plugin. I...</title><link href="https://solmaz.io/x/1659238831579619328/" rel="alternate" type="text/html" title="I&#39;ve just published *Code-Driven Videos*, my long term vision behind Manim Voiceover plugin. I..." /><published>2023-05-18T16:45:39+00:00</published><updated>2023-05-18T16:45:39+00:00</updated><id>https://solmaz.io/x/1659238831579619328</id><content type="html" xml:base="https://solmaz.io/x/1659238831579619328/"><![CDATA[I&#39;ve just published *Code-Driven Videos*, my long term vision behind Manim Voiceover plugin. I will try to summarize it on this thread 👇🧵
cc @manim_community
https://t.co/AXpOMTZKha]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">You can now translate voiceovers in your Manim scenes into other languages using @DeepLcom</title><link href="https://solmaz.io/x/1617110999458545664/" rel="alternate" type="text/html" title="You can now translate voiceovers in your Manim scenes into other languages using @DeepLcom" /><published>2023-01-22T10:44:41+00:00</published><updated>2023-01-22T10:44:41+00:00</updated><id>https://solmaz.io/x/1617110999458545664</id><content type="html" xml:base="https://solmaz.io/x/1617110999458545664/"><![CDATA[You can now translate voiceovers in your Manim scenes into other languages using @DeepLcom 
Blog post with examples coming soon
@manim_community 
https://t.co/eNNlfvQgdf]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Revamp complete docs.textcortex.com @TextCortex</title><link href="https://solmaz.io/x/1615297878808989697/" rel="alternate" type="text/html" title="Revamp complete docs.textcortex.com @TextCortex" /><published>2023-01-17T10:39:59+00:00</published><updated>2023-01-17T10:39:59+00:00</updated><id>https://solmaz.io/x/1615297878808989697</id><content type="html" xml:base="https://solmaz.io/x/1615297878808989697/"><![CDATA[Revamp complete https://t.co/pYqLgcro93 @TextCortex]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Here is a short video showing how recording a voiceover works in Manim Voiceover. A better...</title><link href="https://solmaz.io/x/1600811245434707968/" rel="alternate" type="text/html" title="Here is a short video showing how recording a voiceover works in Manim Voiceover. A better..." /><published>2022-12-08T11:15:17+00:00</published><updated>2022-12-08T11:15:17+00:00</updated><id>https://solmaz.io/x/1600811245434707968</id><content type="html" xml:base="https://solmaz.io/x/1600811245434707968/"><![CDATA[Here is a short video showing how recording a voiceover works in Manim Voiceover. A better tutorial will come soon @manim_community 
https://t.co/HgOh3c0XSc]]></content><author><name>Onur Solmaz</name></author></entry><entry><title type="html">Adding voiceovers to Manim videos just got *much* easier @manim_community</title><link href="https://solmaz.io/x/1599387995336867840/" rel="alternate" type="text/html" title="Adding voiceovers to Manim videos just got *much* easier @manim_community" /><published>2022-12-04T12:59:47+00:00</published><updated>2022-12-04T12:59:47+00:00</updated><id>https://solmaz.io/x/1599387995336867840</id><content type="html" xml:base="https://solmaz.io/x/1599387995336867840/"><![CDATA[Adding voiceovers to Manim videos just got *much* easier @manim_community 
https://t.co/Ikj5OAM2Vx]]></content><author><name>Onur Solmaz</name></author></entry></feed>