COORD_X:005 // POST
Back to blogНазад в блог
ai strategy | 2026-08-04 | 3 minмин | Telegram originalоригинал

$100 on Codex vs. $100 on Claude$100 на Codex против $100 на Claude

$100 on Codex vs. $100 on Claude

A month of work, 40+ automations, and my final setup

Over the past month, I ran an experiment: I moved all my automations and my work with Claude (Opus 4.8) to Codex (GPT-5.5 → Sol 5.6).

Migration in a few hours

I did it in literally a few hours, despite having 10 separate projects and 22 ongoing chats. The switch may have been easy partly because of shared memory, shared skills, and a shared conversation database in Retain.

22 ongoing project chats in Codex. Some names are hidden.

Some of my 40+ automations. Yes, two tasks in the screenshot are currently broken 😅

FYI: I use all models with effort: high. This is less a comparison of the harnesses and more about the overall experience and current models.

First, GPT-5.5

At first, GPT-5.5 made me really happy because it did exactly what I asked, usually in one shot — with a single prompt.

However, on large tasks where it needs to generate a lot of content and keep everything in memory, the quality noticeably degrades. You definitely need to check the result with at least one additional pass.

GPT-5.5 is definitely worse at design than Opus 4.8, so presentations and websites became more painful.

Switching to Sol 5.6

I was thinking about going back to Claude because they extended access to Fable 5 and the design quality is solid. But then OpenAI released the new Luna, Terra, and Sol family, and I started testing it actively.

I integrated Luna and Terra into products where volume matters and speed and cost savings are important. I made Sol 5.6, the strongest model, the default for my 22 pinned chats.

And that’s when it got really good: I felt a jump in response quality, design, and performance on complex, long-running tasks. Overall, my quality concerns disappeared. It still sometimes dips on large tasks, but the result is noticeably better than GPT-5.5.

What I liked

Codex has convenient built-in Russian dictation: it’s fast and doesn’t put any load on local hardware. I use it all the time. Claude only has English dictation. I’ve already written about how much voice speeds up my work.

What annoys me

I was doing product fuzzing tasks for a client, and Sol regularly refused to work because it considered them hacking. OpenAI itself says that additional cyber checks sometimes affect legitimate defensive work. For this authorized task, I switched to GPT-5.5.

I also remotely asked computer use to log into my OpenAI account and reset my limit. This option appears in my account once every two weeks. Sol refused, while GPT-5.5 helped. The reset expires if you don’t use it, so I didn’t want to wait until I got home.

Two available full resets in my account.

Codex isn’t as friendly and doesn’t feel as much like “your” assistant: it mirrors you less effectively than Claude. For creative work, when you want to make something cool, Claude is much more enjoyable to talk to. For cold-blooded execution, Codex is still the best option.

My final setup

The main thing I love about Codex is the limit. It feels several times higher, especially with resets. With my $100 OpenAI subscription, I can do roughly 2.5 times more than with an Anthropic subscription at the same price.

Another interesting shift: I used Claude through the CLI in the CMUX terminal, while I use Codex through the official app.

I’ve already written that I spend $180–250 a month on AI. After this experiment, my ideal setup is two $100 subscriptions: one for OpenAI and one for Anthropic.

Decide for yourself which one makes more sense for you right now.

$100 на Codex против $100 на Claude

Месяц работы, 40+ автоматизаций и мой итоговый сетап

Последний месяц я экспериментировал: перевёл все свои автоматизации и работу с Claude (Opus 4.8) на Codex (GPT-5.5 → Sol 5.6).

Миграция за несколько часов

Сделал это буквально за несколько часов, несмотря на 10 отдельных проектов и 22 постоянных чата. Возможно, переключиться было легко в том числе из-за общей памяти, общих skills и общей базы диалогов в Retain.

22 постоянных проектных чата в Codex. Часть названий скрыта.

Часть моих 40+ автоматизаций. Да, две задачи на скрине сейчас сломаны 😅

FYI: все модели использую на effort: high. Здесь сравниваю не столько harness, сколько общий опыт и актуальные модели.

Сначала GPT-5.5

Поначалу GPT-5.5 очень радовал, потому что делал ровно то, что я просил, и обычно с one shot — одного промпта.

Однако в объёмных задачах, где нужно сгенерировать много контента и удержать всё в памяти, качество заметно деградирует. Итог обязательно нужно проверять хотя бы одним дополнительным проходом.

С дизайном у GPT-5.5 точно хуже, чем у Opus 4.8, поэтому с презентациями и сайтами стало больнее.

Переход на Sol 5.6

Думал вернуться обратно на Claude, потому что там продлили доступ к Fable 5 и дизайн нормальный. Но тут OpenAI выпустила новое семейство Luna, Terra и Sol, и я начал активно тестировать.

Luna и Terra я встроил в продукты, где есть объёмы и важны скорость и экономия. Sol 5.6, самую сильную модель, сделал основной для своих 22 закреплённых чатов.

И тут пошёл кайф: я ощутил скачок в качестве ответов, дизайне и работе со сложными и долгими задачами. В целом вопросы к качеству отпали. На больших задачах оно всё равно иногда проседает, но результат заметно лучше GPT-5.5.

Что зашло

В Codex удобная встроенная диктовка на русском: быстрая и без нагрузки на локальное железо. Использую её постоянно. В Claude диктовка только на английском. Я уже писал, насколько голос ускоряет мою работу.

Что бесит

Я делал для заказчика задачи по фаззингу продукта, и Sol регулярно отказывался работать, потому что считал это хакерством. OpenAI сама пишет, что дополнительные cyber-проверки иногда затрагивают легитимную defensive-работу. Для этой разрешённой задачи я переключался на GPT-5.5.

Ещё я удалённо попросил через computer use зайти в мой OpenAI-аккаунт и сбросить лимит. У меня в аккаунте раз в две недели появляется такая возможность. Sol отказался, а GPT-5.5 помог. Reset сгорает, если его не использовать, поэтому дома ждать не хотелось.

Два доступных full reset в моём аккаунте.

Codex не такой дружелюбный и не такой «свой»: он хуже зеркалит тебя, чем Claude. Для креативной работы, когда хочется создать что-то классное, с Claude намного приятнее общаться. Для хладнокровного выполнения Codex пока лучший вариант.

Итоговый сетап

Основной кайф для меня — лимит Codex. По ощущениям, он кратно выше, особенно с reset. В свою подписку OpenAI за $100 я могу сделать примерно в 2,5 раза больше, чем в такую же по стоимости подписку Anthropic.

Ещё один интересный сдвиг: Claude я запускал через CLI в терминале CMUX, а Codex использую через официальное приложение.

Я уже писал, что трачу $180–250 в месяц на AI. После этого эксперимента мой идеальный сетап — две подписки по $100, на OpenAI и Anthropic.

Делайте выводы сами, что актуальнее для вас.