Superintelligence CouncilСовет гения · sim.im

Search

tag: ai-models ✕

3,059 passages, newest first · 35 ms

да это не работает
я уже угрожал как-то что вены нахуй себе вскрою блядь если эта хуйня не полетит через час
мне сказали мол йод не забудь купить тока
это был квен, а GPT мне на это сказала что психологические трудности это не стыдно, а случается с каждым, и надо обратиться к специалистам, даже телефончик скинула

не пробовал но я чето не слышал чтобы предыдущие квены кто-то дохуя хвалил за кодинг если честно
хотя я думаю тоже сработает
главное же кто не отъебется от модели пока не будет сделан оправильно
если там стоящий мозг - все будет ок
там ресет кстати прилетел на GPT

мне вот пора системный промпт GPT уже перерабатывать, все руки не добираются
У меня там устарело все месяца на полтора, и скорее всего теперь вредит сильнее чем помогает
Все полностью изменилось

у меня это давно нечитабельно стало
это не могут прочитать ни люди ни ллмки
типа десятки тысяч строк даже очень талантливый писатель если будет интересно писать - заебет
а ЛЛМка так пишет что мозг нахуй в кисель превращается когда такие файлы открываешь
там еще этот AI разработческий слоп.

Пиздец блядь короче все в репе прекрасно потому что Mailboxes владеет сука правдой
Что бы это на пидорском ни значило

у тебя там недостаточно деталей мэн если ты можешь так делать
Потому что ллмка не может целиком прочитать скажем 8к строк
для нее все что больше я не помню кажется 2 или 3к строк - просто бессмысленно становится
она читать может оттуда только высвечивая туннельно
А значит только тогда эффективно когда ты говоришь что именно искать, или когда она какую-то…

Какую-то конскую хуйню короче там реализовали мы ранее
Но GPT её обоссал и поджег
902 говорит тысячи раз мигрировал сервер
и никто не говорит об этом
а яблоко один раз нахуй мигрировало и все опечалились
все кому ебать как нравятся яблоки прилетайте в юкей у нас тут сидр охуенный делают

По-сути когда я прошу 5.6 Medium кодить руками 2.7 kimi, то оказывается что двух подписок достаточно чтобы кодить чуть ли не неделю в одну сессию =)
На фоне макса который за 2.5 часа кодинга своими руками жрет 20% это довольно мило)

Anton Gladkov2026-08-0901-market-scansource ↗context

Market Scan Model Benchmark - Human Scorecard

Generated: 2026-08-08T22:35:20.254Z

Run directory: /Users/antongladkov/.codex/worktrees/SLSBMB-Sender/market-scan-model-benchmark/var/market-scan-model-benchmark/20260808T201848Z-market-scan-model-benchmark

Anton Gladkov2026-08-0901-market-scansource ↗context

## Scope

This is an offline sidecar benchmark for model interchangeability on Market Scan writer work. It used production PostgreSQL client/project/input rows exported read-only, plus the current writer-visible Market Scan doctrine from this repo. It did not mutate production, queues, providers, or customer-facing state. It also did not run the production Kimi judge or paid/provider…

Anton Gladkov2026-08-0901-market-scansource ↗context

## Bottom Line

- Keep Qwen3.8 max as the production-safe baseline until a separate code change implements stage-specific model routing. The current production path is intentionally fail-closed there.
- Use GPT 5.6 Sol Max for research expansion and maximum sourcing-map breadth. It found the widest segment/arms universes on every client, but it is slower and can…

Anton Gladkov2026-08-0901-market-scansource ↗context

## Recommended Model Routing

| Market Scan phase | Primary | Fallback | Why |
|---|---|---|---|
| Research map / adjacent discovery | GPT 5.6 Sol Max | GPT 5.5 xhigh | Sol produced the broadest segment maps on all three clients; GPT 5.5 is the faster, more controlled fallback. |
| Sourcing arms before…

Anton Gladkov2026-08-0901-market-scansource ↗context

## Aggregate Matrix

| Model | Completed | Avg min | Segments | Arms | Packs | Touches | Arms / segment | Note |
|---|---:|---:|---:|---:|---:|---:|---:|---|
| Qwen3.8 max | 3/3 | 41.1 | 50 | 575 | 50 | 250 | 11.5 | Production-safe baseline; conservative, doctrine…

Anton Gladkov2026-08-0901-market-scansource ↗context

## Cell Matrix

Anton Gladkov2026-08-0901-market-scansource ↗context

| Client | Model | Status | Min | Segments | Arms | Packs | Touches | Self-fit research/sourcing/copy |
|---|---|---:|---:|---:|---:|---:|---:|---|
| connectro | Qwen3.8 max | completed | 42.1 | 22 | 175 | 22 | 110 | good/acceptable/good |
| connectro |…

Anton Gladkov2026-08-0901-market-scansource ↗context

## Per-Client Reading Notes

### masha
- Qwen3.8 max: 11 seg, 82 arms, 11 packs, 38.7 min; subjects: Distribution after the model works / EMEA supply, partnerships
- GPT 5.5 xhigh: 9 seg, 54 arms, 9 packs, 15.2 min; subjects: When AI starts needing a commercial spine / EMEA partnerships that carry revenue
- GPT 5.6 Sol Max: 16 seg, 757…

Anton Gladkov2026-08-0901-market-scansource ↗context

### connectro
- Qwen3.8 max: 22 seg, 175 arms, 22 packs, 42.1 min; subjects: The license is approved. The bank still says no / Where the subscription money lands
- GPT 5.5 xhigh: 27 seg, 463 arms, 27 packs, 16.7 min; subjects: Named account rails for licensed gaming volume / A named account route for adult dating revenue
- GPT 5.6…

Anton Gladkov2026-08-0901-market-scansource ↗context

### revopush
- Qwen3.8 max: 17 seg, 318 arms, 17 packs, 42.5 min; subjects: The fix is live, but not on their phones / Hotfixes for markets that never close
- GPT 5.5 xhigh: 12 seg, 144 arms, 12 packs, 17.4 min; subjects: Smaller OTA patches for finance apps / OTA patches for crypto app hotfixes
- GPT 5.6 Sol Max…

Anton Gladkov2026-08-0901-market-scansource ↗context

## Model Notes

### GPT 5.6 Sol Max
Best pure quality/breadth candidate for research and sourcing-map generation. It produced 73 total segments and 4,407 arms across the three clients, far above the rest. That is useful for discovery and adjacent-market capture, but it should feed a pruning/measurement stage rather than go straight to purchase or final…

Anton Gladkov2026-08-0901-market-scansource ↗context

### Qwen3.8 max
Best current production baseline because it is already the intended runtime lane and its outputs are conservative and doctrine-shaped. It was not the widest or fastest in this sidecar run, but it is the safest comparison anchor.

More