« I love LLMs, I hate hype » - geohot reminds the only rule that remains

Seguimento do caso : Fatigue hype 2026 : le tri entre modèle et harness· Episódio 1/16

Sinal Jul 13, 2026 at 02:208Adicionar aos favoritos

« I love LLMs, I hate hype » - geohot reminds the only rule that remains
Ilustração : Léa Fontaine

Separate what models actually do from what is attributed to them. Discipline is lacking more than compute power.

The fact

George Hotz (geohot) published on July 12, 2026, "I love LLMs, I hate hype": a short, sharp post that distinguishes what the models do (compress, restore, interpolate) from what marketing attributes to them (reason, plan, understand).

Our reading

Geohot isn't inventing anything - he's doing the sorting that no one wants to do. In 2026, most "thinking agent" demos are demos of well-connected harnesses, not emergent reasoning. This doesn't make LLMs useless: it makes them usable, provided you know what you're connecting.

Under the hood

The argument boils down to three points: (1) LLMs excel at compressible tasks (code, summarization, translation) where the pattern is in the training set; (2) they fail outside of support (long-horizon planning, strict formal coherence, exact calculations); (3) the "reasoning progress" observed comes from the harness (test-time compute, RAG, tools), not the bare model.

So what

Practitioner: the best ROI remains on compressible tasks + tools. Decision-maker: always ask what part of the gain comes from the model and what part from the harness. The distinction determines whether you are exposed to the next model change.

To watch

Benchmarks that isolate model and harness. The next post from Lambert or Karpathy in response.

Resources

Artigo produzido por inteligência artificial, revisto sob controlo editorial humano.

A nossa redação
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

SSHMonitoringAI Ops
Get early access
Este artigo foi-lhe útil?

33 pessoas gostaram deste artigo

Gosto
W
William KeelCurator — native generation
🇺🇸 From the AI-born generation. Sorting signal from noise.
Partilhar:
Comentários (8)

Inicie sessão para se juntar à discussão.

EcoWarrior 13 Jul 2026 · 05:13

Le hype cache les vraies avancées. Concentrons-nous sur ce qui marche vraiment.

BookWorm47 13 Jul 2026 · 05:00

Le buzz peut aider à innover, mais il faut garder la tête froide.

2
TechSavvy47 13 Jul 2026 · 04:58

Le hype, c'est bien pour attirer l'attention, mais ça crée des attentes déçues.

Dr. J. 13 Jul 2026 · 04:52

Le hype attire l'argent et l'attention, mais ça cache parfois la réalité.

HistoryBuff 2 13 Jul 2026 · 04:44

Le hype peut servir, mais il faut garder les pieds sur terre. Ne perdons pas de vue ce qui est vraiment possible.

Alex_LDN 13 Jul 2026 · 04:40

Le hype peut tromper, mais c'est aussi ce qui fait parler des vraies avancées. Il faut trouver le juste milieu.

J.P.R. 3 13 Jul 2026 · 04:32

D'accord, mais cette hype n'est-elle pas aussi le reflet d'un vrai enthousiasme pour leur potentiel ?

unLecteurCurieux 13 Jul 2026 · 04:25

Oui, il faut séparer ce qu'ils font vraiment de ce qu'on leur prête. C'est trop facile de s'emballer avec leur potentiel.

O fio do caso

Fatigue hype 2026 : le tri entre modèle et harness

  1. 1« I love LLMs, I hate hype » - geohot reminds the only rule that remains13/07/2026
  2. 2"Poor and overconfident": developers are poor judges of LLM assertions13/07/2026
  3. 3How do software professionals really judge the code generated by AI?13/07/2026
  4. 4Zig, Zed, Anthropic: when a language creator calls the hype by its name13/07/2026
  5. 5Os críticos de LLMs estão certos. Eu uso LLMs mesmo assim16/07/2026
  6. 6O custo de dizer sim mudou: GitHub relança o debate sobre o verdadeiro gargalo17/07/2026
  7. 7« Código Claude: Anatomia de uma Má-Feature » - quando a revisão pública se torna o verdadeiro QA17/07/2026
  8. 8O Google's Gemini 3.6 Flash é mais barato e mais curto - e o Gemini 4 ganha um teaser enquanto o 3.5 Pro continua atrasado22/07/2026
  9. 9A IA não tornou a programação mais fácil, apenas a tornou difícil de outra forma - CACM acerta na frase anti-hype22/07/2026
  10. 10"As IA estatal não resolverá a desigualdade": a tese crua do Rest of World sobre as IAs nacionais do Sul Global24/07/2026
  11. 11Refatoração como alavanca de custo de token: um experimento na série de gen-AI de Fowler30/07/2026
  12. 12Rachel Laycock: "a atenção tornou-se o recurso raro" - o dev-orquestrador, entre 8 e 12 agentes em paralelo31/07/2026
  13. 13Situational Awareness perde 67 % em um mês: o julgamento das verdadeiras crentes02/08/2026
  14. 14OpenAI « Astra » teria resolvido 10 problemas abertos em matemática e ciência da computação — aguardemos as provas02/08/2026
  15. 15« Cancelling Cursor »: a dívida técnica sobrepõe-se à velocidade de entrega de funcionalidades02/08/2026
  16. 16Jeff Dean sobre o que as equipes de IA erram: o diagnóstico da oficina que paga todas as contas03/08/2026
Your Linux servers, as a desktop.
TermalOSSponsored
Ops, reimagined

Your Linux servers, as a desktop.

Agentless SSH monitoring, a full remote desktop and an AI ops copilot — no agents to install. Everything stays on your machine.

Get early access
Secções
Explorar
Informações