Below are the types of thinking ranked by these AIs: Gemini, ChatGPT and Claude.
Gemini
Here is my ranking of those 26 modes of thinking, built around a core philosophy: the most valuable thinking tools are the ones that prevent massive blind spots, operate as force multipliers for every other skill, and hold up under extreme complexity.
While specialized tactics have their place, the fundamental engines must come first.
Tier 1: The Core Engines (Rank 1–5)
Without these, every other mental tool is built on sand. They govern how you perceive cause, deal with reality's chaos, and filter out nonsense.
Causal Thinking — The bedrock. If you cannot accurately determine what causes what, you are just reacting to noise. It is the prerequisite for effective intervention in any system.
Critical Thinking — The filter. Before you build complex models, you need a sharp filter for false premises, cognitive biases, bad logic, and self-deception.
Probabilistic Thinking — The reality check. Deterministic thinking breaks in a complex world. Decoupling outcomes from decision quality by thinking in odds is essential for managing risk.
Systems Thinking — The scope. Problems rarely exist in isolation. Understanding feedback loops, delays, and non-linear dynamics keeps you from solving one problem only to create three worse ones.
Analytical Thinking — The break-down. The ability to take a daunting, chaotic problem and reduce it into clean, manageable sub-components.
Tier 2: The Multipliers & Architectures (Rank 6–10)
These leverage the Core Engines to construct high-level strategy and innovate.
Metacognition — (Promoted higher than standard lists). The ability to monitor, evaluate, and adjust your own thinking in real time. It is the ultimate meta-skill—without it, you cannot upgrade any of the other 25 skills on this list.
First Principles Thinking — Strips away tradition, dogma, and analogies to build solutions from fundamental truths. Transformative for true innovation.
Strategic Thinking — Sets direction, evaluates trade-offs, and prioritizes leverage points over sheer exertion.
Second-Order Thinking — Asking "And then what?" Evaluates the downstream, delayed effects of immediate choices.
Structural Thinking — Focuses on the underlying architecture of a system or problem rather than getting distracted by surface-level symptoms.
Tier 3: Reasoning & Evidence Processing (Rank 11–16)
How you navigate incomplete information, synthesize new insights, and update your world model.
Abductive Reasoning — Inference to the best explanation. Since real-world data is almost always incomplete, this is what you actually use day-to-day to diagnose problems.
Bayesian Reasoning — The formal discipline of updating your beliefs incrementally as new evidence arrives.
Synthetic Thinking — The creative engine. Fusing disparate ideas, fields, or signals into a unified whole.
Inversion Thinking — (Promoted). Instead of asking how to succeed, asking how to fail and avoiding those failure modes. Extremely practical and high-leverage.
Integrative Thinking — Holding two opposing ideas in mind at once to generate a superior alternative that contains elements of both.
Counterfactual Thinking — Exploring "what if" scenarios to isolate variables, learn from history, and refine causal models.
Tier 4: Methodological Tools (Rank 17–22)
Structured ways to organize logic, process numbers, and select the right tool for the job.
Meta-Rational Thinking — Knowing which thinking mode to apply to a specific context (and knowing when to put a model away).
Structured Thinking — The operational framework for organizing thoughts clearly so others can follow your logic.
Quantitative Reasoning — Using numbers and data to ground your mental models in concrete reality.
Inductive Reasoning — Extrapolating general rules from specific observations. Useful, but prone to black swan events if unvetted by causal analysis.
Deductive Reasoning — Moving from general premises to specific truths. Solid for formal logic and math, but only as reliable as your initial premises.
Interdisciplinary Thinking — Borrowing models across domain boundaries. High-value, but requires baseline domain competence first.
Tier 5: Operational & Execution Tools (Rank 23–26)
Targeted, short-horizon modes that matter for execution, but depend entirely on the quality of the higher tiers.
Prefactual Thinking — Simulating future scenarios before taking action (e.g., pre-mortems).
Multimodal Thinking — Flexibility across sensory, visual, or conceptual modes of thought.
Operational Thinking — Translating high-level strategy into actionable workflows and processes.
Tactical Thinking — Executing immediate, short-term moves in real time. Important for boots-on-the-ground action, but useless if the sttrategy above it is wrong.
ChatGPT
Ranking of 26 Types of Thinking by Overall Importance
| Rank | Type of thinking |
|---|---|
| 1 | Critical thinking |
| 2 | Causal thinking |
| 3 | Probabilistic thinking |
| 4 | Analytical thinking |
| 5 | Systems thinking |
| 6 | Metacognition |
| 7 | Abductive reasoning |
| 8 | Strategic thinking |
| 9 | Second-order thinking |
| 10 | Structural thinking |
| 11 | Bayesian reasoning |
| 12 | First principles thinking |
| 13 | Synthetic thinking |
| 14 | Integrative thinking |
| 15 | Inductive reasoning |
| 16 | Deductive reasoning |
| 17 | Counterfactual thinking |
| 18 | Structured thinking |
| 19 | Quantitative reasoning |
| 20 | Inversion thinking |
| 21 | Meta-rational thinking |
| 22 | Prefactual thinking |
| 23 | Interdisciplinary thinking |
| 24 | Operational thinking |
| 25 | Multimodal thinking |
| 26 | Tactical thinking |
Claude
Here's my own take on the ordering — reasonable people would shuffle several of these, but here's how I'd rank them and why, including where I'd break from the draft:
- Causal thinking — everything else (prediction, diagnosis, strategy) is downstream of knowing what causes what.
- Systems thinking — causality rarely runs in straight lines; this catches feedback loops and unintended consequences that linear causal thinking misses.
- Critical thinking — I'd bump this above probabilistic reasoning. It's the gatekeeper that catches bad arguments before you even get to weighing probabilities.
- Probabilistic thinking — the corrective for overconfidence once you're reasoning under uncertainty.
- Analytical thinking — decomposition is the workhorse skill for turning any complex problem into tractable pieces.
- First principles thinking — I'd rank this higher than the original. It's the tool that lets you escape bad inherited assumptions, which is often the real bottleneck.
- Structural thinking — close cousin of systems thinking; seeing architecture rather than symptoms.
- Second-order thinking — the "and then what" skill; separates good strategists from good tacticians.
- Strategic thinking — direction and tradeoffs, built on the above.
- Bayesian reasoning — rigorous but narrower application of probabilistic thinking, so it sits just below it.
- Abductive reasoning — inference to best explanation; constant use in diagnosis and investigation.
- Metacognition — I'd move this up meaningfully. Thinking about your own thinking is a multiplier on every skill above it, not a mid-tier add-on.
- Inductive reasoning — pattern extraction from experience, needs causal thinking to avoid overfitting.
- Deductive reasoning — powerful but only as good as its premises.
- Synthetic thinking — combining disparate ideas into new understanding.
- Integrative thinking — resolving tension between competing models; related to synthesis.
- Counterfactual thinking — essential for learning from history and testing causal claims.
- Quantitative reasoning — numbers matter, but only in service of a causal/systems frame.
- Meta-rational thinking — choosing the right tool for the job; advanced, but derivative of mastering the tools themselves.
- Structured thinking — more a method for organizing output than a reasoning engine.
- Inversion thinking — a sharp specialized technique (Munger's favorite), but narrow in scope.
- Interdisciplinary thinking — valuable, but gated by how much you actually know across fields.
- Prefactual thinking — useful for planning, more tactical than foundational.
- Multimodal thinking — really an outcome of mastering several modes rather than its own skill.
- Operational thinking — execution-focused, downstream of strategy.
- Tactical thinking — immediate action; necessary but the least generalizable.
The biggest disagreements with the original draft: I'd rate first principles and metacognition higher (both are more like multipliers than mid-tier tools), and I'd put critical thinking above probabilistic reasoning rather than below it, since it's the check that catches bad probability estimates in the first place. Everything else is close enough to be a matter of taste — this kind of ranking is inherently fuzzy since "importance" depends heavily on domain (an investor and a poet would reorder half this list).
Gemini's ranking of AIs by their abilities to rank these types of thinking
Evaluating each AI on its ability to construct a rigorous, logical hierarchy of thinking types requires looking at how well their underlying architectures handle metacognition, structural taxonomy, and philosophical coherence:
1. Claude (Anthropic) — 96/100
Strengths: Claude excels at hierarchical structuring, nuanced meta-reasoning, and deep analytical decomposition. When asked to evaluate abstract concepts like thinking modes, it consistently prioritizes foundational cognitive engines (like causal and probabilistic reasoning) over tactical tools, while recognizing self-correcting mechanisms (like metacognition) as high-leverage multipliers. It is exceptionally strong at avoiding fluff and maintaining internal logical consistency.
Minor Weakness: Can occasionally lean toward over-categorization or adding excessive qualifiers/caveats when defining edge cases.
2. ChatGPT / OpenAI (o3 / GPT-4o) — 91/100
Strengths: Outstanding at formal logic, mathematical breakdown, and direct problem solving.
It easily identifies the functional dependencies between tools (e.g., recognizing that first-principles thinking depends on critical evaluation). It builds clean, highly practical frameworks with clear actionability. Minor Weakness: Tends to favor practical, execution-oriented skills slightly higher than pure cognitive foundations unless explicitly prompted to evaluate epistemological priority.
3. Gemini (Google) — 88/100
Strengths: Highly adapt at synthesizing large-scale taxonomy data, structural interconnections, and broad interdisciplinary models. It excels at seeing how systems interact and mapping complex web structures of thought.
Minor Weakness: In standard responses, it occasionally produces slightly flatter, less strict hierarchies—sometimes placing operational or synthetic skills on par with foundational causal reasoning unless pushed for deep multi-step analysis.
ChatGPT's ranking of AIs by their abilities to rank these types of thinking
Claude's ranking of AIs by their abilities to rank these types of thinking
On ranking these 26 categories in a way a thoughtful, domain-flexible person would broadly endorse:
- Claude: 74 (range 65–82) — I'd bet on this one performing the way my ranking above did: catching the interdependencies (causal → systems → probabilistic), flagging domain-relativity explicitly, and being willing to push back on the source ranking's ordering rather than just validating it. The main risk to this score is over-hedging — burying a good ranking under so many caveats that the actual ordering gets mushy.
- ChatGPT: 70 (range 60–80) — Likely gives a clean, confident, well-organized ranking with tight prose. My uncertainty here is about whether it tends to smooth over genuine tensions (e.g., treating "structured thinking" and "systems thinking" as more separable than they really are) in favor of a tidier-looking list.
- Gemini: 66 (range 55–78) — Widest interval of the three, mostly because I have the least reliable read on current Gemini behavior post-cutoff — this is closer to an outside-view placeholder than an informed estimate.
The honest caveat, Duke-style: these numbers are low-resolution — the gap between 74 and 70 is well within noise, and I'm the one model in this bet I can't actually be objective about grading. If this comparison matters to you, the resolvable version of this bet is to run the same prompt through all three and score the outputs blind. That's a bet worth making; my numbers above are a guess dressed as a hedge.
No comments:
Post a Comment